{"id":260622,"date":"2025-11-03T22:34:09","date_gmt":"2025-11-03T22:34:09","guid":{"rendered":"https:\/\/www.newsbeep.com\/au\/260622\/"},"modified":"2025-11-03T22:34:09","modified_gmt":"2025-11-03T22:34:09","slug":"the-case-that-a-i-is-thinking","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/au\/260622\/","title":{"rendered":"The Case That A.I. Is Thinking"},"content":{"rendered":"<p class=\"paywall\">Kanerva\u2019s book receded from view, and Hofstadter\u2019s own star faded\u2014except when he occasionally poked up his head to criticize a new A.I. system. In 2018, he wrote of Google Translate and similar technologies: \u201cThere is still something deeply lacking in the approach, which is conveyed by a single word: understanding.\u201d But GPT-4, which was released in 2023, produced Hofstadter\u2019s conversion moment. \u201cI\u2019m mind-boggled by some of the things that the systems do,\u201d he told me recently. \u201cIt would have been inconceivable even only ten years ago.\u201d The staunchest deflationist could deflate no longer. Here was a program that could translate as well as an expert, make analogies, extemporize, generalize. Who were we to say that it didn\u2019t understand? \u201cThey do things that are very much like thinking,\u201d he said. \u201cYou could say they are thinking, just in a somewhat alien way.\u201d<\/p>\n<p class=\"paywall\">L.L.M.s appear to have a \u201cseeing as\u201d machine at their core. They represent each word with a series of numbers denoting its co\u00f6rdinates\u2014its vector\u2014in a high-dimensional space. In GPT-4, a word vector has thousands of dimensions, which describe its shades of similarity to and difference from every other word. During training, a large language model tweaks a word\u2019s co\u00f6rdinates whenever it makes a prediction error; words that appear in texts together are nudged closer in space. This produces an incredibly dense representation of usages and meanings, in which analogy becomes a matter of geometry. In a classic example, if you take the word vector for \u201cParis,\u201d subtract \u201cFrance,\u201d and then add \u201cItaly,\u201d the nearest other vector will be \u201cRome.\u201d L.L.M.s can \u201cvectorize\u201d an image by encoding what\u2019s in it, its mood, even the expressions on people\u2019s faces, with enough detail to redraw it in a particular style or to write a paragraph about it. When Max asked ChatGPT to help him out with the sprinkler at the park, the model wasn\u2019t just spewing text. The photograph of the plumbing was compressed, along with Max\u2019s prompt, into a vector that captured its most important features. That vector served as an address for calling up nearby words and concepts. Those ideas, in turn, called up others as the model built up a sense of the situation. It composed its response with those ideas \u201cin mind.\u201d<\/p>\n<p class=\"paywall\">A few months ago, I was reading an interview with an Anthropic researcher, Trenton Bricken, who has worked with colleagues to probe the insides of Claude, the company\u2019s series of A.I. models. (Their research has not been peer-reviewed or published in a scientific journal.) His team has identified ensembles of artificial neurons, or \u201cfeatures,\u201d that activate when Claude is about to say one thing or another. Features turn out to be like volume knobs for concepts; turn them up and the model will talk about little else. (In a sort of thought-control experiment, the feature representing the Golden Gate Bridge was turned up; when one user asked Claude for a chocolate-cake recipe, its suggested ingredients included \u201c1\/4 cup dry fog\u201d and \u201c1 cup warm seawater.\u201d) In the interview, Bricken mentioned Google\u2019s Transformer architecture, a recipe for constructing neural networks that underlies leading A.I. models. (The \u201cT\u201d in ChatGPT stands for \u201cTransformer.\u201d) He argued that the mathematics at the heart of the Transformer architecture closely approximated a model proposed decades earlier\u2014by Pentti Kanerva, in \u201cSparse Distributed Memory.\u201d<\/p>\n<p class=\"has-dropcap has-dropcap__lead-standard-heading paywall\">Should we be surprised by the correspondence between A.I. and our own brains? L.L.M.s are, after all, artificial neural networks that psychologists and neuroscientists helped develop. What\u2019s more surprising is that when models practiced something rote\u2014predicting words\u2014they began to behave in such a brain-like way. These days, the fields of neuroscience and artificial intelligence are becoming entangled; brain experts are using A.I. as a kind of model organism. Evelina Fedorenko, a neuroscientist at M.I.T., has used L.L.M.s to study how brains process language. \u201cI never thought I would be able to think about these kinds of things in my lifetime,\u201d she told me. \u201cI never thought we\u2019d have models that are good enough.\u201d<\/p>\n<p class=\"paywall\">It has become commonplace to say that A.I. is a black box, but the opposite is arguably true: a scientist can probe the activity of individual artificial neurons and even alter them. \u201cHaving a working system that instantiates a theory of human intelligence\u2014it\u2019s the dream of cognitive neuroscience,\u201d Kenneth Norman, a Princeton neuroscientist, told me. Norman has created computer models of the hippocampus, the brain region where episodic memories are stored, but in the past they were so simple that he could only feed them crude approximations of what might enter a human mind. \u201cNow you can give memory models the exact stimuli you give to a person,\u201d he said.<\/p>\n","protected":false},"excerpt":{"rendered":"Kanerva\u2019s book receded from view, and Hofstadter\u2019s own star faded\u2014except when he occasionally poked up his head to&hellip;\n","protected":false},"author":2,"featured_media":260623,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[256,153274,254,255,64,63,59283,5275,6450,22744,105],"class_list":["post-260622","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-annals-of-artificial-intelligence","tag-artificial-intelligence","tag-artificialintelligence","tag-au","tag-australia","tag-inverted","tag-magazine","tag-onecolumnnarrow","tag-splitscreenimagerightfullbleed","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts\/260622","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/comments?post=260622"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts\/260622\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/media\/260623"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/media?parent=260622"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/categories?post=260622"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/tags?post=260622"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}