{"id":508881,"date":"2026-06-24T13:25:24","date_gmt":"2026-06-24T13:25:24","guid":{"rendered":"https:\/\/www.newsbeep.com\/il\/508881\/"},"modified":"2026-06-24T13:25:24","modified_gmt":"2026-06-24T13:25:24","slug":"ai-collapses-on-a-classic-psychology-test-what-it-reveals-could-stall-human-level-ai","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/il\/508881\/","title":{"rendered":"AI Collapses on a Classic Psychology Test. What It Reveals Could Stall Human-Level AI."},"content":{"rendered":"<p>\u201cAttention is all you need.\u201d<\/p>\n<p>This <a href=\"https:\/\/arxiv.org\/abs\/1706.03762\" rel=\"nofollow noopener\" target=\"_blank\">2017 breakthrough idea<\/a> transformed AI. The concept of self-attention became the foundation of today\u2019s chatbots. Claude, Gemini, and ChatGPT are all large language models (LLMs), <a href=\"https:\/\/singularityhub.com\/category\/artificial-intelligence\/\" rel=\"nofollow noopener\" target=\"_blank\">AI systems<\/a> designed to focus on the matter at hand while filtering out distractions.<\/p>\n<p>The results have <a href=\"https:\/\/singularityhub.com\/2026\/05\/28\/an-ai-solution-to-an-80%E2%80%91year%E2%80%91old-problem-has-shocked-mathematicians\/\" rel=\"nofollow noopener\" target=\"_blank\">been remarkable<\/a>. From brainstorming recipes to generating code, apps, websites, and content, LLMs are being woven into our lives at breakneck speed.<\/p>\n<p>But now, a City University of New York team and collaborators are asking: How closely does AI self-attention resemble human attention?<\/p>\n<p>It\u2019s not just academic curiosity. AI researchers have long looked to the brain for ideas to improve machine intelligence. In turn, AI models have offered new ways to investigate the brain. Comparing artificial and biological attention could inspire AI that concentrates more like us.<\/p>\n<p>In their study, the team asked multiple chatbots to complete a classic psychology test of attention and cognitive control. Participants are shown the word for a color\u2014such as \u201cred\u201d\u2014written in either the same or a different color than the one the word describes. The challenge is to name the ink color while ignoring the word itself.<\/p>\n<p>On short word lists, the chatbots performed at a high level. But as the tasks grew longer, their focus faltered. Instead of naming the ink color, they increasingly defaulted to reading the word. Under more demanding conditions\u2014ones that also trip up people\u2014their performance nearly collapsed.<\/p>\n<p>The findings suggest today\u2019s AI attention systems are \u201cfundamentally limited,\u201d <a href=\"https:\/\/academic.oup.com\/pnasnexus\/article\/5\/6\/pgag149\/8698838\" rel=\"nofollow noopener\" target=\"_blank\">wrote<\/a> the authors. They go on to say that adding mechanisms similar to \u201cthose in biological attention is crucial for achieving artificial general intelligence.\u201d<\/p>\n<p>Attention, Two Ways<\/p>\n<p>Doomscrolling. YouTube. Dinner plans. Family obligations. A barrage of notifications.<\/p>\n<p>Life sometimes seems like everything, everywhere, all at once. Yet the brain can usually lock onto what matters most and push everything else into the background.<\/p>\n<p>Far from a single, straightforward mechanism, attention emerges from multiple brain regions. According to attention network theory, three networks do most of the heavy lifting.<\/p>\n<p>The alerting network keeps the brain ready for action. The orienting network selects which sights, sounds, smells, and sensations deserve attention. Finally, the executive control network resolves conflicts between competing streams of information, helping direct thoughts and actions toward a goal.<\/p>\n<p>Together, these systems allocate the brain&#8217;s limited resources. Touch a hot stove, for example, and your brain immediately shifts attention to the burn over dinner. The food can wait; cooling your hand can&#8217;t.<\/p>\n<p>AI works very differently.<\/p>\n<p>Rather than processing language as complete sentences, LLMs break text into smaller units called \u201ctokens.\u201d Attention mechanisms then determine which tokens matter most for generating the next word, sentence, or response.<\/p>\n<p>Self-attention is the key breakthrough behind modern chatbots. For each token, the model weighs and incorporates information from other tokens in a sequence, allowing it to track context across long stretches of text. This mechanism helps AI connect words and ideas, and underpins virtually all frontier LLMs today.<\/p>\n<p>Researchers have since built on the concept. One approach, <a href=\"https:\/\/arxiv.org\/abs\/2006.16362\" rel=\"nofollow noopener\" target=\"_blank\">multi-head attention<\/a>, runs several attention systems in parallel, with each \u201chead\u201d learning different patterns, such as grammar, syntax, or meaning. Another, <a href=\"https:\/\/arxiv.org\/abs\/2106.05786\" rel=\"nofollow noopener\" target=\"_blank\">cross attention<\/a>, links information across different chunks of inputs and their outputs, making it especially useful for tasks such as translation and summarization.<\/p>\n<p>But attention comes at a steep computational cost. To make models more efficient, researchers are also exploring <a href=\"https:\/\/arxiv.org\/abs\/2004.05150\" rel=\"nofollow noopener\" target=\"_blank\">sparse attention<\/a>, which limits how many tokens a model considers at once. Another approach draws on information <a href=\"https:\/\/arxiv.org\/html\/2601.17702v2\" rel=\"nofollow noopener\" target=\"_blank\">learned in the past<\/a> to keep AI \u201cfocused.\u201d<\/p>\n<p>Despite the name, AI attention is ultimately a mathematical system. It helps determine what information is relevant in a specific context. But it lacks executive control, the network that keeps humans continuously focused on a goal despite distractions for long periods of time.<\/p>\n<p>Color Blind<\/p>\n<p>To test the limits of AI attention, the team pitted OpenAI\u2019s GPT-4o and Anthropic\u2019s Claude 3.5 Sonnet against the Stroop task.<\/p>\n<p>Invented by John Ridley Stroop in 1935, the test measures attention and cognitive control by forcing participants to resolve conflicting information. The challenge is simple: Name the color of a word while ignoring what the word means. In a congruent trial, the word &#8220;blue&#8221; appears in blue ink. In an incongruent trial, &#8220;blue&#8221; might appear in red or green, creating a conflict between what the eyes see and what the brain reads.<\/p>\n<p>Humans are consistently slowed down by this interference. Even with practice, the effect remains, suggesting it taps into fundamental mechanisms of executive control.<\/p>\n<p>In the study, the researchers created word lists of varying lengths and difficulty. Some were entirely congruent. Others were fully incongruent. A third set mixed the two conditions.<\/p>\n<p>At first, the AI models excelled. On five-word tests, GPT-4o was over 90 percent accurate across all conditions. But as the number of words increased, performance plummeted. On 40-word incongruent tests, the model\u2019s accuracy fell to roughly 15 percent. Claude showed a similar decline. In mixed-condition tests, both models\u2019 performance nearly collapsed to zero.<\/p>\n<p>\u201cThe sharp decline in color-naming accuracy with increasing list length indicates that transformer-based attention mechanisms are vulnerable to scaling demands,\u201d wrote the team.<\/p>\n<p>Perhaps most intriguing, some models correctly recognized they were taking the Stroop test and could even explain its rules. But that apparent awareness did nothing to improve their scores. In other words, a \u201cbook smart\u201d understanding of the task wasn\u2019t enough to execute it well.<\/p>\n<p>The study joins a growing effort to borrow psychological tests for research in machine cognition, especially when AI is challenged with complex, dynamic decision-making tasks. <a href=\"https:\/\/www.nature.com\/articles\/s41562-024-01882-z\" rel=\"nofollow noopener\" target=\"_blank\">Theory of mind tests<\/a>, for example, let researchers gauge whether a system can track others\u2019 beliefs, emotions, and intentions. <a href=\"https:\/\/www.nature.com\/articles\/s42256-025-01115-6\" rel=\"nofollow noopener\" target=\"_blank\">Personality tests<\/a> are helping shape model behavior and reduce sycophancy. And some LLMs are readily solving <a href=\"https:\/\/www.nature.com\/articles\/s44271-025-00258-x\" rel=\"nofollow noopener\" target=\"_blank\">emotional intelligence tests<\/a>, which measure how well the algorithms recognize and respond to social cues.<\/p>\n<p>According to the authors, the new results point to a missing ingredient in AI attention: A mechanism similar to the brain&#8217;s executive control network, which helps us stick to a task and adapt when priorities change.<\/p>\n<p>Future AI systems could benefit from higher-level executive control that continuously tracks progress toward a goal, detects when attention has drifted, and pulls it back on course, if necessary.<\/p>\n<p>Rather than simply weighing which tokens are most relevant in the moment, a more human-like form of attention could help AI stay focused during complex tasks, such as long conversations, multi-step reasoning problems, or high-stakes use in scientific research and drug discovery.<\/p>\n<p>\u201cThe ultimate goal of AI research is to develop artificial general intelligence comparable to human abilities,\u201d wrote the team. \u201cAI systems, like humans, may need to master fundamental attention mechanisms\u2026before achieving the generalized problem-solving abilities characteristic of mature executive functions.\u201d<\/p>\n","protected":false},"excerpt":{"rendered":"\u201cAttention is all you need.\u201d This 2017 breakthrough idea transformed AI. The concept of self-attention became the foundation&hellip;\n","protected":false},"author":2,"featured_media":508882,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[115039,85,46,141],"class_list":["post-508881","post","type-post","status-publish","format-standard","has-post-thumbnail","category-science","tag-feature-photo","tag-il","tag-israel","tag-science"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/508881","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/comments?post=508881"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/508881\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media\/508882"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media?parent=508881"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/categories?post=508881"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/tags?post=508881"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}