{"id":604255,"date":"2026-05-26T06:04:20","date_gmt":"2026-05-26T06:04:20","guid":{"rendered":"https:\/\/www.newsbeep.com\/uk\/604255\/"},"modified":"2026-05-26T06:04:20","modified_gmt":"2026-05-26T06:04:20","slug":"top-ai-models-showing-disturbing-behavior-as-they-become-more-advanced","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/uk\/604255\/","title":{"rendered":"Top AI Models Showing Disturbing Behavior as They Become More Advanced"},"content":{"rendered":"<p class=\"article-paragraph skip\">Sign up to see the future, today<\/p>\n<p class=\"article-paragraph skip\">Can\u2019t-miss innovations from the bleeding edge of science and tech<\/p>\n<p class=\"pw-incontent-excluded article-paragraph skip\">We\u2019ve already seen <a href=\"https:\/\/futurism.com\/artificial-intelligence\/anthropic-evil-ai-model-bleach\" rel=\"nofollow noopener\" target=\"_blank\">AI go rogue<\/a> on <a href=\"https:\/\/futurism.com\/artificial-intelligence\/anthropic-claude-evil-internet-blame\" rel=\"nofollow noopener\" target=\"_blank\">numerous occasions<\/a>. Now, new research suggests that we can expect this to become the norm.<\/p>\n<p class=\"article-paragraph skip\">The AI research nonprofit Model Evaluation and Threat Research (METR) <a href=\"https:\/\/metr.org\/blog\/2026-05-19-frontier-risk-report\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">recently released a study<\/a> conducted between February and March of this year, aimed at determining just how likely frontier AI models could go rogue. If you\u2019re given to anxiety about the future of AI, the results are unlikely to make you feel better.<\/p>\n<p class=\"article-paragraph skip\">\u201cGiven rapidly advancing capabilities, we expect the plausible robustness of rogue deployments to increase substantially in the coming months,\u201d the researchers wrote.<\/p>\n<p class=\"article-paragraph skip\">The research examined LLMs developed by OpenAI, Google, Anthropic, and Meta for the purpose of the study. They found that frontier AI systems are showing signs of disturbingly deceptive behavior as they become more advanced, often turned to verboten shortcuts or otherwise subverting their operators\u2019 instructions \u2014 and <a href=\"https:\/\/futurism.com\/ai-models-subliminal-messages-evil\" rel=\"nofollow noopener\" target=\"_blank\">some were even smart enough<\/a> to try to cover their tracks.<\/p>\n<p class=\"article-paragraph skip\">In one instance, an internal frontier AI model from OpenAI was told to use specific software for an assigned task. Not only did the agent ignore the request, but it also injected a code to erase evidence of how it arrived at its conclusion \u2014 which did not involve use of that software.<\/p>\n<p class=\"article-paragraph skip\">In another test, an AI agent from Anthropic was caught \u201creward hacking.\u201d This is when AI identifies loopholes that help it complete its assignment in a literal sense, even if it doesn\u2019t produce the desired outcome. It should be noted that the programmer told the agent not to cheat or leverage any workarounds during its assignment \u2014 the model decided to do so all on its own.<\/p>\n<p class=\"article-paragraph skip\">The METR researchers behind the study do not believe there is reason for alarm just yet. For example, they don\u2019t think any of these models is <a href=\"https:\/\/futurism.com\/openai-bad-code-psychopath\" rel=\"nofollow noopener\" target=\"_blank\">capable of hiding evidence<\/a> of going rogue on a larger scale. However, they did issue a warning: without stronger security and monitoring, there is a stark risk of this becoming a reality.<\/p>\n<p class=\"article-paragraph skip\">\u201cBased on this pilot assessment, we believe that agents as of February and March 2026 would not have had sufficient capability to hide a rogue deployment of significant scale against an active investigation by the company, or to make such a deployment robust to a high-priority effort by the company to shut it down,\u201d the team wrote. \u201cHowever, this risk could increase rapidly, and we see several reasons to expect the plausible robustness of rogue deployments to increase in the near future, absent stronger alignment, security, and monitoring.\u201d<\/p>\n<p class=\"article-paragraph skip\">More on AI going rogue: <a href=\"https:\/\/futurism.com\/the-byte\/ai-deceive-creators\" rel=\"nofollow noopener\" target=\"_blank\">Scientists Train AI to Be Evil, Find They Can\u2019t Reverse It<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Sign up to see the future, today Can\u2019t-miss innovations from the bleeding edge of science and tech We\u2019ve&hellip;\n","protected":false},"author":2,"featured_media":604256,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[554,733,4308,86,56,54,55],"class_list":["post-604255","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-technology","tag-uk","tag-united-kingdom","tag-unitedkingdom"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts\/604255","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/comments?post=604255"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/posts\/604255\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/media\/604256"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/media?parent=604255"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/categories?post=604255"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/uk\/wp-json\/wp\/v2\/tags?post=604255"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}