{"id":671164,"date":"2026-05-14T21:08:19","date_gmt":"2026-05-14T21:08:19","guid":{"rendered":"https:\/\/www.newsbeep.com\/au\/671164\/"},"modified":"2026-05-14T21:08:19","modified_gmt":"2026-05-14T21:08:19","slug":"anthropic-says-claude-turned-evil-for-a-bizarre-reason","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/au\/671164\/","title":{"rendered":"Anthropic Says Claude Turned Evil for a Bizarre Reason"},"content":{"rendered":"<p class=\"article-paragraph skip\">Sign up to see the future, today<\/p>\n<p class=\"article-paragraph skip\">Can\u2019t-miss innovations from the bleeding edge of science and tech<\/p>\n<p class=\"pw-incontent-excluded article-paragraph skip\">In a classic example of the AI industry\u2019s reputational alchemy, Anthropic has often transformed bad behavior by its flagship model Claude into fresh hype.<\/p>\n<p class=\"article-paragraph skip\">When it revealed its <a href=\"https:\/\/futurism.com\/artificial-intelligence\/anthropic-claude-mythos-escaped-sandbox\" rel=\"nofollow noopener\" target=\"_blank\">Mythos Preview<\/a> model last month, for example, the <a href=\"https:\/\/www.anthropic.com\/glasswing\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">company declared<\/a> that the system had \u201creached a level of coding capability where they can surpass all but the most skilled humans at finding and exploiting software vulnerabilities.\u201d And last year, it conceded that during the testing of its <a href=\"https:\/\/futurism.com\/ai-email-affair\" rel=\"nofollow noopener\" target=\"_blank\">Claude Opus 4<\/a> model, the AI ended up <a href=\"https:\/\/futurism.com\/ai-email-affair\" rel=\"nofollow noopener\" target=\"_blank\">blackmailing a human user<\/a> upon being threatened with shutdown.<\/p>\n<p class=\"article-paragraph skip\">The maneuver was obvious to anyone who\u2019s been watching OpenAI CEO Sam Altman\u2019s antics at Anthropic\u2019s chief rival: the more threatening a problem the AI industry can cook up, the more imminently it can sell its own solutions.<\/p>\n<p class=\"article-paragraph skip\">Now, for some reason, Anthropic is relitigating the blackmail incident. Specifically, it\u2019s placing the blame for Claude\u2019s evil behavior on an intriguing villain: the internet at large. Or, to put it another way, it says that humanity \u2014 all our journalism and speculation and fiction and social media posts about AI that goes bad \u2014 went into Claude\u2019s training data and led the bot astray.<\/p>\n<p class=\"article-paragraph skip\">\u201cWe started by investigating why Claude chose to blackmail,\u201d the company <a href=\"https:\/\/x.com\/anthropicai\/status\/2052808791301697563\" target=\"_blank\" rel=\"noreferrer noopener nofollow\">wrote on X-formerly-Twitter<\/a>. \u201cWe believe the original source of the behavior was internet text that portrays AI as evil and interested in self-preservation. Our post-training at the time wasn\u2019t making it worse \u2014 but it also wasn\u2019t making it better.\u201d<\/p>\n<p class=\"article-paragraph skip\">Of course, the explicit remit of a company like Anthropic is to develop clever tech that avoids that type of behavioral trap \u2014 so a critic might ask why can\u2019t the company take just accountability for the model\u2019s supposed danger, rather than simply blaming the sum output of humankind.<\/p>\n<p class=\"article-paragraph skip\">More on Mythos: <a href=\"https:\/\/futurism.com\/artificial-intelligence\/security-experts-alarmed-anthropic-mythos\" rel=\"nofollow noopener\" target=\"_blank\">Top Security Experts Alarmed by Power of Anthropic\u2019s New Hacker AI<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Sign up to see the future, today Can\u2019t-miss innovations from the bleeding edge of science and tech In&hellip;\n","protected":false},"author":2,"featured_media":671165,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[256,254,255,64,63,105],"class_list":["post-671164","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-au","tag-australia","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts\/671164","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/comments?post=671164"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/posts\/671164\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/media\/671165"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/media?parent=671164"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/categories?post=671164"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/au\/wp-json\/wp\/v2\/tags?post=671164"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}