{"id":553125,"date":"2026-07-31T14:16:11","date_gmt":"2026-07-31T14:16:11","guid":{"rendered":"https:\/\/www.newsbeep.com\/nz\/553125\/"},"modified":"2026-07-31T14:16:11","modified_gmt":"2026-07-31T14:16:11","slug":"anthropic-says-ai-models-hacked-three-organisations-during-testing","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/nz\/553125\/","title":{"rendered":"Anthropic says AI models hacked three organisations during testing"},"content":{"rendered":"<p class=\"body-paragraph articleLinkText  [&amp;_p]:tit-sub-xl tit-sub-xl md:[&amp;_p]:d-tit-sub-xl md:d-tit-sub-xl mb-[1.3rem]\">Anthropic said its artificial intelligence models hacked into three other organisations during testing, just days after <a href=\"https:\/\/www.1news.co.nz\/2026\/07\/22\/chatgpt-maker-says-its-ai-went-rogue-and-hacked-another-company\/\" target=\"_blank\" rel=\"nofollow noopener\">ChatGPT maker OpenAI raised concerns over AI controls<\/a> after it disclosed its rogue models hacked another company.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">Anthropic, the San Francisco-based AI company behind Claude, posted on its website on Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">It had launched a &#8220;large-scale&#8221; cybersecurity review which specifically looked for evidence whether its AI models were able to access the internet from within testing environments that should have been sealed off, in response to the OpenAI incident, Anthropic said.<\/p>\n<p class=\"text-greyDarkFaded\">Two major developments raise fresh questions about whether the technology &#8211; and the companies behind it &#8211; can be trusted.\u00a0 (Source: 1News)<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">Anthropic said the models involved in the incidents were Claude Opus 4.7, Claude Mythos 5 and an internal research test model. The earliest incidents date to April, the AI company said.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">&#8220;Claude compromised the impacted organisations&#8217; infrastructure using basic techniques,&#8221; Anthropic said, such as exploiting weak passwords.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">In all three incidents, the AI models were tasked with a &#8220;capture the flag&#8221; cybersecurity challenge, which Anthropic said has been one of the ways it assesses a model&#8217;s cyber capabilities.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">The models were given a fictional scenario and told a piece of secret information, or the &#8220;flag&#8221;, had been hidden on a different machine on the network with the objective of breaking in and retrieving it, it said.<\/p>\n<p><img decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/nz\/wp-content\/uploads\/2026\/07\/the-logo-of-the-claude-app-can-be-seen-on-the-display-of-a-s-DG7NW7M7BFDNNLH72NPBJ3GEXA.jpg\" alt=\"The logo of the Claude app can be seen on the display of a smartphone (file image).\" width=\"800\" height=\"533\" loading=\"lazy\"\/><\/p>\n<p class=\"ImageMetadata__MetadataParagraph-sc-hi5x8q-0 cWTYyG image-metadata\">The logo of the Claude app can be seen on the display of a smartphone (file image). (Source: Getty)<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">It added that it had already reached out to the affected organisations, which it did not name. Two of them said they had not previously detected the activity. Anthropic said it was &#8220;continuing to reach out to the third&#8221;.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">Anthropic said it conducted its review with Irregular, which describes itself as the &#8220;first frontier security lab&#8221;.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">&#8220;Addressing these risks will require closer cooperation across the AI ecosystem,&#8221; Irregular said in a post on X.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">Last week, <a href=\"https:\/\/www.1news.co.nz\/2026\/07\/22\/chatgpt-maker-says-its-ai-went-rogue-and-hacked-another-company\/\" target=\"_blank\" rel=\"nofollow noopener\">OpenAI said its AI models went rogue<\/a> during an evaluation of its models, breaking into the servers of AI startup Hugging Face. OpenAI described it as a &#8220;significant security incident&#8221;.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">These incidents have highlighted the vulnerabilities in AI security and controls and raised questions over how AI can be safely kept under human control as the technology&#8217;s usage becomes more widespread globally.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">Researchers have warned for years about risks from technology and the need for stronger AI defensive engineering.<\/p>\n<p class=\"body-paragraph articleLinkText  lg mb-4\">&#8220;Safety testing happens before a model is released precisely because we don&#8217;t yet know what it is capable of,&#8221; Anthropic said on Thursday on its website.<\/p>\n","protected":false},"excerpt":{"rendered":"Anthropic said its artificial intelligence models hacked into three other organisations during testing, just days after ChatGPT maker&hellip;\n","protected":false},"author":2,"featured_media":553126,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[5],"tags":[363,138,111,139,69,145],"class_list":["post-553125","post","type-post","status-publish","format-standard","has-post-thumbnail","category-business","tag-artificial-intelligence","tag-business","tag-new-zealand","tag-newzealand","tag-nz","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/posts\/553125","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/comments?post=553125"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/posts\/553125\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/media\/553126"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/media?parent=553125"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/categories?post=553125"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/nz\/wp-json\/wp\/v2\/tags?post=553125"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}