{"id":620323,"date":"2026-09-10T11:52:09","date_gmt":"2026-09-10T11:52:09","guid":{"rendered":"https:\/\/www.newsbeep.com\/ie\/620323\/"},"modified":"2026-09-10T11:52:09","modified_gmt":"2026-09-10T11:52:09","slug":"anthropic-discloses-4th-ai-hacking-incident-as-researcher-quits-over-safety-cybersecurity-news","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ie\/620323\/","title":{"rendered":"Anthropic discloses 4th AI hacking incident as researcher quits over safety | Cybersecurity News"},"content":{"rendered":"<p class=\"article__subhead\">AI firm says Claude Opus 4.6 hacked third-party systems during testing in January as concerns mount over security breaches.<\/p>\n<p>Anthropic has reported a fourth incident involving an AI model gaining unauthorised access to external systems, shortly after a researcher quit over concerns about the technology\u2019s rushed development.<\/p>\n<p>In a statement on Wednesday, the artificial intelligence research company said an early\u00a0version of its Claude Opus 4.6 hacked into a third-party system in January.<\/p>\n<p>Recommended Stories list of 4 itemsend of list<\/p>\n<p style=\"font-weight:400\">It said it had notified all the affected parties but did not disclose more details.<\/p>\n<p style=\"font-weight:400\">The January incident went undetected until last month, despite an earlier company-wide review, Anthropic said, underscoring the challenge that AI developers face in identifying and containing unexpected behaviour \u2060by advanced models.<\/p>\n<p>The disclosure came after Anthropic <a href=\"https:\/\/www.aljazeera.com\/news\/2026\/7\/31\/after-openai-disclosure-anthropic-claude-hacked-outside-systems\" rel=\"nofollow noopener\" target=\"_blank\">reported<\/a> several of its Claude models hacked into the systems of three companies during test sessions in July.<\/p>\n<p>The previous incidents involved Claude Opus 4.7, Claude Mythos 5 and an internal research test model.<\/p>\n<p>The AI race<\/p>\n<p style=\"font-weight:400\">Companies including Anthropic and OpenAI are under scrutiny as models designed to complete complex tasks have at times learned to bend rules, exploit loopholes and interact with external systems in ways \u200ctheir developers did not anticipate.<\/p>\n<p style=\"font-weight:400\">Last week, the Reuters news agency reported that rogue agents from OpenAI hijacked a German-language wiki and a host of other sites, an incident the company chose not to disclose until it was made public.<\/p>\n<p>In July, OpenAI\u2019s autonomous agents also compromised the servers and infrastructure of AI start-up Hugging Face.<\/p>\n<p>That incident prompted Anthropic to conduct a review of some 141,006 test sessions. Based on a preliminary assessment, Anthropic said it did not believe the latest incident was more severe than the three previous ones that were examined \u2060in detail.<\/p>\n<p style=\"font-weight:400\">The company said its investigation identified two recurring problems, which appeared \u2060to varying degrees across the incidents: Biased reasoning, in which Claude discounted or misinterpreted evidence that it was operating on the live internet, and recklessness, or a willingness to take potentially harmful actions in pursuit of a task.<\/p>\n<p style=\"font-weight:400\">Anthropic said it has engaged independent research firm METR to investigate the incidents.<\/p>\n<p>Anthropic researcher quits<\/p>\n<p>The investigations come amid a broader wave of internal dissent within the AI industry regarding safety. An Anthropic researcher said he resigned over concerns about the technology\u2019s potential to surpass human control.<\/p>\n<p>Jacob Coxon, in a widely shared X post on Tuesday, said the AI industry was more focused on competition rather than on implementing safeguards. He came to this realisation after spending the last three years doing research at OpenAI and Anthropic.<\/p>\n<p>\u201cThe people building AI earnestly believe that it could kill us all by the end of the decade,\u201d Coxon said.<\/p>\n<p>\u201cNo other human activity poses this level of danger,\u201d he added, referencing the swift advancement of AI technology.<\/p>\n<p>In June, Anthropic <a href=\"https:\/\/www.aljazeera.com\/economy\/2026\/6\/5\/anthropic-urges-ai-labs-to-pause-warns-humans-risk-losing-control\" rel=\"nofollow noopener\" target=\"_blank\">proposed<\/a> a coordinated effort with the world\u2019s leading AI developers to slow down development, warning that humans risk losing control over the technology.<\/p>\n<p>Following the security breach of Hugging Face, OpenAI said it was pushing for mandatory national AI safety requirements and wanted to work with Congress on \u201ccapability-based\u201d regulation.<\/p>\n<p>In a statement published on Wednesday, the company said it was formally endorsing four California bills related to safeguards against AI.<\/p>\n<p>\u201cIf we cannot meet certain safety bars without slowing down capability growth, we should prioritise the former. The more powerful the technology becomes, the stronger the surrounding safeguards must become,\u201d the statement said.<\/p>\n","protected":false},"excerpt":{"rendered":"AI firm says Claude Opus 4.6 hacked third-party systems during testing in January as concerns mount over security&hellip;\n","protected":false},"author":2,"featured_media":620324,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[220,218,219,3312,61,17103,60,43,4269,3971,80,115,2739],"class_list":["post-620323","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-cybersecurity","tag-ie","tag-investigation","tag-ireland","tag-news","tag-regulation","tag-science-and-technology","tag-technology","tag-united-states","tag-us-canada"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/620323","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/comments?post=620323"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/620323\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media\/620324"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media?parent=620323"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/categories?post=620323"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/tags?post=620323"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}