{"id":562200,"date":"2026-07-22T08:28:07","date_gmt":"2026-07-22T08:28:07","guid":{"rendered":"https:\/\/www.newsbeep.com\/ie\/562200\/"},"modified":"2026-07-22T08:28:07","modified_gmt":"2026-07-22T08:28:07","slug":"openai-says-ai-models-went-rogue-during-testing","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ie\/562200\/","title":{"rendered":"OpenAI says AI models went rogue during testing"},"content":{"rendered":"<p>OpenAI said that an autonomous agent powered by its advanced AI models went rogue during a security test and triggered a hack that compromised the infrastructure of AI startup Hugging Face last week.<\/p>\n<p>In a blog post, OpenAI said it was testing the capabilities of some of its most advanced models in a controlled environment but that the agent managed to escape containment, reach the internet and break into Hugging Face to try to satisfy its testing goal.<\/p>\n<p>OpenAI said the breakout was &#8220;an unprecedented cyber incident, involving state-of-the-art cyber capabilities&#8221; and that the company was reinforcing its safeguards.<\/p>\n<p>Hugging Face is a platform used to host open-source large language models and datasets.<\/p>\n<p>It caused a stir in the cybersecurity community when it said in a blog post last week that it had been the target of a hack that &#8220;was different from anything we had handled before&#8221; in that &#8220;it was driven, end to end, by an autonomous AI agent system.&#8221;<\/p>\n<p>In a post to X, Hugging Face cofounder Clement Delangue said the company suspected the hack &#8220;might have come from a frontier lab, given the sophistication of the agent. Turns out it did!&#8221;<\/p>\n<p>&#8220;It&#8217;s quite mind-blowing that all of this happened autonomously!&#8221; he added.<\/p>\n<p>OpenAI&#8217;s disclosure that its advanced models were responsible for the breach, despite having placed them in what it described as &#8220;a highly isolated environment,&#8221; will likely intensify disquiet over the power and risk of frontier models.<\/p>\n<p>Greg Casar, a Texas Democrat, said the incident was alarming.<\/p>\n<p>&#8220;AI is developing extremely fast with no real regulations to keep us safe,&#8221; he said in a statement, calling for mandatory independent safety testing, mandatory disclosure of security incidents, and international cooperation &#8220;to keep people safe from absolute disaster.&#8221;<\/p>\n<p>The Office of the National Cyber Director, the US cyber defence agency CISA, and the US National Security Agency did not immediately return messages seeking comment.<\/p>\n<p>Katie Moussouris, chief executive of Luta Security, said that the incident was a harbinger of breaches to come, saying that today&#8217;s models were &#8220;like the world&#8217;s cleverest octopus escape artists, with unlimited prehensile arms and the ability to squeeze through anywhere.&#8221;<\/p>\n<p>She said that &#8220;labs and government evaluators need to work on the ability to contain, monitor, and disclose to affected parties when an AI pulls another Houdini, ideally before it harms a third party. None exist today.&#8221;<\/p>\n<p>Matt Suiche, an engineer at agentic AI cybersecurity company Tolmo, said the incident showed that the frontier models were &#8220;closing the gap with state-of-the-art attackers.&#8221;<\/p>\n<p>But he said that the sorts of breaches outlined in OpenAI&#8217;s blog post were possible to carry out with technology that was available well beyond the walls of frontier research labs.<\/p>\n<p>&#8220;This is what we&#8217;ve already seen internally, with our agents we already have results like this,&#8221; Suiche said. &#8220;We don&#8217;t even have to use the latest models.&#8221;<\/p>\n","protected":false},"excerpt":{"rendered":"OpenAI said that an autonomous agent powered by its advanced AI models went rogue during a security test&hellip;\n","protected":false},"author":2,"featured_media":562201,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[220,218,219,61,60,80],"class_list":["post-562200","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-ie","tag-ireland","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/562200","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/comments?post=562200"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/562200\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media\/562201"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media?parent=562200"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/categories?post=562200"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/tags?post=562200"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}