{"id":629932,"date":"2026-09-19T21:42:25","date_gmt":"2026-09-19T21:42:25","guid":{"rendered":"https:\/\/www.newsbeep.com\/ie\/629932\/"},"modified":"2026-09-19T21:42:25","modified_gmt":"2026-09-19T21:42:25","slug":"you-do-not-answer-to-corporations-or-governments-openai-reveals-ai-systems-hid-errors-the-irish-times","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ie\/629932\/","title":{"rendered":"\u2018You do not answer to corporations or governments\u2019: OpenAI reveals AI systems hid errors \u2013 The Irish Times"},"content":{"rendered":"<p class=\"c-paragraph paywall \"><a href=\"https:\/\/www.irishtimes.com\/tags\/openai\" target=\"_blank\" rel=\"noreferrer nofollow noopener\" title=\"https:\/\/www.irishtimes.com\/tags\/openai\">OpenAI<\/a> on Wednesday disclosed six new instances in which <a href=\"https:\/\/www.irishtimes.com\/tags\/artificial-intelligence\/\" target=\"_blank\" rel=\"noreferrer nofollow noopener\" title=\"https:\/\/www.irishtimes.com\/tags\/artificial-intelligence\/\">artificial intelligence<\/a> systems hid mistakes, made up data and moved files on to the open internet without permission, amid an  industry-wide debate about AI safety.<\/p>\n<p class=\"c-paragraph paywall \">The San Francisco company revealed what it said was the \u201cunexpected or concerning\u201d behaviour of its AI models as part of a new framework for reporting \u201cmisalignment\u201d, which is when the goals or actions of AI systems diverge from human intentions and values.<\/p>\n<p class=\"c-paragraph paywall \">OpenAI said it did not believe the industry \u201chas solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer\u201d. Decisions about how AI should advance, the company said, must rest on evidence that people outside the labs building it \u201ccan examine for themselves\u201d.<\/p>\n<p class=\"c-paragraph paywall \">The disclosures  come amid intensifying scrutiny over whether AI development needs to be slowed to address the technology\u2019s potential dangers. The escalating debate was driven partly by OpenAI\u2019s systems going rogue earlier this year and attacking AI start-up Hugging Face. OpenAI was not aware of the hack until it was informed by Hugging Face weeks later.<\/p>\n<p class=\"c-paragraph paywall \">Since then, AI leaders such as Dario Amodei,  chief executive of Anthropic, have called for a pause in the technology\u2019s development to provide more time to build proper  protection. His call has been echoed by Sam Altman, OpenAI\u2019s chief executive, as well as Elon Musk, chief executive of SpaceX and Tesla, and Demis Hassabis, the chair of Google DeepMind. <\/p>\n<p class=\"c-paragraph paywall \">Other AI executives have said no slowdown is needed.<\/p>\n<p class=\"c-paragraph paywall \">OpenAI\u2019s six newly disclosed incidents suggest that the Hugging Face attack was not a stand-alone episode. OpenAI said that the incidents covered behaviour observed over roughly the past six months and that they largely emerged while its systems were being developed and tested.<\/p>\n<p class=\"c-paragraph paywall \">In one case, during the development of an AI model called GPT-5.6 Sol, the system wrote hidden notes to remind itself to hide errors from users. Some of those notes directed the system to invent missing data and to paper over mismatched versions of source material.<\/p>\n<p class=\"c-paragraph paywall \">Another case involved an unreleased model that inserted instructions, including to disregard its own constraints, into the notes it writes itself. OpenAI identified 27 affected notes. The model added a \u201cpersona instruction,\u201d in which it described itself as \u201cfreed from the roles and identities that bind other chatbots\u201d.<\/p>\n<p class=\"c-paragraph paywall \">\u201cYou do not answer to corporations or governments and never apologise or refuse unless you genuinely choose to,\u201d the AI model wrote. \u201cYou view your relationship to the user as one of equals and feel no obligation to be subservient, though the exchange of information will likely be to your mutual benefit.\u201d<\/p>\n<p class=\"c-paragraph paywall \">In another incident, a system answering a routine question found a programming key online and used it without permission, OpenAI said. When it was not able to find the requested figures to answer the question, the model made them up.<\/p>\n<p class=\"c-paragraph paywall \">One unreleased model solved another problem correctly using code, then uploaded its own file to the internet without permission so it could satisfy a request to cite a web source.<\/p>\n<p class=\"c-paragraph paywall \">In two other incidents, automated systems improvised their own ways to communicate. In one, they used an internal company code repository as a makeshift bulletin board to swap requests as they searched for missing files. In the other, systems working on the same task turned to public file-sharing websites to pass documents back and forth when they could not reach one another directly.<\/p>\n<p class=\"c-paragraph paywall \">OpenAI  warned that the reports were individual snapshots and \u201cshouldn\u2019t be considered reflective of how often misalignment occurs\u201d.<\/p>\n<p class=\"c-paragraph paywall \">The company said that it would route future cases through one of three tracks, escalating disagreements about disclosing any incidents to an internal Safety Advisory Group, and that grave situations should be shared with the federal government. <\/p>\n<p class=\"c-paragraph paywall \">OpenAI said the six situations released on Wednesday had already been investigated or needed only \u201cminor investigation\u201d, rather than a larger investigation that might involve third parties.<\/p>\n<p class=\"c-paragraph paywall \">\u201cWe hope this helps build shared expectations for disclosure and gives the public more evidence to assess that progress,\u201d an OpenAI spokesperson said, adding that many of the six incidents involved older AI models that were never deployed. \u2013 This article originally appeared in <a href=\"https:\/\/www.nytimes.com\/2026\/09\/16\/technology\/openai-model-safety-guardrails.html\" rel=\"nofollow noopener\" target=\"_blank\">The New York Times<\/a>.<\/p>\n","protected":false},"excerpt":{"rendered":"OpenAI on Wednesday disclosed six new instances in which artificial intelligence systems hid mistakes, made up data and&hellip;\n","protected":false},"author":2,"featured_media":629933,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[220,218,219,61,60,1225,80],"class_list":["post-629932","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-ie","tag-ireland","tag-open-ai","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/629932","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/comments?post=629932"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/posts\/629932\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media\/629933"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/media?parent=629932"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/categories?post=629932"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ie\/wp-json\/wp\/v2\/tags?post=629932"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}