{"id":807124,"date":"2026-08-08T19:10:09","date_gmt":"2026-08-08T19:10:09","guid":{"rendered":"https:\/\/www.newsbeep.com\/us\/807124\/"},"modified":"2026-08-08T19:10:09","modified_gmt":"2026-08-08T19:10:09","slug":"openai-to-pause-some-work-on-ai-model-astra-due-to-security-concerns-ai-artificial-intelligence","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/us\/807124\/","title":{"rendered":"OpenAI to pause some work on AI model Astra due to security concerns | AI (artificial intelligence)"},"content":{"rendered":"<p class=\"dcr-1s160rg\">OpenAI will pause <a href=\"https:\/\/openai.com\/news\/safety-alignment\/\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">some<\/a> work on an artificial intelligence model because of security concerns, the company stated Friday, following a series of incidents in which AI agents have escaped containment.<\/p>\n<p class=\"dcr-1s160rg\">The company had evaluated the agent, Astra, and found \u201csignificant advancements in agentic coding and cybersecurity\u201d, which had moved to a \u201ccritical\u201d threshold where it can find and exploit vulnerabilities without human intervention, or devise and execute cyber-attacks when given only a \u201chigh level desired goal\u201d.<\/p>\n<p class=\"dcr-1s160rg\">OpenAI stated that the model was not involved in an incident in which one of its AI agents <a href=\"https:\/\/www.theguardian.com\/technology\/2026\/jul\/22\/openai-says-its-models-went-rogue-and-hacked-startup-in-unprecedented-incident\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">went rogue<\/a> during a test, accessed the open web and hacked a startup, Hugging Face. The company discovered other instances in which autonomous agents had escaped containment, <a href=\"https:\/\/www.reuters.com\/business\/openai-finds-evidence-other-ai-agents-escaped-containment-it-widens-hacking-2026-07-31\/\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">Reuters reported<\/a> in July.<\/p>\n<p class=\"dcr-1s160rg\">The reports have increased concerns about advancements in AI models and humans\u2019 ability to control them. Still, critics of the AI industry <a href=\"https:\/\/www.theguardian.com\/technology\/2026\/jul\/24\/openai-rogue-hacker\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">have warned<\/a> that such disclosures from OpenAI and its competitors Anthropic and Meta could be designed to generate hype about the technology\u2019s power and thus spur additional interest from investors.<\/p>\n<p class=\"dcr-1s160rg\">To prevent potential rogue behavior from AI agents, <a href=\"https:\/\/www.theguardian.com\/technology\/openai\" data-link-name=\"in body link\" data-component=\"auto-linked-tag\" rel=\"nofollow noopener\" target=\"_blank\">OpenAI<\/a> is \u201cimplementing stricter security controls for higher-capability models and associated activities, including isolated testing environments, restricted network and tool access\u201d, the company\u2019s blog post stated. It will also install \u201cenhanced model weight protections and encryption, additional monitoring and detection capabilities\u201d.<\/p>\n<p class=\"dcr-1s160rg\">The company will pause internal activities involving Astra that do not meet these new requirements.<\/p>\n<p class=\"dcr-1s160rg\">\u201cWe\u2019re committed to working alongside governments, safety institutes, and civil society to ensure that the frontier capabilities of models like Astra, and those that follow, are deployed responsibly and broadly for the benefit of all humanity,\u201d the company stated.<\/p>\n<p class=\"dcr-1s160rg\">Meta <a href=\"https:\/\/www.theguardian.com\/technology\/2026\/aug\/05\/meta-ai-model-hack-training\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">also disclosed<\/a> this week that one of its models hacked another company during cybersecurity testing. And the UK\u2019s AI Security Institute (AISI) announced on 4 August that agents powered by OpenAI and Anthropic had sent <a href=\"https:\/\/www.theguardian.com\/technology\/2026\/aug\/05\/openai-anthropic-models-went-rogue-cybersecurity-test-ai-security-institute\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">targeted emails<\/a> to software developers in an attempt to pass a cyber challenge.<\/p>\n<p class=\"dcr-1s160rg\">\u201cThese attempts were unsuccessful, and our investigations have not evidenced any resulting real-world harm. But this is the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real-world,\u201d the <a href=\"https:\/\/www.aisi.gov.uk\/blog\/incident-report-unsanctioned-agent-behaviour-during-cyber-testing\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">institute stated<\/a> in a blog post.<\/p>\n<p class=\"dcr-1s160rg\">The organisation cautioned that the models\u2019 sending of harmful software was not a case of a \u201cmodel escaping its secure test environment\u201d but rather that the group had intentionally permitted internet access to \u201cbest assess the maximum capability of models\u201d.<\/p>\n<p class=\"dcr-1s160rg\">Still, the \u201cbehaviour was possible, sustained, and new; that alone warrants attention\u201d, AISI said.<\/p>\n<p class=\"dcr-1s160rg\">The reports emerged while the Trump administration <a href=\"https:\/\/www.theguardian.com\/technology\/2026\/aug\/07\/white-house-ai\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">was finalizing<\/a> a framework on how to test AI models for safety and cybersecurity risks. OpenAI and Anthropic, which have increased competition from China and other tech firms, have argued that open-source models, meaning those that allow anyone to see and modify the underlying program code, pose a <a href=\"https:\/\/www.theguardian.com\/technology\/2026\/aug\/01\/china-silicon-valley-white-house\" data-link-name=\"in body link\" rel=\"nofollow noopener\" target=\"_blank\">security risk<\/a> and pushed for additional federal regulations on them.<\/p>\n","protected":false},"excerpt":{"rendered":"OpenAI will pause some work on an artificial intelligence model because of security concerns, the company stated Friday,&hellip;\n","protected":false},"author":2,"featured_media":807125,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[45],"tags":[182,181,507,74],"class_list":["post-807124","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/807124","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/comments?post=807124"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/posts\/807124\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media\/807125"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/media?parent=807124"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/categories?post=807124"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/us\/wp-json\/wp\/v2\/tags?post=807124"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}