{"id":729328,"date":"2026-06-11T07:44:08","date_gmt":"2026-06-11T07:44:08","guid":{"rendered":"https:\/\/www.newsbeep.com\/ca\/729328\/"},"modified":"2026-06-11T07:44:08","modified_gmt":"2026-06-11T07:44:08","slug":"anthropic-walks-back-policy-that-could-have-sabotaged-ai-researchers-using-claude","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ca\/729328\/","title":{"rendered":"Anthropic Walks Back Policy That Could Have \u2018Sabotaged\u2019 AI Researchers Using Claude"},"content":{"rendered":"<p>Anthropic is backtracking on a policy that would have covertly limited competitors from using its new AI model, <a href=\"https:\/\/www.wired.com\/story\/anthropic-releases-claude-fable-5-mythos-5\/\" target=\"_blank\" class=\"text link\" rel=\"nofollow noopener\">Claude Fable 5<\/a>, to develop other AI models. The company changed course after the move received significant backlash from the <a href=\"https:\/\/www.wired.com\/tag\/artificial-intelligence\/\" class=\"text link\" rel=\"nofollow noopener\" target=\"_blank\">AI research community<\/a>.<\/p>\n<p class=\"paywall\">\u201cWe\u2019re changing Fable 5\u2019s safeguards for frontier LLM development to make them visible,\u201d Anthropic said in a statement to WIRED. \u201cWe made the wrong trade-off and we apologize for not getting the balance right.\u201d<\/p>\n<p class=\"paywall\">Anthropic released Claude Fable 5, a version of its latest AI model with additional safety guardrails designed to prevent misuse, earlier this week. Some of the safeguards Anthropic decided on were unsurprising: The company said it would reroute users who asked questions about cybersecurity, biology, or chemistry to a less capable AI model to reduce the chances of someone using the advanced AI to carry out a cyberattack or build a bioweapon.<\/p>\n<p class=\"paywall\">But for researchers trying to use Claude Fable 5 for frontier AI development, Anthropic outlined a different approach. The firm would deliberately degrade the model\u2019s performance in ways that were invisible to the user. The move would effectively sabotage researchers trying to use Claude to train competing AI models, which Anthropic explicitly bans in its <a data-offer-url=\"https:\/\/www.anthropic.com\/legal\/consumer-terms\" class=\"external-link text link\" data-event-click=\"{&quot;element&quot;:&quot;ExternalLink&quot;,&quot;outgoingURL&quot;:&quot;https:\/\/www.anthropic.com\/legal\/consumer-terms&quot;}\" href=\"https:\/\/www.anthropic.com\/legal\/consumer-terms\" rel=\"nofollow noopener\" target=\"_blank\">terms of service<\/a>.<\/p>\n<p class=\"paywall\">Anthropic now says it\u2019s changing course, and that Claude Fable 5\u2019s safeguards for AI development will be visible to users. If the company suspects a user is trying to use Claude to build a highly capable AI, it will alert them that it\u2019s either refusing the request or rerouting the user to a less capable model.<\/p>\n<p class=\"paywall\">Anthropic reversed the policy after it received fierce backlash from the AI research community. Anthropic has already taken steps to <a href=\"https:\/\/www.wired.com\/story\/anthropic-revokes-openais-access-to-claude\/\" class=\"text link\" rel=\"nofollow noopener\" target=\"_blank\">limit competitors from using Claude<\/a> to build closed- and open-source AI models, but critics say that quietly degrading the model\u2019s performance for certain users went a step too far. Claude\u2019s coding agent has become a favored tool among developers, including those working on open-source AI research projects, and researchers tell WIRED that the company\u2019s latest policy could have led to a troubling future in which only a handful of leading AI labs could perform advanced AI research.<\/p>\n<p class=\"paywall\">Dean Ball, a senior fellow at the Foundation for American Innovation and a former adviser to the White House on AI, wrote in a <a data-offer-url=\"https:\/\/x.com\/deanwball\/status\/2064434861088395730?s=20\" class=\"external-link text link\" data-event-click=\"{&quot;element&quot;:&quot;ExternalLink&quot;,&quot;outgoingURL&quot;:&quot;https:\/\/x.com\/deanwball\/status\/2064434861088395730?s=20&quot;}\" href=\"https:\/\/x.com\/deanwball\/status\/2064434861088395730?s=20\" rel=\"nofollow noopener\" target=\"_blank\">post<\/a> on X that \u201cdegrading performance on ML research *without telling the user* is shockingly hostile and a terrible look.\u201d He continued in another <a data-offer-url=\"https:\/\/x.com\/deanwball\/status\/2064665679307985244?s=20\" class=\"external-link text link\" data-event-click=\"{&quot;element&quot;:&quot;ExternalLink&quot;,&quot;outgoingURL&quot;:&quot;https:\/\/x.com\/deanwball\/status\/2064665679307985244?s=20&quot;}\" href=\"https:\/\/x.com\/deanwball\/status\/2064665679307985244?s=20\" rel=\"nofollow noopener\" target=\"_blank\">post<\/a> that the \u201csecret sabotage\u201d policy undermines Anthropic\u2019s overall stance, because it limits AI researchers from collaborating on AI safety.<\/p>\n<p class=\"paywall\">\u201cIt felt like Anthropic was saying to the public, \u2018We don&#8217;t trust anybody else to do AI research. We are the only ones who have to do AI research,\u201d says Will Brown, research lead at the open-source AI startup Prime Intellect. \u201cIt feels a bit like they\u2019re starting to pull the ladder up behind them.\u201d<\/p>\n<p class=\"paywall\">Brown said the policy would also have left developers in the dark about whether they were violating Anthropic\u2019s rules, since the company wouldn\u2019t alert them when its safeguards were triggered. He added that the restrictions could have had widespread consequences. For example, he pointed to the growing ecosystem of third-party evaluation firms that test frontier models for safety, performance, and reliability\u2014work that could have been hindered if Anthropic secretly degraded its model.<\/p>\n","protected":false},"excerpt":{"rendered":"Anthropic is backtracking on a policy that would have covertly limited competitors from using its new AI model,&hellip;\n","protected":false},"author":2,"featured_media":729329,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[62,1930,276,277,49,48,5272,10413,2927,61],"class_list":["post-729328","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-anthropic","tag-artificial-intelligence","tag-artificialintelligence","tag-ca","tag-canada","tag-claude","tag-generative-ai","tag-startups","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts\/729328","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/comments?post=729328"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts\/729328\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/media\/729329"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/media?parent=729328"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/categories?post=729328"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/tags?post=729328"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}