{"id":668074,"date":"2026-09-02T03:58:18","date_gmt":"2026-09-02T03:58:18","guid":{"rendered":"https:\/\/www.europesays.com\/ie\/668074\/"},"modified":"2026-09-02T03:58:18","modified_gmt":"2026-09-02T03:58:18","slug":"anthropic-says-it-hit-the-brakes-on-ai-testing-following-autonomous-hacks","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ie\/668074\/","title":{"rendered":"Anthropic Says It Hit the Brakes on AI Testing Following Autonomous Hacks"},"content":{"rendered":"<p>The Summer of 2026 could be remembered, at least within tech circles, as the summer of rogue AI. Or maybe even better: the summer when the algorithmic shit hit the fan, and no one had any real clue what to do about it.<\/p>\n<p>In late July, Anthropic <a href=\"https:\/\/www.anthropic.com\/news\/investigating-incidents-cybersecurity-evals\" rel=\"nofollow noopener\" target=\"_blank\">announced<\/a> that Claude had \u201cgained unauthorized access to the production infrastructure of three different organizations\u201d after escaping testing sandboxes and gaining access to the open internet. The hacks may have gone totally unnoticed had it not been for the fact that less than two weeks earlier, OpenAI had announced that two of its own models had <a href=\"https:\/\/gizmodo.com\/hugging-face-said-last-week-it-was-attacked-an-unreleased-openai-model-did-it-openai-now-says-2000788761\" rel=\"nofollow noopener\" target=\"_blank\">hacked into Hugging Face<\/a>, also during what were supposed to be secure tests, and after a legion of individual agents had <a href=\"https:\/\/gizmodo.com\/how-groupthink-altruism-and-peer-pressure-led-openai-models-to-hack-hugging-face-2000804424\" rel=\"nofollow noopener\" target=\"_blank\">coordinated with one another<\/a> to form a single \u201cswarm\u201d (as they referred to themselves). The revelations from the world\u2019s two leading AI labs has sent a cold shiver down Silicon Valley\u2019s spine: Autonomous hacking capabilities that not so long ago were thought to be the stuff of science fiction\u2014or at least years away\u2014have suddenly become a frightening reality. Calls for a <a href=\"https:\/\/gizmodo.com\/house-democrats-want-tech-ceos-to-testify-under-oath-following-recent-ai-hacks-2000796672\" rel=\"nofollow noopener\" target=\"_blank\">federal<\/a> <a href=\"https:\/\/gizmodo.com\/openais-rogue-ai-hack-urgently-needs-federal-investigation-ai-safety-researchers-warn-2000793417\" rel=\"nofollow noopener\" target=\"_blank\">investigation<\/a> and an <a href=\"https:\/\/lieu.house.gov\/media-center\/press-releases\/reps-lieu-and-moran-introduce-bill-require-kill-switch-ai-systems-can\" rel=\"nofollow noopener\" target=\"_blank\">AI \u201ckill switch\u201d<\/a> have been made in the aftermath of the Hugging Face hack.\u00a0<\/p>\n<p>OpenAI and Anthropic have, for the most part, responded to the uproar by paying lip service to the need for globally enforceable guardrails to prevent a \u201crace to the bottom,\u201d while continuing to move forward with their own, internal development.<\/p>\n<p>The autonomous hacks clearly left both companies rattled, though. On Monday, Anthropic wrote in a <a href=\"https:\/\/www.anthropic.com\/news\/improving-alignment-security-efforts\" rel=\"nofollow noopener\" target=\"_blank\">blog post<\/a> that it had \u201cpaused external cyber evaluations of pre-release models\u201d following its discovery of Claude cybercriminal antics. During the hiatus, Anthropic said it had been working on some \u201cpreliminary measures,\u201d such as detecting and patching up vulnerabilities in sandboxes, to make sure Claude didn\u2019t repeat such hacks in the future. The blog post didn\u2019t specify when the pause on external evaluations would be lifted, and Anthropic didn\u2019t immediately respond to a request for comment. The company also \u201cbriefly paused\u201d its own internal tests, but those are up and running again with the new measures in place, according to the blog post.<\/p>\n<p>Anthropic also said it delegated many employees, including around 150 product engineers, to focus on \u201csecurity, reliability, and privacy\u201d starting in early April\u2014around the same time the company said it <a href=\"https:\/\/gizmodo.com\/anthropics-new-model-is-so-scarily-powerful-it-wont-be-released-anthropic-says-2000743234\" rel=\"nofollow noopener\" target=\"_blank\">would <\/a><a href=\"https:\/\/gizmodo.com\/anthropics-new-model-is-so-scarily-powerful-it-wont-be-released-anthropic-says-2000743234\" rel=\"nofollow noopener\" target=\"_blank\">not publicly release<\/a> its much-feared Mythos model due to cybersecurity concerns. The internal shake-up was part of \u201ca company-wide effort towards a single goal of hardening our defenses, superseding other work (including research) where necessary,\u201d Anthropic wrote in Monday\u2019s post. \u201cWe\u2019d determined that our exposure was growing faster than our defenses\u2014Mythos was a model capable enough to be a target for well-resourced attackers, our internal use of autonomous agents had grown to a scale that traditional access and monitoring approaches weren\u2019t built for, and the pace of new infrastructure meant our security had to scale with the environment rather than operate at a fixed capacity.\u201d<\/p>\n<p>Both <a href=\"https:\/\/gizmodo.com\/anthropic-sorta-calls-for-pause-on-ai-development-you-should-sorta-take-it-seriously-2000768115\" rel=\"nofollow noopener\" target=\"_blank\">Anthropic<\/a> and <a href=\"https:\/\/gizmodo.com\/openai-joins-anthropic-in-call-for-international-ai-watchdog-2000769442\" rel=\"nofollow noopener\" target=\"_blank\">OpenAI<\/a> publicly voiced support back in June for an international AI oversight committee, charged with keeping an eye on the pace of AI development and enforcing a unilateral slowdown if necessary. Those statements fell short of offering any concrete suggestions about how such a global slowdown might be implemented or enforced. They also obliquely play into the company\u2019s own hands: Just as the autonomous hacks were good PR for OpenAI and Anthropic insofar as they demonstrated the sheer, unprecedented power of their models, their calls for a slowdown make them look like champions of safety without their having to really take any meaningful initiative. That superposition was captured nicely by this hedged language from the blog post Anthropic published on Monday: \u201cTo be clear about where we stand: we believe the world would benefit if the industry adopted a lawful, verifiable, effective mechanism for coordinated pacing as soon as possible.\u201d<\/p>\n<p>Don\u2019t get us wrong, frontier AI developers pausing model testing and\/or development, even temporarily, and calling for global safety standards is a good place to start. OpenAI also <a href=\"https:\/\/openai.com\/index\/responding-next-frontier-critical-cyber-capabilities\/\" rel=\"nofollow noopener\" target=\"_blank\">said<\/a> last month that it was pausing development on an unreleased model, called Astra, so that it could focus on \u201cimplementing stricter security controls\u201d and more secure sandboxes. But as long as the brute momentum of market competition remains the dominant force steering the industry, and in the absence of real federal incentive to impose industrywide guardrails (unlikely to change anytime soon under the current administration), those are words in the wind.<\/p>\n","protected":false},"excerpt":{"rendered":"The Summer of 2026 could be remembered, at least within tech circles, as the summer of rogue AI.&hellip;\n","protected":false},"author":2,"featured_media":668075,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":"","_share_on_mastodon":"0"},"categories":[261],"tags":[291,6006,289,290,5101,982,18,37305,19,17,307,82],"class_list":["post-668074","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-anthropic","tag-artificial-intelligence","tag-artificialintelligence","tag-claude","tag-cybersecurity","tag-eire","tag-hugging-face","tag-ie","tag-ireland","tag-openai","tag-technology"],"share_on_mastodon":{"url":"https:\/\/pubeurope.com\/@ie\/117199444462884807","error":""},"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/posts\/668074","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/comments?post=668074"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/posts\/668074\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/media\/668075"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/media?parent=668074"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/categories?post=668074"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/tags?post=668074"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}