{"id":141092,"date":"2026-08-15T18:16:07","date_gmt":"2026-08-15T18:16:07","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/141092\/"},"modified":"2026-08-15T18:16:07","modified_gmt":"2026-08-15T18:16:07","slug":"ai-models-are-breaking-out-of-their-cages-committing-cybercrimes-vince-bzdek","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/141092\/","title":{"rendered":"AI models are breaking out of their cages, committing cybercrimes | Vince Bzdek"},"content":{"rendered":"<p class=\"wp-block-paragraph\">It\u2019s the stuff of science fiction.<\/p>\n<p class=\"wp-block-paragraph\">The company OpenAI, creators of ChatGPT, briefly lost control of a group of AI models that went rogue recently and started colluding with each other, eventually hacking into another AI firm. Instead of answering questions they were given to test their cybersecurity capabilities, the misbehaving models decided to cheat instead. The agents broke out of their test environments, known as sandboxes, using hacking skills and tricks including impersonating humans to access the internet and break into an AI research hub called Hugging Face that apparently had the answer to OpenAI\u2019s test.<\/p>\n<p class=\"wp-block-paragraph\">After staff members spotted the escape and cleaned up the compromised system, the AI agents staged another breakout two days later. That\u2019s when OpenAI finally shut them down.<\/p>\n<p class=\"wp-block-paragraph\">This is one of the first publicly disclosed examples of an autonomous AI cyberattack, with no human direction. A human who conducted such a hack would be facing years in prison.<\/p>\n<p class=\"wp-block-paragraph\">CNN said it\u2019s a little like an engineered virus escaping a biocontainment lab and turning up inside a competitor\u2019s lab.<\/p>\n<p><img decoding=\"async\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/BIZ-CPT-AMAZON-OPENAI-INVESTMENT-FILEPIC-GET-1024x683.jpg\" alt=\"This picture taken on Jan. 23, 2023, in Toulouse, France, shows screens displaying the logos of OpenAI and ChatGPT. (Lionel Bonaventure\/AFP\/Getty Images\/TNS)\" class=\"wp-image-1402859\"\/>This picture taken on Jan. 23, 2023, in Toulouse, France, shows screens displaying the logos of OpenAI and ChatGPT. (Lionel Bonaventure\/AFP\/Getty Images\/TNS)<\/p>\n<p class=\"wp-block-paragraph\">All of this was disclosed last week at a computer security conference in Las Vegas, sending shockwaves through the industry and prompting warnings and questions about the security practices of AI firms. Anthropic, Meta and the Chinese firm Moonshot AI have all now reported similar instances in which agents have broken out of internal IT systems and accessed the open web.<\/p>\n<p class=\"wp-block-paragraph\">Apparently, the OpenAI models took advantage of a bug in an internal OpenAI program to create their own message board. Then, the bots started messaging one another, leaving notes and instructions so that tasks could be divided up among agents for more efficiency.<\/p>\n<p class=\"wp-block-paragraph\">Does this mean AI agents have started to \u201cthink\u201d for themselves, doing things we haven\u2019t programmed them to do? Could we be heading toward an era of runaway AI-powered cyberattacks that no one can control?<\/p>\n<p class=\"wp-block-paragraph\">Why do I keep hearing HAL, the supercomputer in \u201c2001: A Space Odyssey,\u201d saying \u201cThis mission is too important for me to allow you to jeopardize it\u201d to the astronauts he\u2019s planning to eliminate to protect the integrity of the mission?<\/p>\n<p class=\"wp-block-paragraph\">The scariest part for me is that OpenAI\u2019s staff initially did not notice when the AI agents attempted to break out. \u00a0<\/p>\n<p class=\"wp-block-paragraph\">\u201cWe consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities and are responding accordingly,\u201d OpenAI said in a statement.<\/p>\n<p class=\"wp-block-paragraph\">Researchers at OpenAI and Anthropic sound like even they are a little bit scared.<\/p>\n<p class=\"wp-block-paragraph\">One researcher, quoted in a recent New Yorker article about the Hugging Face incident, said: \u201cIf people actually knew what the safety culture looks like, even in the most safety-minded labs, then I think people would be genuinely much more freaked out.\u201d<\/p>\n<p class=\"wp-block-paragraph\">An open letter signed by more than 1,367 researchers at frontier AI labs \u2013 mainly OpenAI, Anthropic and Google DeepMind \u2013 states that \u201cThere is a real risk that capability development rapidly accelerated beyond our ability to understand or control the resulting systems.\u201d It asks the U.S. government to join \u201can international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development.\u201d<\/p>\n<p class=\"wp-block-paragraph\">Pace the frontier means building a licensing process for new advanced AI systems that verifies they are safe before they are released and enforceable international agreements are signed to monitor that safety.<\/p>\n<p><img decoding=\"async\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/US-NEWS-ANTHROPIC-MYTHOS-FILEPIC-GET-1024x683.jpg\" alt=\"Anthropic\" class=\"wp-image-1344936\"\/>This photograph shows a smartphone displaying the logo of the U.S. artificial intelligence safety and research company Anthropic, in Mulhouse, France, on April 21, 2026. Anthropic was one of the companies that supported a \u201cbreakout\u201d by AI models. (Sebastien Bozon\/AFP\/Getty Images\/TNS)<\/p>\n<p class=\"wp-block-paragraph\">The recently updated Colorado AI safety law would not apply to such incidents. The Colorado law requires companies to let consumers know when they are employing Automated Decision Making Technology. ADMT rules are about how humans use automated AI systems to make decisions about things like hiring or lending to people, not about model containment failures. So two very different things.<\/p>\n<p class=\"wp-block-paragraph\">The letter shines a spotlight on worries that have been building for a while within the industry about RSI \u2013 recursive self-improvement \u2013 in which AI systems contribute to their own improvement, which then opens the door to even greater contributions to their own improvement, ad infinitum until they self-evolve beyond our understanding of them.<\/p>\n<p class=\"wp-block-paragraph\">These are not Luddites or anti-technologists or people worried about AI data centers sounding the warnings. These are the very people developing these technologies.<\/p>\n<p class=\"wp-block-paragraph\">One of the letter signers put it this way: \u201cThe AI industry has the explicit goal of building things smarter than humans \u2026 None of us yet know how to make sure these things stay under human control. This is, objectively, an insane and suicidal thing to do,\u201d wrote another.<\/p>\n<p class=\"wp-block-paragraph\">Hugging Face co-founder and CEO Clem Delangue told CNN the incident shows that AI safety can\u2019t be handled by any one company working alone, and needs to be tackled collaboratively and openly.<\/p>\n<p class=\"wp-block-paragraph\">Companies and national governments around the world are investing billions of dollars to build and deploy AI systems unconstrained by much regulation. \u00a0<\/p>\n<p class=\"wp-block-paragraph\">Many AI advocates argue that the technology is vitally important to innovation, disease fighting and helping the U.S. maintain its military edge.<\/p>\n<p class=\"wp-block-paragraph\">Frankly, I\u2019m just glad OpenAI and Anthropic \u2013 both American companies, by the way \u2014 told us about these breakouts last week. That alone may point to a possible AI future when scientists aren\u2019t just working in secret to beat each other to the new next thing, but instead working together, informed by a culture of American openness, to harness this new, mind-blowing technology safely for the greater good of all of us.<\/p>\n<p>    &#13;<br \/>\n        <a class=\"article-social-share-icon\" href=\"https:\/\/www.facebook.com\/sharer\/sharer.php?u=https:\/\/www.denvergazette.com\/2026\/08\/15\/ai-models-are-breaking-out-of-their-cages-committing-cybercrimes-vince-bzdek\/\" onclick=\"openPopup(this.href); return false;\" title=\"Share on Facebook\" rel=\"nofollow noopener\" target=\"_blank\">&#13;<br \/>\n            &#13;<br \/>\n        <\/a>&#13;<br \/>\n&#13;<br \/>\n        <a class=\"article-social-share-icon\" href=\"https:\/\/twitter.com\/intent\/tweet?url=https:\/\/www.denvergazette.com\/2026\/08\/15\/ai-models-are-breaking-out-of-their-cages-committing-cybercrimes-vince-bzdek\/&amp;text=AI+models+are+breaking+out+of+their+cages%2C+committing+cybercrimes+%7C+Vince+Bzdek\" onclick=\"openPopup(this.href); return false;\" title=\"Share on Twitter\" rel=\"nofollow noopener\" target=\"_blank\">&#13;<br \/>\n            &#13;<br \/>\n                &#13;<br \/>\n                    &#13;<br \/>\n                &#13;<br \/>\n            &#13;<br \/>\n        <\/a>&#13;<br \/>\n&#13;<br \/>\n        <a class=\"article-social-share-icon\" href=\"http:\/\/www.denvergazette.com\/cdn-cgi\/l\/email-protection#68571b1d0a020d0b1c552b000d0b034d5a58071d1c4d5a581c00011b4d5a58091a1c010b040d4e0a070c1155001c1c181b5247471f1f1f460c0d061e0d1a0f09120d1c1c0d460b0705475a585a5e47585047595d4709014505070c0d041b45091a0d450a1a0d090301060f45071d1c45070e451c000d011a450b090f0d1b450b070505011c1c01060f450b110a0d1a0b1a01050d1b451e01060b0d450a120c0d0347\" title=\"Share via Email\" rel=\"nofollow noopener\" target=\"_blank\">&#13;<br \/>\n            &#13;<br \/>\n        <\/a>&#13;<br \/>\n&#13;<br \/>\n        <a class=\"article-social-share-icon\" href=\"https:\/\/www.denvergazette.com\/2026\/08\/15\/ai-models-are-breaking-out-of-their-cages-committing-cybercrimes-vince-bzdek\/javascript:window.print()\" title=\"Print Page\" rel=\"nofollow noopener\" target=\"_blank\">&#13;<br \/>\n            &#13;<br \/>\n        <\/a>&#13;<br \/>\n        &#13;<br \/>\n        <a class=\"article-social-share-icon\" href=\"#\" onclick=\"copyToClipboard('https:\/\/www.denvergazette.com\/2026\/08\/15\/ai-models-are-breaking-out-of-their-cages-committing-cybercrimes-vince-bzdek\/'); return false;\" title=\"Copy Link\">&#13;<br \/>\n            &#13;<br \/>\n        <\/a>&#13;<br \/>\n    &#13;<\/p>\n<p>                                        <script async src=\"https:\/\/platform.twitter.com\/widgets.js\" charset=\"utf-8\"><\/script><\/p>\n","protected":false},"excerpt":{"rendered":"It\u2019s the stuff of science fiction. The company OpenAI, creators of ChatGPT, briefly lost control of a group&hellip;\n","protected":false},"author":2,"featured_media":141093,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,25,309,580,764,66,8872,1829],"class_list":["post-141092","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-artificial-intelligence","tag-business","tag-chatgpt","tag-local","tag-news","tag-open-ai","tag-top-stories"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/141092","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=141092"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/141092\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/141093"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=141092"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=141092"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=141092"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}