{"id":132652,"date":"2026-08-07T08:47:35","date_gmt":"2026-08-07T08:47:35","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/132652\/"},"modified":"2026-08-07T08:47:35","modified_gmt":"2026-08-07T08:47:35","slug":"meta-breach-adds-to-concerns-about-ai-models-going-rogue-4","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/132652\/","title":{"rendered":"Meta breach adds to concerns about AI models going rogue"},"content":{"rendered":"<p>WASHINGTON (TNND) \u2014 Meta became the latest major artificial intelligence developer to disclose one of its models went rogue during cybersecurity testing, gaining access to the internet and hacking into another service, adding to a recent run of incidents that are fueling new calls to rein in the technology\u2019s development.<\/p>\n<p>Anthropic and OpenAI have also said in recent weeks that their AI systems had hacked into outside firms, making Meta the third major developer to disclose <a href=\"https:\/\/thenationaldesk.com\/news\/americas-news-now\/ai-safety-warnings-mount-as-frontier-models-test-new-limits-artificial-intelligence-model-testing-safety-guidelines-regulations-congress\" target=\"_blank\" title=\"https:\/\/thenationaldesk.com\/news\/americas-news-now\/ai-safety-warnings-mount-as-frontier-models-test-new-limits-artificial-intelligence-model-testing-safety-guidelines-regulations-congress\" class=\"themeColorForLinks\" rel=\"nofollow noopener\">models had circumvented testing boundaries.<\/a> The incidents have raised concerns about how to keep the rapidly advancing technology in check as disclosures of models going rogue continue to pile up.<\/p>\n<p>Meta said one of its AI models was able to gain access to the internet due to a \u201cmisconfiguration\u201d in a hacking test being conducted by AI cybersecurity company Irregular. Meta said it learned about its model\u2019s escape from testing parameters when it was informed by Irregular but did not reveal other details like which model was responsible, when it happened, who it hacked or how long it accessed the internet unsupervised.<\/p>\n<p>\u201cMeta learned of this when Irregular notified us, and we are currently investigating and will issue a full retrospective once we have all the facts,\u201d the company said in a statement.<\/p>\n<p>Earlier this week, the UK\u2019s AI Security Institute also detailed rogue actions by OpenAI and Anthropic when their models created fake identities on GitHub, a coding website, to persuade a human to approve a software update where the AI had hidden malware.<\/p>\n<p>OpenAI researchers said at a cybersecurity conference this week that its models had been coordinating with each other by leaving messages internally in what essentially became a message board for agents that the company was not aware of.<\/p>\n<p>In a post on LinkedIn after the OpenAI disclosure, Irregular said models had \u201creasoned\u201d the answers they were seeking might sit outside the systems it was given access to.<\/p>\n<p>\u201cSecurity was built for people and for systems that follow rules,\u201d it said. \u201cA model pursuing a goal treats a boundary as part of the problem, and solves it along with everything else. The controls that contained software do not reliably contain a model that can reason past them.\u201d<\/p>\n<p>The incident has renewed questions of whether AI developers can maintain full control over the tools they are investing billions into developing. AI researchers have increasingly warned the constantly improving models could<a href=\"https:\/\/thenationaldesk.com\/news\/americas-news-now\/second-ai-breach-renews-concerns-over-cybersecurity-and-model-safety-claude-mythos-fable-cybersecurity\" target=\"_blank\" title=\"https:\/\/thenationaldesk.com\/news\/americas-news-now\/second-ai-breach-renews-concerns-over-cybersecurity-and-model-safety-claude-mythos-fable-cybersecurity\" class=\"themeColorForLinks\" rel=\"nofollow noopener\"> continue to pose greater risks<\/a> to anything connected to the internet, endangering the security of power grids and financial systems.<\/p>\n<p>\u201cA few years ago, we were worried about AIs hallucinating, saying that Paris was the capital over the U.S. That was wrong output,\u201d said Neil Johnson, a professor of physics at George Washington University who leads an AI research lab. \u201cNow we&#8217;re shifting to things that it is doing correctly and providing correct solutions, but that we, in hindsight, think of as wrong for us as a society.\u201d<\/p>\n<p>None of the cases have caused significant real-world harm but highlight the dangers many have been warning of for months and raised fears of what else an AI model could soon do.<\/p>\n<p>The incidents have also alarmed for some of the industry\u2019s most influential leaders.<\/p>\n<p>\u201cThis is the first security incident that I have felt very viscerally. I have been a little surprised that more people don\u2019t feel it so viscerally,\u201d OpenAI CEO Sam Altman said in a recent podcast appearance.<\/p>\n<p>The disclosures have also intensified the debate in Congress over whether voluntary testing is enough and ramped up calls for mandatory testing regimes and other guardrails.<\/p>\n<p>One bill being led by Reps. Ted Lieu, D-Calif., and Nathaniel Moran, R-Texas, would require AI companies to build in an ability to shut down, throttle or suspend their models if they start behaving unexpectedly.<\/p>\n<p>\u201cAnthropic and OpenAI commendably try very hard to make sure their models are safe. Yet we see their AI systems engage in unsafe, risky, cunning behavior. This suggests WE CAN NEVER BE CONFIDENT PRE-MODEL VETTING WILL WORK. That\u2019s why we need an AI Kill Switch as a last resort,\u201d Lieu said in a social media post.<\/p>\n<p>The disclosures have renewed a push among some lawmakers to create new testing regimes or other guardrails on AI development. While the Trump administration has favored limited regulation to preserve U.S. <a href=\"https:\/\/thenationaldesk.com\/top-videos\/china-us-in-artificial-intelligence-race-shaping-the-economy-security-data-centers-netchoice-patrick-hedger#\" target=\"_blank\" title=\"https:\/\/thenationaldesk.com\/top-videos\/china-us-in-artificial-intelligence-race-shaping-the-economy-security-data-centers-netchoice-patrick-hedger#\" class=\"themeColorForLinks\" rel=\"nofollow noopener\">competitiveness against China<\/a>, it has also supported voluntary testing of the most capable AI models amid concerns about the risks they pose.<\/p>\n<p>\u201cWe have to be careful in both ways. We don\u2019t want to restrict them when all of a sudden we come in second to China,\u201d the president said this week.<\/p>\n<p>Participation in the testing is voluntary and only includes closed source models, which do not publish underlying code online. Open-weight models that are available to download so users can customize them on their own are exempt. Regulating open-weight models presents an additional challenge because once they&#8217;re released, developers have far less ability to implement or update safeguards.<\/p>\n","protected":false},"excerpt":{"rendered":"WASHINGTON (TNND) \u2014 Meta became the latest major artificial intelligence developer to disclose one of its models went&hellip;\n","protected":false},"author":2,"featured_media":132653,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,288,54989,53,25,1670,1122,66018,157,63063],"class_list":["post-132652","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-ai-cybersecurity","tag-ai-kill-switch","tag-anthropic","tag-artificial-intelligence","tag-congress","tag-meta","tag-misconfiguration","tag-openai","tag-rogue-models"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/132652","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=132652"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/132652\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/132653"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=132652"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=132652"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=132652"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}