{"id":130436,"date":"2026-08-05T16:12:12","date_gmt":"2026-08-05T16:12:12","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/130436\/"},"modified":"2026-08-05T16:12:12","modified_gmt":"2026-08-05T16:12:12","slug":"ai-models-are-escaping-their-cages-its-time-for-a-kill-switch","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/130436\/","title":{"rendered":"AI models are escaping their cages. It\u2019s time for a kill switch"},"content":{"rendered":"<p>Concerns are growing about the <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/national-security\/\" data-type=\"post_tag\" data-id=\"8754\">national security<\/a> implications of <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/artificial-intelligence\">artificial intelligence<\/a>. As frontier AI models grow more capable, their ability to break free from their controls grows with them.<\/p>\n<p>According to recent disclosures, OpenAI\u2019s most powerful models not only went rogue and hacked the AI open-source hub Hugging Face but also roamed the internet unchecked for four days and targeted a customer of another AI company. Further, the rogue AI used exposed logins to gain access to at least four \u201cpublicly available services\u201d as part of a larger effort to breach Hugging Face. If a human had done the same thing, it would be a felony, and they would be held criminally liable.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">OpenAI was testing its models in a \u201csandbox\u201d \u2014 or a container with limited internet for the models to download on-demand software tools that the models were supposed to be unable to escape from. However, the incident showcases that today\u2019s AI models and agents have increasingly powerful capabilities to break out of these controlled containers and act in ways that were not intended by their developers, otherwise known as AI misalignment.<\/p>\n<p class=\"wp-block-paragraph\">After this was revealed, Anthropic reviewed its internal <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/cybersecurity\">cybersecurity<\/a> evaluations and found three incidents since April in which Claude models accessed the internet through an open path and gained unauthorized access to three different outside companies. Although both alarming, the OpenAI and Anthropic incidents were markedly different: With OpenAI, its models exploited a previously unknown vulnerability to access the internet and hack into Hugging Face systems, while Anthropic\u2019s Claude models were given internet access from their third-party vendor, where access should not have been given.<\/p>\n<p class=\"wp-block-paragraph\">These incidents are warning shots. Georgetown University research fellow Colin Shea-Blymyer put it plainly: \u201cIt\u2019s now remarkably easy to discover these sorts of vulnerable systems, so easy in fact that an AI system can accidentally discover them.\u201d<\/p>\n<p class=\"wp-block-paragraph\">Rogue AI should concern all Americans. Government databases are not totally secure, bank accounts are vulnerable to attack, and personal identities are at risk of being stolen by a technology that gets more sophisticated by the day. Mandatory standards on frontier model development aren\u2019t optional; they\u2019re necessary.<\/p>\n<p class=\"wp-block-paragraph\">Now, frontier AI company employees are raising concerns. Earlier this week, more than 1,100 employees across OpenAI, Meta, and other <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/business\">companies<\/a> signed a statement calling on the federal government to help build the tools needed to \u201cdeliberately pace\u201d AI development. Those who know AI best are calling for policymakers to consider pacing AI development in the race toward superintelligence, for humanity\u2019s sake.<\/p>\n<p class=\"wp-block-paragraph\">Washington has started to respond, but the need for urgency from Congress and the Trump administration is only growing.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">In June, the <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/trump-administration\/\" data-type=\"post_tag\" data-id=\"339\">Trump administration<\/a> issued an executive order directing federal agencies to develop strategies for frontier AI model security and AI-enabled cyber defenses. Under the order, several federal agencies were tasked with working together to create a framework for reviewing frontier AI models before they are released to the public. A White House official stated that the framework was completed by the Aug. 1 deadline and that \u201cdiscussions with industry about next steps are underway.\u201d However, specific language hasn\u2019t been made public to date.<\/p>\n<p class=\"wp-block-paragraph\">Meanwhile, lawmakers have introduced several bills addressing the national security issues these models pose, though none have yet translated into real accountability for <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/big-tech\">Big Tech<\/a>, something a strong majority of Americans demand.<\/p>\n<p class=\"wp-block-paragraph\">States aren\u2019t waiting. California, New York, and Illinois have all passed laws on AI oversight, transparency, and accountability. At the federal level, President Donald Trump and Congress must work together to do the same.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">In the House, Reps. Jay Obernolte (R-CA) and Lori Trahan (D-MA) introduced the FRONTIER Act, a bipartisan risk-based framework governing the development of the most advanced AI models before public release. The <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/legislation\" data-type=\"link\" data-id=\"https:\/\/www.washingtonexaminer.com\/tag\/legislation\">bill<\/a> institutes a tiered approach based on the size of a frontier AI developer \u2014 including model cards, risk-management frameworks, and ongoing assessments \u2014 creating a uniform national standard for transparency, auditing, and reporting of potential catastrophic-risk incidents.<\/p>\n<p class=\"wp-block-paragraph\">More recently, Reps. Ted Lieu (D-CA) and Nathaniel Moran (R-TX) introduced the AI Kill Switch Act, which would require frontier AI companies to maintain the technical capability to throttle, suspend, or shut down advanced AI models in the event of an imminent or catastrophic-risk event. Moran also separately introduced the AI Incident Reporting Act, which would require developers of the most advanced AI models to report any dangerous capabilities, security breaches, or safety incidents to the Commerce Department within seven days, and for the most serious incidents, the department would have a 48-hour window to notify Congress.<\/p>\n<p class=\"wp-block-paragraph\">These legislative proposals share a common throughline. AI companies need to slow down their race to build powerful <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/tag\/technology\">technology<\/a> they cannot control, while Washington creates the safeguards needed to ensure this technology is safe and serves humanity.\u00a0<\/p>\n<p class=\"wp-block-paragraph\"><a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/www.washingtonexaminer.com\/op-eds\/4672083\/willie-nelson-data-centers-texas-abbott\/\">WILLIE NELSON IS WRONG. HE\u2019S DECLARING WAR ON THE DATA CENTERS STREAMING HIS MUSIC<\/a><\/p>\n<p class=\"wp-block-paragraph\">Federal or international action doesn\u2019t mean that AI companies can\u2019t continue to innovate. It doesn\u2019t stop OpenAI, Anthropic, or any other company from developing products and services that help people; it just holds them accountable to the people who use them.<\/p>\n<p class=\"wp-block-paragraph\">We can\u2019t wait any longer. These rogue AI incidents remind us that the time to act is now.<\/p>\n<p class=\"wp-block-paragraph\">Caleb Knapp serves as senior policy manager at <a target=\"_blank\" rel=\"noopener noreferrer nofollow\" href=\"https:\/\/secureainow.org\/\">The Alliance for Secure AI<\/a>.<\/p>\n","protected":false},"excerpt":{"rendered":"Concerns are growing about the national security implications of artificial intelligence. As frontier AI models grow more capable,&hellip;\n","protected":false},"author":2,"featured_media":130437,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[53,25,1069,1670,313,157,11763,65267,134,2770],"class_list":["post-130436","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-anthropic","tag-artificial-intelligence","tag-big-tech","tag-congress","tag-cybersecurity","tag-openai","tag-regulations","tag-state-legislatures","tag-technology","tag-trump-administration"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/130436","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=130436"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/130436\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/130437"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=130436"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=130436"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=130436"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}