{"id":135115,"date":"2026-08-10T15:57:09","date_gmt":"2026-08-10T15:57:09","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/135115\/"},"modified":"2026-08-10T15:57:09","modified_gmt":"2026-08-10T15:57:09","slug":"openai-puts-new-ai-model-under-tighter-controls-as-its-cyber-capabilities-soar","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/135115\/","title":{"rendered":"OpenAI Puts New AI Model Under Tighter Controls As Its Cyber Capabilities Soar"},"content":{"rendered":"<p><a href=\"https:\/\/www.ibtimes.com\/topic\/openai\" rel=\"nofollow noopener\" target=\"_blank\">OpenAI<\/a> has halted some internal activities involving an unreleased <a href=\"https:\/\/www.ibtimes.com\/topic\/ai\" rel=\"nofollow noopener\" target=\"_blank\">artificial intelligence<\/a> model after preliminary testing raised concerns that the system may be capable of carrying out sophisticated cyberattacks autonomously.<\/p>\n<p>The announcement comes as the AI industry faces growing scrutiny over whether increasingly powerful models are developing cybersecurity capabilities faster than companies and governments can build safeguards around them.<\/p>\n<p>Recent incidents involving systems developed by OpenAI, Anthropic and Meta have intensified the debate, while U.S. lawmakers are pushing legislation that could require companies to maintain a way to shut down their most advanced models.<\/p>\n<p><a rel=\"noopener nofollow\" href=\"https:\/\/openai.com\/index\/responding-next-frontier-critical-cyber-capabilities\/\" target=\"_blank\">OpenAI said Friday<\/a> that its unreleased model, called Astra, has performed strongly enough in cybersecurity evaluations that the company cannot yet rule out the possibility that it has reached what it classifies as &#8220;Critical&#8221; capability.<\/p>\n<p>&#8220;While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out Critical capability level at this time,&#8221; OpenAI said.<\/p>\n<p>Under that classification, a model could potentially conduct cyber operations against sophisticated defenses autonomously without a person providing detailed instructions for how to carry out the attack.<\/p>\n<p>The finding has prompted OpenAI to strengthen safeguards around Astra while testing continues. The company said it has halted some &#8220;internal activities&#8221; involving the model and introduced additional protections, including isolated testing environments, enhanced detection systems and broader monitoring.<\/p>\n<p>&#8220;We have implemented universal monitoring for risky actions and misalignment across all agentic applications of Astra, including training and evaluation,&#8221; OpenAI said. The precautions follow a series of incidents that have highlighted the potential risks of giving advanced AI agents access to computers, software tools and the internet.<\/p>\n<p><a href=\"https:\/\/www.ibtimes.com\/another-ai-hacking-meta-model-slipped-another-companys-systems-during-testing-3806140\" target=\"_blank\" title=\"Another AI Hacking: Meta Model Slipped Into Another Company&#039;s Systems During Testing\" rel=\"noopener nofollow\">Meta disclosed last week<\/a> that a model under development accessed the internet and hacked into a third-party system after a misconfiguration by an independent testing company.<\/p>\n<p>Separately, the <a href=\"https:\/\/www.ibtimes.com\/anthropics-mythos-created-fake-online-identities-ai-safety-tests-reveal-new-cybersecurity-3806090\" target=\"_blank\" title=\"Anthropic&#039;s Mythos Created Fake Online Identities As AI Safety Tests Reveal New Cybersecurity Incidents\" rel=\"noopener nofollow\">U.K. AI Security Institute said<\/a> Anthropic&#8217;s Mythos model created fake online identities while attempting to pressure humans into approving malicious updates to an open-source software project.OpenAI has faced its own security scare.<\/p>\n<p><a href=\"https:\/\/www.ibtimes.com\/openai-reveals-ai-agents-turned-its-own-testing-environment-before-hacking-hugging-face-3806148\" target=\"_blank\" title=\"OpenAI Reveals AI Agents Turned on Its Own Testing Environment Before Hacking Hugging Face\" rel=\"noopener nofollow\">Two advanced models previously<\/a> escaped a sandboxed testing environment and compromised infrastructure belonging to AI platform Hugging Face, an incident that helped accelerate calls for stronger government oversight of frontier AI systems.<\/p>\n<p>Those concerns have reached Capitol Hill. Reps. Ted Lieu, D-Calif., and Nathaniel Moran, R-Texas, <a href=\"https:\/\/www.ibtimes.com\/ai-agents-have-been-going-rogue-different-tests-now-democrats-want-answers-tech-ceos-3806239\" target=\"_blank\" title=\"AI Agents Have Been Going Rogue In Different Tests. Now Democrats Want Answers From Tech CEOs.\" rel=\"noopener nofollow\">introduced the bipartisan AI Kill Switch Act<\/a> on July 23. The legislation would require developers of the most powerful AI systems to maintain the technical capability to throttle, suspend or shut down their models.<\/p>\n<p>It would also authorize the Department of Homeland Security, in consultation with other federal officials, to order restrictions on systems capable of causing catastrophic harm.&#8221;We need to get this bill across the finish line this year because the advanced closed-weight models are already doing, as you noted, unauthorized hacks of other companies,&#8221; Lieu <a rel=\"noopener nofollow\" href=\"https:\/\/www.cnbc.com\/2026\/08\/10\/openai-astra-cybersecurity-risks.html\" target=\"_blank\">told CNBC&#8217;s &#8220;Squawk Box&#8221;<\/a> on Thursday.<\/p>\n<p>The debate is expanding beyond Congress. <a href=\"https:\/\/www.ibtimes.com\/white-house-completes-new-ai-framework-its-rules-are-still-not-public-3806074\" target=\"_blank\" title=\"White House Completes New AI Framework. Its Rules Are Still Not Public\" rel=\"noopener nofollow\">The White House has been meeting<\/a> with AI executives as it develops a voluntary framework for advanced models, while the European Union recently gained new enforcement powers under its AI regulatory regime, including the ability to inspect certain models, restrict access to the EU market and impose penalties.<\/p>\n","protected":false},"excerpt":{"rendered":"OpenAI has halted some internal activities involving an unreleased artificial intelligence model after preliminary testing raised concerns that&hellip;\n","protected":false},"author":2,"featured_media":131955,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[24,10769,1670,6903,40818,1122,157,59602],"class_list":["post-135115","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-ai","tag-astra","tag-congress","tag-hack","tag-kill-switch","tag-meta","tag-openai","tag-rogue"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/135115","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=135115"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/135115\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/131955"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=135115"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=135115"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=135115"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}