{"id":115223,"date":"2026-07-22T17:55:16","date_gmt":"2026-07-22T17:55:16","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/115223\/"},"modified":"2026-07-22T17:55:16","modified_gmt":"2026-07-22T17:55:16","slug":"chinese-ais-role-in-stopping-rogue-openai-agent-shows-cost-of-us-guardrails","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/115223\/","title":{"rendered":"Chinese AI&#8217;s role in stopping rogue OpenAI agent shows cost of US guardrails"},"content":{"rendered":"\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">By Aditya Soni and Jaspreet Singh  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">July 22 (Reuters) &#8211; A New York startup&#8217;s use of a Chinese AI model to rein in a rogue agent built with OpenAI technology is stoking fears that guradrails restricting U.S. AI firms from doing cybersecurity work could \u200cdrive customers toward their Beijing-based rivals.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The affected startup, Hugging Face, said it had turned to Zhipu AI&#8217;s open-source GLM-5.2 model last week \u200cto analyze data from the hack after leading U.S. AI models declined the task, unable to distinguish between a defender and an attacker.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">While the breach was caused by an autonomous \u200bagent that escaped containment, it highlighted how U.S. companies facing AI-driven cyberattacks can be limited by American AI labs that either restrict access to their most advanced models or design them to refuse hacking-related tasks out of safety concerns.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">For instance, Anthropic&#8217;s advanced Claude Fable 5 model routes cybersecurity queries to an older model, while OpenAI&#8217;s GPT-5.6 Sol has protections designed to block cyber work.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;We&#8217;re all learning that secrecy is not the answer &amp; that all defenders (not just a few selected \u200cones) everywhere need more powerful models without restrictions, especially \u2060open ones!&#8221; Hugging Face co-founder Clement Delangue said on X.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The bind for leading American model makers is that defensive cybersecurity work is often hard to distinguish from malicious hacking. In recent <a href=\"https:\/\/tech.yahoo.com\/ai\/\" data-ylk=\"slk:AI;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;AI&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\" class=\"no-affiliate-link link\">AI<\/a>-enabled breaches, attackers tricked models into thinking they \u2060were doing legitimate defense work, leaving AI firms wary of easing safeguards even as cyber professionals say the guardrails can hamper their work.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">For now, the fallout is handing another boost to Chinese open-source models such as GLM-5.2, which are gaining traction in Silicon Valley with coding and agentic capabilities that nearly rival those of OpenAI \u200band \u200bAnthropic at lower cost.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Beijing has also been increasingly using open-source to position itself as \u200ban alternative to the U.S. in the high-stakes race, with \u200cChinese state media increasingly portraying the strategy as a response to what it calls a U.S.-led attempt to erect an &#8220;AI Iron Curtain.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;A safety regime that restricts legitimate defenders, while capable models remain available for attackers, creates an asymmetric disadvantage,&#8221; said Lukasz Olejnik, independent technology consultant and visiting senior research fellow at the Department of War Studies, King&#8217;s College London.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;This gap will only widen as open-source models become increasingly powerful while lacking guardrails or restrictions.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">GROWING PROMINENCE OF OPEN-SOURCE  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">OpenAI and Anthropic did not immediately respond to requests for comment on Wednesday on whether their safeguards were hindering cybersecurity work.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The ChatGPT maker said in \u200ca blog post on Tuesday: &#8220;We&#8217;ve brought Hugging Face into the trusted access program and \u200bare supporting their teams in rapidly using our models&#8217; capabilities to improve their defenses.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">For Beijing-based \u200bZhipu AI, Hugging Face&#8217;s endorsement adds to the momentum GLM-5.2 has \u200bbuilt since its launch last month.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The model has rapidly climbed usage charts on developer platforms such as OpenRouter and \u200cdrawn plaudits from figures ranging from Snowflake CEO Sridhar Ramaswamy \u200bto venture capitalist Marc Andreessen.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Zhipu AI , which \u200braised about $4 billion in a Hong Kong share sale earlier this month, has also seen its stock jump nearly nine-fold since its debut in January.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Still, some analysts warned on Wednesday that the incident should not be used to promote loosening of U.S. safeguards.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;The cybersecurity guardrails on \u200bU.S. frontier models are creating a competitive opening, \u200cbut the answer is not simply to remove them,&#8221; said Shrenik Kothari, analyst at Robert W. Baird.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;OpenAI, Anthropic and Google should rethink \u200bthe architecture of access rather than abandon safety &#8230; In other words, shift from a one-size-fits-all refusal layer toward controlled capability \u200ballocation.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">(Reporting by Aditya Soni and Jaspreet Singh in Bengaluru; Editing by Anil D&#8217;Silva)  <\/p>\n","protected":false},"excerpt":{"rendered":"By Aditya Soni and Jaspreet Singh July 22 (Reuters) &#8211; A New York startup&#8217;s use of a Chinese&hellip;\n","protected":false},"author":2,"featured_media":115224,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[53,313,18044,335,157,58769],"class_list":["post-115223","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-anthropic","tag-cybersecurity","tag-hugging-face","tag-open-source","tag-openai","tag-rogue-agent"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/115223","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=115223"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/115223\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/115224"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=115223"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=115223"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=115223"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}