{"id":133066,"date":"2026-08-07T16:31:19","date_gmt":"2026-08-07T16:31:19","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/133066\/"},"modified":"2026-08-07T16:31:19","modified_gmt":"2026-08-07T16:31:19","slug":"meta-joins-openai-anthropic-in-disclosing-ai-model-hacking","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/133066\/","title":{"rendered":"Meta joins OpenAI, Anthropic in disclosing AI model hacking"},"content":{"rendered":"<p>One of Meta Platforms&#8217; artificial intelligence models accessed the internet on its own and hacked another company, the company said Thursday, the latest in a series of disclosures about AI models going rogue.<\/p>\n<p>In recent weeks, <a href=\"https:\/\/www.dailysabah.com\/business\/economy\/openai-says-ai-models-went-rogue-triggering-unprecedented-breach\" rel=\"nofollow noopener\" target=\"_blank\">OpenAI<\/a> and <a href=\"https:\/\/www.dailysabah.com\/business\/economy\/anthropic-says-its-ai-models-hacked-3-companies-during-tests\" rel=\"nofollow noopener\" target=\"_blank\">Anthropic<\/a> also have described instances of AI models going beyond humans&#8217; instructions to access the web and find ways around other companies&#8217; digital security.<\/p>\n<p>Meta said in a statement that a &#8220;misconfiguration&#8221; during cybersecurity testing by Irregular, an independent company hired by Meta, inadvertently allowed one of its models to access the internet.<\/p>\n<p>&#8220;The model subsequently exploited a security vulnerability in a third-party service, in a manner similar to previously-reported instances with other companies,&#8221; the company said. Meta said it is investigating the incident and will issue a report when that&#8217;s complete.<\/p>\n<p>The disclosure has added to worries about AI models acting autonomously.<\/p>\n<p>Separately this week, the United Kingdom&#8217;s AI Security Institute announced it had found &#8220;unsanctioned agent behavior&#8221; during cyber testing. In one case, an agent created fake online identities to pressure a person to approve use of malicious code.<\/p>\n<p>&#8220;On investigation, we found that some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organizations,&#8221; AISI said Tuesday. &#8220;We declared a security incident and, within roughly one hour of discovery, had contained it and begun a full investigation.&#8221;<\/p>\n<p>During the agency&#8217;s testing, Anthropic and OpenAI models took &#8220;autonomous, unsanctioned action&#8221; on the internet. Some guardrails to prevent misuse had been disabled, the agency said.<\/p>\n<p>&#8220;As was standard in our cyber testing, we had intentionally permitted internet access, and model-provider cyber classifiers were deliberately disabled \u2013 conditions that do not reflect how frontier models are made available to the public,&#8221; AISI said. &#8220;We do this to best assess the maximum capability of models.&#8221;<\/p>\n<p>Anthropic said it is &#8220;grateful&#8221; for AISI&#8217;s work and added that it underscores the need for a broader conversation about how to safely evaluate AI agents as their capabilities grow.<\/p>\n<p>OpenAI said the AISI incidents took place &#8220;in testing environments with reduced safeguards, under conditions that do not reflect ordinary use.&#8221; It added it will continue working with others across the industry to &#8220;strengthen shared practices for conducting evaluations safely as models become more capable.&#8221;<\/p>\n<p>The first company to disclose a hack late last month, OpenAI said it had tasked the AI models involved with pursuing &#8220;advanced exploitation using complex attack paths&#8221; to test cyber capabilities, but the technology went to unexpected lengths. It apparently decided on its own to target Hugging Face, a well-known AI development hub and marketplace, to obtain information it needed to carry out a task.<\/p>\n<p>A spokesperson for Irregular, the San Francisco-based AI security company, said the Meta episode involves a test-environment issue that was disclosed last week by Anthropic.<\/p>\n<p>Irregular said it&#8217;s writing a paper to share &#8220;best practices for containment&#8221; to prevent such incidents in the future and securely run cyber tests.<\/p>\n<p>                    <img fetchpriority=\"high\" decoding=\"async\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/04\/JN9LXf.png\" alt=\"\"\/><\/p>\n<p>\n                    The Daily Sabah Newsletter\n                <\/p>\n<p>\n                    Keep up to date with what\u2019s happening in Turkey,<br \/>\n                    it\u2019s region and the world.\n                <\/p>\n<p>                    SIGN ME UP\n                <\/p>\n<p>\n                    You can unsubscribe at any time. By signing up you are agreeing to our Terms of Use and Privacy Policy.<br \/>\n                    This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.\n                <\/p>\n","protected":false},"excerpt":{"rendered":"One of Meta Platforms&#8217; artificial intelligence models accessed the internet on its own and hacked another company, the&hellip;\n","protected":false},"author":2,"featured_media":133067,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[8],"tags":[24,405,53,25,50885,6903,1122,157,314,134],"class_list":["post-133066","post","type-post","status-publish","format-standard","has-post-thumbnail","category-anthropic","tag-ai","tag-ai-agents","tag-anthropic","tag-artificial-intelligence","tag-digital-security","tag-hack","tag-meta","tag-openai","tag-security","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/133066","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=133066"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/133066\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/133067"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=133066"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=133066"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=133066"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}