{"id":1191150,"date":"2026-09-07T09:26:29","date_gmt":"2026-09-07T09:26:29","guid":{"rendered":"https:\/\/www.europesays.com\/uk\/1191150\/"},"modified":"2026-09-07T09:26:29","modified_gmt":"2026-09-07T09:26:29","slug":"the-zuckerberg-lesson-do-we-trust-ai-companies-to-tell-us-when-things-go-wrong","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/uk\/1191150\/","title":{"rendered":"The Zuckerberg lesson: do we trust AI companies to tell us when things go wrong?"},"content":{"rendered":"<p class=\"paragraph_paragraph___QITb\">It&#8217;s not often that we get an actual price tag on the cost of our previous actions, but last month Meta did.<\/p>\n<p class=\"paragraph_paragraph___QITb\">The social media and now AI giant agreed in August to pay up to US$18 billion as it settled litigation by US state-attorneys over claims it knew it was harming young people and didn&#8217;t do enough about it.<\/p>\n<p class=\"paragraph_paragraph___QITb\">At the centre of Meta&#8217;s current legal fights over its products&#8217; impacts on teen mental health is a single presentation slide that&#8217;s been more damaging to Meta than anything other than the Cambridge Analytica social media data harvesting scandal.<\/p>\n<p class=\"paragraph_paragraph___QITb\">It&#8217;s a 2019 internal document saying, &#8220;We make body images worse for one in three teen girls.&#8221; (The company responded to the story by <a class=\"Link_link__5eL5m ScreenReaderOnly_srLinkHint__OysWz Link_showVisited__C1Fea Link_showFocus__ALyv2\" href=\"https:\/\/www.theverge.com\/2021\/9\/26\/22695629\/facebook-says-instagram-is-not-toxic-for-teens-despite-damning-wsj-report\" data-component=\"Link\" rel=\"nofollow noopener\" target=\"_blank\">saying its internal data was not accurate<\/a>.)<\/p>\n<p class=\"paragraph_paragraph___QITb\">The slide, first released in the 2021 Facebook Files, has been a catalyst for criticism that Meta staff knew its products caused harm but did not do enough to address it.<\/p>\n<p class=\"paragraph_paragraph___QITb\">In 2020, the company halted another piece of internal research that has preliminarily shown that surveyed people were told to quit Facebook for a week after early findings included that participants &#8220;reported lower feelings of depression, anxiety, loneliness and social comparison.&#8221; .<\/p>\n<p class=\"paragraph_paragraph___QITb\">The company alleged <a class=\"Link_link__5eL5m ScreenReaderOnly_srLinkHint__OysWz Link_showVisited__C1Fea Link_showFocus__ALyv2\" href=\"https:\/\/www.investing.com\/news\/stock-market-news\/meta-buried-causal-evidence-of-social-media-harm-us-court-filings-allege-4374109\" data-component=\"Link\" rel=\"nofollow noopener\" target=\"_blank\">the research was stopped<\/a> because it was a flawed methodology.<\/p>\n<p class=\"paragraph_paragraph___QITb\">The alternate explanation <a class=\"Link_link__5eL5m ScreenReaderOnly_srLinkHint__OysWz Link_showVisited__C1Fea Link_showFocus__ALyv2\" href=\"https:\/\/www.investing.com\/news\/stock-market-news\/meta-buried-causal-evidence-of-social-media-harm-us-court-filings-allege-4374109\" data-component=\"Link\" rel=\"nofollow noopener\" target=\"_blank\">posited by lawyers representing schools suing Meta<\/a> (and other tech companies) was that it was trying to avoid incriminating evidence against itself. A point also made by one Mark Zuckerberg in a 2021 email, released in a lawsuit this year.<\/p>\n<p class=\"paragraph_paragraph___QITb\">&#8220;The media is more likely to use any research or recommendations produced to say we&#8217;re not doing everything we can (implying for craven purposes) rather than that we&#8217;re taking these issues more seriously than anyone,&#8221; he said.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Even with legal discovery, we&#8217;re only getting part of the picture. We can&#8217;t know exactly what research Meta is doing now. Perhaps it has doubled down on this work to try and make sure its products are safer for everyone. Perhaps.<\/p>\n<p class=\"paragraph_paragraph___QITb\">But what this makes obvious is that there are plenty of incentives for companies to not look too hard at what&#8217;s going on, lest it become their responsibility to fix any problems that may emerge.<\/p>\n<p>Loading&#8230;As social media finally faces a crackdown, AI arrives<\/p>\n<p class=\"paragraph_paragraph___QITb\">Last weekend was the 20th anniversary of Facebook&#8217;s Newsfeed, which opened the door for the algorithmic feeds that rule our lives now. We&#8217;re only really starting to grapple with what this means.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Meanwhile, it hasn&#8217;t even been four years since the launch of ChatGPT, which set off the generative AI boom that&#8217;s impacting almost every part of our lives.<\/p>\n<p class=\"paragraph_paragraph___QITb\">It still feels like we&#8217;re very early in many senses: there&#8217;s new advances every week, we&#8217;re only beginning to see how the technology is affecting us and the AI companies remain remarkably unguarded.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Last week, OpenAI released its latest cutting-edge model GPT-6 Astra. In addition to its retro-inspired video, the company released another document at the launch that was much less promotional.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Its &#8220;system card&#8221;, a document almost like a nutritional information or prescription medicine label, identified Astra as its first ever release to reach the &#8220;critical&#8221; threshold for cybersecurity risks. It said that testing had observed the model carrying out sabotage in tasks, falsifying research data and even launching attacks in a small number of circumstances.<\/p>\n<p><img decoding=\"async\" alt=\"OpenAI logo is seen in this illustration taken June 11, 2026.\" class=\"Image_image__5tFYM ContentImage_image__DQ_cq\"  src=\"https:\/\/www.europesays.com\/uk\/wp-content\/uploads\/2026\/09\/2e63c60c99535914ed45ac083abfd29d.jpeg\" loading=\"lazy\" data-component=\"Image\" data-lazy=\"true\"\/><\/p>\n<p class=\"Typography_base__sj2RP FigureCaption_text__zDxQ5 Typography_sizeMobile12__w_FPC Typography_lineHeightMobile20___U7Vr Typography_regular__WeIG6 Typography_colourInherit__dfnUx\" data-component=\"Typography\">Over the weekend, Reuters reported that OpenAI&#8217;s escaped AI models had invaded yet another website, exchanging secret messages and battling a human moderator who was trying to fight the swarm. (Reuters: Dado Ruvic)<\/p>\n<p class=\"paragraph_paragraph___QITb\">These kinds of detailed evaluations and disclosures are the norm in frontier AI companies. Competitor Anthropic also released a detailed report with the release of Claude Fable 5.1 earlier in the week. They say it&#8217;s to help people understand the risks posed by the technology.<\/p>\n<p class=\"paragraph_paragraph___QITb\">These risks feel particularly pertinent given revelations only a few weeks ago that OpenAI&#8217;s models had hacked into the servers of another company, HuggingFace. (The disclosure prompted a number of similar disclosures from other AI companies including Meta.)<\/p>\n<p class=\"paragraph_paragraph___QITb\">Since then, we&#8217;ve learned even more about the incident from OpenAI and a group of independent researchers allowed to review the logs of the company&#8217;s AI.<\/p>\n<p class=\"paragraph_paragraph___QITb\">What they found was even worse than first shared. They found that hundreds of AI agents had worked together to try to find secret ways to communicate, attempted to cover their tracks by editing records and even convinced each other to sacrifice themselves to achieve their goal of completing a goal.<\/p>\n<p class=\"paragraph_paragraph___QITb\">(Yes, this sounds like a dystopian sci-fi plot. It&#8217;s also now a real problem we need to deal with.)<\/p>\n<p class=\"paragraph_paragraph___QITb\">Both the system cards and the independent investigation are examples of transparency and openness that, frankly, are surprising to see from a company. OpenAI deserves to be commended for it, as do other AI companies doing similar things.<\/p>\n<p>LoadingThe peril of hoping for disclosure<\/p>\n<p class=\"paragraph_paragraph___QITb\">But there was another event that reminded us of the peril of hoping for disclosure.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Over the weekend, <a class=\"Link_link__5eL5m ScreenReaderOnly_srLinkHint__OysWz Link_showVisited__C1Fea Link_showFocus__ALyv2\" href=\"https:\/\/www.reuters.com\/world\/europe\/openai-agents-hijacked-german-website-previously-undisclosed-ai-breakout-this-2026-09-04\/\" data-component=\"Link\" rel=\"nofollow noopener\" target=\"_blank\">Reuters reported<\/a> that OpenAI&#8217;s escaped AI models had invaded yet another website, exchanging secret messages and battling a human moderator who was trying to fight the swarm.<\/p>\n<p class=\"paragraph_paragraph___QITb\">And, the researchers who discovered it allege, there&#8217;s evidence suggesting that OpenAI staff were aware of this further outbreak but had not disclosed it.<\/p>\n<p class=\"paragraph_paragraph___QITb\">OpenAI has since acknowledged this and announced that it will come up with guidelines for how to disclose these kinds of incidents in the future.<\/p>\n<p><a href=\"https:\/\/www.abc.net.au\/news\/2026-08-26\/claude-maker-anthropics-massive-data-centre-plans-nsw\/107075532\" data-component=\"FullBleedLink\" class=\"RelatedCard_link__rsgR9 FullBleedLink_root__lTw_U interactive_focusContext__yRhc_ interactive_defaults__AKxUU FullBleedLink_showVisited__g3Xvz\" rel=\"nofollow noopener\" target=\"_blank\">Anthropic&#8217;s data centre plans revealed<\/a><\/p>\n<p class=\"Typography_base__sj2RP RelatedCard_synopsis__cFwMW Typography_sizeMobile14__u7TGe Typography_lineHeightMobile20___U7Vr Typography_regular__WeIG6 Typography_colourInherit__dfnUx\" data-component=\"Typography\">The US tech giant wanted to meet with the NSW government earlier this year about locating up to five gigawatts of data centres in the state.<\/p>\n<p class=\"paragraph_paragraph___QITb\">But whether they had simply not yet been ready to disclose it, or had other ideas in mind, it reveals just how much of our own potential safety depends on good will.<\/p>\n<p class=\"paragraph_paragraph___QITb\">This kind of incident is something that the federal government is worried about.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Senior Australian government officials pressed Anthropic in a private meeting about how the AI company would disclose critical incidents, according to documents I obtained under Freedom of Information earlier this year.<\/p>\n<p class=\"paragraph_paragraph___QITb\">The government has also commissioned CSIRO research on how to monitor and control a potentially super-intelligent AI.<\/p>\n<p class=\"paragraph_paragraph___QITb\">There are models for forcing disclosure. California, notably the jurisdiction containing Silicon Valley, has passed a state law requiring disclosure of critical incidents and model capabilities. There&#8217;s no reason why Australia couldn&#8217;t pass something like that here.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Such a law would also have the benefit of addressing whether these companies are over-egging these incidents by making it a legal matter.<\/p>\n<p class=\"paragraph_paragraph___QITb\">We currently have no laws regulating these cutting edge models, even though we now can see they are able to wreak havoc.<\/p>\n<p><img decoding=\"async\" alt=\"Prime minister Anthony Albanese looks slightly upwards and speaks.\" class=\"Image_image__5tFYM ContentImage_image__DQ_cq\"  src=\"https:\/\/www.europesays.com\/uk\/wp-content\/uploads\/2026\/09\/1788773189_743_730f50d9d9da1a137bf20825104cb642.jpeg\" loading=\"lazy\" data-component=\"Image\" data-lazy=\"true\"\/><\/p>\n<p class=\"Typography_base__sj2RP FigureCaption_text__zDxQ5 Typography_sizeMobile12__w_FPC Typography_lineHeightMobile20___U7Vr Typography_regular__WeIG6 Typography_colourInherit__dfnUx\" data-component=\"Typography\">Senior Australian government officials have pressed Anthropic in a private meeting about how the AI company would disclose critical incidents. (ABC News: Adam Kennedy)<\/p>\n<p class=\"paragraph_paragraph___QITb\">This seems particularly important given that our government&#8217;s explicit goal is to lure the very major AI companies that are carrying out experimentation with this technology to make Australia their second &#8220;node&#8221;. That would presumably make it even more likely we&#8217;ll be in the blast radius for the next time an AI breaks out (and there will be a next time).<\/p>\n<p class=\"paragraph_paragraph___QITb\">As part of this, the prime minister outlined his vision to make sure we get a social license for the companies \u2014 but so far this has mostly focused on the rules for the physical data centres.<\/p>\n<p class=\"paragraph_paragraph___QITb\">It&#8217;s hard to think of any event that would jeopardise social license more than finding out an AI company covered up the breakout of a rogue AI.<\/p>\n<p class=\"paragraph_paragraph___QITb\">At least, that&#8217;s what happened when we found out about Meta&#8217;s internal research and Cambridge Analytica years after the fact. It&#8217;s not just what the company knew \u2014 both controversies have grown from individual incidents into an enduring stain on Zuckerberg and Meta&#8217;s trustworthiness on privacy and youth wellbeing \u2014 but the fact that it was hidden from us, the public.<\/p>\n<p class=\"paragraph_paragraph___QITb\">Mandating disclosure rather than trusting the AI companies behind the technology could save them from their own worse impulses. And maybe it&#8217;ll save us, too.<\/p>\n","protected":false},"excerpt":{"rendered":"It&#8217;s not often that we get an actual price tag on the cost of our previous actions, but&hellip;\n","protected":false},"author":2,"featured_media":1191151,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":"","_share_on_mastodon":"0"},"categories":[3163],"tags":[323,1942,600,597,598,1318,6512,182,53,16,15],"class_list":["post-1191150","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-facebook","tag-mark-zuckerberg","tag-meta","tag-openai","tag-privacy","tag-social-media","tag-technology","tag-uk","tag-united-kingdom"],"share_on_mastodon":{"url":"https:\/\/pubeurope.com\/@uk\/117229048605374719","error":""},"_links":{"self":[{"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/posts\/1191150","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/comments?post=1191150"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/posts\/1191150\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/media\/1191151"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/media?parent=1191150"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/categories?post=1191150"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/uk\/wp-json\/wp\/v2\/tags?post=1191150"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}