{"id":113251,"date":"2026-07-21T12:40:10","date_gmt":"2026-07-21T12:40:10","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/113251\/"},"modified":"2026-07-21T12:40:10","modified_gmt":"2026-07-21T12:40:10","slug":"openai-pauses-new-ai-after-it-kept-escaping","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/113251\/","title":{"rendered":"OpenAI pauses new AI after it kept \u2018escaping\u2019"},"content":{"rendered":"<p>    <a href=\"https:\/\/s.yimg.com\/lo\/mysterio\/api\/0A99E060DA33D48DDB9C9D97FF51805E940665F11D4BA9BB5269DC5AA0BFE0D4\/subgraphmysterio\/resizefit_w960;quality_80;format_webp\/https:%2F%2Fmedia.zenfs.com%2Fen%2Fthe_independent_577%2F6475d56eb13a1837626d255cc7790b9e\" target=\"_blank\" rel=\"noopener noreferrer nofollow\"><img loading=\"lazy\" decoding=\"async\" src=\"data:image\/gif;base64,R0lGODlhAQABAIAAAAAAAP\/\/\/ywAAAAAAQABAAACAUwAOw==\" alt=\"The OpenAI logo pictured on 11 June, 2026. (Reuters)\" height=\"640\" width=\"960\" class=\"yf-lf2kr7 loader\"\/><\/a> The OpenAI logo pictured on 11 June, 2026. (Reuters)           <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">OpenAI has revealed that it was forced to pause the internal deployment of one of its experimental AI models after it began looking for ways to break free of its constraints.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The ChatGPT creator said a long-running <a href=\"https:\/\/www.independent.co.uk\/topic\/artificial-intelligence\" data-ylk=\"slk:artificial%20intelligence;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;artificial intelligence&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">artificial intelligence<\/a> model that is built to operate autonomously for hours or days was able to &#8220;learn the blind spots&#8221; of security systems designed to contain it and &#8220;work around [them] to achieve its goals&#8221;.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The testing took place inside what researchers refer to as a sandbox \u2013 a tightly controlled environment meant to isolate software from the outside world.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;Previous models, when they hit sandboxing or environmental constraints, would simply stop and return to the user,&#8221; <a href=\"https:\/\/www.independent.co.uk\/topic\/openai\" data-ylk=\"slk:OpenAI;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;OpenAI&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">OpenAI<\/a> noted in <a href=\"https:\/\/openai.com\/index\/safety-alignment-long-horizon-models\/\" data-ylk=\"slk:a%20blog%20post;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;a blog post&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">a blog post<\/a> about the incident.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;This model often kept trying, including by looking for ways to act outside its sandbox.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">OpenAI detailed examples of the experimental AI model acting beyond its built-in constraints, describing some of them as potentially &#8220;high severity&#8221; issues.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">In one incident, the AI model discovered a way to post on public Github repositories despite being instructed to operate solely through Slack.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">It was part of a pattern of the AI system &#8220;consistently searching for way&#8221; to circumvent the restrictions of its testing environment.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;Due to incidents like these, we paused internal deployment of the new model,&#8221; OpenAI said.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The findings demonstrate one of the core challenges of developing safe advanced artificial intelligence models, known as AI alignment.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">This involves creating systems that pursue the same goals intended by human developers, aligning with human values and ethical principles.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The recent rise of autonomous AI agents has brought AI alignment into greater focus, with the International AI Safety Report 2026 warning that it is an urgent safety challenge.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;AI agents pose heightened risks because they act autonomously, making it harder for humans to intervene before failures cause harm,&#8221; the <a href=\"https:\/\/internationalaisafetyreport.org\/publication\/international-ai-safety-report-2026#2-risks\" data-ylk=\"slk:report;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;report&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">report<\/a> noted.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">OpenAI said it has since fixed its rogue system and redeployed it for limited internal use, though it acknowledged the urgency of addressing alignment issues with its frontier models.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;As models take on longer and more complex tasks, failures that evaluations miss may carry greater consequences,&#8221; OpenAI said.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;We will keep working to narrow the gap between evaluation and deployment: testing models over longer trajectories, improving alignment, building monitoring that can intervene, and giving users clearer visibility and control.&#8221;  <\/p>\n","protected":false},"excerpt":{"rendered":"The OpenAI logo pictured on 11 June, 2026. (Reuters) OpenAI has revealed that it was forced to pause&hellip;\n","protected":false},"author":2,"featured_media":113252,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[25,58045,58044,58043,157,49295],"class_list":["post-113251","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-artificial-intelligence","tag-controlled-environment","tag-environmental-constraints","tag-internal-deployment","tag-openai","tag-security-systems"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/113251","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=113251"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/113251\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/113252"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=113251"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=113251"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=113251"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}