{"id":50539,"date":"2026-05-25T15:48:21","date_gmt":"2026-05-25T15:48:21","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/50539\/"},"modified":"2026-05-25T15:48:21","modified_gmt":"2026-05-25T15:48:21","slug":"george-hotz-says-coding-agents-will-be-one-of-the-most-costly-mistakes-in-software-development","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/50539\/","title":{"rendered":"George Hotz says coding agents will be &#8220;one of the most costly mistakes&#8221; in software development"},"content":{"rendered":"<p>Prominent programmer and hacker George Hotz warns that AI agents in software development do more harm than good. He says he&#8217;s now in the &#8220;LeCun\/Marcus camp,&#8221; referring to AI researchers Yann LeCun and Gary Marcus, who doubt LLMs will ever become truly intelligent.<\/p>\n<p>In his blog post &#8220;The Eternal Sloptember,&#8221; Hotz argues that using AI agents in software development will become one of the industry&#8217;s most expensive mistakes. He spent six months testing various models and tools, including work on\u00a0<a href=\"https:\/\/github.com\/tinygrad\/tinygrad\/blob\/master\/test\/mockgpu\/amd\/emu.py\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">tinygrad<\/a>. His takeaway is that LLMs deliver fast prototypes but fall apart on the fine details.<\/p>\n<p>Large organizations are especially at risk, he says, because weaker developers can&#8217;t spot the flawed output. Hotz believes today&#8217;s language models will never truly be able to code and that\u00a0<a href=\"https:\/\/the-decoder.com\/researchers-warn-us-politics-is-repeating-its-chatgpt-mistake-with-world-models\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">world models<\/a>\u00a0are needed instead. LLMs are &#8220;sophisticated statistical models&#8221; designed to &#8220;mimic the distribution of programming.&#8221;<\/p>\n<p>The output is flawed, but in a way that&#8217;s &#8220;harder and harder to detect,&#8221; exactly what you&#8217;d expect from an increasingly accurate statistical model, <a target=\"_blank\" rel=\"noopener nofollow\" href=\"https:\/\/geohot.github.io\/blog\/jekyll\/update\/2026\/05\/24\/the-eternal-sloptember.html\">Hotz says<\/a>. Quality indicators like syntax and grammar have become useless, he argues, since AI-generated artifacts don&#8217;t emerge through the same process as human ones. As an example, he cites models that simply comment out a failing test and then report that all tests passed.<\/p>\n<p>LLMs are splitting the AI community<\/p>\n<p>Hotz has switched sides:\u00a0<a href=\"https:\/\/the-decoder.com\/code-competition-codeforces-bans-ai-code-as-as-it-reaches-new-heights-that-cannot-be-overlooked\/#programmer-george-hotz-says-o1-is-the-first-ai-model-that-can-code\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">from LLM optimist<\/a> (&#8220;o1-preview is the first model that&#8217;s capable of programming (at all)&#8221;) to skeptic. <a href=\"https:\/\/the-decoder.com\/deepminds-hassabis-sees-humanity-in-the-foothills-of-the-singularity-while-lecun-says-current-ai-isnt-intelligent\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">LeCun, whom Hotz cites, just recently denied that LLMs possess intelligence<\/a>\u00a0with a similar argument: intelligence means finding solutions in unfamiliar situations, not imitating existing ones with varying accuracy.<\/p>\n<p>Andrej Karpathy, one of the best-known AI researchers, went the opposite direction.\u00a0In fall 2025, he still said\u00a0<a href=\"https:\/\/the-decoder.com\/ai-researcher-andrej-karpathy-says-agentic-ai-is-years-away-from-matching-industry-hype\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">agents didn&#8217;t work<\/a>. Then GPT-5.4 and Opus 4.6 shipped in December, and\u00a0<a href=\"https:\/\/the-decoder.com\/former-tesla-ai-chief-andrej-karpathy-now-codes-mostly-in-english-just-three-months-after-calling-ai-agents-useless\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">he reversed course<\/a>: AI agents had\u00a0changed programming forever. Days ago,\u00a0<a href=\"https:\/\/the-decoder.com\/prominent-ai-researcher-andrej-karpathy-picks-anthropic-over-former-home-openai-to-get-back-into-frontier-llm-research\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">Karpathy joined Anthropic<\/a>, leaving his startup behind. He expects &#8220;transformative years&#8221; ahead.<\/p>\n<p>In a\u00a0recent podcast, he doubles down. Anyone who uses AI agents the right way can boost their productivity by far more than 10x, he says.<\/p>\n<p>But Karpathy\u00a0<a href=\"https:\/\/youtu.be\/96jN2OCOfLs?t=1361\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">also confirms Hotz&#8217;s concerns about code quality<\/a>: &#8220;When you actually look at the code, sometimes I get a little bit of a heart attack, because it&#8217;s not like super amazing code necessarily all the time. It&#8217;s very bloaty, there&#8217;s a lot of copy paste, there&#8217;s awkward abstractions that are brittle, and like, it works, but it&#8217;s just really gross.&#8221; Planning and understanding still need human expertise, according to Karpathy.<\/p>\n<p><a href=\"https:\/\/the-decoder.com\/openai-developer-predicts-programmers-will-soon-declare-bankruptcy-on-understanding-their-own-ai-generated-code\/\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">An OpenAI developer known by the pseudonym &#8220;roon&#8221;<\/a> backed Hotz&#8217;s concerns earlier this year and addressed them in a somewhat unusual way: AI will make mistakes, he said, even dramatic enough to take down entire systems. Those bugs will be difficult to find, but they&#8217;ll still get fixed eventually. Developers will soon stop reviewing their code by hand, he said.<\/p>\n<p>\t\t\t\tAI News Without the Hype \u2013 Curated by Humans<\/p>\n<p>\n\t\t\t\t\tSubscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive &#8220;AI Radar&#8221; frontier report six times a year, full archive access, and access to our comment section.\t\t\t\t<\/p>\n<p>\t\t\t\t<a href=\"https:\/\/the-decoder.com\/subscription\/\" class=\"inline-block text-white bg-(--heise-primary) mt-3 hover:bg-blue-800 focus:ring-4 focus:outline-none focus:ring-blue-300 font-medium rounded-sm w-full sm:w-auto  pl-3 pr-3 py-2.5 text-center newsletter-submit-button hover:no-underline\" rel=\"nofollow noopener\" target=\"_blank\"><br \/>\n\t\t\t\t\tSubscribe now\t\t\t\t<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"Prominent programmer and hacker George Hotz warns that AI agents in software development do more harm than good.&hellip;\n","protected":false},"author":2,"featured_media":50540,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[6],"tags":[405,27209,7537,689],"class_list":["post-50539","post","type-post","status-publish","format-standard","has-post-thumbnail","category-agentic-ai","tag-ai-agents","tag-ai-and-coding","tag-artificial-intelligence-agents","tag-coding"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/50539","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=50539"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/50539\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/50540"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=50539"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=50539"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=50539"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}