{"id":137681,"date":"2026-08-12T18:54:21","date_gmt":"2026-08-12T18:54:21","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/137681\/"},"modified":"2026-08-12T18:54:21","modified_gmt":"2026-08-12T18:54:21","slug":"claude-raises-riemann-zeta-zeros-to-67-2-two-papers-no-one-had-combined","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/137681\/","title":{"rendered":"Claude Raises Riemann Zeta Zeros to 67.2%: Two Papers No One Had Combined"},"content":{"rendered":"<p><img loading=\"lazy\" decoding=\"async\" class=\"mapping-embed imgPhoto\" id=\"i472459\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/claymath.jpg\" alt=\"ClayMath\" width=\"836\" height=\"418\"\/><\/p>\n<p>Claymath.org<\/p>\n<p>An unreleased research version of Claude spent 36 hours autonomously coordinating roughly 60 AI subagents \u2014 generating 650 failed ideas before landing on a cross-domain insight that two existing papers could be combined in a way no human mathematician had tried \u2014 and emerged with the largest single-step advance on a famous 167-year-old math problem in decades. On August 10, <a href=\"https:\/\/www.anthropic.com\/research\/riemann-zeta\" rel=\"nofollow noopener\" target=\"_blank\">Anthropic&#8217;s research post<\/a> announced that the model had raised the proven lower bound for the fraction of Riemann zeta function zeros lying on the so-called &#8220;critical line&#8221; from 41.6% to 67.2% \u2014 a 25.6-percentage-point jump accomplished in roughly 36 hours.<\/p>\n<p>According to one observer, a partner and researcher at Menlo Ventures, it may be the <a href=\"https:\/\/www.kucoin.com\/news\/flash\/ai-model-claude-advances-riemann-zeta-function-research-raises-zero-point-bound-to-67-2\" rel=\"nofollow noopener\" target=\"_blank\">most significant number-theory advance since 2013,<\/a> a field that had moved that same lower bound only 0.8 percentage points over the previous 37 years.<\/p>\n<p>The result did not prove the Riemann hypothesis \u2014 one of mathematics&#8217; <a href=\"https:\/\/www.claymath.org\/millennium-problems\/\" rel=\"nofollow noopener\" target=\"_blank\">seven Millennium Prize Problems<\/a>, with a $1 million bounty from the Clay Mathematics Institute for a complete solution. But what the run produced was something specific, verifiable, and already mechanically confirmed: a <a href=\"https:\/\/github.com\/anthropics\/zeta-23-lean\" rel=\"nofollow noopener\" target=\"_blank\">public Lean 4 proof<\/a> on GitHub, that any mathematician in the world can check for themselves. The method itself is available for inspection. The AI model that produced it is not.<\/p>\n<p>What the Riemann Hypothesis Actually Is<\/p>\n<p>To understand why a lower bound on zeta zeros matters, a brief foundation. The <a href=\"https:\/\/en.wikipedia.org\/wiki\/Riemann_zeta_function\" rel=\"nofollow noopener\" target=\"_blank\">Riemann zeta function<\/a> \u2014 first fully studied by Bernhard Riemann in an 1859 paper \u2014 is a mathematical object whose behavior encodes how prime numbers are distributed among the integers. The function has infinitely many &#8220;nontrivial zeros,&#8221; points where it equals zero, and the <a href=\"https:\/\/en.wikipedia.org\/wiki\/Riemann_hypothesis\" rel=\"nofollow noopener\" target=\"_blank\">Riemann hypothesis conjecture page<\/a> describes how every one of those zeros is conjectured to lie on a specific vertical line in the complex plane \u2014 the critical line, where the real part of the complex argument equals exactly 1\/2.<\/p>\n<p>Nobody has proved or disproved it. Because a complete proof remains out of reach, mathematicians have instead spent decades working on the adjacent question: what minimum fraction of those zeros can we rigorously prove lies on the critical line? G.H. Hardy proved in 1914 that infinitely many zeros are on the line; Levinson showed in 1974 that the fraction is at least 1\/3; Conrey improved that to 2\/5 in 1989; and a 2018 paper by Pratt, Robles, Zaharescu, and Zeindler established the <a href=\"https:\/\/kingy.ai\/blog\/claude-riemann-hypothesis-67-percent-result\/\" rel=\"nofollow noopener\" target=\"_blank\">previous state of art<\/a> at 41.6%.<\/p>\n<p>Claude&#8217;s run raised that floor to 67.2%.<\/p>\n<p>From a Dare to a Day and a Half of Subagents<\/p>\n<p>Jarred Sumner \u2014 the creator of the Bun JavaScript runtime, who <a href=\"https:\/\/www.anthropic.com\/news\/anthropic-acquires-bun-as-claude-code-reaches-usd1b-milestone\" rel=\"nofollow noopener\" target=\"_blank\">joined Anthropic when Bun was acquired<\/a> in December 2025 \u2014 gave the model a prompt that a professional mathematician would have found alarming: &#8220;Take a real stab at the Riemann hypothesis.&#8221; Sumner&#8217;s prompt left every mathematical choice to the model. He is not a mathematician.<\/p>\n<p>Claude&#8217;s first response was something resembling productive skepticism. Over the initial session, the model generated and evaluated 650 distinct ideas \u2014 every one of them a dead end. Sumner prompted it to try again, with input that Anthropic describes as &#8220;mostly variants of &#8216;keep going&#8217; or &#8216;believe in yourself.'&#8221; What followed was a day and a half of autonomous multi-agent work inside Claude Code: approximately 60 subagents operating in parallel, running 2,400 shell commands, writing hundreds of Python scripts, performing thousands of numerical checks against known zeta zeros, and having subagents critique one another&#8217;s reasoning. The agents also downloaded and reviewed 54 papers from arXiv to confirm the eventual finding hadn&#8217;t already been published. Total compute across both sessions: 31 million output tokens.<\/p>\n<p>The encouragement pattern is not incidental. The model had internalized from its training both that open mathematical problems are very hard and that AI systems have limitations \u2014 and had initially applied those priors conservatively to its own search. A similar encouragement approach, Anthropic notes, was used earlier in 2026 to help Claude produce <a href=\"https:\/\/kingy.ai\/blog\/claude-fable-jacobian-conjecture-counterexample\/\" rel=\"nofollow noopener\" target=\"_blank\">a counterexample to the Jacobian conjecture<\/a>, another longstanding open problem.<\/p>\n<p>The Specific Cross-Paper Bridge No Human Had Made<\/p>\n<p>The mathematical substance of the result turns on a connection between two bodies of work that no researcher had previously combined.<\/p>\n<p>In 1973, Hugh Montgomery introduced a set of pair correlation techniques for studying the distribution of zeta zeros. Those techniques were powerful \u2014 but they assumed the Riemann hypothesis was true, which meant they could not be used to prove results about zeros unconditionally. More recently, Baluyot, Goldston, Suriajaya, and Turnage-Butterbaugh published <a href=\"https:\/\/arxiv.org\/abs\/2306.04799\" rel=\"nofollow noopener\" target=\"_blank\">unconditional pair correlation results<\/a> that removed that assumption, producing unconditional versions of Montgomery&#8217;s methods that could support work on raising the lower bound.<\/p>\n<p>Separately, <a href=\"https:\/\/eudml.org\/doc\/252338\" rel=\"nofollow noopener\" target=\"_blank\">a 2000 paper by Enrico Bombieri<\/a> \u2014 a Fields Medal winner \u2014 studied a quadratic form associated with Weil&#8217;s explicit formula, the identity linking prime numbers to zeta zeros. Bombieri showed that this quadratic form is positive semidefinite if and only if the Riemann hypothesis holds, and that violations of the hypothesis correspond exactly to negative eigenvalues of the form.<\/p>\n<p>Claude&#8217;s insight was to treat the entire function space simultaneously \u2014 accounting together for the positive-definite subspace (zeros on the critical line) and the negative-definite subspace (zeros off it), allowing the associated quadratic form to be non-diagonal. By writing down a rank inequality using first- and second-moment information, computed via the dual picture over primes and Hilbert transform control, the model derived a bound that improved on what either body of work could produce alone. In Anthropic&#8217;s own description, the result &#8220;emerged as the unintended byproduct&#8221; of attempting the full hypothesis.<\/p>\n<p>Note that Dan Goldston \u2014 one of the external reviewers who examined Claude&#8217;s paper \u2014 is also a co-author of the Baluyot et al. papers that Claude&#8217;s proof relies on. His presence in both capacities strengthens the validation: the researcher who helped produce the upstream machinery reviewed whether the downstream application of it was correct.<\/p>\n<p>Verification: Lean Proof, Four Named Reviewers, No Journal Submission Yet<\/p>\n<p>Two in-house mathematicians at Anthropic \u2014 Levent Alp\u00f6ge and Ralph Furman \u2014 examined Claude&#8217;s paper and produced an <a href=\"https:\/\/www-cdn.anthropic.com\/23455459f8832d06bb175cc0f88d019aed962ef8.pdf\" rel=\"nofollow noopener\" target=\"_blank\">informal note for experts<\/a>, summarizing the proof concisely. Two external analytic number theory specialists, Brian Conrey and Dan Goldston, reviewed the work on short notice. Claude also worked with Anthropic staff member Eric Easley to produce a <a href=\"https:\/\/github.com\/anthropics\/zeta-23-lean\" rel=\"nofollow noopener\" target=\"_blank\">Lean 4 formalization<\/a> of the result \u2014 a mechanically verifiable version of the proof that passes the standard validation tool comparator.<\/p>\n<p>This verification package represents an unusual asymmetry in the epistemic landscape of the result. The mathematical claim is more auditable than almost any comparable human-produced paper: the Lean formalization means that anyone with access to the proof assistant can check the logical derivation step by step, without trusting the author&#8217;s judgment or reading ability. At the same time, the AI capability claim is less reproducible than almost any other published AI benchmark: the model used is unreleased, with no public weights, no accessible checkpoint, and no API. A researcher who wants to verify that Claude produced this result has to take Anthropic&#8217;s word for the process, even as the mathematical output itself is independently checkable. As one independent analysis observed, this puts reviewers in the position of <a href=\"https:\/\/kingy.ai\/blog\/claude-riemann-hypothesis-67-percent-result\/\" rel=\"nofollow noopener\" target=\"_blank\">validating a proof<\/a> without access to the tool that generated the ideas.<\/p>\n<p>As of August 12, the paper has not undergone conventional peer review at a mathematics journal.<\/p>\n<p>What It Doesn&#8217;t Mean \u2014 and What the 32.8% Tells You<\/p>\n<p>Anthropic is direct on one point: the techniques Claude used are not expected to lead to a complete proof of the Riemann hypothesis. Raising the lower bound to 67.2% establishes that at least two-thirds of zeta zeros are on the critical line. It says nothing about the remaining 32.8% \u2014 those zeros may well also be on the line, but the proof does not establish it. The <a href=\"https:\/\/www.claymath.org\/millennium-problems\/\" rel=\"nofollow noopener\" target=\"_blank\">Clay Prize requirements page<\/a> makes clear it requires a complete proof, not an improved lower bound. The hypothesis itself remains open.<\/p>\n<p>What the result does establish is a verifiable advance on a benchmark that the mathematics community had moved only incrementally over 46 years of combined human effort. The analogy to the <a href=\"https:\/\/en.wikipedia.org\/wiki\/Yitang_Zhang\" rel=\"nofollow noopener\" target=\"_blank\">bounded prime gap result<\/a> of 2013 \u2014 when Yitang Zhang proved for the first time that infinitely many pairs of primes are separated by at most 70 million, breaking a decades-long standstill \u2014 is imperfect but instructive. Both results were notable not just for what they proved, but for demonstrating that a long-stalled problem was still accessible to new techniques.<\/p>\n<p>What the Template Implies Beyond This Result<\/p>\n<p>The methodology of the Riemann zeta run \u2014 many autonomous subagents targeting a hard technical problem, numerical verification at scale, Lean formalization as a mechanical audit layer, and named human experts as the closing validation step \u2014 is attracting at least as much attention as the result itself.<\/p>\n<p>This is not the first time Claude has surfaced unexpected mathematics in 2026. Earlier this year, Claude Fable 5 contributed to resolving the Jacobian conjecture. Around the same period, Google&#8217;s Gemini resolved several open Erd\u0151s problems in combinatorics, and OpenAI&#8217;s GPT disproved a longstanding discrete geometry conjecture. The Riemann zeta result fits a pattern: frontier AI models finding non-obvious connections between existing bodies of work across subfields, rather than inventing new mathematical concepts from scratch.<\/p>\n<p>As one analysis of the methodology observed, models are not replacing mathematicians; they are extending the reach of existing human results by finding combinations nobody assembled. That framing \u2014 extension rather than replacement \u2014 fits the Bombieri plus Baluyot architecture precisely. Claude did not invent Weil&#8217;s quadratic form, Bombieri&#8217;s 2000 analysis, or the unconditional Montgomery machinery. It read all of them and noticed they fit together in a way that yielded a new bound.<\/p>\n<p>Whether that capability holds as problems grow harder is unanswered. The Riemann hypothesis itself, and the other six Millennium Prize Problems, remain open. But for the lower-bound question specifically, the record now stands at 67.2% \u2014 set in 36 hours, with a <a href=\"https:\/\/github.com\/anthropics\/zeta-23-lean\" rel=\"nofollow noopener\" target=\"_blank\">public Lean proof<\/a>, by an AI whose weights remain out of reach.<\/p>\n<p>Frequently Asked QuestionsDid Claude prove the Riemann hypothesis?<\/p>\n<p>No. Claude raised the proven lower bound on the fraction of Riemann zeta function zeros lying on the critical line \u2014 from 41.6% to 67.2%. This is a verifiable mathematical advance on a related, more tractable question. The Riemann hypothesis itself \u2014 which predicts that all nontrivial zeros lie on the critical line \u2014 remains unproven and undisproven, as it has since 1859. The <a href=\"https:\/\/www.claymath.org\/millennium-problems\/\" rel=\"nofollow noopener\" target=\"_blank\">Clay Mathematics Institute&#8217;s $1 million prize<\/a> for a complete proof has not been claimed.<\/p>\n<p>What exactly is the Lean 4 formalization, and does it prove the result is correct?<\/p>\n<p><a href=\"https:\/\/en.wikipedia.org\/wiki\/Lean_(proof_assistant)\" rel=\"nofollow noopener\" target=\"_blank\">Lean 4 proof assistant<\/a> is an open-source system that mechanically checks whether a system that mechanically checks whether a mathematical argument&#8217;s logical steps are valid, using the same type-checking that verifies computer programs. Claude&#8217;s Lean formalization of the Riemann zeta result is publicly available on GitHub and passes the standard validation tool comparator. This means the logical derivation can be verified by any researcher with access to the proof assistant \u2014 the result does not require trusting Claude&#8217;s judgment or Anthropic&#8217;s description of the proof. However, as researchers in AI-assisted formal verification note, machine type-checking confirms logical correctness; it does not guarantee that the formalized statement faithfully captures the intended mathematical content. That alignment judgment still requires human expert review, which Conrey, Goldston, Alp\u00f6ge, and Furman provided.<\/p>\n<p>Can this multi-agent approach be applied to other open problems?<\/p>\n<p>That is the live question the result leaves open. The template \u2014 many autonomous subagents, numerical verification at scale, Lean formalization as a mechanical audit layer, named human experts closing the validation \u2014 is in principle generalizable. Anthropic used a similar pattern in the Jacobian conjecture work earlier in 2026, and the same subagent-plus-Lean recipe also underpinned Claude Mythos Preview&#8217;s cryptanalysis results in July. The honest answer is that no one yet knows which open problems are accessible to this approach, or how reliably it produces results of comparable quality across different domains. What is now established is that it worked once on a benchmark that human researchers had barely moved in 46 years. For a broader read on <a href=\"https:\/\/aiweekly.co\/alerts\/anthropic-unreleased-claude-improves-zeta-bound-to-672\" rel=\"nofollow noopener\" target=\"_blank\">this agentic research pattern<\/a>, AI Weekly covers the template&#8217;s implications in detail.<\/p>\n<p>Is the result independently reproducible?<\/p>\n<p>The mathematics is independently auditable \u2014 via the public Lean proof and the released paper. The AI run itself is not reproducible: the model used is an unreleased research version of Claude, with no public weights, no checkpoint, and no API access. Researchers can verify that the proof is correct; they cannot verify that an AI system will produce a comparable result on a different problem, because the specific model is not available to them. This auditability-reproducibility asymmetry is a novel situation for the mathematics and AI research communities alike, and it is one the community will need to develop norms around as AI-assisted mathematical research becomes more common. One independent review covers <a href=\"https:\/\/kingy.ai\/blog\/claude-riemann-hypothesis-67-percent-result\/\" rel=\"nofollow noopener\" target=\"_blank\">reproducibility concerns remain<\/a> in detail.<\/p>\n","protected":false},"excerpt":{"rendered":"Claymath.org An unreleased research version of Claude spent 36 hours autonomously coordinating roughly 60 AI subagents \u2014 generating&hellip;\n","protected":false},"author":2,"featured_media":137682,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[8],"tags":[179,28028,53,3154,182,68088,68089,67329,68090],"class_list":["post-137681","post","type-post","status-publish","format-standard","has-post-thumbnail","category-anthropic","tag-agentic-ai","tag-ai-mathematical-proof","tag-anthropic","tag-anthropic-claude","tag-claude","tag-claude-ai-riemann-hypothesis","tag-lean-4-formal-proof","tag-riemann-hypothesis","tag-riemann-zeta-lower-bound"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/137681","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=137681"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/137681\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/137682"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=137681"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=137681"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=137681"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}