{"id":132226,"date":"2026-08-06T22:50:16","date_gmt":"2026-08-06T22:50:16","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/132226\/"},"modified":"2026-08-06T22:50:16","modified_gmt":"2026-08-06T22:50:16","slug":"ai-agents-are-checking-the-scientific-literature-and-spotting-decades-old-errors","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/132226\/","title":{"rendered":"AI agents are checking the scientific literature \u2014 and spotting decades-old errors"},"content":{"rendered":"<p> <img decoding=\"async\" class=\"figure__image\" alt=\"A laboratory professional supervising a large glass reactor vessel connected to an extensive network of pipes, valves, sensors, and control equipment in a research environment.\" loading=\"lazy\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/d41586-026-02235-8_53042962.jpg\"\/><\/p>\n<p class=\"figure__caption u-sans-serif\">An AI fact-checking tool found errors in molecule boiling points listed in a chemistry reference database.Credit: Monty Rakusen\/Getty<\/p>\n<p>For decades, chemists have relied on handbook values for a molecule\u2019s boiling point to identify substances and plan processes such as distillation. But an artificial-intelligence model has revealed that some trusted numbers in one reference database have been wrong all along.<\/p>\n<p>Sebastian Pios, a theoretical chemist at Zhejiang Lab in Hangzhou, China, was using an AI system to predict the boiling points of several molecules when it began producing values that clashed with long-accepted entries in a 75-year-old reference database. At first, he thought the model was wrong. But when he manually checked the original literature, he found that the reference data were wrong, not the AI model.<\/p>\n<p>In two other cases, Pios\u2019s AI model spotted errors in older papers and reference books \u2014 mistakes that have made their way into the scientific canon. One was a typo in a paper; another was incorrect values of century-old boiling-point measurements. Both errors are likely to have caused researchers using the database \u201ca lot of trouble\u201d, says Pios.<\/p>\n<p>Pios is one of a growing number of scientists using AI as a tool for auditing scientific knowledge. As well as checking databases, researchers are using specialized <a href=\"https:\/\/www.nature.com\/articles\/d41586-025-00648-5\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-025-00648-5\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">AI tools to hunt for errors in papers published in journals and conferences<\/a>.<\/p>\n<p><a href=\"https:\/\/www.nature.com\/articles\/d41586-025-01839-w\" class=\"u-link-inherit\" data-track=\"click\" data-track-label=\"recommended article\" rel=\"nofollow noopener\" target=\"_blank\"><img decoding=\"async\" class=\"recommended__image\" alt=\"\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/d41586-026-02235-8_51208868.png\"\/><\/p>\n<p class=\"recommended__title u-serif\">AI, peer review and the human activity of science<\/p>\n<p><\/a><\/p>\n<p>In an analysis <a href=\"https:\/\/sai.science\/blog\/how-much-science-is-verifiable\" data-track=\"click\" data-label=\"https:\/\/sai.science\/blog\/how-much-science-is-verifiable\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">posted online on 22 July<\/a>, researchers at SAI Labs, a for-profit research-review company in Delaware, used AI agents to assess 168 papers selected for oral presentation at the 2026 International Conference on Machine Learning (ICML). The AI agents extracted the authors\u2019 central claims, downloaded accompanying resources, reran experiments where possible and compared the results with those reported by the authors.<\/p>\n<p>Of the 92 papers that had at least five claims available for assessment, the AI agents were able to reproduce more than two of the five claims from only 34 papers. The agents successfully repeated more than 80% of claims from just eight papers.<\/p>\n<p>But such AI fact-checking tools remain unreliable arbiters of the scientific corpus, says Odd Erik Gundersen, a computer scientist at the Norwegian University of Science and Technology in Trondheim. The tools \u201cmake mistakes like humans do\u201d, he says, which is why the quality of AI fact checkers must be manually processed with human oversight.<\/p>\n<p>AI fact checker<\/p>\n<p>Many researchers already dedicate their time to spotting errors in papers and use tools to check <a href=\"https:\/\/www.nature.com\/articles\/d41586-024-01247-6\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-024-01247-6\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">certain facets of papers<\/a>. But one advantage of using AI is the speed at which it can scan scientific databases and literature compared to humans, says James Zou, a computer scientist at Stanford University, California. \u201cThe biggest difference is to be able to do this at a scale that was not possible before.\u201d<\/p>\n<p>In a study posted on the preprint server arXiv<a href=\"#ref-CR1\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">1<\/a>, Zou and his colleagues used an \u2018AI checker\u2019 to scan papers published at NeurIPS \u2014 a prestigious annual AI research conference \u2014 for errors. Their tool found that errors in papers rose from 3.8 in 2021 to 5.9 in 2025 \u2014 an increase of 55%.<\/p>\n<p><a href=\"https:\/\/www.nature.com\/articles\/d41586-025-00648-5\" class=\"u-link-inherit\" data-track=\"click\" data-track-label=\"recommended article\" rel=\"nofollow noopener\" target=\"_blank\"><img decoding=\"async\" class=\"recommended__image\" alt=\"\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/d41586-026-02235-8_50837996.png\"\/><\/p>\n<p class=\"recommended__title u-serif\">AI tools are spotting errors in research papers: inside a growing movement<\/p>\n<p><\/a><\/p>\n<p>\u201cThese are papers that have been published, so they\u2019re sort of taken as the foundational knowledge for the next generation of research,\u201d says Zou. \u201cIf there are mistakes in these foundations, this can propagate and make the follow-on research shakier,\u201d he adds.<\/p>\n<p>The analysis focused on \u2018objective\u2019 errors, such as those in formulae, calculations and figures, and excluded subjective mistakes about data interpretation and novelty. Co-author Federico Bianchi, a machine-learning scientist at Together AI, based in San Francisco, California, says this was a design choice. \u201cAI should not do everything, and leave choices about novelty and significance to humans,\u201d he says.<\/p>\n","protected":false},"excerpt":{"rendered":"An AI fact-checking tool found errors in molecule boiling points listed in a chemistry reference database.Credit: Monty Rakusen\/Getty&hellip;\n","protected":false},"author":2,"featured_media":132227,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[6],"tags":[405,7537,8206,1743,50,1744,160],"class_list":["post-132226","post","type-post","status-publish","format-standard","has-post-thumbnail","category-agentic-ai","tag-ai-agents","tag-artificial-intelligence-agents","tag-databases","tag-humanities-and-social-sciences","tag-machine-learning","tag-multidisciplinary","tag-science"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/132226","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=132226"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/132226\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/132227"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=132226"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=132226"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=132226"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}