{"id":12432,"date":"2026-04-01T19:34:48","date_gmt":"2026-04-01T19:34:48","guid":{"rendered":"https:\/\/www.europesays.com\/news\/12432\/"},"modified":"2026-04-01T19:34:48","modified_gmt":"2026-04-01T19:34:48","slug":"hallucinated-citations-are-polluting-the-scientific-literature-what-can-be-done","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/news\/12432\/","title":{"rendered":"Hallucinated citations are polluting the scientific literature. What can be done?"},"content":{"rendered":"\n<p>Earlier this year, computer scientist Guillaume Cabanac received a notification from Google Scholar that one of his publications had been cited in a paper published in the International Dental Journal<a href=\"#ref-CR1\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">1<\/a>. That was unexpected, because <a href=\"https:\/\/www.nature.com\/articles\/d41586-021-02134-0\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-021-02134-0\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">his research on spotting fabricated paper<\/a>s doesn\u2019t typically intersect with dentistry. \u201cI was very surprised to see that I couldn\u2019t recognize my own reference,\u201d says Cabanac, who is based at the University of Toulouse in France.<\/p>\n<p>The title in the citation resembled that of a preprint<a href=\"#ref-CR2\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">2<\/a> he had posted in 2021 and never published formally, but the journal was listed as Nature and the DOI \u2014 the unique identifier assigned by publishers and preprint repositories \u2014 did not lead to the original preprint. \u201cI got very concerned,\u201d adds Cabanac, who immediately suspected that the citation had been hallucinated by artificial intelligence.<\/p>\n<p>This is just one example of a rapidly growing problem. Surveys and related studies have shown that<a href=\"https:\/\/www.nature.com\/articles\/d41586-025-00343-5\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-025-00343-5\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\"> researchers are increasingly using large language models<\/a> (LLMs) to help to conduct literature searches, write manuscripts and format bibliographies. And sometimes, these models <a href=\"https:\/\/www.nature.com\/articles\/d41586-025-02853-8\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-025-02853-8\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">generate non-existent academic references<\/a>.<\/p>\n<p><a href=\"https:\/\/www.nature.com\/articles\/d41586-023-03817-6\" class=\"u-link-inherit\" data-track=\"click\" data-track-label=\"recommended article\" rel=\"nofollow noopener\" target=\"_blank\"><img decoding=\"async\" class=\"recommended__image\" alt=\"\" src=\"https:\/\/www.europesays.com\/news\/wp-content\/uploads\/2026\/04\/d41586-026-00969-z_26512702.jpg\"\/><\/p>\n<p class=\"recommended__title u-serif\">Is AI leading to a reproducibility crisis in science?<\/p>\n<p><\/a><\/p>\n<p>Over the past year, efforts have begun turning up such hallucinated citations in the literature. One analysis of nearly 18,000 papers accepted by three <a href=\"https:\/\/www.nature.com\/articles\/d41586-025-03967-9\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-025-03967-9\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">computer-science<\/a> conferences found a sharp increase in references that cannot be traced to actual scholarly publications<a href=\"#ref-CR3\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">3<\/a>. The results, reported in January, indicated that 2.6% of papers in 2025 had a least one potentially hallucinated citation \u2014 up from about 0.3% in 2024. Another analysis, released in February, estimated that 2\u20136% of papers in four other 2025 computer-science conferences included references with rephrased titles or citations of publications that the authors couldn\u2019t verify by searching through databases and journal archives<a href=\"#ref-CR4\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">4<\/a>.<\/p>\n<p>And although the scale of the problem remains uncertain, it\u2019s clear that not only conferences are affected. An exclusive analysis conducted by Nature\u2019s news team, in collaboration with Grounded AI, a company based in Stevenage, UK, suggests that at least tens of thousands of 2025 publications, including journal papers and books, as well as conference proceedings, probably contain invalid references generated by AI.<\/p>\n<p>Grounded AI is among the companies offering publishers tools for screening submissions for problematic references. Several publishers told Nature reporters that they have been exploring such tools or developing in-house versions.<\/p>\n<p>But some researchers are concerned that the problem will soon get out of hand. \u201cWe\u2019re going to see a flood of fake references,\u201d says Alison Johnston, a political scientist at Oregon State University in Corvallis.<\/p>\n<p>Another issue is deciding what to do about hallucinated citations that make it into the published literature. That\u2019s a problem that academic publishers are wrestling with right now.<\/p>\n<p>Sources of error<\/p>\n<p>Citation errors are not new to academic publishing. \u201cEven before generative AI, we already had so many inaccuracies in citations,\u201d says Mohammad Hosseini, who studies research ethics and integrity at Northwestern University Feinberg School of Medicine in Chicago, Illinois. Issues have tended to include misspelling of authors\u2019 names or errors in the year of publication, the title of the journal or the DOI. Another issue has been discrepancies between the information in the cited work and the details given by the paper citing it<a href=\"#ref-CR5\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">5<\/a>,<a href=\"#ref-CR6\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">6<\/a>.<\/p>\n<p>\u201cNow the problem is not just inaccuracy, it\u2019s about fake citations. It\u2019s about fabricated citations, which is a whole different problem,\u201d says Hosseini.<\/p>\n<p>Publishers told Nature that they are seeing increases in the number of fabricated and inaccurate citations in submissions, and they are taking steps to tackle the issue.<\/p>\n<p>Johnston, co-lead editor of the Review of International Political Economy (RIPE), a journal published by the UK-based Taylor &amp; Francis, says that she rejected 25% of some 100 submissions in January \u201cbecause of fake references\u201d. She uses the plagiarism-detection software iThenticate to flag unusual or partial matches between the references in submitted papers and published bibliographies. Then she manually checks the suspicious citations. \u201cI\u2019m doing things now to try and detect hallucinated references that I wasn\u2019t doing prior to 2025,\u201d she says.<\/p>\n<p><a href=\"https:\/\/www.nature.com\/articles\/d41586-025-01512-2\" class=\"u-link-inherit\" data-track=\"click\" data-track-label=\"recommended article\" rel=\"nofollow noopener\" target=\"_blank\"><img decoding=\"async\" class=\"recommended__image\" alt=\"\" src=\"https:\/\/www.europesays.com\/news\/wp-content\/uploads\/2026\/04\/d41586-026-00969-z_50971464.jpg\"\/><\/p>\n<p class=\"recommended__title u-serif\">Take Nature\u2019s AI research test: find out how your ethics compare<\/p>\n<p><\/a><\/p>\n<p>Frontiers, based in Lausanne, Switzerland, has developed an in-house AI tool for flagging integrity issues at the point of submission, including references to irrelevant or retracted work and hallucinated citations. \u201cAround 5% [of manuscripts] show potential reference-related issues flagged through our checks,\u201d says Elena Vicario, Frontiers\u2019 head of research integrity. But \u201cnot all flagged references ultimately turn out to be genuinely problematic\u201d, she adds. That makes it challenging, Vicario says, to come up with a precise measure of the prevalence of any of these types of citation issue.<\/p>\n<p>Experiments using AI chatbots to generate papers have provided insights into how often LLMs produce citation errors and what types of error they tend to make. In one study, researchers prompted OpenAI\u2019s GPT-4o LLM to generate six literature reviews on three mental-health disorders, and analysed the 176 references in those synthetic reviews<a href=\"#ref-CR7\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">7<\/a>. Under these experimental conditions, they found that nearly 20% were fabricated references and could not be linked to actual research. And 45% of the remaining references, which corresponded to genuine publications, contained errors, often incorrect or invalid DOIs.<\/p>\n<p>In some cases, including in references in published articles, all of the component parts are made up, says Kathryn Weber-Boer, director of scientometrics at the London-based company Digital Science. (The firm is operated by the Holtzbrinck Publishing Group, which is the majority shareholder of Springer Nature, which publishes Nature. Nature\u2019s news team is editorially independent of its publisher.) AI also hallucinates DOIs, both in references that are otherwise genuine as well as in fabricated ones, she adds.<\/p>\n<p>AI-generated references commonly combine fragments of genuine publications, say researchers who have studied the issue (see \u2018How fakes can look real\u2019). Joe Shockman, co-founder and chief executive of Grounded AI, calls such references \u2018Frankenstein\u2019 citations, likening their assembly to that of the fictional monster. \u201cIt looks real to a human being, but is not actually a reference to a real thing,\u201d says Shockman, who is based in Ashland, Oregon.<\/p>\n<p><img decoding=\"async\" class=\"figure__image\" alt=\"How fakes can look real. A breakdown of a fake citation generated by AI showing how different elements can appear plausible.\" loading=\"lazy\" src=\"https:\/\/www.europesays.com\/news\/wp-content\/uploads\/2026\/04\/d41586-026-00969-z_52221360.jpg\"\/><\/p>\n<p class=\"figure__caption u-sans-serif\">Source: Ref. 7<\/p>\n<p>Although some types of error seem to implicate AI, others are less clear-cut, say researchers. \u201cIn today\u2019s landscape, we have to recognize that there are human errors and there are machine errors, and those can often overlap,\u201d says Weber-Boer.<\/p>\n<p>Published problems<\/p>\n<p>How many hallucinated citations are showing up in published research remains difficult to discern. To get an estimate, Nature\u2019s news team joined forces with Grounded AI, which has developed an AI tool called Veracity that checks citations against scholarly databases and across the web, flagging ones that are invalid, irrelevant or <a href=\"https:\/\/www.nature.com\/articles\/d41586-024-02719-5\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-024-02719-5\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">cite retracted work<\/a>.<\/p>\n<p>Nature and Grounded AI collaborated to analyse more than 4,000 publications from last year, covering five leading publishers: Elsevier, Sage, Springer Nature, Taylor &amp; Francis and Wiley. Grounded AI randomly sampled these papers from Europe PMC \u2014 a repository of open-access biomedical research articles \u2014 and the bibliometric database Crossref, to include equal number of publications per month from each of the five publishers. The sample included published papers as well as book chapters and conference proceedings, and it cut across all subject areas in these publishers\u2019 portfolios.<\/p>\n<p>Grounded AI\u2019s tool looks for an exact match to a reference or the closest match it can find. It then flags citations with major issues, such as mismatched titles or DOIs, missing authors and incorrect journals, as well as more-minor issues. Citations that pointed to papers that couldn\u2019t be found even though they should be easy to find \u2014 because the journal in question is indexed by scholarly databases, for example \u2014 were marked as especially problematic.<\/p>\n<p>After running the publications through the tool, Grounded AI assigned a risk score to each of the published papers, on the basis of the number of references that had major issues and how likely those issues were to have been generated by AI. Grounded AI determined that likelihood using data gleaned from a separate analysis that used two AI models to generate 20,000 synthetic papers; this allowed the company to identify the most common types of citation error that AI makes.<\/p>\n<p>Nature manually checked the 100 most suspicious publications and confirmed that 65 contained at least one invalid reference, meaning that it pointed to a publication that did not seem to exist (see \u2018Finding the fabrications\u2019). But 22 of the 100 most-suspicious papers had references that did point to genuine publications.<\/p>\n<p><img decoding=\"async\" class=\"figure__image\" alt=\"Finding the fabrications. A flowchart showing the analysis process.\" loading=\"lazy\" src=\"https:\/\/www.europesays.com\/news\/wp-content\/uploads\/2026\/04\/d41586-026-00969-z_52221356.jpg\"\/><\/p>\n<p>For the remaining 13 papers, it was unclear whether all their citations pointed to existing research or not. These 13 papers included references to articles that were said to be published in regional journals <a href=\"https:\/\/www.nature.com\/articles\/d41586-026-00229-0\" data-track=\"click\" data-label=\"https:\/\/www.nature.com\/articles\/d41586-026-00229-0\" data-track-category=\"body text link\" rel=\"nofollow noopener\" target=\"_blank\">in languages other than English<\/a>, and references that had mismatches in metadata that looked like plausible human errors, for example.<\/p>\n<p>The analysis, which looked at reference lists from Crossref and full text from Europe PMC publications, turned up no clear trend across publishers. Each of the selected publishers had more than five publications with references that manual checks couldn\u2019t validate.<\/p>\n<p>As a rough estimate, if the rate of 65 publications with at least one invalid reference out of some 4,000 publications analysed holds across the academic literature, it would suggest that more than 110,000 of the 7 million or so scholarly publications from 2025 contain invalid references.<\/p>\n<p>Nick Morley, Grounded AI\u2019s co-founder and chief product officer, says that the types of citation problem seen in 2025 are different from those found by his team before the proliferation of LLMs. This fact, he says, points to the use of AI as a leading culprit.<\/p>\n<p>The true number of hallucinated references is almost certainly higher, says Weber-Boer, because the analysis focused on big publishers, which have more resources for checking citations systematically than do smaller publishers. Fields such as computer science, which has seen a surge in the use of LLMs to produce manuscripts<a href=\"#ref-CR8\" data-track=\"click\" data-action=\"anchor-link\" data-track-label=\"go to reference\" data-track-category=\"references\">8<\/a>, might be more affected than other fields. What\u2019s more, the Grounded AI analysis turned up a few hundred more publications that had some risk of hallucinated citations, suggesting that extra manual checking would have brought more such citations to light.<\/p>\n<p>Spokespeople for all five publishers said that they check references as part of their screening and editing process, and they intend to investigate the publications flagged by the Nature analysis. A spokesperson for Taylor &amp; Francis said that some of the publications flagged were already under investigation by its ethics and integrity team.<\/p>\n<p><a href=\"https:\/\/www.nature.com\/articles\/d41586-025-01826-1\" class=\"u-link-inherit\" data-track=\"click\" data-track-label=\"recommended article\" rel=\"nofollow noopener\" target=\"_blank\"><img decoding=\"async\" class=\"recommended__image\" alt=\"\" src=\"https:\/\/www.europesays.com\/news\/wp-content\/uploads\/2026\/04\/d41586-026-00969-z_51086026.jpg\"\/><\/p>\n<p class=\"recommended__title u-serif\">How to spot suspicious papers: a sleuthing guide for scientists<\/p>\n<p><\/a><\/p>\n<p>When it comes to hallucinated references, \u201cThere have been cases where authors have been able to clearly document where issues have occurred in the process of producing a manuscript, for example using a translation tool, and demonstrate that the rest of the paper can be relied upon, in which case the paper will be corrected,\u201d says Chris Graf, Springer Nature\u2019s research-integrity director. But, more often, these references reflect broader problems with the content, he says.<\/p>\n<p>Shockman says that the number of potentially problematic citations flagged by Veracity is an order of magnitude greater when it is used in pilot programmes to screen submissions on behalf of publishers than when it analyses publications. This suggests that publishers are catching a large proportion of such citations before they can make it into the literature.<\/p>\n<p>Nature\u2019s collaboration with Grounded AI also highlighted, as many experts have noted, that the detection of invalid citations with automated tools is not error-free. One of the challenges is that journals have various ways of formatting references, and AI tools might fail to recognize references because of how they are styled. These types of problem showed up among citations that manual checks determined to be genuine despite having been flagged by Grounded AI.<\/p>\n<p>Another issue, says Weber-Boer, is that large-scale bibliometric databases might not index references that can\u2019t be verified, meaning their metadata might not match what appears on the publishers\u2019 websites. Some references do not contain their corresponding DOI, which makes it hard for automated tools to identify the cited paper, adds Weber-Boer. \u201cWe\u2019re starting to get a handle on the characteristics of this problem, which are a precursor to understanding the scale of it,\u201d she says.<\/p>\n<p>The Grounded AI team members acknowledge that not all the references their tool flags will be true positives, but they say they are continuing to improve its performance. IOP Publishing, based in Bristol, UK, is now using Grounded AI\u2019s tool to screen submissions for problematic citations across all of its proprietary journals, says Kim Eggleton, head of peer review and research integrity. \u201cWe know it\u2019s a problem, we just don\u2019t know how big the problem is,\u201d she says.<\/p>\n<p>Fake-citation fallout<\/p>\n","protected":false},"excerpt":{"rendered":"Earlier this year, computer scientist Guillaume Cabanac received a notification from Google Scholar that one of his publications&hellip;\n","protected":false},"author":2,"featured_media":12433,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":"","_share_on_mastodon":"0"},"categories":[4],"tags":[8497,8,1372,8498,1373,9,8499,1371,8500,7],"class_list":["post-12432","post","type-post","status-publish","format-standard","has-post-thumbnail","category-top-stories","tag-ethics","tag-headlines","tag-humanities-and-social-sciences","tag-machine-learning","tag-multidisciplinary","tag-news","tag-publishing","tag-science","tag-scientific-community","tag-top-stories"],"share_on_mastodon":{"url":"https:\/\/pubeurope.com\/@news\/116331133182488151","error":""},"_links":{"self":[{"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/posts\/12432","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/comments?post=12432"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/posts\/12432\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/media\/12433"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/media?parent=12432"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/categories?post=12432"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/news\/wp-json\/wp\/v2\/tags?post=12432"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}