{"id":109427,"date":"2026-07-17T11:11:22","date_gmt":"2026-07-17T11:11:22","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/109427\/"},"modified":"2026-07-17T11:11:22","modified_gmt":"2026-07-17T11:11:22","slug":"ai-made-up-a-science-term-now-its-in-dozens-of-papers","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/109427\/","title":{"rendered":"AI Made Up a Science Term \u2014 Now It\u2019s in Dozens of Papers"},"content":{"rendered":"<p><a href=\"https:\/\/cdn.zmescience.com\/wp-content\/uploads\/2025\/04\/andandand0017_A_fossilized_word_embedded_in_a_stone_tablet_ma_a25c50e9-9f4b-4ea1-9a08-20887a1f289a_2.png\" rel=\"nofollow noopener\" target=\"_blank\"><img fetchpriority=\"high\" decoding=\"async\" width=\"1024\" height=\"574\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/07\/andandand0017_A_fossilized_word_embedded_in_a_stone_tablet_ma_a25c50e9-9f4b-4ea1-9a08-20887a1f289a_2.png\" alt=\"AI-generated image of a digital fossil\" class=\"wp-image-281889\"  \/><\/a>AI \u201cdigital fossils\u201d are already polluting the world. They seem authentic and complex, but in fact, they don\u2019t make sense. Kind of like this AI-generated image.<\/p>\n<p class=\"wp-block-paragraph\">In 2023 and 2024, as AI text generators started to become mainstream, a curious trend emerged: the word \u201cdelve\u201d began appearing in a suspicious number of science papers. It became a kind of calling card for AI-generated content \u2014 but it\u2019s far from the weirdest one.<\/p>\n<p class=\"wp-block-paragraph\">Let us introduce you to: \u201cvegetative electron microscopy.\u201d<\/p>\n<p>Vegetative what?<\/p>\n<p class=\"wp-block-paragraph\">If you know basic science, you\u2019re already raising an eyebrow. \u201cVegetative electron microscopy\u201d doesn\u2019t make sense \u2014 and that\u2019s because it isn\u2019t a real thing. It\u2019s what researchers call a \u201cdigital fossil\u201d \u2014 a strange, erroneous term born from a mix of optical scanning errors and an apparent translation mistake, then preserved in the data used to train artificial intelligence.<\/p>\n<p class=\"wp-block-paragraph\">Remarkably, this nonsense phrase appears to have emerged independently through two unrelated errors.<\/p>\n<p class=\"wp-block-paragraph\">Back in the 1950s, <a href=\"https:\/\/journals.asm.org\/doi\/10.1128\/br.20.4.207-242.1956\" rel=\"nofollow noopener\" target=\"_blank\">two papers<\/a> in the journal Bacteriological Reviews were scanned and digitized. In one of them, the word \u201cvegetative\u201d appeared in one column and \u201celectron microscopy\u201d in the adjacent one. The OCR software mistakenly merged the two \u2014 and so, the fossil was born.<\/p>\n<p><a href=\"https:\/\/cdn.zmescience.com\/wp-content\/uploads\/2025\/04\/file-20250414-56-7m3h0n.avif\" rel=\"nofollow noopener\" target=\"_blank\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"212\" alt=\"\" class=\"wp-image-281890 perfmatters-lazy\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/07\/file-20250414-56-7m3h0n-1024x212.jpg\"  data-\/><\/a><\/p>\n<p class=\"wp-block-paragraph\">Then, in  <a href=\"https:\/\/scholar.google.com\/scholar?hl=en&amp;as_sdt=0%2C5&amp;q=Production+of+mesoporous+activated+carbon+from+cone+of+Iranian+pine+tree+(Pinus+eldarica)+using+chemical+activation+for+adsorption+of+sodium+dodecylbenzene+sulfonate+from+aqueous+solution&amp;btnG=\" rel=\"nofollow noopener\" target=\"_blank\">2017<\/a> and <a href=\"https:\/\/web.p.ebscohost.com\/abstract?site=ehost&amp;scope=site&amp;jrnl=20085729&amp;AN=141678734&amp;h=e9Z0lqUsvh1WBhQvCayQkWtMqGcULLWTPrWyrZbI%2bQdCrwycHUHwP0UFo7hX3eLpPU1VEhqXgz4QHsTCrtBAFw%3d%3d&amp;crl=c&amp;resultLocal=ErrCrlNoResults&amp;resultNs=Ehost&amp;crlhashurl=login.aspx%3fdirect%3dtrue%26profile%3dehost%26scope%3dsite%26authtype%3dcrawler%26jrnl%3d20085729%26AN%3d141678734\" rel=\"nofollow noopener\" target=\"_blank\">2019<\/a>, two papers used the term again. Here, this appears to be a <a href=\"https:\/\/retractionwatch.com\/2025\/03\/04\/vegetative-electron-microscopy-phrase-farsi-typo\/\" rel=\"nofollow noopener\" target=\"_blank\">translation error<\/a>. In Farsi, the words for \u201cvegetative\u201d and \u201cscanning\u201d differ by only a single dot. So instead of scanning electron microscopy, you got vegetative electron microscopy.<\/p>\n<p><a href=\"https:\/\/cdn.zmescience.com\/wp-content\/uploads\/2025\/04\/file-20250414-56-pa8vsp.avif\" rel=\"nofollow noopener\" target=\"_blank\"><img loading=\"lazy\" decoding=\"async\" width=\"1024\" height=\"549\" alt=\"\" class=\"wp-image-281891 perfmatters-lazy\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/07\/file-20250414-56-pa8vsp-1024x549.jpg\"  data-\/><\/a><\/p>\n<p class=\"wp-block-paragraph\">All of this came to light thanks to a detailed investigation by Retraction Watch in February. But this wasn\u2019t the end of the story.<\/p>\n<p>Why this matters<\/p>\n<p class=\"wp-block-paragraph\">You\u2019d think this weird glitch wouldn\u2019t matter \u2014 but it turns out, it kind of does.<\/p>\n<p>\u00d7<\/p>\n<p>                        Thank you! One more thing&#8230;<\/p>\n<p>Please check your inbox and confirm your subscription.<\/p>\n<p class=\"wp-block-paragraph\">The term has now appeared in at least <a href=\"https:\/\/scholar.google.com\/scholar?hl=en&amp;as_sdt=0%2C5&amp;q=%22vegetative+electron%22&amp;btnG=\" rel=\"nofollow noopener\" target=\"_blank\">22 different papers<\/a>. Some have been corrected or retracted, but by then, the damage was done. Even El Pa\u00eds, one of Spain\u2019s leading newspapers, <a href=\"https:\/\/english.elpais.com\/science-tech\/2023-04-02\/one-of-the-worlds-most-cited-scientists-rafael-luque-suspended-without-pay-for-13-years.html\" rel=\"nofollow noopener\" target=\"_blank\">quoted it in a story<\/a> in 2023.<\/p>\n<p class=\"wp-block-paragraph\">Why? Blame AI.<\/p>\n<p class=\"wp-block-paragraph\">Modern AI systems are trained on vast troves of data \u2014 <a href=\"https:\/\/www.scientificamerican.com\/article\/your-personal-information-is-probably-being-used-to-train-generative-ai-models\/\" rel=\"nofollow noopener\" target=\"_blank\">essentially everything<\/a> they can scrape. Once \u201cvegetative electron microscopy\u201d appeared in several published sources, the AI models treated it like a legitimate term. So when researchers asked these systems to help write or draft papers, the models sometimes spat it out, blissfully unaware that it was gibberish.<\/p>\n<p class=\"wp-block-paragraph\">According to Aaron J. Snoswell and colleagues, who published a deep dive on <a href=\"https:\/\/theconversation.com\/a-weird-phrase-is-plaguing-scientific-papers-and-we-traced-it-back-to-a-glitch-in-ai-training-data-254463\" rel=\"nofollow noopener\" target=\"_blank\">The Conversation<\/a>, the term began polluting the AI knowledge pool after 2020 \u2014 after those two problematic Farsi translations. And it\u2019s not just a one-time fluke: the error persists in large models like GPT-4o and Claude 3.5.<\/p>\n<p class=\"wp-block-paragraph\">\u201cWe also found the error persists in later models including GPT-4o and Anthropic\u2019s Claude 3.5,\u201d the <a href=\"https:\/\/theconversation.com\/a-weird-phrase-is-plaguing-scientific-papers-and-we-traced-it-back-to-a-glitch-in-ai-training-data-254463\" rel=\"nofollow noopener\" target=\"_blank\">group write<\/a> in a post on The Conversation. \u201cThis suggests the nonsense term may now be permanently embedded in AI knowledge bases.\u201d<\/p>\n<p class=\"wp-block-paragraph\">\u201cPermanently\u201d may be too strong in a literal sense. AI developers can filter outputs, revise training sets and replace old models. But correcting an error buried inside billions or trillions of learned statistical relationships is much harder than fixing a line in a conventional database. Removing the original phrase from the internet wouldn\u2019t necessarily remove the association from models already trained on it.<\/p>\n<p>AI-assisted scientific writing is no longer a marginal phenomenon<\/p>\n<p class=\"wp-block-paragraph\">This bizarre example is more than a fun anecdote \u2014 it highlights real risks.<\/p>\n<p class=\"wp-block-paragraph\">Since the vegetative-electron-microscopy story emerged, researchers have begun measuring AI\u2019s influence across entire bodies of scientific literature. <\/p>\n<p class=\"wp-block-paragraph\">In 2025, <a href=\"https:\/\/www.science.org\/doi\/10.1126\/sciadv.adt3813\" rel=\"nofollow noopener\" target=\"_blank\">researchers analyzed<\/a> more than 15 million biomedical abstracts indexed in PubMed. They tracked vocabulary that became unusually common after the release of ChatGPT and estimated that at least 13.5 percent of abstracts published in 2024 had been processed with a large language model. Rates varied widely between fields, journals and countries, sometimes reaching about 40 percent.<\/p>\n<p class=\"wp-block-paragraph\">Another study, published in Nature Human Behaviour, <a href=\"https:\/\/nlp.stanford.edu\/~manning\/papers\/Liang_et_al-2025-Nature_Human_Behaviour.pdf\" rel=\"nofollow noopener\" target=\"_blank\">examined more than<\/a> 1.1 million papers and preprints. By September 2024, the researchers estimated that large language models had modified 22.5 percent of computer-science abstract sentences and 19.6 percent of introduction sentences. The estimated rates were lower in mathematics and Nature Portfolio journals but had still risen sharply after ChatGPT\u2019s release.<\/p>\n<p class=\"wp-block-paragraph\">In 2026, an analysis of 7.3 million articles published between 2020 and 2025 found evidence of some degree of LLM influence in slightly <a href=\"https:\/\/www.pnas.org\/doi\/10.1073\/pnas.2605754123\" rel=\"nofollow noopener\" target=\"_blank\">more than half of the papers from 2025<\/a>. That did not mean that half of the papers were wholly written by chatbots. The category included everything from light linguistic editing to much more extensive generation.<\/p>\n<p class=\"wp-block-paragraph\">Using AI to improve grammar or phrasing isn\u2019t scientific misconduct. For many researchers, especially those writing in a second language, these tools can make papers clearer and more accessible. The problem arises when authors use them without checking the output, allow them to invent references or scientific claims, or conceal the extent to which a machine produced the paper\u2019s reasoning.<\/p>\n<p>An AI Arms Race<\/p>\n<p class=\"wp-block-paragraph\">Researchers are trying to fight this and detect this sort of issue. The <a href=\"https:\/\/dbrech.irit.fr\/pls\/apex\/f?p=9999:11::::::\" rel=\"nofollow noopener\" target=\"_blank\">Problematic Paper Screener<\/a>, for instance, is an automated tool that <a href=\"https:\/\/theconversation.com\/problematic-paper-screener-trawling-for-fraud-in-the-scientific-literature-246317\" rel=\"nofollow noopener\" target=\"_blank\">combs through 130 million articles<\/a> every week. It uses nine detectors searching for new instances of known fingerprints or improper use of AI. They found <a href=\"https:\/\/pubpeer.com\/publications\/7E79C63A6C566C67D6B46715F8B906#2\" rel=\"nofollow noopener\" target=\"_blank\">78 papers<\/a> in Springer Nature\u2019s Environmental Science and Pollution Research alone.<\/p>\n<p class=\"wp-block-paragraph\">But it\u2019s an uphill battle.<\/p>\n<p class=\"wp-block-paragraph\">There\u2019s already so much AI content everywhere that it\u2019s almost becoming virtually impossible to detect it; and that\u2019s just one part of the problem. Scientific journals are another problem.<\/p>\n<p class=\"wp-block-paragraph\">Journals have every incentive to protect their reputation and avoid retractions, even if it means defending dubious content. Case in point: <a href=\"https:\/\/retractionwatch.com\/2025\/02\/10\/vegetative-electron-microscopy-fingerprint-paper-mill\/\" rel=\"nofollow noopener\" target=\"_blank\">Elsevier initially tried to justify<\/a> the use of \u201cvegetative electron microscopy\u201d before ultimately issuing a correction. They ultimately issued a correction but the response is telling.<\/p>\n<p class=\"wp-block-paragraph\">The problem is that as long as tech companies aren\u2019t transparent about their training data and methods, researchers have to play detective and look for AI needles in the publishing haystack. According <a href=\"https:\/\/news.exeter.ac.uk\/faculty-of-environment-science-and-economy\/avalanche-of-papers-could-erode-trust-in-science\/\" rel=\"nofollow noopener\" target=\"_blank\">to one estimate<\/a>, there are close to 3 million papers published a year, and the use of AI in writing is becoming more and more common.<\/p>\n<p class=\"wp-block-paragraph\">The real danger is that these kinds of accidental errors can become entrenched in our scientific record \u2014 and once embedded, AI systems will keep repeating them. Knowledge is incremental, and if we build on wrong foundations, the consequences can be severe.<\/p>\n<p class=\"wp-block-paragraph\">Ultimately, it seems even nonsense, once digitized and published, can become immortal.<\/p>\n","protected":false},"excerpt":{"rendered":"AI \u201cdigital fossils\u201d are already polluting the world. They seem authentic and complex, but in fact, they don\u2019t&hellip;\n","protected":false},"author":2,"featured_media":109428,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,2720,32051,9439,25,56408,56409,11206,56410,56411,56412,56413,32902],"class_list":["post-109427","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-ai-hallucinations","tag-ai-in-science","tag-ai-training-data","tag-artificial-intelligence","tag-claude-3-5","tag-digital-fossils","tag-gpt-4o","tag-knowledge-integrity","tag-ocr-errors","tag-paper-retractions","tag-retraction-watch","tag-scientific-publishing"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/109427","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=109427"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/109427\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/109428"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=109427"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=109427"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=109427"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}