{"id":671007,"date":"2026-09-03T20:41:13","date_gmt":"2026-09-03T20:41:13","guid":{"rendered":"https:\/\/www.europesays.com\/ie\/671007\/"},"modified":"2026-09-03T20:41:13","modified_gmt":"2026-09-03T20:41:13","slug":"an-18-year-old-high-school-senior-from-pasadena-wrote-machine-learning-code-to-comb-200-billion-entries-of-raw-infrared-sky-data-sorted-what-came-back-and-in-2025-found-1-5-million-new-potential-obj","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ie\/671007\/","title":{"rendered":"An 18-year-old high school senior from Pasadena wrote machine-learning code to comb 200 billion entries of raw infrared sky data, sorted what came back, and in 2025 found 1.5 million new potential objects"},"content":{"rendered":"<p>Let\u2019s start with the number that makes the rest of the story hard to picture: 200 billion. That is roughly how many rows of data a machine-learning program built by a Pasadena high schooler had to work through. Each row is one moment a retired NASA telescope caught a flicker of infrared light somewhere in the sky. Stacked over more than a decade of scanning, they add up to a table nobody could read by hand.<\/p>\n<p>The person who wrote code to comb through it is Matteo Paz. When the flickers were sorted, his program flagged <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=His%20model%20revealed%201.5%20million%20previously%20unknown%20potential%20celestial%20bodies\" target=\"_blank\" rel=\"noopener nofollow\">about 1.5 million<\/a> potential new objects. In 2025 that work won him the top prize in a national science competition.<\/p>\n<p>We want to answer the obvious questions: what was the code actually doing, why \u201cpotential\u201d is the word that matters, and how a teenager ended up doing it inside one of Caltech\u2019s research centers.<\/p>\n<p>What is 200 billion entries of raw infrared data?<\/p>\n<p>The data came from NEOWISE, an infrared space telescope that <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/\" target=\"_blank\" rel=\"noopener nofollow\">scanned the entire sky<\/a> over and over for more than ten years before it was retired. Every pass added more detections. Paz\u2019s mentor, IPAC senior scientist Davy Kirkpatrick, put the scale plainly: the count was <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=200%20billion\" target=\"_blank\" rel=\"noopener nofollow\">creeping up towards 200 billion rows<\/a> in the table of every detection the survey had made.<\/p>\n<p>Raw is the key word. This was not a tidy catalog of stars and galaxies. It was individual snapshots, the kind of firehose most projects trim down before they even start, because the full stream is so unwieldy. The value hiding in it is change over time. If you watch the same patch of sky again and again, some points of light stay steady and some brighten and dim. The ones that vary are often the interesting ones. Kirkpatrick\u2019s original plan for the summer was modest: take <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=variable%20stars\" target=\"_blank\" rel=\"noopener nofollow\">a small piece of the sky and look for variable stars<\/a>.<\/p>\n<p>What was the code actually doing?<\/p>\n<p>Paz\u2019s program was built to catch tiny differences in infrared brightness across those repeated measurements. He described the method in a peer-reviewed <a href=\"https:\/\/arxiv.org\/abs\/2409.15499#:~:text=Fourier%20and%20Wavelet\" target=\"_blank\" rel=\"noopener nofollow\">paper<\/a> in The Astronomical Journal, which he wrote alone. It combines signal-processing math with a neural network that learned to tell one kind of flicker from another.<\/p>\n<p>The output was not just a pile of \u201cthis one changes.\u201d The program sorted sources into a small set of <a href=\"https:\/\/arxiv.org\/abs\/2409.15499#:~:text=classification\" target=\"_blank\" rel=\"noopener nofollow\">categories<\/a>, separating steady sources from the various ways an object\u2019s light can vary. Some things pulse on their own. Some dim on a schedule because a companion star passes in front of them. Some flare once and fade. Sorting candidates into groups is what turns a raw list into something a scientist can actually search, because a researcher hunting for eclipsing pairs of stars does not want to sift through every distant flickering galaxy to find them.<\/p>\n<p>What strikes us is the reach Paz claims for the approach beyond astronomy. He has said the model can be used <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=temporal\" target=\"_blank\" rel=\"noopener nofollow\">for other time domain studies<\/a> in astronomy, and potentially anything else that comes in a temporal format. That \u201cpotentially\u201d is his, and worth keeping. Anything that arrives as a stream of measurements over time is a candidate in principle. Whether the model actually proves useful on, say, financial data or sensor readings is an open question, not a proven result.<\/p>\n<p>1.5 million \u201cpotential\u201d is not 1.5 million discoveries<\/p>\n<p>The distinction the headline number can hide is this: the <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=1.5%20million%20previously%20unknown%20potential%20celestial%20bodies\" target=\"_blank\" rel=\"noopener nofollow\">1.5 million potential new objects<\/a> are candidates, not confirmed finds. Some may turn out to be sources already cataloged. Some will be false alarms, quirks of the data rather than real varying objects. The catalog is a list of things worth a closer look, and the looking is a separate job.<\/p>\n<p>This is not a knock on the work. It is how surveys of this kind function. A program built to trawl 200 billion measurements is valuable precisely because it narrows an impossible search down to a manageable set of leads. Among the candidates are objects whose brightness shifts over time, and the Society for Science\u2019s summary of Paz\u2019s <a href=\"https:\/\/www.societyforscience.org\/regeneron-sts\/2025-student-finalists\/matteo-paz\/#:~:text=including%20supermassive%20black%20holes%2C%20newborn%20stars%20and%20supernovae\" target=\"_blank\" rel=\"noopener nofollow\">census of infrared variable objects<\/a> lists supermassive black holes, newborn stars and supernovae among them. Confirming any single one takes follow-up observation and analysis. Treating the 1.5 million as finished discoveries would overstate what the catalog is, and understate why a filtered list of leads is genuinely useful.<\/p>\n<p>The paper describing the method was <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=published%20in%20November\" target=\"_blank\" rel=\"noopener nofollow\">published in November<\/a>, and Paz and Kirkpatrick have said they plan to release the full catalog so others can begin that follow-up work.<\/p>\n<p>How did a high schooler end up doing this?<\/p>\n<p>The short answer is a chain of outreach programs rather than a single lucky break. Paz\u2019s route ran through Caltech public lectures and a summer research program that paired him with Kirkpatrick in 2023. The project was meant to last six weeks. Paz had something larger in mind from the start, telling Kirkpatrick on the first day that he was <a href=\"https:\/\/www.smithsonianmag.com\/smart-news\/high-school-student-discovers-1-5-million-potential-new-astronomical-objects-by-developing-an-ai-algorithm-180986429\/#:~:text=paper\" target=\"_blank\" rel=\"noopener nofollow\">considering working toward a paper<\/a>, a far bigger goal than six weeks would normally allow. By his account, the mentor did not discourage him.<\/p>\n<p>What the mentorship bought, by Paz\u2019s account, was less technical instruction than room to think. It left space for the ambitious version of the project to survive, and the six weeks became the start of a much longer piece of work.<\/p>\n<p>What happens to a catalogue this size now?<\/p>\n<p>The prize, first place and its <a href=\"https:\/\/www.societyforscience.org\/press-release\/regeneron-sts-top-awards-2025\/\" target=\"_blank\" rel=\"noopener nofollow\">$250,000 award<\/a> in the 2025 Regeneron Science Talent Search, reads more like a starting line than a finish. Paz now works at IPAC as a <a href=\"https:\/\/www.caltech.edu\/about\/news\/exploring-space-with-AI?utm_source=chatgpt.com\" target=\"_blank\" rel=\"noopener nofollow\">Caltech employee<\/a>, and the catalog he built is still a set of leads waiting to be checked.<\/p>\n<p>Which of the 1.5 million candidates are real and new, which are already known, and which are just noise, are questions the release is meant to hand to the wider community. A candidate list is an invitation. What astronomers make of it, and what the method does when pointed at the next flood of survey data, is the work that has not happened yet.<\/p>\n<p class=\"bbm-disclaimer__heading\">About this article<\/p>\n<p class=\"bbm-disclaimer__body\">This article is for general information and reflection. It is not professional advice. For your specific situation, consult a qualified professional.<\/p>\n<p class=\"article__editorial-note-kicker\">Standards<\/p>\n<p class=\"article__editorial-note-text\">\n\t\t\t\tSpace Daily articles are edited and fact-checked before publication. We use AI tools in the newsroom. See our <a href=\"https:\/\/spacedaily.com\/editorial-policy\/\" rel=\"nofollow noopener\" target=\"_blank\">editorial standards<\/a> and <a href=\"https:\/\/spacedaily.com\/masthead\/\" rel=\"nofollow noopener\" target=\"_blank\">masthead<\/a>.\n\t\t\t<\/p>\n","protected":false},"excerpt":{"rendered":"Let\u2019s start with the number that makes the rest of the story hard to picture: 200 billion. That&hellip;\n","protected":false},"author":2,"featured_media":671008,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":"","_share_on_mastodon":"0"},"categories":[270],"tags":[18,19,17,133,451],"class_list":["post-671007","post","type-post","status-publish","format-standard","has-post-thumbnail","category-space","tag-eire","tag-ie","tag-ireland","tag-science","tag-space"],"share_on_mastodon":{"url":"https:\/\/pubeurope.com\/@ie\/117209050906104196","error":""},"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/posts\/671007","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/comments?post=671007"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/posts\/671007\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/media\/671008"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/media?parent=671007"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/categories?post=671007"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ie\/wp-json\/wp\/v2\/tags?post=671007"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}