{"id":110512,"date":"2026-07-18T13:01:08","date_gmt":"2026-07-18T13:01:08","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/110512\/"},"modified":"2026-07-18T13:01:08","modified_gmt":"2026-07-18T13:01:08","slug":"lin-qiaos-fireworks-bets-on-specialized-models-over-general-a-i-hype","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/110512\/","title":{"rendered":"Lin Qiao\u2019s Fireworks Bets on Specialized Models Over General A.I. Hype"},"content":{"rendered":"<p>\t\t<img fetchpriority=\"high\" decoding=\"async\" class=\"size-full-width wp-image-1682830\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/07\/GettyImages-2184020005.jpg\" alt=\"Lin Qiao, CEO &amp; Co Founder, Fireworks AI, on Centre stage during day three of Web Summit 2024 at the MEO Arena in Lisbon, Portugal. \" width=\"970\" height=\"647\"  \/>According to Lin Qiao, the future of A.I. lies in millions of specialized models built on proprietary data. Sam Barnes\/Sportsfile for Web Summit via Getty Images<\/p>\n<p>The soaring demand for A.I. has given rise to a new category of digital utility companies that sell compute power, access to models and developer infrastructure. Among <a href=\"https:\/\/observer.com\/2024\/12\/amd-ai-startup-investment\/\" data-lasso-id=\"2994067\" rel=\"nofollow noopener\" target=\"_blank\">the leaders of this pack<\/a> is <a href=\"https:\/\/observer.com\/company\/fireworks-ai\/\" title=\"Fireworks AI\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Fireworks AI<\/a>, co-founded by former <a href=\"https:\/\/observer.com\/company\/meta\/\" title=\"Meta\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Meta<\/a> executive <a href=\"https:\/\/observer.com\/person\/lin-qiao\/\" title=\"Lin Qiao\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Lin Qiao<\/a>, who led the creation of PyTorch, a popular open-source machine learning framework, and a team of engineers from Meta and <a href=\"https:\/\/observer.com\/company\/google\/\" title=\"Google\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Google<\/a>.<\/p>\n<p>Fireworks AI is a platform for developers to build products faster and at lower cost than with proprietary models, using open-source models. It has access to many of the most capable open-source models on the market, such as Meta\u2019s Llama series, Mistral, Qwen and <a href=\"https:\/\/observer.com\/company\/deepseek\/\" title=\"DeepSeek\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">DeepSeek<\/a>. It also allows enterprises to upload their own data to train and fine-tune these models. Its clients include <a href=\"https:\/\/observer.com\/company\/cursor\/\" title=\"Cursor\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Cursor<\/a>, <a href=\"https:\/\/observer.com\/company\/harvey-ai\/\" title=\"Harvey\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Harvey<\/a>, <a href=\"https:\/\/observer.com\/company\/uber\/\" title=\"Uber\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Uber<\/a> and <a href=\"https:\/\/observer.com\/company\/shopify\/\" title=\"Shopify\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Shopify<\/a>, among others.<\/p>\n<p>Lin describes Fireworks as \u201ca specialized intelligence platform,\u201d as opposed to general intelligence. Specialized intelligence was what A.I. researchers primarily relied on before general intelligence became viable. \u201cBefore generative A.I. was a thing, there was no foundation model holding world knowledge together. GenAI changed that,\u201d Lin explained to Observer. \u201cNow, foundation models learn from the public internet and massive labeled datasets, creating a deeper, more generalized knowledge base that you could use directly as a black-box API.\u201d<\/p>\n<p>But Lin believes that, amid the abundance of public data and general intelligence models, the most valuable uses of A.I. will, counterintuitively, come from specialization.<\/p>\n<p>\u201cBecause foundational models do not have access to the private data locked inside applications and enterprises,\u201d she said. \u201cThe majority of data is private, locked inside enterprises as proprietary IP and information that would never get shared outside the company.\u201d<\/p>\n<p>Training and fine-tuning models with that private data creates an ongoing need for Fireworks\u2019 services. \u201cThis is a continuous process because applications keep evolving, data distribution changes, and base models keep improving,\u201d Lin said. \u201cWe have customers tuning once a week, once a day, or even once every few hours.\u201d She predicted that this tuning process would soon be fully automated.<\/p>\n<p>Once a model is finely tuned, Fireworks helps optimize it for inference speed and cost. The company offers some of the fastest inference\u2014the speed at which an A.I. generates a response\u2014in the industry. For example, Cursor\u2019s code editor uses Fireworks\u2019s speculative decoding to deliver code suggestions up to 13 times faster than traditional setups.<\/p>\n<p>Fireworks processes more than 30 trillion tokens in daily inference traffic (excluding training), more than <a href=\"https:\/\/observer.com\/company\/openai\/\" title=\"OpenAI\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">OpenAI<\/a> and Google\u2019s Gemini, according to the latest published data.<\/p>\n<p>The company makes money by charging users a flat rate per million tokens. Tokens are the basic unit of data that an A.I. reads, processes and generates; in English, a token is roughly four characters, or about three-quarters of a word.<\/p>\n<p>\u201cWe provide one platform covering the whole end-to-end spectrum of model development, from quality to speed and cost. The end result is our customer gets better quality, much faster speed, and five to ten times lower cost, allowing them to go to production at a massive scale quickly,\u201d Lin said.<\/p>\n<p>The new moat<\/p>\n<p>These days, A.I. executives like to talk about \u201cmoat,\u201d or a competitive edge that allows a company to stay ahead of the competition. In a time when it\u2019s easier than ever to turn an idea into an application thanks to A.I. coding tools, the traditional moat of products disappears.<\/p>\n<p>\u201cData is the moat, because it cannot be copied,\u201d Lin declared. \u201cThe data collected to understand user intent, user preferences and user engagement\u2014what works well, what doesn\u2019t work well and where you should optimize\u2014is all your proprietary information, and that creates the asymmetry needed to compete. Whoever can turn this data into their proprietary intelligence can build on top of that. And that can compound.\u201d<\/p>\n<p>Fireworks competes with both closed-model providers (such as OpenAI, <a href=\"https:\/\/observer.com\/company\/anthropic\/\" title=\"Anthropic\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Anthropic<\/a> and Google) and infrastructure platforms like <a href=\"https:\/\/observer.com\/company\/together-ai\/\" title=\"Together AI\" class=\"company-link\" rel=\"nofollow noopener\" target=\"_blank\">Together AI<\/a>, Replicate and AWS Bedrock. Its differentiation lies in focusing on open models while tightly integrating training, fine-tuning and high-performance inference into a single system.<\/p>\n<p>\u201cWe don\u2019t need a Ferrari for grocery shopping.\u201d<\/p>\n<p>Besides the data moat, another argument for open models is unit economics. By allowing developers to choose from a wide range of open-weight models, platforms like Fireworks can match each task with the most cost-efficient level of intelligence. This flexibility is increasingly important as companies look to deploy A.I. at scale. Using a single, frontier model for every task quickly becomes prohibitively expensive.<\/p>\n<p>\u201cWe don\u2019t need to drive a Ferrari to go grocery shopping,\u201d Lin said. \u201cThere are so many tasks we solve day-to-day at varying levels of complexity. Some are extremely hard, requiring beyond-human-level intelligence to solve. Others are not that hard. If you use a vendor who can help you automatically select the best model suitable for solving a particular task, you get the quality you need at the lowest cost.\u201d<\/p>\n<p>When Lin founded Fireworks two years ago, the company initially focused on inference, treating it as \u201cone size fits one.\u201d Now, it is doubling down on training as well, driven by the rapid improvement and release cadence of open models. Open model quality has significantly narrowed the gap with closed models, while release cycles have accelerated from monthly to weekly. New models frequently top benchmarks and approach frontier-level performance.<\/p>\n<p>\u201cThis makes training particularly appealing. With your private data and a little bit of tuning, you can stay on top,\u201d Lin said.<\/p>\n<p>She continued to conclude, \u201cWe believe specialized and generalized intelligence will coexist, but the world will not be dominated by a few generalized models. There will be millions of specialized intelligence models\u2014one per use case.\u201d<\/p>\n<p>\t\t\t\t<img decoding=\"async\" itemprop=\"image\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/07\/GettyImages-2184020005.jpg\" alt=\"Lin Qiao\u2019s Fireworks AI Bets on Specialized Models Over General A.I. Hype\" style=\"display:none;width:0;\"\/><\/p>\n","protected":false},"excerpt":{"rendered":"According to Lin Qiao, the future of A.I. lies in millions of specialized models built on proprietary data.&hellip;\n","protected":false},"author":2,"featured_media":110513,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[2],"tags":[24,52890,53,25,309,21192,6541,6010,35362,132,10736,6561,56880,1122,157,5115,134,53119,56881,5411,52889],"class_list":["post-110512","post","type-post","status-publish","format-standard","has-post-thumbnail","category-ai","tag-ai","tag-all-features","tag-anthropic","tag-artificial-intelligence","tag-business","tag-business-interviews","tag-cursor","tag-deepseek","tag-fireworks-ai","tag-google","tag-harvey","tag-interviews","tag-lin-qiao","tag-meta","tag-openai","tag-shopify","tag-technology","tag-the-new-power-class-weekly-features","tag-together-ai","tag-uber","tag-weekly-features"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/110512","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=110512"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/110512\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/110513"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=110512"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=110512"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=110512"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}