{"id":140490,"date":"2026-08-14T22:07:16","date_gmt":"2026-08-14T22:07:16","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/140490\/"},"modified":"2026-08-14T22:07:16","modified_gmt":"2026-08-14T22:07:16","slug":"speed-becomes-the-product-as-openai-and-google-sell-faster-ai","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/140490\/","title":{"rendered":"Speed Becomes the Product as OpenAI and Google Sell Faster AI"},"content":{"rendered":"<p><a href=\"https:\/\/openai.com\/\" rel=\"nofollow noopener\" target=\"_blank\">OpenAI<\/a> and <a href=\"https:\/\/www.google.com\/\" rel=\"nofollow noopener\" target=\"_blank\">Google<\/a> both released faster artificial intelligence models this week. Both put speed at the center of the pitch, and both are starting to treat response time as something businesses will pay for on its own.<\/p>\n<p>OpenAI\u2019s new tier is called <a href=\"https:\/\/openai.com\/index\/previewing-ultrafast\/\" rel=\"nofollow noopener\" target=\"_blank\">Ultrafast<\/a>, and for now it is a preview open to a small group of customers. It runs the company\u2019s <a href=\"https:\/\/openai.com\/index\/gpt-5-6\/\" rel=\"nofollow noopener\" target=\"_blank\">GPT-5.6 Sol<\/a> model up to 14 times faster than the standard tier, at up to 750 output tokens per second, according to a Thursday (Aug. 13) company announcement. The model itself underneath is the same. It just answers faster, running on chips from a company called <a href=\"https:\/\/www.cerebras.ai\/\" rel=\"nofollow noopener\" target=\"_blank\">Cerebras<\/a>.<\/p>\n<p>Early customers include <a href=\"https:\/\/www.janestreet.com\/\" rel=\"nofollow noopener\" target=\"_blank\">Jane Street<\/a>, <a href=\"https:\/\/www.podium.com\/\" rel=\"nofollow noopener\" target=\"_blank\">Podium<\/a>, <a href=\"https:\/\/www.getbasis.ai\/\" rel=\"nofollow noopener\" target=\"_blank\">Basis<\/a> and <a href=\"https:\/\/rogo.com\/\" rel=\"nofollow noopener\" target=\"_blank\">Rogo<\/a>, which are testing it for coding, financial research, customer support, voice and commerce, per the announcement.<\/p>\n<p>\u201cUntil now, getting real-time speed typically meant choosing a smaller or more specialized model,\u201d the announcement said. \u201cUltrafast points to progress in a new direction: more useful work per second.\u201d<\/p>\n<p>Google launched its own faster model, <a href=\"https:\/\/blog.google\/innovation-and-ai\/models-and-research\/gemini-models\/introducing-gemini-3-7-flash\/\" rel=\"nofollow noopener\" target=\"_blank\">Gemini 3.7 Flash<\/a>, the same day. It costs $0.75 per million input tokens and $3.75 per million output tokens through Dec. 31, half what the previous Flash model cost, then doubles on Jan. 1, 2027, to the price <a href=\"https:\/\/blog.google\/innovation-and-ai\/models-and-research\/gemini-models\/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber\/\" rel=\"nofollow noopener\" target=\"_blank\">Gemini 3.6 Flash<\/a> carried all along, according to a Thursday company blog post.<\/p>\n<p>Google is also claiming a capability gain, calling 3.7 Flash its \u201cmost intelligent workhorse model yet for coding and agents\u201d in the post.<\/p>\n<p>Benchmarking firm <a href=\"https:\/\/artificialanalysis.ai\/articles\/gemini-3-7-time-frontier\" rel=\"nofollow noopener\" target=\"_blank\">Artificial Analysis<\/a> clocked Gemini 3.7 Flash\u2019s output at about 340 tokens per second, nearly three times the speed of <a href=\"https:\/\/openai.com\/index\/gpt-5-6\/\" rel=\"nofollow noopener\" target=\"_blank\">GPT-5.6 Terra<\/a> and <a href=\"https:\/\/z.ai\/blog\/glm-5.2\" rel=\"nofollow\">GLM-5.2<\/a>, and placed it on the frontier of intelligence versus time per task.<\/p>\n<p>Both launches point to the same shift. For two years, AI pricing has mostly come down to how capable a model is and how much a company uses it. Speed is becoming a third feature companies are being asked to pay for on its own.<\/p>\n<p>Speed Matters More for Fraud Checks Than for Overnight Reports<\/p>\n<p>Not every AI task needs to be fast. A bank checking whether a transaction is fraudulent must decide in a fraction of a second. AI-driven <a href=\"https:\/\/www.pymnts.com\/cybersecurity\/fraud-prevention\/2026\/42-percent-of-issuers-say-ai-has-cut-fraud-losses-by-at-least-5-million\/\" rel=\"nofollow noopener\" target=\"_blank\">fraud detection<\/a> has already saved at least $5 million for 42% of card issuers, according to the <a href=\"https:\/\/www.pymnts.com\/pymnts-intelligence\/\" rel=\"nofollow noopener\" target=\"_blank\">PYMNTS Intelligence<\/a> report\u202f\u201c<a href=\"https:\/\/www.pymnts.com\/tracker_posts\/where-payment-decisions-happen-how-issuer-data-is-powering-the-next-era-of-commerce\/\" rel=\"nofollow noopener\" target=\"_blank\">Where Payment Decisions Happen: How Issuer Data Is Powering the Next Era of Commerce<\/a>.\u201d A slow fraud check is not just annoying. It can mean letting a fraudulent charge go through before the system catches up.<\/p>\n<p>In its announcement, OpenAI named voice applications, customer support, commerce, coding and financial research as the kind of work its fast tier is built for, along with incident response, which covers fraud and security threats that need a fast decision.<\/p>\n<p>Compare that to a company running an AI job overnight to sort through old documents. Nobody is waiting on the other end. That job can run on cheaper, slower computers with no real cost to the business. It\u2019s the same AI capability, but two different prices, depending only on whether someone is waiting on the answer.<\/p>\n<p>AI Pricing Is Splitting Into a Fast Lane and a Slow One<\/p>\n<p>This is similar to how companies already buy internet service by paying more for a guaranteed fast connection when it matters, and using a cheaper, slower connection everywhere else. Cerebras, the company powering OpenAI\u2019s fast tier, made the same argument in its Thursday <a href=\"https:\/\/www.globenewswire.com\/news-release\/2026\/08\/13\/3344804\/0\/en\/cerebras-powers-ultrafast-mode-for-openai-s-gpt-5-6-sol.html\" rel=\"nofollow noopener\" target=\"_blank\">announcement<\/a>, comparing the OpenAI launch to earlier tech shifts like the move from dial-up internet to broadband.<\/p>\n<p>If that comparison holds, businesses will start splitting their AI spending into two lanes. Fast, expensive AI will be used for jobs where a delay costs real money, like a customer chatting live with a company or a fraud check happening in real time. Slow, cheaper AI will be used for jobs where nobody notices the wait, like an overnight report or a batch of paperwork.<\/p>\n<p>This week\u2019s launches from OpenAI and Google suggest both companies expect that split to become normal, something businesses will soon plan for deliberately rather than treat as an afterthought.<\/p>\n<p>For all PYMNTS AI coverage, subscribe to the daily <a href=\"https:\/\/pymnts.com\/subscribe\/\" rel=\"nofollow noopener\" target=\"_blank\">AI Newsletter<\/a>.<\/p>\n","protected":false},"excerpt":{"rendered":"OpenAI and Google both released faster artificial intelligence models this week. Both put speed at the center of&hellip;\n","protected":false},"author":2,"featured_media":140491,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[25,313,2858,132,66,157,310,314],"class_list":["post-140490","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-artificial-intelligence","tag-cybersecurity","tag-fraud","tag-google","tag-news","tag-openai","tag-pymnts-news","tag-security"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/140490","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=140490"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/140490\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/140491"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=140490"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=140490"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=140490"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}