{"id":139106,"date":"2026-08-13T19:01:19","date_gmt":"2026-08-13T19:01:19","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/139106\/"},"modified":"2026-08-13T19:01:19","modified_gmt":"2026-08-13T19:01:19","slug":"cerebras-powers-openai-gpt-5-6-sol-at-750-tps","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/139106\/","title":{"rendered":"Cerebras Powers OpenAI GPT-5.6 Sol at 750 TPS"},"content":{"rendered":"<p>Cerebras (NASDAQ: CBRS) announced that its hardware powers Ultrafast mode, a new OpenAI API tier for GPT-5.6 Sol, delivering up to 750 output tokens per second and up to 14\u00d7 faster processing than the Standard tier, with the same model intelligence.<\/p>\n<p>According to Cerebras, Ultrafast is initially in limited preview and targets workloads needing very low latency. Reported comparisons indicate Ultrafast runs about 5\u00d7 faster than Claude Opus 4.8 Fast and 11\u00d7 faster than Claude Fable 5, and achieves a 5.6\u00d7 speedup on GDP-Val benchmarks with no measured quality loss.<\/p>\n<p>\n            Loading&#8230;\n          <\/p>\n<p>          Loading translation&#8230;<\/p>\n<p>          Positive<\/p>\n<p>                    Powers OpenAI GPT-5.6 Sol Ultrafast tier at up to 14\u00d7 speed<\/p>\n<p>                    Supports up to 750 output tokens per second for GPT-5.6 Sol<\/p>\n<p>                    Ultrafast about 5\u00d7 faster than Claude Opus 4.8 Fast mode<\/p>\n<p>                    Ultrafast about 11\u00d7 faster than Claude Fable 5 on output speed<\/p>\n<p>                    5.6\u00d7 end-to-end speedup on GDP-Val with no quality loss reported<\/p>\n<p>                    Wafer-Scale Engine keeps 44 GB of model weights on-chip<\/p>\n<p>          Negative<\/p>\n<p>                    Ultrafast service initially limited to a small group of customers<\/p>\n<p>The limited-preview service uses Cerebras\u2019s Wafer-Scale Engine to keep model weights in 44 GB of on-chip SRAM per wafer-sized chip instead of moving them between on-chip memory and off-chip storage, which the release identifies as the source of its faster inference.<\/p>\n<p class=\"context-narrative-text\">\n      Historical event 1085268 recorded a 0.59% reaction after a CrowdStrike partnership. This platform record places the OpenAI service launch beside mixed partnership outcomes, while high short positioning and recent insider net selling remain risks to monitor.\n    <\/p>\n<p>\n        Ultrafast speed<br \/>\n        750 output tokens per second<\/p>\n<p>        GPT-5.6 Sol Ultrafast limited preview<\/p>\n<p>\n        Speed advantage<br \/>\n        14\u00d7 faster<\/p>\n<p>        vs. Standard processing<\/p>\n<p>\n        Claude Opus speed advantage<br \/>\n        5x faster<\/p>\n<p>        vs. Claude Opus 4.8 in Fast mode<\/p>\n<p>\n        Claude Fable speed advantage<br \/>\n        11x faster<\/p>\n<p>        vs. Claude Fable 5<\/p>\n<p>\n        Benchmark size<br \/>\n        2,500 questions<\/p>\n<p>        Humanity\u2019s Last Exam<\/p>\n<p>\n        Exam completion time<br \/>\n        just over 11 hours<\/p>\n<p>        GPT-5.6 Sol Ultrafast on Humanity\u2019s Last Exam<\/p>\n<p>\n        Accuracy speed advantage<br \/>\n        nearly 7x faster<\/p>\n<p>        vs. Claude Fable 5 at comparable accuracy<\/p>\n<p>\n        End-to-end speedup<br \/>\n        5.6x<\/p>\n<p>        GDP-Val benchmark with no loss in quality<\/p>\n<p>            Date<br \/>\n            Event<br \/>\n            Sentiment<br \/>\n            24h Move<br \/>\n            Catalyst<\/p>\n<p>            Aug 05<\/p>\n<p>                <a href=\"https:\/\/www.stocktitan.net\/news\/CBRS\/lovable-and-cerebras-partner-to-power-ai-software-creation-on-the-5rkwb4qqg9fh.html\" rel=\"nofollow noopener\" target=\"_blank\">AI partnership<\/a><\/p>\n<p>              Positive<\/p>\n<p>              -5.7%<\/p>\n<p>              Lovable partnership announced to deploy Cerebras inference for latency-sensitive software workflows.<\/p>\n<p>            Jul 29<\/p>\n<p>                <a href=\"https:\/\/www.stocktitan.net\/news\/CBRS\/clean-core-solutions-inc-nyse-american-zone-signs-ai-colocation-8kmd63zhjrcw.html\" rel=\"nofollow noopener\" target=\"_blank\">Colocation agreement<\/a><\/p>\n<p>              Positive<\/p>\n<p>              -12.1%<\/p>\n<p>              Cerebras signed a long-term data-center colocation agreement with CleanCore Solutions.<\/p>\n<p>            Jul 22<\/p>\n<p>                <a href=\"https:\/\/www.stocktitan.net\/news\/CBRS\/cerebras-systems-sets-date-of-second-quarter-2026-financial-guzu6dr7nsiu.html\" rel=\"nofollow noopener\" target=\"_blank\">Earnings scheduling<\/a><\/p>\n<p>              Neutral<\/p>\n<p>              +4.9%<\/p>\n<p>              Cerebras scheduled release of second-quarter 2026 financial results and conference call.<\/p>\n<p>            Jul 22<\/p>\n<p>                <a href=\"https:\/\/www.stocktitan.net\/news\/CBRS\/crowd-strike-and-cerebras-partner-to-power-ai-detection-and-response-c0z2mpeoumzb.html\" rel=\"nofollow noopener\" target=\"_blank\">AI partnership<\/a><\/p>\n<p>              Positive<\/p>\n<p>              +0.6%<\/p>\n<p>              CrowdStrike deployed Falcon detection models on Cerebras inference infrastructure.<\/p>\n<p>            Jul 09<\/p>\n<p>                <a href=\"https:\/\/www.stocktitan.net\/news\/CBRS\/flex-and-cerebras-expand-partnership-to-scale-american-manufacturing-soi2ss0tg9xd.html\" rel=\"nofollow noopener\" target=\"_blank\">Manufacturing partnership<\/a><\/p>\n<p>              Positive<\/p>\n<p>              +9.3%<\/p>\n<p>              Flex expanded manufacturing capacity for Cerebras CS-3 AI supercomputers.<\/p>\n<p>        Pattern Detected<\/p>\n<p class=\"context-pattern-text\">Positive partnership and commercial-expansion announcements produced mixed reactions, with two positive outcomes and three divergences in the selected history.<\/p>\n<p>          inference<\/p>\n<p>          technical<\/p>\n<p>&#8220;Cerebras\u2019 inference technology&#8221;<\/p>\n<p>Inference is the process of drawing a conclusion from available evidence or data, like a detective piecing together clues to form a likely story. For investors it matters because these judgments turn raw reports, test results, or market signals into expectations about future performance, risk, or regulatory outcomes\u2014so how someone infers from the same facts can change investment decisions and valuation.<\/p>\n<p>          sram<\/p>\n<p>          technical<\/p>\n<p>&#8220;44 GB of SRAM on each wafer-sized chip&#8221;<\/p>\n<p>Static random-access memory (SRAM) is a fast type of computer memory that stores data instantly and reliably while power is on, similar to a small, very quick notepad a processor keeps next to it. Investors care because SRAM influences product speed, power use, cost and production capacity for electronics and semiconductor companies, affecting profit margins, competitive positioning and supply-chain risk.<\/p>\n<p>          memory-bandwidth bottleneck<\/p>\n<p>          technical<\/p>\n<p>&#8220;This eliminates the memory-bandwidth bottleneck&#8221;<\/p>\n<p>A memory-bandwidth bottleneck occurs when a computer processor cannot get data from memory fast enough to run at full speed, so the CPU or accelerator sits idle waiting for information. Think of it like a highway where too few lanes slow traffic to a trickle even though the cars (processors) are ready to go; this limits overall performance and can reduce the value of hardware or software that depends on fast data flow. Investors care because such bottlenecks affect product competitiveness, server utilization, and the real-world performance gains behind revenue and cost forecasts.<\/p>\n<p>          on-chip memory<\/p>\n<p>          technical<\/p>\n<p>&#8220;between on-chip memory and off-chip storage&#8221;<\/p>\n<p>Memory that is built directly into an integrated circuit or processor die, used to store data and instructions close to where they are processed. Like a notepad kept on a worker\u2019s desk instead of across the room, on-chip memory is much faster and more energy-efficient than external memory but usually smaller; its size and speed affect a chip\u2019s performance, power consumption, and cost, which in turn influence product capabilities and competitiveness in the market.<\/p>\n<p class=\"context-ai-disclaimer\">AI-generated analysis. <a href=\"https:\/\/www.stocktitan.net\/rhea-ai.html\" rel=\"nofollow noopener\" target=\"_blank\">How Rhea-AI works<\/a>. Not financial advice.<\/p>\n<p>    &#13;<br \/>\n&#13;<br \/>\n&#13;<\/p>\n<p>  <img decoding=\"async\" class=\"ps-bar__icon\" src=\"https:\/\/static.stocktitan.net\/img\/icons\/Google_News_icon.svg\" width=\"24\" height=\"24\" alt=\"\" loading=\"lazy\" aria-hidden=\"true\"\/><\/p>\n<p>&#13;<br \/>\n    &#13;<br \/>\n    See more from StockTitan in Google Search and AI answers.&#13;<br \/>\n    Adds StockTitan as a preferred source \u00b7 opens Google&#13;\n  <\/p>\n<p>&#13;<br \/>\n&#13;<br \/>\n&#13;<br \/>\n&#13;<\/p>\n<p>      08\/13\/2026 &#8211; 01:00 PM<\/p>\n<p>Cerebras powers OpenAI&#8217;s most capable model at up to 14x the speed, giving frontier intelligence with no compromise on latency<\/p>\n<p>SUNNYVALE, Calif., Aug.  13, 2026  (GLOBE NEWSWIRE) &#8212; <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=8kV5FpRBRx0ZAbIA8I47HoNGLenI9W9XIzgcAXNBI3M2ms5a8N01MC5m0ReLW8iDPPNBnJei2eRCIiPiqqEokA==\" rel=\"nofollow noopener\" target=\"_blank\">Cerebras<\/a> (NASDAQ: <a href=\"https:\/\/www.stocktitan.net\/overview\/CBRS\/\" title=\"View CBRS stock overview\" class=\"symbol-link\" rel=\"nofollow noopener\" target=\"_blank\">CBRS<\/a>) today announced that it is powering Ultrafast mode, a new service tier in the OpenAI API for GPT-5.6 Sol. Available initially in limited preview to OpenAI customers, Ultrafast runs GPT-5.6 Sol at up to 750 output tokens per second and up to 14\u00d7 faster than Standard processing.<\/p>\n<p>GPT-5.6 Sol Ultrafast powered by Cerebras runs with the same intelligence as GPT-5.6 Sol Standard, enabling frontier intelligence at Cerebras&#8217; blistering fast speed.<\/p>\n<p>Every major computing shift has been unlocked by a leap in speed, not just capability: the PC era needed the jump from kilohertz to gigahertz, and the internet needed the move from dial-up to broadband before it could reach everyone. AI is no different. Now that frontier models are demonstrably capable, the next constraint on adoption is how fast they run.<\/p>\n<p>\u201cGPT-5.6 Sol on Ultrafast is proof that speed and intelligence are no longer mutually exclusive,\u201d said Andrew Feldman, CEO and co-founder, Cerebras. \u201cTogether with OpenAI, we&#8217;re putting frontier intelligence in the hands of users at unprecedented speed and changing what&#8217;s possible with AI.\u201d<\/p>\n<p>\u201cBy combining GPT-5.6 Sol with Cerebras\u2019 inference technology, we\u2019re exploring what becomes possible when customers can get the intelligence of our most capable models with significantly lower latency. We\u2019re starting with a small group of customers to learn where that speed creates meaningful value, and we\u2019ll use those learnings to inform how we expand the service over time,\u201d said Sachin Katti, VP Compute Strategy &amp; GPT-Infra at OpenAI.<\/p>\n<p>A New Speed-Intelligence Frontier<\/p>\n<p>Until now, organizations needed to choose between the capabilities of larger models and the faster response times of smaller models. Ultrafast powered by Cerebras offers access to the full intelligence of GPT-5.6 Sol at speeds suited to work that cannot wait. Based on output speeds for Anthropic models reported by <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=uVS2CrKOR21U-Z4CVYjAX7P_1i9jwyPR1HPNDEtQRuSwWaBAp2fDVLBhgig13eXuahejwqmiHJEMxHNt1mcxzGc8Cs1GTQ6hVmyrzPCGqG-HE6VrazxItwJ3R2n94lRKGuRC9SwYSGJ8vynZOcOfocnvQOw0dWHt3Dg7HF1dlqc-KBGLxM9SLCehktiHRS7I\" rel=\"nofollow noopener\" target=\"_blank\">Artificial Analysis<\/a>, Ultrafast is 5x faster than Claude Opus 4.8 in Fast mode, and 11x faster than Claude Fable 5. In doing so, it opens up a new region of speed-intelligence: frontier intelligence combined with unprecedented speed.<\/p>\n<p>Benchmarks focused on economically valuable work, from programming to drafting legal documents, show how faster token generation translates directly into higher productivity for users.<\/p>\n<p>On Humanity\u2019s Last Exam, a 2,500-question benchmark spanning graduate-level chemistry, economics and literature, GPT-5.6 Sol Ultrafast answered the full question set in just over 11 hours. This compares to more than three days of continuous compute for Claude Fable 5, with GPT-5.6 Sol Ultrafast reaching comparable accuracy nearly 7x faster. On GDP-Val, a benchmark of economically valuable knowledge-work tasks such as legal briefs, financial models, and engineering reports, Ultrafast delivered a 5.6x end-to-end speedup with no loss in quality.<\/p>\n<p>Ultrafast\u2019s speed comes from Cerebras&#8217; Wafer-Scale Engine architecture, which keeps model weights on-chip \u2014 44 GB of SRAM on each wafer-sized chip \u2014 rather than shuttling them between on-chip memory and off-chip storage as GPU-based inference must. This eliminates the memory-bandwidth bottleneck that constrains frontier-model inference speed on conventional hardware.<\/p>\n<p>To get notified when capacity expands to more customers, please visit the <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=8kV5FpRBRx0ZAbIA8I47HvZp7WD9bTj_Pva1FN6vilzV97LBtNJ4fmwie9Q18orInFMb4t2y0riR2S49lhWRBHCDVq9bWd5x9kattLfllzvypQNF4tie7eYge0VAc0nWswwj7vzWqxNaq5pfCwjZE5a-4OalTOke2FLGD4tx-RZLj9RciFsYgWvUdzE7-xsa0Z6ZbLrPfbOQQIfAdAsz84mygo2a9L5uuBQaKMJJECQgOEG_u_8PUL0SY-NMKss0NU39Sj6VOF0uZSowvIvM0Q==\" rel=\"nofollow noopener\" target=\"_blank\">Cerebras website<\/a>.<\/p>\n<p>About Cerebras Systems<br \/><a href=\"https:\/\/www.globenewswire.com\/Tracker?data=8kV5FpRBRx0ZAbIA8I47HltThTcJKz9pJizs6YtpuX-kagmvxA253yk9ZNsdXuAhvDK6P-QqLkBaT9buB22xMVmZC2DWOzWrQ1v_0UtW8X0=\" rel=\"nofollow noopener\" target=\"_blank\">Cerebras Systems<\/a> (NASDAQ: <a href=\"https:\/\/www.stocktitan.net\/overview\/CBRS\/\" title=\"View CBRS stock overview\" class=\"symbol-link\" rel=\"nofollow noopener\" target=\"_blank\">CBRS<\/a>) builds the world\u2019s fastest AI infrastructure. The Cerebras team of pioneering computer architects, computer scientists, AI researchers, and engineers of all types came together to make AI blisteringly fast through innovation and invention. We believe that when AI is fast, it will change the world. Leading global corporations, research institutes, and governments choose Cerebras to run their AI workloads. Cerebras solutions are available on premises and in the cloud. Visit\u00a0<a href=\"https:\/\/www.globenewswire.com\/Tracker?data=pqdDH8vWkLhRBxgi-rySPW-_Bw_MGHSkEjvx7rrhZEe8JEj6JATSRCBDfOz1T1Z_iHiO6_4PUJ6msG_adGGPghIbEGCmY3k1IaTD5PxJtXamotlwVFGpAG1Z12WMIPo-m5UZklIScKVpPJGhJiYLEBi8cLGG8kWor7lPeG5uFqpUWjK4eyFSHNh5CS0BXlYxN_0GwzY6vpGEsqP57jpmjgDmCVWaAgogeqmkf9gisJ4=\" rel=\"nofollow noopener\" target=\"_blank\">cerebras.ai<\/a><a href=\"https:\/\/www.globenewswire.com\/Tracker?data=llM96Gr2VuBBGSTk7_3H2yGraC32Z3kcBJoFH8-izL90xYBXW52CtMrmbSrhJdhEY3GLA4BRUjvXGr29wMoFCM9advSqS-2LwCKLcIrckfKGr9TO_NyJ6gsD5OMSRwL6FQvPhjDTCCcmMBywZmJ-7jlkpEJVz0eKLuMgFyD54q6lDluWtG8l3UA56dsbTZcZUorbJPRC1-VWdtQLYHmu8ygiohvh8oF_z9w3oKH5m8s=\" rel=\"nofollow noopener\" target=\"_blank\">\u00a0for more<\/a>.<\/p>\n<p>Cerebras Disclosure Information <br \/>Cerebras uses its investor relations page (<a href=\"https:\/\/www.globenewswire.com\/Tracker?data=9jvo_HyvrpAQZycORGVvVXxbJ0nXsXOwiYaAbXhl0_4CbHorf9ILtb2hFrF-mtVE5HRivLaMq_8ZnawN4RA3QxqNvdKaZeiicItgbKBAOz4=\" rel=\"nofollow noopener\" target=\"_blank\">investors.cerebras.ai<\/a>), its X account <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=bU0KJ-hGacr9KE29D2jH27wbusaTvMoNbm_tswYT2cPPOepdjaX7U6JAYUrpARaTZResGENJF8NxwTajh5bvyw==\" rel=\"nofollow noopener\" target=\"_blank\">(@cerebras<\/a>), and its LinkedIn page (<a href=\"https:\/\/www.globenewswire.com\/Tracker?data=yCsJ3fSAgTbOU8x0QSJX34-XBum3Sio_VpYoBTQaxqm14YRshLO-JzyFP4KEx7i9w78pSh3I5ySd2-SY9VJ_o5AuJS0-b3u7PBv3zDTb-JezhNXAeGJDMhy39Yyi6vZBtb0x9JKe2X2KeUDNJmvl0A==\" rel=\"nofollow noopener\" target=\"_blank\">linkedin.com\/company\/cerebras-systems\/<\/a>) to disclose material nonpublic information and for complying with its disclosure obligations under Regulation FD. Accordingly, investors should monitor these channels, in addition to following Cerebras&#8217; press releases, Securities and Exchange Commission (SEC) filings, public conference calls and public webcasts.<\/p>\n<p>Forward-Looking Statements<\/p>\n<p>This press release contains &#8220;forward-looking statements&#8221; within the meaning of applicable securities laws. All statements other than statements of historical fact could be deemed to be forward-looking. The words &#8220;may,&#8221; &#8220;will,&#8221; &#8220;shall,&#8221; &#8220;should,&#8221; &#8220;expects,&#8221; &#8220;plans,&#8221; &#8220;anticipates,&#8221; &#8220;could,&#8221; &#8220;intends,&#8221; &#8220;target,&#8221; &#8220;projects,&#8221; &#8220;contemplates,&#8221; &#8220;believes,&#8221; &#8220;estimates,&#8221; &#8220;predicts,&#8221; &#8220;potential,&#8221; &#8220;objective,&#8221; or &#8220;continue,&#8221; or the negative of these words or other similar terms or expressions that concern our expectations, strategy, plans, or intentions are intended to identify forward-looking statements, although not all forward-looking statements contain these identifying words. These forward-looking statements are subject to a number of risks and uncertainties, many of which involve factors or circumstances that are beyond Cerebras&#8217; control. These risks and uncertainties include, but are not limited to: Cerebras&#8217; ability to sustain and manage its growth, access borrowings and other sources of capital on acceptable terms, and deploy available capital to support growth; its history of net losses and ability to achieve and maintain profitability; its limited operating history at its current scale and ability to accurately forecast revenue and appropriately budget and manage expenses; its dependence on a limited number of significant customers, including OpenAI, Group 42 Holding Ltd, Mohamed bin Zayed University of Artificial Intelligence, and AWS, and the potential impact of any reduction in demand from, material adverse development in its relationships with, or failure to meet its obligations to, such customers; the timing, execution and expected benefits of its strategic customer, partner and financing arrangements; its historical reliance on sales of hardware systems and the early-stage, rapidly evolving market for its cloud-based offerings and AI infrastructure; its ability to secure sufficient data center capacity and capital to support its cloud-based offerings; its ability to launch new offerings and add new product capabilities; and its ability to compete effectively in the rapidly evolving and competitive market for AI computing solutions.<\/p>\n<p>Cerebras&#8217; actual results could differ materially from those stated or implied in forward-looking statements due to a number of factors. Accordingly, undue reliance should not be placed on such statements. These forward-looking statements are made as of the date they were first issued and are based on information available to Cerebras together with Cerebras&#8217; expectations, estimates, forecasts, projections, beliefs, and assumptions as of such date. These forward-looking statements should not be relied upon as representing Cerebras&#8217; views as of any date subsequent to the date of this press release. Past performance is not necessarily indicative of future results. Cerebras undertakes no intention or obligation to update or revise any forward-looking statements, whether as a result of new information, future events, or otherwise, except as required by law.<\/p>\n<p>Further information on potential risks that could affect actual results is included in Cerebras&#8217; most recent filings with the SEC, including in Cerebras&#8217; most recent Quarterly Report on Form 10-Q, copies of which may be obtained by visiting Cerebras&#8217; Investor Relations website at investors.cerebras.ai or the SEC&#8217;s website at <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=5dsnomANR0pkhDPRbyNpBfEEcuR2BCttaKdh2kn5ElAkRQqNDrK3C5pSL7JCq8EGgWNVfr4on5l9P-XU6k-DSw==\" rel=\"nofollow noopener\" target=\"_blank\">www.sec.gov<\/a>.<\/p>\n<p align=\"left\">Contacts<\/p>\n<p align=\"left\">Kriselle Laran<br \/>Media Relations<br \/><a href=\"https:\/\/www.globenewswire.com\/Tracker?data=D8uR0-2UPQhVW46vMk-xAxMCj8Qnky1lsxWgoxo68uRVew4Hh_kn9BI4uEKcj51Lsxv6c_v_CL89yubo7OvbrA==\" rel=\"nofollow noopener\" target=\"_blank\">pr@cerebras.ai<\/a><\/p>\n<p align=\"left\">Sean Dorsey<br \/>Investor Relations<br \/><a href=\"https:\/\/www.globenewswire.com\/Tracker?data=9jvo_HyvrpAQZycORGVvVY3XNItP4H8s-h-5oM8Xx-8j89L0jgY9-_EBZbs8uF5KF20X6SY-arVXj3J_dkmgcdJib5tAsZzpX6Y3te8RpNw=\" rel=\"nofollow noopener\" target=\"_blank\">investors@cerebras.ai<\/a><\/p>\n<p> <img decoding=\"async\" loading=\"lazy\" alt=\"\" class=\"__GNW8366DE3E__IMG\" src=\"https:\/\/www.globenewswire.com\/newsroom\/ti?nf=OTgwOTczNiM3Nzc1NTkzIzIyOTQ1OTU=\"\/> <br \/><img decoding=\"async\" loading=\"lazy\" alt=\"\" src=\"https:\/\/www.europesays.com\/ai\/wp-content\/uploads\/2026\/08\/Cerebras-Systems-Inc-.png\" referrerpolicy=\"no-referrer-when-downgrade\"\/><\/p>\n<p>      &#13;<br \/>\n&#13;<br \/>\n&#13;<br \/>\n&#13;<br \/>\n&#13;<br \/>\n&#13;<br \/>\n  &#13;<br \/>\n&#13;<\/p>\n<p>&#13;<br \/>\n    FAQ  &#13;\n  <\/p>\n<p>&#13;<br \/>\n  &#13;<br \/>\n  &#13;<\/p>\n<p>        What did Cerebras (NASDAQ: CBRS) announce about powering OpenAI GPT-5.6 Sol Ultrafast?<\/p>\n<p>&#13;<br \/>\n          Cerebras announced it powers OpenAI\u2019s new GPT-5.6 Sol Ultrafast API tier, delivering up to 750 tokens per second and up to 14\u00d7 faster processing than Standard. According to Cerebras, Ultrafast maintains the same model intelligence while significantly reducing latency for demanding workloads.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<\/p>\n<p>        How fast is GPT-5.6 Sol Ultrafast powered by Cerebras compared with the Standard tier?<\/p>\n<p>&#13;<br \/>\n          According to Cerebras, GPT-5.6 Sol in Ultrafast mode runs up to 14\u00d7 faster than Standard processing and can generate up to 750 output tokens per second. This speed targets applications where response time is critical but users still require the model\u2019s full intelligence.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<\/p>\n<p>        How does Cerebras-powered Ultrafast compare to Anthropic models for speed?<\/p>\n<p>&#13;<br \/>\n          Cerebras reports that GPT-5.6 Sol Ultrafast is about 5\u00d7 faster than Claude Opus 4.8 in Fast mode and about 11\u00d7 faster than Claude Fable 5. These comparisons are based on output speed data reported by Artificial Analysis, according to Cerebras.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<\/p>\n<p>        What benchmark results did Cerebras share for GPT-5.6 Sol Ultrafast on Humanity\u2019s Last Exam?<\/p>\n<p>&#13;<br \/>\n          According to Cerebras, GPT-5.6 Sol Ultrafast completed the 2,500-question Humanity\u2019s Last Exam benchmark in just over 11 hours, reaching comparable accuracy to Claude Fable 5 nearly 7\u00d7 faster. The benchmark spans graduate-level chemistry, economics, literature and other knowledge-work domains.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<\/p>\n<p>        What is the GDP-Val benchmark result for Cerebras-powered GPT-5.6 Sol Ultrafast?<\/p>\n<p>&#13;<br \/>\n          On the GDP-Val benchmark of economically valuable tasks, GPT-5.6 Sol Ultrafast delivered a 5.6\u00d7 end-to-end speedup with no loss in quality, according to Cerebras. GDP-Val includes knowledge-work tasks such as legal briefs, financial models, engineering reports and similar productivity-focused workloads.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<\/p>\n<p>        How does Cerebras hardware achieve higher GPT-5.6 Sol Ultrafast inference speed?<\/p>\n<p>&#13;<br \/>\n          According to Cerebras, its Wafer-Scale Engine architecture keeps 44 GB of model weights in on-chip SRAM, avoiding transfers to off-chip memory. This design removes the memory-bandwidth bottleneck that often constrains frontier-model inference speed on conventional GPU-based hardware platforms.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<\/p>\n<p>        Is OpenAI\u2019s GPT-5.6 Sol Ultrafast powered by Cerebras widely available to CBRS investors\u2019 customers?<\/p>\n<p>&#13;<br \/>\n          Cerebras states that GPT-5.6 Sol Ultrafast is initially available in limited preview to a small group of OpenAI customers. Interested users can register on the Cerebras website to be notified when capacity increases and the service expands to more customers.&#13;\n        <\/p>\n<p>    &#13;<br \/>\n  &#13;<br \/>\n&#13;<br \/>\n&#13;<\/p>\n","protected":false},"excerpt":{"rendered":"Cerebras (NASDAQ: CBRS) announced that its hardware powers Ultrafast mode, a new OpenAI API tier for GPT-5.6 Sol,&hellip;\n","protected":false},"author":2,"featured_media":139107,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[407,45327,24111,46529,68621,47338,157,32131,68622],"class_list":["post-139106","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-ai-infrastructure","tag-cbrs","tag-cerebras","tag-gpt-5-6-sol","tag-inference-speed","tag-limited-preview","tag-openai","tag-openai-api","tag-ultrafast"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/139106","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=139106"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/139106\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/139107"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=139106"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=139106"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=139106"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}