{"id":144385,"date":"2026-08-19T02:49:13","date_gmt":"2026-08-19T02:49:13","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/144385\/"},"modified":"2026-08-19T02:49:13","modified_gmt":"2026-08-19T02:49:13","slug":"cerebras-unveils-cs-4-openai-and-amd-partnerships-to-accelerate-ai-inference","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/144385\/","title":{"rendered":"Cerebras Unveils CS-4, OpenAI and AMD Partnerships to Accelerate AI Inference"},"content":{"rendered":"<p>        Key Points            <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\"><a href=\"https:\/\/www.marketbeat.com\/newsletter\/PDFoffer.aspx?offer=top5&amp;RegistrationCode=YahooFinance\" data-ylk=\"slk:Interested%20in%20Cerebras%20Systems%20Inc.%3F%20Here%20are%20five%20stocks%20we%20like%20better.;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;Interested in Cerebras Systems Inc.? Here are five stocks we like better.&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">Interested in Cerebras Systems Inc.? Here are five stocks we like better.<\/a>  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Cerebras unveiled its CS-4 AI system, which is expected to deliver up to twice the token-generation speed of CS-3, six times higher system-level performance and up to 10 times more tokens per watt in select applications.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The company highlighted partnerships with OpenAI and AMD to accelerate inference, including an AMD-GPU\/Cerebras split in which GPUs handle model prefill and Cerebras handles token decoding.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Cerebras is expanding its data-center footprint, with 600 megawatts of power online or under contract by the end of next year, while targeting AI-agent, design, coding and cybersecurity applications that benefit from lower latency.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Cerebras Systems (NASDAQ:CBRS) outlined a product roadmap centered on faster AI inference, expanded data-center capacity and new partnerships with OpenAI, Arista Networks and Advanced Micro Devices during a company event led by CEO and Co-Founder Andrew Feldman.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Feldman said the company recently went public and is seeking to position its wafer-scale computing technology as an infrastructure platform for real-time AI applications. He argued that inference speed has become a product-level consideration rather than solely a technical benchmark, particularly as AI shifts toward interactive tools and autonomous agents.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">\u2192 <a href=\"https:\/\/www.marketbeat.com\/articles\/amgs-alternatives-boom-powers-record-growth\/\" data-ylk=\"slk:AMG&#039;s%20Alternatives%20Boom%20Powers%20Record%20Growth;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;AMG&#039;s Alternatives Boom Powers Record Growth&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">AMG&#8217;s Alternatives Boom Powers Record Growth<\/a>  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;In AI, speed is productivity,&#8221; Feldman said, asserting that faster response times allow users to run more workloads and address more complex tasks without the traditional trade-off between model intelligence and latency.  <\/p>\n<p>            OpenAI Discusses Ultrafast Service Tier          <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Feldman highlighted OpenAI&#8217;s recently announced GPT-5.6 Sol Ultrafast mode, which he said runs OpenAI&#8217;s most intelligent models at up to 14 times the speed of its standard offering for select customers. He said Cerebras hardware powers the service and presented a comparison based on Humanity&#8217;s Last Exam, a graduate-level reasoning benchmark.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">\u2192 <a href=\"https:\/\/www.marketbeat.com\/articles\/microsofts-maia-300-chip-targets-nvidias-ai-dominance\/\" data-ylk=\"slk:Microsoft&#039;s%20Maia%20300%20Chip%20Targets%20NVIDIA&#039;s%20AI%20Dominance;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;Microsoft&#039;s Maia 300 Chip Targets NVIDIA&#039;s AI Dominance&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">Microsoft&#8217;s Maia 300 Chip Targets NVIDIA&#8217;s AI Dominance<\/a>  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">According to Feldman, the Cerebras-powered system completed the full 2,500-question benchmark in 11 hours, 11 minutes and 26 seconds, while the comparison system required more than three days.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Thibault Sottiaux, a member of the technical staff at OpenAI, said the company&#8217;s strategy has been to build leading models and serve them at scale, with increasing focus on agents. He said ultrafast inference reduces the need to choose between a smaller, lower-latency model and a larger model that takes longer to respond.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">\u2192 <a href=\"https:\/\/www.marketbeat.com\/articles\/the-metals-companys-big-bet-now-comes-down-to-a-license\/\" data-ylk=\"slk:The%20Metals%20Company&#039;s%20Big%20Bet%20Now%20Comes%20Down%20to%20a%20License;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;The Metals Company&#039;s Big Bet Now Comes Down to a License&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">The Metals Company&#8217;s Big Bet Now Comes Down to a License<\/a>  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Sottiaux said OpenAI would like Ultrafast to become the default experience over time, though he stressed that the company is still early in deploying it. He said OpenAI reserves some Ultrafast capacity for incidents, security matters, major internal projects and research efforts, while also allocating capacity for customers through its API.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">He added that ChatGPT has surpassed 1 billion users, while the company&#8217;s more sophisticated agentic workflows remain used by a smaller but rapidly growing segment. Feldman said OpenAI has 15 million weekly users of Codex and ChatGPT agents, a figure Sottiaux referenced as recently published by OpenAI.  <\/p>\n<p>       Data-Center Buildout and Infrastructure Partners         <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Feldman said Cerebras has data centers operating or being developed across North America and Europe, naming locations including Santa Clara, Toronto, Dallas, Minneapolis, Montreal, Oklahoma City, Alabama, Lyon, France, Norway and Mikkeli, Finland. He said the company has brought on 600 megawatts of power that is online or under contract for delivery by the end of next year.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Jayshree Ullal, CEO of Arista, said AI infrastructure requires coordinated networking alongside compute capacity. She described the growing importance of networking across scale-up, scale-out and geographically distributed &#8220;scale-across&#8221; deployments.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Ullal said that AI workloads differ from traditional cloud traffic and require networks that can handle varying traffic patterns, model types and agent workloads. She also said bandwidth requirements could continue to increase rapidly, citing potential networking speeds of 3.2 terabits and 6.4 terabits in the coming years.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Feldman also discussed a disaggregated inference approach with Mark Papermaster, AMD&#8217;s executive vice president and chief technology officer. Under that approach, GPUs handle the &#8220;prefill&#8221; stage of inference, in which an AI model processes an input context, while Cerebras systems handle &#8220;decode,&#8221; the token-by-token generation phase.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Feldman said this combination could provide inference that is 10 times faster than GPUs and five times more throughput than Cerebras alone. Papermaster said the partnership combines GPU strengths in parallel computation with Cerebras&#8217; low-latency decode capabilities.  <\/p>\n<p>          CS-4 System Targets Higher Performance         <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Sean Lie, Cerebras&#8217; CTO and co-founder, introduced the company&#8217;s next-generation CS-4 system, which is in early access and is expected to be generally available later this quarter.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Lie said CS-4 is designed around the company&#8217;s Nexus rack-scale platform and includes three wafer-scale engines in a single system. He said the platform will offer up to twice the token-generation speed of the current CS-3 generation, six times higher system-level performance, three times more density per rack and up to 10 times more tokens per watt in certain solutions.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">According to Lie, the system uses the company&#8217;s WSE-3 Turbo chip, with each wafer providing 43 petabytes per second of memory bandwidth. He said the architecture is intended to reduce bottlenecks associated with moving model weights across multiple chips and to support larger models through lower-latency wafer-to-wafer connections.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">CS-4 is expected to provide up to 30 times faster token generation than GPU solutions, according to Cerebras.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The system is designed with twice the I\/O bandwidth per wafer and lower network latency than the prior generation.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Cerebras said its modular rack architecture is intended to reduce components by 50% and shorten data-center deployment times from days to hours.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Lie said the Nexus architecture is intended to support future CS-5 and CS-6 systems. Cerebras&#8217; roadmap calls for doubling speed annually and delivering solutions with up to 20 times higher throughput by 2027.  <\/p>\n<p>       Focus on AI Agents and Time-Sensitive Applications         <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Jessica Liu, Cerebras&#8217; senior vice president of product, said fast inference can improve the usefulness of AI agents by allowing more model calls, planning loops and tool use within the same time budget. She said slower systems can force users to choose smaller models or wait longer for more capable agents.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Speakers from Figma, Cognition and CrowdStrike described applications for faster model inference in design, software engineering and cybersecurity. Figma AI Research Product Lead Pavi Bhatter said the company&#8217;s Design Agent, currently in beta, uses Cerebras hardware as part of its effort to make design workflows more interactive. Cognition&#8217;s Silas Alberti said the company has deployed coding models on Cerebras and reported faster end-to-end task completion in certain use cases.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Keith Coley, CrowdStrike&#8217;s vice president of engineering, said cybersecurity decisions can be highly time-sensitive and described the company&#8217;s partnership with Cerebras as a way to allow more AI-based inspection within acceptable response windows.  <\/p>\n<p>       About Cerebras Systems (NASDAQ:CBRS)         <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Cerebras Systems is a technology company focused on building artificial intelligence infrastructure, including hardware and software designed to accelerate deep learning and large-scale AI workloads. The company is best known for its wafer-scale processor architecture, which is intended to provide high-performance compute for training and inference applications.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">In addition to its AI chips, Cerebras offers systems and related software tools that support researchers and enterprises working with machine learning models.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">This instant news alert was generated by narrative science technology and financial data from MarketBeat in order to provide readers with the fastest reporting and unbiased coverage. Please send any questions or comments about this story to contact@marketbeat.com.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">The article &#8220;<a href=\"https:\/\/www.marketbeat.com\/instant-alerts\/cerebras-unveils-cs-4-openai-and-amd-partnerships-to-accelerate-ai-inference-2026-08-18\/?utm_source=yahoofinance&amp;utm_medium=yahoofinance\" data-ylk=\"slk:Cerebras%20Unveils%20CS-4%2C%20OpenAI%20and%20AMD%20Partnerships%20to%20Accelerate%20AI%20Inference;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;Cerebras Unveils CS-4, OpenAI and AMD Partnerships to Accelerate AI Inference&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">Cerebras Unveils CS-4, OpenAI and AMD Partnerships to Accelerate AI Inference<\/a>&#8221; was originally published by MarketBeat.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\"><a href=\"https:\/\/www.marketbeat.com\/newsletter\/pdfoffer.aspx?offer=newyear&amp;RegsistrationCode=YahooFinance-NewYear\" data-ylk=\"slk:View%20MarketBeat&#039;s%20top%20stocks%20for%20August%202026;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;View MarketBeat&#039;s top stocks for August 2026&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">View MarketBeat&#8217;s top stocks for August 2026<\/a>.  <\/p>\n","protected":false},"excerpt":{"rendered":"Key Points Interested in Cerebras Systems Inc.? Here are five stocks we like better. Cerebras unveiled its CS-4&hellip;\n","protected":false},"author":2,"featured_media":144386,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[36008,24111,3887,157,71086],"class_list":["post-144385","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-andrew-feldman","tag-cerebras","tag-inference","tag-openai","tag-system-level-performance"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/144385","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=144385"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/144385\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/144386"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=144385"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=144385"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=144385"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}