{"id":85179,"date":"2026-06-25T01:22:20","date_gmt":"2026-06-25T01:22:20","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/85179\/"},"modified":"2026-06-25T01:22:20","modified_gmt":"2026-06-25T01:22:20","slug":"openai-unveils-jalapeno-chip-for-large-scale-inference-workloads","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/85179\/","title":{"rendered":"OpenAI unveils Jalape\u00f1o chip for large-scale inference workloads"},"content":{"rendered":"<p class=\"wp-block-paragraph\">OpenAI has unveiled its first custom AI accelerator, called Jalape\u00f1o, marking the company\u2019s move into chip design as it looks to reduce the cost and improve the efficiency of running large language models (LLMs).<\/p>\n<p class=\"wp-block-paragraph\">Developed in partnership with Broadcom and Celestica, the chip is designed specifically for AI inference \u2014 the process of generating responses from trained AI models. OpenAI said early testing shows Jalape\u00f1o delivers significantly better performance per watt than current state-of-the-art AI accelerators, though detailed benchmarks will be released later.<\/p>\n<p class=\"wp-block-paragraph\">The announcement expands OpenAI\u2019s efforts to control more of the infrastructure behind its products. In addition to building models and applications such as ChatGPT and Codex, the company is now designing the hardware that powers them. Engineering samples of Jalape\u00f1o are already running machine learning workloads in the lab, including GPT-5.3-Codex-Spark, at production target frequency and power levels, according to the company.<\/p>\n<p>Custom chip, faster AI<\/p>\n<p class=\"wp-block-paragraph\">Unlike general-purpose AI accelerators adapted for multiple workloads, Jalape\u00f1o was built specifically for LLM inference. OpenAI said the architecture was designed around the compute, memory, networking, and serving requirements of modern AI models.<\/p>\n<p class=\"wp-block-paragraph\">The company claims the chip reduces data movement while balancing compute, memory, and networking resources to improve hardware utilization. Broadcom contributed silicon implementation and networking technologies, including its Tomahawk networking platform.<\/p>\n<p class=\"wp-block-paragraph\">\u201cJalape\u00f1o is part of our long-term full-stack infrastructure strategy to make compute more abundant, resulting in AI which is faster, more reliable, more affordable for people and businesses, and can be used to solve more important problems,\u201d said Greg Brockman, President and Co-Founder of OpenAI.<\/p>\n<p class=\"wp-block-paragraph\">Richard Ho, who leads OpenAI\u2019s hardware program, said the accelerator was optimized around the workloads most important for frontier AI systems.<\/p>\n<p class=\"wp-block-paragraph\">\u201cBased on early testing, Jalape\u00f1o will efficiently execute our most important workloads close to the hardware\u2019s theoretical limits,\u201d Ho said. The chip is also intended to support future LLMs across the broader AI industry, not just OpenAI\u2019s own models.<\/p>\n<p>Nine-month design sprint<\/p>\n<p class=\"wp-block-paragraph\">According to the companies, Jalape\u00f1o was developed from initial design to manufacturing tape-out in just nine months. OpenAI described the effort as potentially the fastest ASIC development cycle achieved for a high-performance advanced semiconductor.<\/p>\n<p class=\"wp-block-paragraph\">The development process involved extensive software-hardware co-design between OpenAI and Broadcom engineers. OpenAI also said its own <a href=\"https:\/\/interestingengineering.com\/ai-robotics\/china-model-humanoids-robot-arms\" target=\"_blank\" rel=\"dofollow noopener\">AI models<\/a> were used to accelerate portions of the chip design and optimization workflow.<\/p>\n<p class=\"wp-block-paragraph\">\u201cOur collaboration with OpenAI represents a fundamental commitment to scaling the physical infrastructure required for the next decade of AI,\u201d said Hock Tan, President and CEO of Broadcom.<\/p>\n<p class=\"wp-block-paragraph\">The companies plan to deploy the accelerator at gigawatt-scale <a href=\"https:\/\/interestingengineering.com\/science\/nvidias-servers-slash-data-center-energy\" target=\"_blank\" rel=\"dofollow noopener\">data centers<\/a> beginning in 2026. Jalape\u00f1o is the first product in what OpenAI describes as a multi-generation compute platform that will combine OpenAI-designed accelerators with Broadcom networking and connectivity technologies and Celestica\u2019s system integration expertise.<\/p>\n<p class=\"wp-block-paragraph\"><a href=\"https:\/\/interestingengineering.com\/culture\/openai-confidential-ipo-trillion-dollar-ai-race\" target=\"_blank\" rel=\"dofollow noopener\">OpenAI<\/a> said improvements in inference efficiency could translate into faster ChatGPT responses, lower AI operating costs, and more reliable access to advanced AI services as demand continues to grow.<\/p>\n","protected":false},"excerpt":{"rendered":"OpenAI has unveiled its first custom AI accelerator, called Jalape\u00f1o, marking the company\u2019s move into chip design as&hellip;\n","protected":false},"author":2,"featured_media":85180,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[7],"tags":[12529,407,328,8333,39,45612,45788,157],"class_list":["post-85179","post","type-post","status-publish","format-standard","has-post-thumbnail","category-openai","tag-ai-accelerator","tag-ai-infrastructure","tag-broadcom","tag-custom-silicon","tag-data-centers","tag-jalapeno-chip","tag-llm-inference","tag-openai"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/85179","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=85179"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/85179\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/85180"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=85179"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=85179"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=85179"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}