{"id":113920,"date":"2026-07-21T21:02:23","date_gmt":"2026-07-21T21:02:23","guid":{"rendered":"https:\/\/www.europesays.com\/ai\/113920\/"},"modified":"2026-07-21T21:02:23","modified_gmt":"2026-07-21T21:02:23","slug":"benchmark-study-from-deepinfra-confirms-nvidia-vera-cpu-outperforms-leading-cpus-on-agentic-workloads","status":"publish","type":"post","link":"https:\/\/www.europesays.com\/ai\/113920\/","title":{"rendered":"Benchmark Study from DeepInfra Confirms NVIDIA Vera CPU Outperforms Leading CPUs on Agentic Workloads"},"content":{"rendered":"<p>    <a href=\"https:\/\/s.yimg.com\/lo\/mysterio\/api\/9E233089343DB58460B95FE9F4983A7AC032433C7B0DD2EC8D587D097F8281A4\/subgraphmysterio\/resizefit_w960;quality_80;format_webp\/https:%2F%2Fmedia.zenfs.com%2Fen%2Fglobenewswire.com%2F73c5726906987d79230a3d1932fa8fab\" target=\"_blank\" rel=\"noopener noreferrer nofollow\"><img loading=\"lazy\" decoding=\"async\" src=\"data:image\/gif;base64,R0lGODlhAQABAIAAAAAAAP\/\/\/ywAAAAAAQABAAACAUwAOw==\" alt=\"DeepInfra\" height=\"109\" width=\"300\" class=\"yf-lf2kr7 loader\"\/><\/a> DeepInfra           <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Production-scale testing on DeepInfra&#8217;s agent infrastructure showed the NVIDIA Vera CPU hosting up to 1.6x more concurrent AI agents at the same quality of service  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">PALO ALTO, Calif., July 21, 2026 <a href=\"https:\/\/www.globenewswire.com\" data-ylk=\"slk:(GLOBE%20NEWSWIRE);elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;(GLOBE NEWSWIRE)&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">(GLOBE NEWSWIRE)<\/a> &#8212; <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=OhsR6PgUV_nPyPoTnvaWlxclCy4U2huaCBI1QXh2O9ki4L2K1FMw5M2PapSsp1ptFwwZIf4vwJuqgAUgzB0tLA==\" data-ylk=\"slk:DeepInfra;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;DeepInfra&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">DeepInfra<\/a>, a purpose-built cloud platform for high-throughput AI inference, today announced the results of an independent benchmark of the <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=WJbfBDRyizVNd1WT4fZsvPS12WdNCWtPd5Cl_hnFNYO0ExYAIANSz0EoEeQQ1fKl8KGG4VZQz384WLRFM9nPw1NTMi-zQNy162Yjz1CC3I7KrIZwD1kMUD8Rv333UdDC\" data-ylk=\"slk:NVIDIA%20Vera%20CPU;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;NVIDIA Vera CPU&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">NVIDIA Vera CPU<\/a>, NVIDIA&#8217;s next generation CPU purpose-built to power agentic workloads. As one of a select group of collaborators granted early access to Vera hardware through NVIDIA&#8217;s open AI ecosystem, DeepInfra designed and ran its own production-grade benchmark to test the infrastructure with a real-world agent workload. The benchmark found Vera led by up to 2.2x against competitive CPUs.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">DeepInfra processes nearly five trillion tokens per week, with close to 30% driven by agentic systems. At that scale, CPU performance is critical to delivering the cost efficiency, low latency, and throughput that agentic workloads demand.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;We built DeepInfra&#8217;s infrastructure from the ground up for inference. We have tuned every layer of it for cost, latency, and throughput at production scale,&#8221; said Nikola Borisov, co-founder and CEO, DeepInfra. &#8220;NVIDIA is building for where AI workloads are headed, and we believe Vera is exactly the kind of hardware solution the next iteration demands.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">DeepInfra tested Vera using its own production AI agent and real captured traffic, benchmarking against three leading CPUs from AMD, and Intel under identical conditions.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Key results from the benchmark include:  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Fastest in every workload category, doubling NVIDIA&#8217;s claim: Against NVIDIA&#8217;s published claim of 80% faster agentic CPU performance, Vera measured up to 2.2x faster orchestration than the x86 baseline &#8211; and was the fastest of all four architectures tested in every workload category.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Up to 1.6x more concurrent agents at the same quality of service: On identical CPU partitions, Vera sustained 256 concurrent agents within strict response-time and error targets, versus 160\u2013192 for the competing chips.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Spare capacity served a full LLM: With all 256 agents running at full load, Vera&#8217;s leftover cores simultaneously served a 20-billion-parameter open-source model faster than an entire previous-generation CPU socket dedicated solely to that task, without slowing the agents.<\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">&#8220;Bringing Vera to market is an incredible opportunity, but performance is ultimately proven in production,&#8221; said Ian Finder, Director, Data Center CPU Products, NVIDIA. &#8220;DeepInfra is pushing Vera with demanding, real-world agentic workloads that reflect the needs of production environments. These results demonstrate exactly what Vera was designed to deliver.&#8221;  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">DeepInfra also published a technical analysis with the full benchmark methodology and complete results, <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=HM2_dztmyWkl43JybGrPL0Vo25U7L29FtCohsey_jiFrzubuDGon-5BzkDUCxzGr8OSC0sIbcXQeADyqQWkdBdxbHO0fTxGAy74DGB761-3N7Aj8mp2h7xHXwJINPc9l\" data-ylk=\"slk:here;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;here&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">here<\/a>.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">To learn more, visit <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=cU-WK_SC53zum53XFFVTlEKhgtS0VyrGWD1hGynWwqZ0vswX1K32rTBKUCJJcigquIwWUGo2bChjN2H2hGafgw==\" data-ylk=\"slk:deepinfra.com;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;deepinfra.com&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">deepinfra.com<\/a>.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">About DeepInfra<br \/>Founded in 2022, DeepInfra is a purpose-built cloud inference platform for high-throughput AI, enabling companies to run open-source and proprietary AI and agentic models at scale. The company owns and operates its AI Compute infrastructure, supports 200+ open-source models, and processes nearly five trillion tokens per week \u2014 delivering the cost efficiency, low latency, and security that production AI demands. Its fully managed platform includes OpenAI-compatible APIs and enterprise-grade security compliance, including zero data retention and SOC 2 and ISO 27001 certification. For more information, visit <a href=\"https:\/\/www.globenewswire.com\/Tracker?data=cU-WK_SC53zum53XFFVTlB_SgHujqSBYHJnPOpaSGTcleRxi0Mp9JVDBrSTsEgU0Fv9v2mMzx-Gn91RQzpUBiw==\" data-ylk=\"slk:deepinfra.com;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;deepinfra.com&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">deepinfra.com<\/a>.  <\/p>\n<p class=\"text text-block paragraph text-left neo-font-paragraph-xl-reg  yf-18d6y07\" style=\"text-decoration: none; font-style: normal; text-transform: none; text-align: inherit; font-variant-numeric: normal;\">Media Contact:<br \/>Julianna Sheridan<br \/><a href=\"https:\/\/www.globenewswire.com\/Tracker?data=tDd3qJORze8dtsr5LvBYtz-1LCpZmrnPqPfa-OhFr0Wsc2V3oSL-QuNTt7IfyXhv1_ww3QUD8i-ASf-gLymSzu9774ji0xIjouNfsFwgx4Y=\" data-ylk=\"slk:press%40deepinfra.com;elm:context_link;itc:0;sec:content-canvas;source:content-canvas%20default\" data-yga=\"{&quot;yLinkText&quot;:&quot;press@deepinfra.com&quot;,&quot;yLinkElement&quot;:&quot;context_link&quot;,&quot;yModuleName&quot;:&quot;content-canvas&quot;,&quot;yTrafficOrigin&quot;:&quot;content-canvas default&quot;}\" target=\"_blank\" rel=\"noopener noreferrer nofollow\">press@deepinfra.com<\/a>  <\/p>\n","protected":false},"excerpt":{"rendered":"DeepInfra Production-scale testing on DeepInfra&#8217;s agent infrastructure showed the NVIDIA Vera CPU hosting up to 1.6x more concurrent&hellip;\n","protected":false},"author":2,"featured_media":113921,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[6],"tags":[179,7493,17415,23670,58,33316],"class_list":["post-113920","post","type-post","status-publish","format-standard","has-post-thumbnail","category-agentic-ai","tag-agentic-ai","tag-agentic-artificial-intelligence","tag-cost-efficiency","tag-cpu-performance","tag-nvidia","tag-vera"],"_links":{"self":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/113920","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/comments?post=113920"}],"version-history":[{"count":0,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/posts\/113920\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media\/113921"}],"wp:attachment":[{"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/media?parent=113920"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/categories?post=113920"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.europesays.com\/ai\/wp-json\/wp\/v2\/tags?post=113920"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}