{"id":382420,"date":"2026-04-09T02:16:23","date_gmt":"2026-04-09T02:16:23","guid":{"rendered":"https:\/\/www.newsbeep.com\/il\/382420\/"},"modified":"2026-04-09T02:16:23","modified_gmt":"2026-04-09T02:16:23","slug":"intel-and-sambanova-team-up-on-heterogenous-ai-inference-platform-different-hardware-performs-different-workloads","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/il\/382420\/","title":{"rendered":"Intel and SambaNova team up on heterogenous AI inference platform \u2014 different hardware performs different workloads"},"content":{"rendered":"<p id=\"elk-aa011808-3a94-4fd1-bf02-d89e40d5ede2\">Intel and SambaNova on Wednesday <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.businesswire.com\/news\/home\/20260408117878\/en\/SambaNova-and-Intel-Announce-Blueprint-for-Heterogeneous-Inference-GPUs-for-Prefill-SambaNova-RDUs-for-Decode-and-Intel-Xeon-6-CPUs-for-Agentic-Tools\" data-url=\"https:\/\/www.businesswire.com\/news\/home\/20260408117878\/en\/SambaNova-and-Intel-Announce-Blueprint-for-Heterogeneous-Inference-GPUs-for-Prefill-SambaNova-RDUs-for-Decode-and-Intel-Xeon-6-CPUs-for-Agentic-Tools\" target=\"_blank\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" rel=\"nofollow noopener\">announced<\/a> their joint production-ready heterogeneous inference architecture that relies on AI accelerators or GPUs for prefill, <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/tech-industry\/artificial-intelligence\/sambanova-introduces-new-ai-accelerator-partners-with-intel-to-deploy-xeon-cpus-for-inferencing-and-agentic-workloads-sambanova-claims-sn50-chip-is-three-times-more-efficient-than-nvidia-b200\" data-url=\"https:\/\/www.tomshardware.com\/tech-industry\/artificial-intelligence\/sambanova-introduces-new-ai-accelerator-partners-with-intel-to-deploy-xeon-cpus-for-inferencing-and-agentic-workloads-sambanova-claims-sn50-chip-is-three-times-more-efficient-than-nvidia-b200\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/tech-industry\/artificial-intelligence\/sambanova-introduces-new-ai-accelerator-partners-with-intel-to-deploy-xeon-cpus-for-inferencing-and-agentic-workloads-sambanova-claims-sn50-chip-is-three-times-more-efficient-than-nvidia-b200\" rel=\"nofollow noopener\" target=\"_blank\">SambaNova reconfigurable dataflow units (RDUs) SN50<\/a> for decode, and Xeon 6 processors for agentic tools and system orchestration. The platform is designed to address as broad a set of workloads as possible to siphon some of the market share away from Nvidia and other emerging players.<\/p>\n<p>Go deeper with TH Premium: AI and data centers<\/p>\n<p class=\"vanilla-image-block\" style=\"padding-top:56.25%;\">\n<p><img decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2025\/12\/Vh4nY3pMCcmra2ymXah9S7.jpg\" alt=\"Microsoft data center in Mount Pleasant, Wisconsin\"   loading=\"lazy\" data-new-v2-image=\"true\" data-original-mos=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2025\/12\/Vh4nY3pMCcmra2ymXah9S7.jpg\" data-pin-media=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2025\/12\/Vh4nY3pMCcmra2ymXah9S7.jpg\" class=\"pinterest-pin-exclude\"\/>\n<\/p>\n<p>(Image credit: Microsoft)<\/p>\n<p id=\"elk-3ec29ac1-814d-4bc3-88d5-f5645a2d1c4c\" class=\"paywall\" aria-hidden=\"true\">The heterogeneous inference platform by Intel and SambaNova separates inference into distinct stages handled by different silicon: It uses AI GPUs or AI accelerators for ingesting long prompts and building key-value caches; SambaNova&#8217;s SN50 RDU for decoding and generating tokens; and Xeon 6 processors for running agent-related operations (e.g., compiling and executing code and validating outputs) as well as coordinating and distributing workloads across hardware. <\/p>\n<p>Splitting prefill, decode, and token generation stages is similar to Nvidia&#8217;s approach to its Rubin platform, which is based on the <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/nvidias-new-cpx-gpu-aims-to-change-the-game-in-ai-inference-how-the-debut-of-cheaper-and-cooler-gddr7-memory-could-redefine-ai-inference-infrastructure\" data-url=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/nvidias-new-cpx-gpu-aims-to-change-the-game-in-ai-inference-how-the-debut-of-cheaper-and-cooler-gddr7-memory-could-redefine-ai-inference-infrastructure\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/nvidias-new-cpx-gpu-aims-to-change-the-game-in-ai-inference-how-the-debut-of-cheaper-and-cooler-gddr7-memory-could-redefine-ai-inference-infrastructure\" rel=\"nofollow noopener\" target=\"_blank\">Rubin CPX<\/a> and heavy-duty Rubin GPU with HBM4 memory \u2014 with the obvious difference that <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/nvidia-removes-rubin-cpx-accelerators-from-its-roadmap-groq-3-lpus-take-center-stage-as-cpx-is-removed\" data-url=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/nvidia-removes-rubin-cpx-accelerators-from-its-roadmap-groq-3-lpus-take-center-stage-as-cpx-is-removed\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/pc-components\/gpus\/nvidia-removes-rubin-cpx-accelerators-from-its-roadmap-groq-3-lpus-take-center-stage-as-cpx-is-removed\" rel=\"nofollow noopener\" target=\"_blank\">the Rubin CPX is not coming to market<\/a>. But, more importantly for Intel, the new platform will rely on its Xeon 6 processors \u2014 not on competing offerings.<\/p>\n<p class=\"vanilla-image-block\" style=\"padding-top:56.67%;\">\n<p><img decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2026\/04\/zAbAyUGWkTSc2s8VaNnYdT.jpg\" alt=\"SambaNova\"   loading=\"lazy\" data-new-v2-image=\"true\" data-original-mos=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2026\/04\/zAbAyUGWkTSc2s8VaNnYdT.jpg\" data-pin-media=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2026\/04\/zAbAyUGWkTSc2s8VaNnYdT.jpg\" class=\"inline\"\/>\n<\/p>\n<p>(Image credit: SambaNova)<\/p>\n<p id=\"elk-72058ca9-c1a4-4664-8313-21c4af0046b5\">The solution is scheduled to be available in the second half of 2026 to enterprises, cloud operators, and sovereign AI programs seeking scalable inference platforms in general and coding agents, and other agentic workloads in particular, completely in-house.<\/p>\n<p>According to SambaNova&#8217;s internal data, Xeon 6 achieves over 50% faster LLVM compilation compared to Arm-based server CPUs, and delivers up to 70% higher performance in vector database workloads, relative to competing x86 processors \u2014 namely, <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/pc-components\/cpus\/amd-launches-epyc-turin-9005-series-our-benchmarks-of-fifth-gen-zen-5-chips-with-up-to-192-cores-500w-tdp\" data-url=\"https:\/\/www.tomshardware.com\/pc-components\/cpus\/amd-launches-epyc-turin-9005-series-our-benchmarks-of-fifth-gen-zen-5-chips-with-up-to-192-cores-500w-tdp\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/pc-components\/cpus\/amd-launches-epyc-turin-9005-series-our-benchmarks-of-fifth-gen-zen-5-chips-with-up-to-192-cores-500w-tdp\" rel=\"nofollow noopener\" target=\"_blank\">AMD EPYC<\/a>. These gains are intended to shorten end-to-end development cycles for coding agents and similar applications, the two companies claim.<\/p>\n<p>Perhaps the biggest advantage of the joint production-ready heterogeneous inference architecture is that SambaNova SN50 and Xeon-based servers are drop-in compatible with data centers that can handle 30kW \u2014 which is the vast majority of <a data-analytics-id=\"inline-link\" href=\"https:\/\/www.tomshardware.com\/tag\/enterprise\" data-auto-tag-linker=\"true\" data-url=\"https:\/\/www.tomshardware.com\/tag\/enterprise\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" data-before-rewrite-localise=\"https:\/\/www.tomshardware.com\/tag\/enterprise\" rel=\"nofollow noopener\" target=\"_blank\">enterprise<\/a> data centers. <\/p>\n<p>&#8220;The data center software ecosystem is built on x86, and it runs on Xeon \u2014 providing a mature, proven foundation that developers, enterprises, and cloud providers rely on at scale,&#8221; said Kevork Kechichian, Executive Vice President and General Manager of the Data Center Group (DCG) at Intel Corporation. &#8220;Workloads of the future will require a heterogeneous mix of computing, and this collaboration with SambaNova delivers a cost\u2011efficient, high\u2011performance inference architecture designed to meet customer needs at scale \u2014 powered by Xeon 6.&#8221;<\/p>\n<p>Article continues below <\/p>\n<p>            You may like<\/p>\n<p>    <a href=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" id=\"elk-1254cb0e-8a84-43c8-86a8-b3b6c2d99d1e\" data-url=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" target=\"_blank\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" rel=\"nofollow noopener\"><\/p>\n<p class=\"vanilla-image-block\" style=\"padding-top:31.51%;\">\n<p><img decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2025\/10\/7cUTDmN2PHNRiNBVqbKf56.png\" alt=\"Google Preferred Source\"   loading=\"lazy\" data-new-v2-image=\"true\" data-original-mos=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2025\/10\/7cUTDmN2PHNRiNBVqbKf56.png\" data-pin-media=\"https:\/\/www.newsbeep.com\/il\/wp-content\/uploads\/2025\/10\/7cUTDmN2PHNRiNBVqbKf56.png\" class=\"pull-left\"\/>\n<\/p>\n<p><\/a><\/p>\n<p id=\"elk-b01016a9-c0cd-441f-a37c-628d0d4c810c\">Follow<a data-analytics-id=\"inline-link\" href=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" target=\"_blank\" data-url=\"https:\/\/news.google.com\/publications\/CAAqLAgKIiZDQklTRmdnTWFoSUtFSFJ2YlhOb1lYSmtkMkZ5WlM1amIyMG9BQVAB\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" rel=\"nofollow noopener\"> Tom&#8217;s Hardware on Google News<\/a>, or<a data-analytics-id=\"inline-link\" href=\"https:\/\/google.com\/preferences\/source?q=\" target=\"_blank\" data-url=\"https:\/\/google.com\/preferences\/source?q=\" referrerpolicy=\"no-referrer-when-downgrade\" data-hl-processed=\"none\" data-mrf-recirculation=\"inline-link\" rel=\"nofollow noopener\"> add us as a preferred source<\/a>, to get our latest news, analysis, &amp; reviews in your feeds.<\/p>\n<p><a id=\"elk-seasonal\" class=\"paywall\" aria-hidden=\"true\"\/><\/p>\n<p class=\"newsletter-form__strapline\">Get Tom&#8217;s Hardware&#8217;s best news and in-depth reviews, straight to your inbox.<\/p>\n","protected":false},"excerpt":{"rendered":"Intel and SambaNova on Wednesday announced their joint production-ready heterogeneous inference architecture that relies on AI accelerators or&hellip;\n","protected":false},"author":2,"featured_media":382421,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[345,343,344,85,46,125],"class_list":["post-382420","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-il","tag-israel","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/382420","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/comments?post=382420"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/posts\/382420\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media\/382421"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/media?parent=382420"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/categories?post=382420"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/il\/wp-json\/wp\/v2\/tags?post=382420"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}