flâneur — a map of the web's best reading

Another Giant Leap: The Rubin CPX Specialized Accelerator & Rack – SemiAnalysis

semianalysis.com · 7,177 words · saved by 1 readers

Nvidia announced the Rubin CPX, a solution that is specifically designed to be optimized for the prefill phase, with the single-die Rubin CPX heavily emphasizing compute FLOPS over memory bandwidth. This is a game changer for inference, and its significance is surpassed only by the March 2024 announcement of the GB200 NVL72 Oberon rack-scale form factor. Only with hardware specialized to the very different phases of inference, prefill and decode, can disaggregated serving achieve its full potential. As a result, the rack system design gap between Nvidia and its competitors has become canyon-sized. AMD and custom silicon competitors may have made a small step forward in emulating Nvidia’s 72-GPU rack scale design, but Nvidia has just made another Giant Leap, again leaving competitors very distant objects in the rear-view mirror. AMD and ASIC providers have already been investing heavily to catch up in terms of their own rack-scale solutions. AMD in particular has been working tirelessly

Another Giant Leap: The Rubin CPX Specialized Accelerator & Rack – SemiAnalysis Skip to content September 10, 2025 Another Giant Leap: The Rubin CPX Specialized Accelerator & Rack New Prefill Specialized GPU, Rack Architecture, BOM, Disaggregated PD, Higher Perf per TCO, Lower TCO, GDDR7 & HBM Market Trends 23 minutes 2 comments on Another Giant Leap: The Rubin CPX Specialized Accelerator & Rack By Dylan Patel , Daniel Nishball , Kimbo Chen , Myron Xie , Wega Chu , Gerald Wong , Kang Wen Cheang and Ivan Chiam Share using Native tools Share Copied to clipboard Share on LinkedIn (Opens in

Explore this link on the map →

saved by

related reading