INFRASTRUCTURE
Rubin
Overview
Rubin is NVIDIA’s next-generation AI computing platform, following Blackwell, announced for broad cloud distribution starting H2 2026. The platform integrates six chips — Vera CPU, Rubin GPU, NVLink 6 switch, ConnectX-9 SuperNIC, BlueField-4 DPU, and Spectrum-6 ethernet switch — positioned as the primary infrastructure for large-scale AI training and inference deployments across major clouds and neoclouds.
Timeline
- 2026-05-05-AI-Digest — NVIDIA formally opened the Rubin platform — six new chips spanning Vera CPU, Rubin GPU, NVLink 6 switch, ConnectX-9 SuperNIC, BlueField-4 DPU, Spectrum-6 ethernet switch — for distribution starting H2 2026 across AWS, Google Cloud, Microsoft Azure, Oracle Cloud, plus the neocloud tier (CoreWeave, Lambda, Nebius, Nscale). Headline performance claims versus Blackwell: 3.5× training throughput, 5× inference throughput, 8× power efficiency. Microsoft’s Fairwater data centre sites in Wisconsin and Atlanta reported as already operating Vera Rubin NVL72 racks. Distribution piece closed; first GA price point remains open.
Key Developments
-
Six-Chip Integrated Platform: Unlike Blackwell’s primary GPU-centric design, Rubin bundles compute, switching, networking, and DPU infrastructure as a unified system, reducing operational complexity for hyperscalers.
-
3.5×/5×/8× Performance Claims: Training throughput, inference throughput, and power efficiency gains versus Blackwell establish clear generational improvement narrative.
-
Immediate Production Deployment: Microsoft’s Fairwater sites already running Vera Rubin NVL72 racks de-risks cloud-provider adoption timelines and signals high customer confidence.
-
Broad Cloud Distribution: Confirmed partnerships across all major clouds (AWS, GCP, Azure, OCI) plus neocloud tier (CoreWeave, Lambda, Nebius, Nscale) establish Rubin as the industry standard for H2 2026 new AI workload deployments.
-
Deferred Pricing Disclosure: Despite confirmed distribution partnerships, first customer-facing price points remain unpublished, keeping downside customer-acquisition risk open through H2 2026.
Market Position
Rubin’s distribution announcement closes a critical loop in the Blackwell → Rubin generational shift narrative tracked since March 2026. The platform’s integration of switching, networking, and DPU alongside compute reflects NVIDIA’s effort to capture the full inference-infrastructure stack rather than just GPUs, while the immediate Microsoft production deployment suggests cloud providers are sufficiently confident in the roadmap to commit capacity before final pricing.
See Also
- NVIDIA — the vendor
- Blackwell — the predecessor platform
- MOC - AI Infrastructure
Timeline (continued)
-
2026-07-29-AI-Digest — Vera Rubin becomes the compute anchor of Nvidia‘s $5B strategic-partnership equity into Safe Superintelligence — SSI gets access to Vera Rubin CPU-GPU systems for reportedly a ~10× compute increase over its prior stack, with SSI’s cap table sitting at ~$7B raised at ~$32B post-money. Deal reportedly includes rare research-access rights for both Nvidia and Alphabet as part of consideration. The TPU-to-GPU switch is inferred from the pre-existing Google Cloud arrangement being superseded — the Nvidia press release confirms Vera Rubin access and the ~10× compute jump but doesn’t explicitly name TPU displacement. Corpus reads this as vendor-financed frontier-lab silicon lock-in: Nvidia writing nine-to-ten-figure equity checks to anchor Vera Rubin as the compute substrate under a fresh frontier-safety-labelled lab. First frontier-lab strategic-partnership Vera Rubin allocation the corpus tracks after the H2 2026 GA distribution announcement in 2026-05-05-AI-Digest — extends the platform-adoption story from cloud-provider distribution into direct frontier-lab equity-and-access packaging.
-
2026-09-04-AI-Digest — Vera Rubin GPUs anchor Nscale‘s new $3.5B six-year compute-services contract with humanoid-robotics firm Figure — deployed at a Barstow, TX site starting H2 2027, with intent to scale toward a $6B envelope and up to 100,000 Vera Rubin GPUs, plus a separate undisclosed equity investment by Nscale. The $3.5B is a compute-services contract (take-or-pay in shape), not equity or investment financing — the two flows are structurally distinct. Extends the H2 2026 GA distribution announcement from 2026-05-05-AI-Digest and the Anthropic-Nscale $45B / 460 MW West Virginia forward-compute deal (2026-08-27-AI-Digest) with a second frontier-tenant Vera Rubin allocation — the NVIDIA-centered circular-financing pattern hardening across 2026 continues to accrete Vera Rubin capacity on the compute-supplier side.
Key Developments (continued)
- Vera Rubin as Anchor of Nvidia’s $5B SSI Strategic Partnership (July 29, 2026): Vera Rubin CPU-GPU systems are the compute substrate under Nvidia’s up-to-$5B equity investment into Safe Superintelligence, with a reported ~10× compute increase over SSI’s prior stack and SSI cap table at ~$7B raised / ~$32B post-money. Deal reportedly includes rare research-access rights for both Nvidia and Alphabet. TPU-to-GPU switch is inferred from the superseded Google Cloud arrangement; the Nvidia press release confirms Vera Rubin access and the compute jump but doesn’t explicitly name TPU displacement. First direct frontier-lab strategic-partnership Vera Rubin allocation the corpus tracks after the H2 2026 GA cloud distribution — extends the platform-adoption story from cloud-provider distribution into direct frontier-lab equity-and-access packaging.