INFRASTRUCTURE

Rubin

infrastructuretopic-notenvidia-platform

Overview

Rubin is NVIDIA’s next-generation AI computing platform, following Blackwell, announced for broad cloud distribution starting H2 2026. The platform integrates six chips — Vera CPU, Rubin GPU, NVLink 6 switch, ConnectX-9 SuperNIC, BlueField-4 DPU, and Spectrum-6 ethernet switch — positioned as the primary infrastructure for large-scale AI training and inference deployments across major clouds and neoclouds.

Timeline

  • 2026-05-05-AI-DigestNVIDIA formally opened the Rubin platformsix new chips spanning Vera CPU, Rubin GPU, NVLink 6 switch, ConnectX-9 SuperNIC, BlueField-4 DPU, Spectrum-6 ethernet switch — for distribution starting H2 2026 across AWS, Google Cloud, Microsoft Azure, Oracle Cloud, plus the neocloud tier (CoreWeave, Lambda, Nebius, Nscale). Headline performance claims versus Blackwell: 3.5× training throughput, 5× inference throughput, 8× power efficiency. Microsoft’s Fairwater data centre sites in Wisconsin and Atlanta reported as already operating Vera Rubin NVL72 racks. Distribution piece closed; first GA price point remains open.

Key Developments

  1. Six-Chip Integrated Platform: Unlike Blackwell’s primary GPU-centric design, Rubin bundles compute, switching, networking, and DPU infrastructure as a unified system, reducing operational complexity for hyperscalers.

  2. 3.5×/5×/8× Performance Claims: Training throughput, inference throughput, and power efficiency gains versus Blackwell establish clear generational improvement narrative.

  3. Immediate Production Deployment: Microsoft’s Fairwater sites already running Vera Rubin NVL72 racks de-risks cloud-provider adoption timelines and signals high customer confidence.

  4. Broad Cloud Distribution: Confirmed partnerships across all major clouds (AWS, GCP, Azure, OCI) plus neocloud tier (CoreWeave, Lambda, Nebius, Nscale) establish Rubin as the industry standard for H2 2026 new AI workload deployments.

  5. Deferred Pricing Disclosure: Despite confirmed distribution partnerships, first customer-facing price points remain unpublished, keeping downside customer-acquisition risk open through H2 2026.

Market Position

Rubin’s distribution announcement closes a critical loop in the Blackwell → Rubin generational shift narrative tracked since March 2026. The platform’s integration of switching, networking, and DPU alongside compute reflects NVIDIA’s effort to capture the full inference-infrastructure stack rather than just GPUs, while the immediate Microsoft production deployment suggests cloud providers are sufficiently confident in the roadmap to commit capacity before final pricing.

See Also

Timeline (continued)

  • 2026-07-29-AI-DigestVera Rubin becomes the compute anchor of Nvidia‘s $5B strategic-partnership equity into Safe Superintelligence — SSI gets access to Vera Rubin CPU-GPU systems for reportedly a ~10× compute increase over its prior stack, with SSI’s cap table sitting at ~$7B raised at ~$32B post-money. Deal reportedly includes rare research-access rights for both Nvidia and Alphabet as part of consideration. The TPU-to-GPU switch is inferred from the pre-existing Google Cloud arrangement being superseded — the Nvidia press release confirms Vera Rubin access and the ~10× compute jump but doesn’t explicitly name TPU displacement. Corpus reads this as vendor-financed frontier-lab silicon lock-in: Nvidia writing nine-to-ten-figure equity checks to anchor Vera Rubin as the compute substrate under a fresh frontier-safety-labelled lab. First frontier-lab strategic-partnership Vera Rubin allocation the corpus tracks after the H2 2026 GA distribution announcement in 2026-05-05-AI-Digest — extends the platform-adoption story from cloud-provider distribution into direct frontier-lab equity-and-access packaging.

  • 2026-09-04-AI-DigestVera Rubin GPUs anchor Nscale‘s new $3.5B six-year compute-services contract with humanoid-robotics firm Figure — deployed at a Barstow, TX site starting H2 2027, with intent to scale toward a $6B envelope and up to 100,000 Vera Rubin GPUs, plus a separate undisclosed equity investment by Nscale. The $3.5B is a compute-services contract (take-or-pay in shape), not equity or investment financing — the two flows are structurally distinct. Extends the H2 2026 GA distribution announcement from 2026-05-05-AI-Digest and the Anthropic-Nscale $45B / 460 MW West Virginia forward-compute deal (2026-08-27-AI-Digest) with a second frontier-tenant Vera Rubin allocation — the NVIDIA-centered circular-financing pattern hardening across 2026 continues to accrete Vera Rubin capacity on the compute-supplier side.

Key Developments (continued)

  1. Vera Rubin as Anchor of Nvidia’s $5B SSI Strategic Partnership (July 29, 2026): Vera Rubin CPU-GPU systems are the compute substrate under Nvidia’s up-to-$5B equity investment into Safe Superintelligence, with a reported ~10× compute increase over SSI’s prior stack and SSI cap table at ~$7B raised / ~$32B post-money. Deal reportedly includes rare research-access rights for both Nvidia and Alphabet. TPU-to-GPU switch is inferred from the superseded Google Cloud arrangement; the Nvidia press release confirms Vera Rubin access and the compute jump but doesn’t explicitly name TPU displacement. First direct frontier-lab strategic-partnership Vera Rubin allocation the corpus tracks after the H2 2026 GA cloud distribution — extends the platform-adoption story from cloud-provider distribution into direct frontier-lab equity-and-access packaging.