The Information Machine

NVIDIA Launches Vera CPU and Vera Rubin NVL72 at COMPUTEX / GTC Taipei · history

Version 16

2026-06-11 08:28 UTC · 280 items

What

Bloomberg confirmed on June 5 that NVIDIA's CEO certified all three major HBM4 suppliers — Samsung, SK Hynix, and Micron — for Vera Rubin [6], resolving the most contested supply-chain question about the platform. Analyst Rohan Paul subsequently reported market share detail: SK Hynix at approximately 60-70%, Samsung at 25-30%, and Micron with the remainder, all in full production [7]. Rack-level validation has cleared L11 at both Dell/CoreWeave (May 31) [3] and Microsoft/Foxconn (June 1) [4], while L12 cluster validation and the start of rack-level mass production remain unconfirmed.

Why it matters

CEO-level confirmation of three-vendor HBM4 supply removes the most frequently cited constraint on Vera Rubin deployment pace. The remaining structural binding on large-scale deployment is the 600kW per-rack power requirement, which mandates greenfield data center construction and is independent of memory supply [1].

Open questions

  • L12 cluster validation — the first full compute cluster with scale-out networking — has not been reported for either CoreWeave or Microsoft; when does the first L12 milestone occur?

  • Jensen described Vera Rubin as 'in full production' on June 5 [12][6], but at COMPUTEX on June 1 stated rack-level mass production had not yet begun [4] — when does rack-level mass production actually begin, and does this gap create delivery risk against H2 2026 commitments?

  • The SemiAnalysis Vera SOCAMM report attracted accusations of inaccuracy and was defended with physical evidence from the SK Hynix Computex booth [13] — what exactly does the report claim, and what are the architectural or supply-chain implications if correct?

Narrative

NVIDIA's Vera Rubin NVL72 is a high-density AI compute rack integrating 72 Rubin GPUs interconnected via NVLink at the rack level, requiring 600kW per rack and HBM4 memory [1]. The platform reached the first rack-level validation milestone on May 31, when Dell delivered a fully validated VR200 NVL72 rack to CoreWeave, confirming the L11 scale-up domain operational [2][3]. Microsoft completed bring-up of its first Rubin NVL72 via Foxconn on June 1 at COMPUTEX [4]. NVIDIA's validation hierarchy runs from L10 (single-server firmware) through L11 (single-rack scale-up domain) to L12 (full compute cluster with scale-out networking) [5]; neither operator has reported clearing L12.

The central supply-chain question — whether NVIDIA qualified two or three HBM4 suppliers — was answered when Bloomberg reported on June 5 that the NVIDIA CEO confirmed certification of Samsung, SK Hynix, and Micron to supply HBM4 for Vera Rubin [6]. Analyst Rohan Paul subsequently reported market share: SK Hynix at 60-70%, Samsung at 25-30%, and Micron with the remainder, all in full production [7]. This contradicts prior independent analyses that concluded only Samsung and SK Hynix had been designated [8][9], and broadens the effective supply ceiling. Prior reporting had Samsung selling out its entire 2026 HBM4 allocation with rack prices at $8.8M and shortage projections through 2028 [10][11]; the three-vendor picture may revise those projections, though the 600kW per-rack power requirement remains a binding infrastructure constraint independent of memory supply [1].

During the same Seoul visit where Bloomberg captured the three-vendor certification statement, Jensen Huang described Vera Rubin as 'in full production' [12] — language that sits in tension with his COMPUTEX statement four days earlier that wafer-level production had started but rack-level mass production had not yet begun [4]. The distinction between wafer-level and rack-level production is material for delivery timelines against H2 2026 commitments. A separate contested claim comes from SemiAnalysis, which published a report on 'Vera SOCAMM' — a memory module format — that attracted accusations of inaccuracy; SemiAnalysis defended it by citing physical evidence at the SK Hynix Computex booth [13]. The full content and architectural implications of the SOCAMM report remain incompletely described in available sources.

Deployment commitments include Nscale (130,000 Rubin GPUs for Microsoft, with NVIDIA holding approximately $674M in Nscale equity) [14][15] and Verda across Europe, the US, and APAC. SemiAnalysis has rated the GTC Taipei keynote 'F tier' and documents that Rubin FP4/FP8 gains are approximately 3.5x over GB200 while FP16 gains are approximately 1.6x and HBM capacity is flat [16] — making stated efficiency advantages workload-specific rather than uniform.

Timeline

  • 2026-01-05: NVIDIA debuts Rubin chip at CES: 336 billion transistors, 50 petaflops AI performance. [31]
  • 2026-01: Jensen Huang announces at CES 2026 that Vera Rubin NVL72 is in full production. [32][33]
  • 2026-02: SK Hynix begins HBM4 mass production shipments to NVIDIA, holding approximately 70% of NVIDIA's HBM4 orders. [34][28]
  • 2026-03-17: Nscale acquires 8GW Monarch Compute Campus in West Virginia; Microsoft signs 1.35GW LOI co-announced with NVIDIA and Caterpillar. [35][36][23][37]
  • 2026-05: Samsung sells out entire 2026 HBM4 supply; rack prices reach $8.8M; shortage projected until 2028. [38][11][39][10]
  • 2026-05: NVL72's 600kW per-rack power requirement documented as incompatible with existing data centers, requiring greenfield construction. [1][29][30]
  • 2026-05: NVIDIA equity stake in Nscale confirmed at approximately $674M; Microsoft's Rubin GPU deployment via Nscale revised upward to 130,000 units. [40][15][14]
  • 2026-05-18: First Vera CPUs hand-delivered to OpenAI, Anthropic, and other leading AI labs. [41][42][43]
  • 2026-05-21: NVIDIA reports Q1 2026 earnings: $81.6B revenue, up 85% year-over-year. [18][44]
  • 2026-05-21: GTC Taipei: Vera Rubin NVL72 wins Computex Best Choice Golden Award; Meta, Google Cloud, and Microsoft formalize partnerships; $2B NVLink Fusion investment in Marvell announced. [45][46][47][48][22][19]
  • 2026-05-31: Dell delivers world's first fully validated Vera Rubin NVL72 rack to CoreWeave; L11 diagnostics confirm scale-up domain operational. [2][24][3][5][25][49][50]
  • 2026-06-01: Jensen Huang at COMPUTEX announces Microsoft completed bring-up of first Rubin VR200 NVL72 via Foxconn; wafer-level production started, rack-level mass production not yet begun. [4]
  • 2026-06-01: SemiAnalysis rates Jensen's COMPUTEX keynote 'F tier': no new AI datacenter products; Rubin FP4/FP8 gains documented at ~3.5x over GB200 while FP16 gains are ~1.6x. [20][16]
  • 2026-06-04: Jensen Huang identifies agentic AI as a $200B market, framing Vera CPU as an active orchestration layer rather than a scheduler. [17]
  • 2026-06-05: Bloomberg reports NVIDIA CEO certified all three HBM4 suppliers — Samsung, SK Hynix, and Micron — for Vera Rubin; Jensen describes the platform as 'in full production' during Seoul supply-chain visit. [12][6]
  • 2026-06-05: NVIDIA open-sources Rubin NVSwitch Tray BoM; SemiAnalysis reports design includes AMD EPYC 3151 embedded CPUs — nine per VR200 rack. [21]
  • 2026-06-08: Analyst Rohan Paul reports HBM4 market share detail: SK Hynix ~60-70%, Samsung ~25-30%, Micron remainder, all in full production for Vera Rubin. [7]
  • 2026-06-08: SemiAnalysis defends Vera SOCAMM report against critics calling it fake news, citing evidence at SK Hynix Computex booth. [13]

Perspectives

NVIDIA / Jensen Huang

Q1 2026 earnings ($81.6B, +85% YoY) validate AI demand; CEO confirmed three-vendor HBM4 certification per Bloomberg [6]; Vera Rubin described as 'in full production' with supply chain aligned for H2 2026 [12]; NVLink Fusion and a $200B agentic AI framing extend the platform narrative [17].

Evolution: Consistent and strengthened: the CEO-level three-vendor confirmation and 'full production' language both reinforce the bullish framing from earlier in the thread.

SemiAnalysis

Rubin FP4/FP8 gains are ~3.5x over GB200 while FP16 is ~1.6x and HBM capacity is flat — efficiency is workload-specific. COMPUTEX keynote rated F tier. NVSwitch Tray BoM discloses AMD EPYC 3151 embedded CPUs (nine per VR200 rack). A Vera SOCAMM report drew fake-news accusations; SemiAnalysis defends it with Computex booth evidence [13].

Evolution: Consistent critical stance; the SOCAMM dispute adds a contested architectural claim alongside existing performance critiques.

Microsoft / Foxconn

First hyperscaler to complete Rubin VR200 NVL72 bring-up via Foxconn as ODM; also anchor customer via Nscale (130,000 GPUs, 1.35GW LOI for West Virginia).

Evolution: Consistent.

Dell / CoreWeave

Dell delivered the world's first fully validated Vera Rubin NVL72 rack to CoreWeave on May 31, clearing L11 with the scale-up domain confirmed operational.

Evolution: Consistent.

Micron

Bloomberg confirmed on June 5 that NVIDIA's CEO certified Micron as an HBM4 supplier for Vera Rubin [6], resolving the question of whether Micron held an allocation alongside Samsung and SK Hynix.

Evolution: Resolved in Micron's favor: the qualification is now CEO-confirmed rather than inferred from Micron's own IR claims or analyst reports.

Memory and supply chain analysts

HBM4 shortage had been characterized as a binding structural constraint; Bloomberg's June 5 CEO-level confirmation of three-vendor qualification [6] and Rohan Paul's market share detail [7] suggest the supply ceiling is materially higher than the two-vendor picture implied, potentially revising the shortage-until-2028 projection [11].

Evolution: Softened: multi-vendor qualification is now primary-source confirmed rather than analyst-inferred, undermining the prior two-vendor constraint framing.

Data center infrastructure analysts

Vera Rubin NVL72's 600kW per-rack power requirement is incompatible with existing data center infrastructure, establishing greenfield construction as a structural bottleneck independent of memory supply.

Evolution: Consistent.

Tensions

  • SemiAnalysis published a Vera SOCAMM note that critics called fake news; SemiAnalysis defends it by citing physical evidence at SK Hynix's Computex booth [13] — the substantive content is disputed and incompletely described in available sources. [13]
  • Jensen stated at COMPUTEX on June 1 that wafer-level production had started but rack-level mass production had not yet begun [4], then Bloomberg quoted him describing Vera Rubin as 'in full production' four days later in Seoul [6][12] — two statements using different production language that leave actual rack delivery timelines ambiguous. [4][12][6]
  • NVIDIA markets Vera Rubin on a 10x cost-per-token reduction, but SemiAnalysis documents FP4/FP8 gains of ~3.5x over GB200 while FP16 gains are ~1.6x and HBM capacity is flat [16] — the efficiency claim is workload-specific, not uniform. [16]
  • NVIDIA's promotional materials present the NVL72 as a NVIDIA-integrated system, but the open-sourced NVSwitch Tray BoM discloses AMD EPYC 3151 embedded CPUs — nine per VR200 rack [21] — a cross-vendor dependency absent from NVIDIA's public architecture descriptions. [21]
  • NVIDIA holds approximately $674M in Nscale equity [15] while describing Nscale publicly as a commercial partner, making large deployment announcements co-publicized by both companies non-arm's-length transactions [23]. [15][23]

Sources

  1. [1] The Data Center Isn't Ready. NVIDIA's Vera Rubin platform ships in… — reactive:nvidia-vera-computex-launch
  2. [2] BREAKING NEWS: COREWEAVE & DELL IS THE FIRST CLOUD TO ANNOUNCE THAT THEY HAVE RUBIN VR200 NVL72 WITH FULLY PASSING L… — SemiAnalysis Twitter (2026-05-31)
  3. [3] Notably, passing L11 diags means that this rack is up and running, including the IMEX channels on the NVL72 scale-up dom… — SemiAnalysis Twitter (2026-05-31)
  4. [4] BREAKING NEWS: JENSEN JUST ANNOUNCED MICROSOFT HAS FINISHED BRING UP ON THEIR FIRST RUBIN VR200 NVL72 RACK with their OD… — SemiAnalysis Twitter (2026-06-01)
  5. [5] At L10 your Firmware/BIOS and OS works on a single server, at L11 a single rack or scale-up domain works, and then at L1… — SemiAnalysis Twitter (2026-05-31)
  6. [6] Nvidia Clears Memory's Big Three for Vera Rubin HBM4 Supply — reactive:nvidia-vera-computex-launch
  7. [7] Nvidia just cleared the memory bottleneck significantly for Vera Rubin by qualifying HBM4 from Samsung, SK Hynix, and Mi… — Rohan Paul Twitter (2026-06-08)
  8. [8] Micron Is Locked Out of HBM4 in NVIDIA's Vera Rubin Systems — reactive:nvidia-vera-computex-launch
  9. [9] NVIDIA to Use SK hynix and Samsung HBM4 for "Vera Rubin" Without Micron | TechPowerUp — reactive:nvidia-vera-computex-launch
  10. [10] SK Hynix Surges 15% to New High: HBM Shortage Until 2028, How Much Longer Can AI Memory King Rise? — reactive:nvidia-vera-computex-launch
  11. [11] Nvidia's memory costs soar 485%, latest AI systems now cost $7.8 ... — reactive:nvidia-vera-computex-launch
  12. [12] Seoul Purpose: How NVIDIA and South Korea Are Building the Future of AI — NVIDIA Blog (2026-06-05)
  13. [13] Our Vera SOCAMM note is causing a bit of a stir. As always some folks are jumping to the wrong conclusions. Those saying… — SemiAnalysis Twitter (2026-06-08)
  14. [14] 130,000 Rubin GPUs Are Being Deployed at Nscale For Microsoft, Further Showing Massive Interest In NVIDIA's Next-Gen AI Chips — reactive:nvidia-vera-computex-launch
  15. [15] UK AI Infrastructure Startup Nscale Receives $674 Million (£500 ... — reactive:nvidia-vera-computex-launch
  16. [16] for more details on Nvidia's VR NVL72 Oberon and future roadmap, check out our article from February: — SemiAnalysis Twitter (2026-05-31)
  17. [17] Jensen Huang just identified the next $200 billion market (Save this). — Milk Road AI Twitter (2026-06-04)
  18. [18] NVIDIA just dropped $81.6B in Q1 revenue up 85% YoY 🤯 — reactive:nvidia-vera-computex-launch (2026-05-21)
  19. [19] The CEO of NVIDIA, looked at Matt Murphy and said "The next trillion dollar company, ladies and gentlemen." (Save this). — Milk Road AI Twitter (2026-06-02)
  20. [20] F TIER KEYNOTEMAX: Jensen ComputeX presentation was one of the worst keynotes he has done. He announced nothing new on t… — SemiAnalysis Twitter (2026-06-01)
  21. [21] BREAKING NEWS: NVIDIA HAS JUST OPEN SOURCED THEIR RUBIN NVSWITCH TRAY BoM & DIAGRAM & IT INCLUDES AMD EYPC 3151 … — SemiAnalysis Twitter (2026-06-05)
  22. [22] Microsoft's strategic AI datacenter planning enables seamless, large ... — reactive:nvidia-vera-computex-launch
  23. [23] Nscale acquires 8GW Monarch Compute Campus, Microsoft signs on for 1.35GW of compute - DCD — reactive:nvidia-vera-computex-launch
  24. [24] Dell just made history this weekend and it is the culmination of an execution streak that no other company in enterprise… — Milk Road AI Twitter (2026-05-31)
  25. [25] CoreWeave Completes Industry-First Bring-Up And Validation Of NVIDIA Vera Rubin NVL72 — reactive:nvidia-vera-computex-launch
  26. [26] Micron in High-Volume Production of HBM4 Designed for NVIDIA ... — reactive:nvidia-vera-computex-launch
  27. [27] Micron Singapore - Facebook — reactive:nvidia-vera-computex-launch
  28. [28] SK Hynix Secures 70% of Nvidia's HBM4 Orders - Semicon — reactive:nvidia-vera-computex-launch
  29. [29] NVIDIA Vera Rubin: 600kW Racks by 2027 | Introl Blog — reactive:nvidia-vera-computex-launch
  30. [30] Nvidia's Vera Rubin GPU: Redesigning Data Centres for 600kW Racks — reactive:nvidia-vera-computex-launch
  31. [31] Nvidia debuts Rubin chip with 336B transistors and 50 petaflops of AI performance - SiliconANGLE — reactive:nvidia-vera-computex-launch
  32. [32] Nvidia CEO confirms Vera Rubin NVL72 is now in production — reactive:nvidia-vera-computex-launch
  33. [33] NVIDIA Vera Rubin AI Platform Hits Full Production CES 2026 ... — reactive:nvidia-vera-computex-launch
  34. [34] SK Hynix set to ship HBM4 for Nvidia's Vera Rubin this month — reactive:nvidia-vera-computex-launch
  35. [35] Nscale and Microsoft Announce Collaboration with NVIDIA and Caterpillar to Deliver 1.35GW of NVIDIA Vera Rubin NVL72 GPUs at Flagship AI Factory Campus in West Virginia — reactive:nvidia-vera-computex-launch
  36. [36] Nscale and Microsoft Announce Collaboration with NVIDIA and Caterpillar to Deliver 1.35GW of NVIDIA Vera Rubin NVL72 GPUs at Flagship AI Factory Campus in West Virginia — reactive:nvidia-vera-computex-launch
  37. [37] Nscale acquisition includes plan to build AI facility in Mason County — reactive:nvidia-vera-computex-launch
  38. [38] Samsung sells out of 2026 HBM4 supply as memory resurgence ... — reactive:aws-garman-a100-demand
  39. [39] Price of Nvidia's Vera Rubin NVL72 racks skyrockets to as much as $8.8 million apiece, but server makers' margins will be tight — Nvidia is moving closer to shipping entire full-scale systems — reactive:nvidia-vera-computex-launch
  40. [40] Nvidia-backed UK AI firm Nscale raises $1.1 billion funding round — reactive:nvidia-vera-computex-launch
  41. [41] Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs — NVIDIA Blog (2026-05-18)
  42. [42] NVIDIA hand-delivers first 1.2 TB/s Vera CPUs to OpenAI, Anthropic ... — reactive:nvidia-vera-computex-launch
  43. [43] Nvidia unveils details of new 88-core Vera CPUs positioned to compete with AMD and Intel – new Vera CPU rack features 256 liquid-cooled chips that deliver up to a 6X gain in CPU throughput | Tom's Hardware — reactive:nvidia-vera-computex-launch
  44. [44] "Demand has gone parabolic. The reason is simple: Agentic AI has arrived." — reactive:nvidia-vera-computex-launch (2026-05-21)
  45. [45] NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI — NVIDIA Blog (2026-05-21)
  46. [46] NVIDIA Vera Rubin NVL72 wins Computex 2026 awards for AI ... — reactive:nvidia-vera-computex-launch
  47. [47] Meta Builds AI Infrastructure With NVIDIA — reactive:nvidia-vera-computex-launch
  48. [48] NVIDIA GTC 2026: Google Cloud Deepens Partnership for AI ... — reactive:nvidia-vera-computex-launch
  49. [49] HPCwire - Since 1987 – Covering the Fastest Computers in the World and the People Who Run Them — reactive:nvidia-vera-computex-launch
  50. [50] CoreWeave Completes Industry-First Bring-Up and Validation of NVIDIA Vera Rubin NVL72 - Las Vegas Sun News — reactive:nvidia-vera-computex-launch