NVIDIA Launches Vera CPU and Vera Rubin NVL72 at COMPUTEX / GTC Taipei · history
Version 16
2026-06-11 08:28 UTC · 280 items
What
Bloomberg confirmed on June 5 that NVIDIA's CEO certified all three major HBM4 suppliers — Samsung, SK Hynix, and Micron — for Vera Rubin [6], resolving the most contested supply-chain question about the platform. Analyst Rohan Paul subsequently reported market share detail: SK Hynix at approximately 60-70%, Samsung at 25-30%, and Micron with the remainder, all in full production [7]. Rack-level validation has cleared L11 at both Dell/CoreWeave (May 31) [3] and Microsoft/Foxconn (June 1) [4], while L12 cluster validation and the start of rack-level mass production remain unconfirmed.
Why it matters
CEO-level confirmation of three-vendor HBM4 supply removes the most frequently cited constraint on Vera Rubin deployment pace. The remaining structural binding on large-scale deployment is the 600kW per-rack power requirement, which mandates greenfield data center construction and is independent of memory supply [1].
Open questions
L12 cluster validation — the first full compute cluster with scale-out networking — has not been reported for either CoreWeave or Microsoft; when does the first L12 milestone occur?
Jensen described Vera Rubin as 'in full production' on June 5 [12][6], but at COMPUTEX on June 1 stated rack-level mass production had not yet begun [4] — when does rack-level mass production actually begin, and does this gap create delivery risk against H2 2026 commitments?
The SemiAnalysis Vera SOCAMM report attracted accusations of inaccuracy and was defended with physical evidence from the SK Hynix Computex booth [13] — what exactly does the report claim, and what are the architectural or supply-chain implications if correct?
Narrative
NVIDIA's Vera Rubin NVL72 is a high-density AI compute rack integrating 72 Rubin GPUs interconnected via NVLink at the rack level, requiring 600kW per rack and HBM4 memory [1]. The platform reached the first rack-level validation milestone on May 31, when Dell delivered a fully validated VR200 NVL72 rack to CoreWeave, confirming the L11 scale-up domain operational [2][3]. Microsoft completed bring-up of its first Rubin NVL72 via Foxconn on June 1 at COMPUTEX [4]. NVIDIA's validation hierarchy runs from L10 (single-server firmware) through L11 (single-rack scale-up domain) to L12 (full compute cluster with scale-out networking) [5]; neither operator has reported clearing L12.
The central supply-chain question — whether NVIDIA qualified two or three HBM4 suppliers — was answered when Bloomberg reported on June 5 that the NVIDIA CEO confirmed certification of Samsung, SK Hynix, and Micron to supply HBM4 for Vera Rubin [6]. Analyst Rohan Paul subsequently reported market share: SK Hynix at 60-70%, Samsung at 25-30%, and Micron with the remainder, all in full production [7]. This contradicts prior independent analyses that concluded only Samsung and SK Hynix had been designated [8][9], and broadens the effective supply ceiling. Prior reporting had Samsung selling out its entire 2026 HBM4 allocation with rack prices at $8.8M and shortage projections through 2028 [10][11]; the three-vendor picture may revise those projections, though the 600kW per-rack power requirement remains a binding infrastructure constraint independent of memory supply [1].
During the same Seoul visit where Bloomberg captured the three-vendor certification statement, Jensen Huang described Vera Rubin as 'in full production' [12] — language that sits in tension with his COMPUTEX statement four days earlier that wafer-level production had started but rack-level mass production had not yet begun [4]. The distinction between wafer-level and rack-level production is material for delivery timelines against H2 2026 commitments. A separate contested claim comes from SemiAnalysis, which published a report on 'Vera SOCAMM' — a memory module format — that attracted accusations of inaccuracy; SemiAnalysis defended it by citing physical evidence at the SK Hynix Computex booth [13]. The full content and architectural implications of the SOCAMM report remain incompletely described in available sources.
Deployment commitments include Nscale (130,000 Rubin GPUs for Microsoft, with NVIDIA holding approximately $674M in Nscale equity) [14][15] and Verda across Europe, the US, and APAC. SemiAnalysis has rated the GTC Taipei keynote 'F tier' and documents that Rubin FP4/FP8 gains are approximately 3.5x over GB200 while FP16 gains are approximately 1.6x and HBM capacity is flat [16] — making stated efficiency advantages workload-specific rather than uniform.
Timeline
- 2026-01-05: NVIDIA debuts Rubin chip at CES: 336 billion transistors, 50 petaflops AI performance. [31]
- 2026-01: Jensen Huang announces at CES 2026 that Vera Rubin NVL72 is in full production. [32][33]
- 2026-02: SK Hynix begins HBM4 mass production shipments to NVIDIA, holding approximately 70% of NVIDIA's HBM4 orders. [34][28]
- 2026-03-17: Nscale acquires 8GW Monarch Compute Campus in West Virginia; Microsoft signs 1.35GW LOI co-announced with NVIDIA and Caterpillar. [35][36][23][37]
- 2026-05: Samsung sells out entire 2026 HBM4 supply; rack prices reach $8.8M; shortage projected until 2028. [38][11][39][10]
- 2026-05: NVL72's 600kW per-rack power requirement documented as incompatible with existing data centers, requiring greenfield construction. [1][29][30]
- 2026-05: NVIDIA equity stake in Nscale confirmed at approximately $674M; Microsoft's Rubin GPU deployment via Nscale revised upward to 130,000 units. [40][15][14]
- 2026-05-18: First Vera CPUs hand-delivered to OpenAI, Anthropic, and other leading AI labs. [41][42][43]
- 2026-05-21: NVIDIA reports Q1 2026 earnings: $81.6B revenue, up 85% year-over-year. [18][44]
- 2026-05-21: GTC Taipei: Vera Rubin NVL72 wins Computex Best Choice Golden Award; Meta, Google Cloud, and Microsoft formalize partnerships; $2B NVLink Fusion investment in Marvell announced. [45][46][47][48][22][19]
- 2026-05-31: Dell delivers world's first fully validated Vera Rubin NVL72 rack to CoreWeave; L11 diagnostics confirm scale-up domain operational. [2][24][3][5][25][49][50]
- 2026-06-01: Jensen Huang at COMPUTEX announces Microsoft completed bring-up of first Rubin VR200 NVL72 via Foxconn; wafer-level production started, rack-level mass production not yet begun. [4]
- 2026-06-01: SemiAnalysis rates Jensen's COMPUTEX keynote 'F tier': no new AI datacenter products; Rubin FP4/FP8 gains documented at ~3.5x over GB200 while FP16 gains are ~1.6x. [20][16]
- 2026-06-04: Jensen Huang identifies agentic AI as a $200B market, framing Vera CPU as an active orchestration layer rather than a scheduler. [17]
- 2026-06-05: Bloomberg reports NVIDIA CEO certified all three HBM4 suppliers — Samsung, SK Hynix, and Micron — for Vera Rubin; Jensen describes the platform as 'in full production' during Seoul supply-chain visit. [12][6]
- 2026-06-05: NVIDIA open-sources Rubin NVSwitch Tray BoM; SemiAnalysis reports design includes AMD EPYC 3151 embedded CPUs — nine per VR200 rack. [21]
- 2026-06-08: Analyst Rohan Paul reports HBM4 market share detail: SK Hynix ~60-70%, Samsung ~25-30%, Micron remainder, all in full production for Vera Rubin. [7]
- 2026-06-08: SemiAnalysis defends Vera SOCAMM report against critics calling it fake news, citing evidence at SK Hynix Computex booth. [13]
Perspectives
NVIDIA / Jensen Huang
Q1 2026 earnings ($81.6B, +85% YoY) validate AI demand; CEO confirmed three-vendor HBM4 certification per Bloomberg [6]; Vera Rubin described as 'in full production' with supply chain aligned for H2 2026 [12]; NVLink Fusion and a $200B agentic AI framing extend the platform narrative [17].
Evolution: Consistent and strengthened: the CEO-level three-vendor confirmation and 'full production' language both reinforce the bullish framing from earlier in the thread.
SemiAnalysis
Rubin FP4/FP8 gains are ~3.5x over GB200 while FP16 is ~1.6x and HBM capacity is flat — efficiency is workload-specific. COMPUTEX keynote rated F tier. NVSwitch Tray BoM discloses AMD EPYC 3151 embedded CPUs (nine per VR200 rack). A Vera SOCAMM report drew fake-news accusations; SemiAnalysis defends it with Computex booth evidence [13].
Evolution: Consistent critical stance; the SOCAMM dispute adds a contested architectural claim alongside existing performance critiques.
Microsoft / Foxconn
First hyperscaler to complete Rubin VR200 NVL72 bring-up via Foxconn as ODM; also anchor customer via Nscale (130,000 GPUs, 1.35GW LOI for West Virginia).
Evolution: Consistent.
Dell / CoreWeave
Dell delivered the world's first fully validated Vera Rubin NVL72 rack to CoreWeave on May 31, clearing L11 with the scale-up domain confirmed operational.
Evolution: Consistent.
Micron
Bloomberg confirmed on June 5 that NVIDIA's CEO certified Micron as an HBM4 supplier for Vera Rubin [6], resolving the question of whether Micron held an allocation alongside Samsung and SK Hynix.
Evolution: Resolved in Micron's favor: the qualification is now CEO-confirmed rather than inferred from Micron's own IR claims or analyst reports.
Memory and supply chain analysts
HBM4 shortage had been characterized as a binding structural constraint; Bloomberg's June 5 CEO-level confirmation of three-vendor qualification [6] and Rohan Paul's market share detail [7] suggest the supply ceiling is materially higher than the two-vendor picture implied, potentially revising the shortage-until-2028 projection [11].
Evolution: Softened: multi-vendor qualification is now primary-source confirmed rather than analyst-inferred, undermining the prior two-vendor constraint framing.
Tensions
- SemiAnalysis published a Vera SOCAMM note that critics called fake news; SemiAnalysis defends it by citing physical evidence at SK Hynix's Computex booth [13] — the substantive content is disputed and incompletely described in available sources. [13]
- Jensen stated at COMPUTEX on June 1 that wafer-level production had started but rack-level mass production had not yet begun [4], then Bloomberg quoted him describing Vera Rubin as 'in full production' four days later in Seoul [6][12] — two statements using different production language that leave actual rack delivery timelines ambiguous. [4][12][6]
- NVIDIA markets Vera Rubin on a 10x cost-per-token reduction, but SemiAnalysis documents FP4/FP8 gains of ~3.5x over GB200 while FP16 gains are ~1.6x and HBM capacity is flat [16] — the efficiency claim is workload-specific, not uniform. [16]
- NVIDIA's promotional materials present the NVL72 as a NVIDIA-integrated system, but the open-sourced NVSwitch Tray BoM discloses AMD EPYC 3151 embedded CPUs — nine per VR200 rack [21] — a cross-vendor dependency absent from NVIDIA's public architecture descriptions. [21]
- NVIDIA holds approximately $674M in Nscale equity [15] while describing Nscale publicly as a commercial partner, making large deployment announcements co-publicized by both companies non-arm's-length transactions [23]. [15][23]
Sources
- [1] The Data Center Isn't Ready. NVIDIA's Vera Rubin platform ships in… — reactive:nvidia-vera-computex-launch
- [2] BREAKING NEWS: COREWEAVE & DELL IS THE FIRST CLOUD TO ANNOUNCE THAT THEY HAVE RUBIN VR200 NVL72 WITH FULLY PASSING L… — SemiAnalysis Twitter (2026-05-31)
- [3] Notably, passing L11 diags means that this rack is up and running, including the IMEX channels on the NVL72 scale-up dom… — SemiAnalysis Twitter (2026-05-31)
- [4] BREAKING NEWS: JENSEN JUST ANNOUNCED MICROSOFT HAS FINISHED BRING UP ON THEIR FIRST RUBIN VR200 NVL72 RACK with their OD… — SemiAnalysis Twitter (2026-06-01)
- [5] At L10 your Firmware/BIOS and OS works on a single server, at L11 a single rack or scale-up domain works, and then at L1… — SemiAnalysis Twitter (2026-05-31)
- [6] Nvidia Clears Memory's Big Three for Vera Rubin HBM4 Supply — reactive:nvidia-vera-computex-launch
- [7] Nvidia just cleared the memory bottleneck significantly for Vera Rubin by qualifying HBM4 from Samsung, SK Hynix, and Mi… — Rohan Paul Twitter (2026-06-08)
- [8] Micron Is Locked Out of HBM4 in NVIDIA's Vera Rubin Systems — reactive:nvidia-vera-computex-launch
- [9] NVIDIA to Use SK hynix and Samsung HBM4 for "Vera Rubin" Without Micron | TechPowerUp — reactive:nvidia-vera-computex-launch
- [10] SK Hynix Surges 15% to New High: HBM Shortage Until 2028, How Much Longer Can AI Memory King Rise? — reactive:nvidia-vera-computex-launch
- [11] Nvidia's memory costs soar 485%, latest AI systems now cost $7.8 ... — reactive:nvidia-vera-computex-launch
- [12] Seoul Purpose: How NVIDIA and South Korea Are Building the Future of AI — NVIDIA Blog (2026-06-05)
- [13] Our Vera SOCAMM note is causing a bit of a stir. As always some folks are jumping to the wrong conclusions. Those saying… — SemiAnalysis Twitter (2026-06-08)
- [14] 130,000 Rubin GPUs Are Being Deployed at Nscale For Microsoft, Further Showing Massive Interest In NVIDIA's Next-Gen AI Chips — reactive:nvidia-vera-computex-launch
- [15] UK AI Infrastructure Startup Nscale Receives $674 Million (£500 ... — reactive:nvidia-vera-computex-launch
- [16] for more details on Nvidia's VR NVL72 Oberon and future roadmap, check out our article from February: — SemiAnalysis Twitter (2026-05-31)
- [17] Jensen Huang just identified the next $200 billion market (Save this). — Milk Road AI Twitter (2026-06-04)
- [18] NVIDIA just dropped $81.6B in Q1 revenue up 85% YoY 🤯 — reactive:nvidia-vera-computex-launch (2026-05-21)
- [19] The CEO of NVIDIA, looked at Matt Murphy and said "The next trillion dollar company, ladies and gentlemen." (Save this). — Milk Road AI Twitter (2026-06-02)
- [20] F TIER KEYNOTEMAX: Jensen ComputeX presentation was one of the worst keynotes he has done. He announced nothing new on t… — SemiAnalysis Twitter (2026-06-01)
- [21] BREAKING NEWS: NVIDIA HAS JUST OPEN SOURCED THEIR RUBIN NVSWITCH TRAY BoM & DIAGRAM & IT INCLUDES AMD EYPC 3151 … — SemiAnalysis Twitter (2026-06-05)
- [22] Microsoft's strategic AI datacenter planning enables seamless, large ... — reactive:nvidia-vera-computex-launch
- [23] Nscale acquires 8GW Monarch Compute Campus, Microsoft signs on for 1.35GW of compute - DCD — reactive:nvidia-vera-computex-launch
- [24] Dell just made history this weekend and it is the culmination of an execution streak that no other company in enterprise… — Milk Road AI Twitter (2026-05-31)
- [25] CoreWeave Completes Industry-First Bring-Up And Validation Of NVIDIA Vera Rubin NVL72 — reactive:nvidia-vera-computex-launch
- [26] Micron in High-Volume Production of HBM4 Designed for NVIDIA ... — reactive:nvidia-vera-computex-launch
- [27] Micron Singapore - Facebook — reactive:nvidia-vera-computex-launch
- [28] SK Hynix Secures 70% of Nvidia's HBM4 Orders - Semicon — reactive:nvidia-vera-computex-launch
- [29] NVIDIA Vera Rubin: 600kW Racks by 2027 | Introl Blog — reactive:nvidia-vera-computex-launch
- [30] Nvidia's Vera Rubin GPU: Redesigning Data Centres for 600kW Racks — reactive:nvidia-vera-computex-launch
- [31] Nvidia debuts Rubin chip with 336B transistors and 50 petaflops of AI performance - SiliconANGLE — reactive:nvidia-vera-computex-launch
- [32] Nvidia CEO confirms Vera Rubin NVL72 is now in production — reactive:nvidia-vera-computex-launch
- [33] NVIDIA Vera Rubin AI Platform Hits Full Production CES 2026 ... — reactive:nvidia-vera-computex-launch
- [34] SK Hynix set to ship HBM4 for Nvidia's Vera Rubin this month — reactive:nvidia-vera-computex-launch
- [35] Nscale and Microsoft Announce Collaboration with NVIDIA and Caterpillar to Deliver 1.35GW of NVIDIA Vera Rubin NVL72 GPUs at Flagship AI Factory Campus in West Virginia — reactive:nvidia-vera-computex-launch
- [36] Nscale and Microsoft Announce Collaboration with NVIDIA and Caterpillar to Deliver 1.35GW of NVIDIA Vera Rubin NVL72 GPUs at Flagship AI Factory Campus in West Virginia — reactive:nvidia-vera-computex-launch
- [37] Nscale acquisition includes plan to build AI facility in Mason County — reactive:nvidia-vera-computex-launch
- [38] Samsung sells out of 2026 HBM4 supply as memory resurgence ... — reactive:aws-garman-a100-demand
- [39] Price of Nvidia's Vera Rubin NVL72 racks skyrockets to as much as $8.8 million apiece, but server makers' margins will be tight — Nvidia is moving closer to shipping entire full-scale systems — reactive:nvidia-vera-computex-launch
- [40] Nvidia-backed UK AI firm Nscale raises $1.1 billion funding round — reactive:nvidia-vera-computex-launch
- [41] Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs — NVIDIA Blog (2026-05-18)
- [42] NVIDIA hand-delivers first 1.2 TB/s Vera CPUs to OpenAI, Anthropic ... — reactive:nvidia-vera-computex-launch
- [43] Nvidia unveils details of new 88-core Vera CPUs positioned to compete with AMD and Intel – new Vera CPU rack features 256 liquid-cooled chips that deliver up to a 6X gain in CPU throughput | Tom's Hardware — reactive:nvidia-vera-computex-launch
- [44] "Demand has gone parabolic. The reason is simple: Agentic AI has arrived." — reactive:nvidia-vera-computex-launch (2026-05-21)
- [45] NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI — NVIDIA Blog (2026-05-21)
- [46] NVIDIA Vera Rubin NVL72 wins Computex 2026 awards for AI ... — reactive:nvidia-vera-computex-launch
- [47] Meta Builds AI Infrastructure With NVIDIA — reactive:nvidia-vera-computex-launch
- [48] NVIDIA GTC 2026: Google Cloud Deepens Partnership for AI ... — reactive:nvidia-vera-computex-launch
- [49] HPCwire - Since 1987 – Covering the Fastest Computers in the World and the People Who Run Them — reactive:nvidia-vera-computex-launch
- [50] CoreWeave Completes Industry-First Bring-Up and Validation of NVIDIA Vera Rubin NVL72 - Las Vegas Sun News — reactive:nvidia-vera-computex-launch