NVIDIA Launches Vera CPU and Vera Rubin NVL72 at COMPUTEX / GTC Taipei · history
Version 17
2026-06-14 02:29 UTC · 286 items
What
NVIDIA's Vera Rubin NVL72 has cleared L11 rack-level validation at two operators — CoreWeave on May 31 [3] and Microsoft on June 1 [4] — with all three HBM4 suppliers confirmed by the NVIDIA CEO [5]. NVIDIA simultaneously advanced an agentic AI infrastructure narrative, releasing AgentPerf benchmark results on June 12 showing GB300 NVL72 running up to 20x more agents per megawatt than H200 [13]. Rack-level mass production and L12 cluster validation have not been confirmed, leaving the timeline against H2 2026 delivery commitments open.
Why it matters
Cleared L11 validations and three-vendor HBM4 supply remove the two most-cited near-term deployment barriers. The pivot to agentic throughput as the primary benchmark metric gives NVIDIA a framework where it can claim large generation-over-generation gains on workloads that traditional token-speed benchmarks cannot capture — but that framework is currently promoted rather than independently validated.
Open questions
L12 cluster validation — the first full compute cluster with scale-out networking — has not been reported for either CoreWeave or Microsoft; when does the first L12 milestone occur?
Jensen said at COMPUTEX on June 1 that rack-level mass production had not yet begun [4]; when does it start, and does that gap create delivery risk against H2 2026 commitments?
AgentPerf is produced by Artificial Analysis and promoted by NVIDIA [13][14] — does any independent body validate the benchmark methodology or run it on non-NVIDIA hardware?
SemiAnalysis's Vera SOCAMM report drew fake-news accusations; SemiAnalysis defends it with physical evidence at the SK Hynix Computex booth [16] — what exactly does the report claim and what are the architectural or supply-chain implications?
Narrative
NVIDIA's Vera Rubin NVL72 is a high-density AI compute rack integrating 72 Rubin GPUs interconnected via NVLink at the rack level, requiring 600kW per rack and HBM4 memory [1]. NVIDIA's validation hierarchy runs from L10 (single-server firmware) through L11 (single-rack scale-up domain) to L12 (full compute cluster with scale-out networking) [2]. The L11 milestone cleared on May 31 when Dell delivered a fully validated VR200 NVL72 rack to CoreWeave [3], and Microsoft completed bring-up of its first Rubin VR200 NVL72 via Foxconn on June 1 at COMPUTEX [4]. Neither operator has reported clearing L12.
The supply chain question — whether NVIDIA qualified two or three HBM4 suppliers — is now settled. Bloomberg reported on June 5 that the NVIDIA CEO confirmed Samsung, SK Hynix, and Micron are all certified to supply HBM4 for Vera Rubin [5]. Analyst Rohan Paul subsequently reported market share: SK Hynix at 60-70%, Samsung at 25-30%, and Micron with the remainder, all in full production [6]. Earlier analyses had concluded only Samsung and SK Hynix held allocations [7][8], and prior reporting projected Samsung selling out its entire 2026 HBM4 allocation with rack prices at $8.8M and shortages through 2028 [9][10]; three-vendor confirmation may soften those projections, though the 600kW per-rack power requirement remains an independent infrastructure constraint requiring greenfield construction [1].
A gap persists between Jensen Huang's production characterizations. At COMPUTEX on June 1 he stated that wafer-level production had started but rack-level mass production had not yet begun [4]; four days later in Seoul, Bloomberg reported him describing Vera Rubin as "in full production" [5][11]. Milk Road AI's June 11 summary adds that the Vera Rubin supply chain is twice the size of Grace Blackwell's and rack assembly time has dropped from two hours to five minutes per unit [12], though these figures originate from NVIDIA promotional materials rather than independent reporting.
In parallel, NVIDIA published results on June 12 from AgentPerf — a benchmark designed by Artificial Analysis to measure concurrent agent throughput rather than single-call token speed. On that metric, the GB300 NVL72 (Blackwell, not Rubin) runs up to 20x more agents per megawatt than the H200 [13][14]. This is relevant to the Vera Rubin thread because Jensen Huang had framed Vera CPU as an active orchestration layer for a $200B agentic AI market [15] and AgentPerf provides the first generation-over-generation metric suited to that framing. The benchmark is not independently validated. SemiAnalysis, separately, published a Vera SOCAMM report that drew accusations of inaccuracy; SemiAnalysis defended it by citing physical evidence at the SK Hynix Computex booth [16], but the report's full content and architectural implications are not described in available sources.
Timeline
- 2026-01-05: NVIDIA debuts Rubin chip at CES: 336 billion transistors, 50 petaflops AI performance. [32]
- 2026-01: Jensen Huang announces at CES 2026 that Vera Rubin NVL72 is in full production. [33][34]
- 2026-02: SK Hynix begins HBM4 mass production shipments to NVIDIA, holding approximately 70% of NVIDIA's HBM4 orders. [35][28]
- 2026-03-17: Nscale acquires 8GW Monarch Compute Campus in West Virginia; Microsoft signs 1.35GW LOI co-announced with NVIDIA and Caterpillar. [36][37][23][38]
- 2026-05: Samsung sells out entire 2026 HBM4 supply; rack prices reach $8.8M; shortage projected until 2028. [39][10][40][9]
- 2026-05: NVL72's 600kW per-rack power requirement documented as incompatible with existing data centers, requiring greenfield construction. [1][29][30]
- 2026-05: NVIDIA equity stake in Nscale confirmed at approximately $674M; Microsoft's Rubin GPU deployment via Nscale revised upward to 130,000 units. [41][31][24]
- 2026-05-18: First Vera CPUs hand-delivered to OpenAI, Anthropic, and other leading AI labs. [42][43][44]
- 2026-05-21: NVIDIA reports Q1 2026 earnings: $81.6B revenue, up 85% year-over-year. [17][45]
- 2026-05-21: GTC Taipei: Vera Rubin NVL72 wins Computex Best Choice Golden Award; Meta, Google Cloud, and Microsoft formalize partnerships; $2B NVLink Fusion investment in Marvell announced. [46][47][48][49][22][18]
- 2026-05-31: Dell delivers world's first fully validated Vera Rubin NVL72 rack to CoreWeave; L11 diagnostics confirm scale-up domain operational. [25][26][3][2][27][50][51]
- 2026-06-01: Jensen Huang at COMPUTEX announces Microsoft completed bring-up of first Rubin VR200 NVL72 via Foxconn; wafer-level production started, rack-level mass production not yet begun. [4]
- 2026-06-01: SemiAnalysis rates Jensen's COMPUTEX keynote 'F tier': no new AI datacenter products; Rubin FP4/FP8 gains documented at ~3.5x over GB200 while FP16 gains are ~1.6x. [20][19]
- 2026-06-04: Jensen Huang identifies agentic AI as a $200B market, framing Vera CPU as an active orchestration layer rather than a scheduler. [15]
- 2026-06-05: Bloomberg reports NVIDIA CEO certified all three HBM4 suppliers — Samsung, SK Hynix, and Micron — for Vera Rubin; Jensen describes the platform as 'in full production' during Seoul supply-chain visit. [11][5]
- 2026-06-05: NVIDIA open-sources Rubin NVSwitch Tray BoM; SemiAnalysis reports design includes AMD EPYC 3151 embedded CPUs — nine per VR200 rack. [21]
- 2026-06-08: Analyst Rohan Paul reports HBM4 market share: SK Hynix ~60-70%, Samsung ~25-30%, Micron remainder, all in full production for Vera Rubin; SemiAnalysis defends Vera SOCAMM report against fake-news accusations. [6][16]
- 2026-06-12: NVIDIA publishes AgentPerf benchmark results via Artificial Analysis: GB300 NVL72 runs up to 20x more agents per megawatt than H200 on agentic workloads. [13][14]
Perspectives
NVIDIA / Jensen Huang
Q1 2026 earnings ($81.6B, +85% YoY) validate AI demand; CEO confirmed three-vendor HBM4 certification [5]; Vera Rubin described as 'in full production' [11]; AgentPerf benchmark frames GB300 as delivering 20x agent-per-megawatt gains over H200 [13]; NVLink Fusion and $200B agentic AI framing extend the platform narrative [15].
Evolution: Strengthened: the agentic AI benchmark adds a new efficiency metric alongside the production and supply-chain confirmations from the prior pass.
SemiAnalysis
Rubin FP4/FP8 gains are ~3.5x over GB200 while FP16 is ~1.6x and HBM capacity is flat — efficiency is workload-specific. COMPUTEX keynote rated F tier. NVSwitch Tray BoM discloses AMD EPYC 3151 embedded CPUs. A Vera SOCAMM report drew fake-news accusations; SemiAnalysis defends it with Computex booth evidence [16].
Evolution: Consistent critical stance across performance, product framing, and the disputed SOCAMM architectural claim.
Microsoft / Foxconn
First hyperscaler to complete Rubin VR200 NVL72 bring-up via Foxconn as ODM; also anchor customer via Nscale (130,000 GPUs, 1.35GW LOI for West Virginia).
Evolution: Consistent.
Dell / CoreWeave
Dell delivered the world's first fully validated Vera Rubin NVL72 rack to CoreWeave on May 31, clearing L11 with the scale-up domain confirmed operational.
Evolution: Consistent.
Memory and supply chain analysts
HBM4 shortage had been framed as a binding structural constraint under a two-vendor picture; Bloomberg's CEO-level confirmation of three-vendor qualification [5] and Rohan Paul's market share detail [6] suggest the supply ceiling is materially higher, potentially revising the shortage-until-2028 projection [10].
Evolution: Softened: multi-vendor qualification is now primary-source confirmed rather than analyst-inferred.
Tensions
- Jensen stated at COMPUTEX on June 1 that wafer-level production had started but rack-level mass production had not yet begun [4]; Bloomberg reported him describing Vera Rubin as 'in full production' four days later in Seoul [5][11] — the two statements use different production language and leave actual rack delivery timelines ambiguous. [4][11][5]
- NVIDIA promotes AgentPerf as the first agentic AI infrastructure benchmark and claims GB300 NVL72 delivers 20x more agents per megawatt than H200 [13]; the benchmark is produced by Artificial Analysis under NVIDIA direction with no independent validation. [13][14]
- NVIDIA markets Vera Rubin on a 10x cost-per-token reduction, but SemiAnalysis documents FP4/FP8 gains of ~3.5x over GB200 while FP16 gains are ~1.6x and HBM capacity is flat [19] — the efficiency claim is workload-specific, not uniform. [19]
- SemiAnalysis published a Vera SOCAMM note that critics called fake news; SemiAnalysis defends it by citing physical evidence at SK Hynix's Computex booth [16] — the substantive content remains incompletely described in available sources. [16]
- NVIDIA's promotional materials present the NVL72 as a NVIDIA-integrated system, but the open-sourced NVSwitch Tray BoM discloses AMD EPYC 3151 embedded CPUs — nine per VR200 rack [21] — a cross-vendor dependency absent from NVIDIA's public architecture descriptions. [21]
- NVIDIA holds approximately $674M in Nscale equity [31] while describing Nscale publicly as a commercial partner, making large deployment announcements co-publicized by both companies non-arm's-length transactions [23]. [31][23]
Sources
- [1] The Data Center Isn't Ready. NVIDIA's Vera Rubin platform ships in… — reactive:nvidia-vera-computex-launch
- [2] At L10 your Firmware/BIOS and OS works on a single server, at L11 a single rack or scale-up domain works, and then at L1… — SemiAnalysis Twitter (2026-05-31)
- [3] Notably, passing L11 diags means that this rack is up and running, including the IMEX channels on the NVL72 scale-up dom… — SemiAnalysis Twitter (2026-05-31)
- [4] BREAKING NEWS: JENSEN JUST ANNOUNCED MICROSOFT HAS FINISHED BRING UP ON THEIR FIRST RUBIN VR200 NVL72 RACK with their OD… — SemiAnalysis Twitter (2026-06-01)
- [5] Nvidia Clears Memory's Big Three for Vera Rubin HBM4 Supply — reactive:nvidia-vera-computex-launch
- [6] Nvidia just cleared the memory bottleneck significantly for Vera Rubin by qualifying HBM4 from Samsung, SK Hynix, and Mi… — Rohan Paul Twitter (2026-06-08)
- [7] Micron Is Locked Out of HBM4 in NVIDIA's Vera Rubin Systems — reactive:nvidia-vera-computex-launch
- [8] NVIDIA to Use SK hynix and Samsung HBM4 for "Vera Rubin" Without Micron | TechPowerUp — reactive:nvidia-vera-computex-launch
- [9] SK Hynix Surges 15% to New High: HBM Shortage Until 2028, How Much Longer Can AI Memory King Rise? — reactive:nvidia-vera-computex-launch
- [10] Nvidia's memory costs soar 485%, latest AI systems now cost $7.8 ... — reactive:nvidia-vera-computex-launch
- [11] Seoul Purpose: How NVIDIA and South Korea Are Building the Future of AI — NVIDIA Blog (2026-06-05)
- [12] Nvidia announced that Vera Rubin is in full production (Save this). — Milk Road AI Twitter (2026-06-11)
- [13] NVIDIA Blackwell Leads on First Agentic AI Infrastructure Benchmark — NVIDIA Blog (2026-06-12)
- [14] NVIDIA just posted the first agentic AI benchmark results where GB300 NVL72 runs up to 20x more coding agents per megawa… — Rohan Paul Twitter (2026-06-12)
- [15] Jensen Huang just identified the next $200 billion market (Save this). — Milk Road AI Twitter (2026-06-04)
- [16] Our Vera SOCAMM note is causing a bit of a stir. As always some folks are jumping to the wrong conclusions. Those saying… — SemiAnalysis Twitter (2026-06-08)
- [17] NVIDIA just dropped $81.6B in Q1 revenue up 85% YoY 🤯 — reactive:nvidia-vera-computex-launch (2026-05-21)
- [18] The CEO of NVIDIA, looked at Matt Murphy and said "The next trillion dollar company, ladies and gentlemen." (Save this). — Milk Road AI Twitter (2026-06-02)
- [19] for more details on Nvidia's VR NVL72 Oberon and future roadmap, check out our article from February: — SemiAnalysis Twitter (2026-05-31)
- [20] F TIER KEYNOTEMAX: Jensen ComputeX presentation was one of the worst keynotes he has done. He announced nothing new on t… — SemiAnalysis Twitter (2026-06-01)
- [21] BREAKING NEWS: NVIDIA HAS JUST OPEN SOURCED THEIR RUBIN NVSWITCH TRAY BoM & DIAGRAM & IT INCLUDES AMD EYPC 3151 … — SemiAnalysis Twitter (2026-06-05)
- [22] Microsoft's strategic AI datacenter planning enables seamless, large ... — reactive:nvidia-vera-computex-launch
- [23] Nscale acquires 8GW Monarch Compute Campus, Microsoft signs on for 1.35GW of compute - DCD — reactive:nvidia-vera-computex-launch
- [24] 130,000 Rubin GPUs Are Being Deployed at Nscale For Microsoft, Further Showing Massive Interest In NVIDIA's Next-Gen AI Chips — reactive:nvidia-vera-computex-launch
- [25] BREAKING NEWS: COREWEAVE & DELL IS THE FIRST CLOUD TO ANNOUNCE THAT THEY HAVE RUBIN VR200 NVL72 WITH FULLY PASSING L… — SemiAnalysis Twitter (2026-05-31)
- [26] Dell just made history this weekend and it is the culmination of an execution streak that no other company in enterprise… — Milk Road AI Twitter (2026-05-31)
- [27] CoreWeave Completes Industry-First Bring-Up And Validation Of NVIDIA Vera Rubin NVL72 — reactive:nvidia-vera-computex-launch
- [28] SK Hynix Secures 70% of Nvidia's HBM4 Orders - Semicon — reactive:nvidia-vera-computex-launch
- [29] NVIDIA Vera Rubin: 600kW Racks by 2027 | Introl Blog — reactive:nvidia-vera-computex-launch
- [30] Nvidia's Vera Rubin GPU: Redesigning Data Centres for 600kW Racks — reactive:nvidia-vera-computex-launch
- [31] UK AI Infrastructure Startup Nscale Receives $674 Million (£500 ... — reactive:nvidia-vera-computex-launch
- [32] Nvidia debuts Rubin chip with 336B transistors and 50 petaflops of AI performance - SiliconANGLE — reactive:nvidia-vera-computex-launch
- [33] Nvidia CEO confirms Vera Rubin NVL72 is now in production — reactive:nvidia-vera-computex-launch
- [34] NVIDIA Vera Rubin AI Platform Hits Full Production CES 2026 ... — reactive:nvidia-vera-computex-launch
- [35] SK Hynix set to ship HBM4 for Nvidia's Vera Rubin this month — reactive:nvidia-vera-computex-launch
- [36] Nscale and Microsoft Announce Collaboration with NVIDIA and Caterpillar to Deliver 1.35GW of NVIDIA Vera Rubin NVL72 GPUs at Flagship AI Factory Campus in West Virginia — reactive:nvidia-vera-computex-launch
- [37] Nscale and Microsoft Announce Collaboration with NVIDIA and Caterpillar to Deliver 1.35GW of NVIDIA Vera Rubin NVL72 GPUs at Flagship AI Factory Campus in West Virginia — reactive:nvidia-vera-computex-launch
- [38] Nscale acquisition includes plan to build AI facility in Mason County — reactive:nvidia-vera-computex-launch
- [39] Samsung sells out of 2026 HBM4 supply as memory resurgence ... — reactive:aws-garman-a100-demand
- [40] Price of Nvidia's Vera Rubin NVL72 racks skyrockets to as much as $8.8 million apiece, but server makers' margins will be tight — Nvidia is moving closer to shipping entire full-scale systems — reactive:nvidia-vera-computex-launch
- [41] Nvidia-backed UK AI firm Nscale raises $1.1 billion funding round — reactive:nvidia-vera-computex-launch
- [42] Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs — NVIDIA Blog (2026-05-18)
- [43] NVIDIA hand-delivers first 1.2 TB/s Vera CPUs to OpenAI, Anthropic ... — reactive:nvidia-vera-computex-launch
- [44] Nvidia unveils details of new 88-core Vera CPUs positioned to compete with AMD and Intel – new Vera CPU rack features 256 liquid-cooled chips that deliver up to a 6X gain in CPU throughput | Tom's Hardware — reactive:nvidia-vera-computex-launch
- [45] "Demand has gone parabolic. The reason is simple: Agentic AI has arrived." — reactive:nvidia-vera-computex-launch (2026-05-21)
- [46] NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI — NVIDIA Blog (2026-05-21)
- [47] NVIDIA Vera Rubin NVL72 wins Computex 2026 awards for AI ... — reactive:nvidia-vera-computex-launch
- [48] Meta Builds AI Infrastructure With NVIDIA — reactive:nvidia-vera-computex-launch
- [49] NVIDIA GTC 2026: Google Cloud Deepens Partnership for AI ... — reactive:nvidia-vera-computex-launch
- [50] HPCwire - Since 1987 – Covering the Fastest Computers in the World and the People Who Run Them — reactive:nvidia-vera-computex-launch
- [51] CoreWeave Completes Industry-First Bring-Up and Validation of NVIDIA Vera Rubin NVL72 - Las Vegas Sun News — reactive:nvidia-vera-computex-launch