The Information Machine

NVIDIA Cancels 4-Die Rubin Ultra and Faces Structural Market Share Erosion from Trainium, TPUs, and AMD · history

Version 2

2026-07-01 02:57 UTC · 36 items

What

NVIDIA cancelled the original 4-die Rubin Ultra GPU announced at GTC 2026 due to manufacturing concerns; the replacement carries the same name but delivers roughly half the original design's performance [1]. SemiAnalysis, which reported the cancellation, also projects NVIDIA datacenter compute revenue 20% above consensus for 2H FY2027, attributing the Rubin ramp delay to HBM4 supply issues now resolved [5]. WCCFtech reports Rubin and Rubin Ultra face broader design and spec issues, with AMD's MI500 positioned as a competitor for late 2027 [2]. Separately, Anthropic runs substantial Claude Code inference on AWS Trainium and trains Claude models on Google TPUs, backed by a $100 billion-plus, ten-year commitment to AWS technologies [6][8].

Why it matters

NVIDIA faces simultaneous pressure on two fronts: internal product execution difficulties that forced a significant Rubin Ultra redesign, and growing frontier-lab adoption of alternative silicon from AWS and Google. How quickly NVIDIA executes on the revised Rubin roadmap — and whether AMD, Trainium, and TPUs convert product gaps into durable share — will determine whether NVIDIA's datacenter revenue leadership remains intact through 2027.

Open questions

  • Does SemiAnalysis's bullish 2H NVIDIA revenue forecast [5] account for the reduced Rubin Ultra performance, or is demand high enough that a smaller product still fills the revenue gap?

  • Will AMD's MI500 achieve meaningful traction with hyperscalers and frontier labs in 2H 2027 [2]?

  • Is Anthropic's Trainium and TPU adoption a leading indicator for the broader industry, or specific to its AWS and Google relationships [6][8]?

  • What are the downstream effects on HBM memory suppliers from the Rubin Ultra redesign and NVIDIA's cancellation of some planned rack configurations [4]?

Narrative

NVIDIA announced the 4-die Rubin Ultra at GTC 2026 in March, then cancelled the original design roughly three months later due to manufacturing execution concerns [1]. The replacement product keeps the 'Rubin Ultra' name but is approximately half the size and delivers roughly half the real-world performance of the announced configuration [1]. WCCFtech reports that both the Rubin and Rubin Ultra platforms face broader design and specification issues, and identifies AMD's MI500 as a competitor positioned for the second half of 2027 [2]. Tom's Hardware separately reported that Rubin CPX accelerators were removed from NVIDIA's roadmap [3], and NVIDIA also cancelled some planned rack configurations, with SemiAnalysis reporting material downstream implications for HBM memory demand [4].

Despite the Rubin Ultra product setback, SemiAnalysis projects a strong NVIDIA second half overall. In a June 30, 2026 post, SemiAnalysis estimated NVIDIA datacenter compute revenue would run 20% above sell-side consensus for 2H FY2027, explaining that the Rubin ramp had been delayed by HBM4 supply issues that are now resolved and that front-end wafer supply has been built up sufficiently to support the ramp [5]. This is a more nuanced position than straight-line pessimism: SemiAnalysis sees the Rubin Ultra product as diminished but views overall NVIDIA revenue trajectory as well above what analysts expect — consistent with a market where demand exceeds available supply even for a smaller product.

The competitive landscape on the alternatives side centers on Anthropic's departure from NVIDIA-centric infrastructure. SemiAnalysis reports that Anthropic runs a substantial share of Claude Code inference on AWS Trainium and trains Claude models on Google TPUs [6], a shift it characterized as something that would have been unimaginable a year ago. Reddit discussion confirms Claude Opus was trained on AWS Trainium2 [7], consistent with Anthropic's April 2026 announcement committing over $100 billion over ten years to AWS technologies including Trainium generations 2 through 4, with nearly 1 GW of Trainium2 and Trainium3 capacity expected online by end of 2026 [8]. Amazon deepened its financial stake simultaneously, committing up to $20 billion in additional investment on top of a prior $8 billion [8].

The Trainium narrative remains contested. An analysis titled 'Amazon Trainium Is A Disaster; Strategy Reset Needed' argues the platform has not delivered on its promise [9], directly contradicting the bullish framing from both SemiAnalysis and Anthropic's public commitments. Neither view has been validated with published performance benchmarks. SemiAnalysis's dual position — critical of specific NVIDIA product execution while projecting above-consensus NVIDIA revenue — illustrates that product-level setbacks and business-level outcomes are not the same thing, and that even a diminished Rubin Ultra may satisfy enough of the demand backlog to sustain strong revenue growth.

Timeline

  • 2026-03-01: NVIDIA announces the 4-die Rubin Ultra GPU at GTC 2026 (approximately 3 months before the cancellation report). [1]
  • 2026-04-20: Anthropic and Amazon announce a deal securing up to 5 GW of compute capacity; Anthropic commits $100B+ over 10 years to AWS technologies including Trainium2 through Trainium4. [8]
  • 2026-04-20: Amazon announces an additional $5 billion investment in Anthropic, with up to $20 billion more committed on top of a prior $8 billion stake. [8]
  • 2026-04-20: Anthropic reports annualized run-rate revenue exceeding $30 billion, up from approximately $9 billion at end of 2025. [8]
  • 2026-06-29: SemiAnalysis reports NVIDIA cancelled the original 4-die Rubin Ultra due to manufacturing execution concerns; replacement is roughly half the size and performance. [1]
  • 2026-06-29: SemiAnalysis reports Claude Code inference runs on AWS Trainium and Claude training on Google TPUs, framing this as evidence NVIDIA's CUDA moat is eroding. [6]
  • 2026-06-29: SemiAnalysis reports NVIDIA cancelled some planned future rack configurations, with material implications for HBM memory demand. [4]
  • 2026-06-29: Tom's Hardware reports Rubin CPX accelerators were removed from NVIDIA's roadmap. [3]
  • 2026-06-29: WCCFtech reports Rubin and Rubin Ultra platforms face broader design and spec issues; AMD MI500 positioned as a competitor for 2H 2027. [2]
  • 2026-06-30: SemiAnalysis projects NVIDIA datacenter compute revenue 20% above consensus for 2H FY2027, citing HBM4 supply issues now resolved and sufficient front-end wafer supply built up. [5]

Perspectives

SemiAnalysis

NVIDIA's Rubin Ultra cancellation is a manufacturing execution failure compounding market share erosion from Trainium, TPUs, and AMD; simultaneously, SemiAnalysis projects NVIDIA datacenter compute revenue 20% above consensus for 2H FY2027, attributing the Rubin delay primarily to HBM4 supply issues now resolved.

Evolution: Adds a bullish near-term NVIDIA revenue forecast alongside prior critical product-level reporting — a more complex dual position than the earlier framing implied.

Anthropic

The Amazon deal is framed as an infrastructure response to demand that outpaced capacity, with Trainium central to future Claude workloads; multi-cloud availability on AWS, Google, and Azure is presented as a competitive differentiator.

Evolution: No prior stance in this thread; April 2026 announcement is the first public articulation.

Amazon (AWS)

Deepening financial and infrastructure commitment to Anthropic, positioning Trainium as credible for frontier AI training and inference workloads.

Evolution: Consistent with prior investments; this deal substantially expands financial exposure.

Enertuition (Substack analyst)

Trainium is a strategic failure requiring a reset, directly contesting the bullish narrative around Amazon's custom silicon.

Evolution: No prior stance in this thread; introduced as the primary dissenting voice on Trainium.

WCCFtech

NVIDIA's Rubin and Rubin Ultra platforms face broader design and spec issues beyond the Rubin Ultra cancellation; AMD MI500 is positioned as a competitive alternative for the second half of 2027.

Evolution: New voice in this thread, adding the AMD competitive framing.

Tensions

  • SemiAnalysis and Anthropic argue Trainium has achieved real production-scale adoption at a frontier lab; Enertuition argues Trainium is a strategic failure requiring a reset. [6][8][9]
  • SemiAnalysis frames the Rubin Ultra cancellation as a manufacturing execution failure; SemiAnalysis simultaneously projects NVIDIA datacenter revenue 20% above consensus for 2H, showing product-level setbacks do not automatically translate to revenue losses when demand exceeds supply. [1][5]
  • SemiAnalysis argues NVIDIA's CUDA moat is structurally eroding due to Trainium, TPU, and AMD adoption; whether this is specific to Anthropic's cost and supply situation or a broader industry shift remains unresolved. [6][10]
  • WCCFtech reports Rubin platforms face broad design and spec issues with AMD MI500 as an alternative; SemiAnalysis projects strong overall NVIDIA second-half revenue despite the product setbacks. [2][5]

Sources

  1. [1] INTERESTING: Only 3 months after Rubin Ultra was announced at GTC 2026, the original 4-die Rubin Ultra has been cancelle… — SemiAnalysis Twitter (2026-06-29)
  2. [2] NVIDIA Rubin & Rubin Ultra Platforms Facing Design/Spec Issues ... — reactive:aws-garman-a100-demand
  3. [3] Nvidia removes Rubin CPX accelerators from its roadmap — Groq 3 LPUs take center stage as CPX is removed | Tom's Hardware — reactive:nvidia-rubin-execution-failure
  4. [4] Furthermore, check out our latest accelerator model update, which talks more about the HBM memory implications of these … — SemiAnalysis Twitter (2026-06-29)
  5. [5] We are seeing a huge second half ramp for Nvidia this year. Our Accelerator Model estimate has Nvidia DC compute revenue… — SemiAnalysis Twitter (2026-06-30)
  6. [6] A good chunk of inference for the most successful AI agent, Claude Code, is done on Trainium, while Claude training is d… — SemiAnalysis Twitter (2026-06-29)
  7. [7] anthropic's claude opus just trained on aws' trainium2 gpus — reactive:nvidia-rubin-execution-failure
  8. [8] Anthropic and Amazon expand collaboration for up to 5 gigawatts of new compute — Anthropic News (2026-04-20)
  9. [9] Amazon Trainium Is A Disaster; Strategy Reset Needed — reactive:nvidia-rubin-execution-failure
  10. [10] This all comes against the backdrop of NVIDIA’s market share being eroded by Trainium, TPUs, and AMD chips. For NVIDIA t… — SemiAnalysis Twitter (2026-06-29)