The Information Machine

SemiAnalysis: AI Silicon Shortage — HBM Bottleneck and N3 Wafer Dominance · history

Version 16

2026-06-15 02:43 UTC · 184 items

What

HBM memory, TSMC N3 logic, and power/energy infrastructure are three concurrent binding constraints on AI accelerator production. SemiAnalysis documents the GPU rental market reversing from a cooling phase in October 2025 to a hard squeeze by early 2026, with H100 one-year rental prices up 40% [7][8], and argues agentic workloads — not training — are now driving demand [9]. Micron posted Q2 FY2026 revenue of $23.86B (+196% YoY) with 75% gross margins and the most bullish CEO guidance in company history [3]; Goldman Sachs projects 2027 hyperscaler capex at $1.4T against a $920B Wall Street consensus [14]. Intel's foundry role for NVIDIA's Feynman GPU remains unsettled, with sources ranging from 'confirmed chiplet supplier' to 'backup option in early evaluation' [19][20][21].

Why it matters

Three simultaneous supply constraints — silicon, memory, and power — compound each other: adding compute capacity requires all three to expand together. The shift from training to agentic demand suggests token consumption is moving from episodic to continuous, which sustains GPU scarcity at a higher structural floor than the 2023 cycle implied and makes the squeeze harder to relieve through incremental capacity additions.

Open questions

  • Has Intel finalized its foundry commitment from NVIDIA, or is the relationship still in early evaluation as multiple sources characterize it [21][22]?

  • Does the 'backup AI foundry' framing accurately describe Intel's intended role, and how does that scope 18A volume targets for 2028 [20]?

  • Will power and energy infrastructure become a binding constraint alongside HBM and TSMC N3, given that GPU racks are reaching 400kW and public grid delivery takes 3–5 years [11][12]?

  • Does the SK Hynix–NVIDIA co-development partnership give NVIDIA preferential HBM4/HBM5 allocation, structurally disadvantaging AMD and other accelerator makers competing for the same supply [17]?

Narrative

HBM memory and TSMC's N3 process node are the two established binding constraints on AI accelerator production. SemiAnalysis identified HBM wafer supply — not the earlier CoWoS packaging bottleneck — as the primary scarce resource [1]. SK Hynix has sold out its HBM capacity into 2026 [2], and Micron's Q2 FY2026 results confirmed the demand picture: $23.86 billion in revenue, up 196% year-over-year, with 75% non-GAAP gross margins and the most bullish forward guidance in the company's history [3]. Bernstein projects HBM4 pricing will nearly double from $16.6/GB to $37/GB by 2027 [4]. On the logic side, AI is projected to consume approximately 60% of TSMC's N3 output in 2026, rising to 86% in 2027 [5], with TSMC posting +41% year-over-year Q1 2026 revenue at all-time-high margins [6].

The GPU compute rental market has tightened sharply. SemiAnalysis launched an H100 1-Click Rental Index in June 2026 documenting a 40% price increase in one-year H100 rentals from $1.70/hr in October 2025 to $2.35/hr in March 2026, with a 15–20% step occurring in January–February alone [7][8]. The GPU spot market went from a reported cooling phase to a hard squeeze in approximately five months [8]. SemiAnalysis argues this squeeze differs structurally from 2023: demand has shifted from training runs to agentic workloads, with enterprise token spend moving from a curiosity to a real cost line [9]. Public neocloud equities are priced as if the cycle is rolling over, while SemiAnalysis's read is that GPU scarcity is real and the long-dated rental floor is materially higher than equity valuations imply [10].

Power and energy infrastructure are emerging as a third constraint layer. GPU racks are reaching 400kW power density, exceeding what legacy data centers can support, and public grid delivery timelines run 3–5 years [11][12]. Bloom Energy's on-site fuel cells can deliver power in approximately 90 days, which hyperscalers are paying a premium for rather than waiting on grid timelines [12]; Radiant completed an AI-ready data center from groundbreaking to production in 12 months by bypassing the grid entirely [11]. At the supply chain floor, multi-layer ceramic capacitors — required in tens of thousands per rack — have experienced price hikes and extended lead times as AI server demand accelerates [13].

The investment signals are large and widening. Goldman Sachs projects 2027 hyperscaler capex at $1.4 trillion against a $920 billion Wall Street consensus [14]. Jensen Huang has framed AI factory economics as a $50 billion build cost producing $300–400 billion in intelligence output over the facility's life, and stated the buildout is accelerating with H2 2026 expected to exceed H1 and 2027 larger still [15][16]. SK Hynix and NVIDIA formalized a multi-year HBM co-development partnership in June 2026 covering NVIDIA's Vera Rubin platform [17], and Jensen Huang stated NVIDIA consumes essentially all available HBM output [18]. On the foundry side, Intel's relationship with NVIDIA around the Feynman GPU and Intel 18A remains unsettled: earlier reports framed the deal as confirmed, while more recent sources describe NVIDIA as in early testing with Intel working to finalize commitments, and some frame Intel as a backup rather than co-primary supplier [19][20][21][22]. Google's order of 3M+ TPUs from Intel for 2028 delivery remains confirmed [23].

Timeline

  • 2026-03-01: SemiAnalysis identifies HBM wafer supply — not CoWoS packaging — as the binding constraint on AI accelerator production. [1][44]
  • 2026-05-30: SemiAnalysis projects AI consuming 60% of TSMC N3 output in 2026, rising to 86% in 2027. [24][45][5][46]
  • 2026-05-31: SK Hynix confirmed sold out of DRAM, NAND, and HBM into 2026; Micron confirms 2026 HBM sold out and commits roughly $200 billion to long-term memory capacity. [2][33][47][34]
  • 2026-06-01: TSMC posts +41% year-over-year Q1 2026 revenue with all-time-high margins; Arizona Phase 1 fab profitable ahead of schedule. [39][6][40]
  • 2026-06-01: Intel announces Crescent Island GPU targeting AI inference by end of 2026, without HBM, competing on cost and thermal efficiency. [36][37]
  • 2026-06-03: SemiAnalysis clarifies AMD VR200 and MI455 racks at CoreWeave and Microsoft are engineering samples with incomplete software stacks and no production tokens. [25]
  • 2026-06-05: SemiAnalysis flags MLCCs as an overlooked AI server supply constraint; multiple publications corroborate AI-driven shortages of the sub-$1 passive components required in tens of thousands per rack. [13][48][49][50][51]
  • 2026-06-06: Micron crosses $1 trillion in market cap as investor conviction in the HBM supercycle thesis strengthens. [35]
  • 2026-06-08: SK Hynix and NVIDIA formalize a multi-year HBM co-development partnership covering NVIDIA's Vera Rubin platform; Jensen Huang states NVIDIA consumes essentially all available HBM supply. [17][18]
  • 2026-06-08: Bernstein projects HBM4 pricing to nearly double to $37/GB by 2027; Google's 3M+ TPU order from Intel foundry for 2028 delivery confirmed across multiple publications. [4][38][23]
  • 2026-06-10: Multiple publications report NVIDIA's Feynman GPU will use Intel Foundry for some components in a multi-die chiplet design targeting 2028. [19][26][27][28]
  • 2026-06-10: Jensen Huang states the AI buildout is accelerating — H2 2026 expected to exceed H1, and 2027 projected to be very large; frames AI factory economics as 6–8x return on capital. [15][16]
  • 2026-06-11: Additional reporting characterizes the NVIDIA-Intel 18A arrangement as early testing/evaluation with Intel working to finalize commitments; some sources frame Intel as a backup AI foundry. [20][21][22][52]
  • 2026-06-11: SemiAnalysis flags 400kW GPU racks as exceeding legacy data center capacity and warns grid throttling will constrain AI compute; Radiant completed an AI-ready data center in 12 months by bypassing the grid. [11]
  • 2026-06-12: SemiAnalysis launches H100 1-Click Rental Index documenting a 40% price increase from October 2025 to March 2026; GPU spot market moved from cooling to hard squeeze in approximately five months. [7][8]
  • 2026-06-12: SemiAnalysis argues the current GPU squeeze is structurally driven by agentic workloads rather than training, and that neocloud equities are mispriced relative to real GPU scarcity. [10][9]
  • 2026-06-12: Goldman Sachs projects 2027 hyperscaler capex at $1.4 trillion, more than 50% above the $920 billion Wall Street consensus. [14]
  • 2026-06-13: Micron CEO issues most bullish forward guidance in company history; Q2 FY2026 revenue $23.86B (+196% YoY) with 75% non-GAAP gross margins. [3]

Perspectives

SemiAnalysis

AI's dominance of leading-edge semiconductor capacity is structural; the GPU rental market moved from cooling to hard squeeze in five months, with H100 rentals up 40%; agentic workloads — not training — are driving the current demand surge; neocloud equities are mispriced relative to GPU scarcity and long-dated rental floors; MLCCs and 400kW power density are overlooked infrastructure constraints.

Evolution: Expanded materially: the H100 rental index provides quantitative backing for the scarcity thesis, and the agentic-demand framing adds a new structural argument distinct from earlier publications.

NVIDIA / Jensen Huang

AI buildout is accelerating — H2 2026 will exceed H1 and 2027 will be very large; AI factory economics yield $300–400B in output from a $50B build; NVIDIA consumes essentially all available HBM supply; the Feynman GPU reportedly targets Intel Foundry for some components in a 2028 multi-die design, though that commitment remains in evaluation.

Evolution: New items add Jensen Huang's explicit acceleration outlook and AI factory economics framing; the Intel foundry commitment status remains unsettled.

SK Hynix

Committed to HBM leadership through a multi-year co-development partnership with NVIDIA; targets DRAM capacity doubling by 2031; expects memory supply to remain tight until at least 2030; own 2026 market outlook is explicitly bullish on an HBM-led supercycle.

Evolution: Consistent.

Micron

Q2 FY2026 revenue of $23.86B (+196% YoY) with 75% gross margins and the most bullish CEO guidance in company history; 2026 HBM sold out with roughly $200B committed to long-term capacity; crossed $1T market cap.

Evolution: Blowout Q2 earnings and record guidance materially strengthen the prior stance.

Intel

On two foundry tracks targeting 2028 — Google's confirmed 3M+ TPU order and the NVIDIA Feynman GPU multi-die design, though the NVIDIA commitment is still in evaluation — and separately targeting AI inference with HBM-free Crescent Island by end of 2026.

Evolution: More recent items frame Intel as working to finalize the NVIDIA commitment and as a potential backup foundry, adding negotiation uncertainty to earlier framing that treated the deal as more settled.

TSMC

AI demand is structurally robust through at least 2027–2028; Q1 2026 delivered all-time-high margins on +41% growth; TSMC Arizona is profitable ahead of schedule; AI chips projected to consume 86% of N3 output by 2027.

Evolution: Consistent; no new disclosures this pass.

Goldman Sachs and independent analysts

Goldman projects 2027 hyperscaler capex at $1.4T vs. $920B consensus; Bernstein projects HBM4 nearly doubles to $37/GB by 2027; the early-evaluation characterization of the Intel-NVIDIA relationship adds modest uncertainty to the Intel foundry investment thesis.

Evolution: Goldman's $1.4T projection is new this pass and substantially above consensus, widening the bull-bear gap on AI infrastructure spend.

Public neocloud equity market

Neocloud equities are priced as if the AI infrastructure demand cycle is about to end, reflecting skepticism about the durability of GPU scarcity — directly contra SemiAnalysis's read on rental floors and agentic demand.

Evolution: New voice surfaced by SemiAnalysis's contrarian observation; represents the market consensus SemiAnalysis is explicitly arguing against.

Tensions

  • SemiAnalysis argues GPU scarcity is real and neocloud equity valuations are too low relative to the long-dated rental floor [10][8]; public neocloud equity markets are priced as if the AI infrastructure cycle is rolling over — the two positions assign opposite meanings to the same data. [10][8]
  • Earlier reports state NVIDIA's Feynman GPU 'will use Intel Foundry for some components' [19][27], while newer sources describe NVIDIA as in 'early testing/evaluation stages' with Intel still 'working to finalize commitments' [21][22] — the two framings assign different maturity levels to the same deal. [19][27][21][22]
  • Some sources frame Intel as a 'backup AI foundry' for NVIDIA [20] while TSMC is projected to supply 86% of N3 wafers to AI by 2027 [5] — the scope of Intel's intended role (limited chiplet versus meaningful volume) is unspecified and unresolved. [20][5]
  • The NVIDIA–SK Hynix co-development partnership concentrates leading-edge HBM access around NVIDIA [17], while AMD's MI455/VR200 remain engineering samples [25] and Intel's Crescent Island deliberately avoids HBM [37] — neither competitor is positioned to access HBM4 at scale in the near term. [17][25][37]
  • Goldman Sachs projects 2027 hyperscaler capex at $1.4T [14] vs. the $920B Wall Street consensus — the gap between bull and base cases for AI infrastructure spend has widened to over 50%. [14]
  • Investor framing positions Micron as an 'AI gatekeeper' [42][43], but SK Hynix holds dominant HBM market share, leads HBM4 development, and has formalized a co-development partnership with NVIDIA [17] — the two narratives assign structural primacy to different memory suppliers. [42][43][2][17]

Sources

  1. [1] The Great AI Silicon Shortage - SemiAnalysis — reactive:great-ai-silicon-shortage
  2. [2] SK Hynix sells out DRAM, NAND, and HBM capacity into 2026 amid ... — reactive:great-ai-silicon-shortage
  3. [3] Micron's CEO just dropped the most bullish forward guidance in the company's history and the earnings report is 8 tradin… — Milk Road AI Twitter (2026-06-13)
  4. [4] Most investors think memory stocks have peaked but they are completely wrong. (Save this). — Milk Road AI Twitter (2026-06-08)
  5. [5] Our work shows AI taking roughly 60% of N3 family wafers in 2026 and stepping up to about 86% in 2027, which is a regime… — SemiAnalysis Twitter (2026-05-30)
  6. [6] After posting +41% y/y growth with ATH GM and OM in 1Q26, TSMC is tracking to high-30s growth in CY26. We raised our TSM… — SemiAnalysis Twitter (2026-06-01)
  7. [7] The index has H100 one-year rentals running from $1.70 per hour per GPU in October 2025 to about $2.35 in March 2026, wh… — SemiAnalysis Twitter (2026-06-12)
  8. [8] Alongside the launch of our H100 1-Click Rental Index, we wrote up what the GPU rental market actually looks like in ear… — SemiAnalysis Twitter (2026-06-12)
  9. [9] What we walk through in the article is why this isnt a repeat of the 2023 squeeze. The demand side is no longer training… — SemiAnalysis Twitter (2026-06-12)
  10. [10] Interestingly, the public market is positioned in the opposite direction, with neocloud names trading like the cycle is … — SemiAnalysis Twitter (2026-06-12)
  11. [11] GPU Racks hitting 400kW? Legacy data centers wont be able to handle it and the grid WILL get throttled. — SemiAnalysis Twitter (2026-06-11)
  12. [12] AI data centers need huge amounts of electricity, and the public grid takes 3 to 5 years to deliver it. — Milk Road AI Twitter (2026-06-12)
  13. [13] Nobody is asking who makes the <$1 multi-layer ceramic capacitor (MLCC) that keeps voltage stable across every chip i… — SemiAnalysis Twitter (2026-06-05)
  14. [14] This is WILD! — Milk Road AI Twitter (2026-06-11)
  15. [15] Jensen Huang just explained AI in a way that makes the investment thesis for Nvidia almost impossible to argue with (Sav… — Milk Road AI Twitter (2026-06-11)
  16. [16] Jensen Huang just made a statement that every investor in AI infrastructure needs to hear (Save this). — Milk Road AI Twitter (2026-06-10)
  17. [17] SK hynix and NVIDIA just formed a multi-year memory partnership to build the chips behind the next wave of AI factories. — Rohan Paul Twitter (2026-06-08)
  18. [18] In Seoul, Nvidia CEO Jensen Huang handed out SK Hynix x 7-Eleven HBM Chips snack bags while addressing the crowd. — Rohan Paul Twitter (2026-06-08)
  19. [19] NVIDIA Feynman and Intel Foundry: New Report, Old Core – But With an Important Packaging Clue|igor´sLAB — reactive:great-ai-silicon-shortage
  20. [20] Key facts: Intel tests 18A multi‑die; backup AI foundry; Cadence 14A — TradingView News — reactive:great-ai-silicon-shortage
  21. [21] $NVDA Nvidia is in early testing/evaluation stages with $INTC Intel's ... — reactive:great-ai-silicon-shortage
  22. [22] Intel is reportedly 'working to finalize commitments from Nvidia' as a foundry partner, suggesting gaming potential for the 18A node : r/hardware — reactive:great-ai-silicon-shortage
  23. [23] Google orders 3 million TPUs from Intel as TSMC strains - Quartz — reactive:great-ai-silicon-shortage
  24. [24] It also explains why the bottleneck conversation is migrating away from CoWoS, which is finally easing, and onto memory,… — SemiAnalysis Twitter (2026-05-30)
  25. [25] IMPORTANT: it is important to understand that the CoreWeave & Microsoft photos are still Engineering/Quality Samples… — SemiAnalysis Twitter (2026-06-03)
  26. [26] Nvidia's Next-Gen GPU Could be Coming to Intel Foundry — reactive:great-ai-silicon-shortage
  27. [27] Nvidia Feynman GPUs to use Intel Foundry for some components — reactive:great-ai-silicon-shortage
  28. [28] NVIDIA to Build GPUs on Intel Foundry from 2028: Report - Reddit — reactive:great-ai-silicon-shortage
  29. [29] SK hynix Delays HBM4 Mass Production and Capacity Expansion — reactive:aws-garman-a100-demand
  30. [30] SK hynix just said AI memory demand is now so large that it will double wafer capacity within 5 years, yet still expects… — Rohan Paul Twitter (2026-06-02)
  31. [31] SK hynix said to be planning to double DRAM capacity by 2031 - New Electronics — reactive:great-ai-silicon-shortage
  32. [32] 2026 Market Outlook: SK hynix's HBM to Fuel AI Memory Boom — reactive:great-ai-silicon-shortage
  33. [33] Micron's Sold Out 2026 HBM And US$200b Bet On AI Demand — reactive:micron-hbm-bull-case
  34. [34] Micron's AI Supercycle Accelerates (NASDAQ:MU) | Seeking Alpha — reactive:great-ai-silicon-shortage
  35. [35] Micron crossed $1 trillion in market cap and it is still undervalued (Save this). — Milk Road AI Twitter (2026-06-06)
  36. [36] Intel: Our upcoming AI chip will be cheaper, run cooler than Nvidia, AMD options — Ars Technica AI (2026-06-01)
  37. [37] Intel's new inference chip, Crescent Island, doesn't use HBM. — reactive:great-ai-silicon-shortage (2026-06-05)
  38. [38] The Information reports that Google has picked Intel to manufacture 3M+ Google TPUs in 2028. — Rohan Paul Twitter (2026-06-08)
  39. [39] TSMC Arizona surprised. After ramping up strongly in CY25 ($2B+ revenue), Phase 1 net profit in 1Q26 alone exceeded the … — SemiAnalysis Twitter (2026-06-01)
  40. [40] The foundry industry hit a record $48.8B in 1Q26, +32% y/y and +3% q/q in seasonally soft Q1, marking the 9th consecutiv… — SemiAnalysis Twitter (2026-06-01)
  41. [41] 🟢 Intel Surges After Report Google May Use Its Foundry for AI Chips — reactive:great-ai-silicon-shortage (2026-06-08)
  42. [42] FinancialContent - The Memory Supercycle: Why Micron Technology is the New AI Gatekeeper — reactive:great-ai-silicon-shortage
  43. [43] Micron Stock Up 100%: What the HBM Leader Plans for 2026 — reactive:great-ai-silicon-shortage
  44. [44] The Great AI Silicon Shortage — reactive:great-ai-silicon-shortage
  45. [45] The broader implication, which we work through in detail in the piece, is that the supply curve for frontier accelerator… — SemiAnalysis Twitter (2026-05-30)
  46. [46] One of the throughlines in our Great AI Silicon Shortage piece is that the conversation about leading-edge capacity has … — SemiAnalysis Twitter (2026-05-30)
  47. [47] Sold-Out HBM Supply and AI Tailwinds Point to Strong 2026 Growth — reactive:great-ai-silicon-shortage
  48. [48] MLCC Shortages Return as AI Server Demand Strains Capacity - Astute Group — reactive:great-ai-silicon-shortage
  49. [49] MLCC Consider Price Increase as AI Demand Outpaces Supply — reactive:great-ai-silicon-shortage
  50. [50] AI server boom strains tantalum capacitors; MLCC substitution falls ... — reactive:great-ai-silicon-shortage
  51. [51] AI drives MLCC shortage ... — reactive:great-ai-silicon-shortage
  52. [52] NVIDIA to Build GPUs on Intel Foundry from 2028: Report — reactive:great-ai-silicon-shortage