The Information Machine

Chinese AI Models and Products Gain Structural Ground on US Rivals · history

Version 9

2026-06-23 18:24 UTC · 128 items

What

Chinese AI providers are advancing on model capability, open-weight distribution, capital, and physical infrastructure while facing a structural compute spending gap versus the US. GLM-5.2 — a 753B-parameter MIT-licensed model released June 16 by Z.ai — tops the Artificial Analysis Intelligence Index v4.1 [8], costs approximately $4.40 per million output tokens [9], and has drawn sharply divided assessments: Nathan Lambert calls it the first open-weight model to work credibly as a general agent in coding harnesses [10], while Zvi Mowshowitz argues it is heavily distilled from Claude Opus, trails frontier closed models substantially, and sits in an awkward commercial position [12]. DeepSeek's $7.4B raise at a $50B valuation [14] and disclosed MW-to-GW data center plans [15] represent the infrastructure dimension. US hyperscalers are projected to spend approximately 8.3x more than Chinese hyperscalers on AI infrastructure by 2027 [18]. The US-ASML dispute over whether a banned EUV tool reached China remains unresolved: ASML says it actively tracks all ~314–340 EUV units worldwide and none have ever shipped to China [20][21].

Why it matters

The GLM-5.2 debate is a proxy for a larger question: whether Chinese open-weight labs have closed enough of the capability gap to put meaningful commercial pressure on closed-model incumbents. The 8.3x projected compute spending gap suggests US labs retain a structural advantage in scaling, but if distilled open-weight models can deliver near-frontier performance at a fraction of the price, compute scale alone may not maintain commercial separation.

Open questions

  • Did a banned EUV tool actually reach China, or is the US concern unfounded as ASML asserts, claiming it actively tracks all ~314–340 EUV units worldwide and none have ever shipped to China? [19][20][21]

  • Is GLM-5.2 heavily distilled from Claude Opus as Mowshowitz argues [12], and if so, does that mean its benchmark performance overstates how it handles unusual real-world tasks?

  • Will the 8.3x projected US hyperscaler spending advantage by 2027 [18] translate to maintained model quality gaps, or do Chinese labs' efficiency gains offset the compute difference?

  • Will the US government attempt to restrict access to frontier-quality open-weight Chinese models, which Lambert considers a real risk that would dangerously concentrate AI power [10]?

Narrative

Through 2025, US models led weekly token consumption on OpenRouter. From early 2026, Chinese models became the primary growth driver, reaching 61% of the platform's weekly token market share by mid-2026 [1], with China holding 6 of the top 9 spots in a spring 2026 global usage ranking [2]. OpenAI's June 2024 cutoff of API access for Chinese developers created an opening that Baidu, Alibaba, and Moonshot AI actively exploited [3][4]. Observers disagree on causes: Rohan Paul treats the shift as a structural competitive gain [5]; developer-community commentators argue developers route to whatever delivers the best cost-performance ratio at a given moment, making the gains real but price-conditional [6].

The open-weight model pipeline in mid-June 2026 produced two significant releases. MiniMax M3 (428B total parameters, 23B activated, block sparse attention 9x faster than its predecessor) shipped June 13 with day-0 MXFP8 weights [7]. GLM-5.2, released June 16 by Z.ai under an MIT license, is a 753B-parameter mixture-of-experts model that tops the Artificial Analysis Intelligence Index v4.1 with a score of 51, expands context to 1 million tokens, and costs approximately $4.40 per million output tokens [8][9]. Its reception divides sharply. Nathan Lambert describes it as the first open-weight model to work credibly as a general agent in coding harnesses like Claude Code, placing it alongside OpenAI and Anthropic's latest closed models on Arena's agent leaderboard, and argues GLM-5.2 puts economic pressure on Anthropic's Claude Code revenue by offering an open alternative [10]. The Vercel CEO described its coding ability as 'almost shocking' [11]. Zvi Mowshowitz credits the achievement — calling GLM-5.2 the strongest currently available open-weight model, scoring around Opus 4.7 on traditional benchmarks — but argues it is very likely heavily distilled from Claude Opus, citing its behavioral voice and reliance on a Claude harness; he notes distilled models tend to overperform on benchmarks and underperform on less common tasks, and that GLM-5.2 sits in an awkward commercial position — not cheap enough for bulk tasks nor strong enough for the hardest tasks compared to closed-model alternatives [12]. Milk Road AI argues open-source models will not win the AI race and the market narrative around DeepSeek's early 2025 launch misread the episode as validation of open-weight approaches [13].

At the infrastructure and financial layer, DeepSeek completed a $7.4B financing round at a $50B valuation on June 16, becoming China's most valuable AI startup; founder Liang Wenfeng held approximately 90% pre-round and personally contributed around $3B [14]. Job postings for MW-to-GW scale data center design signal a move from software-first to vertical integration of physical compute [15]. China established a government-chartered Space Computing Industry Innovation Center led by Beijing University of Posts and Telecommunications before Elon Musk announced the AI1 satellite [16]. A hardware teardown of SMIC's N+3 process found it reaches TSMC N6-class logic density via DUV multi-patterning but at higher cost, with the resulting Kirin 9030 Pro trailing current flagship SoCs [17]. Against this, Rohan Paul notes that US hyperscalers are projected to spend approximately 8.3x more than Chinese hyperscalers on AI infrastructure by 2027, arguing compute advantage remains firmly on the US side [18].

US Commerce Secretary Howard Lutnick personally raised with ASML leadership the concern that a banned EUV lithography machine may have reached China [19]. ASML denied the claim publicly, saying it has never shipped an EUV machine to China, that it actively tracks all ~314–340 EUV units worldwide, and that the reports are 'inaccurate and damaging to our reputation' [20][21][22]. The dispute is unresolved. US export controls apply to EUV machines in part because they incorporate US intellectual property [23]; ASML separately sells older DUV machines to China, a distinction central to the current disagreement.

Timeline

  • 2024-06: OpenAI cuts API access for developers in China; Baidu, Alibaba, and Moonshot AI actively recruit displaced users. [3][4]
  • 2026-01: Chinese models begin outpacing US models as the primary growth driver of token consumption on OpenRouter. [5]
  • 2026-03: China holds 6 of the top 9 spots in a global AI model usage ranking. [2]
  • 2026-06: Chinese models reach 61% of OpenRouter weekly token market share. [1]
  • 2026-06-10: SemiAnalysis reports DeepSeek posted a role for MW-to-GW scale data center design, interpreting it as a shift to heavy-asset vertical integration. [15]
  • 2026-06-13: MiniMax M3 (428B parameters, 23B activated) released with day-0 MXFP8 weights and block sparse attention 9x faster than M2.7. [7][24]
  • 2026-06-14: SemiAnalysis SMIC N+3 teardown finds TSMC N6-class logic density via DUV multi-patterning but at higher cost; Kirin 9030 Pro trails current flagship SoCs. [17]
  • 2026-06-16: DeepSeek raises $7.4B at a $50B valuation, becoming China's most valuable AI startup; founder Liang Wenfeng personally contributed ~$3B. [14][26]
  • 2026-06-16: Z.ai releases GLM-5.2 (753B parameters, MIT license), topping the Artificial Analysis Intelligence Index v4.1 with a score of 51 and a 1M-token context window. [8]
  • 2026-06-19: US Commerce Secretary Lutnick raises concern with ASML leadership that a banned EUV tool reached China; ASML publicly denies it, saying it tracks all ~314–340 EUV units worldwide. [19][20][25]
  • 2026-06-19: SemiAnalysis notes China launched a government-chartered Space Computing Industry Innovation Center (led by BUPT) before Musk announced the AI1 satellite. [16]
  • 2026-06-20: Milk Road AI argues open-source models will not win the AI race and the market narrative around DeepSeek's early 2025 launch was misread. [13]
  • 2026-06-21: Vercel CEO describes GLM-5.2's coding ability as 'almost shocking', drawing broad attention in developer communities. [11]
  • 2026-06-22: Nathan Lambert (Interconnects) calls GLM-5.2 the first open-weight model to work credibly as a general agent in coding harnesses, arguing it puts economic pressure on Anthropic's Claude Code revenue. [10]
  • 2026-06-22: Zvi Mowshowitz credits GLM-5.2 as the best open-weight model but argues it is heavily distilled from Claude Opus, trails frontier closed models substantially, and occupies an awkward commercial niche. [12]
  • 2026-06-22: Rohan Paul notes US hyperscalers are projected to spend approximately 8.3x more than Chinese hyperscalers on AI infrastructure by 2027. [18]

Perspectives

Nathan Lambert (Interconnects)

GLM-5.2 is the first open-weight model to perform credibly as a general agent in coding harnesses, comparable in significance to DeepSeek R1; it puts economic pressure on Anthropic's Claude Code revenue and raises the prospect of US government restrictions on open models that Lambert views as a dangerous concentration of AI power.

Evolution: New voice this pass.

Zvi Mowshowitz

GLM-5.2 is the strongest open-weight model available but is very likely heavily distilled from Claude Opus, trails frontier closed models (Opus 4.8, GPT-5.5) substantially, and sits in an awkward commercial niche — not cheap enough for bulk tasks, not strong enough for the hardest ones.

Evolution: New voice this pass.

Rohan Paul

Tracks Chinese AI competitive gains across OpenRouter traffic, DeepSeek's capital raise, and the US-ASML EUV dispute, while also noting the 8.3x projected US-China hyperscaler spending gap as a structural US advantage.

Evolution: Added the compute spending counter-data point, which qualifies the competitive-rise framing he has otherwise consistently maintained.

SemiAnalysis

Covers Chinese AI infrastructure moves with technical qualification: SMIC N+3 reaches TSMC N6 density via DUV multi-patterning but at higher cost, resulting SoCs trail current flagships, and China's space-based compute initiative preceded Musk's AI1.

Evolution: Consistent.

ASML

Categorically denies ever shipping an EUV lithography machine to China, saying it actively tracks all ~314–340 EUV units worldwide and that the reports are 'inaccurate and damaging to our reputation.'

Evolution: Denial has hardened with the additional detail of active unit tracking across all deployed machines.

Milk Road AI

Open-source models will not win the AI race, and the dominant narrative around DeepSeek's early 2025 launch misread the episode as validation of open-weight approaches.

Evolution: Consistent; directly challenges the open-weight dominance thesis running through the thread.

Grant Harvey (The Neuron)

GLM-5.2's open-weight availability at ~$4.40 per million output tokens gives developers strategic optionality against closed-model vendor risk; the emerging AI stack will likely use frontier closed models for the hardest tasks and cheaper open models for repeatable or privacy-sensitive work.

Evolution: New voice this pass.

Developer/inference-market commentators (ollobrains)

The LLM market functions as an inference spot market — developers route to whatever model delivers the best cost-performance ratio at a given moment, making Chinese models' gains real but conditional on price.

Evolution: Consistent; structurally different interpretation of the OpenRouter data than the competitive-shift framing.

Tensions

  • Nathan Lambert argues GLM-5.2 is the first open-weight model to match closed frontier performance in coding agent harnesses [10]; Zvi Mowshowitz argues it is very likely heavily distilled from Claude Opus, trails frontier closed models substantially, and occupies an awkward commercial niche [12]. [10][12]
  • The US government believes a banned EUV lithography tool reached China; ASML categorically denies having ever shipped one, saying it tracks all ~314–340 EUV units worldwide [19][20][21]. [19][20][21]
  • Rohan Paul presents a 8.3x projected US hyperscaler spending advantage by 2027 as a structural US edge [18]; the broader thread treats Chinese AI model gains as real and accumulating across capability, open-weight distribution, and capital [1][14][8]. [18][1][14][8]
  • Lambert and open-weight advocates treat GLM-5.2's release as evidence that open-weight models are closing the frontier gap [10]; Milk Road AI argues open-source models will not win the AI race and the dominant open-weight narrative was built on a misread of DeepSeek's launch [13]. [10][13]
  • Rohan Paul treats Chinese models' OpenRouter gains as a structural competitive shift [5]; developer-community commentators argue the gains reflect spot-market price routing and are conditional rather than lasting preference [6]. [5][6]
  • SemiAnalysis's SMIC N+3 teardown confirms China reached TSMC N6-class logic density [17]; the same report finds the Kirin 9030 Pro trails current flagships, showing density achievement has not translated to comparable end-product performance. [17]

Sources

  1. [1] Chinese AI models hit 61% market share on OpenRouter - LinkedIn — reactive:chinese-ai-competitive-rise
  2. [2] China holds 6 out of top 9 spots in global AI model usage ranking ... — reactive:chinese-ai-competitive-rise
  3. [3] Chinese AI startups confront challenges as OpenAI ends API services in China — reactive:chinese-ai-competitive-rise
  4. [4] Chinese AI firms woo OpenAI users as US company plans ... - Reuters — reactive:chinese-ai-competitive-rise
  5. [5] American AI startups are routing far more app traffic to Chinese LLMs. — Rohan Paul Twitter (2026-06-08)
  6. [6] The AI model market is turning into an inference spot market. Developers are routing to whatever model gives the best co... — reactive:chinese-ai-competitive-rise (2026-06-08)
  7. [7] DAY 0 ALERT: @MiniMax_AI M3 is now available on HuggingFace & has been added to InferenceX. The M3 architecture has … — SemiAnalysis Twitter (2026-06-13)
  8. [8] GLM-5.2 is probably the most powerful text-only open weights LLM — Simon Willison (2026-06-17)
  9. [9] 😺 GLM 5.2 brings 1M context — The Neuron (2026-06-22)
  10. [10] GLM-5.2 is the step change for open agents — Interconnects (2026-06-22)
  11. [11] Vercel CEO: "Almost shocked" by GLM-5.2's coding ability. — Rohan Paul Twitter (2026-06-21)
  12. [12] GLM-5.2 Is The New Best Open Model — Zvi's AI Roundups (2026-06-22)
  13. [13] Open source models will not win the AI race (Save this). — Milk Road AI Twitter (2026-06-20)
  14. [14] DeepSeek takes the crown as China’s most valuable AI startup after a massive $7.4B raise at a $50B valuation. — Rohan Paul Twitter (2026-06-16)
  15. [15] DeepSeek is going heavy-asset. — SemiAnalysis Twitter (2026-06-10)
  16. [16] Everyone's talking about Elon Musk's AI1 satellite this week. Almost nobody noticed: China moved on space-based AI compu… — SemiAnalysis Twitter (2026-06-19)
  17. [17] Is SMIC N+3’s Metal Pitch Smaller than Intel 18A’s? — SemiAnalysis Twitter (2026-06-14)
  18. [18] China is growing very quickly in AI, but the scale difference is brutal, spending gap is enormous. — Rohan Paul Twitter (2026-06-22)
  19. [19] ASML just became the center of a US-China chip fight after Washington said it fears a banned EUV lithography tool may ha… — Rohan Paul Twitter (2026-06-19)
  20. [20] ASML denies US government report that its EUV chipmaking tool ... — reactive:chinese-ai-competitive-rise
  21. [21] If an EUV machine reached China, three years of export controls failed silently. ASML denies it - they track all 314 in ... — reactive:chinese-ai-competitive-rise (2026-06-21)
  22. [22] The US told ASML a banned EUV machine may be in China. ASML says it tracks all 340 units — none have ever reached China. — reactive:chinese-ai-competitive-rise (2026-06-21)
  23. [23] @TitanUranus @Nike_Noesis @econ_escape @NandoDF US export controls apply to ASML's EUV machines due to US IP from nation... — reactive:chinese-ai-competitive-rise (2026-06-14)
  24. [24] Congrats to @vllm_project & @lmsysorg for releasing MiniMax M3 428B on both the CUDA & ROCm stack on day 0! Mini… — SemiAnalysis Twitter (2026-06-13)
  25. [25] ASML said on Friday it has never shipped an extreme ultraviolet lithography machine to China, responding to a Bloomberg ... — reactive:chinese-ai-competitive-rise (2026-06-20)
  26. [26] DeepSeek has completed over 50 billion RMB in financing at a valuation exceeding $50 billion, per The Information. Found... — reactive:chinese-ai-competitive-rise (2026-06-16)