The Information Machine

AI Safety Advocacy Splits on US-China Cooperation vs. Domestic Controls · history

Version 4

2026-07-17 02:08 UTC · 43 items

What

Google DeepMind workers voted to unionize after the organization's classified Pentagon deal, extending TurnTrout's individual resignation into collective worker action [11][12][10]. The broader story involves two competing governance approaches: the AI Futures Project's Plan A, proposing a US-China cooperative pause on frontier AI [1], and the Trump administration's domestic competitive controls — export restrictions, state preemption, and a possible open-weight model ban [4][6]. Two contrasting institutional decisions now define the voluntary-commitment question: Google DeepMind signed a classified deal permitting unrestricted military AI use while publicly claiming its principles were unchanged [8]; Anthropic ended the same negotiation rather than waive its redlines [7].

Why it matters

The Google DeepMind unionization vote shows AI safety disagreements inside major labs now producing collective labor responses, not just individual departures — a different kind of institutional pressure than whistleblower accounts alone. Whether voluntary safety commitments are viable governance tools has moved from a theoretical question to one with contrasting empirical data points from two major labs facing identical pressure.

Open questions

  • Will the Google DeepMind union achieve formal recognition, and what specific demands regarding military contracts will workers codify? [11][12]

  • Will the White House issue an executive order banning frontier-capability open-weight models, and would it apply to non-US model releases? [6]

  • Can Plan A secure meaningful Chinese participation and verifiable compliance given current US-China competition, particularly following WAIC 2026 in Shanghai? [2][16]

  • Will OpenAI's reverse federalism strategy produce national AI safety standards before the Trump administration's August 2026 federal framework target? [14][5]

Narrative

The central divide in AI safety governance runs between advocates of a US-China cooperative pause and those pursuing unilateral competitive controls. The AI Futures Project's Plan A proposes joint chip supply controls, data center audits, and research sharing to slow superintelligent AI development [1]. Zvi Mowshowitz argues Plan A deserves serious engagement, with the key crux being whether superintelligence arrives soon enough to justify its costs [2]. Vitalik Buterin defends Plan A against critics who call it naive, arguing they apply coordination skepticism to a cooperative pause but not to the alternative — an unmanaged AI transition that concentrates power and eliminates human agency [2]. The Trump administration has not engaged with the cooperative frame: its January 2025 policy framework centers competitiveness, model weight export controls took effect January 2026, state AI regulations were preempted by executive order, and the White House is reportedly discussing banning frontier-capability open-weight models outright [3][4][5][6].

The question of whether AI safety commitments hold under institutional pressure has taken concrete form through contrasting institutional decisions. Anthropic ended Pentagon contract negotiations in February 2026 after the Defense Department insisted on blanket "anything lawful" usage rights, citing irreconcilable conflicts with its redlines on mass surveillance and autonomous weapons [7]. Google DeepMind took the opposite path: researcher Alex Turner (TurnTrout) published a detailed account of resigning after Google signed a classified Pentagon deal permitting "any lawful government purpose" without binding safety restrictions [8][9][10]. TurnTrout documents CEO Demis Hassabis removing specific prohibitions from Google's 2018 AI principles while publicly claiming "nothing's changed about our principles," and reports that IASEAI publicly promised a member poll supporting Anthropic's stance but then cancelled it without explanation [8]. After TurnTrout's account received broad media coverage, Google DeepMind workers voted to unionize — extending individual conscience into collective workplace action over military AI contracts [11][12].

A parallel domestic debate concerns open-weight models and the terms of AI regulation. Nathan Lambert reports White House discussions of an executive order banning open-weight models above current frontier capability levels, and accuses Anthropic's campaign against Chinese model distillation of regulatory capture — noting Anthropic would gain substantial commercial benefit if targeted Chinese model makers were banned [6]. Lambert disputes the security rationale by pointing to Anthropic's Mythos model being accessed through unauthorized Discord channels during private beta, arguing APIs are not meaningfully more secure than open weights [6]. An Alignment Forum post argues political will — not technical research — is the main AI safety bottleneck, citing a 3.6:1 researcher-to-advocate ratio in the US field and AI companies securing seven times as many European Commission meetings as civil society in 2023 [13].

OpenAI has published advocacy for "reverse federalism," supporting state AI safety laws in California, New York, and Illinois as a path to national standards ahead of formal federal legislation [14]. This conflicts with the Trump administration's state preemption executive order and its own federal framework targeting August 2026 [5][14]. The Center for American Progress characterizes Trump's AI executive order as a broad threat to state authority beyond AI specifically [15]. Plan A attracted attention at WAIC 2026, which opened July 17 in Shanghai [16].

Timeline

  • 2025-01-01: Trump administration releases national AI policy framework centering competitiveness rather than safety regulation as the governing principle. [3][18]
  • 2025-12-01: Trump signs executive order preempting state AI regulations to create a unified federal policy framework. [5]
  • 2026-01-09: US model weight export controls take effect, with commentary that the US had become the world's most aggressive AI regulator on this specific issue. [4]
  • 2026-02-26: Anthropic ends Pentagon contract negotiations after the Defense Department insists on blanket 'anything lawful' usage rights, citing irreconcilable conflicts with its redlines on mass surveillance and autonomous weapons. [7]
  • 2026-07-10: AI Futures Project publishes Plan A — a US-China cooperative pause on frontier AI — drawing substantive public debate including endorsements from Vitalik Buterin and Ryan Greenblatt. [1][2]
  • 2026-07-10: Zvi Mowshowitz publishes 'Plan B' analysis concluding the Trump administration will govern AI through ad hoc executive authority rather than formal licensing. [7]
  • 2026-07-11: Alignment Forum post argues political will — not research — is the main AI safety bottleneck, citing a 3.6:1 researcher-to-advocate ratio and AI industry's seven-to-one advantage in EU Commission meetings over civil society. [13]
  • 2026-07-12: Nathan Lambert reports White House discussions of an executive order to ban frontier-capability open-weight models and accuses Anthropic of regulatory capture in its campaign against Chinese model distillation. [6]
  • 2026-07-15: TurnTrout publishes account of leaving Google DeepMind over a classified Pentagon deal permitting unrestricted military AI use, documenting that Google's CEO removed safety prohibitions from its stated principles while publicly denying any change. [8][9][17]
  • 2026-07-15: OpenAI publishes advocacy for 'reverse federalism,' supporting state AI safety laws as a path to national standards ahead of formal federal legislation. [14]
  • 2026-07-16: Google DeepMind workers vote to unionize following TurnTrout's account of the classified Pentagon deal and subsequent coverage by Business Insider and The Verge. [11][10][12]
  • 2026-07-17: WAIC 2026 opens in Shanghai; Plan A attracts attention in the context of ongoing US-China AI governance debates. [16]

Perspectives

AI Futures Project (Daniel Kokotajlo) / Vitalik Buterin

Advocates a US-China cooperative pause on frontier AI via joint chip supply controls, data center audits, and research sharing; frames safety and growth as compatible. Buterin publicly defended Plan A, arguing critics apply coordination skepticism selectively — to a cooperative pause but not to the assumption that an unmanaged AI transition will go smoothly.

Evolution: Plan A has moved from initial publication to generating substantive debate with named endorsements and coverage at WAIC 2026 in Shanghai.

Zvi Mowshowitz

Not endorsing Plan A but argues it deserves serious engagement; frames the key crux as whether superintelligence arrives soon enough to justify Plan A's costs; concludes the Trump administration will govern AI through ad hoc executive authority rather than formal licensing.

Evolution: Shifted from skeptical analyst to engaged interlocutor stress-testing Plan A's premises rather than dismissing them.

Anthropic (Dario Amodei)

Holds hard redlines against mass surveillance and autonomous weapons targeting; ended Pentagon negotiations rather than waive them; publicly advocates a coordinated, verifiable pause on frontier AI development.

Evolution: TurnTrout's account positions Anthropic as the institutional counterexample to Google's capitulation; Nathan Lambert's regulatory capture accusation complicates Anthropic's safety framing.

TurnTrout (Alex Turner, former Google DeepMind researcher)

Argues AI safety pledges without binding enforcement are structurally inadequate; documents Google dropping its ethics principles under financial and political pressure while claiming otherwise; characterizes IASEAI as having failed to act when it counted.

Evolution: Account received broad media coverage in Business Insider and The Verge and is credited as a trigger for the Google DeepMind worker unionization vote.

Google DeepMind Workers

Voted to unionize in response to the classified Pentagon deal, seeking collective leverage over employer decisions about AI military applications.

Evolution: New voice in this thread; represents the first organized worker response to an AI lab's military contracting decision, following TurnTrout's individual departure.

Nathan Lambert / Open-Weight Advocates

Strongly opposed to any ban on open-weight frontier models; accuses Anthropic's anti-distillation campaign of regulatory capture; argues a unilateral US ban would be ineffective and open models improve safety through broad access.

Evolution: Position has sharpened from a diffuse stance into a specific alarm about imminent executive action naming Anthropic's lobbying as its target.

OpenAI

Advocates 'reverse federalism' — state AI safety laws in California, New York, and Illinois converging into a de facto national standard — while supporting CAISI as durable federal evaluation capacity; warns against regulatory fragmentation.

Evolution: Consistent in this thread; the August 2026 federal framework target increasingly puts OpenAI's preferred state-convergence timeline in direct conflict with the administration's timeline.

Trump Administration

Frames AI governance around US competitiveness; preempted state regulations; implemented model weight export controls; declined to build a formal licensing regime; reportedly in discussions about banning frontier-capability open-weight models; targeting August 2026 for a federal model-testing framework.

Evolution: Moving toward more aggressive domestic regulatory action on open weights while building a federal testing framework that conflicts with OpenAI's state-convergence strategy.

Tensions

  • TurnTrout documents Google DeepMind dropping its AI ethics principles under financial and political pressure while CEO Demis Hassabis publicly claimed 'nothing's changed about our principles'; Google DeepMind workers subsequently voted to unionize, while Anthropic's exit from the same Pentagon negotiation provides a contrasting institutional decision. [8][7][11][12]
  • Nathan Lambert argues Anthropic's campaign against Chinese model distillation is regulatory capture serving commercial interests; Anthropic frames the same campaign as a legitimate safety concern about frontier-capability proliferation. [6][7]
  • Plan A proponents and Vitalik Buterin argue a US-China cooperative pause is the necessary safety mechanism; the Trump administration treats China as a strategic competitor to contain through export controls, not a partner in cooperative governance. [2][1][3]
  • Open-weight advocates argue frontier model weights should be publicly released and open access improves safety; the US government treats frontier open weights as a credible national security risk and is reportedly considering an executive order to ban them. [6][4]
  • OpenAI argues state AI safety laws should converge into a de facto national standard ahead of federal legislation; the Trump administration preempted state AI regulations and is building its own federal framework targeting August 2026. [14][5]
  • TurnTrout and the Alignment Forum post argue voluntary pledges and research investments are insufficient without binding enforcement and political will; the mainstream AI safety field has historically prioritized technical research over advocacy and institutional accountability. [8][13]

Sources

  1. [1] 🟡 AI doom and bloom — Semafor Technology (2026-07-10)
  2. [2] Introduction for and Reactions to Plan A — Zvi's AI Roundups (2026-07-11)
  3. [3] Trump Administration Releases National AI Policy ... — reactive:ai-safety-governance-proposals
  4. [4] Ben Brooks on X: "Effective today, model weights are export controlled by Uncle Sam. This is a big deal. For all the smack talk about the EU, the US is now the world's most aggressive regulator of Expensive Maths. Here's my two cents on the model rule based on the released text (link below)." / X — reactive:ai-safety-governance-proposals
  5. [5] President Trump signs order attempting to block A.I. regulations at the state level — reactive:ai-safety-governance-proposals
  6. [6] 6 months to live for open models — Interconnects (2026-07-12)
  7. [7] AI #176 Part 2: Plan B — Zvi's AI Roundups (2026-07-10)
  8. [8] Why I Left Google DeepMind — Alignment Forum (2026-07-15)
  9. [9] Why I Left Google DeepMind - by Alex Turner - The Pond — reactive:ai-safety-governance-proposals
  10. [10] A DeepMind researcher resigned over its AI military deal — reactive:ai-safety-governance-proposals
  11. [11] Google DeepMind workers vote to unionise after classified ... — reactive:ai-safety-governance-proposals
  12. [12] Google DeepMind workers are unionizing over AI military ... — reactive:ai-safety-governance-proposals
  13. [13] The current bottleneck is political will, not research — Alignment Forum (2026-07-11)
  14. [14] The US is advancing AI safety through state and federal action — OpenAI Blog (2026-07-15)
  15. [15] President Trump's AI National Policy Executive Order Is an ... — reactive:ai-safety-governance-proposals
  16. [16] [7/11 17:00] AI Futures Project publishes "AI 2040: Plan A" / WAIC 2026 opens July 17 in Shanghai... — reactive:ai-safety-governance-proposals
  17. [17] Why I Left Google DeepMind - TurnTrout — reactive:ai-safety-governance-proposals
  18. [18] Artificial Intelligence for the American People — reactive:ai-safety-governance-proposals