The Information Machine

AI Labs Simultaneously Acknowledge Recursive Self-Improvement Threshold · history

Version 13

2026-06-22 18:26 UTC · 237 items

What

In June 2026, Anthropic and OpenAI independently published documents acknowledging that recursive self-improvement may be active in deployed AI systems and called for international coordination to slow frontier development [1][3][5]. The US government then issued an export control directive on June 12, 2026, requiring Anthropic to block global access to its two most advanced deployed models — Fable 5 and Mythos 5 — which Anthropic disabled for all customers on June 13 [15][16]. SK Telecom's inclusion in Anthropic's Project Glasswing program has been identified as the trigger for the directive [18]. Separately, Dario Amodei disclosed that internal testers of Anthropic's most powerful unreleased model recommended against releasing it, calling it a 'super weapon' — characterized as a metaphor from early testers, not a literal licensing requirement [20][21].

Why it matters

The forced global suspension of Fable 5 and Mythos 5 is the first known instance of a government compelling an AI lab to pull frontier models from commercial access worldwide, creating a real precedent for export-control-based AI governance [19]. The US regulatory posture toward Anthropic is incoherent: the same government that embedded Anthropic engineers at the NSA for offensive cyber simultaneously ordered Anthropic's most capable deployed models offline globally [7][16].

Open questions

  • What specific capabilities of Fable 5 or Mythos 5 triggered the June 12 directive, and does SK Telecom's Project Glasswing involvement explain the timing or only the jurisdiction? [18][16]

  • Will Anthropic be permitted to re-enable Fable 5 and Mythos 5 after satisfying export control requirements, or does the suspension reflect a capability threshold the government considers permanently off-limits for global access? [17][19]

  • Anthropic has not responded to Jeremy Howard's argument that its continued use of frontier models for its own research internally negates its slowdown call [11] — does the Fable 5/Mythos 5 suspension change that calculus?

  • If Altman sees a major RSI breakthrough as reason to stay private [13][14], does the forced Anthropic model suspension change OpenAI's IPO or governance stance?

Narrative

In June 2026, Anthropic and OpenAI independently published documents acknowledging that recursive self-improvement may already be underway in deployed AI systems. Anthropic's 'When AI Builds Itself' disclosed that Claude authored more than 80% of Anthropic's production code merged in May 2026, that per-engineer code output reached 8x the 2024 baseline, and that Claude Mythos Preview accelerated model-training code by approximately 52x [1][2]. Claude's open-ended coding success rate reached 76%, up 50 points in six months [2]. OpenAI's concurrent policy blueprint called RSI 'potentially the most consequential frontier safety issue of the coming decade' and cited 'early signs of recursive self-improvement in today's systems' [3]. Google DeepMind researchers separately identified four technical pathways through which AGI could transition to ASI [4]. All three labs endorsed international coordination to slow frontier AI development [5].

Each lab's public stance sits in tension with other conduct. OpenAI proposed a federal oversight body (CAISI) with mandatory evaluation authority; the US government subsequently directed CAISI to stop publishing public AI model evaluations [6]. Anthropic's safety messaging centers on a global slowdown call, but Anthropic engineers are reported embedded at the NSA using Claude Mythos for offensive cyber [7][8], and critics have argued the slowdown call is competitive strategy timed to Anthropic's confidential S-1 filing at a roughly $965 billion valuation [9][10]. Jeremy Howard argued that the only internally consistent form of Anthropic's slowdown call would prohibit the lab with the top-ranked model from using it for frontier research — a test Anthropic fails — and Anthropic has not responded [11]. Sam Altman endorsed coordinated slowdown mechanisms publicly [12] while telling staff that a major RSI breakthrough would favor OpenAI staying private since some work is easier without public-market pressure [13][14].

The most concrete government action arrived on June 12, 2026, when the US issued an export control directive requiring Anthropic to block global access to Fable 5 and Mythos 5, its two most advanced deployed models. Anthropic disabled both globally for all customers on June 13 [15][16][17]. SK Telecom's inclusion in Anthropic's Project Glasswing program has been identified as the proximate trigger [18]. This created a potential precedent for export-control-based AI governance — a government compelling a lab to pull frontier models from global commercial access [19]. Separately, Dario Amodei disclosed on June 19 that internal testers of Anthropic's most powerful unreleased model recommended against releasing it at all, describing it as a 'super weapon' that should require a license to operate; that characterization has since been clarified as a metaphor from early testers rather than a literal licensing position [20][21].

Jack Clark's Import AI declared 'alignment is not on track,' arguing that RSI without adequate alignment means 'rolling very scary dice' [22]. Geoffrey Irving founded Sequent — backed by researchers from the UK AI Security Institute and Timaeus, targeting $100-150M and 40-80 employees — on the argument that current lab alignment approaches are too empirical and reactive to provide principled safety confidence before ASI [23][22]. Altman predicted AI will conduct a significant fraction of OpenAI's own research by March 2028 [24]; OpenAI confidentially filed for an IPO and Altman told staff the company expects to go public within the next year [25][26].

Timeline

  • 2026-05: Claude authors 80%+ of Anthropic's production code; per-engineer output reaches 8x the 2024 baseline; Claude Mythos Preview accelerates model-training code ~52x; Claude's coding success rate reaches 76%, up 50 points in six months. [2][34][1]
  • 2026-06-03: OpenAI publishes 'Democratic Governance of Frontier AI,' acknowledging early RSI signs and proposing CAISI as a federal oversight body with mandatory evaluation authority. [33][3][30]
  • 2026-06-04: Anthropic publishes 'When AI Builds Itself,' disclosing Claude's 80%+ code authorship and calling for a global coordinated slowdown in frontier AI development. [27][35][36]
  • 2026-06-05: Critics note timing of Anthropic's slowdown call relative to its confidential S-1 filing at roughly $965 billion valuation. [9][35][10]
  • 2026-06-08: Sam Altman predicts AI will conduct a significant fraction of OpenAI's research by March 2028; OpenAI confidentially files for an IPO. [24][25]
  • 2026-06-09: All three major labs publicly endorse international coordination to slow frontier AI development; NSA reported using Claude Mythos for offensive cyber with approximately six embedded Anthropic engineers. [5][7]
  • 2026-06-10: Geoffrey Irving founds Sequent; Jeremy Howard argues Anthropic's use of its own frontier models for research negates its slowdown call; Altman tells staff a major RSI breakthrough would favor staying private over pursuing the IPO. [23][11][13][14][26]
  • 2026-06-11: US government directs CAISI to stop publishing public AI model evaluations; Amodei tells Bloomberg AI progress is now in the sharp-acceleration phase of an exponential curve. [6][28]
  • 2026-06-12: Google DeepMind identifies four distinct technical pathways through which AGI could transition to ASI. [4]
  • 2026-06-12: US government issues export control directive requiring Anthropic to block global access to Fable 5 and Mythos 5; SK Telecom's inclusion in Project Glasswing identified as the trigger. [37][16][18]
  • 2026-06-13: Anthropic disables Fable 5 and Mythos 5 globally for all customers. [15][16][17]
  • 2026-06-15: Jack Clark's Import AI 461 declares 'alignment is not on track'; Sequent confirms plans to raise $100-150M and scale to 40-80 employees. [22]
  • 2026-06-17: Tech Policy Press asks whether the US has set an AI export control precedent by blocking Mythos from global access. [19]
  • 2026-06-19: Dario Amodei discloses that testers of Anthropic's most powerful unreleased model recommended against release, describing it as a 'super weapon'; the characterization is clarified as a metaphor from early testers, not a literal licensing requirement. [20][21]

Perspectives

Anthropic / Dario Amodei

Believes current models may be approaching the RSI threshold, calls for a global coordinated slowdown, and disclosed that testers of Anthropic's most powerful unreleased model recommended against releasing it; Amodei says AI progress is in the sharp-acceleration phase of an exponential. The forced global suspension of Fable 5 and Mythos 5 adds a government-enforcement dimension to the safety narrative.

Evolution: The concrete model suspension and 'super weapon' disclosure add specificity to the abstract RSI call, but Anthropic has still not responded to Howard's consistency challenge or the NSA offensive cyber reporting.

OpenAI / Sam Altman

Acknowledges 'early signs' of RSI and calls it the top frontier safety issue; official blog endorses coordinated slowdown mechanisms; Altman targets AI-conducted research at OpenAI by March 2028 and expects an IPO within a year, but told staff a major RSI breakthrough would favor staying private.

Evolution: IPO filing and Altman's conditional private framing sit in tension with the public slowdown endorsement; CAISI's suppression of public evaluations undercuts the transparency mechanism in OpenAI's own proposal.

Google DeepMind

Supports an international organization to enable coordinated slowdowns; researchers publish technical work identifying four pathways from AGI to ASI.

Evolution: Governance position consistent with prior reporting.

Jack Clark (Import AI)

Declares 'alignment is not on track' and argues RSI without adequate alignment means 'rolling very scary dice'; treats Anthropic's 8x code increase as preliminary RSI evidence.

Evolution: Stance has sharpened from observational to declarative: prior coverage noted the RSI productivity data; Import AI 461 makes a direct claim that alignment is insufficient relative to capability development.

Jeremy Howard

Argues the only internally consistent form of Anthropic's slowdown call would prohibit the lab with the top-ranked model from using it for frontier research; Anthropic fails this test and has not responded.

Evolution: Argument stands uncontested.

Geoffrey Irving / Sequent

Founded Sequent on the argument that current lab alignment approaches are too empirical and reactive to provide principled safety confidence before ASI; plans to raise $100-150M and employ 40-80 researchers across scalable oversight and learning theory.

Evolution: Funding targets and research scope now specified, giving Sequent more concrete shape than at founding.

Zvi Mowshowitz

Guardedly positive on OpenAI's blueprint but warns federal preemption of state safety laws is its most dangerous element; treats CAISI evaluation suppression as a significant transparency setback.

Evolution: Coverage consistent; concerned that governance mechanisms are losing enforcement teeth.

Critics and observers (NY Post, social media, decentralized-AI commenters)

Argue Anthropic's and OpenAI's slowdown calls are competitive strategy timed to entrench incumbents; the 'perfect pre-IPO narrative' frame continues to spread; the Fable 5/Mythos 5 suspension prompted observers to note that centralized AI can be 'shut off overnight' [31], strengthening arguments for decentralized alternatives.

Evolution: No new analytical positions; the incumbent-entrenchment frame and the new centralization-risk frame are both spreading without new evidence.

Tensions

  • Howard argues Anthropic's continued use of its top-ranked model for frontier research negates its slowdown call; Anthropic has not responded. [11]
  • Altman publicly endorses coordinated slowdown mechanisms while telling staff that a major RSI breakthrough would favor staying private — the two positions point in different directions on what OpenAI actually wants from RSI timing. [12][13][14]
  • OpenAI's blueprint proposed CAISI as a federal body with public evaluation authority; the US government directed CAISI to stop publishing public AI model evaluations, removing the transparency mechanism the proposal depended on. [33][6]
  • The US simultaneously embeds Anthropic engineers at the NSA for offensive cyber and issues export control directives that force Anthropic's most capable models offline globally — an incoherent regulatory posture toward the same company. [7][16][18]
  • Anthropic frames its slowdown call as a genuine safety response; critics argue it is regulatory strategy timed to a confidential S-1 filing at roughly $965 billion, designed to entrench the incumbent's market position. [27][9][10]
  • Anthropic withheld its most powerful unreleased model after testers called it a 'super weapon,' while simultaneously deploying Fable 5 and Mythos 5 commercially until the government forced them offline — suggesting the internal threshold for withholding a model differs from what Anthropic considered safe for public deployment. [20][15]

Sources

  1. [1] Today’s edition of my newsletter just went out. — Rohan Paul Twitter (2026-06-05)
  2. [2] 😺 Anthropic: AI Is Building AI now — The Neuron (2026-06-05)
  3. [3] Peter Wildeford🇺🇸🚀 on X: "OPENAI: "We also see early signs of recursive self-improvement in today's systems". RSI is "potentially the most consequential frontier safety issue of the coming decade."" / X — reactive:rsi-governance-moment
  4. [4] Beautiful paper from Google DeepMind. — Rohan Paul Twitter (2026-06-12)
  5. [5] Three Labs With a Plan and A Memorandum — Zvi's AI Roundups (2026-06-09)
  6. [6] AI #172: The First Fable — Zvi's AI Roundups (2026-06-11)
  7. [7] National Security Presidential Memorandum/NSPM-11 — reactive:rsi-governance-moment
  8. [8] NSA using Claude Mythos for 'offensive cyber operations,' report claims — says 'half-a-dozen' Anthropic engineers embedded inside the agency — reactive:rsi-governance-moment
  9. [9] The company that just confidentially filed its S-1 for a trillion-dollar IPO published a blog post four days later askin... — reactive:rsi-governance-moment (2026-06-05)
  10. [10] Anthropic has discovered the perfect pre-IPO narrative: its product is so powerful that it justifies a trillion-dollar v... — reactive:rsi-governance-moment (2026-06-07)
  11. [11] Quoting Jeremy Howard — Simon Willison (2026-06-10)
  12. [12] OpenAI's latest official blog says the world may need a way to coordinate "slowing frontier development when needed." ht… — Rohan Paul Twitter (2026-06-09)
  13. [13] Sam Altman tells OpenAI staff an IPO is planned next year, but recursive self-improvement would favor staying private · Digg — reactive:rsi-governance-moment
  14. [14] Sam Altman tells OpenAI staff an IPO is planned next year, but recursive self-improvement would favor staying private · Digg — reactive:rsi-governance-moment
  15. [15] Anthropic disabled its two most advanced models, Fable 5 and Mythos 5, globally for all customers on June 13, 2026, afte... — reactive:rsi-governance-moment (2026-06-15)
  16. [16] Anthropic to disable its most advanced AI models after US order ... — reactive:rsi-governance-moment
  17. [17] Anthropic suspends top AI models after U.S. export control order — reactive:claude-fable-5-mythos-launch
  18. [18] SK Telecom has been identified as the South Korean company whose inclusion in Anthropic's Project Glasswing programme tr... — reactive:rsi-governance-moment (2026-06-19)
  19. [19] Did the US Government Just Set An AI Export Precedent by Blocking ... — reactive:claude-fable-5-mythos-launch
  20. [20] Anthropic's CEO just went on record saying the people who tested their most powerful AI model came back asking them not … — Milk Road AI Twitter (2026-06-19)
  21. [21] @TannerLeidy @Polymarket Mythos doesn't literally require a gun license—it's a metaphor from early testers. — reactive:rsi-governance-moment (2026-06-17)
  22. [22] Import AI 461: "Alignment is not on track"; FrontierCode; and synthetic research interns — Import AI (2026-06-15)
  23. [23] Sequent: scale and automation for higher confidence in alignment — Alignment Forum (2026-06-10)
  24. [24] Sam Altman's new blog about OpenAI's future path says by March-2028 a significant fraction of its own research will be d… — Rohan Paul Twitter (2026-06-08)
  25. [25] @skaas777 @iamai_omni 不是终止IPO啦,OpenAI 6月8日刚 confidentially filed for IPO,只是 timing 还没定(他们自己说有些事 private 公司做更方便)。 — reactive:rsi-governance-moment (2026-06-11)
  26. [26] OpenAI CEO Sam Altman told employees this week that he expects the company to go public "sometime over the next year," ... — reactive:rsi-governance-moment (2026-06-11)
  27. [27] Anthropic just called for a global way to slow frontier AI because its own models may be approaching recursive self-impr… — Rohan Paul Twitter (2026-06-05)
  28. [28] Dario Amodei's new interview, says AI progress suddenly going crazy. — Rohan Paul Twitter (2026-06-11)
  29. [29] Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing — Import AI (2026-06-08)
  30. [30] OpenAI Offers A New Policy Blueprint — Zvi's AI Roundups (2026-06-05)
  31. [31] 👽 bittensor:native IS EXPLODING BECAUSE CENTRALIZED AI JUST PROVED IT CAN BE SHUT OFF OVERNIGHT — reactive:rsi-governance-moment (2026-06-16)
  32. [32] The real AI infrastructure story this week isn't a new model. It's the split between models you rent and systems you can... — reactive:rsi-governance-moment (2026-06-15)
  33. [33] [PDF] Democratic Governance of Frontier AI - OpenAI — reactive:rsi-governance-moment
  34. [34] Anthropic just disclosed that Claude now writes more than 80% of the production code it merges. — Rohan Paul Twitter (2026-06-05)
  35. [35] Anthropic calls for global AI slowdown after $965B valuation. Critics claim it's just to hobble competition. — reactive:rsi-governance-moment
  36. [36] Anthropic calls for pause of global AI development — reactive:rsi-governance-moment
  37. [37] In a US government export control directive issued June 12, 2026 ... — reactive:rsi-governance-moment