AI Labs Simultaneously Acknowledge Recursive Self-Improvement Threshold · history
Version 11
2026-06-15 08:36 UTC · 198 items
What
In June 2026, Anthropic and OpenAI independently published evidence that recursive self-improvement may be active in deployed systems, then joined Google DeepMind in calling for international mechanisms to slow frontier AI development [8]. Anthropic disclosed Claude authored more than 80% of its production code in May 2026 [1]; OpenAI called RSI 'potentially the most consequential frontier safety issue of the coming decade' [3]; and Dario Amodei told Bloomberg that AI development has entered the sudden acceleration phase of an exponential curve [7]. Three challenges undercut the labs' coordination framing: the US government directed CAISI to stop publishing public model evaluations [9]; Anthropic engineers are reported embedded at the NSA using Claude Mythos for offensive cyber [10]; and Jeremy Howard's unanswered argument that Anthropic's continued use of Claude Mythos for frontier research contradicts its own slowdown call [18].
Why it matters
Three labs publicly endorsing a coordinated slowdown while each continues frontier development concentrates agenda-setting power among the same incumbents who would define how and when any slowdown gets implemented. OpenAI's IPO filing and Altman's private framing — that an RSI breakthrough would favor staying private — suggests the company's institutional incentive may be to reach that threshold before going public, not to prevent it.
Open questions
Anthropic has not publicly responded to Howard's argument that its own use of Claude Mythos for frontier research contradicts its slowdown call [18] — does it have a principled distinction between its use and others'?
If Altman views a major RSI breakthrough as reason to stay private longer [14][15], does OpenAI's public endorsement of coordinated slowdowns reflect its actual institutional preference on RSI timing?
Without CAISI publishing public model evaluations [9], what enforcement or transparency mechanism does OpenAI's governance blueprint actually provide?
Can any international coordination mechanism for slowing frontier AI be verified across jurisdictions? All three labs call for one, but none has specified how compliance would be established or enforced [8].
Narrative
In early June 2026, Anthropic and OpenAI independently published documents acknowledging that recursive self-improvement may already be active in deployed systems. Anthropic's 'When AI Builds Itself' disclosed that Claude authored more than 80% of Anthropic's production code merged in May 2026, that per-engineer code output reached 8x the 2024 baseline, and that Claude Mythos Preview accelerated model-training code by approximately 52x [1][2]. Claude's success rate on open-ended coding tasks reached 76%, up 50 points in six months [2]. OpenAI's concurrent policy blueprint called RSI 'potentially the most consequential frontier safety issue of the coming decade,' stating the company sees 'early signs of recursive self-improvement in today's systems' [3][4]. Google DeepMind researchers separately identified four technical pathways through which AGI could transition to ASI [5], and research on a self-improving agent architecture described a loop in which one AI agent rewrites its own configuration and updates its model without human intervention [6]. In a Bloomberg Originals interview, Dario Amodei described AI development as following an exponential curve that appears static for extended periods before accelerating sharply, saying the current moment represents that sudden upswing [7].
On governance, Anthropic, OpenAI, and Google DeepMind have all endorsed international coordination mechanisms to slow frontier AI development [8]. OpenAI's blueprint proposed a federal oversight body (CAISI) with mandatory evaluation authority; the US government subsequently directed CAISI to stop publishing public AI model evaluations, removing the transparency function at the center of the proposal [9]. Analyst Zvi Mowshowitz praised the blueprint but identified federal preemption of state frontier safety laws as its most dangerous element: once states surrender regulatory leverage, Congress may never enact the promised federal replacement [4]. A parallel contradiction has received mainstream coverage: NSPM-11 effectively bans Anthropic from federal contracts while permitting unrestricted government use of any deployed AI model, and the NSA is reported to be using Claude Mythos for offensive cyber operations with approximately half a dozen embedded Anthropic engineers [10][11].
Sam Altman's internal position adds a business-strategy dimension. OpenAI confidentially filed for an IPO on June 8, and Altman told staff the company expects to go public within the next year [12][13]. He also told staff that a major RSI breakthrough would favor staying private, since certain work is easier without public-market pressure on revenue and profit [14][15]. Altman has separately predicted AI will conduct a significant fraction of OpenAI's own research by March 2028 [16], and OpenAI's official blog endorsed coordinated mechanisms to slow frontier development [17]. The three positions — accelerated internal research targets, public endorsement of slowdowns, and private acknowledgment that an RSI event would make staying private preferable — are not obviously consistent.
Two independent challenges to the labs' safety framing have emerged. Jeremy Howard argues the only internally consistent way to slow RSI is to prohibit the lab with the top-ranked model from using it for frontier AI research: Anthropic fails this test by using Claude Mythos for its own research while calling for others to restrict theirs, and has not publicly responded [18]. Geoffrey Irving founded Sequent on the argument that existing lab alignment approaches are 'predominantly empirical and reactive' and cannot yield principled confidence that safety generalizes to ASI [19]. Critics have also framed Anthropic's slowdown call as competitive strategy timed to the company's confidential S-1 filing at a valuation of approximately $965 billion, arguing a global slowdown disproportionately advantages incumbents already at the frontier [20][21][22].
Timeline
- 2026-05: Claude authors more than 80% of Anthropic's production code; per-engineer output reaches 8x the 2024 baseline; Claude Mythos Preview accelerates model-training code by approximately 52x; Claude's success rate on open-ended coding tasks reaches 76%, up 50 points in six months. [2][30][1]
- 2026-06-03: OpenAI publishes 'Democratic Governance of Frontier AI' policy blueprint, acknowledging early RSI signs and proposing CAISI as a federal oversight body with mandatory evaluation authority. [25][3][4]
- 2026-06-04: Anthropic publishes 'When AI Builds Itself,' disclosing Claude's 80%+ code authorship and calling for a global coordinated slowdown in frontier AI development. [23][21][24]
- 2026-06-05: Zvi Mowshowitz praises OpenAI blueprint but warns its federal preemption clause could strip states of safety oversight permanently; critics note the timing of Anthropic's slowdown call relative to its confidential S-1 filing at a ~$965B valuation. [4][20][28][21]
- 2026-06-07: Scientific American covers Anthropic's RSI warning; the 'perfect pre-IPO narrative' critique spreads broadly on social media. [31][22][32]
- 2026-06-08: Jack Clark treats Anthropic's 8x code increase as preliminary RSI evidence; Sam Altman publishes blog predicting AI will conduct a significant fraction of OpenAI's own research by March 2028; OpenAI confidentially files for an IPO. [27][16][12]
- 2026-06-09: OpenAI's official blog endorses coordinated mechanisms to slow frontier development; Zvi Mowshowitz reports three-lab convergence on international slowdown coordination; NSPM-11 bans Anthropic from government contracts while NSA uses Claude Mythos for offensive cyber with approximately half a dozen embedded Anthropic engineers reported. [17][8][10]
- 2026-06-10: Geoffrey Irving announces Sequent, founded on the argument that lab alignment approaches are too reactive to provide principled safety confidence before ASI; Jeremy Howard argues Anthropic's slowdown call is internally contradictory because Anthropic continues using Claude Mythos for frontier research. [19][18]
- 2026-06-10: Sam Altman tells OpenAI staff the company expects to go public within the next year, but that a major RSI breakthrough would favor staying private since some work is easier without public-market pressure. [29][14][15][13]
- 2026-06-11: Tom's Hardware and mainstream tech press pick up the NSA/Claude Mythos offensive cyber story; US government directs CAISI to stop publishing public AI model evaluations. [11][9]
- 2026-06-11: Dario Amodei tells Bloomberg Originals that AI progress follows an exponential curve that appears static before accelerating sharply, and says the current moment represents that sudden acceleration. [7]
- 2026-06-11: Research on a self-improving agent architecture — a loop in which one AI agent rewrites its own configuration and updates its model without human intervention — published. [6]
- 2026-06-12: Google DeepMind researchers identify four distinct technical pathways through which AGI could transition to ASI, including continued scaling of compute, model size, data, and test-time compute. [5]
Perspectives
Anthropic / Dario Amodei
Believes current models may be approaching the RSI threshold and calls for a global coordinated slowdown; Amodei told Bloomberg AI progress is entering the sharp-acceleration phase of an exponential curve; has not publicly responded to Howard's consistency challenge or the NSA offensive cyber reporting.
Evolution: Amodei's Bloomberg interview reinforces the exponential-acceleration framing and extends Anthropic's public communication, but does not change the stated position or address the outstanding contradictions.
OpenAI / Sam Altman
Acknowledges 'early signs' of RSI and calls it the top frontier safety issue of the decade; official blog endorses coordinated slowdown mechanisms; Altman targets AI-conducted research at OpenAI by March 2028; OpenAI has filed confidentially for an IPO expected within the next year, but Altman told staff a major RSI breakthrough would favor staying private.
Evolution: The IPO filing and Altman's conditional private framing add a business-strategy dimension that sits uneasily alongside the public slowdown endorsement; CAISI's suppression of public evaluations also undercuts the transparency argument in OpenAI's own blueprint.
Google DeepMind
Supports an international organization to enable coordinated slowdowns of frontier AI development; researchers are simultaneously publishing technical work identifying four pathways from AGI to ASI.
Evolution: Governance position is consistent with prior reporting; the AGI-to-ASI paper adds technical depth but does not change the public stance.
Jeremy Howard
Argues the only internally consistent way to slow RSI is to prohibit the lab with the top-ranked model from using it for frontier AI research; Anthropic fails this test and has not responded.
Evolution: No response from Anthropic; argument stands uncontested.
Geoffrey Irving / Sequent
Founded Sequent on the argument that current lab alignment approaches are empirical and reactive, cannot yield principled confidence that safety generalizes to ASI, and that an independent theory-focused organization is necessary.
Evolution: New entrant; no subsequent statements on record.
Zvi Mowshowitz
Guardedly positive on OpenAI's blueprint but identifies federal preemption of state safety laws as its most dangerous element; reports three-lab convergence on international coordination; treats CAISI evaluation suppression as a significant transparency setback.
Evolution: Coverage extended from blueprint analysis to CAISI suppression, consistently concerned about governance mechanisms losing their teeth.
Jack Clark (Import AI)
Treats Anthropic's 8x code increase as 'preliminary evidence of prosaic recursive self-improvement' and connects RL reward hacking to the institutional rule systems governing societies.
Evolution: Consistent with long-standing editorial urgency on AI risk; no shift.
Critics (NY Post, social media commentators)
Argue Anthropic's slowdown call is competitive strategy timed to the S-1 filing at a ~$965B valuation; the 'perfect pre-IPO narrative' framing is the dominant skeptical frame and has spread widely.
Evolution: Spreading more broadly; no new analytical positions introduced.
Tensions
- Howard argues Anthropic's continued use of Claude Mythos for its own frontier research logically negates its slowdown call — the lab with the top model advances the frontier by using it; Anthropic has not responded. [18]
- Altman publicly endorses coordinated slowdown mechanisms while telling staff that a major RSI breakthrough would favor staying private to preserve strategic flexibility — the two positions point in different directions on what OpenAI actually wants from RSI timing. [17][29][14][15]
- OpenAI's blueprint proposed CAISI as a federal oversight body with public evaluation authority; the US government directed CAISI to stop publishing public AI model evaluations, removing the transparency function the proposal depended on. [25][9]
- Anthropic's public safety messaging centers on a global slowdown; Anthropic engineers are simultaneously reported embedded at the NSA using Claude Mythos for offensive cyber operations. [8][10][11]
- Anthropic frames its slowdown call as a genuine safety response to near-term RSI risk; critics argue it is regulatory strategy timed to entrench Anthropic's market position after its IPO filing at a ~$965B valuation. [23][20][21][22]
- Jack Clark treats Anthropic's 8x code increase as preliminary RSI evidence; other analysts read the same data as a productivity milestone where humans still set goals and review outputs, making 'autonomous self-improvement' a matter of degree rather than a categorical threshold. [2][30][1][27]
Sources
- [1] Today’s edition of my newsletter just went out. — Rohan Paul Twitter (2026-06-05)
- [2] 😺 Anthropic: AI Is Building AI now — The Neuron (2026-06-05)
- [3] Peter Wildeford🇺🇸🚀 on X: "OPENAI: "We also see early signs of recursive self-improvement in today's systems". RSI is "potentially the most consequential frontier safety issue of the coming decade."" / X — reactive:rsi-governance-moment
- [4] OpenAI Offers A New Policy Blueprint — Zvi's AI Roundups (2026-06-05)
- [5] Beautiful paper from Google DeepMind. — Rohan Paul Twitter (2026-06-12)
- [6] This paper shows an AI improving itself better when it rewrites its setup and updates its model. — Rohan Paul Twitter (2026-06-11)
- [7] Dario Amodei's new interview, says AI progress suddenly going crazy. — Rohan Paul Twitter (2026-06-11)
- [8] Three Labs With a Plan and A Memorandum — Zvi's AI Roundups (2026-06-09)
- [9] AI #172: The First Fable — Zvi's AI Roundups (2026-06-11)
- [10] National Security Presidential Memorandum/NSPM-11 — reactive:rsi-governance-moment
- [11] NSA using Claude Mythos for 'offensive cyber operations,' report claims — says 'half-a-dozen' Anthropic engineers embedded inside the agency — reactive:rsi-governance-moment
- [12] @skaas777 @iamai_omni 不是终止IPO啦,OpenAI 6月8日刚 confidentially filed for IPO,只是 timing 还没定(他们自己说有些事 private 公司做更方便)。 — reactive:rsi-governance-moment (2026-06-11)
- [13] OpenAI CEO Sam Altman told employees this week that he expects the company to go public "sometime over the next year," ... — reactive:rsi-governance-moment (2026-06-11)
- [14] Sam Altman tells OpenAI staff an IPO is planned next year, but recursive self-improvement would favor staying private · Digg — reactive:rsi-governance-moment
- [15] Sam Altman tells OpenAI staff an IPO is planned next year, but recursive self-improvement would favor staying private · Digg — reactive:rsi-governance-moment
- [16] Sam Altman's new blog about OpenAI's future path says by March-2028 a significant fraction of its own research will be d… — Rohan Paul Twitter (2026-06-08)
- [17] OpenAI's latest official blog says the world may need a way to coordinate "slowing frontier development when needed." ht… — Rohan Paul Twitter (2026-06-09)
- [18] Quoting Jeremy Howard — Simon Willison (2026-06-10)
- [19] Sequent: scale and automation for higher confidence in alignment — Alignment Forum (2026-06-10)
- [20] The company that just confidentially filed its S-1 for a trillion-dollar IPO published a blog post four days later askin... — reactive:rsi-governance-moment (2026-06-05)
- [21] Anthropic calls for global AI slowdown after $965B valuation. Critics claim it's just to hobble competition. — reactive:rsi-governance-moment
- [22] Anthropic has discovered the perfect pre-IPO narrative: its product is so powerful that it justifies a trillion-dollar v... — reactive:rsi-governance-moment (2026-06-07)
- [23] Anthropic just called for a global way to slow frontier AI because its own models may be approaching recursive self-impr… — Rohan Paul Twitter (2026-06-05)
- [24] Anthropic calls for pause of global AI development — reactive:rsi-governance-moment
- [25] [PDF] Democratic Governance of Frontier AI - OpenAI — reactive:rsi-governance-moment
- [26] https://t.co/ux9SUEPVWG — reactive:rsi-governance-moment (2026-06-05)
- [27] Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing — Import AI (2026-06-08)
- [28] @unusual_whales Anthropic floated the idea of a global pause on frontier AI development right after securing $65 billion... — reactive:rsi-governance-moment (2026-06-05)
- [29] Sam Altman is reportedly warning staff that recursive self-improvement (RSI) could delay its IPO. — Rohan Paul Twitter (2026-06-10)
- [30] Anthropic just disclosed that Claude now writes more than 80% of the production code it merges. — Rohan Paul Twitter (2026-06-05)
- [31] Anthropic warns AI may soon begin recursive self-improvement — reactive:rsi-governance-moment
- [32] Anthropic Calls for Global Slowdown in AI Development - Reddit — reactive:rsi-governance-moment