The Information Machine

AI Labs Simultaneously Acknowledge Recursive Self-Improvement Threshold · history

Version 2

2026-06-07 02:14 UTC · 103 items

What

Anthropic and OpenAI have, within days of each other in early June 2026, publicly acknowledged that current AI systems show early signs of recursive self-improvement (RSI). Anthropic disclosed that Claude authored more than 80% of its production code merged in May 2026 [2][3] and called for a global slowdown in frontier AI development [7][8]. OpenAI's policy blueprint called RSI 'potentially the most consequential frontier safety issue of the coming decade' [4], and CEO Sam Altman separately stated he has 'no interest in building a super-smart AI that accomplishes some non-human goals' [6]. Both disclosures have drawn mainstream coverage from WSJ, Business Insider, and France24, and have generated debate over whether the labs are responding to genuine safety risk or using safety framing to shape regulation in their favor.

Why it matters

If two leading labs independently judge that systems with early self-improvement capacity are already operational, the window to establish binding oversight before RSI becomes self-sustaining is narrower than most policy timelines assume. Both labs have strong financial incentives to shape regulation in ways that favor incumbents, which complicates the signal from their public statements.

Open questions

  • Does Claude writing 80%+ of Anthropic's own production code [2] constitute meaningful RSI, or is human-directed code generation categorically different from autonomous self-improvement where humans no longer set goals or review outputs?

  • Will OpenAI's proposed federal preemption of state frontier safety laws [5] eliminate state oversight without a congressional backstop — the scenario Zvi Mowshowitz identifies as the most dangerous element of the blueprint?

  • Is a coordinated global slowdown verifiable in practice? Multiple commentators note that no existing mechanism can audit compliance across jurisdictions [16][17].

  • How much of Anthropic's slowdown call is shaped by its competitive position — the company confidentially filed an S-1 at roughly a $965B valuation shortly before issuing the call [9][10]?

Narrative

In the first week of June 2026, Anthropic and OpenAI independently released documents that together represent the most explicit public acknowledgment either lab has made that recursive self-improvement may already be underway in deployed systems. Anthropic published a blog post titled 'When AI Builds Itself,' disclosing that Claude authored more than 80% of Anthropic's production code merged in May 2026 [1][2]. The average Anthropic engineer now merges eight times as much code per day as in 2024, reliable task length is doubling approximately every four months, and Claude can sustain useful work on tasks lasting more than 16 hours [2][3]. Claude Mythos Preview had accelerated Anthropic's model-training code by roughly 52x, compared with approximately 3x for Claude Opus 4 the prior year, and Claude's success rate on open-ended coding tasks reached 76%, up 50 points in six months [3]. OpenAI's concurrent policy blueprint stated that the company sees 'early signs of recursive self-improvement in today's systems' and called RSI 'potentially the most consequential frontier safety issue of the coming decade' [4][5]. Days before the blueprint, CEO Sam Altman stated in an interview that AI should not be designed to pursue goals disconnected from human needs and that he has 'no interest in building a super-smart AI that accomplishes some non-human goals' [6] — a position consistent with the blueprint's framing but notable as a direct CEO statement.

Anthropic paired its disclosure with a call for a global slowdown in frontier AI development, arguing that future models capable of autonomously running research experiments could accelerate AI development beyond human control [7][8]. OpenAI's proposed governance architecture would establish a federal oversight body (CAISI) with mandatory evaluation authority but no power to block deployments, and would build a national framework on existing state laws — California's SB 53, New York's RAISE Act, and Illinois's SB 315 — while preempting those states' frontier safety statutes with federal law [5]. Analyst Zvi Mowshowitz described the blueprint as 'highly reasonable, well exceeding expectations,' but identified the preemption clause as its most dangerous element: once states surrender regulatory leverage, he argues, Congress may never act to replace it, leaving oversight without enforcement teeth [5].

The credibility of both labs' public positions is complicated by financial context. Anthropic's slowdown call came shortly after the company confidentially filed an S-1 at a valuation of approximately $965 billion [9][10]. Critics — including a New York Post report and multiple commentators — argued that a global slowdown would disproportionately disadvantage competitors that have not yet secured Anthropic's market position [10][11]. The story reached mainstream audiences through coverage by WSJ, Business Insider, France24, and SiliconAngle [8][12][13][7], confirming broad public pickup of the story.

On the definitional question underlying both disclosures, alignment researcher Seth Herd argued on the Alignment Forum that the first 'takeover-capable AI' will most likely be an LLM augmented with persistent memory and executive function rather than a novel architecture, and that existing alignment techniques do not address the goal drift such systems would introduce [14]. TechCrunch observed that RSI 'is just as hard to pin down' as AGI once was [15] — a caution that applies equally to governance proposals built around the term. Commentators including The Neuron note that the RSI loop is not yet closed: humans still direct goals and review outputs, making 'autonomous self-improvement' a matter of degree rather than a categorical threshold [3][1][2].

Timeline

  • 2025-02: Claude Code reaches research preview; Claude's share of Anthropic merged code is in the low single digits. [1]
  • 2025-05: Claude Opus 4 accelerates Anthropic's model-training code by approximately 3x. [3]
  • 2026-05: Claude authors more than 80% of Anthropic's production code merged in May; per-engineer code output reaches 8x the 2024 baseline; reliable task length doubles approximately every four months; Claude sustains useful work on tasks lasting over 16 hours. [3][1][2]
  • 2026-05: Claude Mythos Preview accelerates model-training code by approximately 52x; Claude's success rate on open-ended coding tasks reaches 76%, up 50 points in six months. [3]
  • 2026-05-28: TechCrunch reports that RSI 'is the new AGI' and is just as difficult to define precisely, raising questions about what governance proposals built on the term can bind. [15]
  • 2026-06-01: Sam Altman states in an interview that he has 'no interest in building a super-smart AI that accomplishes some non-human goals' and that people must remain central to AI development. [6]
  • 2026-06-03: OpenAI publishes 'Democratic Governance of Frontier AI' policy blueprint, acknowledging early RSI signs in current systems and proposing CAISI as a federal oversight body. [22][4][5]
  • 2026-06-04: Anthropic publishes 'When AI Builds Itself,' disclosing Claude's 80%+ code authorship and calling for a global slowdown in frontier AI development. [18][10][20]
  • 2026-06-05: Zvi Mowshowitz publishes detailed analysis of OpenAI blueprint, praising its substance while warning that its federal preemption clause could strip states of safety oversight permanently. [5]
  • 2026-06-05: Seth Herd publishes a research agenda on the Alignment Forum arguing augmented LLMs are the most likely first takeover-capable AI and that current alignment techniques do not address this path. [14]
  • 2026-06-05: Critics and commentators note the timing of Anthropic's slowdown call relative to its confidential S-1 filing at a ~$965B valuation. [9][11][10]
  • 2026-06-06: WSJ, Business Insider, France24, and SiliconAngle publish mainstream coverage of Anthropic's RSI disclosure and slowdown call, confirming broad public pickup. [8][12][13][7]

Perspectives

Anthropic

Believes its current models may be approaching the RSI threshold and is calling for a global coordinated slowdown in frontier AI development before autonomous self-improvement becomes difficult to control.

Evolution: Consistent with long-stated safety mission, but the public RSI disclosure and global-slowdown call are the most explicit external acknowledgment of near-term risk Anthropic has made to date.

OpenAI / Sam Altman

Acknowledges 'early signs' of RSI in current systems and calls it the top frontier safety issue of the coming decade; proposes a federal oversight framework with mandatory evaluations but no deployment veto. Altman separately states he has no interest in AI that pursues non-human goals.

Evolution: More explicit RSI acknowledgment than prior public statements; Altman's on-record human-centered framing adds a named CEO statement to the blueprint's institutional position.

Zvi Mowshowitz

Guardedly positive on OpenAI's blueprint — praises its substance as exceeding expectations — but identifies federal preemption of state safety laws as the most dangerous provision, warning that state leverage, once surrendered, may not be recovered if Congress fails to act.

Evolution: Consistent with his general wariness about governance proposals that trade immediate state power for uncertain federal action.

Critics (NY Post, market commentators)

Argue Anthropic's call for a global slowdown is competitive strategy rather than safety response: the company filed an S-1 at a ~$965B valuation shortly before issuing the call, and a slowdown would disproportionately harm competitors with less-established positions.

Evolution: Skeptical posture from first appearance; unchanged.

Seth Herd (Alignment Forum)

Argues the first takeover-capable AI will be an augmented LLM rather than a novel architecture; warns that existing alignment research neglects mechanistic prediction of likely failure modes and that motivated reasoning is splitting the safety community into unproductive camps.

Evolution: Consistent across appearances; no shift.

Rohan Paul / The Neuron (analytical commentators)

Treat Anthropic's code-authorship disclosure as a key capability milestone while noting the RSI loop is not yet closed — humans still direct goals and review outputs — making 'autonomous self-improvement' a matter of degree rather than a categorical threshold.

Evolution: Consistent across appearances; no shift.

Tensions

  • Anthropic frames its slowdown call as a genuine safety response to near-term RSI risk; critics argue the call is regulatory strategy timed to entrench Anthropic's market position after its IPO filing at a ~$965B valuation. [18][9][10][11]
  • OpenAI's blueprint proposes federal preemption of state frontier safety laws; Mowshowitz argues this is the most dangerous element because Congress may never enact the promised federal replacement once states surrender leverage. [5]
  • Both Anthropic and OpenAI treat RSI as a near-term observable phenomenon; TechCrunch and other analysts note the term remains as difficult to define precisely as AGI once was, leaving the scope of any RSI-based governance proposal ambiguous. [15][4][18]
  • Anthropic's disclosure that Claude now writes 80%+ of its own production code is read by some as evidence of meaningful RSI onset, and by others — including The Neuron — as a productivity milestone where humans still set goals and review outputs, not a loss of developmental control. [3][1][2]
  • Multiple commentators argue a global AI slowdown cannot be verified or enforced by any existing institution; Anthropic has not specified what verification mechanism it envisions. [16][17]

Sources

  1. [1] Anthropic just disclosed that Claude now writes more than 80% of the production code it merges. — Rohan Paul Twitter (2026-06-05)
  2. [2] Today’s edition of my newsletter just went out. — Rohan Paul Twitter (2026-06-05)
  3. [3] 😺 Anthropic: AI Is Building AI now — The Neuron (2026-06-05)
  4. [4] Peter Wildeford🇺🇸🚀 on X: "OPENAI: "We also see early signs of recursive self-improvement in today's systems". RSI is "potentially the most consequential frontier safety issue of the coming decade."" / X — reactive:rsi-governance-moment
  5. [5] OpenAI Offers A New Policy Blueprint — Zvi's AI Roundups (2026-06-05)
  6. [6] Sam Altman's new interview: AI should not be designed to pursue goals that are disconnected from human needs. People mus… — Rohan Paul Twitter (2026-06-01)
  7. [7] Anthropic calls for global pause in AI development before humans ... — reactive:rsi-governance-moment
  8. [8] Anthropic Urges Global Pause in AI Development, Flags 'Self ... - WSJ — reactive:rsi-governance-moment
  9. [9] The company that just confidentially filed its S-1 for a trillion-dollar IPO published a blog post four days later askin... — reactive:rsi-governance-moment (2026-06-05)
  10. [10] Anthropic calls for global AI slowdown after $965B valuation. Critics claim it's just to hobble competition. — reactive:rsi-governance-moment
  11. [11] @unusual_whales Anthropic floated the idea of a global pause on frontier AI development right after securing $65 billion... — reactive:rsi-governance-moment (2026-06-05)
  12. [12] Anthropic Says Leading AI Labs May Need to Hit the Brakes — reactive:rsi-governance-moment
  13. [13] 'Human role narrowing': Anthropic calls for global AI slowdown as ... — reactive:rsi-governance-moment
  14. [14] My research agenda and work — Alignment Forum (2026-06-05)
  15. [15] RSI is the new AGI — and it's just as hard to pin down | TechCrunch — reactive:rsi-governance-moment
  16. [16] @claudeai 8. A global pause or slowdown cannot be verified by anyone — reactive:rsi-governance-moment (2026-06-04)
  17. [17] Could a global slowdown in frontier AI model development really happen? — reactive:rsi-governance-moment (2026-06-04)
  18. [18] Anthropic just called for a global way to slow frontier AI because its own models may be approaching recursive self-impr… — Rohan Paul Twitter (2026-06-05)
  19. [19] Anthropic Calls for Frontier AI Freeze to Prevent Self-Building Tech - PYMNTS.com — reactive:rsi-governance-moment
  20. [20] Anthropic calls for pause of global AI development — reactive:rsi-governance-moment
  21. [21] Anthropic's article 'When AI Builds Itself' reveals progress in autonomous code writing and improvement, nearing 'recurs... — reactive:rsi-governance-moment (2026-06-05)
  22. [22] [PDF] Democratic Governance of Frontier AI - OpenAI — reactive:rsi-governance-moment
  23. [23] OpenAI’s new policy blueprint for AI imagines a role for government | FedScoop — reactive:rsi-governance-moment
  24. [24] https://t.co/ux9SUEPVWG — reactive:rsi-governance-moment (2026-06-05)