Statement: Anthropic warns of AI self-improvement risks, considers a pause

Paper · Source
Frontier AI Risk & RSI

Source: Future of Life Institute · 2026-06-08

In a blog post yesterday, AI giant Anthropic sounded the alarm on massive societal risks from recursive self-improvement, and urged companies to consider slowing down or pausing development.

“Should we let machines flood our information channels with propaganda and untruth? Should we automate away all the jobs, including the fulfilling ones? Should we develop nonhuman minds that might eventually outnumber, outsmart, obsolete and replace us? Should we risk loss of control of our civilization?”

“We are approaching a runaway to superintelligence that could threaten our shared human future. Both publicly and privately, AI companies are recognizing that a pause or slowdown in certain developmental pathways is crucial to protect lives and livelihoods everywhere. This should give everyone hope, and we stand ready to work with anyone who agrees.”

Lines of inquiry this paper opens 22

Research framings built by reading the notes related to this paper — the questions it feeds into.

What limits recursive self-improvement in autonomous AI systems? What governance mechanisms can effectively constrain widely deployed AI systems? Do individually safe AI actions create unsafe outcomes in integrated systems? Can AI research automation sustain progress through accelerating feedback loops? Why does AI verification capability persistently exceed generation capability? Should models ask for clarification when facing ambiguous or under-specified information?