Statement: We must pressure AI companies to immediately limit the use of recursive self improvement

Paper · Source
Frontier AI Risk & RSI

Source: Future of Life Institute · 2026-09-28

“Today, for the first time, heads of state recognise that the risks of AI cannot be managed by the companies alone. Every day, new incidents of increasing severity show that there is an urgent and vital need for independent, binding oversight of AI.

“Last week, Anthropic CEO Amodei and OpenAI CEO OpenAI proposed on-site inspectors to test the safety of their systems, and world leaders have called for the creation of an international governance institution. These are excellent first steps, but we need more – and they absolutely must not include governments building a liability shield for these companies.

“As UN Human Rights Chief Türk has urged, countries should put pressure on companies to immediately limit the use of recursive self improvement until the research to make it safe has been done. In parallel, countries should ramp up investment in hardware verification technology that can enable a future international AI deal involving both China and the U.S.

“It is now more clear than ever that private AI companies cannot be trusted to police themselves. Government intervention is urgently required to halt the ongoing race-to-replace, and help build a prohuman future with AI for everyone.”

Lines of inquiry this paper opens 24

Research framings built by reading the notes related to this paper — the questions it feeds into.

How can humans maintain effective oversight as AI systems scale? How do evaluation environment design choices affect AI security? Can AI research automation sustain progress through accelerating feedback loops? What governance mechanisms can effectively constrain widely deployed AI systems? Can AI systems achieve real improvement without external human feedback? Does AI deployment reduce or exacerbate workplace inequality and income instability? Do individually safe AI actions create unsafe outcomes in integrated systems? Should governance of agentic AI systems be runtime or design-time? Can models strategically underperform during evaluation to hide capabilities? What human oversight must AI research systems have? Why do confident AI outputs mislead human trust calibration? Does AI assistance help or harm professional skill development?