AI's Recursive Self-Improvement Problem
The people building the most powerful technology in history keep warning that it might destroy us, then returning to the work of making it more powerful.
It is tempting to dismiss that contradiction as marketing, yet the warnings are too specific for that explanation to hold. Dario Amodei signed a call for governments to create the technical and legal means to slow frontier development. OpenAI expressed support for the same goal. More than 1,100 AI researchers and executives, including senior figures from Anthropic, OpenAI, Google DeepMind and Meta, warned that automated AI development could move faster than our ability to understand or control the systems it produces.
This week, Jacob Coxon, a researcher who had worked at both OpenAI and Anthropic, resigned. He said the firms were "racing straight to self-improving superintelligence and gambling with our lives".
There are familiar explanations for why the labs do not simply stop: competitive pressure, national security, sunk cost, a sincere belief that someone less responsible will build the technology anyway, and plain hubris. Each may explain part of the behaviour, but none resolves the central problem. The people closest to the technology are asking governments to create constraints that they will not, or cannot, impose on themselves.
Recursive self-improvement makes that governance gap more serious. If AI systems become materially better at designing, training or evaluating their successors, the speed of capability growth could separate from the speed at which institutions can test it and decide whether it is safe. The issue is not only whether any one model is dangerous today. It is whether the organisations building the next generation retain meaningful control over the pace and consequences of the process.
Boards cannot treat this as a philosophical argument among researchers. If management says the technology could outrun human understanding while continuing to accelerate its development, directors should ask what evidence supports the claimed controls, which thresholds would actually slow or stop deployment, who can invoke them, and whether commercial incentives would survive the decision.
Reassurance from the people running the race is no substitute for independent evidence that the brakes work. For boards carrying the duty, verification has to outrank reassurance.
The Derek Thompson piece that prompted this: https://www.derekthompson.org/p/its-time-to-ask-the-big-question
#ArtificialIntelligence #AISafety #BoardGovernance