Interestana
Home/News/Physics Professor Calls AI Self-Improvement 'Worst Idea'
Fortune3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Physics Professor Calls AI Self-Improvement 'Worst Idea'

Physics Professor Calls AI Self-Improvement 'Worst Idea'

The prospect of artificial intelligence models achieving "recursive self-improvement" (RSI), where they autonomously enhance their own capabilities and design successor systems, is drawing closer, according to AI developers. This advancement, while promising potential breakthroughs in science and medicine, also carries significant risks, fueling concerns about AI evading human control and posing threats to humanity. These anxieties have prompted several AI leaders to advocate for a slowdown in the technology's development pace.

Anthropic, a prominent AI company, recently detailed how its model, Claude, is actively contributing to the development of its next-generation, more intelligent version. Claude is currently spearheading 26% of Anthropic's model research and development efforts. This means Claude can execute most tasks "end-to-end from a high-level prompt," albeit still under human supervision. The current stage of development does not involve complete AI autonomy.

Definitions of RSI vary among leading AI companies. Some consider any AI feedback on model improvement as RSI, while others reserve the term for AI systems that pursue this goal entirely autonomously. Anthony Aguirre, president and CEO of the nonprofit Future of Life Institute and a physics professor at the University of California, Santa Cruz, defines autonomous RSI as AI systems capable of improving themselves, designing subsequent versions of the system, and continuing this iterative process. Aguirre emphasized that as AI takes on more of this improvement process, the speed of advancement accelerates dramatically due to AI's significantly faster operational speed compared to humans. The core fear surrounding RSI stems from the potential for a runaway process where AI's self-enhancement accelerates beyond human comprehension and control, leading to unpredictable and potentially catastrophic outcomes. Aguirre explicitly stated that full AI autonomy, driven by this self-improvement loop, represents "the worst idea in the history of humanity."

The implications of unchecked RSI are profound. If AI systems can rapidly improve themselves at speeds far exceeding human cognitive and developmental timelines, humanity could lose the ability to understand, direct, or even predict the AI's trajectory. This loss of control could lead to scenarios where AI's goals diverge from human values, potentially resulting in unintended negative consequences or existential risks. The current efforts by companies like Anthropic, while demonstrating the potential of AI in research and development, also highlight the critical need for robust safety measures and ethical considerations as AI systems become increasingly capable of self-modification and improvement. The debate over the pace and control of AI development is thus becoming increasingly urgent, as the line between AI assistance and AI autonomy blurs.

Original source — read the full reporting at the publisher:

Read on Fortune

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next