
Anthropic warns AI may soon improve itself
Anthropic has warned that artificial intelligence development is advancing toward a stage where AI systems could eventually improve themselves without direct human involvement, raising concerns about safety and oversight.
In a blog post, Anthropic Institute lead Marina Favaro and Anthropic co-founder Jack Clark said the company is increasingly delegating AI development tasks to AI systems, accelerating research and shortening development cycles.
“For most of AI’s history, humans drove every step in its development cycle, but at Anthropic, we are delegating a growing share of AI development to AI systems themselves, which is speeding up our work,”
Favaro and Clark said.
The executives said AI models can already write code, delegate work to other agents and assist in developing future systems, creating a pathway toward fully autonomous successor models if computing power and capabilities continue improving.
Anthropic reported that AI model improvement is currently doubling roughly every four months, while its Claude model now authors about 80% of the code merged into the company's codebase, reducing the role of human developers.
The company argued that slowing frontier AI development could provide more time to address security, alignment and governance challenges, though it acknowledged that unilateral slowdowns could allow less cautious competitors to gain an advantage.
The warning comes as AI agents gain traction across industries, including crypto, where firms are exploring autonomous transaction settlement, with Circle CEO Jeremy Allaire previously predicting billions of AI agents could operate on behalf of users within the next five years.