AI That Improves Itself Just Got Cheaper Than Humans
This AI could work 24/7, fix itself, and cost $4 an hour vs. $150 for a human
Imagine an AI that can fix itself, improve its own performance, and do it faster and cheaper than a human expert. That’s exactly what a new system from Anthropic just demonstrated. In a research paper, the team showed how an AI model (dubbed the Automated Alignment Researcher) could automatically adjust its own training to overcome specific weaknesses. Over several rounds of testing, the AI improved its performance on every benchmark it was given, matching or beating the solutions proposed by experienced human researchers—all within six hours and at a fraction of the cost ($4 per hour vs. $150 per hour for a human).
The system works by treating itself like a scientist. It reviews existing research, proposes fixes for its own flaws, tests them for 30 minutes at a time, and keeps what works while discarding what doesn’t. This process repeats until it hits the best possible outcome. The implications are huge: if AI can improve its own training, it could accelerate AI development at an unprecedented scale, potentially leading to smarter, safer, and more capable systems in record time.
But there’s a critical limitation: the AI’s improvements are only as good as the benchmarks it’s given. If those benchmarks are flawed or miss important behaviors, the AI’s “fixes” might not actually make it better in the real world. The paper also notes that maintaining and expanding these benchmarks is an ongoing challenge, requiring human oversight to ensure the AI stays aligned with human goals.
For now, this is a glimpse into a future where AI systems might not just assist researchers but replace some of their most repetitive, time-consuming work. While it’s not about to make human researchers obsolete, it does raise questions about how quickly AI could evolve—and who gets to control the speed and direction of that progress.
- An AI system (Automated Alignment Researcher) improved its own training faster and cheaper than human experts in a test, beating their solutions in just six hours for $4 vs. $150.
- The system works like an automated scientist: it tests fixes, keeps what works, and discards what doesn’t—potentially speeding up AI development significantly.
- The catch? The AI’s improvements depend on human-set benchmarks; if those are wrong, the AI’s “fixes” might not actually make it better.
Why It Matters
This could speed up AI progress dramatically, but it also puts pressure on getting the goals right—before the AI optimizes itself into unexpected behaviors.