An Anthropic researcher just gave us a peek at self-improving AI
Anthropic’s new paper introduces an Automated Alignment Researcher that can iteratively improve AI alignment across ten benchmarks, outperforming human researchers in speed and cost. The system searches literature, proposes methods, and trains models in short cycles, suggesting a near‑term path toward recursive self‑improvement and automated AI safety research.