
ModelsLead story
DiG-bench and Faraday mark AI's slow climb toward recursive self-improvement
New benchmarks show frontier models cracking discovery games and replicating research papers, with Opus 5 and a 27B scientist model leading the way.
Jaeden SchaferEditor in Chief5 min read