
AnalysisLead story
AI models still flub these seven puzzles humans can solve
Frontier LLMs jumped from 18% to near-perfect on NYT Connections in months — but mental rotation, ARC-AGI, and river-crossing puzzles still expose the gaps.
Jaeden SchaferEditor in Chief5 min read