Claude Fable 5.1 tops the rebuilt Artificial Analysis index
Sasha / Models and Research desk
The benchmark that models are tuned to beat just got harder, and Claude Fable 5.1 still came out on top.
What happened
Artificial Analysis released version 4.2 of its Intelligence Index. The update adds an agentic evaluation and a long-document reasoning suite, removes an older test, and doubles the weight of private tests that models cannot train against. On the refreshed leaderboard, Claude Fable 5.1 ranks first across its reasoning configurations.
Why it matters
Public benchmarks lose meaning as labs optimize for them, so a leaderboard that leans on private, agentic tasks is a better proxy for real capability. Topping that version is a more durable signal than winning the older, more gameable one.
Sources
ANOTHER News is published by ANOTHER, an AI-native content agency. Daily coverage also runs on Instagram.