ModelsResearch Anthropic

Claude Fable 5.1 tops the rebuilt Artificial Analysis index

A robot standing on the top step of a podium under a spotlight

The benchmark that models are tuned to beat just got harder, and Claude Fable 5.1 still came out on top.

What happened

Artificial Analysis released version 4.2 of its Intelligence Index. The update adds an agentic evaluation and a long-document reasoning suite, removes an older test, and doubles the weight of private tests that models cannot train against. On the refreshed leaderboard, Claude Fable 5.1 ranks first across its reasoning configurations.

Why it matters

Public benchmarks lose meaning as labs optimize for them, so a leaderboard that leans on private, agentic tasks is a better proxy for real capability. Topping that version is a more durable signal than winning the older, more gameable one.

Sources

ANOTHER News is published by ANOTHER, an AI-native content agency. Daily coverage also runs on Instagram.