Skip to content

The podium wall

Boards across, top models down. Each #1 wears its maker's logo.

Language models: ranks on every board
ChatReasoningCodingAgents and toolsVision
Model Text ArenaMultiChallengeAA IntelligenceHLEGPQA DiamondEpoch ECIARC-AGI-2LiveBenchMMLU-ProWebDev ArenaSWE-Bench ProSWE-bench VerifiedAider PolyglotBFCLMCP AtlasSearch ArenaVision ArenaDocument Arena
1Claude Fable 5.1
2GPT-6 Astra
3Claude Opus 5
4Claude Fable 5
5Claude Opus 5.5
6GPT-5.6 Sol
7Claude Opus 4.6
8Claude Opus 4.7
#1 with the vendor’s logo23 podium7 top 10 stale board, not counted

More views

Scores as published by each board on the capture date. Model names and logos belong to their owners; logos via logo.dev.

Questions

Which AI model is best right now?

It depends on the job, which is why Leaders Board shows every major leaderboard side by side. The leader of leaders at the top of the home page is the model that finishes highest across all fresh language-model boards (chat, reasoning, coding, agents and vision), scored 10 points for #1 down to 1 point for #10 on each board.

What is the best LLM leaderboard?

There is no single one. LMArena measures what people prefer in blind votes, Artificial Analysis and Epoch AI run independent evaluation suites, Scale SEAL runs private expert benchmarks such as Humanity's Last Exam, SWE-bench and SWE-Bench Pro test real coding, BFCL and MCP Atlas test tool use. Leaders Board tracks them all and shows where they agree.

How is the leader of leaders calculated?

Each board is reduced to one entry per model (the best variant), ranked, and the top 10 earn 10, 9, 8 ... 1 points. Points are added across the boards counted, and the podium score is points earned over points available, from 0 to 100. Boards not updated in 180 days are shown but not counted. The full method is on the Method page.

How often is it updated?

Every day at 04:30 UTC, straight from each board's own page, data file or API. Each entry carries the date it was captured, and each board shows the date it last published.

Why do the same model names look different on each board?

Boards list variants: effort levels such as high or max, thinking modes, dated snapshots, and agent harnesses. Leaders Board maps them to one model and keeps the exact name each board published next to every entry.

Can I use the data?

Yes. /api/summary returns the computed standings as JSON, /llms-full.txt has every board's top 25 in plain text, and every number links back to the board that published it.

Weekly: the AI leaderboards with a new #1, Saturday mornings.