Tracked AI Labs
OpenAIAnthropicGoogle DeepMindMeta AIxAIDeepSeekMistral AIMoonshot AIAlibaba CloudCohereMicrosoftAmazonZhipu AIBaiduOpenAIAnthropicGoogle DeepMindMeta AIxAIDeepSeekMistral AIMoonshot AIAlibaba CloudCohereMicrosoftAmazonZhipu AIBaidu
01 · Real-Time Rankings

Top Ranked Models

View All →

Top 10 Models by Overall Podium Score

View All →
#ModelPodium ScoreArena EloCodingSpeedPrice (out)
01
Claude Mythos Preview
Anthropic · Proprietary
97.4
80.980 t/s$15/M
02
Claude Fable 5
Anthropic · Proprietary
93.4
150984.171 t/s$50/M
03
Claude Fable 5.1
Anthropic · Proprietary
88.4
95.585 t/s$15/M
4
GPT-6 Astra
OpenAI · Proprietary
86.2
149867.282 t/s$50/M
5
Kimi K3
Moonshot AI · Open Weights
83.3
148561.537 t/s$15/M
6
Claude Opus 5
Anthropic · Proprietary
81.4
149270.460 t/s$25/M
7
GPT-5.6 Sol
OpenAI · Proprietary
80.8
148364.872 t/s$30/M
8
Qwen3.8 Max
Alibaba · Proprietary
79.8
149650.355 t/s$2/M
9
Claude Opus 4.8
Anthropic · Proprietary
76.0
148458.050 t/s$15/M
10
Claude Opus 4 6 Thinking
Anthropic · Unknown
75.0
50 t/s$15/M
02 · Evaluation Domains

Best by Use Case

View All →
03 · Research & Analysis

Latest Articles

All Articles

Open LLM Intelligence Platform

Find the perfect model for your workflow.

Compare accuracy, price-per-token, latency, and real-world agent performance across 700 tracked language models in seconds.

04 · FAQ

Frequently Asked Questions

All FAQ
Every model gets a Podium Score (0–100): a weighted blend of four signals — 35% arena Elo from head-to-head human battles, 30% benchmark average, 20% Artificial Analysis Intelligence Index and 15% LLM Stats score. All signals are normalized to a common scale, so the score is a plausible middle ground between five independent leaderboards.
Claude Mythos Preview leads the podium with a score of 97.4. Claude Fable 5 (93.4) and Claude Fable 5.1 (88.4) follow. Rankings shift weekly as new arena votes and benchmark results arrive.
We curate 700 flagship models across 18 categories (coding, math, reasoning, agentic, knowledge, multimodal, long context, speed, value and open weights), plus the full arena table with 726 models.
Five public leaderboards: Arena.ai, Artificial Analysis, LLM Stats, Vellum Leaderboard, LLMBase. Each source is re-synced weekly; see the Methodology page for the exact formula.
0 / 4