📚 Best AI Models for Knowledge
Ranking of language models by factual accuracy and knowledge depth. These models minimize hallucinations and provide reliable information.Updated August 2026. Based on independent benchmark data.
Why This Matters
Choosing the right model for knowledge is about more than raw benchmark scores. The models below are ranked using a weighted blend of category-specific benchmarks, arena Elo, and the intelligence index — but real-world performance depends on your context window, latency budget, and integration path.
For knowledge, we weighted the benchmarks that most strongly predict success in production: task accuracy first, then speed and cost efficiency. Leaders on this page consistently outperform on both standardized tests and real-world workloads.
If you operate under a strict budget, check the Best Value LLMs ranking. If latency matters more than raw quality, see Fastest LLMs. Both pages use the same underlying dataset with different weightings.
| # | Model | Provider | Score | Speed | Price (output) |
|---|---|---|---|---|---|
| 1 | Claude Mythos Preview | Anthropic | 100.0 | 80 t/s | $15/M |
| 2 | Gemini 3.1 Pro | 97.6 | 136 t/s | $12/M | |
| 3 | Gemini 3 Pro | 92.9 | 120 t/s | $5/M | |
| 4 | Claude Mythos 5 | Anthropic | 75.0 | 50 t/s | $50/M |
| 5 | Qwen3.7 Max | Alibaba | 69.3 | 203 t/s | $7.5/M |
| 6 | Claude Sonnet 4.6 (max) | Anthropic | 67.0 | 50 t/s | $15/M |
| 7 | GPT-5.5 Pro | OpenAI | 50.0 | 60 t/s | $10/M |
| 8 | Qwen3.6 Plus | Alibaba | 48.4 | 55 t/s | $3/M |
| 9 | Gpt 5 6 Terra High | OpenAI | 45.0 | 90 t/s | $10/M |
| 10 | Claude Sonnet 4 6 Adaptive | Anthropic | 45.0 | 80 t/s | $15/M |
| 11 | Gpt 5 6 Luna High | OpenAI | 45.0 | 90 t/s | $10/M |
| 12 | Gpt 5 6 Terra Medium | OpenAI | 45.0 | 90 t/s | $10/M |
| 13 | Deepseek V4 Pro 0424 | DeepSeek | 45.0 | 60 t/s | $1.1/M |
| 14 | Claude Opus 4 6 Adaptive | Anthropic | 45.0 | 80 t/s | $15/M |
| 15 | Gpt 5 5 Low | OpenAI | 45.0 | 90 t/s | $10/M |
| 16 | Deepseek V4 Pro 0424 High | DeepSeek | 45.0 | 60 t/s | $1.1/M |
| 17 | Claude Opus 4 5 Thinking | Anthropic | 45.0 | 50 t/s | $15/M |
| 18 | Gpt 5 6 Terra Low | OpenAI | 45.0 | 90 t/s | $10/M |
| 19 | Gpt 5 2 Codex | OpenAI | 45.0 | 90 t/s | $10/M |
| 20 | Qwen3 6 Max | Alibaba | 45.0 | 85 t/s | $2/M |
Top 5 Knowledge — Quick Comparison
Methodology
Rankings are based on independent benchmark evaluations. The knowledge score is computed from multiple standardized benchmarks and normalized to 0–100. Data is updated regularly from public benchmark results.