LLM-Rangliste
Die definitive
LLM-Rangliste
Vergleichen Sie 25+ Sprachmodelle across 10 Benchmarks. Rankings, Preise, Geschwindigkeit und Reasoning — alles an einem Ort.
GPT-4oGPT-4o Minio3o4-miniClaude 3.5 SonnetClaude 3.5 HaikuClaude Opus 4Claude Sonnet 4Gemini 2.5 ProGemini 2.5 FlashGemini 2.0 FlashLlama 3.1 405BLlama 3.1 70BLlama 3.3 70BLlama 4 MaverickMistral LargeMixtral 8x22BDeepSeek V3DeepSeek R1Qwen 2.5 72BQwen 3 235BCommand R+Grok 2Grok 3Phi-4GPT-4oGPT-4o Minio3o4-miniClaude 3.5 SonnetClaude 3.5 HaikuClaude Opus 4Claude Sonnet 4Gemini 2.5 ProGemini 2.5 FlashGemini 2.0 FlashLlama 3.1 405BLlama 3.1 70BLlama 3.3 70BLlama 4 MaverickMistral LargeMixtral 8x22BDeepSeek V3DeepSeek R1Qwen 2.5 72BQwen 3 235BCommand R+Grok 2Grok 3Phi-4
25+
Verfolgte Modelle
10
Benchmarks
25+
Anbieter
Wöchentlich
Datenaktualisierung
Top-Modelle
Alle ansehen →| # | Modell | Geschwindigkeit |
|---|---|---|
| 1 | Claude Opus 4 Anthropic | 45 t/s |
| 2 | Gemini 2.5 Pro Google | 85 t/s |
| 3 | o3 OpenAI | 58 t/s |
| 4 | DeepSeek R1 DeepSeek | 32 t/s |
| 5 | o4-mini OpenAI | 145 t/s |
| 6 | Claude Sonnet 4 Anthropic | 79 t/s |
| 7 | GPT-4o OpenAI | 110 t/s |
| 8 | Grok 3 xAI | 55 t/s |
| 9 | Claude 3.5 Sonnet Anthropic | 82 t/s |
| 10 | Llama 4 Maverick Meta | 120 t/s |
Kategorien
Nach Aufgabe browsen
Chat
General conversation, instruction following, and helpfulness.
24 models
Code
Code generation, debugging, and software engineering tasks.
18 models
Reasoning
Mathematical reasoning, science, and logical problem solving.
20 models
Image
Image understanding, generation, and visual reasoning.
12 models
Video
Video understanding and generation capabilities.
8 models
Agent
Tool use, multi-step planning, and autonomous task completion.
15 models
FAQ