🤖 Agentic · llm-stats.com
Apex Agents
Long-horizon professional agent tasks with multi-app workflows.
4
Modelli testati
Kimi K3
Modello di punta
37.6%
Punteggio migliore
Moonshot AI
Provider
| # | Modello | Apex Agents |
|---|---|---|
| 1 | Kimi K3 Moonshot AI | 37.6% |
| 2 | Gemini 3.1 Pro Google | 33.5% |
| 3 | Kimi K2.6 Moonshot AI | 27.9% |
| 4 | Hunyuan Hy3 Tencent | 25.6% |
Dati benchmark aggregati da llm-stats.com. Vedi Metodologia.