Agnes

Agnes 2.5 Pro Beta

Agnes 2.5 Pro Beta frontier model with improved reasoning, speed and instruction following.

#363发布于 2025年2月15日·Proprietary·Not disclosed
text131K ctx
API providers排行榜
Intelligence
42.5 #363
AAQI 46.5
Output speed
65 t/s
TTFT 320 ms · ITL 15.4 ms/tok
Blended price (3:1)
$1.5 /1M
In $1/M · Out $3/M
Context
131K
≈ 328 A4 pages · max out 8192

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

Agnes 2.5 Pro Beta
Claude Mythos Preview (#1)
Average (68)
Coding (38)Mathematics (38.5)Hard Science (GPQA) (41.5)Agentic & Tool Use (37)Knowledge (MMLU) (48)Vision / Multimodal (25)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see 评分方法.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

各榜单得分

模型详细参数

发布日期
2025年2月15日
参数量
Not disclosed
开源协议
Proprietary
上下文窗口
131,000 tokens
输入价格
$1/M
输出价格
$3/M
输出速度
65 t/s
首Token延迟
320 ms

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/agnes-2-5-pro-beta/rank.svg)](https://llmpodium.com/models/agnes-2-5-pro-beta)
BibTeX
@misc{llmpodium2026agnes25probeta,
  title = {Benchmark Analysis and Intelligence Evaluation of Agnes 2.5 Pro Beta},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/agnes-2-5-pro-beta}
}

More from Agnes

替代模型推荐

基于能力、定价和基准测试得分与 Agnes 2.5 Pro Beta 最接近的替代模型。

0 / 4