Agnes
Agnes 3.0 Flash
Agnes 3.0 Flash by Sapiens AI is a high-speed frontier reasoning model with 1.0M context window and 234 tokens/sec throughput.
Intelligence
48.2 #344
AAQI 36.0
Output speed
234.7 t/s
TTFT 1.83 s · ITL 4.3 ms/tok
Blended price (3:1)
$0.03 /1M
In $0.05/M · Out $0.15/M
Context
1.0M
≈ 2,500 A4 pages · max out 16384
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Agnes 3.0 Flash
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see 評価方法.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
ソース別スコア
カテゴリー別スコア
ベンチマーク結果
モデル詳細
リリース日
2026年9月11日
パラメータ数
Not disclosed
ライセンス
Proprietary
コンテキストウィンドウ
1,000,000 tokens
入力価格
$0.05/M
出力価格
$0.15/M
出力速度
235 t/s
最初のトークンまでの時間
1.83 s
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/agnes-3-0-flash)BibTeX
@misc{llmpodium2026agnes30flash,
title = {Benchmark Analysis and Intelligence Evaluation of Agnes 3.0 Flash},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/agnes-3-0-flash}
}More from Agnes
代替モデル
機能、価格、ベンチマークスコアに基づく Agnes 3.0 Flash の最適な代替モデル。