Agnes
Agnes 3.0 Flash
Agnes 3.0 Flash by Sapiens AI is a high-speed frontier reasoning model with 1.0M context window and 234 tokens/sec throughput.
Intelligence
48.2 #344
AAQI 36.0
Output speed
234.7 t/s
TTFT 1.83 s · ITL 4.3 ms/tok
Blended price (3:1)
$0.03 /1M
In $0.05/M · Out $0.15/M
Context
1.0M
≈ 2,500 A4 pages · max out 16384
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Agnes 3.0 Flash
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see ระเบียบวิธี.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
คะแนนแยกตามแหล่งข้อมูล
คะแนนแยกตามหมวดหมู่
ผลการทดสอบ Benchmark
รายละเอียดโมเดล
วันที่เปิดตัว
11 ก.ย. 2569
พารามิเตอร์
Not disclosed
สัญญาอนุญาต
Proprietary
ขนาดคอนเทกซ์
1,000,000 tokens
ราคาขาเข้า
$0.05/M
ราคาขาออก
$0.15/M
ความเร็วขาออก
235 t/s
เวลาถึงโทเค็นแรก
1.83 s
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/agnes-3-0-flash)BibTeX
@misc{llmpodium2026agnes30flash,
title = {Benchmark Analysis and Intelligence Evaluation of Agnes 3.0 Flash},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/agnes-3-0-flash}
}More from Agnes
โมเดลทางเลือก
ทางเลือกที่ใกล้เคียงที่สุดสำหรับ Agnes 3.0 Flash โดยอิงตามความสามารถ ราคา และคะแนนเบนช์มาร์ก