Agnes

Agnes 3.0 Flash

Agnes 3.0 Flash by Sapiens AI is a high-speed frontier reasoning model with 1.0M context window and 234 tokens/sec throughput.

#344출시일: 2026년 9월 11일·Proprietary·Not disclosed
textimage1.0M ctxreasoning
API providers순위표
Intelligence
48.2 #344
AAQI 36.0
Output speed
234.7 t/s
TTFT 1.83 s · ITL 4.3 ms/tok
Blended price (3:1)
$0.03 /1M
In $0.05/M · Out $0.15/M
Context
1.0M
≈ 2,500 A4 pages · max out 16384

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

Agnes 3.0 Flash
Claude Mythos Preview (#1)
Average (68)
Coding (40)Mathematics (42)Hard Science (GPQA) (48)Agentic & Tool Use (40)Knowledge (MMLU) (42)Vision / Multimodal (45)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see 평가 방법.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

출처별 점수

모델 상세 정보

출시일
2026년 9월 11일
파라미터 수
Not disclosed
라이선스
Proprietary
컨텍스트 창
1,000,000 tokens
입력 가격
$0.05/M
출력 가격
$0.15/M
출력 속도
235 t/s
첫 토큰 생성 시간
1.83 s

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/agnes-3-0-flash/rank.svg)](https://llmpodium.com/models/agnes-3-0-flash)
BibTeX
@misc{llmpodium2026agnes30flash,
  title = {Benchmark Analysis and Intelligence Evaluation of Agnes 3.0 Flash},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/agnes-3-0-flash}
}

More from Agnes

대체 모델

성능, 가격 및 벤치마크 점수를 기반으로 한 Agnes 3.0 Flash의 주요 대체 모델.

0 / 4