Google
Gemini 3.8 Flash
Gemini 3.8 Flash (high reasoning) by Google. High intelligence index (59), output throughput of 302 TPS, and native 1M context window with 90% prompt caching discount.
Intelligence
68.5 #127
Top 18% tier
Output speed
302.1 t/s
TTFT 180 ms · ITL 3.3 ms/tok
Blended price (3:1)
$1.5 /1M
In $0.75/M · Out $3.75/M
Context
1M
≈ 2,500 A4 pages · max out 65536
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Gemini 3.8 Flash
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see Methodik.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
Ergebnisse nach Quelle
Kategorie-Punkte
Benchmark-Ergebnisse
Modelldetails
Veröffentlichungsdatum
1. Sept. 2026
Parameter
Not disclosed
Lizenz
Proprietary
Kontextfenster
1,000,000 tokens
Eingabepreis
$0.75/M
Ausgabepreis
$3.75/M
Ausgabegeschwindigkeit
302 t/s
Zeit bis zum ersten Token
180 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/gemini-3-8-flash)BibTeX
@misc{llmpodium2026gemini38flash,
title = {Benchmark Analysis and Intelligence Evaluation of Gemini 3.8 Flash},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/gemini-3-8-flash}
}More from Google
Alternative Modelle
Nächste Alternativen zu Gemini 3.8 Flash basierend auf Leistung, Preis und Benchmarks.