Google
Gemini 3.8 Flash (Medium Reasoning)
Balanced reasoning effort variant of Gemini 3.8 Flash for agentic workflows.
Intelligence
65.2 #190
Top 27% tier
Output speed
315 t/s
TTFT 160 ms · ITL 3.2 ms/tok
Blended price (3:1)
$1.5 /1M
In $0.75/M · Out $3.75/M
Context
1M
≈ 2,500 A4 pages · max out 65536
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Gemini 3.8 Flash (Medium Reasoning)
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see Metodologia.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
Pontuações por fonte
Pontuações por categoria
Resultados de benchmarks
Detalhes do modelo
Data de lançamento
1 de set. de 2026
Parâmetros
Not disclosed
Licença
Proprietary
Janela de contexto
1,000,000 tokens
Preço de entrada
$0.75/M
Preço de saída
$3.75/M
Velocidade de saída
315 t/s
Tempo até o primeiro token
160 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/gemini-3-8-flash-medium)BibTeX
@misc{llmpodium2026gemini38flashmedium,
title = {Benchmark Analysis and Intelligence Evaluation of Gemini 3.8 Flash (Medium Reasoning)},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/gemini-3-8-flash-medium}
}More from Google
Modelos alternativos
Principais alternativas para Gemini 3.8 Flash (Medium Reasoning) com base em recursos, preço e benchmarks.