Google
Gemini 3.8 Flash (Low Reasoning)
Low reasoning effort variant of Gemini 3.8 Flash optimized for sub-150ms TTFT and maximum throughput.
Intelligence
62.0 #258
Top 37% tier
Output speed
340 t/s
TTFT 140 ms · ITL 2.9 ms/tok
Blended price (3:1)
$1.5 /1M
In $0.75/M · Out $3.75/M
Context
1M
≈ 2,500 A4 pages · max out 32768
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Gemini 3.8 Flash (Low Reasoning)
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see ระเบียบวิธี.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
คะแนนแยกตามแหล่งข้อมูล
คะแนนแยกตามหมวดหมู่
ผลการทดสอบ Benchmark
รายละเอียดโมเดล
วันที่เปิดตัว
1 ก.ย. 2569
พารามิเตอร์
Not disclosed
สัญญาอนุญาต
Proprietary
ขนาดคอนเทกซ์
1,000,000 tokens
ราคาขาเข้า
$0.75/M
ราคาขาออก
$3.75/M
ความเร็วขาออก
340 t/s
เวลาถึงโทเค็นแรก
140 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/gemini-3-8-flash-low)BibTeX
@misc{llmpodium2026gemini38flashlow,
title = {Benchmark Analysis and Intelligence Evaluation of Gemini 3.8 Flash (Low Reasoning)},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/gemini-3-8-flash-low}
}More from Google
โมเดลทางเลือก
ทางเลือกที่ใกล้เคียงที่สุดสำหรับ Gemini 3.8 Flash (Low Reasoning) โดยอิงตามความสามารถ ราคา และคะแนนเบนช์มาร์ก