Google

Gemini 3.8 Flash (Low Reasoning)

Low reasoning effort variant of Gemini 3.8 Flash optimized for sub-150ms TTFT and maximum throughput.

#258रिलीज़ 1 सित॰ 2026·Proprietary·Not disclosed
textimageaudiovideo1M ctxreasoning
API providersरैंकिंग
Intelligence
62.0 #258
Top 37% tier
Output speed
340 t/s
TTFT 140 ms · ITL 2.9 ms/tok
Blended price (3:1)
$1.5 /1M
In $0.75/M · Out $3.75/M
Context
1M
≈ 2,500 A4 pages · max out 32768

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

Gemini 3.8 Flash (Low Reasoning)
Claude Mythos Preview (#1)
Average (68)
Coding (54)Mathematics (67)Hard Science (GPQA) (56)Agentic & Tool Use (66)Knowledge (MMLU) (58)Vision / Multimodal (62)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see कार्यप्रणाली.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

स्रोत के अनुसार स्कोर

श्रेणी के अनुसार स्कोर

बेंचमार्क परिणाम

मॉडल विवरण

रिलीज़ तिथि
1 सित॰ 2026
पैरामीटर
Not disclosed
लाइसेंस
Proprietary
कॉन्टेक्स्ट विंडो
1,000,000 tokens
इनपुट कीमत
$0.75/M
आउटपुट कीमत
$3.75/M
आउटपुट गति
340 t/s
पहले टोकन तक का समय
140 ms

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/gemini-3-8-flash-low/rank.svg)](https://llmpodium.com/models/gemini-3-8-flash-low)
BibTeX
@misc{llmpodium2026gemini38flashlow,
  title = {Benchmark Analysis and Intelligence Evaluation of Gemini 3.8 Flash (Low Reasoning)},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/gemini-3-8-flash-low}
}

More from Google

वैकल्पिक मॉडल

क्षमता, मूल्य और बेंचमार्क स्कोर के आधार पर Gemini 3.8 Flash (Low Reasoning) के निकटतम विकल्प।

0 / 4