Google
Gemini 3.8 Flash (Low Reasoning)
Low reasoning effort variant of Gemini 3.8 Flash optimized for sub-150ms TTFT and maximum throughput.
Intelligence
62.0 #258
Top 37% tier
Output speed
340 t/s
TTFT 140 ms · ITL 2.9 ms/tok
Blended price (3:1)
$1.5 /1M
In $0.75/M · Out $3.75/M
Context
1M
≈ 2,500 A4 pages · max out 32768
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Gemini 3.8 Flash (Low Reasoning)
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see পদ্ধতি.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
উৎস অনুযায়ী স্কোর
বিভাগ অনুযায়ী স্কোর
বেঞ্চমার্ক ফলাফল
মডেল বিবরণ
প্রকাশের তারিখ
১ সেপ, ২০২৬
প্যারামিটার
Not disclosed
লাইসেন্স
Proprietary
কনটেক্সট উইন্ডো
1,000,000 tokens
ইনপুট দাম
$0.75/M
আউটপুট দাম
$3.75/M
আউটপুট গতি
340 t/s
প্রথম টোকেনের সময়
140 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/gemini-3-8-flash-low)BibTeX
@misc{llmpodium2026gemini38flashlow,
title = {Benchmark Analysis and Intelligence Evaluation of Gemini 3.8 Flash (Low Reasoning)},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/gemini-3-8-flash-low}
}More from Google
বিকল্প মডেলসমূহ
সক্ষমতা, মূল্য এবং বেঞ্চমার্ক স্কোরের ভিত্তিতে Gemini 3.8 Flash (Low Reasoning)-এর নিকটতম বিকল্প।