Google

Gemini 3.8 Flash

Gemini 3.8 Flash (high reasoning) by Google. High intelligence index (59), output throughput of 302 TPS, and native 1M context window with 90% prompt caching discount.

#127صدر 1 سبتمبر 2026·Proprietary·Not disclosed
textimageaudiovideo1M ctxreasoning
API providersالترتيب
Intelligence
68.5 #127
Top 18% tier
Output speed
302.1 t/s
TTFT 180 ms · ITL 3.3 ms/tok
Blended price (3:1)
$1.5 /1M
In $0.75/M · Out $3.75/M
Context
1M
≈ 2,500 A4 pages · max out 65536

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

Gemini 3.8 Flash
Claude Mythos Preview (#1)
Average (68)
Coding (54)Mathematics (73.5)Hard Science (GPQA) (56)Agentic & Tool Use (72.5)Knowledge (MMLU) (58)Vision / Multimodal (62)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see المنهجية.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

الدرجات حسب المصدر

تفاصيل النموذج

تاريخ الإصدار
1 سبتمبر 2026
المعلمات
Not disclosed
الترخيص
Proprietary
نافذة السياق
1,000,000 tokens
سعر الإدخال
$0.75/M
سعر الإخراج
$3.75/M
سرعة الإخراج
302 t/s
الزمن حتى أول رمز
180 ms

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/gemini-3-8-flash/rank.svg)](https://llmpodium.com/models/gemini-3-8-flash)
BibTeX
@misc{llmpodium2026gemini38flash,
  title = {Benchmark Analysis and Intelligence Evaluation of Gemini 3.8 Flash},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/gemini-3-8-flash}
}

More from Google

النماذج البديلة

أقرب البدائل لـ Gemini 3.8 Flash استناداً إلى القدرات والأسعار ونتائج الاختبارات.

0 / 4