IBM
Granite 4.2 3B InstructTerbuka
IBM Granite 4.2 3B compact SLM optimized for on-device and low-latency inference.
Intelligence
42.0 #370
AAQI 36.0
Output speed
148 t/s
TTFT 160 ms · ITL 6.8 ms/tok
Blended price (3:1)
$0.11 /1M
In $0.08/M · Out $0.2/M
Context
128K
≈ 320 A4 pages · max out 8192
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Granite 4.2 3B Instruct
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see Metodologi.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
Skor per sumber
Skor per kategori
Hasil benchmark
Detail model
Tanggal rilis
20 Jul 2026
Parameter
3 billion
Lisensi
Open Weights
Jendela konteks
128,000 tokens
Harga input
$0.08/M
Harga output
$0.2/M
Kecepatan output
148 t/s
Waktu ke token pertama
160 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/granite-4-2-3b)BibTeX
@misc{llmpodium2026granite423b,
title = {Benchmark Analysis and Intelligence Evaluation of Granite 4.2 3B Instruct},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/granite-4-2-3b}
}More from IBM
Model Alternatif
Alternatif terdekat untuk Granite 4.2 3B Instruct berdasarkan kemampuan, harga, dan skor benchmark.