IBM
Granite 4.2 3B InstructOffen
IBM Granite 4.2 3B compact SLM optimized for on-device and low-latency inference.
Intelligence
42.0 #370
AAQI 36.0
Output speed
148 t/s
TTFT 160 ms · ITL 6.8 ms/tok
Blended price (3:1)
$0.11 /1M
In $0.08/M · Out $0.2/M
Context
128K
≈ 320 A4 pages · max out 8192
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
Granite 4.2 3B Instruct
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see Methodik.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
Ergebnisse nach Quelle
Kategorie-Punkte
Benchmark-Ergebnisse
Modelldetails
Veröffentlichungsdatum
20. Juli 2026
Parameter
3 billion
Lizenz
Open Weights
Kontextfenster
128,000 tokens
Eingabepreis
$0.08/M
Ausgabepreis
$0.2/M
Ausgabegeschwindigkeit
148 t/s
Zeit bis zum ersten Token
160 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/granite-4-2-3b)BibTeX
@misc{llmpodium2026granite423b,
title = {Benchmark Analysis and Intelligence Evaluation of Granite 4.2 3B Instruct},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/granite-4-2-3b}
}More from IBM
Alternative Modelle
Nächste Alternativen zu Granite 4.2 3B Instruct basierend auf Leistung, Preis und Benchmarks.