Z.ai
GLM-5.3 Flashओपन
High-speed, low-cost GLM 5.3 Flash model by Z.ai with 180+ tokens/sec throughput.
Intelligence
64.2 #224
AAQI 58.0
Output speed
182 t/s
TTFT 220 ms · ITL 5.5 ms/tok
Blended price (3:1)
$0.15 /1M
In $0.1/M · Out $0.3/M
Context
262K
≈ 655 A4 pages · max out 8192
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
GLM-5.3 Flash
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see कार्यप्रणाली.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
स्रोत के अनुसार स्कोर
श्रेणी के अनुसार स्कोर
बेंचमार्क परिणाम
मॉडल विवरण
रिलीज़ तिथि
1 अग॰ 2026
पैरामीटर
Not disclosed
लाइसेंस
Open Weights
कॉन्टेक्स्ट विंडो
262,000 tokens
इनपुट कीमत
$0.1/M
आउटपुट कीमत
$0.3/M
आउटपुट गति
182 t/s
पहले टोकन तक का समय
220 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/glm-5-3-flash)BibTeX
@misc{llmpodium2026glm53flash,
title = {Benchmark Analysis and Intelligence Evaluation of GLM-5.3 Flash},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/glm-5-3-flash}
}More from Z.ai
वैकल्पिक मॉडल
क्षमता, मूल्य और बेंचमार्क स्कोर के आधार पर GLM-5.3 Flash के निकटतम विकल्प।