Z.ai

GLM-5.3 Flashمفتوح

High-speed, low-cost GLM 5.3 Flash model by Z.ai with 180+ tokens/sec throughput.

#224صدر 1 أغسطس 2026·Open Weights·Not disclosed
textimage262K ctx
API providersالترتيب
Intelligence
64.2 #224
AAQI 58.0
Output speed
182 t/s
TTFT 220 ms · ITL 5.5 ms/tok
Blended price (3:1)
$0.15 /1M
In $0.1/M · Out $0.3/M
Context
262K
≈ 655 A4 pages · max out 8192

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

GLM-5.3 Flash
Claude Mythos Preview (#1)
Average (68)
Coding (58.4)Mathematics (64)Hard Science (GPQA) (62.5)Agentic & Tool Use (61)Knowledge (MMLU) (60.5)Vision / Multimodal (55)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see المنهجية.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

الدرجات حسب المصدر

تفاصيل النموذج

تاريخ الإصدار
1 أغسطس 2026
المعلمات
Not disclosed
الترخيص
Open Weights
نافذة السياق
262,000 tokens
سعر الإدخال
$0.1/M
سعر الإخراج
$0.3/M
سرعة الإخراج
182 t/s
الزمن حتى أول رمز
220 ms

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/glm-5-3-flash/rank.svg)](https://llmpodium.com/models/glm-5-3-flash)
BibTeX
@misc{llmpodium2026glm53flash,
  title = {Benchmark Analysis and Intelligence Evaluation of GLM-5.3 Flash},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/glm-5-3-flash}
}

More from Z.ai

النماذج البديلة

أقرب البدائل لـ GLM-5.3 Flash استناداً إلى القدرات والأسعار ونتائج الاختبارات.

0 / 4