Z.ai
GLM-5.3 Flashโอเพ่นซอร์ส
High-speed, low-cost GLM 5.3 Flash model by Z.ai with 180+ tokens/sec throughput.
Intelligence
64.2 #224
AAQI 58.0
Output speed
182 t/s
TTFT 220 ms · ITL 5.5 ms/tok
Blended price (3:1)
$0.15 /1M
In $0.1/M · Out $0.3/M
Context
262K
≈ 655 A4 pages · max out 8192
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
GLM-5.3 Flash
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see ระเบียบวิธี.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
คะแนนแยกตามแหล่งข้อมูล
คะแนนแยกตามหมวดหมู่
ผลการทดสอบ Benchmark
รายละเอียดโมเดล
วันที่เปิดตัว
1 ส.ค. 2569
พารามิเตอร์
Not disclosed
สัญญาอนุญาต
Open Weights
ขนาดคอนเทกซ์
262,000 tokens
ราคาขาเข้า
$0.1/M
ราคาขาออก
$0.3/M
ความเร็วขาออก
182 t/s
เวลาถึงโทเค็นแรก
220 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/glm-5-3-flash)BibTeX
@misc{llmpodium2026glm53flash,
title = {Benchmark Analysis and Intelligence Evaluation of GLM-5.3 Flash},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/glm-5-3-flash}
}More from Z.ai
โมเดลทางเลือก
ทางเลือกที่ใกล้เคียงที่สุดสำหรับ GLM-5.3 Flash โดยอิงตามความสามารถ ราคา และคะแนนเบนช์มาร์ก