Z.ai
GLM-5.3 Flashオープン
High-speed, low-cost GLM 5.3 Flash model by Z.ai with 180+ tokens/sec throughput.
Intelligence
64.2 #224
AAQI 58.0
Output speed
182 t/s
TTFT 220 ms · ITL 5.5 ms/tok
Blended price (3:1)
$0.15 /1M
In $0.1/M · Out $0.3/M
Context
262K
≈ 655 A4 pages · max out 8192
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
GLM-5.3 Flash
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see 評価方法.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
ソース別スコア
カテゴリー別スコア
ベンチマーク結果
モデル詳細
リリース日
2026年8月1日
パラメータ数
Not disclosed
ライセンス
Open Weights
コンテキストウィンドウ
262,000 tokens
入力価格
$0.1/M
出力価格
$0.3/M
出力速度
182 t/s
最初のトークンまでの時間
220 ms
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/glm-5-3-flash)BibTeX
@misc{llmpodium2026glm53flash,
title = {Benchmark Analysis and Intelligence Evaluation of GLM-5.3 Flash},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/glm-5-3-flash}
}More from Z.ai
代替モデル
機能、価格、ベンチマークスコアに基づく GLM-5.3 Flash の最適な代替モデル。