GLM 5.3 Flash API PricingNEW
Z AI · released 2026-08-26
Pricing
0.015¢ / 1K
0.050¢ / 1K
0.024¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.0225¢ |
| RAG query | 4,000 in / 500 out | 0.085¢ |
| Long document summary | 50,000 in / 1,000 out | 0.8¢ |
| Code review of a 2k-line file | 30,000 in / 2,000 out | 0.55¢ |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $6.75 |
Cost = (input tokens ÷ 1,000,000 × $0.15) + (output tokens ÷ 1,000,000 × $0.50).
Benchmarks
#16 of 533
#21 of 197
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Qwen3.8-Flash-Next | 39.8 | $0.23 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Kimi K3 (max) | 43.6 | $6.00 |
| GPT-5.6 Terra (max) | 42.1 | $4.50 |
| Claude Opus 4.8 (Adaptive Reasoning, Max Effort) | 41.8 | $10.00 |
| GLM 5.3 Flash | 41.8 | $0.237 |
| Gemini 3.8 Flash (high) | 40.9 | $1.50 |
| Claude Opus 4.7 (Adaptive Reasoning, Max Effort) | 40.7 | $10.00 |
| Qwen3.8 Max | 40.2 | $3.00 |
More from Z AI
| Model | Intelligence | Blended $/1M |
|---|---|---|
| GLM-5.3 (max) | 44.8 | $2.15 |
| GLM-5.2 (max) | 33.7 | $2.15 |
| GLM-5 (Reasoning) | 27.9 | $1.55 |
| GLM-5-Turbo | 26.6 | — |
| GLM-5.1 (Reasoning) | 26.1 | $2.00 |
| GLM 5V Turbo (Reasoning) | 23.5 | — |
| GLM-4.7 (Reasoning) | 22.2 | $1.00 |
| GLM-4.6 (Reasoning) | 18.5 | $0.963 |
Count tokens & estimate cost for GLM 5.3 Flash See GLM 5.3 Flash on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops GLM 5.3 Flash?
GLM 5.3 Flash is developed by Z AI, released 2026-08-26.
How much does GLM 5.3 Flash cost per million tokens?
GLM 5.3 Flash costs $0.15 per 1M input tokens and $0.50 per 1M output tokens ($0.237 blended), per Artificial Analysis's official pricing data.
How does GLM 5.3 Flash rank on benchmarks?
GLM 5.3 Flash scores 41.8 on the Artificial Analysis Intelligence Index, ranking #16 of 533 models we track.
Is there a cheaper model with similar intelligence to GLM 5.3 Flash?
Yes — Qwen3.8-Flash-Next scores similar intelligence (39.8) at $0.23/1M blended, versus $0.237/1M for GLM 5.3 Flash.
Where does this data come from?
Pricing and benchmark data for GLM 5.3 Flash comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.