GLM-5.2 (max) API Pricing
Z AI · released 2026-06-16 · merges 2 reasoning-effort tiers
Benchmarks shown for the highest-scoring reasoning tier. Merged tiers: GLM-5.2 (max), GLM-5.2 (Non-reasoning).
Pricing
0.140¢ / 1K
0.440¢ / 1K
0.215¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.202¢ |
| RAG query | 4,000 in / 500 out | 0.78¢ |
| Long document summary | 50,000 in / 1,000 out | $0.0744 |
| Code review of a 2k-line file | 30,000 in / 2,000 out | $0.0508 |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $60.60 |
Cost = (input tokens ÷ 1,000,000 × $1.40) + (output tokens ÷ 1,000,000 × $4.40).
Benchmarks
#36 of 533
#30 of 197
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Agnes 3.0 Flash | 35.5 | $0.075 |
| Agnes 2.5 Pro Beta | 35.2 | $0.15 |
| DeepSeek V4 Flash Vision (Reasoning, Max Effort) | 34.8 | $0.66 |
| DeepSeek V4 Flash 0731 (Reasoning, Max Effort) | 34.3 | $0.66 |
| Qwen3.8 27B (xhigh) | 33.7 | $1.125 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| DeepSeek V4 Flash Vision (Reasoning, Max Effort) | 34.8 | $0.66 |
| DeepSeek V4 Flash 0731 (Reasoning, Max Effort) | 34.3 | $0.66 |
| Gemini 3.6 Flash (high) | 34.0 | $1.50 |
| GLM-5.2 (max) | 33.7 | $2.15 |
| Muse Spark 1.1 (xhigh) | 33.7 | $2.00 |
| Qwen3.8 27B (xhigh) | 33.7 | $1.125 |
| Gemini 3.5 Flash (medium) | 33.6 | $3.375 |
More from Z AI
| Model | Intelligence | Blended $/1M |
|---|---|---|
| GLM-5.3 (max) | 44.8 | $2.15 |
| GLM 5.3 Flash | 41.8 | $0.237 |
| GLM-5 (Reasoning) | 27.9 | $1.55 |
| GLM-5-Turbo | 26.6 | — |
| GLM-5.1 (Reasoning) | 26.1 | $2.00 |
| GLM 5V Turbo (Reasoning) | 23.5 | — |
| GLM-4.7 (Reasoning) | 22.2 | $1.00 |
| GLM-4.6 (Reasoning) | 18.5 | $0.963 |
Count tokens & estimate cost for GLM-5.2 (max) See GLM-5.2 (max) on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops GLM-5.2 (max)?
GLM-5.2 (max) is developed by Z AI, released 2026-06-16.
How much does GLM-5.2 (max) cost per million tokens?
GLM-5.2 (max) costs $1.40 per 1M input tokens and $4.40 per 1M output tokens ($2.15 blended), per Artificial Analysis's official pricing data.
How does GLM-5.2 (max) rank on benchmarks?
GLM-5.2 (max) scores 33.7 on the Artificial Analysis Intelligence Index, ranking #36 of 533 models we track.
Is there a cheaper model with similar intelligence to GLM-5.2 (max)?
Yes — Agnes 3.0 Flash scores similar intelligence (35.5) at $0.075/1M blended, versus $2.15/1M for GLM-5.2 (max).
Where does this data come from?
Pricing and benchmark data for GLM-5.2 (max) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.