GLM-4.7-Flash (Reasoning) API Pricing

Z AI · released 2026-01-19 · merges 2 reasoning-effort tiers

Benchmarks shown for the highest-scoring reasoning tier. Merged tiers: GLM-4.7-Flash (Reasoning), GLM-4.7-Flash (Non-reasoning).

Pricing

$0.07
input / 1M tokens
< 0.01¢ / 1K
$0.40
output / 1M tokens
0.040¢ / 1K
$0.153
blended / 1M tokens
0.015¢ / 1K

What it actually costs

WorkloadTokensCost
Short chat reply500 in / 300 out0.0155¢
RAG query4,000 in / 500 out0.048¢
Long document summary50,000 in / 1,000 out0.39¢
Code review of a 2k-line file30,000 in / 2,000 out0.29¢
1,000 chat replies/day for 30 days30,000 × (500 in / 300 out)$4.65

Cost = (input tokens ÷ 1,000,000 × $0.07) + (output tokens ÷ 1,000,000 × $0.40).

Benchmarks

14.9
Intelligence
#175 of 533
Coding
Math
Tokens/sec
Time to first token

Similar intelligence, lower price

ModelIntelligenceBlended $/1M
Qwen3.5 4B (Reasoning)13.1$0.06
Nemotron 3.5 Lightning12.9$0.095
GPT-5 nano (high)13.0$0.138
Step 3.5 Flash 260317.0$0.15
Step 3.5 Flash16.6$0.15

Leaderboard neighbours

ModelIntelligenceBlended $/1M
o115.2$26.25
Gemini 2.5 Pro Preview (Mar' 25)15.0
GLM-4.6 (Non-reasoning)14.9$0.981
GLM-4.7-Flash (Reasoning)14.9$0.153
DeepSeek V3.1 Terminus (Reasoning)14.8$1.914
Granite 4.2 30B14.8$0.282
Grok 3 mini Reasoning (high)14.6$0.35

More from Z AI

ModelIntelligenceBlended $/1M
GLM-5.3 (max)44.8$2.15
GLM 5.3 Flash41.8$0.237
GLM-5.2 (max)33.7$2.15
GLM-5 (Reasoning)27.9$1.55
GLM-5-Turbo26.6
GLM-5.1 (Reasoning)26.1$2.00
GLM 5V Turbo (Reasoning)23.5
GLM-4.7 (Reasoning)22.2$1.00

Count tokens & estimate cost for GLM-4.7-Flash (Reasoning) See GLM-4.7-Flash (Reasoning) on the leaderboard

Source: Artificial Analysis · last synced 2026-09-22

FAQ

Who develops GLM-4.7-Flash (Reasoning)?

GLM-4.7-Flash (Reasoning) is developed by Z AI, released 2026-01-19.

How much does GLM-4.7-Flash (Reasoning) cost per million tokens?

GLM-4.7-Flash (Reasoning) costs $0.07 per 1M input tokens and $0.40 per 1M output tokens ($0.153 blended), per Artificial Analysis's official pricing data.

How does GLM-4.7-Flash (Reasoning) rank on benchmarks?

GLM-4.7-Flash (Reasoning) scores 14.9 on the Artificial Analysis Intelligence Index, ranking #175 of 533 models we track.

Is there a cheaper model with similar intelligence to GLM-4.7-Flash (Reasoning)?

Yes — Qwen3.5 4B (Reasoning) scores similar intelligence (13.1) at $0.06/1M blended, versus $0.153/1M for GLM-4.7-Flash (Reasoning).

Where does this data come from?

Pricing and benchmark data for GLM-4.7-Flash (Reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.