Qwen3 VL 32B (Reasoning) API Pricing
Alibaba · released 2025-10-21
Pricing
0.016¢ / 1K
0.064¢ / 1K
0.028¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.0272¢ |
| RAG query | 4,000 in / 500 out | 0.096¢ |
| Long document summary | 50,000 in / 1,000 out | 0.864¢ |
| Code review of a 2k-line file | 30,000 in / 2,000 out | 0.608¢ |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $8.16 |
Cost = (input tokens ÷ 1,000,000 × $0.16) + (output tokens ÷ 1,000,000 × $0.64).
Benchmarks
#221 of 533
#44 of 249
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Gemma 4 E4B (Reasoning) | 8.9 | $0.04 |
| Granite 4.2 3B | 9.1 | $0.052 |
| Qwen3.5 4B (Reasoning) | 13.1 | $0.06 |
| NVIDIA Nemotron 3 Nano 30B A3B (Reasoning) | 8.9 | $0.088 |
| Nemotron 3.5 Lightning | 12.9 | $0.095 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Seed-OSS-36B-Instruct | 12.1 | $0.30 |
| Qwen3 235B A22B 2507 Instruct | 12.0 | $0.403 |
| Qwen3 Coder 480B A35B Instruct | 11.9 | $3.00 |
| Qwen3 VL 32B (Reasoning) | 11.9 | $0.28 |
| Magistral Medium 1.2 | 11.8 | — |
| Sonar Reasoning Pro | 11.8 | — |
| Gemini 2.5 Flash Preview (Reasoning) | 11.7 | — |
More from Alibaba
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Qwen3.8 Max (0902) | 45.4 | $3.00 |
| Qwen3.8 Max | 40.2 | $3.00 |
| Qwen3.8 2.4T A95B | 39.9 | $3.00 |
| Qwen3.8-Flash-Next | 39.8 | $0.23 |
| Qwen3.8 27B (xhigh) | 33.7 | $1.125 |
| Qwen3.7 Max | 29.5 | $3.75 |
| Qwen3.6 Max Preview | 28.4 | $2.925 |
| Qwen3.6 Plus | 27.0 | $1.125 |
Count tokens & estimate cost for Qwen3 VL 32B (Reasoning) See Qwen3 VL 32B (Reasoning) on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops Qwen3 VL 32B (Reasoning)?
Qwen3 VL 32B (Reasoning) is developed by Alibaba, released 2025-10-21.
How much does Qwen3 VL 32B (Reasoning) cost per million tokens?
Qwen3 VL 32B (Reasoning) costs $0.16 per 1M input tokens and $0.64 per 1M output tokens ($0.28 blended), per Artificial Analysis's official pricing data.
How does Qwen3 VL 32B (Reasoning) rank on benchmarks?
Qwen3 VL 32B (Reasoning) scores 11.9 on the Artificial Analysis Intelligence Index, ranking #221 of 533 models we track.
Is there a cheaper model with similar intelligence to Qwen3 VL 32B (Reasoning)?
Yes — Gemma 4 E4B (Reasoning) scores similar intelligence (8.9) at $0.04/1M blended, versus $0.28/1M for Qwen3 VL 32B (Reasoning).
Where does this data come from?
Pricing and benchmark data for Qwen3 VL 32B (Reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.