Qwen3 VL 32B (Reasoning) API Pricing

Alibaba · released 2025-10-21

Pricing

$0.16
input / 1M tokens
0.016¢ / 1K
$0.64
output / 1M tokens
0.064¢ / 1K
$0.28
blended / 1M tokens
0.028¢ / 1K

What it actually costs

WorkloadTokensCost
Short chat reply500 in / 300 out0.0272¢
RAG query4,000 in / 500 out0.096¢
Long document summary50,000 in / 1,000 out0.864¢
Code review of a 2k-line file30,000 in / 2,000 out0.608¢
1,000 chat replies/day for 30 days30,000 × (500 in / 300 out)$8.16

Cost = (input tokens ÷ 1,000,000 × $0.16) + (output tokens ÷ 1,000,000 × $0.64).

Benchmarks

11.9
Intelligence
#221 of 533
Coding
84.7
Math
#44 of 249
Tokens/sec
Time to first token

Similar intelligence, lower price

ModelIntelligenceBlended $/1M
Gemma 4 E4B (Reasoning)8.9$0.04
Granite 4.2 3B9.1$0.052
Qwen3.5 4B (Reasoning)13.1$0.06
NVIDIA Nemotron 3 Nano 30B A3B (Reasoning)8.9$0.088
Nemotron 3.5 Lightning12.9$0.095

Leaderboard neighbours

ModelIntelligenceBlended $/1M
Seed-OSS-36B-Instruct12.1$0.30
Qwen3 235B A22B 2507 Instruct12.0$0.403
Qwen3 Coder 480B A35B Instruct11.9$3.00
Qwen3 VL 32B (Reasoning)11.9$0.28
Magistral Medium 1.211.8
Sonar Reasoning Pro11.8
Gemini 2.5 Flash Preview (Reasoning)11.7

More from Alibaba

ModelIntelligenceBlended $/1M
Qwen3.8 Max (0902)45.4$3.00
Qwen3.8 Max40.2$3.00
Qwen3.8 2.4T A95B39.9$3.00
Qwen3.8-Flash-Next39.8$0.23
Qwen3.8 27B (xhigh)33.7$1.125
Qwen3.7 Max29.5$3.75
Qwen3.6 Max Preview28.4$2.925
Qwen3.6 Plus27.0$1.125

Count tokens & estimate cost for Qwen3 VL 32B (Reasoning) See Qwen3 VL 32B (Reasoning) on the leaderboard

Source: Artificial Analysis · last synced 2026-09-22

FAQ

Who develops Qwen3 VL 32B (Reasoning)?

Qwen3 VL 32B (Reasoning) is developed by Alibaba, released 2025-10-21.

How much does Qwen3 VL 32B (Reasoning) cost per million tokens?

Qwen3 VL 32B (Reasoning) costs $0.16 per 1M input tokens and $0.64 per 1M output tokens ($0.28 blended), per Artificial Analysis's official pricing data.

How does Qwen3 VL 32B (Reasoning) rank on benchmarks?

Qwen3 VL 32B (Reasoning) scores 11.9 on the Artificial Analysis Intelligence Index, ranking #221 of 533 models we track.

Is there a cheaper model with similar intelligence to Qwen3 VL 32B (Reasoning)?

Yes — Gemma 4 E4B (Reasoning) scores similar intelligence (8.9) at $0.04/1M blended, versus $0.28/1M for Qwen3 VL 32B (Reasoning).

Where does this data come from?

Pricing and benchmark data for Qwen3 VL 32B (Reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.