Qwen3 VL 30B A3B (Reasoning) API Pricing
Alibaba · released 2025-10-03
Pricing
0.020¢ / 1K
0.240¢ / 1K
0.075¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.082¢ |
| RAG query | 4,000 in / 500 out | 0.2¢ |
| Long document summary | 50,000 in / 1,000 out | $0.0124 |
| Code review of a 2k-line file | 30,000 in / 2,000 out | $0.0108 |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $24.60 |
Cost = (input tokens ÷ 1,000,000 × $0.20) + (output tokens ÷ 1,000,000 × $2.40).
Benchmarks
#271 of 533
#52 of 249
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Llama 3.1 Instruct 8B | 6.9 | $0.028 |
| Gemma 4 E4B (Reasoning) | 8.9 | $0.04 |
| Sarvam 30B (high) | 6.6 | $0.047 |
| Granite 4.2 3B | 9.1 | $0.052 |
| Granite 4.1 8B | 6.6 | $0.063 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| DiffusionGemma 26B A4B | 9.5 | — |
| QwQ 32B | 9.5 | $0.745 |
| Qwen3 235B A22B (Reasoning) | 9.5 | $2.625 |
| Qwen3 VL 30B A3B (Reasoning) | 9.5 | $0.75 |
| Gemini 2.0 Flash Thinking Experimental (Jan '25) | 9.4 | — |
| Gemini 2.5 Flash-Lite Preview (Sep '25) (Non-reasoning) | 9.3 | $0.175 |
| Mistral Large 3 | 9.3 | $0.75 |
More from Alibaba
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Qwen3.8 Max (0902) | 45.4 | $3.00 |
| Qwen3.8 Max | 40.2 | $3.00 |
| Qwen3.8 2.4T A95B | 39.9 | $3.00 |
| Qwen3.8-Flash-Next | 39.8 | $0.23 |
| Qwen3.8 27B (xhigh) | 33.7 | $1.125 |
| Qwen3.7 Max | 29.5 | $3.75 |
| Qwen3.6 Max Preview | 28.4 | $2.925 |
| Qwen3.6 Plus | 27.0 | $1.125 |
Count tokens & estimate cost for Qwen3 VL 30B A3B (Reasoning) See Qwen3 VL 30B A3B (Reasoning) on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops Qwen3 VL 30B A3B (Reasoning)?
Qwen3 VL 30B A3B (Reasoning) is developed by Alibaba, released 2025-10-03.
How much does Qwen3 VL 30B A3B (Reasoning) cost per million tokens?
Qwen3 VL 30B A3B (Reasoning) costs $0.20 per 1M input tokens and $2.40 per 1M output tokens ($0.75 blended), per Artificial Analysis's official pricing data.
How does Qwen3 VL 30B A3B (Reasoning) rank on benchmarks?
Qwen3 VL 30B A3B (Reasoning) scores 9.5 on the Artificial Analysis Intelligence Index, ranking #271 of 533 models we track.
Is there a cheaper model with similar intelligence to Qwen3 VL 30B A3B (Reasoning)?
Yes — Llama 3.1 Instruct 8B scores similar intelligence (6.9) at $0.028/1M blended, versus $0.75/1M for Qwen3 VL 30B A3B (Reasoning).
Where does this data come from?
Pricing and benchmark data for Qwen3 VL 30B A3B (Reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.