NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) API Pricing
NVIDIA · released 2025-10-28
Pricing
0.020¢ / 1K
0.060¢ / 1K
0.030¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.028¢ |
| RAG query | 4,000 in / 500 out | 0.11¢ |
| Long document summary | 50,000 in / 1,000 out | $0.0106 |
| Code review of a 2k-line file | 30,000 in / 2,000 out | 0.72¢ |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $8.40 |
Cost = (input tokens ÷ 1,000,000 × $0.20) + (output tokens ÷ 1,000,000 × $0.60).
Benchmarks
#450 of 533
#173 of 249
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Llama 3.1 Instruct 8B | 6.9 | $0.028 |
| Sarvam 30B (high) | 6.6 | $0.047 |
| Nova Micro | 5.9 | $0.061 |
| Granite 4.1 8B | 6.6 | $0.063 |
| NVIDIA Nemotron Nano 9B V2 (Reasoning) | 7.4 | $0.07 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Gemma 3n E4B Instruct Preview (May '25) | 5.8 | — |
| Mistral Large (Feb '24) | 5.8 | $6.00 |
| Mistral Small (Sep '24) | 5.8 | $0.30 |
| NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) | 5.8 | $0.30 |
| Phi-3 Mini Instruct 3.8B | 5.8 | — |
| Phi-4 Multimodal Instruct | 5.8 | — |
| Qwen2.5 Coder Instruct 7B | 5.8 | — |
More from NVIDIA
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Nemotron 3 Ultra 550B A55B (Reasoning) | 22.9 | $1.05 |
| Nemotron 3.5 Lightning | 12.9 | $0.095 |
| Nemotron 3 Super 120B A12B (Reasoning) | 12.8 | $0.45 |
| Nemotron Cascade 2 30B A3B | 11.7 | — |
| Nemotron 3 Nano Omni 30B A3B Reasoning | 10.3 | $0.42 |
| Llama Nemotron Super 49B v1.5 (Reasoning) | 9.0 | $0.40 |
| Llama 3.3 Nemotron Super 49B v1 (Reasoning) | 8.9 | — |
| NVIDIA Nemotron 3 Nano 30B A3B (Reasoning) | 8.9 | $0.088 |
Count tokens & estimate cost for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) See NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning)?
NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) is developed by NVIDIA, released 2025-10-28.
How much does NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) cost per million tokens?
NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) costs $0.20 per 1M input tokens and $0.60 per 1M output tokens ($0.30 blended), per Artificial Analysis's official pricing data.
How does NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) rank on benchmarks?
NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) scores 5.8 on the Artificial Analysis Intelligence Index, ranking #450 of 533 models we track.
Is there a cheaper model with similar intelligence to NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning)?
Yes — Llama 3.1 Instruct 8B scores similar intelligence (6.9) at $0.028/1M blended, versus $0.30/1M for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning).
Where does this data come from?
Pricing and benchmark data for NVIDIA Nemotron Nano 12B v2 VL (Non-reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.