NVIDIA Nemotron Nano 9B V2 (Reasoning) API Pricing
NVIDIA · released 2025-08-18
Pricing
< 0.01¢ / 1K
0.016¢ / 1K
< 0.01¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.0068¢ |
| RAG query | 4,000 in / 500 out | 0.024¢ |
| Long document summary | 50,000 in / 1,000 out | 0.216¢ |
| Code review of a 2k-line file | 30,000 in / 2,000 out | 0.152¢ |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $2.04 |
Cost = (input tokens ÷ 1,000,000 × $0.04) + (output tokens ÷ 1,000,000 × $0.16).
Benchmarks
#360 of 533
#90 of 249
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Llama 3.1 Instruct 8B | 6.9 | $0.028 |
| Gemma 4 E4B (Reasoning) | 8.9 | $0.04 |
| Sarvam 30B (high) | 6.6 | $0.047 |
| Granite 4.2 3B | 9.1 | $0.052 |
| Nova Micro | 5.9 | $0.061 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Hermes 4 - Llama-3.1 405B (Non-reasoning) | 7.4 | $1.50 |
| Llama Nemotron Super 49B v1.5 (Non-reasoning) | 7.4 | $0.40 |
| NVIDIA Nemotron 3 Nano 4B | 7.4 | — |
| NVIDIA Nemotron Nano 9B V2 (Reasoning) | 7.4 | $0.07 |
| GPT-4o (May '24) | 7.3 | $7.50 |
| Gemini 2.0 Flash-Lite (Preview) | 7.3 | — |
| Kimi Linear 48B A3B Instruct | 7.3 | — |
More from NVIDIA
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Nemotron 3 Ultra 550B A55B (Reasoning) | 22.9 | $1.05 |
| Nemotron 3.5 Lightning | 12.9 | $0.095 |
| Nemotron 3 Super 120B A12B (Reasoning) | 12.8 | $0.45 |
| Nemotron Cascade 2 30B A3B | 11.7 | — |
| Nemotron 3 Nano Omni 30B A3B Reasoning | 10.3 | $0.42 |
| Llama Nemotron Super 49B v1.5 (Reasoning) | 9.0 | $0.40 |
| Llama 3.3 Nemotron Super 49B v1 (Reasoning) | 8.9 | — |
| NVIDIA Nemotron 3 Nano 30B A3B (Reasoning) | 8.9 | $0.088 |
Count tokens & estimate cost for NVIDIA Nemotron Nano 9B V2 (Reasoning) See NVIDIA Nemotron Nano 9B V2 (Reasoning) on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops NVIDIA Nemotron Nano 9B V2 (Reasoning)?
NVIDIA Nemotron Nano 9B V2 (Reasoning) is developed by NVIDIA, released 2025-08-18.
How much does NVIDIA Nemotron Nano 9B V2 (Reasoning) cost per million tokens?
NVIDIA Nemotron Nano 9B V2 (Reasoning) costs $0.04 per 1M input tokens and $0.16 per 1M output tokens ($0.07 blended), per Artificial Analysis's official pricing data.
How does NVIDIA Nemotron Nano 9B V2 (Reasoning) rank on benchmarks?
NVIDIA Nemotron Nano 9B V2 (Reasoning) scores 7.4 on the Artificial Analysis Intelligence Index, ranking #360 of 533 models we track.
Is there a cheaper model with similar intelligence to NVIDIA Nemotron Nano 9B V2 (Reasoning)?
Yes — Llama 3.1 Instruct 8B scores similar intelligence (6.9) at $0.028/1M blended, versus $0.07/1M for NVIDIA Nemotron Nano 9B V2 (Reasoning).
Where does this data come from?
Pricing and benchmark data for NVIDIA Nemotron Nano 9B V2 (Reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.