Llama 3.1 Instruct 70B API Pricing
Meta · released 2024-07-23
Pricing
0.056¢ / 1K
0.056¢ / 1K
0.056¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.0448¢ |
| RAG query | 4,000 in / 500 out | 0.252¢ |
| Long document summary | 50,000 in / 1,000 out | $0.0286 |
| Code review of a 2k-line file | 30,000 in / 2,000 out | $0.0179 |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $13.44 |
Cost = (input tokens ÷ 1,000,000 × $0.56) + (output tokens ÷ 1,000,000 × $0.56).
Benchmarks
#408 of 533
#231 of 249
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Llama 3.1 Instruct 8B | 6.9 | $0.028 |
| Gemma 4 E4B (Reasoning) | 8.9 | $0.04 |
| Sarvam 30B (high) | 6.6 | $0.047 |
| Granite 4.2 3B | 9.1 | $0.052 |
| Nova Micro | 5.9 | $0.061 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| DeepSeek-V2.5 (Dec '24) | 6.6 | — |
| Gemini 2.0 Flash Thinking Experimental (Dec '24) | 6.6 | — |
| Granite 4.1 8B | 6.6 | $0.063 |
| Llama 3.1 Instruct 70B | 6.6 | $0.56 |
| Qwen3 30B A3B (Non-reasoning) | 6.6 | $0.35 |
| Qwen3 4B (Non-reasoning) | 6.6 | — |
| Sarvam 30B (high) | 6.6 | $0.047 |
More from Meta
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Muse Spark 1.3 (max) | 48.1 | $2.00 |
| Muse Spark 1.2 (xhigh) | 39.6 | $2.00 |
| Muse Spark 1.1 (xhigh) | 33.7 | $2.00 |
| Muse Spark | 31.3 | — |
| Muse Glimmer (high) | 17.5 | $0.637 |
| Llama 4 Maverick | 10.0 | $0.422 |
| Llama 4 Scout | 8.1 | $0.313 |
| Llama 3.3 Instruct 70B | 7.7 | $0.712 |
Count tokens & estimate cost for Llama 3.1 Instruct 70B See Llama 3.1 Instruct 70B on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops Llama 3.1 Instruct 70B?
Llama 3.1 Instruct 70B is developed by Meta, released 2024-07-23.
How much does Llama 3.1 Instruct 70B cost per million tokens?
Llama 3.1 Instruct 70B costs $0.56 per 1M input tokens and $0.56 per 1M output tokens ($0.56 blended), per Artificial Analysis's official pricing data.
How does Llama 3.1 Instruct 70B rank on benchmarks?
Llama 3.1 Instruct 70B scores 6.6 on the Artificial Analysis Intelligence Index, ranking #408 of 533 models we track.
Is there a cheaper model with similar intelligence to Llama 3.1 Instruct 70B?
Yes — Llama 3.1 Instruct 8B scores similar intelligence (6.9) at $0.028/1M blended, versus $0.56/1M for Llama 3.1 Instruct 70B.
Where does this data come from?
Pricing and benchmark data for Llama 3.1 Instruct 70B comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.