Gemini 3.1 Flash-Lite API Pricing
Google · released 2026-03-03
Pricing
0.025¢ / 1K
0.150¢ / 1K
0.056¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.0575¢ |
| RAG query | 4,000 in / 500 out | 0.175¢ |
| Long document summary | 50,000 in / 1,000 out | $0.0140 |
| Code review of a 2k-line file | 30,000 in / 2,000 out | $0.0105 |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $17.25 |
Cost = (input tokens ÷ 1,000,000 × $0.25) + (output tokens ÷ 1,000,000 × $1.50).
Benchmarks
#164 of 533
#103 of 197
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Qwen3.5 4B (Reasoning) | 13.1 | $0.06 |
| Nemotron 3.5 Lightning | 12.9 | $0.095 |
| GPT-5 nano (high) | 13.0 | $0.138 |
| Step 3.5 Flash 2603 | 17.0 | $0.15 |
| Step 3.5 Flash | 16.6 | $0.15 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Gemini 2.5 Pro | 16.1 | $3.438 |
| DeepSeek V3.2 (Non-reasoning) | 16.0 | $0.315 |
| MiMo-V2-Flash (Non-reasoning) | 16.0 | — |
| Gemini 3.1 Flash-Lite | 15.6 | $0.563 |
| K2 Horizon 3.7B | 15.6 | — |
| Qwen3 Max | 15.6 | $2.40 |
| Gemini 2.5 Flash Preview (Sep '25) (Reasoning) | 15.5 | — |
More from Google
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Gemini 3.8 Flash (high) | 40.9 | $1.50 |
| Gemini 3.7 Flash (medium) | 39.6 | $1.50 |
| Gemini 3.6 Flash (high) | 34.0 | $1.50 |
| Gemini 3.5 Flash (medium) | 33.6 | $3.375 |
| Gemini 3.1 Pro Preview | 29.7 | $4.50 |
| Gemini 3 Pro Preview (high) | 28.0 | $4.50 |
| Gemini 3 Flash Preview (Reasoning) | 26.3 | $1.125 |
| Gemini 3.5 Flash-Lite | 22.2 | $0.85 |
Count tokens & estimate cost for Gemini 3.1 Flash-Lite See Gemini 3.1 Flash-Lite on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops Gemini 3.1 Flash-Lite?
Gemini 3.1 Flash-Lite is developed by Google, released 2026-03-03.
How much does Gemini 3.1 Flash-Lite cost per million tokens?
Gemini 3.1 Flash-Lite costs $0.25 per 1M input tokens and $1.50 per 1M output tokens ($0.563 blended), per Artificial Analysis's official pricing data.
How does Gemini 3.1 Flash-Lite rank on benchmarks?
Gemini 3.1 Flash-Lite scores 15.6 on the Artificial Analysis Intelligence Index, ranking #164 of 533 models we track.
Is there a cheaper model with similar intelligence to Gemini 3.1 Flash-Lite?
Yes — Qwen3.5 4B (Reasoning) scores similar intelligence (13.1) at $0.06/1M blended, versus $0.563/1M for Gemini 3.1 Flash-Lite.
Where does this data come from?
Pricing and benchmark data for Gemini 3.1 Flash-Lite comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.