Gemini 2.5 Flash-Lite (Reasoning) API Pricing
Google · released 2025-06-17
Pricing
0.010¢ / 1K
0.040¢ / 1K
0.017¢ / 1K
What it actually costs
| Workload | Tokens | Cost |
|---|---|---|
| Short chat reply | 500 in / 300 out | 0.017¢ |
| RAG query | 4,000 in / 500 out | 0.06¢ |
| Long document summary | 50,000 in / 1,000 out | 0.54¢ |
| Code review of a 2k-line file | 30,000 in / 2,000 out | 0.38¢ |
| 1,000 chat replies/day for 30 days | 30,000 × (500 in / 300 out) | $5.10 |
Cost = (input tokens ÷ 1,000,000 × $0.10) + (output tokens ÷ 1,000,000 × $0.40).
Benchmarks
#306 of 533
#124 of 249
Similar intelligence, lower price
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Llama 3.1 Instruct 8B | 6.9 | $0.028 |
| Gemma 4 E4B (Reasoning) | 8.9 | $0.04 |
| Sarvam 30B (high) | 6.6 | $0.047 |
| Granite 4.2 3B | 9.1 | $0.052 |
| Nova Micro | 5.9 | $0.061 |
Leaderboard neighbours
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Magistral Small 1.2 | 8.6 | $0.75 |
| Qwen3 32B (Reasoning) | 8.6 | $0.28 |
| DeepSeek V3 (Dec '24) | 8.5 | $0.463 |
| Gemini 2.5 Flash-Lite (Reasoning) | 8.5 | $0.175 |
| DeepSeek R1 Distill Qwen 32B | 8.4 | — |
| GLM-4.6V (Non-reasoning) | 8.4 | $0.45 |
| GPT-4o (Nov '24) | 8.4 | $4.375 |
More from Google
| Model | Intelligence | Blended $/1M |
|---|---|---|
| Gemini 3.8 Flash (high) | 40.9 | $1.50 |
| Gemini 3.7 Flash (medium) | 39.6 | $1.50 |
| Gemini 3.6 Flash (high) | 34.0 | $1.50 |
| Gemini 3.5 Flash (medium) | 33.6 | $3.375 |
| Gemini 3.1 Pro Preview | 29.7 | $4.50 |
| Gemini 3 Pro Preview (high) | 28.0 | $4.50 |
| Gemini 3 Flash Preview (Reasoning) | 26.3 | $1.125 |
| Gemini 3.5 Flash-Lite | 22.2 | $0.85 |
Count tokens & estimate cost for Gemini 2.5 Flash-Lite (Reasoning) See Gemini 2.5 Flash-Lite (Reasoning) on the leaderboard
Source: Artificial Analysis · last synced 2026-09-22
FAQ
Who develops Gemini 2.5 Flash-Lite (Reasoning)?
Gemini 2.5 Flash-Lite (Reasoning) is developed by Google, released 2025-06-17.
How much does Gemini 2.5 Flash-Lite (Reasoning) cost per million tokens?
Gemini 2.5 Flash-Lite (Reasoning) costs $0.10 per 1M input tokens and $0.40 per 1M output tokens ($0.175 blended), per Artificial Analysis's official pricing data.
How does Gemini 2.5 Flash-Lite (Reasoning) rank on benchmarks?
Gemini 2.5 Flash-Lite (Reasoning) scores 8.5 on the Artificial Analysis Intelligence Index, ranking #306 of 533 models we track.
Is there a cheaper model with similar intelligence to Gemini 2.5 Flash-Lite (Reasoning)?
Yes — Llama 3.1 Instruct 8B scores similar intelligence (6.9) at $0.028/1M blended, versus $0.175/1M for Gemini 2.5 Flash-Lite (Reasoning).
Where does this data come from?
Pricing and benchmark data for Gemini 2.5 Flash-Lite (Reasoning) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.