Gemini 3.8 Flash (high) API PricingNEW

Google · released 2026-09-02 · merges 3 reasoning-effort tiers

Benchmarks shown for the highest-scoring reasoning tier. Merged tiers: Gemini 3.8 Flash (high), Gemini 3.8 Flash (medium), Gemini 3.8 Flash (low).

Pricing

$0.75
input / 1M tokens
0.075¢ / 1K
$3.75
output / 1M tokens
0.375¢ / 1K
$1.50
blended / 1M tokens
0.150¢ / 1K

What it actually costs

WorkloadTokensCost
Short chat reply500 in / 300 out0.15¢
RAG query4,000 in / 500 out0.487¢
Long document summary50,000 in / 1,000 out$0.0413
Code review of a 2k-line file30,000 in / 2,000 out$0.0300
1,000 chat replies/day for 30 days30,000 × (500 in / 300 out)$45.00

Cost = (input tokens ÷ 1,000,000 × $0.75) + (output tokens ÷ 1,000,000 × $3.75).

Benchmarks

40.9
Intelligence
#17 of 533
76.3
Coding
#8 of 197
Math
316
Tokens/sec
10.59s
Time to first token

Similar intelligence, lower price

ModelIntelligenceBlended $/1M
Qwen3.8-Flash-Next39.8$0.23
GLM 5.3 Flash41.8$0.237
DeepSeek V4.1 Flash (Reasoning, Max Effort)39.5$0.525
Step 5 Preview43.7$1.425

Leaderboard neighbours

ModelIntelligenceBlended $/1M
GPT-5.6 Terra (max)42.1$4.50
Claude Opus 4.8 (Adaptive Reasoning, Max Effort)41.8$10.00
GLM 5.3 Flash41.8$0.237
Gemini 3.8 Flash (high)40.9$1.50
Claude Opus 4.7 (Adaptive Reasoning, Max Effort)40.7$10.00
Qwen3.8 Max40.2$3.00
Qwen3.8 2.4T A95B39.9$3.00

More from Google

ModelIntelligenceBlended $/1M
Gemini 3.7 Flash (medium)39.6$1.50
Gemini 3.6 Flash (high)34.0$1.50
Gemini 3.5 Flash (medium)33.6$3.375
Gemini 3.1 Pro Preview29.7$4.50
Gemini 3 Pro Preview (high)28.0$4.50
Gemini 3 Flash Preview (Reasoning)26.3$1.125
Gemini 3.5 Flash-Lite22.2$0.85
Gemma 4 31B (Reasoning)19.0

Count tokens & estimate cost for Gemini 3.8 Flash (high) See Gemini 3.8 Flash (high) on the leaderboard

Source: Artificial Analysis · last synced 2026-09-22

FAQ

Who develops Gemini 3.8 Flash (high)?

Gemini 3.8 Flash (high) is developed by Google, released 2026-09-02.

How much does Gemini 3.8 Flash (high) cost per million tokens?

Gemini 3.8 Flash (high) costs $0.75 per 1M input tokens and $3.75 per 1M output tokens ($1.50 blended), per Artificial Analysis's official pricing data.

How does Gemini 3.8 Flash (high) rank on benchmarks?

Gemini 3.8 Flash (high) scores 40.9 on the Artificial Analysis Intelligence Index, ranking #17 of 533 models we track.

Is there a cheaper model with similar intelligence to Gemini 3.8 Flash (high)?

Yes — Qwen3.8-Flash-Next scores similar intelligence (39.8) at $0.23/1M blended, versus $1.50/1M for Gemini 3.8 Flash (high).

Where does this data come from?

Pricing and benchmark data for Gemini 3.8 Flash (high) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.