Google vs Google

Gemini 3.1 Pro vs Gemini 3.8 Flash: pricing comparison

Gemini 3.8 Flash costs 3.0× less than Gemini 3.1 Pro for a typical 1,000-input, 500-output request. Here is how the two compare on price, context and discounts.

Prices last updated:

Side-by-side prices

Gemini 3.1 ProGemini 3.8 Flash
ProviderGoogleGoogle
Input per 1M tokens$2.00$0.75
Output per 1M tokens$12.00$3.75
Context window1.05M1.05M
Batch discount50% off50% off
Typical request (1,000 in / 500 out)$0.008$0.002625
Monthly at 1,000 requests/day$240.00$78.75

Gemini 3.1 Pro: Price shown is for prompts up to 200K tokens; longer prompts are billed at $4 / $18.

Gemini 3.8 Flash: Promotional rate through 31 December 2026; scheduled to rise to $1.50 / $7.50.

Compare with your own numbers

Gemini 3.1 Pro per request
$0.008
Per month
$240.00
Gemini 3.8 Flash per request
$0.002625
Per month
$78.75

Cost by task

Cost per 1,000 requests at standard prices.
TaskTokens (in / out)Gemini 3.1 ProGemini 3.8 Flash
Short chat reply500 / 300$4.60$1.50
RAG answer6,000 / 600$19.20$6.75
Summarize a long document12,000 / 1,000$36.00$12.75
Output-heavy generation1,000 / 4,000$50.00$15.75

Output tokens usually dominate the bill — see input vs output token cost.

Frequently asked questions

Is Gemini 3.1 Pro or Gemini 3.8 Flash cheaper?
Gemini 3.8 Flash is cheaper for a typical request of 1,000 input and 500 output tokens: $0.002625 versus $0.008 for Gemini 3.1 Pro, about 67% less.
What are the per-token prices of Gemini 3.1 Pro and Gemini 3.8 Flash?
Gemini 3.1 Pro costs $2.00 input and $12.00 output per million tokens. Gemini 3.8 Flash costs $0.75 input and $3.75 output per million tokens.
Which has the larger context window, Gemini 3.1 Pro or Gemini 3.8 Flash?
Both accept 1.05M tokens of context.

More comparisons

All comparisons →