TL;DR — Gemini 3 Pro runs ~17× Gemini 3 Flash ($1.25/$5 vs $0.075/$0.30), and Flash holds the cheapest input price we track. Every number below is from our weekly-verified pricing tracker (verified 2026-08-26).
Pricing side by side (per 1M tokens, USD)
| Model | Provider | Input | Output | |-------|----------|-------|--------| | Gemini 3 Pro | Google | $1.25 | $5 | | Gemini 3 Flash | Google | $0.075 | $0.3 |
Verified 2026-08-26 against official rate cards — the tracker keeps the change log.
What a real month costs
At a standard workload of 10M input + 2M output tokens/month:
- Gemini 3 Pro: $22.50
- Gemini 3 Flash: $1.35
On this mix, Gemini 3 Flash is the cheaper column (16.7× gap). Your ratio will differ — run your own mix through the LLM cost calculator.
Choose Gemini 3 Pro if
you want Google's value-frontier tier for real reasoning at prices that undercut most Western frontier rates.
Choose Gemini 3 Flash if
volume is the game: at $0.075 input, Flash is the floor of the tracked market — background jobs, summarization, and classification cost near-nothing here.
Routing decision card
- Google's internal spread mirrors OpenAI's: Pro for quality, Flash for bulk, ~17× apart
- Flash's input floor makes it the strongest candidate for input-heavy workloads (long-context reads, RAG scans) of any tracked model
- Pair Flash-for-bulk with Pro-for-answers and the blended rate stays close to Flash's
- The router answer: most teams shouldn't hard-pick one — ClawRouters routes each request to whichever side fits it, so the comparison becomes a per-call decision instead of a bet.
Part of the LLM model comparison index. Prices move fast in 2026 — the tracker logs every change with dates.