Gemini 3.1 Pro vs Gemini 3.6 Flash
The two Gemini models people actually search for, and how far Flash gets you.
Side by side
| Attribute | Gemini 3.1 Pro | Gemini 3.6 Flash |
|---|---|---|
| Input, per 1M tokens | $2.00 | $1.50 |
| Output, per 1M tokens | $12.00 | $7.50 |
| Context window | 1,000,000 | 1,000,000 |
| Provider | ||
| Tokenizer | Google tokenizer | Google tokenizer |
| Long-context tier | above 200,000 | none |
| Retires | no date announced | no date announced |
Which to choose
- Reading more than you write
- Gemini 3.6 Flash $1.50 against Gemini 3.1 Pro $2 per million input tokens — a modest gap of 1.3x.
- Writing more than you read
- Gemini 3.1 Pro prices output at 6x its input rate, Gemini 3.6 Flash at 5x — so Gemini 3.6 Flash's lead widens as answers lengthen.
- Past the long-context threshold
- Gemini 3.1 Pro above 200,000 bills the whole request higher — $1.50 and $4. Gemini 3.6 Flash stays ahead either side of that line.
- 10,000 in, 1,000 out
- $0.0320 on Gemini 3.1 Pro, $0.0225 on Gemini 3.6 Flash. Over 10,000 such calls that is $95.00 in Gemini 3.6 Flash's favour.
- One rate here is provisional
- Gemini 3.1 Pro ships as a preview, so $2 is not a committed rate.
What this page does not tell you. Which model produces better answers. That depends on your task, and this site has no benchmark data — everything above is drawn from published pricing and specifications. Compare quality by running your own prompts through both.