Skip to content

Gemini 3.1 Pro vs Gemini 3.6 Flash

The two Gemini models people actually search for, and how far Flash gets you.

Side by side

AttributeGemini 3.1 ProGemini 3.6 Flash
Input, per 1M tokens$2.00$1.50
Output, per 1M tokens$12.00$7.50
Context window1,000,0001,000,000
ProviderGoogleGoogle
TokenizerGoogle tokenizerGoogle tokenizer
Long-context tierabove 200,000none
Retiresno date announcedno date announced

Which to choose

Reading more than you write
Gemini 3.6 Flash $1.50 against Gemini 3.1 Pro $2 per million input tokens — a modest gap of 1.3x.
Writing more than you read
Gemini 3.1 Pro prices output at 6x its input rate, Gemini 3.6 Flash at 5x — so Gemini 3.6 Flash's lead widens as answers lengthen.
Past the long-context threshold
Gemini 3.1 Pro above 200,000 bills the whole request higher — $1.50 and $4. Gemini 3.6 Flash stays ahead either side of that line.
10,000 in, 1,000 out
$0.0320 on Gemini 3.1 Pro, $0.0225 on Gemini 3.6 Flash. Over 10,000 such calls that is $95.00 in Gemini 3.6 Flash's favour.
One rate here is provisional
Gemini 3.1 Pro ships as a preview, so $2 is not a committed rate.

What this page does not tell you. Which model produces better answers. That depends on your task, and this site has no benchmark data — everything above is drawn from published pricing and specifications. Compare quality by running your own prompts through both.