Skip to content

Gemini 3.6 Flash vs GPT-4o

The fast tier from each provider, which is where most everyday API traffic actually goes.

Side by side

AttributeGemini 3.6 FlashGPT-4o
Input, per 1M tokens$1.50$2.50
Output, per 1M tokens$7.50$10.00
Context window1,000,000128,000
ProviderGoogleOpenAI
TokenizerGoogle tokenizero200k_base
Long-context tiernonenone
Retiresno date announcedno date announced

Which to choose

Reading more than you write
Gemini 3.6 Flash $1.50 against GPT-4o $2.50 per million input tokens — a modest gap of 1.7x.
Writing more than you read
Gemini 3.6 Flash prices output at 5x its input rate, GPT-4o at four times — so Gemini 3.6 Flash's lead widens as answers lengthen.
Whole documents in one request
GPT-4o stops at 128,000; Gemini 3.6 Flash takes 1,000,000. Past the lower ceiling only Gemini 3.6 Flash accepts the request, at any price.
10,000 in, 1,000 out
$0.0225 on Gemini 3.6 Flash, $0.0350 on GPT-4o. Over 10,000 such calls that is $125.00 in Gemini 3.6 Flash's favour.
How these counts were arrived at
GPT-4o is counted exactly in your browser; Gemini 3.6 Flash is estimated from a characters-per-token ratio. Close, not exact.

What this page does not tell you. Which model produces better answers. That depends on your task, and this site has no benchmark data — everything above is drawn from published pricing and specifications. Compare quality by running your own prompts through both.