Skip to content

Gemini 3.1 Flash-Lite vs Claude Haiku 4.5

Budget against budget, where the per-token gap decides the whole bill.

Side by side

AttributeGemini 3.1 Flash-LiteClaude Haiku 4.5
Input, per 1M tokens$0.25$1.00
Output, per 1M tokens$1.50$5.00
Context window1,000,000200,000
ProviderGoogleAnthropic
TokenizerGoogle tokenizerAnthropic tokenizer
Long-context tiernonenone
Retiresno date announcedno date announced

Which to choose

Reading more than you write
Gemini 3.1 Flash-Lite $0.25 against Claude Haiku 4.5 $1 per million input tokens — a wide gap of four times.
Writing more than you read
Gemini 3.1 Flash-Lite prices output at 6x its input rate, Claude Haiku 4.5 at 5x — so Gemini 3.1 Flash-Lite's lead widens as answers lengthen.
Whole documents in one request
Claude Haiku 4.5 stops at 200,000; Gemini 3.1 Flash-Lite takes 1,000,000. Past the lower ceiling only Gemini 3.1 Flash-Lite accepts the request, at any price.
10,000 in, 1,000 out
$0.004 on Gemini 3.1 Flash-Lite, $0.0150 on Claude Haiku 4.5. Over 10,000 such calls that is $110.00 in Gemini 3.1 Flash-Lite's favour.
How these counts were arrived at
Both estimated, and on different ratios (gemini, claude-legacy). Small differences between them are noise.

What this page does not tell you. Which model produces better answers. That depends on your task, and this site has no benchmark data — everything above is drawn from published pricing and specifications. Compare quality by running your own prompts through both.