Gemini 3.1 Flash-Lite vs Claude Haiku 4.5
Budget against budget, where the per-token gap decides the whole bill.
Side by side
| Attribute | Gemini 3.1 Flash-Lite | Claude Haiku 4.5 |
|---|---|---|
| Input, per 1M tokens | $0.25 | $1.00 |
| Output, per 1M tokens | $1.50 | $5.00 |
| Context window | 1,000,000 | 200,000 |
| Provider | Anthropic | |
| Tokenizer | Google tokenizer | Anthropic tokenizer |
| Long-context tier | none | none |
| Retires | no date announced | no date announced |
Which to choose
- Reading more than you write
- Gemini 3.1 Flash-Lite $0.25 against Claude Haiku 4.5 $1 per million input tokens — a wide gap of four times.
- Writing more than you read
- Gemini 3.1 Flash-Lite prices output at 6x its input rate, Claude Haiku 4.5 at 5x — so Gemini 3.1 Flash-Lite's lead widens as answers lengthen.
- Whole documents in one request
- Claude Haiku 4.5 stops at 200,000; Gemini 3.1 Flash-Lite takes 1,000,000. Past the lower ceiling only Gemini 3.1 Flash-Lite accepts the request, at any price.
- 10,000 in, 1,000 out
- $0.004 on Gemini 3.1 Flash-Lite, $0.0150 on Claude Haiku 4.5. Over 10,000 such calls that is $110.00 in Gemini 3.1 Flash-Lite's favour.
- How these counts were arrived at
- Both estimated, and on different ratios (gemini, claude-legacy). Small differences between them are noise.
What this page does not tell you. Which model produces better answers. That depends on your task, and this site has no benchmark data — everything above is drawn from published pricing and specifications. Compare quality by running your own prompts through both.