Skip to content

Gemini 3.1 Flash-Lite pricing and token costs

Gemini 3.1 Flash-Lite is a Google model with a 1,000,000-token context window. Rates below are per million tokens, and are the ones in effect today.

In practice

At $0.25 per million input with a flat 1,000,000-token window and no long-context tier, this is the cheapest way here to send very large prompts without hitting a rate step. GPT-5.6 Luna is slightly cheaper per token at $0.20, but doubles past 272,000 tokens — above that line, Flash-Lite is the cheaper of the two. As with Gemini 2.5 Flash the quoted figure is the text rate; audio input is $0.50 and out of scope for this calculator.

Pricing

Input, per 1M tokens
$0.25
Output, per 1M tokens
$1.50
Context window
1,000,000 tokens

Input and output combined.

Tokenizer
Google tokenizer

Estimated from character length, in your browser.

Pricing data last verified: .

Provider pricing can change. Check current rates with OpenAI, Anthropic, and Google before making billing decisions.

Alternatives

Head to head

Articles that quote these figures

Common questions

How much does Gemini 3.1 Flash-Lite cost per million tokens?
$0.25 per million input tokens and $1.50 per million output tokens.
What is the context window of Gemini 3.1 Flash-Lite?
1,000,000 tokens, covering input and output combined.
Which tokenizer does Gemini 3.1 Flash-Lite use?
Gemini 3.1 Flash-Lite uses Google's own tokenizer, which is not published in a form that can run in a browser. Counts here are a character-based estimate and are labelled "est".

Figures come from Google’s published list prices and are checked against every article on this site that quotes them at build time. They are still estimates of what a request costs — verify against your own billing before committing to them. See the methodology page for how these numbers are actually calculated.