Model pricing · Google
Gemini 3.1 Flash-Lite pricing
Gemini 3.1 Flash-Lite costs $0.25 per million input tokens and $1.5 per million output tokens at list price. Best for: cheap high-volume.
| Input | $0.25 / 1M tokens |
|---|---|
| Output | $1.5 / 1M tokens |
| Cached input (read) | $0.025 / 1M tokens |
| Context window | Not published |
Source: Google pricing, cross-checked against models.dev on 2026-09-15. Your provider invoice is authoritative.
What it costs in practice
| Workload | List price | If input is cached |
|---|---|---|
| One request: 10,000 tokens in, 1,000 out | $0.0040 | $0.0018 |
| A month of that request, 100 times a day | $12.00 | $5.25 |
| A 3,000-token instruction file on 100 requests a day, for a month | $2.25 | $0.225 |
The last row is the always-on cost of an instruction file like CLAUDE.md: paid on every request, whether the task needs it or not. Check how much of yours is waste.
Price history
From tokencheat's dated pricing snapshots. Only changes are listed.
| Since | Input / 1M | Output / 1M |
|---|---|---|
| 2026-01-01 | $0.25 | $1.5 |
Similarly priced models
GPT-5.4 nano
$0.2 in · $1.25 out · OpenAI
Gemini 2.5 Flash
$0.3 in · $2.5 out · Google
Codestral
$0.3 in · $0.9 out · Mistral
Cohere Command R (08-2024)
$0.15 in · $0.6 out · Cohere
Compare every model side by side or price your own workload.