Model pricing · Google
Gemini 3.5 Flash pricing
Gemini 3.5 Flash costs $1.5 per million input tokens and $9 per million output tokens at list price. Best for: google flagship-speed tier.
| Input | $1.5 / 1M tokens |
|---|---|
| Output | $9 / 1M tokens |
| Cached input (read) | $0.15 / 1M tokens |
| Context window | Not published |
Context caching bills $1.00/hr storage.
Source: Google pricing, cross-checked against models.dev on 2026-09-15. Your provider invoice is authoritative.
What it costs in practice
| Workload | List price | If input is cached |
|---|---|---|
| One request: 10,000 tokens in, 1,000 out | $0.024 | $0.011 |
| A month of that request, 100 times a day | $72.00 | $31.50 |
| A 3,000-token instruction file on 100 requests a day, for a month | $13.50 | $1.35 |
The last row is the always-on cost of an instruction file like CLAUDE.md: paid on every request, whether the task needs it or not. Check how much of yours is waste.
Price history
From tokencheat's dated pricing snapshots. Only changes are listed.
| Since | Input / 1M | Output / 1M |
|---|---|---|
| 2026-01-01 | $1.5 | $9 |
Similarly priced models
Mistral Medium 3.5
$1.5 in · $7.5 out · Mistral
GLM-5.2
$1.4 in · $4.4 out · Zhipu (Z.ai)
GPT-5.3 Codex
$1.75 in · $14 out · OpenAI
Gemini 2.5 Pro
$1.25 in · $10 out · Google
Compare every model side by side or price your own workload.