Back to models
available
Google: Gemini Flash Latest
google/gemini-flashPrices
- Input
$0.75 per 1M tokens
medium- Input (audio)
$0.75 per 1M tokens
low- Input (batch)
$0.375 per 1M tokens
low- Input (cached read, flex)
$0.0375 per 1M tokens
low- Input (cached read, priority)
$0.135 per 1M tokens
low- Input (cached read)
$0.075 per 1M tokens
medium- Input (flex)
$0.375 per 1M tokens
low- Input (priority)
$1.35 per 1M tokens
low- Output
$3.75 per 1M tokens
medium- Output (batch)
$1.88 per 1M tokens
low- Output (flex)
$1.88 per 1M tokens
low- Output (priority)
$6.75 per 1M tokens
low- Output (reasoning)
$3.75 per 1M tokens
low- Web search (context size: high)
$0.014 per query
low- Web search (context size: low)
$0.014 per query
low- Web search (context size: medium)
$0.014 per query
low
Also available on
The same model, sold by someone else. The price above is the maker's own.
| Channel | What | Price |
|---|---|---|
| Google Vertex AI | Input | $0.75 per 1M tokens |
| Input (audio) | $0.75 per 1M tokens | |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| OpenRouter · ~google | Image (input) | $0.000001 per image |
| Input | $0.75 per 1M tokens | |
| Input (audio, cached read) | $0.075 per 1M tokens | |
| Input (audio) | $0.75 per 1M tokens | |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| Query | $0.014 per query | |
| OpenRouter · google | Image (input) | $0.000001 per image |
| Input | $0.75 per 1M tokens | |
| Input (audio, cached read) | $0.075 per 1M tokens | |
| Input (audio) | $0.75 per 1M tokens | |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| Query | $0.014 per query | |
| OpenRouter | Input | $0.75 per 1M tokens |
| Input (audio, cached read) | $0.075 per 1M tokens | |
| Input (audio) | $0.75 per 1M tokens | |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| Other · edenai | Input | $0.75 per 1M tokens |
| Input (audio) | $0.75 per 1M tokens | |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| Other · kilo | Input | $0.75 per 1M tokens |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| Other · merge-gateway | Input | $1.50 per 1M tokens |
| Input (audio) | $1.50 per 1M tokens | |
| Input (cached read) | $0.15 per 1M tokens | |
| Output | $9.00 per 1M tokens | |
| Other · nano-gpt | Input | $0.75 per 1M tokens |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Other · openrouter | Input | $0.75 per 1M tokens |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens | |
| Output (reasoning) | $3.75 per 1M tokens | |
| Other · orcarouter | Input | $0.50 per 1M tokens |
| Input (cached read) | $0.10 per 1M tokens | |
| Output | $3.00 per 1M tokens | |
| Other · tempr | Input | $0.75 per 1M tokens |
| Input (audio) | $0.75 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $3.75 per 1M tokens |
Specifications
- Provider
- Context window
- 1,048,576
- Max output tokens
- 65,536
- Open weights
- no
- Released
- 08/13/2026
- Retires
- no date announced
- First seen by this tracker
Price history
Input (cached read) — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 0.03 |
| 09/17/2026 | 0.075 |
Input (audio) — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 1 |
| 09/17/2026 | 0.75 |
Input — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 0.3 |
| 09/17/2026 | 0.75 |
Output (reasoning) — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 2.5 |
| 09/17/2026 | 3.75 |
Output — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 2.5 |
| 09/17/2026 | 3.75 |
Web search (context size: low) — per query
| From | Amount |
|---|---|
| 09/14/2026 | 0.035 |
| 09/17/2026 | 0.014 |
Web search (context size: medium) — per query
| From | Amount |
|---|---|
| 09/14/2026 | 0.035 |
| 09/17/2026 | 0.014 |
Web search (context size: high) — per query
| From | Amount |
|---|---|
| 09/14/2026 | 0.035 |
| 09/17/2026 | 0.014 |
No changes recorded
- Input (cached read, flex)
- 0.0375 per 1M tokens, unchanged since 09/17/2026
- Input (cached read, priority)
- 0.135 per 1M tokens, unchanged since 09/17/2026
- Input (batch)
- 0.375 per 1M tokens, unchanged since 09/17/2026
- Input (flex)
- 0.375 per 1M tokens, unchanged since 09/17/2026
- Input (priority)
- 1.35 per 1M tokens, unchanged since 09/17/2026
- Output (batch)
- 1.88 per 1M tokens, unchanged since 09/17/2026
- Output (flex)
- 1.88 per 1M tokens, unchanged since 09/17/2026
- Output (priority)
- 6.75 per 1M tokens, unchanged since 09/17/2026