Back to models
available
Google: Gemini 3.1 Flash Lite
google/gemini-3-1-flash-litePrices
- Input
$0.25 per 1M tokens
verified- Input (audio tokens)
$0.25 per 1M tokens
low- Input (audio, cached read)
$0.05 per 1M tokens
medium- Input (audio)
$0.50 per 1M tokens
medium- Input (batch)
$0.125 per 1M tokens
verified- Input (cached read, batch)
$0.0125 per 1M tokens
verified- Input (cached read, flex)
$0.0125 per 1M tokens
verified- Input (cached read, priority)
$0.045 per 1M tokens
verified- Input (cached read)
$0.025 per 1M tokens
verified- Input (flex)
$0.125 per 1M tokens
verified- Input (priority)
$0.45 per 1M tokens
verified- Output
$1.50 per 1M tokens
verified- Output (batch)
$0.75 per 1M tokens
verified- Output (flex)
$0.75 per 1M tokens
verified- Output (priority)
$2.70 per 1M tokens
verified- Output (reasoning)
$1.50 per 1M tokens
low- Web search (context size: high)
$0.014 per query
low- Web search (context size: low)
$0.014 per query
low- Web search (context size: medium)
$0.014 per query
low
Also available on
The same model, sold by someone else. The price above is the maker's own.
| Channel | What | Price |
|---|---|---|
| Google Vertex AI | Input | $0.25 per 1M tokens |
| Input (audio tokens) | $0.90 per 1M tokens | |
| Input (audio tokens) | $0.25 per 1M tokens | |
| Input (audio, cached read) | $0.05 per 1M tokens | |
| Input (audio) | $0.50 per 1M tokens | |
| Input (batch) | $0.125 per 1M tokens | |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read, batch) | $0.0125 per 1M tokens | |
| Input (cached read, flex) | $0.0125 per 1M tokens | |
| Input (cached read, priority) | $0.045 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Input (flex) | $0.125 per 1M tokens | |
| Input (priority) | $0.45 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Output (batch) | $0.75 per 1M tokens | |
| Output (flex) | $0.75 per 1M tokens | |
| Output (priority) | $2.70 per 1M tokens | |
| Output (reasoning) | $1.50 per 1M tokens | |
| Web search (context size: high) | $0.014 per query | |
| Web search (context size: low) | $0.014 per query | |
| Web search (context size: medium) | $0.014 per query | |
| Databricks | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cache write) | $0.3125 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| OpenRouter · google | Image (input) | $0.00 per image |
| Input | $0.25 per 1M tokens | |
| Input (audio, cached read) | $0.05 per 1M tokens | |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Output (reasoning) | $1.50 per 1M tokens | |
| Query | $0.014 per query | |
| OpenRouter | Input | $0.25 per 1M tokens |
| Input (audio, cached read) | $0.05 per 1M tokens | |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Output (reasoning) | $1.50 per 1M tokens | |
| Other · edenai | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Output (reasoning) | $1.50 per 1M tokens | |
| Other · google-ai-studio | Input | $0.25 per 1M tokens |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · impossibl | Input | $0.25 per 1M tokens |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · kilo | Input | $0.125 per 1M tokens |
| Input (cache write) | $0.0417 per 1M tokens | |
| Input (cached read) | $0.0125 per 1M tokens | |
| Output | $0.75 per 1M tokens | |
| Output (reasoning) | $0.75 per 1M tokens | |
| Other · merge-gateway | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · nano-gpt | Input | $0.25 per 1M tokens |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · nearai | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · ofox | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cache write) | $1.00 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · openrouter | Input | $0.25 per 1M tokens |
| Input (cache write) | $0.0833 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Output (reasoning) | $1.50 per 1M tokens | |
| Other · orcarouter | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · poe | Input | $0.25 per 1M tokens |
| Output | $1.50 per 1M tokens | |
| Other · tempr | Input | $0.25 per 1M tokens |
| Input (audio) | $0.50 per 1M tokens | |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · vercel | Input | $0.25 per 1M tokens |
| Input (cached read) | $0.03 per 1M tokens | |
| Output | $1.50 per 1M tokens | |
| Other · zenmux | Input | $0.25 per 1M tokens |
| Input (cached read) | $0.025 per 1M tokens | |
| Output | $1.50 per 1M tokens |
Specifications
- Provider
- Context window
- 1,048,576
- Max output tokens
- 65,536
- Open weights
- no
- Released
- 05/07/2026
- Retires
- 05/07/2027
- First seen by this tracker
Price history
No price of this model has changed since we started tracking it.
- Input
- 0.25 per 1M tokens, unchanged since 09/14/2026
- Output
- 1.5 per 1M tokens, unchanged since 09/14/2026
- Input (cached read)
- 0.025 per 1M tokens, unchanged since 09/14/2026
- Input (batch)
- 0.125 per 1M tokens, unchanged since 09/14/2026
- Output (batch)
- 0.75 per 1M tokens, unchanged since 09/14/2026
- Input (cached read, batch)
- 0.0125 per 1M tokens, unchanged since 09/14/2026
- Input (flex)
- 0.125 per 1M tokens, unchanged since 09/14/2026
- Output (flex)
- 0.75 per 1M tokens, unchanged since 09/14/2026
- Input (cached read, flex)
- 0.0125 per 1M tokens, unchanged since 09/14/2026
- Input (priority)
- 0.45 per 1M tokens, unchanged since 09/14/2026
- Output (priority)
- 2.7 per 1M tokens, unchanged since 09/14/2026
- Input (cached read, priority)
- 0.045 per 1M tokens, unchanged since 09/14/2026
- Input (audio)
- 0.5 per 1M tokens, unchanged since 09/14/2026
- Input (audio, cached read)
- 0.05 per 1M tokens, unchanged since 09/14/2026
- Output (reasoning)
- 1.5 per 1M tokens, unchanged since 09/14/2026
- Web search (context size: low)
- 0.014 per query, unchanged since 09/14/2026
- Web search (context size: medium)
- 0.014 per query, unchanged since 09/14/2026
- Web search (context size: high)
- 0.014 per query, unchanged since 09/14/2026
- Input (audio tokens)
- 0.25 per 1M tokens, unchanged since 09/17/2026