Back to models
deprecated
OpenAI: GPT-4o-mini (2024-07-18)
openai/gpt-4o-miniPrices
- Input
$0.15 per 1M tokens
high- Input (batch)
$0.075 per 1M tokens
low- Input (cache write)
$0.15 per 1M tokens
low- Input (cached read, priority)
$0.125 per 1M tokens
low- Input (cached read)
$0.075 per 1M tokens
high- Input (priority)
$0.25 per 1M tokens
low- Output
$0.60 per 1M tokens
high- Output (batch)
$0.30 per 1M tokens
low- Output (priority)
$1.00 per 1M tokens
low- Web search (context size: high)
$0.025 per query
low- Web search (context size: low)
$0.025 per query
low- Web search (context size: medium)
$0.025 per query
low
Also available on
The same model, sold by someone else. The price above is the maker's own.
| Channel | What | Price |
|---|---|---|
| Microsoft Azure | Input | $0.15 per 1M tokens |
| Input (batch) | $0.075 per 1M tokens | |
| Input (batch) (region: eu) | $0.083 per 1M tokens | |
| Input (batch) (region: us) | $0.083 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Input (cached read) (region: eu) | $0.083 per 1M tokens | |
| Input (cached read) (region: us) | $0.083 per 1M tokens | |
| Input (deployment: global-standard) | $0.15 per 1M tokens | |
| Input (region: eu) | $0.165 per 1M tokens | |
| Input (region: us) | $0.165 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Output (batch) | $0.30 per 1M tokens | |
| Output (deployment: global-standard) | $0.60 per 1M tokens | |
| Output (region: eu) | $0.66 per 1M tokens | |
| Output (region: us) | $0.66 per 1M tokens | |
| Microsoft Azure · azure-ai | Input | $0.15 per 1M tokens |
| Input (cache write) | $0.15 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| OpenRouter · openai | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| OpenRouter | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · cloudflare-ai-gateway | Input | $0.075 per 1M tokens |
| Input (cached read) | $0.0375 per 1M tokens | |
| Output | $0.30 per 1M tokens | |
| Other · crossmodel | Input | $0.15 per 1M tokens |
| Input (cache write) | $0.15 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · edenai | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · impossibl | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · kilo | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · llmgateway-providers | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · merge-gateway | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · nano-gpt | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · ofox | Input | $0.12 per 1M tokens |
| Input (cached read) | $0.06 per 1M tokens | |
| Output | $0.48 per 1M tokens | |
| Other · openrouter | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · orcarouter | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · poe | Input | $0.14 per 1M tokens |
| Input (cached read) | $0.068 per 1M tokens | |
| Output | $0.54 per 1M tokens | |
| Other · vercel | Input | $0.15 per 1M tokens |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens |
Specifications
- Provider
- openai
- Context window
- 128,000
- Max output tokens
- 16,384
- Open weights
- no
- Released
- 07/18/2024
- Retires
- 04/14/2027
- First seen by this tracker
Price history
No price of this model has changed since we started tracking it.
- Input
- 0.15 per 1M tokens, unchanged since 09/14/2026
- Input (cached read)
- 0.075 per 1M tokens, unchanged since 09/14/2026
- Output
- 0.6 per 1M tokens, unchanged since 09/14/2026
- Input (cached read, priority)
- 0.125 per 1M tokens, unchanged since 09/14/2026
- Input (batch)
- 0.075 per 1M tokens, unchanged since 09/14/2026
- Input (priority)
- 0.25 per 1M tokens, unchanged since 09/14/2026
- Output (batch)
- 0.3 per 1M tokens, unchanged since 09/14/2026
- Output (priority)
- 1 per 1M tokens, unchanged since 09/14/2026
- Web search (context size: high)
- 0.025 per query, unchanged since 09/14/2026
- Web search (context size: low)
- 0.025 per query, unchanged since 09/14/2026
- Web search (context size: medium)
- 0.025 per query, unchanged since 09/14/2026
- Input (cache write)
- 0.15 per 1M tokens, unchanged since 09/17/2026