Back to models
available
Llama 4 Maverick 17B 128E Instruct
meta/llama-4-maverick-17b-128e-instruct-fp8Prices
- Input
$0.25 per 1M tokens
medium- Output
$1.00 per 1M tokens
medium
Also available on
The same model, sold by someone else. The price above is the maker's own.
| Channel | What | Price |
|---|---|---|
| Microsoft Azure | Input | $0.25 per 1M tokens |
| Output | $1.00 per 1M tokens | |
| IBM watsonx.ai | Input | $0.371 per 1M tokens |
| Output | $1.48 per 1M tokens | |
| Other · abacus | Input | $0.14 per 1M tokens |
| Output | $0.59 per 1M tokens | |
| Other · deepinfra | Input | $0.20 per 1M tokens |
| Output | $0.80 per 1M tokens | |
| Other · huggingface-novita | Input | $0.27 per 1M tokens |
| Output | $0.85 per 1M tokens | |
| Other · huggingface-together | Input | $0.27 per 1M tokens |
| Output | $0.85 per 1M tokens | |
| Other · io-net | Input | $0.15 per 1M tokens |
| Input (cache write) | $0.30 per 1M tokens | |
| Input (cached read) | $0.075 per 1M tokens | |
| Output | $0.60 per 1M tokens | |
| Other · novita-ai | Input | $0.27 per 1M tokens |
| Output | $0.85 per 1M tokens | |
| Other · together | Input | $0.27 per 1M tokens |
| Output | $0.85 per 1M tokens | |
| Other · watsonx | Input | $0.371 per 1M tokens |
| Output | $1.48 per 1M tokens |
Specifications
- Provider
- meta
- Context window
- 1,000,000
- Max output tokens
- 4,028
- Open weights
- yes
- Released
- 01/15/2025
- Retires
- no date announced
- First seen by this tracker
Price history
Input — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 1.41 |
| 09/17/2026 | 0.25 |
Output — per 1M tokens
| From | Amount |
|---|---|
| 09/14/2026 | 0.35 |
| 09/17/2026 | 1 |