Back to models
deprecated

OpenAI: GPT-4.1

openai/gpt-4-1

Prices

Input

$2.00 per 1M tokens

medium
Input (batch)

$1.00 per 1M tokens

low
Input (cached read, priority)

$0.875 per 1M tokens

low
Input (cached read)

$0.50 per 1M tokens

medium
Input (priority)

$3.50 per 1M tokens

low
Output

$8.00 per 1M tokens

medium
Output (batch)

$4.00 per 1M tokens

low
Output (priority)

$14.00 per 1M tokens

low
Web search (context size: high)

$0.025 per query

low
Web search (context size: low)

$0.025 per query

low
Web search (context size: medium)

$0.025 per query

low

Also available on

The same model, sold by someone else. The price above is the maker's own.

18 sales channels for this model.
ChannelWhatPrice
Microsoft AzureInput$2.00 per 1M tokens
Input (batch)$1.00 per 1M tokens
Input (batch) (region: eu)$1.10 per 1M tokens
Input (batch) (region: us)$1.10 per 1M tokens
Input (cached read, priority)$0.875 per 1M tokens
Input (cached read, priority) (region: eu)$0.963 per 1M tokens
Input (cached read, priority) (region: us)$0.963 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Input (cached read) (region: eu)$0.55 per 1M tokens
Input (cached read) (region: us)$0.55 per 1M tokens
Input (priority)$3.50 per 1M tokens
Input (priority) (region: eu)$3.85 per 1M tokens
Input (priority) (region: us)$3.85 per 1M tokens
Input (region: eu)$2.20 per 1M tokens
Input (region: us)$2.20 per 1M tokens
Output$8.00 per 1M tokens
Output (batch)$4.00 per 1M tokens
Output (batch) (region: eu)$4.40 per 1M tokens
Output (batch) (region: us)$4.40 per 1M tokens
Output (priority)$14.00 per 1M tokens
Output (priority) (region: eu)$15.40 per 1M tokens
Output (priority) (region: us)$15.40 per 1M tokens
Output (region: eu)$8.80 per 1M tokens
Output (region: us)$8.80 per 1M tokens
Microsoft Azure · azure-aiInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
OpenRouter · openaiInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Query$0.01 per query
OpenRouterInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · cloudflare-ai-gatewayInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · edenaiInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · fastrouterInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · impossiblInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · kiloInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · llmgateway-providersInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · merge-gatewayInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · nano-gptInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · nearaiInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · ofoxInput$1.60 per 1M tokens
Input (cached read)$0.40 per 1M tokens
Output$6.40 per 1M tokens
Other · openrouterInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · orcarouterInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens
Other · poeInput$1.80 per 1M tokens
Input (cached read)$0.45 per 1M tokens
Output$7.20 per 1M tokens
Other · vercelInput$2.00 per 1M tokens
Input (cached read)$0.50 per 1M tokens
Output$8.00 per 1M tokens

Specifications

Provider
openai
Context window
1,047,576
Max output tokens
32,768
Open weights
no
Released
04/14/2025
Retires
04/14/2027
First seen by this tracker

Price history

No price of this model has changed since we started tracking it.

Input
2 per 1M tokens, unchanged since 09/14/2026
Input (cached read)
0.5 per 1M tokens, unchanged since 09/14/2026
Output
8 per 1M tokens, unchanged since 09/14/2026
Input (cached read, priority)
0.875 per 1M tokens, unchanged since 09/14/2026
Input (batch)
1 per 1M tokens, unchanged since 09/14/2026
Input (priority)
3.5 per 1M tokens, unchanged since 09/14/2026
Output (batch)
4 per 1M tokens, unchanged since 09/14/2026
Output (priority)
14 per 1M tokens, unchanged since 09/14/2026
Web search (context size: high)
0.025 per query, unchanged since 09/14/2026
Web search (context size: low)
0.025 per query, unchanged since 09/14/2026
Web search (context size: medium)
0.025 per query, unchanged since 09/14/2026

Last updated 10/03/2026, 02:10