Voltar aos modelos
obsoleto

OpenAI: GPT-4o-mini (2024-07-18)

openai/gpt-4o-mini

Preços

Input

US$ 0,15 por 1M tokens

alta
Input (batch)

US$ 0,075 por 1M tokens

baixa
Input (cache write)

US$ 0,15 por 1M tokens

baixa
Input (cached read, priority)

US$ 0,125 por 1M tokens

baixa
Input (cached read)

US$ 0,075 por 1M tokens

alta
Input (priority)

US$ 0,25 por 1M tokens

baixa
Output

US$ 0,60 por 1M tokens

alta
Output (batch)

US$ 0,30 por 1M tokens

baixa
Output (priority)

US$ 1,00 por 1M tokens

baixa
Web search (context size: high)

US$ 0,025 por query

baixa
Web search (context size: low)

US$ 0,025 por query

baixa
Web search (context size: medium)

US$ 0,025 por query

baixa

Também disponível em

O mesmo modelo, vendido por outros. O preço acima é o de quem o fabrica.

17 canais de venda para este modelo.
CanalItemPreço
Microsoft AzureInputUS$ 0,15 por 1M tokens
Input (batch)US$ 0,075 por 1M tokens
Input (batch) (region: eu)US$ 0,083 por 1M tokens
Input (batch) (region: us)US$ 0,083 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
Input (cached read) (region: eu)US$ 0,083 por 1M tokens
Input (cached read) (region: us)US$ 0,083 por 1M tokens
Input (deployment: global-standard)US$ 0,15 por 1M tokens
Input (region: eu)US$ 0,165 por 1M tokens
Input (region: us)US$ 0,165 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Output (batch)US$ 0,30 por 1M tokens
Output (deployment: global-standard)US$ 0,60 por 1M tokens
Output (region: eu)US$ 0,66 por 1M tokens
Output (region: us)US$ 0,66 por 1M tokens
Microsoft Azure · azure-aiInputUS$ 0,15 por 1M tokens
Input (cache write)US$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
OpenRouter · openaiInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
OpenRouterInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · cloudflare-ai-gatewayInputUS$ 0,075 por 1M tokens
Input (cached read)US$ 0,0375 por 1M tokens
OutputUS$ 0,30 por 1M tokens
Outro · crossmodelInputUS$ 0,15 por 1M tokens
Input (cache write)US$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · edenaiInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · impossiblInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · kiloInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · llmgateway-providersInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · merge-gatewayInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · nano-gptInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · ofoxInputUS$ 0,12 por 1M tokens
Input (cached read)US$ 0,06 por 1M tokens
OutputUS$ 0,48 por 1M tokens
Outro · openrouterInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · orcarouterInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens
Outro · poeInputUS$ 0,14 por 1M tokens
Input (cached read)US$ 0,068 por 1M tokens
OutputUS$ 0,54 por 1M tokens
Outro · vercelInputUS$ 0,15 por 1M tokens
Input (cached read)US$ 0,075 por 1M tokens
OutputUS$ 0,60 por 1M tokens

Especificações

Provedor
openai
Janela de contexto
128.000
Máximo de tokens de saída
16.384
Pesos abertos
não
Lançado
18/07/2024
Descontinuado
14/04/2027
Visto pela primeira vez por este rastreador

Histórico de preços

Nenhum preço deste modelo mudou desde que começamos a segui-lo.

Input
0,15 por 1M tokens, sem alterações desde 14/09/2026
Input (cached read)
0,075 por 1M tokens, sem alterações desde 14/09/2026
Output
0,6 por 1M tokens, sem alterações desde 14/09/2026
Input (cached read, priority)
0,125 por 1M tokens, sem alterações desde 14/09/2026
Input (batch)
0,075 por 1M tokens, sem alterações desde 14/09/2026
Input (priority)
0,25 por 1M tokens, sem alterações desde 14/09/2026
Output (batch)
0,3 por 1M tokens, sem alterações desde 14/09/2026
Output (priority)
1 por 1M tokens, sem alterações desde 14/09/2026
Web search (context size: high)
0,025 por query, sem alterações desde 14/09/2026
Web search (context size: low)
0,025 por query, sem alterações desde 14/09/2026
Web search (context size: medium)
0,025 por query, sem alterações desde 14/09/2026
Input (cache write)
0,15 por 1M tokens, sem alterações desde 17/09/2026

Última atualização 02/10/2026, 02:10