Voltar aos modelos
disponível

Google: Gemini 3.1 Flash Lite

google/gemini-3-1-flash-lite

Preços

Input

US$ 0,25 por 1M tokens

verificado
Input (audio tokens)

US$ 0,25 por 1M tokens

baixa
Input (audio, cached read)

US$ 0,05 por 1M tokens

média
Input (audio)

US$ 0,50 por 1M tokens

média
Input (batch)

US$ 0,125 por 1M tokens

verificado
Input (cached read, batch)

US$ 0,0125 por 1M tokens

verificado
Input (cached read, flex)

US$ 0,0125 por 1M tokens

verificado
Input (cached read, priority)

US$ 0,045 por 1M tokens

verificado
Input (cached read)

US$ 0,025 por 1M tokens

verificado
Input (flex)

US$ 0,125 por 1M tokens

verificado
Input (priority)

US$ 0,45 por 1M tokens

verificado
Output

US$ 1,50 por 1M tokens

verificado
Output (batch)

US$ 0,75 por 1M tokens

verificado
Output (flex)

US$ 0,75 por 1M tokens

verificado
Output (priority)

US$ 2,70 por 1M tokens

verificado
Output (reasoning)

US$ 1,50 por 1M tokens

baixa
Web search (context size: high)

US$ 0,014 por query

baixa
Web search (context size: low)

US$ 0,014 por query

baixa
Web search (context size: medium)

US$ 0,014 por query

baixa

Também disponível em

O mesmo modelo, vendido por outros. O preço acima é o de quem o fabrica.

18 canais de venda para este modelo.
CanalItemPreço
Google Vertex AIInputUS$ 0,25 por 1M tokens
Input (audio tokens)US$ 0,90 por 1M tokens
Input (audio tokens)US$ 0,25 por 1M tokens
Input (audio, cached read)US$ 0,05 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (batch)US$ 0,125 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read, batch)US$ 0,0125 por 1M tokens
Input (cached read, flex)US$ 0,0125 por 1M tokens
Input (cached read, priority)US$ 0,045 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
Input (flex)US$ 0,125 por 1M tokens
Input (priority)US$ 0,45 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Output (batch)US$ 0,75 por 1M tokens
Output (flex)US$ 0,75 por 1M tokens
Output (priority)US$ 2,70 por 1M tokens
Output (reasoning)US$ 1,50 por 1M tokens
Web search (context size: high)US$ 0,014 por query
Web search (context size: low)US$ 0,014 por query
Web search (context size: medium)US$ 0,014 por query
DatabricksInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cache write)US$ 0,3125 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
OpenRouter · googleImage (input)US$ 0,00 por image
InputUS$ 0,25 por 1M tokens
Input (audio, cached read)US$ 0,05 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Output (reasoning)US$ 1,50 por 1M tokens
QueryUS$ 0,014 por query
OpenRouterInputUS$ 0,25 por 1M tokens
Input (audio, cached read)US$ 0,05 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Output (reasoning)US$ 1,50 por 1M tokens
Outro · edenaiInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Output (reasoning)US$ 1,50 por 1M tokens
Outro · google-ai-studioInputUS$ 0,25 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · impossiblInputUS$ 0,25 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · kiloInputUS$ 0,125 por 1M tokens
Input (cache write)US$ 0,0417 por 1M tokens
Input (cached read)US$ 0,0125 por 1M tokens
OutputUS$ 0,75 por 1M tokens
Output (reasoning)US$ 0,75 por 1M tokens
Outro · merge-gatewayInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · nano-gptInputUS$ 0,25 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · nearaiInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · ofoxInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cache write)US$ 1,00 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · openrouterInputUS$ 0,25 por 1M tokens
Input (cache write)US$ 0,0833 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Output (reasoning)US$ 1,50 por 1M tokens
Outro · orcarouterInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · poeInputUS$ 0,25 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · temprInputUS$ 0,25 por 1M tokens
Input (audio)US$ 0,50 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · vercelInputUS$ 0,25 por 1M tokens
Input (cached read)US$ 0,03 por 1M tokens
OutputUS$ 1,50 por 1M tokens
Outro · zenmuxInputUS$ 0,25 por 1M tokens
Input (cached read)US$ 0,025 por 1M tokens
OutputUS$ 1,50 por 1M tokens

Especificações

Provedor
google
Janela de contexto
1.048.576
Máximo de tokens de saída
65.536
Pesos abertos
não
Lançado
07/05/2026
Descontinuado
07/05/2027
Visto pela primeira vez por este rastreador

Histórico de preços

Nenhum preço deste modelo mudou desde que começamos a segui-lo.

Input
0,25 por 1M tokens, sem alterações desde 14/09/2026
Output
1,5 por 1M tokens, sem alterações desde 14/09/2026
Input (cached read)
0,025 por 1M tokens, sem alterações desde 14/09/2026
Input (batch)
0,125 por 1M tokens, sem alterações desde 14/09/2026
Output (batch)
0,75 por 1M tokens, sem alterações desde 14/09/2026
Input (cached read, batch)
0,0125 por 1M tokens, sem alterações desde 14/09/2026
Input (flex)
0,125 por 1M tokens, sem alterações desde 14/09/2026
Output (flex)
0,75 por 1M tokens, sem alterações desde 14/09/2026
Input (cached read, flex)
0,0125 por 1M tokens, sem alterações desde 14/09/2026
Input (priority)
0,45 por 1M tokens, sem alterações desde 14/09/2026
Output (priority)
2,7 por 1M tokens, sem alterações desde 14/09/2026
Input (cached read, priority)
0,045 por 1M tokens, sem alterações desde 14/09/2026
Input (audio)
0,5 por 1M tokens, sem alterações desde 14/09/2026
Input (audio, cached read)
0,05 por 1M tokens, sem alterações desde 14/09/2026
Output (reasoning)
1,5 por 1M tokens, sem alterações desde 14/09/2026
Web search (context size: low)
0,014 por query, sem alterações desde 14/09/2026
Web search (context size: medium)
0,014 por query, sem alterações desde 14/09/2026
Web search (context size: high)
0,014 por query, sem alterações desde 14/09/2026
Input (audio tokens)
0,25 por 1M tokens, sem alterações desde 17/09/2026

Última atualização 03/10/2026, 02:10