Models

12 models shown. Amounts are per the unit stated in each row.
ModelProviderContextPrices
gpt-oss:20bollama-cloud131,072
Input

$0.07 per 1M tokens

low
Input (cached read)

$0.035 per 1M tokens

low
Output

$0.30 per 1M tokens

low
DeepSeek V4 Flash 0731ollama-cloud1,048,576
Input

$0.22 per 1M tokens

low
Input (cached read)

$0.007 per 1M tokens

low
Output

$0.66 per 1M tokens

low
MiniMax-M2.7ollama-cloud204,800
Input

$0.819 per 1M tokens

low
Input (cache write)

$0.375 per 1M tokens

low
Input (cached read)

$0.06 per 1M tokens

low
Output

$3.28 per 1M tokens

low
MiniMax-M3ollama-cloud1,048,576
Input

$0.225 per 1M tokens

low
Input (>512k ctx)

$0.45 per 1M tokens

low
Input (>524.288k ctx)

$1.20 per 1M tokens

low
Input (cached read, >512k ctx)

$0.09 per 1M tokens

low

…5 more prices

qwen3.5:397bollama-cloud262,144
Input

$0.60 per 1M tokens

low
Output

$3.60 per 1M tokens

low
gpt-oss:120bollama-cloud131,072
Input

$0.00 per 1M tokens

low
Input (cached read)

$0.014 per 1M tokens

low
Output

$0.00 per 1M tokens

low
nemotron-3-ultraollama-cloud262,144
Input

$0.10 per 1M tokens

low
Input (cached read)

$0.10 per 1M tokens

low
Output

$3.00 per 1M tokens

low
DeepSeek V4 Pro 0813ollama-cloud1,048,576
Input

$0.66 per 1M tokens

low
Input (cached read)

$0.022 per 1M tokens

low
Output

$1.98 per 1M tokens

low
nemotron-3-nano:30bollama-cloud1,048,576
Input

$0.06 per 1M tokens

low
Output

$0.24 per 1M tokens

low
mistral-large-3:675bollama-cloud262,144
Input

$0.50 per 1M tokens

low
Output

$1.50 per 1M tokens

low
gemma4:31bollama-cloud262,144
Input

$0.14 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.40 per 1M tokens

low
nemotron-3-superollama-cloud262,144
Input

$0.015 per 1M tokens

low
Input (cached read)

$0.015 per 1M tokens

low
Output

$0.60 per 1M tokens

low

Last updated 10/04/2026, 02:10