Models

5,671 models shown. Amounts are per the unit stated in each row.
ModelProviderContextPrices
Codestral-22B-v0.1mistral128,000
Input

$0.30 per 1M tokens

low
Input (cache write)

$0.30 per 1M tokens

low
Input (cached read)

$0.30 per 1M tokens

low
Output

$0.90 per 1M tokens

low
Mistral 7B Instruct v0.3mistral32,768
Input

$0.20 per 1M tokens

low
Input (cache write)

$0.20 per 1M tokens

low
Input (cached read)

$0.20 per 1M tokens

low
Output

$0.20 per 1M tokens

low
Ministral 8B Instructmistral128,000
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
Nemotron 3 Nano 30B A3Bnvidia262,144
Input

$0.05 per 1M tokens

low
Input (cache write)

$0.05 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.20 per 1M tokens

low
Nemotron 3 Ultra 550B A55Bnvidia1,000,000
Input

$0.50 per 1M tokens

low
Input (cache write)

$0.50 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$2.50 per 1M tokens

low
Nemotron 3 Super 120B A12Bnvidia256,000
Input

$0.09 per 1M tokens

low
Input (cache write)

$0.09 per 1M tokens

low
Input (cached read)

$0.09 per 1M tokens

low
Output

$0.45 per 1M tokens

low
Nemotron 3.5 Lightning 30B A3Bnvidia262,144
Input

$0.50 per 1M tokens

low
Input (cache write)

$0.50 per 1M tokens

low
Input (cached read)

$0.50 per 1M tokens

low
Output

$0.50 per 1M tokens

low
GLM-5.2z-ai1,000,000
Input

$2.10 per 1M tokens

low
Input (cache write)

$2.10 per 1M tokens

low
Input (cached read)

$0.21 per 1M tokens

low
Output

$6.60 per 1M tokens

low
LFM2-24B-A2Bliquid32,768
Input

$0.03 per 1M tokens

low
Input (cache write)

$0.03 per 1M tokens

low
Input (cached read)

$0.03 per 1M tokens

low
Output

$0.12 per 1M tokens

low
GLiGuard LLM Guardrails 300Mfastino8,192
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
GLiNER2 Basefastino8,192
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
GLiNER2 Multifastino8,192
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
GLiNER2 Multi Largefastino8,192
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
GLiNER2 Privacy Filter PII (Multi)fastino8,192
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
GLiNER2 Largefastino8,192
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
Qwen3 4B Instructqwen262,144
Input

$0.20 per 1M tokens

low
Output

$0.20 per 1M tokens

low
Qwen3 1.7B Baseqwen32,768
Input

$0.10 per 1M tokens

low
Input (cache write)

$0.10 per 1M tokens

low
Input (cached read)

$0.10 per 1M tokens

low
Output

$0.10 per 1M tokens

low
Qwen3 4B Baseqwen32,768
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.15 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.15 per 1M tokens

low
Qwen2.5-Coder-0.5Bqwen32,768
Input

$0.10 per 1M tokens

low
Input (cache write)

$0.10 per 1M tokens

low
Input (cached read)

$0.10 per 1M tokens

low
Output

$0.10 per 1M tokens

low
Llama-3.2-3Bmeta131,072
Input

$0.10 per 1M tokens

low
Input (cache write)

$0.10 per 1M tokens

low
Input (cached read)

$0.10 per 1M tokens

low
Output

$0.10 per 1M tokens

low
Llama-3.2-1Bmeta131,072
Input

$0.10 per 1M tokens

low
Input (cache write)

$0.10 per 1M tokens

low
Input (cached read)

$0.10 per 1M tokens

low
Output

$0.10 per 1M tokens

low
Kimi K3moonshotai1,048,576
Input

$4.50 per 1M tokens

medium
Input (cached read)

$0.45 per 1M tokens

medium
Output

$22.50 per 1M tokens

medium
Meta Llama 3.1 8B Instruct Turbohelicone128,000
Input

$0.02 per 1M tokens

low
Output

$0.03 per 1M tokens

low
xAI Grok 3 Minihelicone131,072
Input

$0.30 per 1M tokens

low
Input (cached read)

$0.075 per 1M tokens

low
Output

$0.50 per 1M tokens

low
xAI Grok 4.1 Fast Non-Reasoninghelicone2,000,000
Input

$0.20 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.50 per 1M tokens

low

Last updated 10/07/2026, 02:23