Models

5,671 models shown. Amounts are per the unit stated in each row.
ModelProviderContextPrices
Embed v1 0.6bperplexity32,768
Input

$0.004 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Qwen3.5 0.8Bqvac32,768
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Qwen3.5 4Bqvac32,768
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Nemotron 3.5 Lightningnvidia262,144
Input

$0.07 per 1M tokens

low
Input (cached read)

$0.04 per 1M tokens

low
Output

$0.20 per 1M tokens

low
Qwen3 14B Instructopenpipe32,768
Input

$0.05 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.22 per 1M tokens

low
Granite 4.1 8Bibm131,072
Input

$0.05 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.10 per 1M tokens

low
Mellum2 12B A2.5Bjetbrains131,072
Input

$0.05 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.10 per 1M tokens

low
GLM-5.3 (free)z-ai1,000,000
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Inkling (256K)thinkingmachines262,144
Input

$3.74 per 1M tokens

low
Input (cached read)

$0.748 per 1M tokens

low
Output

$9.36 per 1M tokens

low
Standard Computestandardcompute1,000,000
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
DeepSeek: DeepSeek V3.1deepseek163,840
Input

$0.25 per 1M tokens

medium
Input (cached read)

$0.13 per 1M tokens

unverified · disputed
Output

$0.95 per 1M tokens

medium
Google Gemma 3 27B Instructvenice198,000
Input

$0.12 per 1M tokens

low
Output

$0.20 per 1M tokens

low
GLM 5.2venice1,000,000
Input

$1.40 per 1M tokens

low
Input (cached read)

$0.26 per 1M tokens

low
Output

$4.40 per 1M tokens

low
DeepSeek V4 Flash 0731 Fastvenice1,000,000
Input

$0.35 per 1M tokens

low
Input (cached read)

$0.0875 per 1M tokens

low
Output

$0.70 per 1M tokens

low
Qwen 3.5 9Bvenice256,000
Input

$0.09 per 1M tokens

low
Input (cached read)

$0.045 per 1M tokens

low
Output

$0.13 per 1M tokens

low
GPT-5.5 Provenice1,000,000
Input

$37.50 per 1M tokens

low
Output

$225.00 per 1M tokens

low
GLM 4.7 Flashvenice128,000
Input

$0.06 per 1M tokens

low
Input (cached read)

$0.01 per 1M tokens

low
Output

$0.40 per 1M tokens

low
Mistral Small 3.2 24B Instructvenice256,000
Input

$0.0938 per 1M tokens

low
Output

$0.25 per 1M tokens

low
Venice Uncensored 1.2venice128,000
Input

$0.20 per 1M tokens

low
Output

$0.90 per 1M tokens

low
Gemma 4 Uncensoredvenice256,000
Input

$0.1625 per 1M tokens

low
Output

$0.50 per 1M tokens

low
GLM 5.1venice200,000
Input

$1.54 per 1M tokens

low
Input (cached read)

$0.286 per 1M tokens

low
Output

$4.84 per 1M tokens

low
GPT-5.5venice1,000,000
Input

$6.25 per 1M tokens

low
Input (>272k ctx)

$12.50 per 1M tokens

low
Input (cached read, >272k ctx)

$1.25 per 1M tokens

low
Input (cached read)

$0.625 per 1M tokens

low

…2 more prices

MiniMax M2.5venice198,000
Input

$0.27 per 1M tokens

low
Input (cached read)

$0.03 per 1M tokens

low
Output

$0.95 per 1M tokens

low
Aion 3.0 Minivenice128,000
Input

$0.875 per 1M tokens

low
Input (cached read)

$0.225 per 1M tokens

low
Output

$1.75 per 1M tokens

low
GPT-5.2 Codexvenice256,000
Input

$2.19 per 1M tokens

low
Input (cached read)

$0.219 per 1M tokens

low
Output

$17.50 per 1M tokens

low

Last updated 10/07/2026, 02:23