Models

5,658 models shown. Amounts are per the unit stated in each row.
ModelProviderContextPrices
MiniMax-M2.7-highspeedauriko204,800
Input

$0.30 per 1M tokens

low
Input (cache write)

$0.375 per 1M tokens

low
Input (cached read)

$0.03 per 1M tokens

low
Output

$1.20 per 1M tokens

low
Grok 4.3auriko1,000,000
Input

$1.25 per 1M tokens

low
Input (>200k ctx)

$2.50 per 1M tokens

low
Input (cache write, >200k ctx)

$2.50 per 1M tokens

low
Input (cache write)

$1.25 per 1M tokens

low

…4 more prices

Qwen3.6 Plusauriko1,000,000
Input

$0.50 per 1M tokens

low
Input (>256k ctx)

$2.00 per 1M tokens

low
Input (cache write, >256k ctx)

$2.50 per 1M tokens

low
Input (cached read, >256k ctx)

$0.20 per 1M tokens

low

…3 more prices

Claude Opus 4.6auriko1,000,000
Input

$5.00 per 1M tokens

low
Input (>200k ctx)

$10.00 per 1M tokens

low
Input (cache write, >200k ctx)

$12.50 per 1M tokens

low
Input (cache write)

$6.25 per 1M tokens

low

…4 more prices

Claude Opus 4.7auriko1,000,000
Input

$5.00 per 1M tokens

low
Input (>200k ctx)

$10.00 per 1M tokens

low
Input (cache write, >200k ctx)

$12.50 per 1M tokens

low
Input (cache write)

$6.25 per 1M tokens

low

…4 more prices

MiniMax-M2.7auriko204,800
Input

$0.15 per 1M tokens

low
Input (cache write)

$0.375 per 1M tokens

low
Input (cached read)

$0.03 per 1M tokens

low
Output

$0.60 per 1M tokens

low
baidu/ERNIE-4.5-300B-A47Bbaidu131,000
Input

$0.28 per 1M tokens

low
Output

$1.10 per 1M tokens

low
inclusionAI/Ling-flash-2.0inclusionai131,000
Input

$0.14 per 1M tokens

low
Output

$0.57 per 1M tokens

low
Qwen: Qwen3 235B A22B Thinking 2507qwen131,072
Input

$0.22 per 1M tokens

low
Output

$0.88 per 1M tokens

low
Qwen: Qwen3 VL 30B A3B Instructqwen262,144
Input

$0.15 per 1M tokens

low
Output

$0.60 per 1M tokens

low
Qwen: Qwen3 VL 8B Instructqwen262,144
Input

$0.20 per 1M tokens

low
Output

$0.20 per 1M tokens

low
Qwen: Qwen3 VL 32B Instructqwen131,072
Input

$0.90 per 1M tokens

low
Output

$0.90 per 1M tokens

low
Qwen2.5 72B Instructqwen131,072
Input

$0.574 per 1M tokens

low
Output

$1.72 per 1M tokens

low
Qwen2.5 7B Instructqwen131,072
Input

$0.072 per 1M tokens

low
Output

$0.144 per 1M tokens

low
Qwen3-Coder 480B-A35B Instructqwen262,144
Input

$0.45 per 1M tokens

low
Output

$1.80 per 1M tokens

low
Qwen/Qwen3-VL-32B-Thinkingqwen131,072
Input

$0.16 per 1M tokens

low
Output

$2.87 per 1M tokens

low
Qwen: Qwen3 VL 30B A3B Thinkingqwen262,144
Input

$0.15 per 1M tokens

low
Output

$0.60 per 1M tokens

low
Tencent: Hunyuan A13B Instructtencent131,072
Input

$0.14 per 1M tokens

unverified
Output

$0.57 per 1M tokens

unverified
PaddlePaddle/PaddleOCR-VL-1.5paddlepaddle16,384
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Nova 2 Litenova1,000,000
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Output (reasoning)

$0.00 per 1M tokens

low
Nova 2 Pronova1,000,000
Input

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low
Output (reasoning)

$0.00 per 1M tokens

low
DeepSeek V3.2nvidia131,072
Input

$0.55 per 1M tokens

low
Output

$1.65 per 1M tokens

low
NVIDIA Nemotron 3 Nano Omninvidia262,144
Input

$0.13 per 1M tokens

low
Output

$0.38 per 1M tokens

low
NVIDIA Nemotron Cascade 2nvidia262,144
Input

$0.15 per 1M tokens

low
Output

$0.60 per 1M tokens

low
gpt-oss:20bollama-cloud131,072
Input

$0.07 per 1M tokens

low
Input (cached read)

$0.035 per 1M tokens

low
Output

$0.30 per 1M tokens

low

Last updated 10/05/2026, 02:10