Models

38 models shown. Amounts are per the unit stated in each row.
ModelProviderContextPrices
wandb/nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3Bwandb262,000
Input

$0.07 per 1M tokens

low
Input (cached read)

$0.04 per 1M tokens

low
Output

$0.20 per 1M tokens

low
wandb/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55Bwandb262,000
Input

$0.50 per 1M tokens

low
Input (cached read)

$0.10 per 1M tokens

low
Output

$2.15 per 1M tokens

low
wandb/OpenPipe/Qwen3-14B-Instructwandb32,800
Input

$0.05 per 1M tokens

low
Output

$0.22 per 1M tokens

low
wandb/Qwen/Qwen3.8-27Bwandb262,000
Input

$0.40 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$3.00 per 1M tokens

low
wandb/Qwen/Qwen3.6-35B-A3Bwandb262,000
Input

$0.25 per 1M tokens

low
Output

$1.25 per 1M tokens

low
wandb/Qwen/Qwen3.6-27Bwandb262,000
Input

$0.60 per 1M tokens

low
Input (cached read)

$0.12 per 1M tokens

low
Output

$3.60 per 1M tokens

low
wandb/Qwen/Qwen3.5-35B-A3Bwandb262,000
Input

$0.25 per 1M tokens

low
Output

$1.25 per 1M tokens

low
wandb/Qwen/Qwen3-30B-A3B-Instruct-2507wandb262,000
Input

$0.10 per 1M tokens

low
Output

$0.30 per 1M tokens

low
wandb/ibm-granite/granite-4.2-8bwandb131,000
Input

$0.10 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.15 per 1M tokens

low
wandb/zai-org/GLM-5.2wandb1,049,000
Input

$0.76 per 1M tokens

low
Input (cached read)

$0.14 per 1M tokens

low
Output

$2.42 per 1M tokens

low
wandb/zai-org/GLM-5.3-Flashwandb1,049,000
Input

$0.15 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.50 per 1M tokens

low
wandb/deepseek-ai/DeepSeek-V4.1-Flashwandb1,049,000
Input

$0.20 per 1M tokens

low
Input (cached read)

$0.03 per 1M tokens

low
Output

$0.65 per 1M tokens

low
wandb/google/gemma-4-26B-A4B-itwandb262,000
Input

$0.10 per 1M tokens

low
Input (cached read)

$0.05 per 1M tokens

low
Output

$0.30 per 1M tokens

low

Last updated 10/04/2026, 02:10