Modelle

5.658 Modelle angezeigt. Die Beträge gelten pro der in jeder Zeile angegebenen Einheit.
ModellAnbieterKontextPreise
MiniMax-M2.7-highspeedauriko204.800
Input

0,30 $ pro 1M tokens

niedrig
Input (cache write)

0,375 $ pro 1M tokens

niedrig
Input (cached read)

0,03 $ pro 1M tokens

niedrig
Output

1,20 $ pro 1M tokens

niedrig
Grok 4.3auriko1.000.000
Input

1,25 $ pro 1M tokens

niedrig
Input (>200k ctx)

2,50 $ pro 1M tokens

niedrig
Input (cache write, >200k ctx)

2,50 $ pro 1M tokens

niedrig
Input (cache write)

1,25 $ pro 1M tokens

niedrig

…4 weitere Preise

Qwen3.6 Plusauriko1.000.000
Input

0,50 $ pro 1M tokens

niedrig
Input (>256k ctx)

2,00 $ pro 1M tokens

niedrig
Input (cache write, >256k ctx)

2,50 $ pro 1M tokens

niedrig
Input (cached read, >256k ctx)

0,20 $ pro 1M tokens

niedrig

…3 weitere Preise

Claude Opus 4.6auriko1.000.000
Input

5,00 $ pro 1M tokens

niedrig
Input (>200k ctx)

10,00 $ pro 1M tokens

niedrig
Input (cache write, >200k ctx)

12,50 $ pro 1M tokens

niedrig
Input (cache write)

6,25 $ pro 1M tokens

niedrig

…4 weitere Preise

Claude Opus 4.7auriko1.000.000
Input

5,00 $ pro 1M tokens

niedrig
Input (>200k ctx)

10,00 $ pro 1M tokens

niedrig
Input (cache write, >200k ctx)

12,50 $ pro 1M tokens

niedrig
Input (cache write)

6,25 $ pro 1M tokens

niedrig

…4 weitere Preise

MiniMax-M2.7auriko204.800
Input

0,15 $ pro 1M tokens

niedrig
Input (cache write)

0,375 $ pro 1M tokens

niedrig
Input (cached read)

0,03 $ pro 1M tokens

niedrig
Output

0,60 $ pro 1M tokens

niedrig
baidu/ERNIE-4.5-300B-A47Bbaidu131.000
Input

0,28 $ pro 1M tokens

niedrig
Output

1,10 $ pro 1M tokens

niedrig
inclusionAI/Ling-flash-2.0inclusionai131.000
Input

0,14 $ pro 1M tokens

niedrig
Output

0,57 $ pro 1M tokens

niedrig
Qwen: Qwen3 235B A22B Thinking 2507qwen131.072
Input

0,22 $ pro 1M tokens

niedrig
Output

0,88 $ pro 1M tokens

niedrig
Qwen: Qwen3 VL 30B A3B Instructqwen262.144
Input

0,15 $ pro 1M tokens

niedrig
Output

0,60 $ pro 1M tokens

niedrig
Qwen: Qwen3 VL 8B Instructqwen262.144
Input

0,20 $ pro 1M tokens

niedrig
Output

0,20 $ pro 1M tokens

niedrig
Qwen: Qwen3 VL 32B Instructqwen131.072
Input

0,90 $ pro 1M tokens

niedrig
Output

0,90 $ pro 1M tokens

niedrig
Qwen2.5 72B Instructqwen131.072
Input

0,574 $ pro 1M tokens

niedrig
Output

1,72 $ pro 1M tokens

niedrig
Qwen2.5 7B Instructqwen131.072
Input

0,072 $ pro 1M tokens

niedrig
Output

0,144 $ pro 1M tokens

niedrig
Qwen3-Coder 480B-A35B Instructqwen262.144
Input

0,45 $ pro 1M tokens

niedrig
Output

1,80 $ pro 1M tokens

niedrig
Qwen/Qwen3-VL-32B-Thinkingqwen131.072
Input

0,16 $ pro 1M tokens

niedrig
Output

2,87 $ pro 1M tokens

niedrig
Qwen: Qwen3 VL 30B A3B Thinkingqwen262.144
Input

0,15 $ pro 1M tokens

niedrig
Output

0,60 $ pro 1M tokens

niedrig
Tencent: Hunyuan A13B Instructtencent131.072
Input

0,14 $ pro 1M tokens

unbestätigt
Output

0,57 $ pro 1M tokens

unbestätigt
PaddlePaddle/PaddleOCR-VL-1.5paddlepaddle16.384
Input

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Nova 2 Litenova1.000.000
Input

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Output (reasoning)

0,00 $ pro 1M tokens

niedrig
Nova 2 Pronova1.000.000
Input

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Output (reasoning)

0,00 $ pro 1M tokens

niedrig
DeepSeek V3.2nvidia131.072
Input

0,55 $ pro 1M tokens

niedrig
Output

1,65 $ pro 1M tokens

niedrig
NVIDIA Nemotron 3 Nano Omninvidia262.144
Input

0,13 $ pro 1M tokens

niedrig
Output

0,38 $ pro 1M tokens

niedrig
NVIDIA Nemotron Cascade 2nvidia262.144
Input

0,15 $ pro 1M tokens

niedrig
Output

0,60 $ pro 1M tokens

niedrig
gpt-oss:20bollama-cloud131.072
Input

0,07 $ pro 1M tokens

niedrig
Input (cached read)

0,035 $ pro 1M tokens

niedrig
Output

0,30 $ pro 1M tokens

niedrig

Zuletzt aktualisiert 05.10.2026, 02:10