Modelle

5.697 Modelle angezeigt. Die Beträge gelten pro der in jeder Zeile angegebenen Einheit.
ModellAnbieterKontextPreise
Grok 4.1 Fast (Reasoning)xai128.000
Input

1,25 $ pro 1M tokens

unbestätigt · umstritten
Input (>200k ctx)

2,50 $ pro 1M tokens

niedrig
Input (cached read, >200k ctx)

0,40 $ pro 1M tokens

niedrig
Input (cached read)

0,20 $ pro 1M tokens

unbestätigt · umstritten

…2 weitere Preise

GPT-5.6 Sol (Official API)infer271.999
Input

2,50 $ pro 1M tokens

niedrig
Input (cache write)

3,13 $ pro 1M tokens

niedrig
Input (cached read)

0,25 $ pro 1M tokens

niedrig
Output

12,50 $ pro 1M tokens

niedrig
GPT-6 Astra (Official API)infer271.999
Input

12,50 $ pro 1M tokens

niedrig
Input (cache write)

15,63 $ pro 1M tokens

niedrig
Input (cached read)

1,25 $ pro 1M tokens

niedrig
Output

62,50 $ pro 1M tokens

niedrig
Step 1 (32K)stepfun32.768
Input

2,05 $ pro 1M tokens

niedrig
Input (cached read)

0,41 $ pro 1M tokens

niedrig
Output

9,59 $ pro 1M tokens

niedrig
Step 2 (16K)stepfun16.384
Input

5,21 $ pro 1M tokens

niedrig
Input (cached read)

1,04 $ pro 1M tokens

niedrig
Output

16,44 $ pro 1M tokens

niedrig
Qwen3-Coder 30B-A3B Instructpendra262.144
Input

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Qwen3.6 27Bpendra262.144
Input

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Llama-3.3-70B-Instructpendra128.000
Input

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
GLM-4.7-Flashpendra203.000
Input

0,00 $ pro 1M tokens

niedrig
Input (cache write)

0,00 $ pro 1M tokens

niedrig
Input (cached read)

0,00 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Gemma 4 26B A4B ITscaleway262.000
Input

0,25 $ pro 1M tokens

niedrig
Input (cached read)

0,049 $ pro 1M tokens

niedrig
Output

0,50 $ pro 1M tokens

niedrig
BGE Multilingual Gemma2scaleway8.191
Input

0,10 $ pro 1M tokens

niedrig
Output

0,00 $ pro 1M tokens

niedrig
Qwen2.5-Omni 7Bqwen32.768
Input

0,087 $ pro 1M tokens

niedrig
Input (audio)

5,45 $ pro 1M tokens

niedrig
Output

0,345 $ pro 1M tokens

niedrig
Qwen-MT Turboqwen16.384
Input

0,101 $ pro 1M tokens

niedrig
Output

0,28 $ pro 1M tokens

niedrig
Qwen Deep Researchqwen1.000.000
Input

7,74 $ pro 1M tokens

niedrig
Output

23,37 $ pro 1M tokens

niedrig
Moonshot Kimi K2 Instructmoonshotai131.072
Input

0,574 $ pro 1M tokens

niedrig
Output

2,29 $ pro 1M tokens

niedrig
Qwen-Omni Turbo Realtimeqwen32.768
Input

0,23 $ pro 1M tokens

niedrig
Input (audio)

3,58 $ pro 1M tokens

niedrig
Output

0,918 $ pro 1M tokens

niedrig
Output (audio)

7,17 $ pro 1M tokens

niedrig
Qwen3-VL 30B-A3Bqwen131.072
Input

0,108 $ pro 1M tokens

niedrig
Output

0,431 $ pro 1M tokens

niedrig
Output (reasoning)

1,08 $ pro 1M tokens

niedrig
Qwen-VL OCRqwen34.096
Input

0,043 $ pro 1M tokens

niedrig
Output

0,072 $ pro 1M tokens

niedrig
Tongyi Intent Detect V3qwen8.192
Input

0,058 $ pro 1M tokens

niedrig
Output

0,144 $ pro 1M tokens

niedrig
Qwen2.5-Math 7B Instructqwen4.096
Input

0,144 $ pro 1M tokens

niedrig
Output

0,287 $ pro 1M tokens

niedrig
Qwen3-ASR Flashqwen53.248
Input

0,032 $ pro 1M tokens

niedrig
Output

0,032 $ pro 1M tokens

niedrig
Qwen3-Omni Flashqwen65.536
Input

0,058 $ pro 1M tokens

niedrig
Input (audio)

3,58 $ pro 1M tokens

niedrig
Output

0,23 $ pro 1M tokens

niedrig
Output (audio)

7,17 $ pro 1M tokens

niedrig
Qwen2.5-Math 72B Instructqwen4.096
Input

0,574 $ pro 1M tokens

niedrig
Output

1,72 $ pro 1M tokens

niedrig
Qwen3-Omni Flash Realtimeqwen65.536
Input

0,23 $ pro 1M tokens

niedrig
Input (audio)

3,58 $ pro 1M tokens

niedrig
Output

0,918 $ pro 1M tokens

niedrig
Output (audio)

7,17 $ pro 1M tokens

niedrig
Qwen2.5 14B Instructqwen131.072
Input

0,144 $ pro 1M tokens

niedrig
Output

0,431 $ pro 1M tokens

niedrig

Zuletzt aktualisiert 08.10.2026, 02:10