| Grok 4.1 Fast (Reasoning) | xai | 128.000 | - Input
1,25 $ pro 1M tokens unbestätigt · umstritten- Input (>200k ctx)
2,50 $ pro 1M tokens niedrig- Input (cached read, >200k ctx)
0,40 $ pro 1M tokens niedrig- Input (cached read)
0,20 $ pro 1M tokens unbestätigt · umstritten
…2 weitere Preise |
| GPT-5.6 Sol (Official API) | infer | 271.999 | - Input
2,50 $ pro 1M tokens niedrig- Input (cache write)
3,13 $ pro 1M tokens niedrig- Input (cached read)
0,25 $ pro 1M tokens niedrig- Output
12,50 $ pro 1M tokens niedrig
|
| GPT-6 Astra (Official API) | infer | 271.999 | - Input
12,50 $ pro 1M tokens niedrig- Input (cache write)
15,63 $ pro 1M tokens niedrig- Input (cached read)
1,25 $ pro 1M tokens niedrig- Output
62,50 $ pro 1M tokens niedrig
|
| Step 1 (32K) | stepfun | 32.768 | - Input
2,05 $ pro 1M tokens niedrig- Input (cached read)
0,41 $ pro 1M tokens niedrig- Output
9,59 $ pro 1M tokens niedrig
|
| Step 2 (16K) | stepfun | 16.384 | - Input
5,21 $ pro 1M tokens niedrig- Input (cached read)
1,04 $ pro 1M tokens niedrig- Output
16,44 $ pro 1M tokens niedrig
|
| Qwen3-Coder 30B-A3B Instruct | pendra | 262.144 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Qwen3.6 27B | pendra | 262.144 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Llama-3.3-70B-Instruct | pendra | 128.000 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| GLM-4.7-Flash | pendra | 203.000 | - Input
0,00 $ pro 1M tokens niedrig- Input (cache write)
0,00 $ pro 1M tokens niedrig- Input (cached read)
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Gemma 4 26B A4B IT | scaleway | 262.000 | - Input
0,25 $ pro 1M tokens niedrig- Input (cached read)
0,049 $ pro 1M tokens niedrig- Output
0,50 $ pro 1M tokens niedrig
|
| BGE Multilingual Gemma2 | scaleway | 8.191 | - Input
0,10 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Qwen2.5-Omni 7B | qwen | 32.768 | - Input
0,087 $ pro 1M tokens niedrig- Input (audio)
5,45 $ pro 1M tokens niedrig- Output
0,345 $ pro 1M tokens niedrig
|
| Qwen-MT Turbo | qwen | 16.384 | - Input
0,101 $ pro 1M tokens niedrig- Output
0,28 $ pro 1M tokens niedrig
|
| Qwen Deep Research | qwen | 1.000.000 | - Input
7,74 $ pro 1M tokens niedrig- Output
23,37 $ pro 1M tokens niedrig
|
| Moonshot Kimi K2 Instruct | moonshotai | 131.072 | - Input
0,574 $ pro 1M tokens niedrig- Output
2,29 $ pro 1M tokens niedrig
|
| Qwen-Omni Turbo Realtime | qwen | 32.768 | - Input
0,23 $ pro 1M tokens niedrig- Input (audio)
3,58 $ pro 1M tokens niedrig- Output
0,918 $ pro 1M tokens niedrig- Output (audio)
7,17 $ pro 1M tokens niedrig
|
| Qwen3-VL 30B-A3B | qwen | 131.072 | - Input
0,108 $ pro 1M tokens niedrig- Output
0,431 $ pro 1M tokens niedrig- Output (reasoning)
1,08 $ pro 1M tokens niedrig
|
| Qwen-VL OCR | qwen | 34.096 | - Input
0,043 $ pro 1M tokens niedrig- Output
0,072 $ pro 1M tokens niedrig
|
| Tongyi Intent Detect V3 | qwen | 8.192 | - Input
0,058 $ pro 1M tokens niedrig- Output
0,144 $ pro 1M tokens niedrig
|
| Qwen2.5-Math 7B Instruct | qwen | 4.096 | - Input
0,144 $ pro 1M tokens niedrig- Output
0,287 $ pro 1M tokens niedrig
|
| Qwen3-ASR Flash | qwen | 53.248 | - Input
0,032 $ pro 1M tokens niedrig- Output
0,032 $ pro 1M tokens niedrig
|
| Qwen3-Omni Flash | qwen | 65.536 | - Input
0,058 $ pro 1M tokens niedrig- Input (audio)
3,58 $ pro 1M tokens niedrig- Output
0,23 $ pro 1M tokens niedrig- Output (audio)
7,17 $ pro 1M tokens niedrig
|
| Qwen2.5-Math 72B Instruct | qwen | 4.096 | - Input
0,574 $ pro 1M tokens niedrig- Output
1,72 $ pro 1M tokens niedrig
|
| Qwen3-Omni Flash Realtime | qwen | 65.536 | - Input
0,23 $ pro 1M tokens niedrig- Input (audio)
3,58 $ pro 1M tokens niedrig- Output
0,918 $ pro 1M tokens niedrig- Output (audio)
7,17 $ pro 1M tokens niedrig
|
| Qwen2.5 14B Instruct | qwen | 131.072 | - Input
0,144 $ pro 1M tokens niedrig- Output
0,431 $ pro 1M tokens niedrig
|