| Nemotron 2 Nano 9B | aws | — | - Input
0,06 $ pro 1M tokens niedrig- Output
0,23 $ pro 1M tokens niedrig
|
| Nemotron 3 Super 120B | aws | — | - Input
0,15 $ pro 1M tokens niedrig- Output
0,65 $ pro 1M tokens niedrig
|
| GPT-5.4 | aws | — | - Input
2,75 $ pro 1M tokens niedrig- Input (cached read)
0,275 $ pro 1M tokens niedrig- Output
16,50 $ pro 1M tokens niedrig
|
| GPT-5.5 | aws | — | - Input
5,50 $ pro 1M tokens niedrig- Input (cached read)
0,55 $ pro 1M tokens niedrig- Output
33,00 $ pro 1M tokens niedrig
|
| gpt-oss-120b | aws | — | - Input
0,15 $ pro 1M tokens niedrig- Output
0,60 $ pro 1M tokens niedrig
|
| gpt-oss-20b | aws | — | - Input
0,07 $ pro 1M tokens niedrig- Output
0,30 $ pro 1M tokens niedrig
|
| Qwen3 32B (dense) | aws | — | - Input
0,15 $ pro 1M tokens niedrig- Output
0,60 $ pro 1M tokens niedrig
|
| Qwen3-Coder-30B-A3B-Instruct | aws | — | - Input
0,15 $ pro 1M tokens niedrig- Output
0,60 $ pro 1M tokens niedrig
|
| Qwen3-Coder-480B-A35B-Instruct | aws | — | - Input
0,45 $ pro 1M tokens niedrig- Output
1,80 $ pro 1M tokens niedrig
|
| Qwen3-VL-235B-A22B-Instruct | aws | — | - Input
0,53 $ pro 1M tokens niedrig- Output
2,66 $ pro 1M tokens niedrig
|
| Llama 3.3 70B | cerebras | 128.000 | - Input
0,85 $ pro 1M tokens mittel- Output
1,20 $ pro 1M tokens mittel
|
| Llama 3.1 8B | cerebras | 128.000 | - Input
0,10 $ pro 1M tokens mittel- Output
0,10 $ pro 1M tokens mittel
|
| Qwen 3 32B | cerebras | 128.000 | - Input
0,40 $ pro 1M tokens mittel- Output
0,80 $ pro 1M tokens mittel
|
| Command | cohere | 4.096 | - Input
1,00 $ pro 1M tokens mittel- Output
2,00 $ pro 1M tokens mittel
|
| Command R | cohere | 128.000 | - Input
0,15 $ pro 1M tokens niedrig- Output
0,60 $ pro 1M tokens niedrig
|
| Command R+ | cohere | 128.000 | - Input
2,50 $ pro 1M tokens niedrig- Output
10,00 $ pro 1M tokens niedrig
|
| Command R7B | cohere | — | - Input
0,0375 $ pro 1M tokens niedrig- Output
0,15 $ pro 1M tokens niedrig
|
| Qwen3 14B | qwen | — | - Input
0,05 $ pro 1M tokens niedrig- Output
0,60 $ pro 1M tokens niedrig
|
| Qwen3 VL 30B A3B Instruct | qwen | — | - Input
0,16 $ pro 1M tokens niedrig- Output
0,80 $ pro 1M tokens niedrig
|
| Qwen3.5 35B A3B | qwen | — | - Input
0,25 $ pro 1M tokens niedrig- Output
2,00 $ pro 1M tokens niedrig
|
| Nemotron 3 Super 120B A12B | nvidia | 262.144 | - Input
0,30 $ pro 1M tokens niedrig- Input (cached read)
0,30 $ pro 1M tokens niedrig- Output
1,00 $ pro 1M tokens niedrig
|
| Deepseek V3.2 | deepseek | 163.840 | - Input
0,56 $ pro 1M tokens mittel- Input (cached read)
0,28 $ pro 1M tokens niedrig- Output
1,68 $ pro 1M tokens mittel
|
| GLM-4.7 | z-ai | 202.800 | - Input
0,60 $ pro 1M tokens mittel- Input (cached read)
0,30 $ pro 1M tokens niedrig- Output
2,20 $ pro 1M tokens mittel
|
| GLM-5.1 | z-ai | 202.800 | - Input
1,40 $ pro 1M tokens mittel- Input (cached read)
0,26 $ pro 1M tokens mittel- Output
4,40 $ pro 1M tokens mittel
|
| GLM 5.2 | z-ai | 1.048.576 | - Input
1,40 $ pro 1M tokens mittel- Input (cached read, priority)
0,175 $ pro 1M tokens niedrig- Input (cached read)
0,14 $ pro 1M tokens mittel- Input (priority)
1,75 $ pro 1M tokens niedrig
…2 weitere Preise |