| GLM-4.7-FlashX | llmgateway | 200.000 | - Input
0,07 $ pro 1M tokens niedrig- Input (cache write)
0,00 $ pro 1M tokens niedrig- Input (cached read)
0,01 $ pro 1M tokens niedrig- Output
0,40 $ pro 1M tokens niedrig
|
| Fugu Max | llmgateway | 1.000.000 | - Input
2,00 $ pro 1M tokens niedrig- Input (cached read)
0,25 $ pro 1M tokens niedrig- Output
6,00 $ pro 1M tokens niedrig
|
| GLM-4 32B (0414-128k) | llmgateway | 128.000 | - Input
0,10 $ pro 1M tokens niedrig- Output
0,10 $ pro 1M tokens niedrig
|
| Seed 1.6 (250615) | llmgateway | 256.000 | - Input
0,25 $ pro 1M tokens niedrig- Input (cached read)
0,05 $ pro 1M tokens niedrig- Output
2,00 $ pro 1M tokens niedrig
|
| GPT-3.5-turbo | llmgateway | 16.385 | - Input
0,50 $ pro 1M tokens niedrig- Input (cached read)
0,00 $ pro 1M tokens niedrig- Output
1,50 $ pro 1M tokens niedrig
|
| Qwen3 VL 235B A22B Thinking | llmgateway | 131.072 | - Input
0,98 $ pro 1M tokens niedrig- Output
3,95 $ pro 1M tokens niedrig
|
| Llama 3.1 70B Instruct | llmgateway | 128.000 | - Input
0,72 $ pro 1M tokens niedrig- Output
0,72 $ pro 1M tokens niedrig
|
| Grok 4.20 (Reasoning) | llmgateway | 2.000.000 | - Input
2,00 $ pro 1M tokens niedrig- Input (>200k ctx)
2,50 $ pro 1M tokens niedrig- Input (cached read, >200k ctx)
0,40 $ pro 1M tokens niedrig- Input (cached read)
0,20 $ pro 1M tokens niedrig
…2 weitere Preise |
| Grok 4.20 (Non-Reasoning) | llmgateway | 2.000.000 | - Input
2,00 $ pro 1M tokens niedrig- Input (>200k ctx)
2,50 $ pro 1M tokens niedrig- Input (cached read, >200k ctx)
0,40 $ pro 1M tokens niedrig- Input (cached read)
0,20 $ pro 1M tokens niedrig
…2 weitere Preise |
| ministral-8b-2512 | llmgateway | 256.000 | - Input
0,15 $ pro 1M tokens niedrig- Input (cached read)
0,015 $ pro 1M tokens niedrig- Output
0,15 $ pro 1M tokens niedrig
|
| Seed 1.6 (250915) | llmgateway | 256.000 | - Input
0,25 $ pro 1M tokens niedrig- Input (cached read)
0,05 $ pro 1M tokens niedrig- Output
2,00 $ pro 1M tokens niedrig
|
| GPT-4 | llmgateway | 8.192 | - Input
30,00 $ pro 1M tokens niedrig- Output
60,00 $ pro 1M tokens niedrig
|
| Grok 4.20 (Non-Reasoning) | llmgateway | 2.000.000 | - Input
1,25 $ pro 1M tokens niedrig- Input (>200k ctx)
2,50 $ pro 1M tokens niedrig- Input (cached read, >200k ctx)
0,40 $ pro 1M tokens niedrig- Input (cached read)
0,20 $ pro 1M tokens niedrig
…2 weitere Preise |
| Qwen Plus Latest | llmgateway | 1.000.000 | - Input
0,40 $ pro 1M tokens niedrig- Input (cache write)
0,50 $ pro 1M tokens niedrig- Input (cached read)
0,08 $ pro 1M tokens niedrig- Output
1,20 $ pro 1M tokens niedrig
|
| Fugu Ultra v2.0 | llmgateway | 1.000.000 | - Input
5,00 $ pro 1M tokens niedrig- Input (cached read)
0,50 $ pro 1M tokens niedrig- Output
30,00 $ pro 1M tokens niedrig
|
| Llama 3 70B Instruct | llmgateway | 8.192 | - Input
0,51 $ pro 1M tokens niedrig- Output
0,74 $ pro 1M tokens niedrig
|
| Grok 4.20 (Reasoning) | llmgateway | 2.000.000 | - Input
1,25 $ pro 1M tokens niedrig- Input (>200k ctx)
2,50 $ pro 1M tokens niedrig- Input (cached read, >200k ctx)
0,40 $ pro 1M tokens niedrig- Input (cached read)
0,20 $ pro 1M tokens niedrig
…2 weitere Preise |
| BGE Multilingual Gemma2 | infomaniak | 8.000 | - Input
0,08 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| All-MiniLM-L12-v2 | infomaniak | 128 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Apertus v1.5 70B | swiss-ai | 100.000 | - Input
0,87 $ pro 1M tokens niedrig- Output
3,10 $ pro 1M tokens niedrig
|
| Ministral 3 14B Instruct | mistral | 256.000 | - Input
0,20 $ pro 1M tokens mittel- Input (flex)
0,10 $ pro 1M tokens niedrig- Output
0,20 $ pro 1M tokens mittel- Output (flex)
0,10 $ pro 1M tokens niedrig
|
| Nemotron 3 Nano 30B A3B FP8 | nvidia | 1.000.000 | - Input
0,06 $ pro 1M tokens niedrig- Output
0,25 $ pro 1M tokens niedrig
|
| Qwen3.5 122B-A10B FP8 | qwen | 200.000 | - Input
0,50 $ pro 1M tokens niedrig- Output
3,97 $ pro 1M tokens niedrig
|
| Qwen3.5 397B-A17B FP8 | qwen | 200.000 | - Input
0,99 $ pro 1M tokens niedrig- Output
4,46 $ pro 1M tokens niedrig
|
| Mercury Edit 2 | inception | 32.000 | - Input
0,25 $ pro 1M tokens niedrig- Input (cached read)
0,025 $ pro 1M tokens niedrig- Output
0,75 $ pro 1M tokens niedrig
|