| Mistral: Mistral Nemo | mistral | 131.072 | - Input
0,15 $ pro 1M tokens niedrig- Output
0,15 $ pro 1M tokens niedrig
|
| Codestral Embed | mistral | 8.192 | - Input
0,15 $ pro 1M tokens niedrig- Input (cached read)
0,015 $ pro 1M tokens niedrig
|
| Mistral: Mistral Medium 3.5 | mistral | 262.144 | - Input
1,50 $ pro 1M tokens niedrig- Input (cached read)
0,15 $ pro 1M tokens niedrig- Output
7,50 $ pro 1M tokens niedrig
|
| Pixtral 12B | mistral | 128.000 | - Input
0,15 $ pro 1M tokens mittel- Output
0,15 $ pro 1M tokens mittel
|
| Embed v1 4b | perplexity | 32.768 | - Input
0,03 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Embed v1 0.6b | perplexity | 32.768 | - Input
0,004 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Qwen3.5 0.8B | qvac | 32.768 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Qwen3.5 4B | qvac | 32.768 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Nemotron 3.5 Lightning | nvidia | 262.144 | - Input
0,07 $ pro 1M tokens niedrig- Input (cached read)
0,04 $ pro 1M tokens niedrig- Output
0,20 $ pro 1M tokens niedrig
|
| Qwen3 14B Instruct | openpipe | 32.768 | - Input
0,05 $ pro 1M tokens niedrig- Input (cached read)
0,05 $ pro 1M tokens niedrig- Output
0,22 $ pro 1M tokens niedrig
|
| Granite 4.1 8B | ibm | 131.072 | - Input
0,05 $ pro 1M tokens niedrig- Input (cached read)
0,05 $ pro 1M tokens niedrig- Output
0,10 $ pro 1M tokens niedrig
|
| Mellum2 12B A2.5B | jetbrains | 131.072 | - Input
0,05 $ pro 1M tokens niedrig- Input (cached read)
0,05 $ pro 1M tokens niedrig- Output
0,10 $ pro 1M tokens niedrig
|
| GLM-5.3 (free) | z-ai | 1.000.000 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| Inkling (256K) | thinkingmachines | 262.144 | - Input
3,74 $ pro 1M tokens niedrig- Input (cached read)
0,748 $ pro 1M tokens niedrig- Output
9,36 $ pro 1M tokens niedrig
|
| Standard Compute | standardcompute | 1.000.000 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| DeepSeek: DeepSeek V3.1 | deepseek | 163.840 | - Input
0,25 $ pro 1M tokens mittel- Input (cached read)
0,13 $ pro 1M tokens unbestätigt · umstritten- Output
0,95 $ pro 1M tokens mittel
|
| Google Gemma 3 27B Instruct | venice | 198.000 | - Input
0,12 $ pro 1M tokens niedrig- Output
0,20 $ pro 1M tokens niedrig
|
| GLM 5.2 | venice | 1.000.000 | - Input
1,40 $ pro 1M tokens niedrig- Input (cached read)
0,26 $ pro 1M tokens niedrig- Output
4,40 $ pro 1M tokens niedrig
|
| DeepSeek V4 Flash 0731 Fast | venice | 1.000.000 | - Input
0,35 $ pro 1M tokens niedrig- Input (cached read)
0,0875 $ pro 1M tokens niedrig- Output
0,70 $ pro 1M tokens niedrig
|
| Qwen 3.5 9B | venice | 256.000 | - Input
0,09 $ pro 1M tokens niedrig- Input (cached read)
0,045 $ pro 1M tokens niedrig- Output
0,13 $ pro 1M tokens niedrig
|
| GPT-5.5 Pro | venice | 1.000.000 | - Input
37,50 $ pro 1M tokens niedrig- Output
225,00 $ pro 1M tokens niedrig
|
| GLM 4.7 Flash | venice | 128.000 | - Input
0,06 $ pro 1M tokens niedrig- Input (cached read)
0,01 $ pro 1M tokens niedrig- Output
0,40 $ pro 1M tokens niedrig
|
| Mistral Small 3.2 24B Instruct | venice | 256.000 | - Input
0,0938 $ pro 1M tokens niedrig- Output
0,25 $ pro 1M tokens niedrig
|
| Venice Uncensored 1.2 | venice | 128.000 | - Input
0,20 $ pro 1M tokens niedrig- Output
0,90 $ pro 1M tokens niedrig
|
| Gemma 4 Uncensored | venice | 256.000 | - Input
0,1625 $ pro 1M tokens niedrig- Output
0,50 $ pro 1M tokens niedrig
|