| Kimi K2 0711 Instruct FP4 | baseten | 131.072 | - Input
0,40 $ pro 1M tokens niedrig- Input (cached read)
0,20 $ pro 1M tokens niedrig- Output
1,80 $ pro 1M tokens niedrig
|
| baseten/MiniMaxAI/MiniMax-M2.5 | baseten | — | - Input
0,30 $ pro 1M tokens niedrig- Output
1,20 $ pro 1M tokens niedrig
|
| baseten/nvidia/Nemotron-120B-A12B | baseten | — | - Input
0,30 $ pro 1M tokens niedrig- Output
0,75 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-5 | baseten | — | - Input
0,95 $ pro 1M tokens niedrig- Output
3,15 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-4.7 | baseten | 200.000 | - Input
0,60 $ pro 1M tokens niedrig- Input (cached read)
0,12 $ pro 1M tokens niedrig- Output
2,20 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-4.6 | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
2,20 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2.5 | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
3,00 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2-Thinking | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
2,50 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2-Instruct-0905 | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
2,50 $ pro 1M tokens niedrig
|
| baseten/openai/gpt-oss-120b | baseten | 128.072 | - Input
0,10 $ pro 1M tokens niedrig- Input (cached read)
0,10 $ pro 1M tokens niedrig- Output
0,50 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V3.1 | baseten | — | - Input
0,50 $ pro 1M tokens niedrig- Output
1,50 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V3-0324 | baseten | — | - Input
0,77 $ pro 1M tokens niedrig- Output
0,77 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-5.3 | baseten | 1.048.576 | - Input
1,40 $ pro 1M tokens niedrig- Input (cached read)
0,14 $ pro 1M tokens niedrig- Output
4,40 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V4.1-Flash | baseten | 1.048.576 | - Input
0,30 $ pro 1M tokens niedrig- Input (cached read)
0,007 $ pro 1M tokens niedrig- Output
1,20 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2.6 | baseten | 262.000 | - Input
0,95 $ pro 1M tokens niedrig- Input (cached read)
0,16 $ pro 1M tokens niedrig- Output
4,00 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2.7-Code | baseten | 262.000 | - Input
0,95 $ pro 1M tokens niedrig- Input (cached read)
0,16 $ pro 1M tokens niedrig- Output
4,00 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K3 | baseten | 1.048.576 | - Input
3,00 $ pro 1M tokens niedrig- Input (cached read)
0,30 $ pro 1M tokens niedrig- Output
15,00 $ pro 1M tokens niedrig
|
| baseten/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B | baseten | 202.800 | - Input
0,60 $ pro 1M tokens niedrig- Input (cached read)
0,12 $ pro 1M tokens niedrig- Output
2,40 $ pro 1M tokens niedrig
|
| baseten/thinkingmachines/inkling | baseten | 1.048.576 | - Input
1,00 $ pro 1M tokens niedrig- Input (cached read)
0,17 $ pro 1M tokens niedrig- Output
4,05 $ pro 1M tokens niedrig
|
| baseten/thinkingmachines/inkling-small | baseten | 1.048.576 | - Input
0,50 $ pro 1M tokens niedrig- Input (cached read)
0,10 $ pro 1M tokens niedrig- Output
1,20 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-5.2 | baseten | 1.048.576 | - Input
1,40 $ pro 1M tokens niedrig- Input (cached read)
0,14 $ pro 1M tokens niedrig- Output
4,40 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-5.3-Flash | baseten | 1.048.576 | - Input
0,15 $ pro 1M tokens niedrig- Input (cached read)
0,03 $ pro 1M tokens niedrig- Output
0,50 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V4-Flash-0731 | baseten | 1.048.576 | - Input
0,13 $ pro 1M tokens niedrig- Input (cached read)
0,028 $ pro 1M tokens niedrig- Output
0,26 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V4-Pro | baseten | 1.048.576 | - Input
1,74 $ pro 1M tokens niedrig- Input (cached read)
0,145 $ pro 1M tokens niedrig- Output
3,48 $ pro 1M tokens niedrig
|
| baseten/zai-org/GLM-5.2-Fast | baseten | 1.048.576 | - Input
2,10 $ pro 1M tokens niedrig- Input (cached read)
0,21 $ pro 1M tokens niedrig- Output
6,60 $ pro 1M tokens niedrig
|