| baseten/zai-org/GLM-4.6 | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
2,20 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2.5 | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
3,00 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2-Thinking | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
2,50 $ pro 1M tokens niedrig
|
| baseten/moonshotai/Kimi-K2-Instruct-0905 | baseten | — | - Input
0,60 $ pro 1M tokens niedrig- Output
2,50 $ pro 1M tokens niedrig
|
| baseten/openai/gpt-oss-120b | baseten | 128.072 | - Input
0,10 $ pro 1M tokens niedrig- Input (cached read)
0,10 $ pro 1M tokens niedrig- Output
0,50 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V3.1 | baseten | — | - Input
0,50 $ pro 1M tokens niedrig- Output
1,50 $ pro 1M tokens niedrig
|
| baseten/deepseek-ai/DeepSeek-V3-0324 | baseten | — | - Input
0,77 $ pro 1M tokens niedrig- Output
0,77 $ pro 1M tokens niedrig
|
| gmi/Qwen/Qwen3-VL-235B-A22B-Instruct-FP8 | gmi | 262.144 | - Input
0,30 $ pro 1M tokens niedrig- Output
1,40 $ pro 1M tokens niedrig
|
| gmi/zai-org/GLM-4.7-FP8 | gmi | 202.752 | - Input
0,40 $ pro 1M tokens niedrig- Output
2,00 $ pro 1M tokens niedrig
|
| google_pse/search | google-pse | — | |
| gradient_ai/anthropic-claude-3-opus | gradient-ai | 200.000 | - Input
15,00 $ pro 1M tokens niedrig- Output
75,00 $ pro 1M tokens niedrig
|
| gradient_ai/anthropic-claude-3.5-haiku | gradient-ai | 200.000 | - Input
0,80 $ pro 1M tokens niedrig- Output
4,00 $ pro 1M tokens niedrig
|
| gradient_ai/anthropic-claude-3.5-sonnet | gradient-ai | 200.000 | - Input
3,00 $ pro 1M tokens niedrig- Output
15,00 $ pro 1M tokens niedrig
|
| gradient_ai/anthropic-claude-3.7-sonnet | gradient-ai | 200.000 | - Input
3,00 $ pro 1M tokens niedrig- Output
15,00 $ pro 1M tokens niedrig
|
| gradient_ai/deepseek-r1-distill-llama-70b | gradient-ai | 32.768 | - Input
0,99 $ pro 1M tokens niedrig- Output
0,99 $ pro 1M tokens niedrig
|
| gradient_ai/llama3-8b-instruct | gradient-ai | 8.192 | - Input
0,20 $ pro 1M tokens niedrig- Output
0,20 $ pro 1M tokens niedrig
|
| gradient_ai/llama3.3-70b-instruct | gradient-ai | 128.000 | - Input
0,65 $ pro 1M tokens niedrig- Output
0,65 $ pro 1M tokens niedrig
|
| gradient_ai/mistral-nemo-instruct-2407 | gradient-ai | 128.000 | - Input
0,30 $ pro 1M tokens niedrig- Output
0,30 $ pro 1M tokens niedrig
|
| gradient_ai/openai-o3 | gradient-ai | 200.000 | - Input
2,00 $ pro 1M tokens niedrig- Output
8,00 $ pro 1M tokens niedrig
|
| gradient_ai/openai-o3-mini | gradient-ai | 200.000 | - Input
1,10 $ pro 1M tokens niedrig- Output
4,40 $ pro 1M tokens niedrig
|
| lemonade/Qwen3-Coder-30B-A3B-Instruct-GGUF | lemonade | 262.144 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| lemonade/gpt-oss-20b-mxfp4-GGUF | lemonade | 131.072 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| lemonade/gpt-oss-120b-mxfp-GGUF | lemonade | 131.072 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| lemonade/Gemma-3-4b-it-GGUF | lemonade | 128.000 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|
| lemonade/Qwen3-4B-Instruct-2507-GGUF | lemonade | 262.144 | - Input
0,00 $ pro 1M tokens niedrig- Output
0,00 $ pro 1M tokens niedrig
|