| GLM 5.3 Fast | z-ai | 1 048 576 | - Input
2,10 $US par 1M tokens moyenne- Input (cached read)
0,39 $US par 1M tokens moyenne- Output
6,60 $US par 1M tokens moyenne
|
| together_ai/arcee-ai/trinity-mini | together-ai | 128 000 | - Input
0,045 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| together_ai/deepseek-ai/deepseek-coder-33b-instruct | together-ai | 16 384 | - Input
0,80 $US par 1M tokens faible- Output
0,80 $US par 1M tokens faible
|
| together_ai/deepseek-ai/DeepSeek-R1-Distill-Llama-70B | together-ai | 131 072 | - Input
2,00 $US par 1M tokens faible- Output
2,00 $US par 1M tokens faible
|
| together_ai/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B | together-ai | 131 072 | - Input
0,18 $US par 1M tokens faible- Output
0,18 $US par 1M tokens faible
|
| together_ai/deepseek-ai/DeepSeek-R1-Distill-Qwen-14B | together-ai | 131 072 | - Input
1,60 $US par 1M tokens faible- Output
1,60 $US par 1M tokens faible
|
| together_ai/google/gemma-2-27b-it | together-ai | 8 192 | - Input
0,80 $US par 1M tokens faible- Output
0,80 $US par 1M tokens faible
|
| together_ai/meta-llama/Llama-3-8b-chat-hf | together-ai | 8 192 | - Input
0,20 $US par 1M tokens faible- Output
0,20 $US par 1M tokens faible
|
| together_ai/meta-llama/Llama-3.1-405B-Instruct | together-ai | 4 096 | - Input
3,50 $US par 1M tokens faible- Output
3,50 $US par 1M tokens faible
|
| together_ai/meta-llama/Llama-3.2-1B-Instruct | together-ai | 131 072 | - Input
0,06 $US par 1M tokens faible- Output
0,06 $US par 1M tokens faible
|
| together_ai/meta-llama/Llama-3.2-3B-Instruct | together-ai | 131 072 | - Input
0,06 $US par 1M tokens faible- Output
0,06 $US par 1M tokens faible
|
| together_ai/meta-llama/Meta-Llama-3-70B-Instruct-Turbo | together-ai | 8 192 | - Input
0,88 $US par 1M tokens faible- Output
0,88 $US par 1M tokens faible
|
| together_ai/meta-llama/Meta-Llama-3-8B-Instruct | together-ai | 8 192 | - Input
0,20 $US par 1M tokens faible- Output
0,20 $US par 1M tokens faible
|
| together_ai/NousResearch/Nous-Hermes-2-Mixtral-8x7B-DPO | together-ai | 32 768 | - Input
0,60 $US par 1M tokens faible- Output
0,60 $US par 1M tokens faible
|
| together_ai/nvidia/Llama-3.1-Nemotron-70B-Instruct-HF | together-ai | 32 768 | - Input
0,88 $US par 1M tokens faible- Output
0,88 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2-1.5B-Instruct | together-ai | 32 768 | - Input
0,02 $US par 1M tokens faible- Output
0,02 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2-72B-Instruct | together-ai | 32 768 | - Input
0,90 $US par 1M tokens faible- Output
0,90 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2-VL-72B-Instruct | together-ai | 32 768 | - Input
1,20 $US par 1M tokens faible- Output
1,20 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2.5-14B-Instruct | together-ai | 32 768 | - Input
0,80 $US par 1M tokens faible- Output
0,80 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2.5-72B-Instruct | together-ai | 32 768 | - Input
1,20 $US par 1M tokens faible- Output
1,20 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2.5-Coder-32B-Instruct | together-ai | 16 384 | - Input
0,80 $US par 1M tokens faible- Output
0,80 $US par 1M tokens faible
|
| together_ai/Qwen/Qwen2.5-VL-72B-Instruct | together-ai | 32 768 | - Input
1,95 $US par 1M tokens faible- Output
8,00 $US par 1M tokens faible
|
| DeepSeek V4.1 Flash (Runware) | runware | 1 048 576 | - Input
0,15 $US par 1M tokens faible- Input (cached read)
0,01 $US par 1M tokens faible- Output
0,60 $US par 1M tokens faible
|
| aihubmix/agnes-2.5-flash | aihubmix | 512 000 | - Input
0,03 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| aihubmix/agnes-2.5-pro | aihubmix | 1 000 000 | - Input
0,45 $US par 1M tokens faible- Input (cached read)
0,00378 $US par 1M tokens faible- Output
0,90 $US par 1M tokens faible
|