| Nemotron 3 Super 120B | aws | — | - Input
US$ 0,15 por 1M tokens baixa- Output
US$ 0,65 por 1M tokens baixa
|
| GPT-5.4 | aws | — | - Input
US$ 2,75 por 1M tokens baixa- Input (cached read)
US$ 0,275 por 1M tokens baixa- Output
US$ 16,50 por 1M tokens baixa
|
| GPT-5.5 | aws | — | - Input
US$ 5,50 por 1M tokens baixa- Input (cached read)
US$ 0,55 por 1M tokens baixa- Output
US$ 33,00 por 1M tokens baixa
|
| gpt-oss-120b | aws | — | - Input
US$ 0,15 por 1M tokens baixa- Output
US$ 0,60 por 1M tokens baixa
|
| gpt-oss-20b | aws | — | - Input
US$ 0,07 por 1M tokens baixa- Output
US$ 0,30 por 1M tokens baixa
|
| Qwen3 32B (dense) | aws | — | - Input
US$ 0,15 por 1M tokens baixa- Output
US$ 0,60 por 1M tokens baixa
|
| Qwen3-Coder-30B-A3B-Instruct | aws | — | - Input
US$ 0,15 por 1M tokens baixa- Output
US$ 0,60 por 1M tokens baixa
|
| Qwen3-Coder-480B-A35B-Instruct | aws | — | - Input
US$ 0,45 por 1M tokens baixa- Output
US$ 1,80 por 1M tokens baixa
|
| Qwen3-VL-235B-A22B-Instruct | aws | — | - Input
US$ 0,53 por 1M tokens baixa- Output
US$ 2,66 por 1M tokens baixa
|
| Llama 3.3 70B | cerebras | 128.000 | - Input
US$ 0,85 por 1M tokens média- Output
US$ 1,20 por 1M tokens média
|
| Llama 3.1 8B | cerebras | 128.000 | - Input
US$ 0,10 por 1M tokens média- Output
US$ 0,10 por 1M tokens média
|
| Qwen 3 32B | cerebras | 128.000 | - Input
US$ 0,40 por 1M tokens média- Output
US$ 0,80 por 1M tokens média
|
| Command | cohere | 4.096 | - Input
US$ 1,00 por 1M tokens média- Output
US$ 2,00 por 1M tokens média
|
| Command R | cohere | 128.000 | - Input
US$ 0,15 por 1M tokens baixa- Output
US$ 0,60 por 1M tokens baixa
|
| Command R+ | cohere | 128.000 | - Input
US$ 2,50 por 1M tokens baixa- Output
US$ 10,00 por 1M tokens baixa
|
| Command R7B | cohere | — | - Input
US$ 0,0375 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| Qwen3 14B | qwen | — | - Input
US$ 0,05 por 1M tokens baixa- Output
US$ 0,60 por 1M tokens baixa
|
| Qwen3 VL 30B A3B Instruct | qwen | — | - Input
US$ 0,16 por 1M tokens baixa- Output
US$ 0,80 por 1M tokens baixa
|
| Qwen3.5 35B A3B | qwen | — | - Input
US$ 0,25 por 1M tokens baixa- Output
US$ 2,00 por 1M tokens baixa
|
| Nemotron 3 Super 120B A12B | nvidia | 262.144 | - Input
US$ 0,30 por 1M tokens baixa- Input (cached read)
US$ 0,30 por 1M tokens baixa- Output
US$ 1,00 por 1M tokens baixa
|
| Deepseek V3.2 | deepseek | 163.840 | - Input
US$ 0,56 por 1M tokens média- Input (cached read)
US$ 0,28 por 1M tokens baixa- Output
US$ 1,68 por 1M tokens média
|
| GLM-4.7 | z-ai | 202.800 | - Input
US$ 0,60 por 1M tokens média- Input (cached read)
US$ 0,30 por 1M tokens baixa- Output
US$ 2,20 por 1M tokens média
|
| GLM-5.1 | z-ai | 202.800 | - Input
US$ 1,40 por 1M tokens média- Input (cached read)
US$ 0,26 por 1M tokens média- Output
US$ 4,40 por 1M tokens média
|
| GLM 5.2 | z-ai | 1.048.576 | - Input
US$ 1,40 por 1M tokens média- Input (cached read, priority)
US$ 0,175 por 1M tokens baixa- Input (cached read)
US$ 0,14 por 1M tokens média- Input (priority)
US$ 1,75 por 1M tokens baixa
…2 preços a mais |
| Kimi K2.5 | moonshotai | 262.144 | - Input
US$ 0,60 por 1M tokens média- Input (cached read)
US$ 0,10 por 1M tokens média- Output
US$ 3,00 por 1M tokens média
|