| Codestral-22B-v0.1 | mistral | 128.000 | - Input
US$ 0,30 por 1M tokens baixa- Input (cache write)
US$ 0,30 por 1M tokens baixa- Input (cached read)
US$ 0,30 por 1M tokens baixa- Output
US$ 0,90 por 1M tokens baixa
|
| Mistral 7B Instruct v0.3 | mistral | 32.768 | - Input
US$ 0,20 por 1M tokens baixa- Input (cache write)
US$ 0,20 por 1M tokens baixa- Input (cached read)
US$ 0,20 por 1M tokens baixa- Output
US$ 0,20 por 1M tokens baixa
|
| Ministral 8B Instruct | mistral | 128.000 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| Nemotron 3 Nano 30B A3B | nvidia | 262.144 | - Input
US$ 0,05 por 1M tokens baixa- Input (cache write)
US$ 0,05 por 1M tokens baixa- Input (cached read)
US$ 0,05 por 1M tokens baixa- Output
US$ 0,20 por 1M tokens baixa
|
| Nemotron 3 Ultra 550B A55B | nvidia | 1.000.000 | - Input
US$ 0,50 por 1M tokens baixa- Input (cache write)
US$ 0,50 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 2,50 por 1M tokens baixa
|
| Nemotron 3 Super 120B A12B | nvidia | 256.000 | - Input
US$ 0,09 por 1M tokens baixa- Input (cache write)
US$ 0,09 por 1M tokens baixa- Input (cached read)
US$ 0,09 por 1M tokens baixa- Output
US$ 0,45 por 1M tokens baixa
|
| Nemotron 3.5 Lightning 30B A3B | nvidia | 262.144 | - Input
US$ 0,50 por 1M tokens baixa- Input (cache write)
US$ 0,50 por 1M tokens baixa- Input (cached read)
US$ 0,50 por 1M tokens baixa- Output
US$ 0,50 por 1M tokens baixa
|
| GLM-5.2 | z-ai | 1.000.000 | - Input
US$ 2,10 por 1M tokens baixa- Input (cache write)
US$ 2,10 por 1M tokens baixa- Input (cached read)
US$ 0,21 por 1M tokens baixa- Output
US$ 6,60 por 1M tokens baixa
|
| LFM2-24B-A2B | liquid | 32.768 | - Input
US$ 0,03 por 1M tokens baixa- Input (cache write)
US$ 0,03 por 1M tokens baixa- Input (cached read)
US$ 0,03 por 1M tokens baixa- Output
US$ 0,12 por 1M tokens baixa
|
| GLiGuard LLM Guardrails 300M | fastino | 8.192 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| GLiNER2 Base | fastino | 8.192 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| GLiNER2 Multi | fastino | 8.192 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| GLiNER2 Multi Large | fastino | 8.192 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| GLiNER2 Privacy Filter PII (Multi) | fastino | 8.192 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| GLiNER2 Large | fastino | 8.192 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| Qwen3 4B Instruct | qwen | 262.144 | - Input
US$ 0,20 por 1M tokens baixa- Output
US$ 0,20 por 1M tokens baixa
|
| Qwen3 1.7B Base | qwen | 32.768 | - Input
US$ 0,10 por 1M tokens baixa- Input (cache write)
US$ 0,10 por 1M tokens baixa- Input (cached read)
US$ 0,10 por 1M tokens baixa- Output
US$ 0,10 por 1M tokens baixa
|
| Qwen3 4B Base | qwen | 32.768 | - Input
US$ 0,15 por 1M tokens baixa- Input (cache write)
US$ 0,15 por 1M tokens baixa- Input (cached read)
US$ 0,15 por 1M tokens baixa- Output
US$ 0,15 por 1M tokens baixa
|
| Qwen2.5-Coder-0.5B | qwen | 32.768 | - Input
US$ 0,10 por 1M tokens baixa- Input (cache write)
US$ 0,10 por 1M tokens baixa- Input (cached read)
US$ 0,10 por 1M tokens baixa- Output
US$ 0,10 por 1M tokens baixa
|
| Llama-3.2-3B | meta | 131.072 | - Input
US$ 0,10 por 1M tokens baixa- Input (cache write)
US$ 0,10 por 1M tokens baixa- Input (cached read)
US$ 0,10 por 1M tokens baixa- Output
US$ 0,10 por 1M tokens baixa
|
| Llama-3.2-1B | meta | 131.072 | - Input
US$ 0,10 por 1M tokens baixa- Input (cache write)
US$ 0,10 por 1M tokens baixa- Input (cached read)
US$ 0,10 por 1M tokens baixa- Output
US$ 0,10 por 1M tokens baixa
|
| Kimi K3 | moonshotai | 1.048.576 | - Input
US$ 4,50 por 1M tokens média- Input (cached read)
US$ 0,45 por 1M tokens média- Output
US$ 22,50 por 1M tokens média
|
| Meta Llama 3.1 8B Instruct Turbo | helicone | 128.000 | - Input
US$ 0,02 por 1M tokens baixa- Output
US$ 0,03 por 1M tokens baixa
|
| xAI Grok 3 Mini | helicone | 131.072 | - Input
US$ 0,30 por 1M tokens baixa- Input (cached read)
US$ 0,075 por 1M tokens baixa- Output
US$ 0,50 por 1M tokens baixa
|
| xAI Grok 4.1 Fast Non-Reasoning | helicone | 2.000.000 | - Input
US$ 0,20 por 1M tokens baixa- Input (cached read)
US$ 0,05 por 1M tokens baixa- Output
US$ 0,50 por 1M tokens baixa
|