| Mistral 7B Instruct v0.3 | mistral | 32 768 | - Input
0,20 $US par 1M tokens faible- Input (cache write)
0,20 $US par 1M tokens faible- Input (cached read)
0,20 $US par 1M tokens faible- Output
0,20 $US par 1M tokens faible
|
| Ministral 8B Instruct | mistral | 128 000 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| Nemotron 3 Nano 30B A3B | nvidia | 262 144 | - Input
0,05 $US par 1M tokens faible- Input (cache write)
0,05 $US par 1M tokens faible- Input (cached read)
0,05 $US par 1M tokens faible- Output
0,20 $US par 1M tokens faible
|
| Nemotron 3 Ultra 550B A55B | nvidia | 1 000 000 | - Input
0,50 $US par 1M tokens faible- Input (cache write)
0,50 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
2,50 $US par 1M tokens faible
|
| Nemotron 3 Super 120B A12B | nvidia | 256 000 | - Input
0,09 $US par 1M tokens faible- Input (cache write)
0,09 $US par 1M tokens faible- Input (cached read)
0,09 $US par 1M tokens faible- Output
0,45 $US par 1M tokens faible
|
| Nemotron 3.5 Lightning 30B A3B | nvidia | 262 144 | - Input
0,50 $US par 1M tokens faible- Input (cache write)
0,50 $US par 1M tokens faible- Input (cached read)
0,50 $US par 1M tokens faible- Output
0,50 $US par 1M tokens faible
|
| GLM-5.2 | z-ai | 1 000 000 | - Input
2,10 $US par 1M tokens faible- Input (cache write)
2,10 $US par 1M tokens faible- Input (cached read)
0,21 $US par 1M tokens faible- Output
6,60 $US par 1M tokens faible
|
| LFM2-24B-A2B | liquid | 32 768 | - Input
0,03 $US par 1M tokens faible- Input (cache write)
0,03 $US par 1M tokens faible- Input (cached read)
0,03 $US par 1M tokens faible- Output
0,12 $US par 1M tokens faible
|
| GLiGuard LLM Guardrails 300M | fastino | 8 192 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| GLiNER2 Base | fastino | 8 192 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| GLiNER2 Multi | fastino | 8 192 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| GLiNER2 Multi Large | fastino | 8 192 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| GLiNER2 Privacy Filter PII (Multi) | fastino | 8 192 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| GLiNER2 Large | fastino | 8 192 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| Qwen3 4B Instruct | qwen | 262 144 | - Input
0,20 $US par 1M tokens faible- Output
0,20 $US par 1M tokens faible
|
| Qwen3 1.7B Base | qwen | 32 768 | - Input
0,10 $US par 1M tokens faible- Input (cache write)
0,10 $US par 1M tokens faible- Input (cached read)
0,10 $US par 1M tokens faible- Output
0,10 $US par 1M tokens faible
|
| Qwen3 4B Base | qwen | 32 768 | - Input
0,15 $US par 1M tokens faible- Input (cache write)
0,15 $US par 1M tokens faible- Input (cached read)
0,15 $US par 1M tokens faible- Output
0,15 $US par 1M tokens faible
|
| Qwen2.5-Coder-0.5B | qwen | 32 768 | - Input
0,10 $US par 1M tokens faible- Input (cache write)
0,10 $US par 1M tokens faible- Input (cached read)
0,10 $US par 1M tokens faible- Output
0,10 $US par 1M tokens faible
|
| Llama-3.2-3B | meta | 131 072 | - Input
0,10 $US par 1M tokens faible- Input (cache write)
0,10 $US par 1M tokens faible- Input (cached read)
0,10 $US par 1M tokens faible- Output
0,10 $US par 1M tokens faible
|
| Llama-3.2-1B | meta | 131 072 | - Input
0,10 $US par 1M tokens faible- Input (cache write)
0,10 $US par 1M tokens faible- Input (cached read)
0,10 $US par 1M tokens faible- Output
0,10 $US par 1M tokens faible
|
| Kimi K3 | moonshotai | 1 048 576 | - Input
4,50 $US par 1M tokens moyenne- Input (cached read)
0,45 $US par 1M tokens moyenne- Output
22,50 $US par 1M tokens moyenne
|
| Meta Llama 3.1 8B Instruct Turbo | helicone | 128 000 | - Input
0,02 $US par 1M tokens faible- Output
0,03 $US par 1M tokens faible
|
| xAI Grok 3 Mini | helicone | 131 072 | - Input
0,30 $US par 1M tokens faible- Input (cached read)
0,075 $US par 1M tokens faible- Output
0,50 $US par 1M tokens faible
|
| xAI Grok 4.1 Fast Non-Reasoning | helicone | 2 000 000 | - Input
0,20 $US par 1M tokens faible- Input (cached read)
0,05 $US par 1M tokens faible- Output
0,50 $US par 1M tokens faible
|
| Llama-3.1-8B-Instruct | helicone | 128 000 | - Input
0,02 $US par 1M tokens faible- Output
0,05 $US par 1M tokens faible
|