Models
| Model | Provider | Context | Prices |
|---|---|---|---|
| Embed v1 0.6b | perplexity | 32,768 |
|
| Qwen3.5 0.8B | qvac | 32,768 |
|
| Qwen3.5 4B | qvac | 32,768 |
|
| Nemotron 3.5 Lightning | nvidia | 262,144 |
|
| Qwen3 14B Instruct | openpipe | 32,768 |
|
| Granite 4.1 8B | ibm | 131,072 |
|
| Mellum2 12B A2.5B | jetbrains | 131,072 |
|
| GLM-5.3 (free) | z-ai | 1,000,000 |
|
| Inkling (256K) | thinkingmachines | 262,144 |
|
| Standard Compute | standardcompute | 1,000,000 |
|
| DeepSeek: DeepSeek V3.1 | deepseek | 163,840 |
|
| Google Gemma 3 27B Instruct | venice | 198,000 |
|
| GLM 5.2 | venice | 1,000,000 |
|
| DeepSeek V4 Flash 0731 Fast | venice | 1,000,000 |
|
| Qwen 3.5 9B | venice | 256,000 |
|
| GPT-5.5 Pro | venice | 1,000,000 |
|
| GLM 4.7 Flash | venice | 128,000 |
|
| Mistral Small 3.2 24B Instruct | venice | 256,000 |
|
| Venice Uncensored 1.2 | venice | 128,000 |
|
| Gemma 4 Uncensored | venice | 256,000 |
|
| GLM 5.1 | venice | 200,000 |
|
| GPT-5.5 | venice | 1,000,000 |
…2 more prices |
| MiniMax M2.5 | venice | 198,000 |
|
| Aion 3.0 Mini | venice | 128,000 |
|
| GPT-5.2 Codex | venice | 256,000 |
|