| GLM-5.3 Flash (Runware) | runware | 1,000,000 | - Input
- Input (cached read)
- Output
|
| Gemma 4 31B IT (Runware) | runware | 262,144 | - Input
- Input (cached read)
- Output
|
| DeepSeek V4 Pro (Runware) | runware | 1,048,576 | - Input
- Input (cached read)
- Output
|
| GPT OSS 120B (Runware) | runware | 131,072 | - Input
- Input (cached read)
- Output
|
| GLM-5.3 (Runware) | runware | 1,000,000 | - Input
- Input (cached read)
- Output
|
| Cerebras-Llama-4-Scout-17B-16E-Instruct | llama | 128,000 | |
| Llama-4-Maverick-17B-128E-Instruct-FP8 | llama | 128,000 | - Input
- Input (cached read)
- Output
|
| Groq-Llama-4-Maverick-17B-128E-Instruct | llama | 128,000 | |
| Llama-4-Scout-17B-16E-Instruct-FP8 | llama | 128,000 | |
| Cerebras-Llama-4-Maverick-17B-128E-Instruct | llama | 128,000 | |
| Llama-3.3-8B-Instruct | llama | 128,000 | |
| HappyHorse 1.1 Reference-to-Video | alibaba-token-plan | 0 | |
| Wan2.7 Image Pro | qwen | 8,192 | |
| HappyHorse 1.1 Text-to-Video | alibaba-token-plan | 0 | |
| Qwen Image 2.0 Pro | qwen | 8,192 | |
| HappyHorse 1.1 Image-to-Video | alibaba-token-plan | 0 | |
| Qwen Image 2.0 | qwen | 8,192 | |
| Wan2.7 Image | qwen | 8,192 | |
| GLM 5.2 Short Fast Flex | neuralwatt | 199,984 | - Input
- Input (cached read)
- Output
|
| GLM 5.2 Short Flex | neuralwatt | 199,984 | - Input
- Input (cached read)
- Output
|
| GLM 5.2 Flex | neuralwatt | 1,048,560 | - Input
- Input (cached read)
- Output
|
| Kimi K2.7 Code Flex | neuralwatt | 262,128 | - Input
- Input (cached read)
- Output
|
| Kimi K2.7 Code Fast | neuralwatt | 262,128 | - Input
- Input (cached read)
- Output
|
| GLM 5.2 Short Fast | neuralwatt | 199,984 | - Input
- Input (cached read)
- Output
|
| Kimi K3 | neuralwatt | 1,040,384 | - Input
- Input (cached read)
- Output
|