C
Cloudflare
cloudflare
· 18 models · synced never
Workers AI inference at the Cloudflare edge.
| Model | Type | Input $/MTok | Output $/MTok | Context | Upstream ID |
|---|---|---|---|---|---|
| Granite 4.0 Micro ibm-granite | LLM | $0.017 | $0.112 | 131K | ibm-granite/granite-4.0-h-micro |
| Llama 3.2 1B Instruct meta-llama | LLM | $0.027 | $0.201 | 60K | meta-llama/llama-3.2-1b-instruct |
| Llama 3.2 3B Instruct meta-llama | LLM | $0.0509 | $0.335 | 80K | meta-llama/llama-3.2-3b-instruct |
| GLM 4.7 Flash z-ai | LLM | $0.0605 | $0.4 | 131K | z-ai/glm-4.7-flash |
| Gemma 4 26B A4B (free) google | LLM | $0.1 | $0.3 | 256K | google/gemma-4-26b-a4b-it |
| Llama 3.1 8B Instruct meta-llama | LLM | $0.152 | $0.287 | 32K | meta-llama/llama-3.1-8b-instruct |
| Llama 3.3 70B Instruct meta-llama | LLM | $0.293 | $2.25 | 24K | meta-llama/llama-3.3-70b-instruct |
| GLM 5.3 Flash (batch) z-ai | LLM | $0.3 | $1.00 | 1.3M | z-ai/glm-5.3-flash |
| Mistral Small 3.1 24B mistralai | LLM | $0.351 | $0.555 | 128K | mistralai/mistral-small-3.1-24b-instruct |
| DeepSeek V4 Flash 0731 deepseek | LLM | $0.44 | $1.32 | 1.3M | deepseek/deepseek-v4-flash-0731 |
| Qwen3.8 27B (free) qwen | LLM | $0.45 | $3.20 | 262K | qwen/qwen3.8-27b |
| Qwen2.5 Coder 32B Instruct qwen | LLM | $0.66 | $1.00 | 33K | qwen/qwen-2.5-coder-32b-instruct |
| Kimi K2.6 moonshotai | LLM | $0.95 | $4.00 | 262K | moonshotai/kimi-k2.6 |
| Kimi K2.7 Code moonshotai | LLM | $0.95 | $4.00 | 262K | moonshotai/kimi-k2.7-code |
| DeepSeek V4 Pro 0423 deepseek | LLM | $1.15 | $2.55 | 1.0M | deepseek/deepseek-v4-pro |
| DeepSeek V4 Pro 0813 deepseek | LLM | $1.32 | $3.96 | 1.0M | deepseek/deepseek-v4-pro-0813 |
| GLM 5.2 (free) z-ai | LLM | $1.40 | $4.40 | 262K | z-ai/glm-5.2 |
| GLM 5.3 (batch) z-ai | LLM | $1.40 | $4.40 | 1.3M | z-ai/glm-5.3 |