Cloudflare LLM API pricing
Every model Cloudflare serves, with its pricing per 1M tokens and the exact model ID to copy. Sorted cheapest input first. Verified against LiteLLM and OpenRouter daily.
Models served
30
on this provider
Cheapest input
$0.00
per 1M tokens
Cheapest output
$0.00
per 1M tokens
Max context
10.5M
tokens
| Model ID | |||||
|---|---|---|---|---|---|
| Llama 2 7B Chat HF Lora | Meta | 8K | $0.00 | $0.00 | cloudflare/@cf/meta-llama/llama-2-7b-chat-hf-lora |
| Gemma 7B It Lora | 4K | $0.00 | $0.00 | cloudflare/@cf/google/gemma-7b-it-lora | |
| Gemma 2B It Lora | 8K | $0.00 | $0.00 | cloudflare/@cf/google/gemma-2b-it-lora | |
| Mistral 7B Instruct V0 2 Lora | Mistral | 15K | $0.00 | $0.00 | cloudflare/@cf/mistral/mistral-7b-instruct-v0.2-lora |
| Granite 4.0 Micro | IBM | 131K | $0.02 | $0.11 | cloudflare/@cf/ibm-granite/granite-4.0-h-micro |
| Llama 3.2 1B Instruct | Meta | 131K | $0.03 | $0.20 | cloudflare/@cf/meta/llama-3.2-1b-instruct |
| Llama 3.2 11B Vision Instruct | Meta | 131K | $0.05 | $0.68 | cloudflare/@cf/meta/llama-3.2-11b-vision-instruct |
| Llama 3.2 3B Instruct | Meta | 131K | $0.05 | $0.34 | cloudflare/@cf/meta/llama-3.2-3b-instruct |
| Qwen3 30B A3b FP8 | Alibaba | 41K | $0.05 | $0.34 | cloudflare/@cf/qwen/qwen3-30b-a3b-fp8 |
| GLM 4.7 Flash | Zhipu | 203K | $0.06 | $0.40 | cloudflare/@cf/zai-org/glm-4.7-flash |
| Gemma 4 26B A4B | 262K | $0.10 | $0.30 | cloudflare/@cf/google/gemma-4-26b-a4b-it | |
| Llama 3.1 8B Instruct FP8 | Meta | 32K | $0.15 | $0.29 | cloudflare/@cf/meta/llama-3.1-8b-instruct-fp8 |
| gpt-oss-20b | OpenAI | 131K | $0.20 | $0.30 | cloudflare/@cf/openai/gpt-oss-20b |
| Llama 4 Scout 17B 16e Instruct | Meta | 10.5M | $0.27 | $0.85 | cloudflare/@cf/meta/llama-4-scout-17b-16e-instruct |
| Llama 3.3 70B Instruct FP8 Fast | Meta | 24K | $0.29 | $2.25 | cloudflare/@cf/meta/llama-3.3-70b-instruct-fp8-fast |
| gpt-oss-120b | OpenAI | 131K | $0.35 | $0.75 | cloudflare/@cf/openai/gpt-oss-120b |
| Mistral Small 3.1 24B | Mistral | 131K | $0.35 | $0.56 | cloudflare/@cf/mistralai/mistral-small-3.1-24b-instruct |
| Gemma Sea Lion V4 27B It | 128K | $0.35 | $0.56 | cloudflare/@cf/aisingapore/gemma-sea-lion-v4-27b-it | |
| Llama Guard 3 8B | Meta | 131K | $0.48 | $0.03 | cloudflare/@cf/meta/llama-guard-3-8b |
| DeepSeek R1 Distill Qwen 32B | DeepSeek | 131K | $0.50 | $4.88 | cloudflare/@cf/deepseek-ai/deepseek-r1-distill-qwen-32b |
| Nemotron 3 120B A12b | NVIDIA | 256K | $0.50 | $1.50 | cloudflare/@cf/nvidia/nemotron-3-120b-a12b |
| Qwen2 5 Coder 32B Instruct | Alibaba | 33K | $0.66 | $1.00 | cloudflare/@cf/qwen/qwen2.5-coder-32b-instruct |
| QwQ 32B | Alibaba | 131K | $0.66 | $1.00 | cloudflare/@cf/qwen/qwq-32b |
| Kimi K2.7 Code | Moonshot | 262K | $0.95 | $4.00 | cloudflare/@cf/moonshotai/kimi-k2.7-code |
| Kimi K2.6 | Moonshot | 262K | $0.95 | $4.00 | cloudflare/@cf/moonshotai/kimi-k2.6 |
| GLM 5.2 | Zhipu | 1M | $1.40 | $4.40 | cloudflare/@cf/zai-org/glm-5.2 |
| Codellama 7B Instruct Awq | Meta | 4K | $1.92 | $1.92 | cloudflare/@hf/thebloke/codellama-7b-instruct-awq |
| Mistral 7B Instruct V0 1 | Mistral | 16K | $1.92 | $1.92 | cloudflare/@cf/mistral/mistral-7b-instruct-v0.1 |
| Llama 2 7B Chat Int8 | Meta | 2K | $1.92 | $1.92 | cloudflare/@cf/meta/llama-2-7b-chat-int8 |
| Llama 2 7B Chat FP16 | Meta | 3K | $1.92 | $1.92 | cloudflare/@cf/meta/llama-2-7b-chat-fp16 |
Pricing verified against LiteLLM + OpenRouter + provider pages · refreshed daily
Every weekday
AI moves fast. Here's your debrief.
News, analysis, tools, and more.
For people who build with AI