Inference provider
Cloudflare
15 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.
Models served
15
Median output speed
89 t/s
Median latency
337ms
Median price / 1M
$0.145
Catalogue
Strongest and fastest on this provider
What the catalogue looks like at the top end, on the two axes that usually decide the choice.
Highest Intelligence Index
Among models this provider serves
- GLM 5.2Z.ai52.6
- V4 Flash 0731DeepSeek51.8
- V4 Flash 0423DeepSeek51.8
- V4 ProDeepSeek45.3
- Kimi K2.6Moonshot AI45.1
- Kimi K2.7 CodeMoonshot AI43.0
- GLM 4.7 FlashZ.ai23.3
- 6.9
View as table
| Model | Intelligence |
|---|---|
| GLM 5.2 (Z.ai) | 52.6 |
| V4 Flash 0731 (DeepSeek) | 51.8 |
| V4 Flash 0423 (DeepSeek) | 51.8 |
| V4 Pro (DeepSeek) | 45.3 |
| Kimi K2.6 (Moonshot AI) | 45.1 |
| Kimi K2.7 Code (Moonshot AI) | 43.0 |
| GLM 4.7 Flash (Z.ai) | 23.3 |
| Qwen2.5 Coder 32B Instruct (Qwen) | 6.9 |
Fastest output
Median tokens per second
- 234 t/s
- V4 Flash 0731DeepSeek148 t/s
- GLM 5.2Z.ai142 t/s
- Kimi K2.7 CodeMoonshot AI140 t/s
- 137 t/s
- 136 t/s
- 119 t/s
- Kimi K2.6Moonshot AI89 t/s
- V4 Flash 0423DeepSeek87 t/s
- V4 ProDeepSeek71 t/s
View as table
| Model | Tokens/s |
|---|---|
| Llama 3.3 70B Instruct (Meta) | 234 |
| V4 Flash 0731 (DeepSeek) | 148 |
| GLM 5.2 (Z.ai) | 142 |
| Kimi K2.7 Code (Moonshot AI) | 140 |
| Llama 3.2 3B Instruct (Meta) | 137 |
| Llama 3.2 1B Instruct (Meta) | 136 |
| Llama 3.1 8B Instruct (Meta) | 119 |
| Kimi K2.6 (Moonshot AI) | 89 |
| V4 Flash 0423 (DeepSeek) | 87 |
| V4 Pro (DeepSeek) | 71 |
Full catalogue
All 15 models on Cloudflare
Filter by creator
15 of 15 models| GLM 5.2 Z.aiopen | 52.6 | $0.23 |
| V4 Flash 0731 DeepSeekopen | 51.8 | $0.02 |
| V4 Flash 0423 DeepSeekopen | 51.8 | $0.02 |
| V4 Pro DeepSeekopen | 45.3 | $0.09 |
| Kimi K2.6 Moonshot AIopen | 45.1 | $0.23 |
| Kimi K2.7 Code Moonshot AIopen | 43.0 | $0.32 |
| GLM 4.7 Flash Z.aiopen | 23.3 | $0.04 |
| Qwen2.5 Coder 32B Instruct Qwenopen | 6.9 | $0.06 |
| Gemma 4 26B A4B Googleopen | — | $0.03 |
| Granite 4.0 Micro IBM Graniteopen | — | $0.004 |
| Mistral Small 3.1 24B Mistral AIopen | — | $0.03 |
| Llama 3.3 70B Instruct Metaopen | — | $0.01 |
| Llama 3.2 1B Instruct Metaopen | — | $0.006 |
| Llama 3.2 3B Instruct Metaopen | — | $0.01 |
| Llama 3.1 8B Instruct Metaopen | — | $0.005 |