Skip to content
llmwaves

Inference provider

Cloudflare

15 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.

Models served

15

Median output speed

89 t/s

Median latency

337ms

Median price / 1M

$0.145

Catalogue

Strongest and fastest on this provider

What the catalogue looks like at the top end, on the two axes that usually decide the choice.

Highest Intelligence Index

Among models this provider serves

View as table
ModelIntelligence
GLM 5.2 (Z.ai)52.6
V4 Flash 0731 (DeepSeek)51.8
V4 Flash 0423 (DeepSeek)51.8
V4 Pro (DeepSeek)45.3
Kimi K2.6 (Moonshot AI)45.1
Kimi K2.7 Code (Moonshot AI)43.0
GLM 4.7 Flash (Z.ai)23.3
Qwen2.5 Coder 32B Instruct (Qwen)6.9

Fastest output

Median tokens per second

View as table
ModelTokens/s
Llama 3.3 70B Instruct (Meta)234
V4 Flash 0731 (DeepSeek)148
GLM 5.2 (Z.ai)142
Kimi K2.7 Code (Moonshot AI)140
Llama 3.2 3B Instruct (Meta)137
Llama 3.2 1B Instruct (Meta)136
Llama 3.1 8B Instruct (Meta)119
Kimi K2.6 (Moonshot AI)89
V4 Flash 0423 (DeepSeek)87
V4 Pro (DeepSeek)71

Full catalogue

All 15 models on Cloudflare

Filter by creator
15 of 15 models
GLM 5.2
Z.aiopen
52.6$0.23
V4 Flash 0731
DeepSeekopen
51.8$0.02
V4 Flash 0423
DeepSeekopen
51.8$0.02
V4 Pro
DeepSeekopen
45.3$0.09
Kimi K2.6
Moonshot AIopen
45.1$0.23
Kimi K2.7 Code
Moonshot AIopen
43.0$0.32
GLM 4.7 Flash
Z.aiopen
23.3$0.04
Qwen2.5 Coder 32B Instruct
Qwenopen
6.9$0.06
Gemma 4 26B A4B
Googleopen
$0.03
Granite 4.0 Micro
IBM Graniteopen
$0.004
Mistral Small 3.1 24B
Mistral AIopen
$0.03
Llama 3.3 70B Instruct
Metaopen
$0.01
Llama 3.2 1B Instruct
Metaopen
$0.006
Llama 3.2 3B Instruct
Metaopen
$0.01
Llama 3.1 8B Instruct
Metaopen
$0.005