Skip to content
llmwaves

Inference provider

Cerebras

3 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.

Models served

3

Median output speed

502 t/s

Median latency

311ms

Median price / 1M

$0.16

Catalogue

Strongest and fastest on this provider

What the catalogue looks like at the top end, on the two axes that usually decide the choice.

Highest Intelligence Index

Among models this provider serves

View as table
ModelIntelligence
GLM 4.7 (Z.ai)34.5
gpt-oss-120b (OpenAI)24.1

Fastest output

Median tokens per second

View as table
ModelTokens/s
gpt-oss-120b (OpenAI)773
GLM 4.7 (Z.ai)502
Gemma 4 31B (Google)232

Full catalogue

All 3 models on Cerebras

Filter by creator
3 of 3 models
GLM 4.7
Z.aiopen
34.5$0.16
gpt-oss-120b
OpenAIopen
24.1$0.02
Gemma 4 31B
Googleopen
$0.03