Skip to content
llmwaves

Inference provider

OpenInference

2 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.

Models served

2

Median output speed

160 t/s

Median latency

444ms

Median price / 1M

$0.135

Catalogue

Strongest and fastest on this provider

What the catalogue looks like at the top end, on the two axes that usually decide the choice.

Highest Intelligence Index

Among models this provider serves

View as table
ModelIntelligence
V4 Flash 0423 (DeepSeek)51.8

Fastest output

Median tokens per second

View as table
ModelTokens/s
Gemma 4 31B (Google)232
V4 Flash 0423 (DeepSeek)87

Full catalogue

All 2 models on OpenInference

Filter by creator
2 of 2 models
V4 Flash 0423
DeepSeekopen
51.8$0.02
Gemma 4 31B
Googleopen
$0.03