Inference provider
OpenInference
2 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.
Models served
2
Median output speed
160 t/s
Median latency
444ms
Median price / 1M
$0.135
Catalogue
Strongest and fastest on this provider
What the catalogue looks like at the top end, on the two axes that usually decide the choice.
Highest Intelligence Index
Among models this provider serves
- V4 Flash 0423DeepSeek51.8
View as table
| Model | Intelligence |
|---|---|
| V4 Flash 0423 (DeepSeek) | 51.8 |
Fastest output
Median tokens per second
- Gemma 4 31BGoogle232 t/s
- V4 Flash 0423DeepSeek87 t/s
View as table
| Model | Tokens/s |
|---|---|
| Gemma 4 31B (Google) | 232 |
| V4 Flash 0423 (DeepSeek) | 87 |
Full catalogue
All 2 models on OpenInference
Filter by creator
2 of 2 models| V4 Flash 0423 DeepSeekopen | 51.8 | $0.02 |
| Gemma 4 31B Googleopen | — | $0.03 |