Inference provider
Morph
9 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.
Models served
9
Median output speed
142 t/s
Median latency
476ms
Median price / 1M
$0.9
Catalogue
Strongest and fastest on this provider
What the catalogue looks like at the top end, on the two axes that usually decide the choice.
Highest Intelligence Index
Among models this provider serves
- Kimi K3Moonshot AI59.7
- GLM 5.2Z.ai52.6
- V4 Flash 0731DeepSeek51.8
- V4 Flash 0423DeepSeek51.8
- MiniMax M3MiniMax45.4
- Qwen3.6 27BQwen37.7
View as table
| Model | Intelligence |
|---|---|
| Kimi K3 (Moonshot AI) | 59.7 |
| GLM 5.2 (Z.ai) | 52.6 |
| V4 Flash 0731 (DeepSeek) | 51.8 |
| V4 Flash 0423 (DeepSeek) | 51.8 |
| MiniMax M3 (MiniMax) | 45.4 |
| Qwen3.6 27B (Qwen) | 37.7 |
Fastest output
Median tokens per second
- V3 FastMorph3,502 t/s
- V3 LargeMorph3,129 t/s
- Gemma 4 31BGoogle232 t/s
- V4 Flash 0731DeepSeek148 t/s
- GLM 5.2Z.ai142 t/s
- V4 Flash 0423DeepSeek87 t/s
- MiniMax M3MiniMax84 t/s
- Qwen3.6 27BQwen77 t/s
- Kimi K3Moonshot AI71 t/s
View as table
| Model | Tokens/s |
|---|---|
| V3 Fast (Morph) | 3502 |
| V3 Large (Morph) | 3129 |
| Gemma 4 31B (Google) | 232 |
| V4 Flash 0731 (DeepSeek) | 148 |
| GLM 5.2 (Z.ai) | 142 |
| V4 Flash 0423 (DeepSeek) | 87 |
| MiniMax M3 (MiniMax) | 84 |
| Qwen3.6 27B (Qwen) | 77 |
| Kimi K3 (Moonshot AI) | 71 |
Full catalogue
All 9 models on Morph
Filter by creator
9 of 9 models| Kimi K3 Moonshot AIopen | 59.7 | $1.35 |
| GLM 5.2 Z.aiopen | 52.6 | $0.23 |
| V4 Flash 0731 DeepSeekopen | 51.8 | $0.02 |
| V4 Flash 0423 DeepSeekopen | 51.8 | $0.02 |
| MiniMax M3 MiniMaxopen | 45.4 | $0.11 |
| Qwen3.6 27B Qwenopen | 37.7 | $0.32 |
| Gemma 4 31B Googleopen | — | $0.03 |
| V3 Large Morph | — | $0.09 |
| V3 Fast Morph | — | $0.07 |