Skip to content
llmwaves

Inference provider

GMICloud

14 models served. Speed, latency and price below are medians across this provider's catalogue, not per-endpoint measurements.

Models served

14

Median output speed

81 t/s

Median latency

457ms

Median price / 1M

$0.499

Catalogue

Strongest and fastest on this provider

What the catalogue looks like at the top end, on the two axes that usually decide the choice.

Highest Intelligence Index

Among models this provider serves

View as table
ModelIntelligence
GLM 5.2 (Z.ai)52.6
V4 Flash 0731 (DeepSeek)51.8
V4 Flash 0423 (DeepSeek)51.8
MiniMax M3 (MiniMax)45.4
V4 Pro (DeepSeek)45.3
MiMo-V2.5-Pro (Xiaomi)42.9
Hy3 (Tencent)42.2
GLM 5.1 (Z.ai)41.0
GLM 5 (Z.ai)40.6
M2.7 (MiniMax)38.9

Fastest output

Median tokens per second

View as table
ModelTokens/s
M2.7 (MiniMax)307
V4 Flash 0731 (DeepSeek)148
GLM 5.2 (Z.ai)142
Qwen3.5 397B A17B (Qwen)116
V3.2 (DeepSeek)100
V4 Flash 0423 (DeepSeek)87
MiniMax M3 (MiniMax)84
Hy3 (Tencent)78
MiMo-V2.5-Pro (Xiaomi)73
V4 Pro (DeepSeek)71

Full catalogue

All 14 models on GMICloud

Filter by creator
14 of 14 models
GLM 5.2
Z.aiopen
52.6$0.23
V4 Flash 0731
DeepSeekopen
51.8$0.02
V4 Flash 0423
DeepSeekopen
51.8$0.02
MiniMax M3
MiniMaxopen
45.4$0.11
V4 Pro
DeepSeekopen
45.3$0.09
MiMo-V2.5-Pro
Xiaomiopen
42.9$0.09
Hy3
Tencentopen
42.2$0.05
GLM 5.1
Z.aiopen
41.0$0.29
GLM 5
Z.aiopen
40.6$0.25
M2.7
MiniMaxopen
38.9$0.10
MiMo-V2.5
Xiaomiopen
38.0$0.03
Hy3 preview
Tencentopen
34.4$0.02
Qwen3.5 397B A17B
Qwenopen
34.3$0.21
V3.2
DeepSeekopen
25.1$0.05