Models
Every model, one table
Sort by any column, filter by capability or creator, and open a model for its full benchmark profile. Missing values always sink to the bottom rather than pretending to be zero.
Models indexed
323
With benchmark scores
161
Open weights
153
Model creators
51
Highlights
Leaders at a glance
The same three rankings that lead the front page, so you can start from the top of a list rather than the middle of a table.
Intelligence
Intelligence Index · higher is better
- Claude Opus 5Anthropic63.1
- Claude Fable 5Anthropic62.1
- GPT-5.6 SolOpenAI60.9
- Kimi K3Moonshot AI59.7
- Qwen3.8 MaxQwen58.1
- Claude Opus 4.8Anthropic57.3
- Muse Spark 1.2Meta56.8
- GPT-5.6 TerraOpenAI56.6
- GPT-5.5OpenAI56.3
- Grok 4.5xAI55.8
Composite of ten reasoning, knowledge, coding and agentic evaluations.
View as table
| Model | Index |
|---|---|
| Claude Opus 5 (Anthropic) | 63.1 |
| Claude Fable 5 (Anthropic) | 62.1 |
| GPT-5.6 Sol (OpenAI) | 60.9 |
| Kimi K3 (Moonshot AI) | 59.7 |
| Qwen3.8 Max (Qwen) | 58.1 |
| Claude Opus 4.8 (Anthropic) | 57.3 |
| Muse Spark 1.2 (Meta) | 56.8 |
| GPT-5.6 Terra (OpenAI) | 56.6 |
| GPT-5.5 (OpenAI) | 56.3 |
| Grok 4.5 (xAI) | 55.8 |
Speed
Output tokens per second · higher is better
- V3 FastMorph3,502 t/s
- V3 LargeMorph3,129 t/s
- Apply 3Relace3,080 t/s
- gpt-oss-120bOpenAI773 t/s
- GLM 4.7Z.ai502 t/s
- 443 t/s
- Qwen3 32BQwen423 t/s
- gpt-oss-safeguard-20bOpenAI422 t/s
- 337 t/s
- M2.7MiniMax307 t/s
Median output tokens per second across serving providers.
View as table
| Model | Tokens/s |
|---|---|
| V3 Fast (Morph) | 3502 |
| V3 Large (Morph) | 3129 |
| Apply 3 (Relace) | 3080 |
| gpt-oss-120b (OpenAI) | 773 |
| GLM 4.7 (Z.ai) | 502 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) (Google) | 443 |
| Qwen3 32B (Qwen) | 423 |
| gpt-oss-safeguard-20b (OpenAI) | 422 |
| Grok 4.20 Multi-Agent (xAI) | 337 |
| M2.7 (MiniMax) | 307 |
Cost per task
Estimated USD per task · lower is better
- Ling-3.0-flashInclusionAI$0.006
- V4 Flash 0423DeepSeek$0.02
- V4 Flash 0731DeepSeek$0.02
- Hy3 previewTencent$0.02
- MiMo-V2.5Xiaomi$0.03
- KAT-Coder-Pro V2KwaiPilot$0.04
- Hy3Tencent$0.05
- GPT-5.6 LunaOpenAI$0.05
- Ring-2.6-1TInclusionAI$0.05
- M2.5MiniMax$0.08
Estimated from list pricing: 50K input tokens plus 80K output tokens for reasoning models (25K for non-reasoning). Limited to models scoring 30+ on the Intelligence Index, so the ranking compares like with like.
View as table
| Model | USD |
|---|---|
| Ling-3.0-flash (InclusionAI) | $0.006 |
| V4 Flash 0423 (DeepSeek) | $0.02 |
| V4 Flash 0731 (DeepSeek) | $0.02 |
| Hy3 preview (Tencent) | $0.02 |
| MiMo-V2.5 (Xiaomi) | $0.03 |
| KAT-Coder-Pro V2 (KwaiPilot) | $0.04 |
| Hy3 (Tencent) | $0.05 |
| GPT-5.6 Luna (OpenAI) | $0.05 |
| Ring-2.6-1T (InclusionAI) | $0.05 |
| M2.5 (MiniMax) | $0.08 |
Trade-off
What the index costs
Every benchmarked model with a published price, plotted against what it charges.
Intelligence vs price
Blended USD per 1M tokens on a log scale · up and to the right is a better deal
- Proprietary
- Open weights
Price axis is inverted so cheaper sits right. Labelled points are the Pareto frontier — nothing in the index is both smarter and cheaper. Only models with both a published price and an Intelligence Index score appear.
View as table
| Model | Intelligence | Blended / 1M |
|---|---|---|
| Claude Opus 5 (Anthropic) | 63.1 | $10 |
| Claude Fable 5 (Anthropic) | 62.1 | $20 |
| GPT-5.6 Sol (OpenAI) | 60.9 | $11.25 |
| Kimi K3 (Moonshot AI) | 59.7 | $6 |
| Qwen3.8 Max (Qwen) | 58.1 | $3 |
| Claude Opus 4.8 (Anthropic) | 57.3 | $10 |
| Muse Spark 1.2 (Meta) | 56.8 | $2 |
| GPT-5.6 Terra (OpenAI) | 56.6 | $2.25 |
| GPT-5.5 (OpenAI) | 56.3 | $11.25 |
| Grok 4.5 (xAI) | 55.8 | $3 |
| Claude Sonnet 5 (Anthropic) | 55.3 | $4 |
| Claude Opus 4.7 (Anthropic) | 55.0 | $10 |
| Muse Spark 1.1 (Meta) | 53.2 | $2 |
| GPT-5.4 (OpenAI) | 53.1 | $5.63 |
| GLM 5.2 (Z.ai) | 52.6 | $1.18 |
| GPT-5.6 Luna (OpenAI) | 52.3 | $0.225 |
| Gemini 3.5 Flash (Google) | 52.0 | $3.38 |
| V4 Flash 0731 (DeepSeek) | 51.8 | $0.113 |
| V4 Flash 0423 (DeepSeek) | 51.8 | $0.11 |
| Gemini 3.6 Flash (Google) | 51.6 | $3 |
| Gemini 3.1 Pro Preview (Google) | 47.7 | $4.5 |
| Qwen3.7 Max (Qwen) | 46.7 | $2.21 |
| GPT-5.3-Codex (OpenAI) | 45.5 | $4.81 |
| MiniMax M3 (MiniMax) | 45.4 | $0.525 |
| V4 Pro (DeepSeek) | 45.3 | $0.544 |
Full index
All 323 models
Batch-pricing endpoints are folded out of this list so a single model cannot occupy two rows.
| Claude Opus 5 Anthropic | 63.1 | $2.25 |
| Claude Fable 5 Anthropic | 62.1 | $4.50 |
| GPT-5.6 Sol OpenAI | 60.9 | $2.65 |
| Kimi K3 Moonshot AIopen | 59.7 | $1.35 |
| Qwen3.8 Max Qwen | 58.1 | $0.58 |
| Claude Opus 4.8 Anthropic | 57.3 | $2.25 |
| Muse Spark 1.2 Meta | 56.8 | $0.40 |
| GPT-5.6 Terra OpenAI | 56.6 | $0.53 |
| GPT-5.5 OpenAI | 56.3 | $2.65 |
| Grok 4.5 xAI | 55.8 | $0.58 |
| Claude Sonnet 5 Anthropic | 55.3 | $0.90 |
| Claude Opus 4.7 Anthropic | 55.0 | $2.25 |
| Muse Spark 1.1 Meta | 53.2 | $0.40 |
| GPT-5.4 OpenAI | 53.1 | $1.32 |
| GLM 5.2 Z.aiopen | 52.6 | $0.23 |
| GPT-5.6 Luna OpenAI | 52.3 | $0.05 |
| Gemini 3.5 Flash Google | 52.0 | $0.79 |
| V4 Flash 0731 DeepSeekopen | 51.8 | $0.02 |
| V4 Flash 0423 DeepSeekopen | 51.8 | $0.02 |
| Gemini 3.6 Flash Google | 51.6 | $0.68 |
| Gemini 3.1 Pro Preview Google | 47.7 | $1.06 |
| Qwen3.7 Max Qwen | 46.7 | $0.43 |
| GPT-5.3-Codex OpenAI | 45.5 | $1.21 |
| MiniMax M3 MiniMaxopen | 45.4 | $0.11 |
| V4 Pro DeepSeekopen | 45.3 | $0.09 |
| Kimi K2.6 Moonshot AIopen | 45.1 | $0.23 |
| GPT-5.2 OpenAI | 43.3 | $1.21 |
| Kimi K2.7 Code Moonshot AIopen | 43.0 | $0.32 |
| MiMo-V2.5-Pro Xiaomiopen | 42.9 | $0.09 |
| Inkling Thinking Machinesopen | 42.3 | $0.37 |
| Hy3 Tencentopen | 42.2 | $0.05 |
| Nex-N2-Pro Nex AGIopen | 41.7 | $0.09 |
| Inkling Small Thinking Machinesopen | 41.2 | $0.12 |
| GPT-5.2-Codex OpenAI | 41.2 | $1.21 |
| Qwen3.6 Max Preview Qwen | 41.1 | $0.54 |
| GLM 5.1 Z.aiopen | 41.0 | $0.29 |
| GPT-5.4 Mini OpenAI | 40.9 | $0.40 |
| GLM 5 Z.aiopen | 40.6 | $0.25 |
| Qwen3.6 Plus Qwen | 40.5 | $0.17 |
| GPT-5.4 Nano OpenAI | 39.7 | $0.11 |