Compare
Four models, one page
Pick the models you are actually choosing between. Every row is scaled within itself, so a bar means something next to its neighbours and nothing across rows.
| Specification | Nemotron 3 Nano 30B A3BNVIDIA | Claude Opus 5Anthropic |
|---|---|---|
| Released | Dec 14, 2025 | Jul 24, 2026 |
| Context window | 262K | 1M |
| Max output | 262K | 128K |
| Input / 1M | $0.05 | $5 |
| Output / 1M | $0.2 | $25 |
| Cost per task | $0.02 | $2.25 |
| Arena Elo | — | 1,393 |
| Serving providers | 3 | 5 |
| Parameters | 31.6B | Undisclosed |
| Licence | other | Proprietary |
| Capabilities |
|
|
Metrics side by side
Each row is scaled to the largest value in that row — bars compare within a row, never across rows
- Nemotron 3 Nano 30B A3B
- Claude Opus 5
Intelligence Index
Nemotron 3 Nano 30B A3B7.2Claude Opus 563.1Coding Index
Nemotron 3 Nano 30B A3Bnot measuredClaude Opus 578.0Agentic Index
Nemotron 3 Nano 30B A3B2.0Claude Opus 559.2Output speed
Nemotron 3 Nano 30B A3B138 t/sClaude Opus 587 t/sContext window
Nemotron 3 Nano 30B A3B262KClaude Opus 51MLatency · lower is better
Nemotron 3 Nano 30B A3B353msClaude Opus 51.12sBlended price / 1M · lower is better
Nemotron 3 Nano 30B A3B$0.088Claude Opus 5$10Cost per task · lower is better
Nemotron 3 Nano 30B A3B$0.02Claude Opus 5$2.25
Rows marked “lower is better” still draw a longer bar for a larger number — read the value, not just the length. Arena Elo is in the specification table above instead: it has no meaningful zero, so a bar would flatten the gaps.
View as table
| Metric | Nemotron 3 Nano 30B A3B | Claude Opus 5 |
|---|---|---|
| Intelligence Index | 7.2 | 63.1 |
| Coding Index | — | 78.0 |
| Agentic Index | 2.0 | 59.2 |
| Output speed | 138 t/s | 87 t/s |
| Context window | 262K | 1M |
| Latency · lower is better | 353ms | 1.12s |
| Blended price / 1M · lower is better | $0.088 | $10 |
| Cost per task · lower is better | $0.02 | $2.25 |
Evaluation scores
Percentage correct on a common 0–100% scale
- Nemotron 3 Nano 30B A3B
- Claude Opus 5
GPQA Diamond
Nemotron 3 Nano 30B A3B39.9%Claude Opus 593.2%Humanity's Last Exam
Nemotron 3 Nano 30B A3B4.6%Claude Opus 554.9%SciCode
Nemotron 3 Nano 30B A3B23.0%Claude Opus 555.7%τ²-bench
Nemotron 3 Nano 30B A3B25.4%Claude Opus 542.1%Terminal-Bench Hard
Nemotron 3 Nano 30B A3B12.1%Claude Opus 5not measuredLiveCodeBench
Nemotron 3 Nano 30B A3B36.0%Claude Opus 5not measuredAA-LCR long context
Nemotron 3 Nano 30B A3B8.7%Claude Opus 575.7%AIME 2025
Nemotron 3 Nano 30B A3B13.3%Claude Opus 5not measured
A missing bar means that evaluation was not run for that model — it is not a zero.
View as table
| Evaluation | Nemotron 3 Nano 30B A3B | Claude Opus 5 |
|---|---|---|
| GPQA Diamond | 39.9% | 93.2% |
| Humanity's Last Exam | 4.6% | 54.9% |
| SciCode | 23.0% | 55.7% |
| τ²-bench | 25.4% | 42.1% |
| Terminal-Bench Hard | 12.1% | — |
| LiveCodeBench | 36.0% | — |
| AA-LCR long context | 8.7% | 75.7% |
| AIME 2025 | 13.3% | — |
Estimated from list pricing: 50K input tokens plus 80K output tokens for reasoning models (25K for non-reasoning).