NVIDIA, open weights
Llama 3.1 Nemotron Ultra 253b v1
Llama 3.1 Nemotron Ultra 253b v1 by NVIDIA ranks 213th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.7. Its strongest category is reasoning, where it ranks 134th.
Last verified
Specifications
- Noometry rank
- #213 of 354
- Index score
- 36.7
- Evidence
- Confirmed 11 results
- Provider
NVIDIA
- Released
- Unknown
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 38.4
- Agentic & Tool Use 15.7
- Reasoning 26.3
- Math 37.5
- Multilingual 43.1
- Instruction Following 69.0
- Long Context 39.5
- Writing & Preference 52.2
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 38.4 | #177 | 1 |
| Agentic & Tool Use | 15.7 | #149 | 1 |
| Reasoning | 26.3 | #134 | 1 |
| Math | 37.5 | #152 | 1 |
| Multilingual | 43.1 | #187 | 1 |
| Instruction Following | 69.0 | #178 | 1 |
| Long Context | 39.5 | #177 | 1 |
| Writing & Preference | 52.2 | #175 | 3 |
Strengths and weaknesses
Categories where Llama 3.1 Nemotron Ultra 253b v1 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Agentic & Tool Use | 15.7 | −14.7 | #149 of 154, top 97% |
| Multilingual | 43.1 | −4.3 | #187 of 297, top 63% |
| Long Context | 39.5 | −1.5 | #177 of 296, top 60% |
Closest competitors
The models ranked just above and below Llama 3.1 Nemotron Ultra 253b v1. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Yi-Lightning | #209 | 37.1 | — | — | Compare |
| Qwen Plus | #210 | 37.1 | $0.60 | 37 | Compare |
| Gemini 2.5 Flash-Lite | #211 | 37.0 | $0.18 | 172 | Compare |
| o3-mini | #212 | 36.7 | $1.93 | — | Compare |
| Granite 4.0 H Small | #214 | 36.5 | — | — | Compare |
| Command A | #215 | 36.5 | $4.38 | 28 | Compare |
| Grok Build 0.1 | #216 | 36.4 | $1.25 | — | Compare |
| gpt-oss-120b | #217 | 36.3 | $0.0703 | 55 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1312 | #191 of 294, top 65% | LMArena | 2026-10-08 |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Berkeley Function Calling Leaderboard | 10% | #48 of 49, top 98% | fc | Berkeley Function Calling Leaderboard |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1316 | #181 of 297, top 61% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1360 | #156 of 285, top 55% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1282 | #187 of 297, top 63% | LMArena | 2026-10-08 | |
| LMArena Russian | 1284 | #188 of 283, top 67% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1308 | #171 of 298, top 58% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1299 | #190 of 291, top 66% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1320 | #180 of 297, top 61% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1314 | #159 of 295, top 54% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1317 | #177 of 295, top 60% | LMArena | 2026-10-08 |
Compare Llama 3.1 Nemotron Ultra 253b v1
- Llama 3.1 Nemotron Ultra 253b v1 vs o3-mini
- Llama 3.1 Nemotron Ultra 253b v1 vs Granite 4.0 H Small
- Llama 3.1 Nemotron Ultra 253b v1 vs Gemini 2.5 Flash-Lite
- Llama 3.1 Nemotron Ultra 253b v1 vs Command A
- Llama 3.1 Nemotron Ultra 253b v1 vs Qwen Plus
- Llama 3.1 Nemotron Ultra 253b v1 vs Grok Build 0.1
- Llama 3.1 Nemotron Ultra 253b v1 vs GPT-6 Astra
- Llama 3.1 Nemotron Ultra 253b v1 vs Claude Fable 5.1
- Llama 3.1 Nemotron Ultra 253b v1 vs Gemini 3.8 Flash
- Llama 3.1 Nemotron Ultra 253b v1 vs Kimi K3
- Llama 3.1 Nemotron Ultra 253b v1 vs Grok 4.6
- Llama 3.1 Nemotron Ultra 253b v1 vs Qwen3.8 Max
- Llama 3.1 Nemotron Ultra 253b v1 vs GLM-5.3
- Llama 3.1 Nemotron Ultra 253b v1 vs Muse Spark 1.3
Frequently asked questions
How good is Llama 3.1 Nemotron Ultra 253b v1?
Llama 3.1 Nemotron Ultra 253b v1 by NVIDIA ranks 213th of 354 ranked models on the Noometry Index as of October 2026, with a score of 36.7. Its strongest category is reasoning, where it ranks 134th.
Is Llama 3.1 Nemotron Ultra 253b v1 open source?
Yes. Llama 3.1 Nemotron Ultra 253b v1's weights are downloadable; check the license for commercial terms.
What are Llama 3.1 Nemotron Ultra 253b v1's strengths and weaknesses?
Relative to other ranked models, Llama 3.1 Nemotron Ultra 253b v1 places best in reasoning, math, coding and lowest in agentic & tool use, multilingual, long context.
What is Llama 3.1 Nemotron Ultra 253b v1 best at?
Its best category is reasoning, where it ranks 134th on Noometry.