DeepSeek, open weights
DeepSeek-R1-Distill-Llama-70B
DeepSeek-R1-Distill-Llama-70B by DeepSeek ranks 195th of 354 ranked models on the Noometry Index as of October 2026, with a score of 37.8. Its strongest category is reasoning, where it ranks 156th.
Last verified
Specifications
- Noometry rank
- #195 of 354
- Index score
- 37.8
- Evidence
- Confirmed 13 results
- Provider
DeepSeek
- Released
- January 20, 2025
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- 18 tokens/s Kagi
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 36.8
- Reasoning 24.9
- Math 36.0
- Knowledge 30.7
- Instruction Following 68.2
- Writing & Preference 49.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 36.8 | #202 | 3 |
| Reasoning | 24.9 | #156 | 3 |
| Math | 36.0 | #176 | 3 |
| Knowledge | 30.7 | #225 | 1 |
| Instruction Following | 68.2 | #190 | 1 |
| Writing & Preference | 49.0 | #194 | 1 |
Strengths and weaknesses
Categories where DeepSeek-R1-Distill-Llama-70B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 30.7 | −6.7 | #225 of 314, top 72% |
| Instruction Following | 68.2 | −3.1 | #190 of 305, top 63% |
| Writing & Preference | 49.0 | −4.8 | #194 of 312, top 63% |
Closest competitors
The models ranked just above and below DeepSeek-R1-Distill-Llama-70B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Olmo 3.1 32b Think | #191 | 37.9 | — | — | Compare |
| GPT-5-Codex | #192 | 37.9 | $3.44 | 95 | Compare |
| Hunyuan Standard 2025 02 10 | #193 | 37.9 | — | — | Compare |
| Gemini 2.0 Flash-Lite | #194 | 37.8 | — | — | Compare |
| MiniMax-M2.7 | #196 | 37.7 | $0.52 | — | Compare |
| Gemini Advanced 0514 | #197 | 37.7 | — | — | Compare |
| Grok 2 Mini 2024 08 13 | #198 | 37.7 | — | — | Compare |
| Mercury | #199 | 37.6 | — | 35 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| BigCodeBench Instruct | 35.3% | #45 of 64, top 71% | BigCodeBench | 2025-01-20 | |
| LiveBench Coding | 51.6% | #15 of 39, top 39% | Epoch AI | ||
| BigCodeBench Complete | 49.9% | #35 of 66, top 54% | BigCodeBench | 2025-01-20 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Kagi LLM Benchmark | 52.3% | #57 of 99, top 58% | Kagi LLM Benchmark | ||
| LiveBench Reasoning | 67.6% | #11 of 39, top 29% | Epoch AI | ||
| LiveBench Data Analysis | 55.9% | #17 of 39, top 44% | Epoch AI | ||
| LiveBench | 54.5% | #16 of 39, top 42% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| OTIS Mock AIME 2024-2025 | 51.4% | #112 of 173, top 65% | Epoch AI | 2025-03-07 | |
| LiveBench Math | 58.1% | #15 of 39, top 39% | Epoch AI | ||
| MATH Level 5 | 89.9% | #15 of 79, top 19% | Epoch AI | 2025-03-10 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 55.7% | #121 of 186, top 66% | Epoch AI | 2025-03-10 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Instruction Following | 69.9% | #17 of 39, top 44% | Epoch AI |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Language | 23.8% | #33 of 39, top 85% | Epoch AI |
Compare DeepSeek-R1-Distill-Llama-70B
- DeepSeek-R1-Distill-Llama-70B vs Gemini 2.0 Flash-Lite
- DeepSeek-R1-Distill-Llama-70B vs MiniMax-M2.7
- DeepSeek-R1-Distill-Llama-70B vs Hunyuan Standard 2025 02 10
- DeepSeek-R1-Distill-Llama-70B vs Gemini Advanced 0514
- DeepSeek-R1-Distill-Llama-70B vs GPT-5-Codex
- DeepSeek-R1-Distill-Llama-70B vs Grok 2 Mini 2024 08 13
- DeepSeek-R1-Distill-Llama-70B vs GPT-6 Astra
- DeepSeek-R1-Distill-Llama-70B vs Claude Fable 5.1
- DeepSeek-R1-Distill-Llama-70B vs Gemini 3.8 Flash
- DeepSeek-R1-Distill-Llama-70B vs Kimi K3
- DeepSeek-R1-Distill-Llama-70B vs Grok 4.6
- DeepSeek-R1-Distill-Llama-70B vs Qwen3.8 Max
- DeepSeek-R1-Distill-Llama-70B vs GLM-5.3
- DeepSeek-R1-Distill-Llama-70B vs Muse Spark 1.3
Other DeepSeek models
Frequently asked questions
How good is DeepSeek-R1-Distill-Llama-70B?
DeepSeek-R1-Distill-Llama-70B by DeepSeek ranks 195th of 354 ranked models on the Noometry Index as of October 2026, with a score of 37.8. Its strongest category is reasoning, where it ranks 156th.
Is DeepSeek-R1-Distill-Llama-70B open source?
Yes. DeepSeek-R1-Distill-Llama-70B's weights are downloadable; check the license for commercial terms.
How fast is DeepSeek-R1-Distill-Llama-70B?
DeepSeek-R1-Distill-Llama-70B generated about 18 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.
What are DeepSeek-R1-Distill-Llama-70B's strengths and weaknesses?
Relative to other ranked models, DeepSeek-R1-Distill-Llama-70B places best in reasoning, math, coding and lowest in knowledge, instruction following, writing & preference.
What is DeepSeek-R1-Distill-Llama-70B best at?
Its best category is reasoning, where it ranks 156th on Noometry.