DeepSeek, open weights

DeepSeek-V3.1-Terminus

DeepSeek-V3.1-Terminus by DeepSeek ranks 97th of 354 ranked models on the Noometry Index as of October 2026, with a score of 43.1. Its strongest category is multilingual, where it ranks 92nd. API pricing starts at $0.27 per million input tokens and $1 per million output tokens, with a 164K-token context window.

Last verified

Specifications

Noometry rank
#97 of 354
Index score
43.1
Evidence
Confirmed 16 results
Provider
DeepSeek
Released
September 22, 2025
Weights
Open weights
Reasoning
Yes
Context window
164K
Max output
147K
Input price
$0.27 / M
Output price
$1 / M
Blended price
$0.45 / M
Output speed
30 tokens/s Kagi
Value
#64 of 219
Knowledge cutoff
Unknown
Input
text

Category scores

Each category score combines every public result we have in that category.

DeepSeek-V3.1-Terminus category scores
  1. Coding 42.0
  2. Reasoning 26.4
  3. Math 38.5
  4. Multilingual 52.1
  5. Instruction Following 74.0
  6. Long Context 43.4
  7. Writing & Preference 61.0
DeepSeek-V3.1-Terminus category ranks
CategoryScoreRankResults
Coding42.0#1132
Reasoning26.4#1335
Math38.5#1371
Multilingual52.1#921
Instruction Following74.0#1061
Long Context43.4#971
Writing & Preference61.0#923

Strengths and weaknesses

Categories where DeepSeek-V3.1-Terminus places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

DeepSeek-V3.1-Terminus: strongest categories
CategoryScorevs medianRank
Writing & Preference61.0+7.2#92 of 312, top 30%
Multilingual52.1+4.7#92 of 297, top 31%
Long Context43.4+2.5#97 of 296, top 33%

Weakest categories

DeepSeek-V3.1-Terminus: weakest categories
CategoryScorevs medianRank
Math38.5+1.9#137 of 327, top 42%
Reasoning26.4+2.8#133 of 350, top 38%
Instruction Following74.0+2.8#106 of 305, top 35%

Closest competitors

The models ranked just above and below DeepSeek-V3.1-Terminus. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to DeepSeek-V3.1-Terminus
ModelRankScoreBlended $/MSpeed
MiMo-V2.5#9343.4$0.18—Compare
Kimi K2.7 Code#9443.3$1.71—Compare
Qwen3-VL 235B-A22B#9543.2$1.22—Compare
Seed 2.0 Pro#9643.2$1.13—Compare
Hunyuan Vision 1.5#9843.1——Compare
Mistral Large 4#9943.1$1.03—Compare
Claude Opus 4#10043.1$3029Compare
Amazon Nova Experimental Chat 11 10#10143.0——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

DeepSeek-V3.1-Terminus Coding benchmark results
BenchmarkScorePositionSettingSourceDate
SciCode40.6%#75 of 121, top 62%Epoch AI
LMArena Coding1426#115 of 294, top 40%thinkingLMArena2026-10-08
ALE-Bench745.17#66 of 105, top 63%Epoch AI

Reasoning

DeepSeek-V3.1-Terminus Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark57.4%#47 of 99, top 48%Kagi LLM Benchmark
CritPt1.7%#73 of 134, top 55%Epoch AI
LMArena Hard Prompts1426#96 of 297, top 33%thinkingLMArena2026-10-08
DTBench81.3%#69 of 151, top 46%Epoch AI
LMCA28.6%#80 of 125, top 64%Epoch AI

Math

DeepSeek-V3.1-Terminus Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1402#124 of 285, top 44%LMArena2026-10-08

Multilingual

DeepSeek-V3.1-Terminus Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1407#92 of 297, top 31%LMArena2026-10-08
LMArena Russian1436#55 of 283, top 20%thinkingLMArena2026-10-08

Instruction Following

DeepSeek-V3.1-Terminus Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1404#99 of 298, top 34%thinkingLMArena2026-10-08

Long Context

DeepSeek-V3.1-Terminus Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1421#91 of 291, top 32%thinkingLMArena2026-10-08

Writing & Preference

DeepSeek-V3.1-Terminus Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1419#98 of 297, top 33%thinkingLMArena2026-10-08
LMArena Creative Writing1403#73 of 295, top 25%LMArena2026-10-08
LMArena Multi-Turn1411#109 of 295, top 37%thinkingLMArena2026-10-08

API pricing by provider

DeepSeek-V3.1-Terminus API prices
RouteInput $/MOutput $/MCached input $/MChecked
openrouter$0.27$1—2026-10-10

Compare DeepSeek-V3.1-Terminus

Other DeepSeek models

Frequently asked questions

How good is DeepSeek-V3.1-Terminus?

DeepSeek-V3.1-Terminus by DeepSeek ranks 97th of 354 ranked models on the Noometry Index as of October 2026, with a score of 43.1. Its strongest category is multilingual, where it ranks 92nd. API pricing starts at $0.27 per million input tokens and $1 per million output tokens, with a 164K-token context window.

How much does DeepSeek-V3.1-Terminus cost?

DeepSeek-V3.1-Terminus costs $0.27 per million input tokens and $1 per million output tokens on openrouter.

What is DeepSeek-V3.1-Terminus's context window?

DeepSeek-V3.1-Terminus accepts up to 164K tokens of input and can write up to 147K tokens in one response.

Is DeepSeek-V3.1-Terminus open source?

Yes. DeepSeek-V3.1-Terminus's weights are downloadable from Hugging Face (deepseek-ai/DeepSeek-V3.1-Terminus); check the license for commercial terms.

How fast is DeepSeek-V3.1-Terminus?

DeepSeek-V3.1-Terminus generated about 30 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are DeepSeek-V3.1-Terminus's strengths and weaknesses?

Relative to other ranked models, DeepSeek-V3.1-Terminus places best in writing & preference, multilingual, long context and lowest in math, reasoning, instruction following.

What is DeepSeek-V3.1-Terminus best at?

Its best category is multilingual, where it ranks 92nd on Noometry.