Alibaba (Qwen), open weights

Qwen2.5-Coder-32B

Qwen2.5-Coder-32B by Alibaba (Qwen) ranks 245th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.4. Its strongest category is knowledge, where it ranks 203rd. API pricing starts at $0.66 per million input tokens and $1 per million output tokens, with a 33K-token context window.

Last verified

Specifications

Noometry rank
#245 of 354
Index score
33.4
Evidence
Confirmed 31 results
Released
September 18, 2024
Weights
Open weights
Reasoning
Unknown
Context window
33K
Max output
29K
Input price
$0.66 / M
Output price
$1 / M
Blended price
$0.74 / M
Output speed
Not measured
Value
#106 of 219
Knowledge cutoff
Unknown
Input
text

Category scores

Each category score combines every public result we have in that category.

Qwen2.5-Coder-32B category scores
  1. Coding 22.6
  2. Reasoning 21.2
  3. Math 33.3
  4. Knowledge 33.4
  5. Multilingual 37.8
  6. Instruction Following 61.4
  7. Long Context 38.0
  8. Writing & Preference 41.6
Qwen2.5-Coder-32B category ranks
CategoryScoreRankResults
Coding22.6#3336
Reasoning21.2#2253
Math33.3#2042
Knowledge33.4#2031
Multilingual37.8#2351
Instruction Following61.4#2452
Long Context38.0#2081
Writing & Preference41.6#2404

Strengths and weaknesses

Categories where Qwen2.5-Coder-32B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen2.5-Coder-32B: strongest categories
CategoryScorevs medianRank
Math33.3−3.3#204 of 327, top 63%
Reasoning21.2−2.4#225 of 350, top 65%
Knowledge33.4−3.9#203 of 314, top 65%

Weakest categories

Qwen2.5-Coder-32B: weakest categories
CategoryScorevs medianRank
Coding22.6−16.1#333 of 340, top 98%
Instruction Following61.4−9.9#245 of 305, top 81%
Multilingual37.8−9.6#235 of 297, top 80%

Closest competitors

The models ranked just above and below Qwen2.5-Coder-32B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen2.5-Coder-32B
ModelRankScoreBlended $/MSpeed
GPT-5 Nano#24133.5$0.144Compare
Mercury 2.5#24233.5$0.0675—Compare
Mistral Small#24333.4$0.26120Compare
Nova 2.0 Pro Preview#24433.4——Compare
Gemini 1.5 Flash (May 2024)#24633.2——Compare
Granite 3.1 2b Instruct#24733.2——Compare
Gemma 2 2b IT#24833.1——Compare
Wizardlm 70b#24933.0——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen2.5-Coder-32B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
SWE-bench Verified (bash only)9%#39 of 39, top 100%SWE-bench2025-08-03
Aider Polyglot16.4%#37 of 44, top 85%Epoch AI
BigCodeBench Instruct49%#4 of 64, top 7%BigCodeBench2024-09-19
LiveBench Coding56.9%#14 of 39, top 36%Epoch AI
LMArena Coding1276#212 of 294, top 73%LMArena2026-10-08
BigCodeBench Complete58%#10 of 66, top 16%BigCodeBench2024-09-19
HumanEval+87.2%#4 of 45, top 9%EvalPlus
MBPP+77%#3 of 38, top 8%EvalPlus

Reasoning

Qwen2.5-Coder-32B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Reasoning42.1%#26 of 39, top 67%Epoch AI
LMArena Hard Prompts1251#219 of 297, top 74%LMArena2026-10-08
LiveBench Data Analysis49.9%#24 of 39, top 62%Epoch AI
Epoch Capabilities Index119.49#170 of 213, top 80%Epoch AI2024-09-18
HellaSwag83%#11 of 29, top 38%Epoch AI
LiveBench46.2%#24 of 39, top 62%Epoch AI
WinoGrande80.8%#14 of 43, top 33%Epoch AI

Math

Qwen2.5-Coder-32B Math benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Math46.6%#21 of 39, top 54%Epoch AI
LMArena Math1251#213 of 285, top 75%LMArena2026-10-08
GSM8K91.1%Epoch AI
GSM8K93%#2 of 38, top 6%Epoch AI

Knowledge

Qwen2.5-Coder-32B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1221#214 of 273, top 79%LMArena2026-10-08
ARC (AI2) Challenge70.5%#18 of 39, top 47%Epoch AI
MMLU79.1%#22 of 81, top 28%Epoch AI

Multilingual

Qwen2.5-Coder-32B Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1205#235 of 297, top 80%LMArena2026-10-08
LMArena Chinese1222#222 of 285, top 78%LMArena2026-10-08
LMArena Russian1228#222 of 283, top 79%LMArena2026-10-08

Instruction Following

Qwen2.5-Coder-32B Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Instruction Following58.7%#28 of 39, top 72%Epoch AI
LMArena Instruction Following1223#228 of 298, top 77%LMArena2026-10-08

Long Context

Qwen2.5-Coder-32B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1251#221 of 291, top 76%LMArena2026-10-08

Writing & Preference

Qwen2.5-Coder-32B Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1230#230 of 297, top 78%LMArena2026-10-08
LMArena Creative Writing1174#243 of 295, top 83%LMArena2026-10-08
LMArena Multi-Turn1222#231 of 295, top 79%LMArena2026-10-08
LiveBench Language23.3%#34 of 39, top 88%Epoch AI

API pricing by provider

Qwen2.5-Coder-32B API prices
RouteInput $/MOutput $/MCached input $/MChecked
openrouter$0.66$1—2026-10-10

Compare Qwen2.5-Coder-32B

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen2.5-Coder-32B?

Qwen2.5-Coder-32B by Alibaba (Qwen) ranks 245th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.4. Its strongest category is knowledge, where it ranks 203rd. API pricing starts at $0.66 per million input tokens and $1 per million output tokens, with a 33K-token context window.

How much does Qwen2.5-Coder-32B cost?

Qwen2.5-Coder-32B costs $0.66 per million input tokens and $1 per million output tokens on openrouter.

What is Qwen2.5-Coder-32B's context window?

Qwen2.5-Coder-32B accepts up to 33K tokens of input and can write up to 29K tokens in one response.

Is Qwen2.5-Coder-32B open source?

Yes. Qwen2.5-Coder-32B's weights are downloadable from Hugging Face (Qwen/Qwen2.5-Coder-32B-Instruct); check the license for commercial terms.

What are Qwen2.5-Coder-32B's strengths and weaknesses?

Relative to other ranked models, Qwen2.5-Coder-32B places best in math, reasoning, knowledge and lowest in coding, instruction following, multilingual.

What is Qwen2.5-Coder-32B best at?

Its best category is knowledge, where it ranks 203rd on Noometry.