Allen Institute for AI (Ai2), open weights

Olmo 2 0325 32b Instruct

Olmo 2 0325 32b Instruct by Allen Institute for AI (Ai2) ranks 254th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.7. Its strongest category is reasoning, where it ranks 175th.

Last verified

Specifications

Noometry rank
#254 of 354
Index score
32.7
Evidence
Confirmed 16 results
Released
Unknown
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Olmo 2 0325 32b Instruct category scores
  1. Coding 35.2
  2. Reasoning 23.6
  3. Math 26.8
  4. Knowledge 19.5
  5. Multilingual 34.8
  6. Instruction Following 61.5
  7. Long Context 36.2
  8. Writing & Preference 42.1
Olmo 2 0325 32b Instruct category ranks
CategoryScoreRankResults
Coding35.2#2271
Reasoning23.6#1751
Math26.8#2552
Knowledge19.5#2792
Multilingual34.8#2481
Instruction Following61.5#2442
Long Context36.2#2341
Writing & Preference42.1#2364

Strengths and weaknesses

Categories where Olmo 2 0325 32b Instruct places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Olmo 2 0325 32b Instruct: strongest categories
CategoryScorevs medianRank
Reasoning23.6−0.0#175 of 350, top 50%
Coding35.2−3.5#227 of 340, top 67%
Writing & Preference42.1−11.6#236 of 312, top 76%

Weakest categories

Olmo 2 0325 32b Instruct: weakest categories
CategoryScorevs medianRank
Knowledge19.5−17.8#279 of 314, top 89%
Multilingual34.8−12.6#248 of 297, top 84%
Instruction Following61.5−9.7#244 of 305, top 80%

Closest competitors

The models ranked just above and below Olmo 2 0325 32b Instruct. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Olmo 2 0325 32b Instruct
ModelRankScoreBlended $/MSpeed
Phi 3 Medium 4k Instruct#25033.0——Compare
Tulu 3 (Tülu 3) 70B#25133.0——Compare
DeepSeek-R1-Distill-Qwen-14B#25232.7——Compare
Qwen1.5-14B#25332.7——Compare
gpt-oss-20b#25532.5$0.03696Compare
Laguna M.1#25632.5——Compare
Command R+#25732.4$4.38—Compare
Granite 3.1 8b Instruct#25832.4——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Olmo 2 0325 32b Instruct Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1210#237 of 294, top 81%LMArena2026-10-08

Reasoning

Olmo 2 0325 32b Instruct Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1208#233 of 297, top 79%LMArena2026-10-08

Math

Olmo 2 0325 32b Instruct Math benchmark results
BenchmarkScorePositionSettingSourceDate
Omni-MATH16.1%#53 of 57, top 93%HELM Capabilities
LMArena Math1208#229 of 285, top 81%LMArena2026-10-08

Knowledge

Olmo 2 0325 32b Instruct Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
MMLU-Pro41.4%#53 of 58, top 92%HELM Capabilities
GPQA (HELM)28.7%#56 of 57, top 99%HELM Capabilities

Multilingual

Olmo 2 0325 32b Instruct Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1160#248 of 297, top 84%LMArena2026-10-08
LMArena Chinese1192#235 of 285, top 83%LMArena2026-10-08
LMArena Russian1187#240 of 283, top 85%LMArena2026-10-08

Instruction Following

Olmo 2 0325 32b Instruct Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
IFEval78%#47 of 57, top 83%HELM Capabilities
LMArena Instruction Following1186#241 of 298, top 81%LMArena2026-10-08

Long Context

Olmo 2 0325 32b Instruct Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1194#243 of 291, top 84%LMArena2026-10-08

Writing & Preference

Olmo 2 0325 32b Instruct Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1218#239 of 297, top 81%LMArena2026-10-08
LMArena Creative Writing1199#234 of 295, top 80%LMArena2026-10-08
WildBench73.4%#49 of 57, top 86%HELM Capabilities
LMArena Multi-Turn1221#232 of 295, top 79%LMArena2026-10-08

Compare Olmo 2 0325 32b Instruct

Other Allen Institute for AI (Ai2) models

Frequently asked questions

How good is Olmo 2 0325 32b Instruct?

Olmo 2 0325 32b Instruct by Allen Institute for AI (Ai2) ranks 254th of 354 ranked models on the Noometry Index as of October 2026, with a score of 32.7. Its strongest category is reasoning, where it ranks 175th.

Is Olmo 2 0325 32b Instruct open source?

Yes. Olmo 2 0325 32b Instruct's weights are downloadable; check the license for commercial terms.

What are Olmo 2 0325 32b Instruct's strengths and weaknesses?

Relative to other ranked models, Olmo 2 0325 32b Instruct places best in reasoning, coding, writing & preference and lowest in knowledge, multilingual, instruction following.

What is Olmo 2 0325 32b Instruct best at?

Its best category is reasoning, where it ranks 175th on Noometry.