Allen Institute for AI (Ai2), open weights

Olmo 3 32b Think

Olmo 3 32b Think by Allen Institute for AI (Ai2) ranks 183rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.7. Its strongest category is reasoning, where it ranks 140th.

Last verified

Specifications

Noometry rank
#183 of 354
Index score
38.7
Evidence
Confirmed 14 results
Released
Unknown
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Olmo 3 32b Think category scores
  1. Coding 38.6
  2. Reasoning 25.9
  3. Math 36.5
  4. Knowledge 35.0
  5. Multilingual 41.2
  6. Instruction Following 67.2
  7. Long Context 39.4
  8. Writing & Preference 49.1
Olmo 3 32b Think category ranks
CategoryScoreRankResults
Coding38.6#1721
Reasoning25.9#1401
Math36.5#1651
Knowledge35.0#1901
Multilingual41.2#2101
Instruction Following67.2#1981
Long Context39.4#1821
Writing & Preference49.1#1933

Strengths and weaknesses

Categories where Olmo 3 32b Think places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Olmo 3 32b Think: strongest categories
CategoryScorevs medianRank
Reasoning25.9+2.3#140 of 350, top 40%
Math36.5−0.0#165 of 327, top 51%
Coding38.6−0.1#172 of 340, top 51%

Weakest categories

Olmo 3 32b Think: weakest categories
CategoryScorevs medianRank
Multilingual41.2−6.2#210 of 297, top 71%
Instruction Following67.2−4.0#198 of 305, top 65%
Writing & Preference49.1−4.7#193 of 312, top 62%

Closest competitors

The models ranked just above and below Olmo 3 32b Think. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Olmo 3 32b Think
ModelRankScoreBlended $/MSpeed
Qwen3-30B-A3B#17938.9$0.2142Compare
GLM-4.7-Flash#18038.8$0.15—Compare
Qwen2.5 Plus 1127#18138.8——Compare
Qwen3.6 Flash#18238.8$0.42—Compare
Hunyuan Large 2025 02 10#18438.6——Compare
Trinity Large Thinking#18538.6$0.39—Compare
GPT-5.1-Codex#18638.6$3.44—Compare
Sonar#18738.5$1—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Olmo 3 32b Think Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Coding1319#186 of 294, top 64%LMArena2026-10-08

Reasoning

Olmo 3 32b Think Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Hard Prompts1302#188 of 297, top 64%LMArena2026-10-08

Math

Olmo 3 32b Think Math benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Math1316#173 of 285, top 61%LMArena2026-10-08

Knowledge

Olmo 3 32b Think Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1273#186 of 273, top 69%LMArena2026-10-08

Multilingual

Olmo 3 32b Think Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1255#210 of 297, top 71%LMArena2026-10-08
LMArena Chinese1300#188 of 285, top 66%LMArena2026-10-08
LMArena French1291#165 of 223, top 74%LMArena2026-10-08
LMArena German1290#153 of 231, top 67%LMArena2026-10-08
LMArena Russian1254#209 of 283, top 74%LMArena2026-10-08

Instruction Following

Olmo 3 32b Think Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1275#192 of 298, top 65%LMArena2026-10-08

Long Context

Olmo 3 32b Think Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1296#193 of 291, top 67%LMArena2026-10-08

Writing & Preference

Olmo 3 32b Think Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1300#191 of 297, top 65%LMArena2026-10-08
LMArena Creative Writing1256#200 of 295, top 68%LMArena2026-10-08
LMArena Multi-Turn1290#195 of 295, top 67%LMArena2026-10-08

Compare Olmo 3 32b Think

Other Allen Institute for AI (Ai2) models

Frequently asked questions

How good is Olmo 3 32b Think?

Olmo 3 32b Think by Allen Institute for AI (Ai2) ranks 183rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.7. Its strongest category is reasoning, where it ranks 140th.

Is Olmo 3 32b Think open source?

Yes. Olmo 3 32b Think's weights are downloadable; check the license for commercial terms.

What are Olmo 3 32b Think's strengths and weaknesses?

Relative to other ranked models, Olmo 3 32b Think places best in reasoning, math, coding and lowest in multilingual, instruction following, writing & preference.

What is Olmo 3 32b Think best at?

Its best category is reasoning, where it ranks 140th on Noometry.