Allen Institute for AI (Ai2), open weights
Olmo 3 32b Think
Olmo 3 32b Think by Allen Institute for AI (Ai2) ranks 183rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.7. Its strongest category is reasoning, where it ranks 140th.
Last verified
Specifications
- Noometry rank
- #183 of 354
- Index score
- 38.7
- Evidence
- Confirmed 14 results
- Provider
Allen Institute for AI (Ai2)
- Released
- Unknown
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 38.6
- Reasoning 25.9
- Math 36.5
- Knowledge 35.0
- Multilingual 41.2
- Instruction Following 67.2
- Long Context 39.4
- Writing & Preference 49.1
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 38.6 | #172 | 1 |
| Reasoning | 25.9 | #140 | 1 |
| Math | 36.5 | #165 | 1 |
| Knowledge | 35.0 | #190 | 1 |
| Multilingual | 41.2 | #210 | 1 |
| Instruction Following | 67.2 | #198 | 1 |
| Long Context | 39.4 | #182 | 1 |
| Writing & Preference | 49.1 | #193 | 3 |
Strengths and weaknesses
Categories where Olmo 3 32b Think places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 41.2 | −6.2 | #210 of 297, top 71% |
| Instruction Following | 67.2 | −4.0 | #198 of 305, top 65% |
| Writing & Preference | 49.1 | −4.7 | #193 of 312, top 62% |
Closest competitors
The models ranked just above and below Olmo 3 32b Think. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Qwen3-30B-A3B | #179 | 38.9 | $0.21 | 42 | Compare |
| GLM-4.7-Flash | #180 | 38.8 | $0.15 | — | Compare |
| Qwen2.5 Plus 1127 | #181 | 38.8 | — | — | Compare |
| Qwen3.6 Flash | #182 | 38.8 | $0.42 | — | Compare |
| Hunyuan Large 2025 02 10 | #184 | 38.6 | — | — | Compare |
| Trinity Large Thinking | #185 | 38.6 | $0.39 | — | Compare |
| GPT-5.1-Codex | #186 | 38.6 | $3.44 | — | Compare |
| Sonar | #187 | 38.5 | $1 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1319 | #186 of 294, top 64% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1302 | #188 of 297, top 64% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1316 | #173 of 285, top 61% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Expert | 1273 | #186 of 273, top 69% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1255 | #210 of 297, top 71% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1300 | #188 of 285, top 66% | LMArena | 2026-10-08 | |
| LMArena French | 1291 | #165 of 223, top 74% | LMArena | 2026-10-08 | |
| LMArena German | 1290 | #153 of 231, top 67% | LMArena | 2026-10-08 | |
| LMArena Russian | 1254 | #209 of 283, top 74% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1275 | #192 of 298, top 65% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1296 | #193 of 291, top 67% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1300 | #191 of 297, top 65% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1256 | #200 of 295, top 68% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1290 | #195 of 295, top 67% | LMArena | 2026-10-08 |
Compare Olmo 3 32b Think
- Olmo 3 32b Think vs Qwen3.6 Flash
- Olmo 3 32b Think vs Hunyuan Large 2025 02 10
- Olmo 3 32b Think vs Qwen2.5 Plus 1127
- Olmo 3 32b Think vs Trinity Large Thinking
- Olmo 3 32b Think vs GLM-4.7-Flash
- Olmo 3 32b Think vs GPT-5.1-Codex
- Olmo 3 32b Think vs GPT-6 Astra
- Olmo 3 32b Think vs Claude Fable 5.1
- Olmo 3 32b Think vs Gemini 3.8 Flash
- Olmo 3 32b Think vs Kimi K3
- Olmo 3 32b Think vs Grok 4.6
- Olmo 3 32b Think vs Qwen3.8 Max
- Olmo 3 32b Think vs GLM-5.3
- Olmo 3 32b Think vs Muse Spark 1.3
Other Allen Institute for AI (Ai2) models
Frequently asked questions
How good is Olmo 3 32b Think?
Olmo 3 32b Think by Allen Institute for AI (Ai2) ranks 183rd of 354 ranked models on the Noometry Index as of October 2026, with a score of 38.7. Its strongest category is reasoning, where it ranks 140th.
Is Olmo 3 32b Think open source?
Yes. Olmo 3 32b Think's weights are downloadable; check the license for commercial terms.
What are Olmo 3 32b Think's strengths and weaknesses?
Relative to other ranked models, Olmo 3 32b Think places best in reasoning, math, coding and lowest in multilingual, instruction following, writing & preference.
What is Olmo 3 32b Think best at?
Its best category is reasoning, where it ranks 140th on Noometry.