Allen Institute for AI (Ai2), open weights
Olmo 3.1 32b Instruct
Olmo 3.1 32b Instruct by Allen Institute for AI (Ai2) ranks 168th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.4. Its strongest category is reasoning, where it ranks 132nd.
Last verified
Specifications
- Noometry rank
- #168 of 354
- Index score
- 39.4
- Evidence
- Confirmed 16 results
- Provider
Allen Institute for AI (Ai2)
- Released
- Unknown
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 39.5
- Reasoning 26.4
- Math 36.3
- Knowledge 36.1
- Multilingual 42.6
- Instruction Following 68.6
- Long Context 39.9
- Writing & Preference 50.2
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 39.5 | #157 | 1 |
| Reasoning | 26.4 | #132 | 1 |
| Math | 36.3 | #167 | 1 |
| Knowledge | 36.1 | #175 | 1 |
| Multilingual | 42.6 | #191 | 1 |
| Instruction Following | 68.6 | #187 | 1 |
| Long Context | 39.9 | #166 | 1 |
| Writing & Preference | 50.2 | #185 | 3 |
Strengths and weaknesses
Categories where Olmo 3.1 32b Instruct places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 42.6 | −4.8 | #191 of 297, top 65% |
| Instruction Following | 68.6 | −2.7 | #187 of 305, top 62% |
| Writing & Preference | 50.2 | −3.6 | #185 of 312, top 60% |
Closest competitors
The models ranked just above and below Olmo 3.1 32b Instruct. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Claude 3.7 Sonnet | #164 | 39.5 | — | — | Compare |
| Claude Haiku 4.5 | #165 | 39.5 | $2 | — | Compare |
| DeepSeek-V3 | #166 | 39.5 | $0.41 | 73 | Compare |
| Grok 4 Fast | #167 | 39.4 | — | 577 | Compare |
| Granite 4.2 3b | #169 | 39.4 | — | — | Compare |
| Gemini 2.5 Flash | #170 | 39.3 | $0.85 | 152 | Compare |
| Step 2 16k Exp 202412 | #171 | 39.2 | — | — | Compare |
| Qwen3 32B | #172 | 39.2 | $1.22 | 86 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1347 | #175 of 294, top 60% | LMArena | 2026-10-08 |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Hard Prompts | 1322 | #179 of 297, top 61% | LMArena | 2026-10-08 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Math | 1305 | #180 of 285, top 64% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Expert | 1308 | #174 of 273, top 64% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1275 | #191 of 297, top 65% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1304 | #187 of 285, top 66% | LMArena | 2026-10-08 | |
| LMArena French | 1328 | #154 of 223, top 70% | LMArena | 2026-10-08 | |
| LMArena German | 1282 | #159 of 231, top 69% | LMArena | 2026-10-08 | |
| LMArena Korean | 1206 | #158 of 213, top 75% | LMArena | 2026-10-08 | |
| LMArena Russian | 1268 | #197 of 283, top 70% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1336 | #151 of 226, top 67% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1299 | #180 of 298, top 61% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1312 | #178 of 291, top 62% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1311 | #185 of 297, top 63% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1264 | #196 of 295, top 67% | LMArena | 2026-10-08 | |
| LMArena Multi-Turn | 1309 | #183 of 295, top 63% | LMArena | 2026-10-08 |
Compare Olmo 3.1 32b Instruct
- Olmo 3.1 32b Instruct vs Grok 4 Fast
- Olmo 3.1 32b Instruct vs Granite 4.2 3b
- Olmo 3.1 32b Instruct vs DeepSeek-V3
- Olmo 3.1 32b Instruct vs Gemini 2.5 Flash
- Olmo 3.1 32b Instruct vs Claude Haiku 4.5
- Olmo 3.1 32b Instruct vs Step 2 16k Exp 202412
- Olmo 3.1 32b Instruct vs GPT-6 Astra
- Olmo 3.1 32b Instruct vs Claude Fable 5.1
- Olmo 3.1 32b Instruct vs Gemini 3.8 Flash
- Olmo 3.1 32b Instruct vs Kimi K3
- Olmo 3.1 32b Instruct vs Grok 4.6
- Olmo 3.1 32b Instruct vs Qwen3.8 Max
- Olmo 3.1 32b Instruct vs GLM-5.3
- Olmo 3.1 32b Instruct vs Muse Spark 1.3
Other Allen Institute for AI (Ai2) models
Frequently asked questions
How good is Olmo 3.1 32b Instruct?
Olmo 3.1 32b Instruct by Allen Institute for AI (Ai2) ranks 168th of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.4. Its strongest category is reasoning, where it ranks 132nd.
Is Olmo 3.1 32b Instruct open source?
Yes. Olmo 3.1 32b Instruct's weights are downloadable; check the license for commercial terms.
What are Olmo 3.1 32b Instruct's strengths and weaknesses?
Relative to other ranked models, Olmo 3.1 32b Instruct places best in reasoning, coding, math and lowest in multilingual, instruction following, writing & preference.
What is Olmo 3.1 32b Instruct best at?
Its best category is reasoning, where it ranks 132nd on Noometry.