Allen Institute for AI (Ai2), open weights
OLMo 2 Furious 13B
OLMo 2 Furious 13B by Allen Institute for AI (Ai2) ranks 304th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.7. Its strongest category is writing & preference, where it ranks 230th.
Last verified
Specifications
- Noometry rank
- #304 of 354
- Index score
- 29.7
- Evidence
- Confirmed 12 results
- Provider
Allen Institute for AI (Ai2)
- Released
- December 31, 2024
- Weights
- Open weights
- Reasoning
- Unknown
- Context window
- —
- Max output
- —
- Input price
- Not listed
- Output price
- Not listed
- Blended price
- Not listed
- Output speed
- Not measured
- Value
- Not ranked
- Knowledge cutoff
- Unknown
Category scores
Each category score combines every public result we have in that category.
- Coding 28.6
- Reasoning 15.5
- Math 23.1
- Knowledge 18.3
- Instruction Following 60.9
- Writing & Preference 42.9
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 28.6 | #313 | 1 |
| Reasoning | 15.5 | #314 | 2 |
| Math | 23.1 | #273 | 2 |
| Knowledge | 18.3 | #283 | 2 |
| Instruction Following | 60.9 | #248 | 2 |
| Writing & Preference | 42.9 | #230 | 2 |
Strengths and weaknesses
Categories where OLMo 2 Furious 13B places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Writing & Preference | 42.9 | −10.8 | #230 of 312, top 74% |
| Instruction Following | 60.9 | −10.4 | #248 of 305, top 82% |
| Math | 23.1 | −13.5 | #273 of 327, top 84% |
Closest competitors
The models ranked just above and below OLMo 2 Furious 13B. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Qwen2-72B | #300 | 30.0 | — | — | Compare |
| Gemini 1.5 Flash 8B | #301 | 29.9 | — | — | Compare |
| Qwen2.5-VL 72B Instruct | #302 | 29.9 | $4.20 | 43 | Compare |
| Mistral | #303 | 29.9 | — | — | Compare |
| Phi 3 Mini 128k Instruct | #305 | 29.7 | — | — | Compare |
| phi-3-medium 14B | #306 | 29.7 | — | — | Compare |
| Gemma 2B | #307 | 29.6 | — | — | Compare |
| Llama 3.1-70B | #308 | 29.6 | $0.40 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Coding | 10.4% | #39 of 39, top 100% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Reasoning | 16.3% | #37 of 39, top 95% | Epoch AI | ||
| LiveBench Data Analysis | 20.6% | #39 of 39, top 100% | Epoch AI | ||
| LiveBench | 22.1% | #39 of 39, top 100% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Omni-MATH | 15.6% | #54 of 57, top 95% | HELM Capabilities | ||
| LiveBench Math | 13.6% | #39 of 39, top 100% | Epoch AI |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| MMLU-Pro | 31% | #57 of 58, top 99% | HELM Capabilities | ||
| GPQA (HELM) | 31.6% | #51 of 57, top 90% | HELM Capabilities |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LiveBench Instruction Following | 60.6% | #27 of 39, top 70% | Epoch AI | ||
| IFEval | 73% | #54 of 57, top 95% | HELM Capabilities |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| WildBench | 68.9% | #52 of 57, top 92% | HELM Capabilities | ||
| LiveBench Language | 11.2% | #38 of 39, top 98% | Epoch AI |
Compare OLMo 2 Furious 13B
- OLMo 2 Furious 13B vs Mistral
- OLMo 2 Furious 13B vs Phi 3 Mini 128k Instruct
- OLMo 2 Furious 13B vs Qwen2.5-VL 72B Instruct
- OLMo 2 Furious 13B vs phi-3-medium 14B
- OLMo 2 Furious 13B vs Gemini 1.5 Flash 8B
- OLMo 2 Furious 13B vs Gemma 2B
- OLMo 2 Furious 13B vs GPT-6 Astra
- OLMo 2 Furious 13B vs Claude Fable 5.1
- OLMo 2 Furious 13B vs Gemini 3.8 Flash
- OLMo 2 Furious 13B vs Kimi K3
- OLMo 2 Furious 13B vs Grok 4.6
- OLMo 2 Furious 13B vs Qwen3.8 Max
- OLMo 2 Furious 13B vs GLM-5.3
- OLMo 2 Furious 13B vs Muse Spark 1.3
Other Allen Institute for AI (Ai2) models
Frequently asked questions
How good is OLMo 2 Furious 13B?
OLMo 2 Furious 13B by Allen Institute for AI (Ai2) ranks 304th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.7. Its strongest category is writing & preference, where it ranks 230th.
Is OLMo 2 Furious 13B open source?
Yes. OLMo 2 Furious 13B's weights are downloadable; check the license for commercial terms.
What are OLMo 2 Furious 13B's strengths and weaknesses?
Relative to other ranked models, OLMo 2 Furious 13B places best in writing & preference, instruction following, math and lowest in coding, knowledge, reasoning.
What is OLMo 2 Furious 13B best at?
Its best category is writing & preference, where it ranks 230th on Noometry.