Model comparison
OLMo 2 Furious 13B vs Qwen2.5-VL 72B Instruct
OLMo 2 Furious 13B and Qwen2.5-VL 72B Instruct score almost the same on the Noometry Index (29.7 vs 29.9), so choose on price, context window or the category you care about most.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in reasoning, where Qwen2.5-VL 72B Instruct leads 20.7 to 15.5.
Side by side
| OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct | |
|---|---|---|
| Provider | Allen Institute for AI (Ai2) | Alibaba (Qwen) |
| Noometry Index | 29.7 | 29.9 |
| Released | 2024-12-31 | 2024-09 |
| Weights | Open | Open |
| Context window | — | 131K |
| Max output | — | 8K |
| Input $ / M tokens | — | $2.80 |
| Output $ / M tokens | — | $8.40 |
| Results tracked | 12 | 6 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
OLMo 2 Furious 13B: 28.6 (#313), Qwen2.5-VL 72B Instruct: —
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| LiveBench Coding | 10.4% | — |
Agentic & Tool Use Not comparable
OLMo 2 Furious 13B: —, Qwen2.5-VL 72B Instruct: 18.6 (#144)
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| OSWorld | — | 5% |
Reasoning Qwen2.5-VL 72B Instruct leads
OLMo 2 Furious 13B: 15.5 (#314), Qwen2.5-VL 72B Instruct: 20.7 (#233)
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| Kagi LLM Benchmark | — | 36% |
| LiveBench Reasoning | 16.3% | — |
| LiveBench Data Analysis | 20.6% | — |
| LiveBench | 22.1% | — |
Math Not comparable
OLMo 2 Furious 13B: 23.1 (#273), Qwen2.5-VL 72B Instruct: —
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| Omni-MATH | 15.6% | — |
| LiveBench Math | 13.6% | — |
Knowledge Not comparable
OLMo 2 Furious 13B: 18.3 (#283), Qwen2.5-VL 72B Instruct: —
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| MMLU-Pro | 31% | — |
| GPQA (HELM) | 31.6% | — |
Multimodal Not comparable
OLMo 2 Furious 13B: —, Qwen2.5-VL 72B Instruct: 33.5 (#97)
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| LMArena Vision | — | 1107 |
| Video-MME | — | 73.5% |
| GeoBench | — | 62% |
| SpatialViz-Bench | — | 33.3% |
Instruction Following Not comparable
OLMo 2 Furious 13B: 60.9 (#248), Qwen2.5-VL 72B Instruct: —
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| LiveBench Instruction Following | 60.6% | — |
| IFEval | 73% | — |
Writing & Preference Not comparable
OLMo 2 Furious 13B: 42.9 (#230), Qwen2.5-VL 72B Instruct: —
| Benchmark | OLMo 2 Furious 13B | Qwen2.5-VL 72B Instruct |
|---|---|---|
| WildBench | 68.9% | — |
| LiveBench Language | 11.2% | — |
Frequently asked questions
Is OLMo 2 Furious 13B better than Qwen2.5-VL 72B Instruct?
OLMo 2 Furious 13B and Qwen2.5-VL 72B Instruct score almost the same on the Noometry Index (29.7 vs 29.9), so choose on price, context window or the category you care about most.
How many benchmarks do OLMo 2 Furious 13B and Qwen2.5-VL 72B Instruct share?
0 benchmarks have published results for both models. OLMo 2 Furious 13B has 12 scored results on Noometry and Qwen2.5-VL 72B Instruct has 6.