Alibaba (Qwen), open weights

Qwen2.5-VL 72B Instruct

Qwen2.5-VL 72B Instruct by Alibaba (Qwen) ranks 302nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.9. Its strongest category is multimodal, where it ranks 97th. API pricing starts at $2.80 per million input tokens and $8.40 per million output tokens, with a 131K-token context window.

Last verified

Specifications

Noometry rank
#302 of 354
Index score
29.9
Evidence
Confirmed 6 results
Released
September 1, 2024
Weights
Open weights
Reasoning
No
Context window
131K
Max output
8K
Input price
$2.80 / M
Output price
$8.40 / M
Blended price
$4.20 / M
Output speed
43 tokens/s Kagi
Value
#197 of 219
Knowledge cutoff
April 2024
Input
text, image

Category scores

Each category score combines every public result we have in that category.

Qwen2.5-VL 72B Instruct category scores
  1. Agentic & Tool Use 18.6
  2. Reasoning 20.7
  3. Multimodal 33.5
Qwen2.5-VL 72B Instruct category ranks
CategoryScoreRankResults
Agentic & Tool Use18.6#1441
Reasoning20.7#2331
Multimodal33.5#973

Strengths and weaknesses

Categories where Qwen2.5-VL 72B Instruct places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen2.5-VL 72B Instruct: strongest categories
CategoryScorevs medianRank
Reasoning20.7−2.9#233 of 350, top 67%

Weakest categories

Qwen2.5-VL 72B Instruct: weakest categories
CategoryScorevs medianRank
Agentic & Tool Use18.6−11.7#144 of 154, top 94%

Closest competitors

The models ranked just above and below Qwen2.5-VL 72B Instruct. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen2.5-VL 72B Instruct
ModelRankScoreBlended $/MSpeed
Palm 2#29830.0——Compare
Gemma 7B#29930.0——Compare
Qwen2-72B#30030.0——Compare
Gemini 1.5 Flash 8B#30129.9——Compare
Mistral#30329.9——Compare
OLMo 2 Furious 13B#30429.7——Compare
Phi 3 Mini 128k Instruct#30529.7——Compare
phi-3-medium 14B#30629.7——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Agentic & Tool Use

Qwen2.5-VL 72B Instruct Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
OSWorld5%#8 of 8, top 100%Epoch AI

Reasoning

Qwen2.5-VL 72B Instruct Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark36%#87 of 99, top 88%Kagi LLM Benchmark

Multimodal

Qwen2.5-VL 72B Instruct Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1107#106 of 122, top 87%LMArena2026-10-09
Video-MME73.5%#4 of 15, top 27%Epoch AI
GeoBench62%#17 of 25, top 68%Epoch AI
SpatialViz-Bench33.3%#6 of 8, top 75%Epoch AI

API pricing by provider

Qwen2.5-VL 72B Instruct API prices
RouteInput $/MOutput $/MCached input $/MChecked
alibaba$2.80$8.40—2026-10-10
openrouter$0.80$1$0.402026-10-10

Compare Qwen2.5-VL 72B Instruct

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen2.5-VL 72B Instruct?

Qwen2.5-VL 72B Instruct by Alibaba (Qwen) ranks 302nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.9. Its strongest category is multimodal, where it ranks 97th. API pricing starts at $2.80 per million input tokens and $8.40 per million output tokens, with a 131K-token context window.

How much does Qwen2.5-VL 72B Instruct cost?

Qwen2.5-VL 72B Instruct costs $2.80 per million input tokens and $8.40 per million output tokens on Alibaba (Qwen)'s own API.

What is Qwen2.5-VL 72B Instruct's context window?

Qwen2.5-VL 72B Instruct accepts up to 131K tokens of input and can write up to 8K tokens in one response.

Is Qwen2.5-VL 72B Instruct open source?

Yes. Qwen2.5-VL 72B Instruct's weights are downloadable from Hugging Face (Qwen/Qwen2.5-VL-72B-Instruct); check the license for commercial terms.

How fast is Qwen2.5-VL 72B Instruct?

Qwen2.5-VL 72B Instruct generated about 43 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are Qwen2.5-VL 72B Instruct's strengths and weaknesses?

Relative to other ranked models, Qwen2.5-VL 72B Instruct places best in reasoning and lowest in agentic & tool use.

What is Qwen2.5-VL 72B Instruct best at?

Its best category is multimodal, where it ranks 97th on Noometry.