Model comparison

Gemma 7B vs Qwen2.5-VL 72B Instruct

Gemma 7B and Qwen2.5-VL 72B Instruct score almost the same on the Noometry Index (30.0 vs 29.9), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Side by side

Gemma 7B and Qwen2.5-VL 72B Instruct specifications
Gemma 7BQwen2.5-VL 72B Instruct
ProviderGoogleAlibaba (Qwen)
Noometry Index30.029.9
Released2024-02-212024-09
WeightsOpenOpen
Context window—131K
Max output—8K
Input $ / M tokens—$2.80
Output $ / M tokens—$8.40
Results tracked276

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 7B: 30.5 (#294), Qwen2.5-VL 72B Instruct: —

Coding benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Coding1048—
HumanEval+28.7%—
MBPP+43.4%—

Agentic & Tool Use Not comparable

Gemma 7B: —, Qwen2.5-VL 72B Instruct: 18.6 (#144)

Agentic & Tool Use benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
OSWorld—5%

Reasoning Too close to call

Gemma 7B: 19.9 (#249), Qwen2.5-VL 72B Instruct: 20.7 (#233)

Reasoning benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
Kagi LLM Benchmark—36%
LMArena Hard Prompts1042—
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Not comparable

Gemma 7B: 31.2 (#228), Qwen2.5-VL 72B Instruct: —

Math benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Math1066—
GSM8K46.4%—

Knowledge Not comparable

Gemma 7B: 27.3 (#252), Qwen2.5-VL 72B Instruct: —

Knowledge benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Expert1001—
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multimodal Not comparable

Gemma 7B: —, Qwen2.5-VL 72B Instruct: 33.5 (#97)

Multimodal benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Vision—1107
Video-MME—73.5%
GeoBench—62%
SpatialViz-Bench—33.3%

Multilingual Not comparable

Gemma 7B: 25.1 (#287), Qwen2.5-VL 72B Instruct: —

Multilingual benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Non-English999—
LMArena Chinese1035—
LMArena French1025—
LMArena Russian993—

Instruction Following Not comparable

Gemma 7B: 51.5 (#295), Qwen2.5-VL 72B Instruct: —

Instruction Following benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Instruction Following1017—

Long Context Not comparable

Gemma 7B: 31.1 (#282), Qwen2.5-VL 72B Instruct: —

Long Context benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Longer Query1022—

Writing & Preference Not comparable

Gemma 7B: 27.1 (#302), Qwen2.5-VL 72B Instruct: —

Writing & Preference benchmarks
BenchmarkGemma 7BQwen2.5-VL 72B Instruct
LMArena Text1056—
LMArena Creative Writing1024—
LMArena Multi-Turn963—

Frequently asked questions

Is Gemma 7B better than Qwen2.5-VL 72B Instruct?

Gemma 7B and Qwen2.5-VL 72B Instruct score almost the same on the Noometry Index (30.0 vs 29.9), so choose on price, context window or the category you care about most.

How many benchmarks do Gemma 7B and Qwen2.5-VL 72B Instruct share?

0 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Qwen2.5-VL 72B Instruct has 6.

Related comparisons

Go deeper