Model comparison

Gemma 1.1 7b IT vs Qwen1.5-7B

Gemma 1.1 7b IT and Qwen1.5-7B score almost the same on the Noometry Index (31.3 vs 31.4), so choose on price, context window or the category you care about most.

Last verified . 12 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Qwen1.5-7B Alibaba (Qwen)

31.4

Rank #273 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 3 categories and Qwen1.5-7B in 5 categories; one gap is clear of the uncertainty.

Side by side

Gemma 1.1 7b IT and Qwen1.5-7B specifications
Gemma 1.1 7b ITQwen1.5-7B
ProviderGoogleAlibaba (Qwen)
Noometry Index31.331.4
Released—2024-02-04
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1913

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemma 1.1 7b IT: 31.5 (#284), Qwen1.5-7B: 32.2 (#276)

Coding benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Coding10841107
HumanEval+35.4%—
MBPP+45%—

Reasoning Too close to call

Gemma 1.1 7b IT: 20.5 (#238), Qwen1.5-7B: 20.4 (#240)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Hard Prompts10711065

Math Too close to call

Gemma 1.1 7b IT: 32.0 (#220), Qwen1.5-7B: 31.4 (#224)

Math benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Math11071080

Knowledge Too close to call

Gemma 1.1 7b IT: 28.3 (#247), Qwen1.5-7B: 28.7 (#243)

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Expert10391055
MMLU—62.6%

Multilingual Too close to call

Gemma 1.1 7b IT: 28.1 (#273), Qwen1.5-7B: 28.5 (#271)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Non-English10521058
LMArena Chinese10611141
LMArena Russian10461006
LMArena French1065—
LMArena German1054—
LMArena Japanese971—
LMArena Korean988—
LMArena Spanish1049—

Instruction Following Too close to call

Gemma 1.1 7b IT: 54.0 (#283), Qwen1.5-7B: 54.1 (#281)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Instruction Following10571058

Long Context Qwen1.5-7B leads

Gemma 1.1 7b IT: 32.1 (#272), Qwen1.5-7B: 33.1 (#266)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Longer Query10561090

Writing & Preference Too close to call

Gemma 1.1 7b IT: 30.4 (#288), Qwen1.5-7B: 29.6 (#293)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITQwen1.5-7B
LMArena Text10941083
LMArena Creative Writing10601035
LMArena Multi-Turn10401062

Frequently asked questions

Is Gemma 1.1 7b IT better than Qwen1.5-7B?

Gemma 1.1 7b IT and Qwen1.5-7B score almost the same on the Noometry Index (31.3 vs 31.4), so choose on price, context window or the category you care about most.

Is Gemma 1.1 7b IT or Qwen1.5-7B better for coding?

They score almost the same on coding (31.5 vs 32.2); test both on your own repository before choosing.

How many benchmarks do Gemma 1.1 7b IT and Qwen1.5-7B share?

12 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Qwen1.5-7B has 13.

Related comparisons

Go deeper