Model comparison

Gemma 1.1 7b IT vs Qwen-14B

Gemma 1.1 7b IT and Qwen-14B score almost the same on the Noometry Index (31.3 vs 31.4), so choose on price, context window or the category you care about most.

Last verified . 10 shared benchmarks.

Gemma 1.1 7b IT Google

31.3

Rank #277 Confirmed

Qwen-14B Alibaba (Qwen)

31.4

Rank #275 Confirmed

Summary

  • They share 10 benchmarks with published results for both. Gemma 1.1 7b IT scores higher in 7 categories and Qwen-14B in 0 categories; 2 gaps are clear of the uncertainty.

Side by side

Gemma 1.1 7b IT and Qwen-14B specifications
Gemma 1.1 7b ITQwen-14B
ProviderGoogleAlibaba (Qwen)
Noometry Index31.331.4
Released—2023-09-24
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked1918

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemma 1.1 7b IT: 31.5 (#284), Qwen-14B: 31.2 (#288)

Coding benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Coding10841071
HumanEval+35.4%—
MBPP+45%—

Reasoning Too close to call

Gemma 1.1 7b IT: 20.5 (#238), Qwen-14B: 19.6 (#257)

Reasoning benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Hard Prompts10711027
BIG-Bench Hard—55%
Epoch Capabilities Index—113.03
LAMBADA—71.1%
PIQA—79.9%

Math Too close to call

Gemma 1.1 7b IT: 32.0 (#220), Qwen-14B: 31.2 (#227)

Math benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Math11071068
GSM8K—61.3%

Knowledge Not comparable

Gemma 1.1 7b IT: 28.3 (#247), Qwen-14B: —

Knowledge benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Expert1039—
ARC (AI2) Challenge—84.4%
BoolQ—86.2%
MMLU—66.3%

Multilingual Too close to call

Gemma 1.1 7b IT: 28.1 (#273), Qwen-14B: 27.5 (#275)

Multilingual benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Non-English10521041
LMArena Chinese10611077
LMArena French1065—
LMArena German1054—
LMArena Japanese971—
LMArena Korean988—
LMArena Russian1046—
LMArena Spanish1049—

Instruction Following Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 54.0 (#283), Qwen-14B: 52.4 (#289)

Instruction Following benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Instruction Following10571031

Long Context Too close to call

Gemma 1.1 7b IT: 32.1 (#272), Qwen-14B: 31.3 (#280)

Long Context benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Longer Query10561028

Writing & Preference Gemma 1.1 7b IT leads

Gemma 1.1 7b IT: 30.4 (#288), Qwen-14B: 27.6 (#299)

Writing & Preference benchmarks
BenchmarkGemma 1.1 7b ITQwen-14B
LMArena Text10941051
LMArena Creative Writing10601028
LMArena Multi-Turn10401022

Frequently asked questions

Is Gemma 1.1 7b IT better than Qwen-14B?

Gemma 1.1 7b IT and Qwen-14B score almost the same on the Noometry Index (31.3 vs 31.4), so choose on price, context window or the category you care about most.

Is Gemma 1.1 7b IT or Qwen-14B better for coding?

They score almost the same on coding (31.5 vs 31.2); test both on your own repository before choosing.

How many benchmarks do Gemma 1.1 7b IT and Qwen-14B share?

10 benchmarks have published results for both models. Gemma 1.1 7b IT has 19 scored results on Noometry and Qwen-14B has 18.

Related comparisons

Go deeper