Model comparison

Gemma 7B vs Qwen Max

Qwen Max is the stronger model overall, scoring 34.7 to 30.0 on the Noometry Index.

Last verified . 13 shared benchmarks.

Gemma 7B Google

30.0

Rank #299 Confirmed

Qwen Max Alibaba (Qwen)

34.7

Rank #230 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Gemma 7B scores higher in 1 category and Qwen Max in 7 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Qwen Max leads 47.8 to 27.1.
  • Gemma 7B has downloadable open weights; the other is API-only.

Side by side

Gemma 7B and Qwen Max specifications
Gemma 7BQwen Max
ProviderGoogleAlibaba (Qwen)
Noometry Index30.034.7
Released2024-02-212024-04-03
WeightsOpenProprietary
Context window—33K
Max output—8K
Input $ / M tokens—$1.60
Output $ / M tokens—$6.40
Results tracked2723

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemma 7B: 30.5 (#294), Qwen Max: 30.7 (#292)

Coding benchmarks
BenchmarkGemma 7BQwen Max
LMArena Coding10481288
Aider Polyglot—21.8%
HumanEval+28.7%—
MBPP+43.4%—

Reasoning Qwen Max leads

Gemma 7B: 19.9 (#249), Qwen Max: 25.1 (#151)

Reasoning benchmarks
BenchmarkGemma 7BQwen Max
LMArena Hard Prompts10421269
Adversarial NLI48.7%—
BIG-Bench Hard55.1%—
Epoch Capabilities Index111.99—
HellaSwag82.2%—
PIQA81.2%—
WinoGrande79%—

Math Gemma 7B leads

Gemma 7B: 31.2 (#228), Qwen Max: 22.3 (#276)

Math benchmarks
BenchmarkGemma 7BQwen Max
LMArena Math10661275
OTIS Mock AIME 2024-2025—16.1%
MATH Level 5—67.2%
FrontierMath (Feb 2025 set)—1%
GSM8K46.4%—

Knowledge Qwen Max leads

Gemma 7B: 27.3 (#252), Qwen Max: 30.3 (#228)

Knowledge benchmarks
BenchmarkGemma 7BQwen Max
LMArena Expert10011248
GPQA Diamond—56.1%
ARC (AI2) Challenge78.3%—
BoolQ83.2%—
MMLU66.1%—
OpenBookQA78.6%—
TriviaQA72.3%—

Multilingual Qwen Max leads

Gemma 7B: 25.1 (#287), Qwen Max: 41.8 (#202)

Multilingual benchmarks
BenchmarkGemma 7BQwen Max
LMArena Non-English9991263
LMArena Chinese10351254
LMArena French10251330
LMArena Russian9931274
LMArena German—1254
LMArena Japanese—1205
LMArena Korean—1142
LMArena Spanish—1290

Instruction Following Qwen Max leads

Gemma 7B: 51.5 (#295), Qwen Max: 66.5 (#208)

Instruction Following benchmarks
BenchmarkGemma 7BQwen Max
LMArena Instruction Following10171262

Long Context Qwen Max leads

Gemma 7B: 31.1 (#282), Qwen Max: 39.4 (#180)

Long Context benchmarks
BenchmarkGemma 7BQwen Max
LMArena Longer Query10221288
Fiction.LiveBench—66.7%

Writing & Preference Qwen Max leads

Gemma 7B: 27.1 (#302), Qwen Max: 47.8 (#205)

Writing & Preference benchmarks
BenchmarkGemma 7BQwen Max
LMArena Text10561282
LMArena Creative Writing10241248
LMArena Multi-Turn9631277

Frequently asked questions

Is Gemma 7B better than Qwen Max?

Qwen Max is the stronger model overall, scoring 34.7 to 30.0 on the Noometry Index.

Is Gemma 7B or Qwen Max better for coding?

They score almost the same on coding (30.5 vs 30.7); test both on your own repository before choosing.

How many benchmarks do Gemma 7B and Qwen Max share?

13 benchmarks have published results for both models. Gemma 7B has 27 scored results on Noometry and Qwen Max has 23.

Related comparisons

Go deeper