Model comparison

Gemini 2.0 Pro vs Qwen3 32B

Gemini 2.0 Pro and Qwen3 32B score almost the same on the Noometry Index (39.1 vs 39.2), so choose on price, context window or the category you care about most.

Last verified . 4 shared benchmarks.

Gemini 2.0 Pro Google

39.1

Rank #173 Confirmed

Qwen3 32B Alibaba (Qwen)

39.2

Rank #172 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Gemini 2.0 Pro scores higher in 3 categories and Qwen3 32B in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in long context, where Qwen3 32B leads 43.8 to 29.2.
  • The biggest single-benchmark swing is Fiction.LiveBench: 41.7% for Gemini 2.0 Pro and 74.2% for Qwen3 32B.
  • Qwen3 32B has downloadable open weights; the other is API-only.

Side by side

Gemini 2.0 Pro and Qwen3 32B specifications
Gemini 2.0 ProQwen3 32B
ProviderGoogleAlibaba (Qwen)
Noometry Index39.139.2
Released2025-02-052025-04
WeightsProprietaryOpen
Context window—131K
Max output—16K
Input $ / M tokens—$0.70
Output $ / M tokens—$2.80
Results tracked1426

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Gemini 2.0 Pro: 37.8 (#187), Qwen3 32B: 37.7 (#190)

Coding benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
Aider Polyglot35.6%40%
SciCode—35.4%
LiveBench Coding63.5%—
LMArena Coding—1358

Agentic & Tool Use Not comparable

Gemini 2.0 Pro: —, Qwen3 32B: 32.6 (#62)

Agentic & Tool Use benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
Berkeley Function Calling Leaderboard—48.7%

Reasoning Gemini 2.0 Pro leads

Gemini 2.0 Pro: 22.3 (#198), Qwen3 32B: 20.2 (#241)

Reasoning benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
Epoch Capabilities Index135.06138.51
Kagi LLM Benchmark—54.9%
CritPt—0.3%
Chess Puzzles—5%
EnigmaEval0.7%—
LiveBench Reasoning60.1%—
LMArena Hard Prompts—1334
DTBench—67.5%
LiveBench Data Analysis68%—
LMCA—17.3%
LiveBench65.1%—

Math Too close to call

Gemini 2.0 Pro: 39.7 (#100), Qwen3 32B: 39.7 (#99)

Math benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
OTIS Mock AIME 2024-2025—66.9%
LiveBench Math71%—
LMArena Math—1399
MATH Level 583.5%—

Knowledge Qwen3 32B leads

Gemini 2.0 Pro: 36.5 (#167), Qwen3 32B: 40.0 (#125)

Knowledge benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
GPQA Diamond65.7%65.7%
Confabulations18.4%—
Vectara Hallucination Rate—5.9%
LMArena Expert—1362

Multilingual Not comparable

Gemini 2.0 Pro: —, Qwen3 32B: 45.6 (#167)

Multilingual benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
LMArena Non-English—1317
LMArena Chinese—1357
LMArena German—1341
LMArena Russian—1311

Instruction Following Gemini 2.0 Pro leads

Gemini 2.0 Pro: 75.5 (#59), Qwen3 32B: 68.9 (#179)

Instruction Following benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
LiveBench Instruction Following83.4%—
LMArena Instruction Following—1305

Long Context Qwen3 32B leads

Gemini 2.0 Pro: 29.2 (#292), Qwen3 32B: 43.8 (#87)

Long Context benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
Fiction.LiveBench41.7%74.2%
LMArena Longer Query—1327

Writing & Preference Too close to call

Gemini 2.0 Pro: 52.7 (#165), Qwen3 32B: 52.9 (#163)

Writing & Preference benchmarks
BenchmarkGemini 2.0 ProQwen3 32B
LMArena Text—1340
LMArena Creative Writing—1297
LMArena Multi-Turn—1331
LiveBench Language44.9%—

Frequently asked questions

Is Gemini 2.0 Pro better than Qwen3 32B?

Gemini 2.0 Pro and Qwen3 32B score almost the same on the Noometry Index (39.1 vs 39.2), so choose on price, context window or the category you care about most.

Is Gemini 2.0 Pro or Qwen3 32B better for coding?

They score almost the same on coding (37.8 vs 37.7); test both on your own repository before choosing.

How many benchmarks do Gemini 2.0 Pro and Qwen3 32B share?

4 benchmarks have published results for both models. Gemini 2.0 Pro has 14 scored results on Noometry and Qwen3 32B has 26.

Related comparisons

Go deeper