Model comparison

Gemma 3 1B vs Llama 3.2 1B

Gemma 3 1B and Llama 3.2 1B score almost the same on the Noometry Index (21.1 vs 20.1), so choose on price, context window or the category you care about most.

Last verified . 4 shared benchmarks.

Gemma 3 1B Google

21.1

Rank #353 Reported

Llama 3.2 1B Meta

20.1

Rank #354 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Gemma 3 1B scores higher in 1 category and Llama 3.2 1B in 3 categories; one gap is clear of the uncertainty.
  • The widest gap is in reasoning, where Gemma 3 1B leads 19.2 to 16.2.

Side by side

Gemma 3 1B and Llama 3.2 1B specifications
Gemma 3 1BLlama 3.2 1B
ProviderGoogleMeta
Noometry Index21.120.1
Released2025-03-122024-09-24
WeightsOpenOpen
Context window—60K
Max output—54K
Input $ / M tokens—$0.027
Output $ / M tokens—$0.20
Results tracked422

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Gemma 3 1B: —, Llama 3.2 1B: 21.1 (#338)

Coding benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
BigCodeBench Instruct—8.2%
LMArena Coding—1070
BigCodeBench Complete—11.3%

Agentic & Tool Use Too close to call

Gemma 3 1B: 13.7 (#152), Llama 3.2 1B: 14.6 (#150)

Agentic & Tool Use benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
Berkeley Function Calling Leaderboard7.2%10.8%
BALROG—6.6%

Reasoning Gemma 3 1B leads

Gemma 3 1B: 19.2 (#264), Llama 3.2 1B: 16.2 (#308)

Reasoning benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
Chess Puzzles0%0%
LMArena Hard Prompts—1044
Epoch Capabilities Index—101.99

Math Too close to call

Gemma 3 1B: 10.2 (#316), Llama 3.2 1B: 10.4 (#313)

Math benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
OTIS Mock AIME 2024-20251.1%0.6%
LMArena Math—1086

Knowledge Too close to call

Gemma 3 1B: 7.0 (#314), Llama 3.2 1B: 7.2 (#312)

Knowledge benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
GPQA Diamond19.9%23.9%
LMArena Expert—1007

Multilingual Not comparable

Gemma 3 1B: —, Llama 3.2 1B: 23.8 (#292)

Multilingual benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
LMArena Non-English—973
LMArena Chinese—959
LMArena German—1014
LMArena Russian—941

Instruction Following Not comparable

Gemma 3 1B: —, Llama 3.2 1B: 52.4 (#290)

Instruction Following benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
LMArena Instruction Following—1031

Long Context Not comparable

Gemma 3 1B: —, Llama 3.2 1B: 31.9 (#274)

Long Context benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
LMArena Longer Query—1050

Writing & Preference Not comparable

Gemma 3 1B: —, Llama 3.2 1B: 21.3 (#310)

Writing & Preference benchmarks
BenchmarkGemma 3 1BLlama 3.2 1B
LMArena Text—1055
LMArena Creative Writing—1033
EQ-Bench Creative Writing—200
LMArena Multi-Turn—1030

Frequently asked questions

Is Gemma 3 1B better than Llama 3.2 1B?

Gemma 3 1B and Llama 3.2 1B score almost the same on the Noometry Index (21.1 vs 20.1), so choose on price, context window or the category you care about most.

How many benchmarks do Gemma 3 1B and Llama 3.2 1B share?

4 benchmarks have published results for both models. Gemma 3 1B has 4 scored results on Noometry and Llama 3.2 1B has 22.

Related comparisons

Go deeper