Model comparison

Granite 4.1 8b vs Qwen3.6 35B-A3B

Granite 4.1 8b and Qwen3.6 35B-A3B score almost the same on the Noometry Index (37.4 vs 37.6), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Granite 4.1 8b IBM

37.4

Rank #205 Confirmed

Qwen3.6 35B-A3B Alibaba (Qwen)

37.6

Rank #201 Confirmed

Summary

  • The widest gap is in knowledge, where Qwen3.6 35B-A3B leads 51.3 to 36.1.

Side by side

Granite 4.1 8b and Qwen3.6 35B-A3B specifications
Granite 4.1 8bQwen3.6 35B-A3B
ProviderIBMAlibaba (Qwen)
Noometry Index37.437.6
Released—2026-04-01
WeightsOpenOpen
Context window—262K
Max output—66K
Input $ / M tokens—$0.25
Output $ / M tokens—$1.49
Results tracked1314

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 35B-A3B leads

Granite 4.1 8b: 30.1 (#297), Qwen3.6 35B-A3B: 37.2 (#196)

Coding benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
LMArena WebDev1192—
SciCode—35.8%
WeirdML—34.5%
LMArena Coding1312—

Agentic & Tool Use Not comparable

Granite 4.1 8b: —, Qwen3.6 35B-A3B: 22.1 (#134)

Agentic & Tool Use benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
Terminal-Bench—23%

Reasoning Qwen3.6 35B-A3B leads

Granite 4.1 8b: 25.7 (#143), Qwen3.6 35B-A3B: 28.0 (#109)

Reasoning benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
NYT Connections (extended)—41.6%
CritPt—0.3%
Chess Puzzles—26%
LMArena Hard Prompts1293—
Mystery Game Puzzles—22%
DTBench—73.9%
LMCA—29.7%
Surface Evolver Bench—44.4%
Epoch Capabilities Index—143.93

Math Qwen3.6 35B-A3B leads

Granite 4.1 8b: 36.4 (#166), Qwen3.6 35B-A3B: 38.9 (#121)

Math benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
FrontierMath (Tiers 1-3)—20.4%
OTIS Mock AIME 2024-2025—86.7%
LMArena Math1312—

Knowledge Qwen3.6 35B-A3B leads

Granite 4.1 8b: 36.1 (#174), Qwen3.6 35B-A3B: 51.3 (#68)

Knowledge benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
GPQA Diamond—84.8%
LMArena Expert1309—

Multilingual Not comparable

Granite 4.1 8b: 41.7 (#204), Qwen3.6 35B-A3B: —

Multilingual benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
LMArena Non-English1261—
LMArena Chinese1337—
LMArena Russian1240—

Instruction Following Not comparable

Granite 4.1 8b: 66.9 (#203), Qwen3.6 35B-A3B: —

Instruction Following benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
LMArena Instruction Following1269—

Long Context Not comparable

Granite 4.1 8b: 38.7 (#193), Qwen3.6 35B-A3B: —

Long Context benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
LMArena Longer Query1275—

Writing & Preference Not comparable

Granite 4.1 8b: 48.1 (#204), Qwen3.6 35B-A3B: —

Writing & Preference benchmarks
BenchmarkGranite 4.1 8bQwen3.6 35B-A3B
LMArena Text1290—
LMArena Creative Writing1251—
LMArena Multi-Turn1268—

Frequently asked questions

Is Granite 4.1 8b better than Qwen3.6 35B-A3B?

Granite 4.1 8b and Qwen3.6 35B-A3B score almost the same on the Noometry Index (37.4 vs 37.6), so choose on price, context window or the category you care about most.

Is Granite 4.1 8b or Qwen3.6 35B-A3B better for coding?

Qwen3.6 35B-A3B scores higher on coding benchmarks: 37.2 versus 30.1 in the Noometry coding category.

How many benchmarks do Granite 4.1 8b and Qwen3.6 35B-A3B share?

0 benchmarks have published results for both models. Granite 4.1 8b has 13 scored results on Noometry and Qwen3.6 35B-A3B has 14.

Related comparisons

Go deeper