Model comparison

Qwen3 32B vs Sonar

Qwen3 32B and Sonar score almost the same on the Noometry Index (39.2 vs 38.5), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Qwen3 32B Alibaba (Qwen)

39.2

Rank #172 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in math, where Qwen3 32B leads 39.7 to 33.7.
  • Sonar is cheaper at $1 / $1 per million input/output tokens, against $0.70 / $2.80 for Qwen3 32B.
  • Qwen3 32B accepts more context: 131K tokens versus 128K.
  • Qwen3 32B has downloadable open weights; the other is API-only.

Side by side

Qwen3 32B and Sonar specifications
Qwen3 32BSonar
ProviderAlibaba (Qwen)Perplexity
Noometry Index39.238.5
Released2025-042024-01-01
WeightsOpenProprietary
Context window131K128K
Max output16K4K
Input $ / M tokens$0.70$1
Output $ / M tokens$2.80$1
Results tracked267

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3 32B leads

Qwen3 32B: 37.7 (#190), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkQwen3 32BSonar
Aider Polyglot40%—
SciCode35.4%—
LiveBench Coding—35.1%
LMArena Coding1358—

Agentic & Tool Use Not comparable

Qwen3 32B: 32.6 (#62), Sonar: —

Agentic & Tool Use benchmarks
BenchmarkQwen3 32BSonar
Berkeley Function Calling Leaderboard48.7%—

Reasoning Too close to call

Qwen3 32B: 20.2 (#241), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkQwen3 32BSonar
Kagi LLM Benchmark54.9%—
CritPt0.3%—
Chess Puzzles5%—
LiveBench Reasoning—46.3%
LMArena Hard Prompts1334—
DTBench67.5%—
LiveBench Data Analysis—37.9%
LMCA17.3%—
Epoch Capabilities Index138.51—
LiveBench—46.9%

Math Qwen3 32B leads

Qwen3 32B: 39.7 (#99), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkQwen3 32BSonar
OTIS Mock AIME 2024-202566.9%—
LiveBench Math—41.6%
LMArena Math1399—

Knowledge Not comparable

Qwen3 32B: 40.0 (#125), Sonar: —

Knowledge benchmarks
BenchmarkQwen3 32BSonar
GPQA Diamond65.7%—
Vectara Hallucination Rate5.9%—
LMArena Expert1362—

Multilingual Not comparable

Qwen3 32B: 45.6 (#167), Sonar: —

Multilingual benchmarks
BenchmarkQwen3 32BSonar
LMArena Non-English1317—
LMArena Chinese1357—
LMArena German1341—
LMArena Russian1311—

Instruction Following Sonar leads

Qwen3 32B: 68.9 (#179), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkQwen3 32BSonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1305—

Long Context Not comparable

Qwen3 32B: 43.8 (#87), Sonar: —

Long Context benchmarks
BenchmarkQwen3 32BSonar
Fiction.LiveBench74.2%—
LMArena Longer Query1327—

Writing & Preference Too close to call

Qwen3 32B: 52.9 (#163), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkQwen3 32BSonar
LMArena Text1340—
LMArena Creative Writing1297—
LMArena Multi-Turn1331—
LiveBench Language—44.1%

Frequently asked questions

Is Qwen3 32B better than Sonar?

Qwen3 32B and Sonar score almost the same on the Noometry Index (39.2 vs 38.5), so choose on price, context window or the category you care about most.

Which is cheaper, Qwen3 32B or Sonar?

Sonar is cheaper. It lists at $1 per million input tokens and $1 per million output tokens; Qwen3 32B lists at $0.70 and $2.80.

Is Qwen3 32B or Sonar better for coding?

Qwen3 32B scores higher on coding benchmarks: 37.7 versus 35.7 in the Noometry coding category.

Which has the bigger context window?

Qwen3 32B does, with 131K tokens against 128K.

How many benchmarks do Qwen3 32B and Sonar share?

0 benchmarks have published results for both models. Qwen3 32B has 26 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper