Model comparison

Laguna M.1 vs Qwen Max

Qwen Max is the stronger model overall, scoring 34.7 to 32.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Laguna M.1 Poolside

32.5

Rank #256 Reported

Qwen Max Alibaba (Qwen)

34.7

Rank #230 Confirmed

Summary

  • The widest gap is in coding, where Laguna M.1 leads 36.6 to 30.7.
  • Laguna M.1 accepts more context: 262K tokens versus 33K.
  • Laguna M.1 has downloadable open weights; the other is API-only.

Side by side

Laguna M.1 and Qwen Max specifications
Laguna M.1Qwen Max
ProviderPoolsideAlibaba (Qwen)
Noometry Index32.534.7
Released2026-04-282024-04-03
WeightsOpenProprietary
Context window262K33K
Max output33K8K
Input $ / M tokens—$1.60
Output $ / M tokens—$6.40
Results tracked323

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Laguna M.1 leads

Laguna M.1: 36.6 (#204), Qwen Max: 30.7 (#292)

Coding benchmarks
BenchmarkLaguna M.1Qwen Max
Aider Polyglot—21.8%
LMArena WebDev1349—
LMArena Coding—1288

Reasoning Qwen Max leads

Laguna M.1: 23.1 (#184), Qwen Max: 25.1 (#151)

Reasoning benchmarks
BenchmarkLaguna M.1Qwen Max
LMArena Hard Prompts—1269
Surface Evolver Bench15.6%—

Math Qwen Max leads

Laguna M.1: 21.1 (#283), Qwen Max: 22.3 (#276)

Math benchmarks
BenchmarkLaguna M.1Qwen Max
OTIS Mock AIME 2024-2025—16.1%
ProofBench0%—
LMArena Math—1275
MATH Level 5—67.2%
FrontierMath (Feb 2025 set)—1%

Knowledge Not comparable

Laguna M.1: —, Qwen Max: 30.3 (#228)

Knowledge benchmarks
BenchmarkLaguna M.1Qwen Max
GPQA Diamond—56.1%
LMArena Expert—1248

Multilingual Not comparable

Laguna M.1: —, Qwen Max: 41.8 (#202)

Multilingual benchmarks
BenchmarkLaguna M.1Qwen Max
LMArena Non-English—1263
LMArena Chinese—1254
LMArena French—1330
LMArena German—1254
LMArena Japanese—1205
LMArena Korean—1142
LMArena Russian—1274
LMArena Spanish—1290

Instruction Following Not comparable

Laguna M.1: —, Qwen Max: 66.5 (#208)

Instruction Following benchmarks
BenchmarkLaguna M.1Qwen Max
LMArena Instruction Following—1262

Long Context Not comparable

Laguna M.1: —, Qwen Max: 39.4 (#180)

Long Context benchmarks
BenchmarkLaguna M.1Qwen Max
Fiction.LiveBench—66.7%
LMArena Longer Query—1288

Writing & Preference Not comparable

Laguna M.1: —, Qwen Max: 47.8 (#205)

Writing & Preference benchmarks
BenchmarkLaguna M.1Qwen Max
LMArena Text—1282
LMArena Creative Writing—1248
LMArena Multi-Turn—1277

Frequently asked questions

Is Laguna M.1 better than Qwen Max?

Qwen Max is the stronger model overall, scoring 34.7 to 32.5 on the Noometry Index.

Is Laguna M.1 or Qwen Max better for coding?

Laguna M.1 scores higher on coding benchmarks: 36.6 versus 30.7 in the Noometry coding category.

Which has the bigger context window?

Laguna M.1 does, with 262K tokens against 33K.

How many benchmarks do Laguna M.1 and Qwen Max share?

0 benchmarks have published results for both models. Laguna M.1 has 3 scored results on Noometry and Qwen Max has 23.

Related comparisons

Go deeper