Model comparison

Qwen3.5 397B-A17B vs Qwen3.5 Max Preview

Qwen3.5 397B-A17B and Qwen3.5 Max Preview score almost the same on the Noometry Index (46.0 vs 45.3), so choose on price, context window or the category you care about most.

Last verified . 17 shared benchmarks.

Qwen3.5 397B-A17B Alibaba (Qwen)

46.0

Rank #67 Confirmed

Qwen3.5 Max Preview Alibaba (Qwen)

45.3

Rank #71 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Qwen3.5 397B-A17B scores higher in 3 categories and Qwen3.5 Max Preview in 5 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Qwen3.5 397B-A17B leads 53.3 to 41.8.
  • Qwen3.5 397B-A17B has downloadable open weights; the other is API-only.

Side by side

Qwen3.5 397B-A17B and Qwen3.5 Max Preview specifications
Qwen3.5 397B-A17BQwen3.5 Max Preview
ProviderAlibaba (Qwen)Alibaba (Qwen)
Noometry Index46.045.3
Released2026-02-01—
WeightsOpenProprietary
Context window262K—
Max output66K—
Input $ / M tokens$0.60—
Output $ / M tokens$3.60—
Results tracked3617

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.5 Max Preview leads

Qwen3.5 397B-A17B: 42.0 (#114), Qwen3.5 Max Preview: 44.0 (#77)

Coding benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Coding14651487
LMArena WebDev1400—

Agentic & Tool Use Not comparable

Qwen3.5 397B-A17B: 33.3 (#53), Qwen3.5 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
APEX-Agents24.9%—
τ²-bench Airline81.5%—
τ²-bench Banking9.8%—
τ²-bench Retail84.4%—
τ²-bench Telecom97.8%—

Reasoning Qwen3.5 397B-A17B leads

Qwen3.5 397B-A17B: 34.5 (#70), Qwen3.5 Max Preview: 30.8 (#84)

Reasoning benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Hard Prompts14481483
Kagi LLM Benchmark73.7%—
NYT Connections (extended)58.9%—
Chess Puzzles13%—
Thematic Generalization65.1%—
Mystery Game Puzzles18%—
DTBench87.5%—
LMCA37.9%—
Epoch Capabilities Index146.65—

Math Qwen3.5 397B-A17B leads

Qwen3.5 397B-A17B: 46.1 (#73), Qwen3.5 Max Preview: 40.1 (#94)

Math benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Math14541474
FrontierMath (Tiers 1-3)31.2%—
OTIS Mock AIME 2024-202588.9%—

Knowledge Qwen3.5 397B-A17B leads

Qwen3.5 397B-A17B: 53.3 (#58), Qwen3.5 Max Preview: 41.8 (#107)

Knowledge benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Expert14621489
GPQA Diamond86.4%—

Multimodal Not comparable

Qwen3.5 397B-A17B: 40.7 (#44), Qwen3.5 Max Preview: —

Multimodal benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Vision1263—

Multilingual Qwen3.5 Max Preview leads

Qwen3.5 397B-A17B: 53.7 (#59), Qwen3.5 Max Preview: 56.2 (#22)

Multilingual benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Non-English14301465
LMArena Chinese15001534
LMArena French14611484
LMArena German14471487
LMArena Japanese14261495
LMArena Korean13841438
LMArena Russian14291471
LMArena Spanish14411470

Instruction Following Qwen3.5 Max Preview leads

Qwen3.5 397B-A17B: 75.0 (#77), Qwen3.5 Max Preview: 77.0 (#31)

Instruction Following benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Instruction Following14241467

Long Context Qwen3.5 Max Preview leads

Qwen3.5 397B-A17B: 44.1 (#74), Qwen3.5 Max Preview: 45.2 (#45)

Long Context benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Longer Query14421476

Writing & Preference Qwen3.5 Max Preview leads

Qwen3.5 397B-A17B: 62.3 (#79), Qwen3.5 Max Preview: 66.0 (#41)

Writing & Preference benchmarks
BenchmarkQwen3.5 397B-A17BQwen3.5 Max Preview
LMArena Text14381470
LMArena Creative Writing14011464
LMArena Multi-Turn14461478
EQ-Bench Creative Writing1478—

Frequently asked questions

Is Qwen3.5 397B-A17B better than Qwen3.5 Max Preview?

Qwen3.5 397B-A17B and Qwen3.5 Max Preview score almost the same on the Noometry Index (46.0 vs 45.3), so choose on price, context window or the category you care about most.

Is Qwen3.5 397B-A17B or Qwen3.5 Max Preview better for coding?

Qwen3.5 Max Preview scores higher on coding benchmarks: 44.0 versus 42.0 in the Noometry coding category.

How many benchmarks do Qwen3.5 397B-A17B and Qwen3.5 Max Preview share?

17 benchmarks have published results for both models. Qwen3.5 397B-A17B has 36 scored results on Noometry and Qwen3.5 Max Preview has 17.

Related comparisons

Go deeper