Model comparison

MiMo-V2-Pro vs Qwen3.5 Plus

MiMo-V2-Pro and Qwen3.5 Plus score almost the same on the Noometry Index (43.0 vs 42.9), so choose on price, context window or the category you care about most.

Last verified . 3 shared benchmarks.

MiMo-V2-Pro Xiaomi

43.0

Rank #103 Confirmed

Qwen3.5 Plus Alibaba (Qwen)

42.9

Rank #106 Confirmed

Summary

  • They share 3 benchmarks with published results for both. MiMo-V2-Pro scores higher in 0 categories and Qwen3.5 Plus in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Qwen3.5 Plus leads 32.8 to 22.1.
  • The biggest single-benchmark swing is CL-bench Life: 6.9% for MiMo-V2-Pro and 12.4% for Qwen3.5 Plus.
  • MiMo-V2-Pro is cheaper at $0.43 / $0.87 per million input/output tokens, against $0.40 / $2.40 for Qwen3.5 Plus.
  • MiMo-V2-Pro accepts more context: 1.05M tokens versus 1M.

Side by side

MiMo-V2-Pro and Qwen3.5 Plus specifications
MiMo-V2-ProQwen3.5 Plus
ProviderXiaomiAlibaba (Qwen)
Noometry Index43.042.9
Released2026-03-182026-02-16
WeightsProprietaryProprietary
Context window1.05M1M
Max output131K66K
Input $ / M tokens$0.43$0.40
Output $ / M tokens$0.87$2.40
Results tracked2315

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

MiMo-V2-Pro: 43.8 (#83), Qwen3.5 Plus: —

Coding benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
ALE-Bench785.17621.92
LMArena WebDev1433—
LMArena Coding1476—

Agentic & Tool Use Not comparable

MiMo-V2-Pro: —, Qwen3.5 Plus: —

Agentic & Tool Use benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
Vending-Bench 2—0.54

Reasoning Qwen3.5 Plus leads

MiMo-V2-Pro: 22.1 (#206), Qwen3.5 Plus: 32.8 (#74)

Reasoning benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
NYT Connections (extended)25.8%—
Chess Puzzles—22%
Thematic Generalization45.9%—
LMArena Hard Prompts1457—
Mystery Game Puzzles—17%
DTBench—80.5%
LMCA—36.4%
Epoch Capabilities Index—146.78

Math Qwen3.5 Plus leads

MiMo-V2-Pro: 39.5 (#102), Qwen3.5 Plus: 49.6 (#61)

Math benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
OTIS Mock AIME 2024-2025—86.7%
LMArena Math1447—
FrontierMath (Feb 2025 set)—21%
FrontierMath Tier 4 (v1)—2.1%

Knowledge Qwen3.5 Plus leads

MiMo-V2-Pro: 41.4 (#111), Qwen3.5 Plus: 46.0 (#83)

Knowledge benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
GPQA Diamond—84.8%
SimpleQA Verified—25.4%
Vectara Hallucination Rate—10.7%
LMArena Expert1478—

Multilingual Not comparable

MiMo-V2-Pro: 52.7 (#81), Qwen3.5 Plus: —

Multilingual benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
LMArena Non-English1416—
LMArena Chinese1456—
LMArena French1469—
LMArena German1417—
LMArena Japanese1366—
LMArena Korean1400—
LMArena Russian1427—
LMArena Spanish1457—

Instruction Following Not comparable

MiMo-V2-Pro: 76.0 (#49), Qwen3.5 Plus: —

Instruction Following benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
LMArena Instruction Following1445—

Long Context Qwen3.5 Plus leads

MiMo-V2-Pro: 41.5 (#138), Qwen3.5 Plus: 43.0 (#113)

Long Context benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
CL-bench15.7%19.8%
CL-bench Life6.9%12.4%
LMArena Longer Query1455—

Writing & Preference Not comparable

MiMo-V2-Pro: 62.8 (#70), Qwen3.5 Plus: —

Writing & Preference benchmarks
BenchmarkMiMo-V2-ProQwen3.5 Plus
LMArena Text1436—
LMArena Creative Writing1415—
LMArena Multi-Turn1456—

Frequently asked questions

Is MiMo-V2-Pro better than Qwen3.5 Plus?

MiMo-V2-Pro and Qwen3.5 Plus score almost the same on the Noometry Index (43.0 vs 42.9), so choose on price, context window or the category you care about most.

Which is cheaper, MiMo-V2-Pro or Qwen3.5 Plus?

MiMo-V2-Pro is cheaper. It lists at $0.43 per million input tokens and $0.87 per million output tokens; Qwen3.5 Plus lists at $0.40 and $2.40.

Which has the bigger context window?

MiMo-V2-Pro does, with 1.05M tokens against 1M.

How many benchmarks do MiMo-V2-Pro and Qwen3.5 Plus share?

3 benchmarks have published results for both models. MiMo-V2-Pro has 23 scored results on Noometry and Qwen3.5 Plus has 15.

Related comparisons

Go deeper