Model comparison

MiniMax-M2 vs Qwen Plus

MiniMax-M2 and Qwen Plus score almost the same on the Noometry Index (37.4 vs 37.1), so choose on price, context window or the category you care about most.

Last verified . 13 shared benchmarks.

MiniMax-M2 MiniMax

37.4

Rank #204 Confirmed

Qwen Plus Alibaba (Qwen)

37.1

Rank #210 Confirmed

Summary

  • They share 13 benchmarks with published results for both. MiniMax-M2 scores higher in 7 categories and Qwen Plus in 1 category; 4 gaps are clear of the uncertainty.
  • The widest gap is in math, where MiniMax-M2 leads 37.3 to 23.3.
  • The biggest single-benchmark swing is Kagi LLM Benchmark: 57.8% for MiniMax-M2 and 63.3% for Qwen Plus.
  • MiniMax-M2 is cheaper at $0.30 / $1.20 per million input/output tokens, against $0.40 / $1.20 for Qwen Plus.
  • Qwen Plus accepts more context: 1M tokens versus 205K.
  • MiniMax-M2 has downloadable open weights; the other is API-only.

Side by side

MiniMax-M2 and Qwen Plus specifications
MiniMax-M2Qwen Plus
ProviderMiniMaxAlibaba (Qwen)
Noometry Index37.437.1
Released2025-10-272024-01-25
WeightsOpenProprietary
Context window205K1M
Max output131K33K
Input $ / M tokens$0.30$0.40
Output $ / M tokens$1.20$1.20
Results tracked2120

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

MiniMax-M2: 39.3 (#159), Qwen Plus: 38.9 (#167)

Coding benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Coding13701328
SWE-bench Verified (bash only)61%—
LMArena WebDev1297—

Agentic & Tool Use Not comparable

MiniMax-M2: 25.1 (#109), Qwen Plus: —

Agentic & Tool Use benchmarks
BenchmarkMiniMax-M2Qwen Plus
Terminal-Bench30%—
Vending-Bench 2160.6—

Reasoning Qwen Plus leads

MiniMax-M2: 19.4 (#258), Qwen Plus: 28.4 (#107)

Reasoning benchmarks
BenchmarkMiniMax-M2Qwen Plus
Kagi LLM Benchmark57.8%63.3%
LMArena Hard Prompts13571317
NYT Connections (extended)14.8%—
DTBench—81.1%
LMCA—24%

Math MiniMax-M2 leads

MiniMax-M2: 37.3 (#160), Qwen Plus: 23.3 (#271)

Math benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Math13521326
OTIS Mock AIME 2024-2025—17.8%
MATH Level 5—65.3%
FrontierMath (Feb 2025 set)—1.7%

Knowledge MiniMax-M2 leads

MiniMax-M2: 37.0 (#163), Qwen Plus: 27.4 (#251)

Knowledge benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Expert13371328
GPQA Diamond—48.1%

Multilingual Too close to call

MiniMax-M2: 45.3 (#171), Qwen Plus: 45.1 (#175)

Multilingual benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Non-English13131310
LMArena Chinese13661347
LMArena Russian13311323
LMArena French1335—
LMArena German1355—
LMArena Japanese—1251
LMArena Spanish1326—

Instruction Following MiniMax-M2 leads

MiniMax-M2: 70.2 (#166), Qwen Plus: 68.8 (#181)

Instruction Following benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Instruction Following13281303

Long Context Too close to call

MiniMax-M2: 40.5 (#153), Qwen Plus: 40.3 (#158)

Long Context benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Longer Query13311324

Writing & Preference Too close to call

MiniMax-M2: 53.0 (#162), Qwen Plus: 52.2 (#176)

Writing & Preference benchmarks
BenchmarkMiniMax-M2Qwen Plus
LMArena Text13401326
LMArena Creative Writing12861293
LMArena Multi-Turn13611336

Frequently asked questions

Is MiniMax-M2 better than Qwen Plus?

MiniMax-M2 and Qwen Plus score almost the same on the Noometry Index (37.4 vs 37.1), so choose on price, context window or the category you care about most.

Which is cheaper, MiniMax-M2 or Qwen Plus?

MiniMax-M2 is cheaper. It lists at $0.30 per million input tokens and $1.20 per million output tokens; Qwen Plus lists at $0.40 and $1.20.

Is MiniMax-M2 or Qwen Plus better for coding?

They score almost the same on coding (39.3 vs 38.9); test both on your own repository before choosing.

Which has the bigger context window?

Qwen Plus does, with 1M tokens against 205K.

How many benchmarks do MiniMax-M2 and Qwen Plus share?

13 benchmarks have published results for both models. MiniMax-M2 has 21 scored results on Noometry and Qwen Plus has 20.

Related comparisons

Go deeper