Model comparison

Claude Sonnet 4.6 vs MiMo-V2.6-Pro

Claude Sonnet 4.6 and MiMo-V2.6-Pro score almost the same on the Noometry Index (50.3 vs 50.3), so choose on price, context window or the category you care about most.

Last verified . 19 shared benchmarks.

Claude Sonnet 4.6 Anthropic

50.3

Rank #50 Confirmed

MiMo-V2.6-Pro Xiaomi

50.3

Rank #49 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Claude Sonnet 4.6 scores higher in 4 categories and MiMo-V2.6-Pro in 6 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in coding, where MiMo-V2.6-Pro leads 55.5 to 46.3.
  • The biggest single-benchmark swing is ProofBench: 45% for Claude Sonnet 4.6 and 70% for MiMo-V2.6-Pro.
  • MiMo-V2.6-Pro is cheaper at $0.43 / $0.87 per million input/output tokens, against $3 / $15 for Claude Sonnet 4.6.
  • MiMo-V2.6-Pro accepts more context: 1.05M tokens versus 1M.
  • MiMo-V2.6-Pro has downloadable open weights; the other is API-only.

Side by side

Claude Sonnet 4.6 and MiMo-V2.6-Pro specifications
Claude Sonnet 4.6MiMo-V2.6-Pro
ProviderAnthropicXiaomi
Noometry Index50.350.3
Released2026-02-172026-09-21
WeightsProprietaryOpen
Context window1M1.05M
Max output128K131K
Input $ / M tokens$3$0.43
Output $ / M tokens$15$0.87
Results tracked5719

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiMo-V2.6-Pro leads

Claude Sonnet 4.6: 46.3 (#67), MiMo-V2.6-Pro: 55.5 (#23)

Coding benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena WebDev15221629
SciCode46.8%60.9%
LMArena Coding15041534
ALE-Bench1,3271,158
SWE-bench Verified75.2%—
DeepSWE29.9%—
FrontierCode24.3%—
WeirdML66.1%—

Agentic & Tool Use Claude Sonnet 4.6 leads

Claude Sonnet 4.6: 39.1 (#28), MiMo-V2.6-Pro: 37.5 (#35)

Agentic & Tool Use benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
APEX-Agents43%59.5%
Terminal-Bench53.4%—
OSWorld 2.09.3%—
DeepResearch Bench54.9%—
OSWorld72.1%—
ExploitBench23.6%—
GBAEval48.8%—
GDP.pdf18%—
LMArena Search1221—
Vending-Bench 27,204—

Reasoning Claude Sonnet 4.6 leads

Claude Sonnet 4.6: 46.1 (#45), MiMo-V2.6-Pro: 43.1 (#50)

Reasoning benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
CritPt3.1%26.6%
LMArena Hard Prompts14841512
ARC-AGI-260.4%—
NYT Connections (extended)80.9%—
ARC-AGI-186.5%—
Chess Puzzles13%—
Thematic Generalization76.3%—
Mystery Game Puzzles16%—
DTBench89.9%—
LMCA46.5%—
Epoch Capabilities Index152.24—
ForecastBench62—

Math MiMo-V2.6-Pro leads

Claude Sonnet 4.6: 52.9 (#49), MiMo-V2.6-Pro: 54.5 (#45)

Math benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
ProofBench45%70%
LMArena Math14621494
OTIS Mock AIME 2024-202585.8%—
FrontierMath (Feb 2025 set)32.4%—
FrontierMath Tier 4 (v1)8.3%—

Knowledge Claude Sonnet 4.6 leads

Claude Sonnet 4.6: 51.7 (#65), MiMo-V2.6-Pro: 43.5 (#92)

Knowledge benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena Expert15001543
GPQA Diamond87.4%—
SimpleQA Verified35.5%—
Vectara Hallucination Rate10.6%—

Multimodal MiMo-V2.6-Pro leads

Claude Sonnet 4.6: 38.0 (#68), MiMo-V2.6-Pro: 40.8 (#43)

Multimodal benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena Vision12831264
Blueprint-Bench 26.7%—
LMArena Document1482—

Multilingual MiMo-V2.6-Pro leads

Claude Sonnet 4.6: 54.4 (#41), MiMo-V2.6-Pro: 56.9 (#14)

Multilingual benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena Non-English14401474
LMArena Chinese14911529
LMArena Russian14401480
LMArena French1465—
LMArena German1428—
LMArena Japanese1420—
LMArena Korean1411—
LMArena Spanish1464—

Instruction Following Too close to call

Claude Sonnet 4.6: 77.4 (#25), MiMo-V2.6-Pro: 78.2 (#12)

Instruction Following benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena Instruction Following14751493

Long Context Too close to call

Claude Sonnet 4.6: 45.3 (#44), MiMo-V2.6-Pro: 46.0 (#27)

Long Context benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena Longer Query14791501

Writing & Preference Claude Sonnet 4.6 leads

Claude Sonnet 4.6: 70.2 (#22), MiMo-V2.6-Pro: 66.8 (#33)

Writing & Preference benchmarks
BenchmarkClaude Sonnet 4.6MiMo-V2.6-Pro
LMArena Text14581492
LMArena Creative Writing14351468
LMArena Multi-Turn14641464
EQ-Bench Creative Writing1810—
EQ-Bench 41207—

Frequently asked questions

Is Claude Sonnet 4.6 better than MiMo-V2.6-Pro?

Claude Sonnet 4.6 and MiMo-V2.6-Pro score almost the same on the Noometry Index (50.3 vs 50.3), so choose on price, context window or the category you care about most.

Which is cheaper, Claude Sonnet 4.6 or MiMo-V2.6-Pro?

MiMo-V2.6-Pro is cheaper. It lists at $0.43 per million input tokens and $0.87 per million output tokens; Claude Sonnet 4.6 lists at $3 and $15.

Is Claude Sonnet 4.6 or MiMo-V2.6-Pro better for coding?

MiMo-V2.6-Pro scores higher on coding benchmarks: 55.5 versus 46.3 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2.6-Pro does, with 1.05M tokens against 1M.

How many benchmarks do Claude Sonnet 4.6 and MiMo-V2.6-Pro share?

19 benchmarks have published results for both models. Claude Sonnet 4.6 has 57 scored results on Noometry and MiMo-V2.6-Pro has 19.

Related comparisons

Go deeper