Model comparison

Hy3 vs MiMo-V2.5

Hy3 and MiMo-V2.5 score almost the same on the Noometry Index (44.2 vs 43.4), so choose on price, context window or the category you care about most.

Last verified . 18 shared benchmarks.

Hy3 Tencent

44.2

Rank #79 Confirmed

MiMo-V2.5 Xiaomi

43.4

Rank #93 Confirmed

Summary

  • They share 18 benchmarks with published results for both. Hy3 scores higher in 4 categories and MiMo-V2.5 in 4 categories; 4 gaps are clear of the uncertainty.
  • The widest gap is in math, where Hy3 leads 40.1 to 36.8.
  • MiMo-V2.5 is cheaper at $0.14 / $0.28 per million input/output tokens, against $0.13 / $0.53 for Hy3.
  • MiMo-V2.5 accepts more context: 1.05M tokens versus 262K.

Side by side

Hy3 and MiMo-V2.5 specifications
Hy3MiMo-V2.5
ProviderTencentXiaomi
Noometry Index44.243.4
Released2026-07-062026-04-22
WeightsOpenOpen
Context window262K1.05M
Max output128K131K
Input $ / M tokens$0.13$0.14
Output $ / M tokens$0.53$0.28
Results tracked1923

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Hy3 leads

Hy3: 46.8 (#63), MiMo-V2.5: 43.9 (#81)

Coding benchmarks
BenchmarkHy3MiMo-V2.5
LMArena WebDev15081438
LMArena Coding14641469
SciCode—43.1%
ALE-Bench—513.95

Reasoning MiMo-V2.5 leads

Hy3: 26.1 (#136), MiMo-V2.5: 28.6 (#101)

Reasoning benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Hard Prompts14471450
NYT Connections (extended)41.2%—
CritPt—3.7%

Math Hy3 leads

Hy3: 40.1 (#93), MiMo-V2.5: 36.8 (#163)

Math benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Math14751436
ProofBench—16%

Knowledge Too close to call

Hy3: 40.8 (#114), MiMo-V2.5: 40.8 (#115)

Knowledge benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Expert14601460

Multimodal Not comparable

Hy3: —, MiMo-V2.5: 39.8 (#54)

Multimodal benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Vision—1247

Multilingual Hy3 leads

Hy3: 53.5 (#65), MiMo-V2.5: 51.9 (#99)

Multilingual benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Non-English14261404
LMArena Chinese14931468
LMArena French14611447
LMArena German14391421
LMArena Japanese13921306
LMArena Korean13951363
LMArena Russian14321395
LMArena Spanish14561416

Instruction Following Too close to call

Hy3: 75.1 (#70), MiMo-V2.5: 75.5 (#60)

Instruction Following benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Instruction Following14261434

Long Context Too close to call

Hy3: 44.1 (#75), MiMo-V2.5: 44.2 (#73)

Long Context benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Longer Query14421445

Writing & Preference Too close to call

Hy3: 62.2 (#81), MiMo-V2.5: 61.6 (#86)

Writing & Preference benchmarks
BenchmarkHy3MiMo-V2.5
LMArena Text14391428
LMArena Creative Writing14021393
LMArena Multi-Turn14361445

Frequently asked questions

Is Hy3 better than MiMo-V2.5?

Hy3 and MiMo-V2.5 score almost the same on the Noometry Index (44.2 vs 43.4), so choose on price, context window or the category you care about most.

Which is cheaper, Hy3 or MiMo-V2.5?

MiMo-V2.5 is cheaper. It lists at $0.14 per million input tokens and $0.28 per million output tokens; Hy3 lists at $0.13 and $0.53.

Is Hy3 or MiMo-V2.5 better for coding?

Hy3 scores higher on coding benchmarks: 46.8 versus 43.9 in the Noometry coding category.

Which has the bigger context window?

MiMo-V2.5 does, with 1.05M tokens against 262K.

How many benchmarks do Hy3 and MiMo-V2.5 share?

18 benchmarks have published results for both models. Hy3 has 19 scored results on Noometry and MiMo-V2.5 has 23.

Related comparisons

Go deeper