Model comparison

GPT-6 Sol vs MiniMax-M2.7

GPT-6 Sol is the stronger model overall, scoring 61.8 to 37.7 on the Noometry Index. MiniMax-M2.7 costs 7.6× less per token, which makes it the better buy when GPT-6 Sol's lead doesn't matter for your workload.

Last verified . 25 shared benchmarks.

GPT-6 Sol OpenAI

61.8

Rank #12 Confirmed

MiniMax-M2.7 MiniMax

37.7

Rank #196 Confirmed

Summary

  • They share 25 benchmarks with published results for both. GPT-6 Sol scores higher in 8 categories and MiniMax-M2.7 in 1 category; 6 gaps are clear of the uncertainty.
  • The widest gap is in math, where GPT-6 Sol leads 87.2 to 25.9.
  • The biggest single-benchmark swing is ProofBench: 83% for GPT-6 Sol and 3% for MiniMax-M2.7.
  • MiniMax-M2.7 is cheaper at $0.30 / $1.20 per million input/output tokens, against $2 / $10 for GPT-6 Sol.
  • GPT-6 Sol accepts more context: 1.05M tokens versus 205K.
  • MiniMax-M2.7 has downloadable open weights; the other is API-only.

Side by side

GPT-6 Sol and MiniMax-M2.7 specifications
GPT-6 SolMiniMax-M2.7
ProviderOpenAIMiniMax
Noometry Index61.837.7
Released2026-09-222026-03-18
WeightsProprietaryOpen
Context window1.05M205K
Max output128K131K
Input $ / M tokens$2$0.30
Output $ / M tokens$10$1.20
Results tracked4530

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding GPT-6 Sol leads

GPT-6 Sol: 60.1 (#11), MiniMax-M2.7: 41.8 (#120)

Coding benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
LMArena WebDev16881398
SciCode57.6%47%
LMArena Coding14471454
ALE-Bench2,462599.25
DeepSWE68.8%—
FrontierCode49.3%—
WeirdML—37%

Agentic & Tool Use GPT-6 Sol leads

GPT-6 Sol: 37.2 (#36), MiniMax-M2.7: 25.1 (#111)

Agentic & Tool Use benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
Terminal-Bench—45.1%
APEX-Agents54.3%—
ExploitBench—13.3%
GBAEval—0%
GDP.pdf26.4%—
Vending-Bench 214,428—

Reasoning GPT-6 Sol leads

GPT-6 Sol: 74.0 (#9), MiniMax-M2.7: 19.7 (#253)

Reasoning benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
NYT Connections (extended)90.1%24.7%
CritPt30.9%0.6%
LMArena Hard Prompts14181422
Epoch Capabilities Index162.72145.85
ARC-AGI-289.6%—
ARC-AGI-195.5%—
Thematic Generalization—39.3%
EBR-Bench53.3%—
Mystery Game Puzzles56%—
DTBench97.3%—
LMCA59.1%—

Math GPT-6 Sol leads

GPT-6 Sol: 87.2 (#7), MiniMax-M2.7: 25.9 (#263)

Math benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
ProofBench83%3%
LMArena Math14021420
FrontierMath (Tiers 1-3)89.8%—
FrontierMath Tier 490%—
OTIS Mock AIME 2024-2025100%—

Knowledge GPT-6 Sol leads

GPT-6 Sol: 64.8 (#15), MiniMax-M2.7: 37.7 (#152)

Knowledge benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
Vectara Hallucination Rate6.5%12.9%
LMArena Expert14391444
GPQA Diamond94.3%—
SimpleQA Verified60.7%—

Multimodal Not comparable

GPT-6 Sol: 47.6 (#10), MiniMax-M2.7: —

Multimodal benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
LMArena Vision1245—
Blueprint-Bench 236.9%—
Furniture Assembly58.3%—

Multilingual Too close to call

GPT-6 Sol: 50.5 (#118), MiniMax-M2.7: 50.3 (#123)

Multilingual benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
LMArena Non-English13851382
LMArena Chinese14051441
LMArena French14101421
LMArena German13901398
LMArena Japanese13851262
LMArena Korean13411313
LMArena Russian14011383
LMArena Spanish13841403

Instruction Following Too close to call

GPT-6 Sol: 74.5 (#94), MiniMax-M2.7: 74.1 (#103)

Instruction Following benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
LMArena Instruction Following14121405

Long Context Too close to call

GPT-6 Sol: 43.1 (#108), MiniMax-M2.7: 43.3 (#99)

Long Context benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
LMArena Longer Query14111419

Writing & Preference GPT-6 Sol leads

GPT-6 Sol: 71.9 (#18), MiniMax-M2.7: 58.9 (#112)

Writing & Preference benchmarks
BenchmarkGPT-6 SolMiniMax-M2.7
LMArena Text13951405
LMArena Creative Writing13781354
LMArena Multi-Turn14121412
EQ-Bench Creative Writing2125—

Frequently asked questions

Is GPT-6 Sol better than MiniMax-M2.7?

GPT-6 Sol is the stronger model overall, scoring 61.8 to 37.7 on the Noometry Index. MiniMax-M2.7 costs 7.6× less per token, which makes it the better buy when GPT-6 Sol's lead doesn't matter for your workload.

Which is cheaper, GPT-6 Sol or MiniMax-M2.7?

MiniMax-M2.7 is cheaper. It lists at $0.30 per million input tokens and $1.20 per million output tokens; GPT-6 Sol lists at $2 and $10.

Is GPT-6 Sol or MiniMax-M2.7 better for coding?

GPT-6 Sol scores higher on coding benchmarks: 60.1 versus 41.8 in the Noometry coding category.

Which has the bigger context window?

GPT-6 Sol does, with 1.05M tokens against 205K.

How many benchmarks do GPT-6 Sol and MiniMax-M2.7 share?

25 benchmarks have published results for both models. GPT-6 Sol has 45 scored results on Noometry and MiniMax-M2.7 has 30.

Related comparisons

Go deeper