Model comparison

Llama 3.1 Nemotron 70b Instruct vs MiniMax M1

MiniMax M1 is the stronger model overall, scoring 40.3 to 37.6 on the Noometry Index.

Last verified . 12 shared benchmarks.

MiniMax M1 MiniMax

40.3

Rank #150 Confirmed

Summary

  • They share 12 benchmarks with published results for both. Llama 3.1 Nemotron 70b Instruct scores higher in 0 categories and MiniMax M1 in 8 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in multilingual, where MiniMax M1 leads 45.8 to 40.5.

Side by side

Llama 3.1 Nemotron 70b Instruct and MiniMax M1 specifications
Llama 3.1 Nemotron 70b InstructMiniMax M1
ProviderNVIDIAMiniMax
Noometry Index37.640.3
Released2024-12-182025-06-13
WeightsOpenOpen
Context window—1M
Max output—40K
Input $ / M tokens—$0.55
Output $ / M tokens—$2.20
Results tracked1418

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 35.9 (#216), MiniMax M1: 39.9 (#153)

Coding benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Coding12721359
BigCodeBench Instruct38.7%—
BigCodeBench Complete48.2%—

Reasoning MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 25.0 (#152), MiniMax M1: 26.9 (#126)

Reasoning benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Hard Prompts12661339

Math MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 35.5 (#182), MiniMax M1: 37.5 (#151)

Math benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Math12711361

Knowledge MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 34.1 (#199), MiniMax M1: 36.4 (#170)

Knowledge benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Expert12421317

Multilingual MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 40.5 (#217), MiniMax M1: 45.8 (#163)

Multilingual benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Non-English12451319
LMArena Chinese12631360
LMArena Russian12271329
LMArena French—1370
LMArena German—1350
LMArena Japanese—1217
LMArena Korean—1266
LMArena Spanish—1353

Instruction Following MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 65.9 (#213), MiniMax M1: 69.3 (#174)

Instruction Following benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Instruction Following12521312

Long Context MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 37.6 (#215), MiniMax M1: 41.4 (#141)

Long Context benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Longer Query12381326
Fiction.LiveBench—69.4%

Writing & Preference MiniMax M1 leads

Llama 3.1 Nemotron 70b Instruct: 48.4 (#203), MiniMax M1: 53.1 (#161)

Writing & Preference benchmarks
BenchmarkLlama 3.1 Nemotron 70b InstructMiniMax M1
LMArena Text12831343
LMArena Creative Writing12691298
LMArena Multi-Turn12751335

Frequently asked questions

Is Llama 3.1 Nemotron 70b Instruct better than MiniMax M1?

MiniMax M1 is the stronger model overall, scoring 40.3 to 37.6 on the Noometry Index.

Is Llama 3.1 Nemotron 70b Instruct or MiniMax M1 better for coding?

MiniMax M1 scores higher on coding benchmarks: 39.9 versus 35.9 in the Noometry coding category.

How many benchmarks do Llama 3.1 Nemotron 70b Instruct and MiniMax M1 share?

12 benchmarks have published results for both models. Llama 3.1 Nemotron 70b Instruct has 14 scored results on Noometry and MiniMax M1 has 18.

Related comparisons

Go deeper