Model comparison

Mistral Large 4 vs Sonar

Mistral Large 4 is the stronger model overall, scoring 43.1 to 38.5 on the Noometry Index.

Last verified . 0 shared benchmarks.

Mistral Large 4 Mistral AI

43.1

Rank #99 Confirmed

Sonar Perplexity

38.5

Rank #187 Confirmed

Summary

  • The widest gap is in coding, where Mistral Large 4 leads 48.6 to 35.7.
  • Both cost about the same: $0.68 input and $2.09 output per million tokens.
  • Mistral Large 4 accepts more context: 1.05M tokens versus 128K.

Side by side

Mistral Large 4 and Sonar specifications
Mistral Large 4Sonar
ProviderMistral AIPerplexity
Noometry Index43.138.5
Released2026-10-062024-01-01
WeightsProprietaryProprietary
Context window1.05M128K
Max output262K4K
Input $ / M tokens$0.68$1
Output $ / M tokens$2.09$1
Results tracked157

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Mistral Large 4 leads

Mistral Large 4: 48.6 (#57), Sonar: 35.7 (#221)

Coding benchmarks
BenchmarkMistral Large 4Sonar
LMArena WebDev1541—
LiveBench Coding—35.1%
LMArena Coding1475—

Reasoning Mistral Large 4 leads

Mistral Large 4: 22.5 (#192), Sonar: 21.1 (#227)

Reasoning benchmarks
BenchmarkMistral Large 4Sonar
NYT Connections (extended)27.4%—
LiveBench Reasoning—46.3%
LMArena Hard Prompts1444—
LiveBench Data Analysis—37.9%
LiveBench—46.9%

Math Mistral Large 4 leads

Mistral Large 4: 40.4 (#91), Sonar: 33.7 (#200)

Math benchmarks
BenchmarkMistral Large 4Sonar
LiveBench Math—41.6%
LMArena Math1488—

Knowledge Not comparable

Mistral Large 4: 36.6 (#166), Sonar: —

Knowledge benchmarks
BenchmarkMistral Large 4Sonar
SimpleQA Verified20%—
LMArena Expert1447—

Multilingual Not comparable

Mistral Large 4: 52.6 (#82), Sonar: —

Multilingual benchmarks
BenchmarkMistral Large 4Sonar
LMArena Non-English1415—
LMArena Chinese1491—
LMArena Russian1414—

Instruction Following Mistral Large 4 leads

Mistral Large 4: 75.0 (#76), Sonar: 71.4 (#150)

Instruction Following benchmarks
BenchmarkMistral Large 4Sonar
LiveBench Instruction Following—76.2%
LMArena Instruction Following1424—

Long Context Not comparable

Mistral Large 4: 43.6 (#89), Sonar: —

Long Context benchmarks
BenchmarkMistral Large 4Sonar
LMArena Longer Query1429—

Writing & Preference Mistral Large 4 leads

Mistral Large 4: 60.4 (#97), Sonar: 52.6 (#167)

Writing & Preference benchmarks
BenchmarkMistral Large 4Sonar
LMArena Text1427—
LMArena Creative Writing1361—
LMArena Multi-Turn1424—
LiveBench Language—44.1%

Frequently asked questions

Is Mistral Large 4 better than Sonar?

Mistral Large 4 is the stronger model overall, scoring 43.1 to 38.5 on the Noometry Index.

Which is cheaper, Mistral Large 4 or Sonar?

Sonar is cheaper. It lists at $1 per million input tokens and $1 per million output tokens; Mistral Large 4 lists at $0.68 and $2.09.

Is Mistral Large 4 or Sonar better for coding?

Mistral Large 4 scores higher on coding benchmarks: 48.6 versus 35.7 in the Noometry coding category.

Which has the bigger context window?

Mistral Large 4 does, with 1.05M tokens against 128K.

How many benchmarks do Mistral Large 4 and Sonar share?

0 benchmarks have published results for both models. Mistral Large 4 has 15 scored results on Noometry and Sonar has 7.

Related comparisons

Go deeper