Model comparison

DeepSeek LLM 67B vs Pixtral Large

Pixtral Large is the stronger model overall, scoring 32.2 to 24.9 on the Noometry Index.

Last verified . 0 shared benchmarks.

DeepSeek LLM 67B DeepSeek

24.9

Rank #347 Confirmed

Pixtral Large Mistral AI

32.2

Rank #259 Reported

Summary

  • The widest gap is in reasoning, where Pixtral Large leads 21.7 to 16.5.

Side by side

DeepSeek LLM 67B and Pixtral Large specifications
DeepSeek LLM 67BPixtral Large
ProviderDeepSeekMistral AI
Noometry Index24.932.2
Released2023-11-292024-11-01
WeightsOpenOpen
Context window—128K
Max output—128K
Input $ / M tokens—$2
Output $ / M tokens—$6
Results tracked153

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

DeepSeek LLM 67B: 31.9 (#278), Pixtral Large: —

Coding benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
LMArena Coding1096—

Reasoning Pixtral Large leads

DeepSeek LLM 67B: 16.5 (#304), Pixtral Large: 21.7 (#218)

Reasoning benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
Chess Puzzles0%—
EnigmaEval—0.8%
LMArena Hard Prompts1070—
Epoch Capabilities Index110.5—

Math Not comparable

DeepSeek LLM 67B: 8.7 (#324), Pixtral Large: —

Math benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
OTIS Mock AIME 2024-20250.8%—
LMArena Math1108—
MATH Level 56.4%—

Knowledge Not comparable

DeepSeek LLM 67B: 7.0 (#313), Pixtral Large: —

Knowledge benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
GPQA Diamond24.6%—

Multimodal Not comparable

DeepSeek LLM 67B: —, Pixtral Large: 30.6 (#111)

Multimodal benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
LMArena Vision—1089

Multilingual Not comparable

DeepSeek LLM 67B: 29.4 (#267), Pixtral Large: —

Multilingual benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
LMArena Non-English1073—
LMArena Chinese1132—

Instruction Following Not comparable

DeepSeek LLM 67B: 55.4 (#277), Pixtral Large: —

Instruction Following benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
LMArena Instruction Following1079—

Long Context Not comparable

DeepSeek LLM 67B: 33.1 (#265), Pixtral Large: —

Long Context benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
LMArena Longer Query1092—

Writing & Preference Pixtral Large leads

DeepSeek LLM 67B: 31.6 (#282), Pixtral Large: 32.9 (#278)

Writing & Preference benchmarks
BenchmarkDeepSeek LLM 67BPixtral Large
LMArena Text1105—
LMArena Creative Writing1067—
EQ-Bench Creative Writing—988
LMArena Multi-Turn1082—

Frequently asked questions

Is DeepSeek LLM 67B better than Pixtral Large?

Pixtral Large is the stronger model overall, scoring 32.2 to 24.9 on the Noometry Index.

How many benchmarks do DeepSeek LLM 67B and Pixtral Large share?

0 benchmarks have published results for both models. DeepSeek LLM 67B has 15 scored results on Noometry and Pixtral Large has 3.

Related comparisons

Go deeper