Model comparison

Mixtral 8x22B vs Yi-34B

Mixtral 8x22B and Yi-34B score almost the same on the Noometry Index (27.1 vs 27.8), so choose on price, context window or the category you care about most.

Last verified . 21 shared benchmarks.

Mixtral 8x22B Mistral AI

27.1

Rank #333 Confirmed

Yi-34B 01.AI

27.8

Rank #329 Confirmed

Summary

  • They share 21 benchmarks with published results for both. Mixtral 8x22B scores higher in 6 categories and Yi-34B in 2 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Yi-34B leads 32.3 to 24.2.
  • The biggest single-benchmark swing is GPQA Diamond: 34.1% for Mixtral 8x22B and 14.7% for Yi-34B.

Side by side

Mixtral 8x22B and Yi-34B specifications
Mixtral 8x22BYi-34B
ProviderMistral AI01.AI
Noometry Index27.127.8
Released2024-04-172023-11-02
WeightsOpenOpen
Context window64K—
Max output64K—
Input $ / M tokens$2—
Output $ / M tokens$6—
Results tracked3423

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Yi-34B leads

Mixtral 8x22B: 24.2 (#329), Yi-34B: 32.3 (#274)

Coding benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Coding11661112
WeirdML3.2%—
BigCodeBench Instruct40.6%—
BigCodeBench Complete50.2%—
HumanEval+72%—
MBPP+64.3%—

Agentic & Tool Use Not comparable

Mixtral 8x22B: 23.1 (#127), Yi-34B: —

Agentic & Tool Use benchmarks
BenchmarkMixtral 8x22BYi-34B
Cybench7.5%—

Reasoning Yi-34B leads

Mixtral 8x22B: 19.9 (#248), Yi-34B: 21.2 (#226)

Reasoning benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Hard Prompts11501104
Epoch Capabilities Index122.03117.39
DTBench55.1%—
BIG-Bench Hard—71.7%
ForecastBench56.3—

Math Mixtral 8x22B leads

Mixtral 8x22B: 22.9 (#275), Yi-34B: 21.6 (#282)

Math benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Math11841114
MATH Level 524.2%5.1%
Omni-MATH16.3%—
GSM8K—76%

Knowledge Mixtral 8x22B leads

Mixtral 8x22B: 15.1 (#293), Yi-34B: 7.5 (#309)

Knowledge benchmarks
BenchmarkMixtral 8x22BYi-34B
GPQA Diamond34.1%14.7%
LMArena Expert11131061
MMLU77.8%76.3%
MMLU-Pro46%—
GPQA (HELM)33.4%—

Multilingual Mixtral 8x22B leads

Mixtral 8x22B: 32.8 (#255), Yi-34B: 29.7 (#264)

Multilingual benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Non-English11281079
LMArena Chinese11161176
LMArena French11661081
LMArena German11411042
LMArena Japanese1037993
LMArena Korean1057959
LMArena Russian11581050
LMArena Spanish11511070

Instruction Following Mixtral 8x22B leads

Mixtral 8x22B: 57.7 (#266), Yi-34B: 56.2 (#274)

Instruction Following benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Instruction Following11471091
IFEval72.4%—

Long Context Mixtral 8x22B leads

Mixtral 8x22B: 34.7 (#247), Yi-34B: 33.2 (#264)

Long Context benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Longer Query11441094

Writing & Preference Mixtral 8x22B leads

Mixtral 8x22B: 36.9 (#262), Yi-34B: 34.1 (#273)

Writing & Preference benchmarks
BenchmarkMixtral 8x22BYi-34B
LMArena Text11621129
LMArena Creative Writing11411108
LMArena Multi-Turn11301113
WildBench71.1%—

Frequently asked questions

Is Mixtral 8x22B better than Yi-34B?

Mixtral 8x22B and Yi-34B score almost the same on the Noometry Index (27.1 vs 27.8), so choose on price, context window or the category you care about most.

Is Mixtral 8x22B or Yi-34B better for coding?

Yi-34B scores higher on coding benchmarks: 32.3 versus 24.2 in the Noometry coding category.

How many benchmarks do Mixtral 8x22B and Yi-34B share?

21 benchmarks have published results for both models. Mixtral 8x22B has 34 scored results on Noometry and Yi-34B has 23.

Related comparisons

Go deeper