Model comparison

Yi-1.5-34B vs Yi-34B

Yi-1.5-34B is the stronger model overall, scoring 30.6 to 27.8 on the Noometry Index.

Last verified . 19 shared benchmarks.

Yi-1.5-34B 01.AI

30.6

Rank #289 Confirmed

Yi-34B 01.AI

27.8

Rank #329 Confirmed

Summary

  • They share 19 benchmarks with published results for both. Yi-1.5-34B scores higher in 8 categories and Yi-34B in 0 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Yi-1.5-34B leads 14.8 to 7.5.
  • The biggest single-benchmark swing is MATH Level 5: 25.5% for Yi-1.5-34B and 5.1% for Yi-34B.

Side by side

Yi-1.5-34B and Yi-34B specifications
Yi-1.5-34BYi-34B
Provider01.AI01.AI
Noometry Index30.627.8
Released2024-05-132023-11-02
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked2123

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Too close to call

Yi-1.5-34B: 32.4 (#272), Yi-34B: 32.3 (#274)

Coding benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Coding11691112
BigCodeBench Instruct33.9%—
BigCodeBench Complete43.8%—

Reasoning Yi-1.5-34B leads

Yi-1.5-34B: 22.5 (#191), Yi-34B: 21.2 (#226)

Reasoning benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Hard Prompts11601104
BIG-Bench Hard—71.7%
Epoch Capabilities Index—117.39

Math Yi-1.5-34B leads

Yi-1.5-34B: 27.5 (#249), Yi-34B: 21.6 (#282)

Math benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Math11821114
MATH Level 525.5%5.1%
GSM8K—76%

Knowledge Yi-1.5-34B leads

Yi-1.5-34B: 14.8 (#295), Yi-34B: 7.5 (#309)

Knowledge benchmarks
BenchmarkYi-1.5-34BYi-34B
GPQA Diamond32%14.7%
LMArena Expert11441061
MMLU—76.3%

Multilingual Yi-1.5-34B leads

Yi-1.5-34B: 32.3 (#256), Yi-34B: 29.7 (#264)

Multilingual benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Non-English11211079
LMArena Chinese12131176
LMArena French11561081
LMArena German11111042
LMArena Japanese1021993
LMArena Korean1005959
LMArena Russian10911050
LMArena Spanish11211070

Instruction Following Yi-1.5-34B leads

Yi-1.5-34B: 59.2 (#257), Yi-34B: 56.2 (#274)

Instruction Following benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Instruction Following11391091

Long Context Yi-1.5-34B leads

Yi-1.5-34B: 34.6 (#248), Yi-34B: 33.2 (#264)

Long Context benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Longer Query11431094

Writing & Preference Yi-1.5-34B leads

Yi-1.5-34B: 37.4 (#257), Yi-34B: 34.1 (#273)

Writing & Preference benchmarks
BenchmarkYi-1.5-34BYi-34B
LMArena Text11731129
LMArena Creative Writing11351108
LMArena Multi-Turn11531113

Frequently asked questions

Is Yi-1.5-34B better than Yi-34B?

Yi-1.5-34B is the stronger model overall, scoring 30.6 to 27.8 on the Noometry Index.

Is Yi-1.5-34B or Yi-34B better for coding?

They score almost the same on coding (32.4 vs 32.3); test both on your own repository before choosing.

How many benchmarks do Yi-1.5-34B and Yi-34B share?

19 benchmarks have published results for both models. Yi-1.5-34B has 21 scored results on Noometry and Yi-34B has 23.

Related comparisons

Go deeper