Model comparison

Llama 3.2 90B vs Yi-34B

Llama 3.2 90B and Yi-34B score almost the same on the Noometry Index (27.5 vs 27.8), so choose on price, context window or the category you care about most.

Last verified . 4 shared benchmarks.

Llama 3.2 90B Meta

27.5

Rank #331 Confirmed

Yi-34B 01.AI

27.8

Rank #329 Confirmed

Summary

  • They share 4 benchmarks with published results for both. Llama 3.2 90B scores higher in 2 categories and Yi-34B in 1 category; 2 gaps are clear of the uncertainty.
  • The widest gap is in knowledge, where Llama 3.2 90B leads 21.7 to 7.5.
  • The biggest single-benchmark swing is MATH Level 5: 39.4% for Llama 3.2 90B and 5.1% for Yi-34B.

Side by side

Llama 3.2 90B and Yi-34B specifications
Llama 3.2 90BYi-34B
ProviderMeta01.AI
Noometry Index27.527.8
Released2024-09-242023-11-02
WeightsOpenOpen
Context window——
Max output——
Input $ / M tokens——
Output $ / M tokens——
Results tracked923

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Llama 3.2 90B: —, Yi-34B: 32.3 (#274)

Coding benchmarks
BenchmarkLlama 3.2 90BYi-34B
LMArena Coding—1112

Agentic & Tool Use Not comparable

Llama 3.2 90B: 30.0 (#80), Yi-34B: —

Agentic & Tool Use benchmarks
BenchmarkLlama 3.2 90BYi-34B
BALROG27.3%—

Reasoning Too close to call

Llama 3.2 90B: 21.7 (#217), Yi-34B: 21.2 (#226)

Reasoning benchmarks
BenchmarkLlama 3.2 90BYi-34B
Epoch Capabilities Index125.5117.39
EnigmaEval0.4%—
LMArena Hard Prompts—1104
BIG-Bench Hard—71.7%

Math Yi-34B leads

Llama 3.2 90B: 11.1 (#308), Yi-34B: 21.6 (#282)

Math benchmarks
BenchmarkLlama 3.2 90BYi-34B
MATH Level 539.4%5.1%
OTIS Mock AIME 2024-20252.6%—
LMArena Math—1114
GSM8K—76%

Knowledge Llama 3.2 90B leads

Llama 3.2 90B: 21.7 (#274), Yi-34B: 7.5 (#309)

Knowledge benchmarks
BenchmarkLlama 3.2 90BYi-34B
GPQA Diamond41%14.7%
MMLU80.3%76.3%
LMArena Expert—1061

Multimodal Not comparable

Llama 3.2 90B: 25.4 (#124), Yi-34B: —

Multimodal benchmarks
BenchmarkLlama 3.2 90BYi-34B
LMArena Vision1000—
GeoBench52%—

Multilingual Not comparable

Llama 3.2 90B: —, Yi-34B: 29.7 (#264)

Multilingual benchmarks
BenchmarkLlama 3.2 90BYi-34B
LMArena Non-English—1079
LMArena Chinese—1176
LMArena French—1081
LMArena German—1042
LMArena Japanese—993
LMArena Korean—959
LMArena Russian—1050
LMArena Spanish—1070

Instruction Following Not comparable

Llama 3.2 90B: —, Yi-34B: 56.2 (#274)

Instruction Following benchmarks
BenchmarkLlama 3.2 90BYi-34B
LMArena Instruction Following—1091

Long Context Not comparable

Llama 3.2 90B: —, Yi-34B: 33.2 (#264)

Long Context benchmarks
BenchmarkLlama 3.2 90BYi-34B
LMArena Longer Query—1094

Writing & Preference Not comparable

Llama 3.2 90B: —, Yi-34B: 34.1 (#273)

Writing & Preference benchmarks
BenchmarkLlama 3.2 90BYi-34B
LMArena Text—1129
LMArena Creative Writing—1108
LMArena Multi-Turn—1113

Frequently asked questions

Is Llama 3.2 90B better than Yi-34B?

Llama 3.2 90B and Yi-34B score almost the same on the Noometry Index (27.5 vs 27.8), so choose on price, context window or the category you care about most.

How many benchmarks do Llama 3.2 90B and Yi-34B share?

4 benchmarks have published results for both models. Llama 3.2 90B has 9 scored results on Noometry and Yi-34B has 23.

Related comparisons

Go deeper