Model comparison

Codellama 34b Instruct vs Qwen Turbo

Codellama 34b Instruct is the stronger model overall, scoring 30.8 to 27.1 on the Noometry Index.

Last verified . 0 shared benchmarks.

Codellama 34b Instruct Meta

30.8

Rank #287 Confirmed

Qwen Turbo Alibaba (Qwen)

27.1

Rank #335 Reported

Summary

  • The widest gap is in math, where Codellama 34b Instruct leads 31.0 to 15.3.
  • Codellama 34b Instruct has downloadable open weights; the other is API-only.

Side by side

Codellama 34b Instruct and Qwen Turbo specifications
Codellama 34b InstructQwen Turbo
ProviderMetaAlibaba (Qwen)
Noometry Index30.827.1
Released—2024-11-01
WeightsOpenProprietary
Context window—1M
Max output—16K
Input $ / M tokens—$0.05
Output $ / M tokens—$0.20
Results tracked143

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Codellama 34b Instruct: 28.5 (#314), Qwen Turbo: —

Coding benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
BigCodeBench Instruct29%—
LMArena Coding1046—
BigCodeBench Complete37.1%—
HumanEval+43.9%—
MBPP+56.3%—

Reasoning Not comparable

Codellama 34b Instruct: 19.6 (#255), Qwen Turbo: —

Reasoning benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
LMArena Hard Prompts1032—

Math Codellama 34b Instruct leads

Codellama 34b Instruct: 31.0 (#230), Qwen Turbo: 15.3 (#297)

Math benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
OTIS Mock AIME 2024-2025—6.1%
LMArena Math1056—
MATH Level 5—56.2%

Knowledge Not comparable

Codellama 34b Instruct: —, Qwen Turbo: 22.2 (#272)

Knowledge benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
GPQA Diamond—41.8%

Multilingual Not comparable

Codellama 34b Instruct: 25.8 (#284), Qwen Turbo: —

Multilingual benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
LMArena Non-English1011—
LMArena Chinese976—

Instruction Following Not comparable

Codellama 34b Instruct: 52.2 (#291), Qwen Turbo: —

Instruction Following benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
LMArena Instruction Following1028—

Long Context Not comparable

Codellama 34b Instruct: 30.9 (#284), Qwen Turbo: —

Long Context benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
LMArena Longer Query1013—

Writing & Preference Not comparable

Codellama 34b Instruct: 28.2 (#297), Qwen Turbo: —

Writing & Preference benchmarks
BenchmarkCodellama 34b InstructQwen Turbo
LMArena Text1066—
LMArena Creative Writing1032—
LMArena Multi-Turn1015—

Frequently asked questions

Is Codellama 34b Instruct better than Qwen Turbo?

Codellama 34b Instruct is the stronger model overall, scoring 30.8 to 27.1 on the Noometry Index.

How many benchmarks do Codellama 34b Instruct and Qwen Turbo share?

0 benchmarks have published results for both models. Codellama 34b Instruct has 14 scored results on Noometry and Qwen Turbo has 3.

Related comparisons

Go deeper