Meta, open weights

Llama 3.2 90B

Llama 3.2 90B by Meta ranks 331st of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.5. Its strongest category is agentic & tool use, where it ranks 80th.

Last verified

Specifications

Noometry rank
#331 of 354
Index score
27.5
Evidence
Confirmed 9 results
Provider
Meta
Released
September 24, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Llama 3.2 90B category scores
  1. Agentic & Tool Use 30.0
  2. Reasoning 21.7
  3. Math 11.1
  4. Knowledge 21.7
  5. Multimodal 25.4
Llama 3.2 90B category ranks
CategoryScoreRankResults
Agentic & Tool Use30.0#801
Reasoning21.7#2171
Math11.1#3082
Knowledge21.7#2741
Multimodal25.4#1242

Strengths and weaknesses

Categories where Llama 3.2 90B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Llama 3.2 90B: strongest categories
CategoryScorevs medianRank
Agentic & Tool Use30.0−0.3#80 of 154, top 52%
Reasoning21.7−1.9#217 of 350, top 62%

Weakest categories

Llama 3.2 90B: weakest categories
CategoryScorevs medianRank
Multimodal25.4−13.1#124 of 128, top 97%
Math11.1−25.5#308 of 327, top 95%

Closest competitors

The models ranked just above and below Llama 3.2 90B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Llama 3.2 90B
ModelRankScoreBlended $/MSpeed
GPT-4.1 nano#32727.9$0.18135Compare
Phi 3 Mini 4k Instruct#32827.9——Compare
Yi-34B#32927.8——Compare
Llama 4 Scout#33027.7$0.15272Compare
Gemini 1.0 Pro#33227.3——Compare
Mixtral 8x22B#33327.1$3—Compare
Mixtral 8x7B#33427.1$0.70—Compare
Qwen Turbo#33527.1$0.0875—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Agentic & Tool Use

Llama 3.2 90B Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
BALROG27.3%#21 of 35, top 60%Epoch AI

Reasoning

Llama 3.2 90B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
EnigmaEval0.4%#38 of 38, top 100%Epoch AI
Epoch Capabilities Index125.5#156 of 213, top 74%Epoch AI2024-09-24

Math

Llama 3.2 90B Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-20252.6%#156 of 173, top 91%Epoch AI2025-02-25
MATH Level 539.4%#52 of 79, top 66%Epoch AI2025-01-27

Knowledge

Llama 3.2 90B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond41%#147 of 186, top 80%Epoch AI2025-01-27
MMLU80.3%#15 of 81, top 19%Epoch AI

Multimodal

Llama 3.2 90B Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1000#117 of 122, top 96%LMArena2026-10-09
GeoBench52%#19 of 25, top 76%Epoch AI

Compare Llama 3.2 90B

Other Meta models

Frequently asked questions

How good is Llama 3.2 90B?

Llama 3.2 90B by Meta ranks 331st of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.5. Its strongest category is agentic & tool use, where it ranks 80th.

Is Llama 3.2 90B open source?

Yes. Llama 3.2 90B's weights are downloadable; check the license for commercial terms.

What are Llama 3.2 90B's strengths and weaknesses?

Relative to other ranked models, Llama 3.2 90B places best in agentic & tool use, reasoning and lowest in multimodal, math.

What is Llama 3.2 90B best at?

Its best category is agentic & tool use, where it ranks 80th on Noometry.