Microsoft, open weights

phi-3-medium 14B

phi-3-medium 14B by Microsoft ranks 306th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.7. Its strongest category is coding, where it ranks 201st.

Last verified

Specifications

Noometry rank
#306 of 354
Index score
29.7
Evidence
Reported 13 results
Provider
Microsoft
Released
April 23, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

phi-3-medium 14B category scores
  1. Coding 36.8
  2. Math 27.3
  3. Knowledge 9.1
phi-3-medium 14B category ranks
CategoryScoreRankResults
Coding36.8#2012
Math27.3#2501
Knowledge9.1#3061

Strengths and weaknesses

Categories where phi-3-medium 14B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

phi-3-medium 14B: strongest categories
CategoryScorevs medianRank
Coding36.8−1.9#201 of 340, top 60%

Weakest categories

phi-3-medium 14B: weakest categories
CategoryScorevs medianRank
Knowledge9.1−28.2#306 of 314, top 98%

Closest competitors

The models ranked just above and below phi-3-medium 14B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to phi-3-medium 14B
ModelRankScoreBlended $/MSpeed
Qwen2.5-VL 72B Instruct#30229.9$4.2043Compare
Mistral#30329.9——Compare
OLMo 2 Furious 13B#30429.7——Compare
Phi 3 Mini 128k Instruct#30529.7——Compare
Gemma 2B#30729.6——Compare
Llama 3.1-70B#30829.6$0.40—Compare
Llama 2-13B#30929.6——Compare
Claude 3 Opus#31029.5——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

phi-3-medium 14B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
BigCodeBench Instruct37.6%#40 of 64, top 63%BigCodeBench2024-05-21
BigCodeBench Complete48.7%#37 of 66, top 57%BigCodeBench2024-05-21

Reasoning

phi-3-medium 14B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Adversarial NLI55.8%#4 of 9, top 45%Epoch AI
BIG-Bench Hard81.4%#4 of 27, top 15%Epoch AI
Epoch Capabilities Index121.23#165 of 213, top 78%Epoch AI2024-04-23
HellaSwag82.4%#14 of 29, top 49%Epoch AI
WinoGrande81.5%#12 of 43, top 28%Epoch AI

Math

phi-3-medium 14B Math benchmark results
BenchmarkScorePositionSettingSourceDate
MATH Level 517.6%#65 of 79, top 83%Epoch AI2025-01-31

Knowledge

phi-3-medium 14B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond27.6%#174 of 186, top 94%Epoch AI2025-01-31
ARC (AI2) Challenge91.6%#5 of 39, top 13%Epoch AI
MMLU78%#25 of 81, top 31%Epoch AI
OpenBookQA87.4%#3 of 19, top 16%Epoch AI
TriviaQA73.9%#16 of 25, top 64%Epoch AI

Compare phi-3-medium 14B

Other Microsoft models

Frequently asked questions

How good is phi-3-medium 14B?

phi-3-medium 14B by Microsoft ranks 306th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.7. Its strongest category is coding, where it ranks 201st.

Is phi-3-medium 14B open source?

Yes. phi-3-medium 14B's weights are downloadable; check the license for commercial terms.

What are phi-3-medium 14B's strengths and weaknesses?

Relative to other ranked models, phi-3-medium 14B places best in coding and lowest in knowledge.

What is phi-3-medium 14B best at?

Its best category is coding, where it ranks 201st on Noometry.