Microsoft, open weights

Phi 3 Mini 4k Instruct

Phi 3 Mini 4k Instruct by Microsoft ranks 328th of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.9. Its strongest category is knowledge, where it ranks 246th.

Last verified

Specifications

Noometry rank
#328 of 354
Index score
27.9
Evidence
Confirmed 35 results
Provider
Microsoft
Released
April 23, 2024
Weights
Open weights
Reasoning
Unknown
Context window
—
Max output
—
Input price
Not listed
Output price
Not listed
Blended price
Not listed
Output speed
Not measured
Value
Not ranked
Knowledge cutoff
Unknown

Category scores

Each category score combines every public result we have in that category.

Phi 3 Mini 4k Instruct category scores
  1. Coding 26.6
  2. Reasoning 14.1
  3. Math 26.6
  4. Knowledge 28.5
  5. Multilingual 26.3
  6. Instruction Following 47.7
  7. Long Context 31.7
  8. Writing & Preference 27.6
Phi 3 Mini 4k Instruct category ranks
CategoryScoreRankResults
Coding26.6#3232
Reasoning14.1#3284
Math26.6#2572
Knowledge28.5#2461
Multilingual26.3#2801
Instruction Following47.7#3032
Long Context31.7#2761
Writing & Preference27.6#3004

Strengths and weaknesses

Categories where Phi 3 Mini 4k Instruct places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Phi 3 Mini 4k Instruct: strongest categories
CategoryScorevs medianRank
Knowledge28.5−8.8#246 of 314, top 79%
Math26.6−9.9#257 of 327, top 79%
Long Context31.7−9.2#276 of 296, top 94%

Weakest categories

Phi 3 Mini 4k Instruct: weakest categories
CategoryScorevs medianRank
Instruction Following47.7−23.5#303 of 305, top 100%
Writing & Preference27.6−26.2#300 of 312, top 97%
Coding26.6−12.1#323 of 340, top 95%

Closest competitors

The models ranked just above and below Phi 3 Mini 4k Instruct. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Phi 3 Mini 4k Instruct
ModelRankScoreBlended $/MSpeed
GPT-4o#32428.6$4.38—Compare
Ministral 8B#32528.2$0.15—Compare
Gemma 3 4B#32628.1$0.0572Compare
GPT-4.1 nano#32727.9$0.18135Compare
Yi-34B#32927.8——Compare
Llama 4 Scout#33027.7$0.15272Compare
Llama 3.2 90B#33127.5——Compare
Gemini 1.0 Pro#33227.3——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Phi 3 Mini 4k Instruct Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Coding15.5%#38 of 39, top 98%Epoch AI
LMArena Coding1093#273 of 294, top 93%LMArena2026-10-08
HumanEval+59.1%#29 of 45, top 65%EvalPlus
MBPP+54.2%#32 of 38, top 85%EvalPlus

Reasoning

Phi 3 Mini 4k Instruct Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Chess Puzzles0%#125 of 129, top 97%Epoch AI2026-08-30
LiveBench Reasoning26.8%#33 of 39, top 85%Epoch AI
LMArena Hard Prompts1072#275 of 297, top 93%LMArena2026-10-08
LiveBench Data Analysis34.7%#35 of 39, top 90%Epoch AI
Adversarial NLI52.8%#6 of 9, top 67%Epoch AI
BIG-Bench Hard71.7%#9 of 27, top 34%Epoch AI
HellaSwag76.7%#23 of 29, top 80%Epoch AI
LiveBench22.4%#38 of 39, top 98%Epoch AI
WinoGrande70.8%#31 of 43, top 73%Epoch AI

Math

Phi 3 Mini 4k Instruct Math benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Math15.7%#38 of 39, top 98%Epoch AI
LMArena Math1111#264 of 285, top 93%LMArena2026-10-08

Knowledge

Phi 3 Mini 4k Instruct Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Expert1045#263 of 273, top 97%LMArena2026-10-08
ARC (AI2) Challenge84.9%#10 of 39, top 26%Epoch AI
MMLU68.8%#49 of 81, top 61%Epoch AI
OpenBookQA88%Best of 19Epoch AI
TriviaQA64%#22 of 25, top 88%Epoch AI

Multilingual

Phi 3 Mini 4k Instruct Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1021#280 of 297, top 95%LMArena2026-10-08
LMArena Chinese1021#272 of 285, top 96%LMArena2026-10-08
LMArena French1076#217 of 223, top 98%LMArena2026-10-08
LMArena German1044#220 of 231, top 96%LMArena2026-10-08
LMArena Japanese935#206 of 211, top 98%LMArena2026-10-08
LMArena Korean905#209 of 213, top 99%LMArena2026-10-08
LMArena Russian1022#271 of 283, top 96%LMArena2026-10-08
LMArena Spanish1085#220 of 226, top 98%LMArena2026-10-08

Instruction Following

Phi 3 Mini 4k Instruct Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LiveBench Instruction Following39.1%#39 of 39, top 100%Epoch AI
LMArena Instruction Following1053#281 of 298, top 95%LMArena2026-10-08

Long Context

Phi 3 Mini 4k Instruct Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1044#280 of 291, top 97%LMArena2026-10-08

Writing & Preference

Phi 3 Mini 4k Instruct Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1073#282 of 297, top 95%LMArena2026-10-08
LMArena Creative Writing1037#281 of 295, top 96%LMArena2026-10-08
LMArena Multi-Turn1018#284 of 295, top 97%LMArena2026-10-08
LiveBench Language9.2%#39 of 39, top 100%Epoch AI

Compare Phi 3 Mini 4k Instruct

Other Microsoft models

Frequently asked questions

How good is Phi 3 Mini 4k Instruct?

Phi 3 Mini 4k Instruct by Microsoft ranks 328th of 354 ranked models on the Noometry Index as of October 2026, with a score of 27.9. Its strongest category is knowledge, where it ranks 246th.

Is Phi 3 Mini 4k Instruct open source?

Yes. Phi 3 Mini 4k Instruct's weights are downloadable; check the license for commercial terms.

What are Phi 3 Mini 4k Instruct's strengths and weaknesses?

Relative to other ranked models, Phi 3 Mini 4k Instruct places best in knowledge, math, long context and lowest in instruction following, writing & preference, coding.

What is Phi 3 Mini 4k Instruct best at?

Its best category is knowledge, where it ranks 246th on Noometry.