Alibaba (Qwen), open weights

Qwen3 14B

Qwen3 14B by Alibaba (Qwen) ranks 225th of 354 ranked models on the Noometry Index as of October 2026, with a score of 35.5. Its strongest category is agentic & tool use, where it ranks 83rd. API pricing starts at $0.35 per million input tokens and $1.40 per million output tokens, with a 131K-token context window.

Last verified

Specifications

Noometry rank
#225 of 354
Index score
35.5
Evidence
Confirmed 12 results
Released
April 1, 2025
Weights
Open weights
Reasoning
Yes
Context window
131K
Max output
8K
Input price
$0.35 / M
Output price
$1.40 / M
Blended price
$0.61 / M
Output speed
79 tokens/s Kagi
Value
#91 of 219
Knowledge cutoff
April 2025
Input
text
Hugging Face
Qwen/Qwen3-14B

Category scores

Each category score combines every public result we have in that category.

Qwen3 14B category scores
  1. Coding 37.3
  2. Agentic & Tool Use 29.6
  3. Reasoning 18.5
  4. Math 38.6
  5. Knowledge 39.3
  6. Long Context 38.1
Qwen3 14B category ranks
CategoryScoreRankResults
Coding37.3#1951
Agentic & Tool Use29.6#831
Reasoning18.5#2805
Math38.6#1331
Knowledge39.3#1342
Long Context38.1#2041

Strengths and weaknesses

Categories where Qwen3 14B places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen3 14B: strongest categories
CategoryScorevs medianRank
Math38.6+2.1#133 of 327, top 41%
Knowledge39.3+1.9#134 of 314, top 43%
Agentic & Tool Use29.6−0.7#83 of 154, top 54%

Weakest categories

Qwen3 14B: weakest categories
CategoryScorevs medianRank
Reasoning18.5−5.1#280 of 350, top 80%
Long Context38.1−2.8#204 of 296, top 69%
Coding37.3−1.5#195 of 340, top 58%

Closest competitors

The models ranked just above and below Qwen3 14B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen3 14B
ModelRankScoreBlended $/MSpeed
C4ai Aya Expanse 32b#22135.9——Compare
Llama 3.1 Nemotron 51b Instruct#22235.9——Compare
Nemotron 4 340b Instruct#22335.9——Compare
Llama 3.1 Tulu 3 8b#22435.7——Compare
DeepSeek-R1-Distill-Qwen-32B#22635.5——Compare
Magistral Medium#22735.2$2.750Compare
Gemini 2.0 Flash (Feb 2025)#22835.1—92Compare
C4ai Aya Expanse 8b#22934.9——Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen3 14B Coding benchmark results
BenchmarkScorePositionSettingSourceDate
SciCode31.6%#105 of 121, top 87%Epoch AI

Agentic & Tool Use

Qwen3 14B Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Berkeley Function Calling Leaderboard41%#23 of 49, top 47%fcBerkeley Function Calling Leaderboard

Reasoning

Qwen3 14B Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Kagi LLM Benchmark49.1%#67 of 99, top 68%Kagi LLM Benchmark
CritPt0%#130 of 134, top 98%Epoch AI
Chess Puzzles4%#99 of 129, top 77%Epoch AI2026-08-30
Chess Puzzles0%noneEpoch AI2026-08-30
DTBench64%#104 of 151, top 69%Epoch AI
LMCA18.2%#95 of 125, top 76%Epoch AI
Epoch Capabilities Index138.23#115 of 213, top 54%Epoch AI2025-04-29

Math

Qwen3 14B Math benchmark results
BenchmarkScorePositionSettingSourceDate
OTIS Mock AIME 2024-202566.4%#99 of 173, top 58%Epoch AI2026-08-28
OTIS Mock AIME 2024-202525.8%noneEpoch AI2026-08-30

Knowledge

Qwen3 14B Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond63.8%#111 of 186, top 60%Epoch AI2026-08-28
GPQA Diamond53.4%noneEpoch AI2026-08-28
Vectara Hallucination Rate (lower is better)5.4%#14 of 96, top 15%Vectara Hallucination Leaderboard

Long Context

Qwen3 14B Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
Fiction.LiveBench62.5%#26 of 47, top 56%Epoch AI

API pricing by provider

Qwen3 14B API prices
RouteInput $/MOutput $/MCached input $/MChecked
alibaba$0.35$1.40—2026-10-10
openrouter$0.12$0.24—2026-10-10

Compare Qwen3 14B

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen3 14B?

Qwen3 14B by Alibaba (Qwen) ranks 225th of 354 ranked models on the Noometry Index as of October 2026, with a score of 35.5. Its strongest category is agentic & tool use, where it ranks 83rd. API pricing starts at $0.35 per million input tokens and $1.40 per million output tokens, with a 131K-token context window.

How much does Qwen3 14B cost?

Qwen3 14B costs $0.35 per million input tokens and $1.40 per million output tokens on Alibaba (Qwen)'s own API.

What is Qwen3 14B's context window?

Qwen3 14B accepts up to 131K tokens of input and can write up to 8K tokens in one response.

Is Qwen3 14B open source?

Yes. Qwen3 14B's weights are downloadable from Hugging Face (Qwen/Qwen3-14B); check the license for commercial terms.

How fast is Qwen3 14B?

Qwen3 14B generated about 79 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

What are Qwen3 14B's strengths and weaknesses?

Relative to other ranked models, Qwen3 14B places best in math, knowledge, agentic & tool use and lowest in reasoning, long context, coding.

What is Qwen3 14B best at?

Its best category is agentic & tool use, where it ranks 83rd on Noometry.