Alibaba (Qwen), proprietary

Qwen3.5-Flash

Qwen3.5-Flash by Alibaba (Qwen) ranks 112th of 354 ranked models on the Noometry Index as of October 2026, with a score of 42.5. Its strongest category is reasoning, where it ranks 72nd. API pricing starts at $0.10 per million input tokens and $0.40 per million output tokens, with a 1M-token context window.

Last verified

Specifications

Noometry rank
#112 of 354
Index score
42.5
Evidence
Confirmed 32 results
Released
February 23, 2026
Weights
Proprietary
Reasoning
Yes
Context window
1M
Max output
66K
Input price
$0.10 / M
Output price
$0.40 / M
Blended price
$0.18 / M
Output speed
Not measured
Value
#32 of 219
Knowledge cutoff
Unknown
Input
text, image, video

Category scores

Each category score combines every public result we have in that category.

Qwen3.5-Flash category scores
  1. Coding 34.2
  2. Reasoning 33.7
  3. Math 37.4
  4. Knowledge 43.2
  5. Multilingual 50.5
  6. Instruction Following 72.6
  7. Long Context 42.4
  8. Writing & Preference 57.9
Qwen3.5-Flash category ranks
CategoryScoreRankResults
Coding34.2#2422
Reasoning33.7#725
Math37.4#1583
Knowledge43.2#934
Multilingual50.5#1211
Instruction Following72.6#1391
Long Context42.4#1241
Writing & Preference57.9#1223

Strengths and weaknesses

Categories where Qwen3.5-Flash places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Qwen3.5-Flash: strongest categories
CategoryScorevs medianRank
Reasoning33.7+10.1#72 of 350, top 21%
Knowledge43.2+5.9#93 of 314, top 30%
Writing & Preference57.9+4.2#122 of 312, top 40%

Weakest categories

Qwen3.5-Flash: weakest categories
CategoryScorevs medianRank
Coding34.2−4.5#242 of 340, top 72%
Math37.4+0.9#158 of 327, top 49%
Instruction Following72.6+1.3#139 of 305, top 46%

Closest competitors

The models ranked just above and below Qwen3.5-Flash. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen3.5-Flash
ModelRankScoreBlended $/MSpeed
DeepSeek-V3.1#10842.8$0.42328Compare
GPT-5.3 Chat#10942.8$4.81—Compare
GPT-5.5 Instant#11042.7——Compare
GPT-5.2 Codex#11142.6$4.81—Compare
Nemotron 3 Ultra#11342.5$0.93—Compare
Hunyuan T1 20250711#11442.5——Compare
DeepSeek-R1#11542.3$0.9110Compare
Step 3.5 Flash#11642.3$0.15—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Qwen3.5-Flash Coding benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena WebDev1244#103 of 113, top 92%LMArena2026-10-08
LMArena Coding1412#126 of 294, top 43%LMArena2026-10-08
ALE-Bench221.8#102 of 105, top 98%Epoch AI

Agentic & Tool Use

Qwen3.5-Flash Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
Vending-Bench 2462.69#50 of 60, top 84%Epoch AI

Reasoning

Qwen3.5-Flash Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
Chess Puzzles21%#56 of 129, top 44%Epoch AI2026-08-07
LMArena Hard Prompts1403#123 of 297, top 42%LMArena2026-10-08
Mystery Game Puzzles20%#42 of 74, top 57%Epoch AI2026-08-27
Mystery Game Puzzles16%noneEpoch AI2026-08-27
DTBench82.9%#60 of 151, top 40%Epoch AI
LMCA29.1%#77 of 125, top 62%Epoch AI
Epoch Capabilities Index143.98#88 of 213, top 42%Epoch AI2026-02-25

Math

Qwen3.5-Flash Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)9.5%Epoch AI2026-08-28
FrontierMath (Tiers 1-3)18.2%#73 of 81, top 91%noneEpoch AI2026-08-28
OTIS Mock AIME 2024-202584.4%#72 of 173, top 42%Epoch AI2026-08-07
LMArena Math1407#118 of 285, top 42%LMArena2026-10-08
FrontierMath (Feb 2025 set)6.2%#42 of 68, top 62%Epoch AI2026-05-12
FrontierMath Tier 4 (v1)0%#54 of 55, top 99%Epoch AI2026-05-12

Knowledge

Qwen3.5-Flash Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond82.3%#75 of 186, top 41%Epoch AI2026-08-07
SimpleQA Verified20.3%#64 of 77, top 84%Epoch AI2026-08-27
Vectara Hallucination Rate (lower is better)10.5%#60 of 96, top 63%Vectara Hallucination Leaderboard
LMArena Expert1407#117 of 273, top 43%LMArena2026-10-08

Multilingual

Qwen3.5-Flash Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1385#121 of 297, top 41%LMArena2026-10-08
LMArena Chinese1446#107 of 285, top 38%LMArena2026-10-08
LMArena French1412#108 of 223, top 49%LMArena2026-10-08
LMArena German1390#102 of 231, top 45%LMArena2026-10-08
LMArena Japanese1368#88 of 211, top 42%LMArena2026-10-08
LMArena Korean1344#108 of 213, top 51%LMArena2026-10-08
LMArena Russian1379#129 of 283, top 46%LMArena2026-10-08
LMArena Spanish1400#110 of 226, top 49%LMArena2026-10-08

Instruction Following

Qwen3.5-Flash Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1374#130 of 298, top 44%LMArena2026-10-08

Long Context

Qwen3.5-Flash Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1392#124 of 291, top 43%LMArena2026-10-08

Writing & Preference

Qwen3.5-Flash Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1397#120 of 297, top 41%LMArena2026-10-08
LMArena Creative Writing1343#136 of 295, top 47%LMArena2026-10-08
LMArena Multi-Turn1393#127 of 295, top 44%LMArena2026-10-08

API pricing by provider

Qwen3.5-Flash API prices
RouteInput $/MOutput $/MCached input $/MChecked
alibaba$0.10$0.40$0.012026-10-10
openrouter$0.065$0.26—2026-10-10

Compare Qwen3.5-Flash

Other Alibaba (Qwen) models

Frequently asked questions

How good is Qwen3.5-Flash?

Qwen3.5-Flash by Alibaba (Qwen) ranks 112th of 354 ranked models on the Noometry Index as of October 2026, with a score of 42.5. Its strongest category is reasoning, where it ranks 72nd. API pricing starts at $0.10 per million input tokens and $0.40 per million output tokens, with a 1M-token context window.

How much does Qwen3.5-Flash cost?

Qwen3.5-Flash costs $0.10 per million input tokens and $0.40 per million output tokens on Alibaba (Qwen)'s own API, with cached input at $0.01.

What is Qwen3.5-Flash's context window?

Qwen3.5-Flash accepts up to 1M tokens of input and can write up to 66K tokens in one response.

Is Qwen3.5-Flash open source?

No. Qwen3.5-Flash is proprietary and available only through Alibaba (Qwen)'s API and partner platforms.

What are Qwen3.5-Flash's strengths and weaknesses?

Relative to other ranked models, Qwen3.5-Flash places best in reasoning, knowledge, writing & preference and lowest in coding, math, instruction following.

What is Qwen3.5-Flash best at?

Its best category is reasoning, where it ranks 72nd on Noometry.