Alibaba (Qwen), proprietary
Qwen3 Max
Qwen3 Max by Alibaba (Qwen) ranks 87th of 354 ranked models on the Noometry Index as of October 2026, with a score of 43.7. Its strongest category is multilingual, where it ranks 62nd. API pricing starts at $1.20 per million input tokens and $6 per million output tokens, with a 262K-token context window.
Last verified
Specifications
- Noometry rank
- #87 of 354
- Index score
- 43.7
- Evidence
- Confirmed 33 results
- Provider
Alibaba (Qwen)
- Released
- September 23, 2025
- Weights
- Proprietary
- Reasoning
- No
- Context window
- 262K
- Max output
- 66K
- Input price
- $1.20 / M
- Output price
- $6 / M
- Blended price
- $2.40 / M
- Output speed
- 48 tokens/s Kagi
- Value
- #156 of 219
- Knowledge cutoff
- April 2025
- Input
- text
Category scores
Each category score combines every public result we have in that category.
- Coding 43.0
- Reasoning 22.6
- Math 38.7
- Knowledge 48.1
- Multilingual 53.7
- Instruction Following 74.8
- Long Context 41.6
- Writing & Preference 62.4
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 43.0 | #93 | 1 |
| Reasoning | 22.6 | #190 | 7 |
| Math | 38.7 | #131 | 4 |
| Knowledge | 48.1 | #78 | 3 |
| Multilingual | 53.7 | #62 | 1 |
| Instruction Following | 74.8 | #87 | 1 |
| Long Context | 41.6 | #134 | 3 |
| Writing & Preference | 62.4 | #76 | 3 |
Strengths and weaknesses
Categories where Qwen3 Max places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 53.7 | +6.3 | #62 of 297, top 21% |
| Writing & Preference | 62.4 | +8.6 | #76 of 312, top 25% |
| Knowledge | 48.1 | +10.8 | #78 of 314, top 25% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Reasoning | 22.6 | −1.0 | #190 of 350, top 55% |
| Long Context | 41.6 | +0.7 | #134 of 296, top 46% |
| Math | 38.7 | +2.1 | #131 of 327, top 41% |
Closest competitors
The models ranked just above and below Qwen3 Max. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| ERNIE 5.1 | #83 | 43.8 | — | — | Compare |
| GLM-5V-Turbo | #84 | 43.8 | $1.90 | — | Compare |
| MiniMax-M3 | #85 | 43.8 | $0.52 | — | Compare |
| Grok 4.3 | #86 | 43.8 | $1.56 | — | Compare |
| MiMo-V2-Omni | #88 | 43.6 | $0.18 | — | Compare |
| Kimi K2.5 Instant | #89 | 43.6 | — | — | Compare |
| Gemma 4 31B IT | #90 | 43.5 | $0.15 | 3 | Compare |
| Qwen3 235B-A22B | #91 | 43.5 | $1.22 | 85 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Coding | 1439 | LMArena | 2026-10-08 | ||
| LMArena Coding | 1456 | #76 of 294, top 26% | LMArena | 2026-10-08 | |
| ALE-Bench | 370.45 | #92 of 105, top 88% | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Vending-Bench 2 | 71.56 | #54 of 60, top 90% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Kagi LLM Benchmark | 72.5% | #20 of 99, top 21% | Kagi LLM Benchmark | ||
| Kagi LLM Benchmark | 55.9% | Kagi LLM Benchmark | |||
| NYT Connections (extended) | 11.8% | Lech Mazur benchmarks | |||
| NYT Connections (extended) | 30.1% | #72 of 91, top 80% | 2026-01-23 | Lech Mazur benchmarks | |
| Chess Puzzles | 4% | #101 of 129, top 79% | Epoch AI | 2025-12-10 | |
| LMArena Hard Prompts | 1423 | LMArena | 2026-10-08 | ||
| LMArena Hard Prompts | 1448 | #66 of 297, top 23% | LMArena | 2026-10-08 | |
| Mystery Game Puzzles | 5% | #72 of 74, top 98% | Epoch AI | 2026-08-27 | |
| DTBench | 82.1% | #66 of 151, top 44% | Epoch AI | ||
| LMCA | 28.3% | #81 of 125, top 65% | Epoch AI | ||
| Epoch Capabilities Index | 142.38 | #99 of 213, top 47% | Epoch AI | 2025-09-24 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 18.9% | #71 of 81, top 88% | Epoch AI | 2026-08-30 | |
| OTIS Mock AIME 2024-2025 | 73.3% | #87 of 173, top 51% | Epoch AI | 2025-10-06 | |
| LMArena Math | 1446 | #66 of 285, top 24% | LMArena | 2026-10-08 | |
| LMArena Math | 1434 | LMArena | 2026-10-08 | ||
| MATH Level 5 | 97.1% | #6 of 79, top 8% | Epoch AI | 2025-10-09 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 72.6% | #96 of 186, top 52% | Epoch AI | 2025-10-06 | |
| SimpleQA Verified | 48.7% | #28 of 77, top 37% | Epoch AI | 2026-08-27 | |
| LMArena Expert | 1455 | #65 of 273, top 24% | LMArena | 2026-10-08 | |
| LMArena Expert | 1397 | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1429 | #62 of 297, top 21% | LMArena | 2026-10-08 | |
| LMArena Non-English | 1400 | LMArena | 2026-10-08 | ||
| LMArena Chinese | 1430 | LMArena | 2026-10-08 | ||
| LMArena Chinese | 1478 | #65 of 285, top 23% | LMArena | 2026-10-08 | |
| LMArena French | 1449 | #71 of 223, top 32% | LMArena | 2026-10-08 | |
| LMArena German | 1463 | #33 of 231, top 15% | LMArena | 2026-10-08 | |
| LMArena German | 1437 | LMArena | 2026-10-08 | ||
| LMArena Japanese | 1397 | #61 of 211, top 29% | LMArena | 2026-10-08 | |
| LMArena Korean | 1376 | LMArena | 2026-10-08 | ||
| LMArena Korean | 1399 | #56 of 213, top 27% | LMArena | 2026-10-08 | |
| LMArena Russian | 1417 | LMArena | 2026-10-08 | ||
| LMArena Russian | 1428 | #73 of 283, top 26% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1462 | #33 of 226, top 15% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1421 | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1419 | #76 of 298, top 26% | LMArena | 2026-10-08 | |
| LMArena Instruction Following | 1398 | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| Fiction.LiveBench | 66.7% | #23 of 47, top 49% | Epoch AI | ||
| CL-bench | 14.5% | #17 of 19, top 90% | Epoch AI | ||
| LMArena Longer Query | 1410 | LMArena | 2026-10-08 | ||
| LMArena Longer Query | 1438 | #68 of 291, top 24% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1413 | LMArena | 2026-10-08 | ||
| LMArena Text | 1439 | #63 of 297, top 22% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1402 | #74 of 295, top 26% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1381 | LMArena | 2026-10-08 | ||
| LMArena Multi-Turn | 1434 | LMArena | 2026-10-08 | ||
| LMArena Multi-Turn | 1446 | #61 of 295, top 21% | LMArena | 2026-10-08 |
API pricing by provider
Compare Qwen3 Max
- Qwen3 Max vs Qwen2.5-Max
- Qwen3 Max vs Grok 4.3
- Qwen3 Max vs MiMo-V2-Omni
- Qwen3 Max vs MiniMax-M3
- Qwen3 Max vs Kimi K2.5 Instant
- Qwen3 Max vs GLM-5V-Turbo
- Qwen3 Max vs Gemma 4 31B IT
- Qwen3 Max vs GPT-6 Astra
- Qwen3 Max vs Claude Fable 5.1
- Qwen3 Max vs Gemini 3.8 Flash
- Qwen3 Max vs Kimi K3
- Qwen3 Max vs Grok 4.6
- Qwen3 Max vs GLM-5.3
- Qwen3 Max vs Muse Spark 1.3
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.7 Max51.5
- Qwen3.6 Max Preview51.5
- Qwen3.6 Plus47.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
Frequently asked questions
How good is Qwen3 Max?
Qwen3 Max by Alibaba (Qwen) ranks 87th of 354 ranked models on the Noometry Index as of October 2026, with a score of 43.7. Its strongest category is multilingual, where it ranks 62nd. API pricing starts at $1.20 per million input tokens and $6 per million output tokens, with a 262K-token context window.
How much does Qwen3 Max cost?
Qwen3 Max costs $1.20 per million input tokens and $6 per million output tokens on Alibaba (Qwen)'s own API.
What is Qwen3 Max's context window?
Qwen3 Max accepts up to 262K tokens of input and can write up to 66K tokens in one response.
Is Qwen3 Max open source?
No. Qwen3 Max is proprietary and available only through Alibaba (Qwen)'s API and partner platforms.
How fast is Qwen3 Max?
Qwen3 Max generated about 48 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.
What are Qwen3 Max's strengths and weaknesses?
Relative to other ranked models, Qwen3 Max places best in multilingual, writing & preference, knowledge and lowest in reasoning, long context, math.
What is Qwen3 Max best at?
Its best category is multilingual, where it ranks 62nd on Noometry.