Alibaba (Qwen), proprietary
Qwen3.7 Max
Qwen3.7 Max by Alibaba (Qwen) ranks 42nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 51.5. Its strongest category is multilingual, where it ranks 15th. API pricing starts at $2.50 per million input tokens and $7.50 per million output tokens, with a 1M-token context window.
Last verified
Specifications
- Noometry rank
- #42 of 354
- Index score
- 51.5
- Evidence
- Confirmed 33 results
- Provider
Alibaba (Qwen)
- Released
- May 19, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 1M
- Max output
- 131K
- Input price
- $2.50 / M
- Output price
- $7.50 / M
- Blended price
- $3.75 / M
- Output speed
- Not measured
- Value
- #166 of 219
- Knowledge cutoff
- Unknown
- Input
- text
Category scores
Each category score combines every public result we have in that category.
- Coding 50.4
- Agentic & Tool Use 22.1
- Reasoning 49.2
- Math 62.4
- Knowledge 61.6
- Multilingual 56.9
- Instruction Following 76.7
- Long Context 45.4
- Writing & Preference 65.0
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 50.4 | #45 | 4 |
| Agentic & Tool Use | 22.1 | #135 | 1 |
| Reasoning | 49.2 | #38 | 9 |
| Math | 62.4 | #32 | 5 |
| Knowledge | 61.6 | #28 | 3 |
| Multilingual | 56.9 | #15 | 1 |
| Instruction Following | 76.7 | #38 | 1 |
| Long Context | 45.4 | #40 | 1 |
| Writing & Preference | 65.0 | #54 | 4 |
Strengths and weaknesses
Categories where Qwen3.7 Max places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 56.9 | +9.5 | #15 of 297, top 6% |
| Knowledge | 61.6 | +24.3 | #28 of 314, top 9% |
| Math | 62.4 | +25.9 | #32 of 327, top 10% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Agentic & Tool Use | 22.1 | −8.3 | #135 of 154, top 88% |
| Writing & Preference | 65.0 | +11.2 | #54 of 312, top 18% |
| Long Context | 45.4 | +4.5 | #40 of 296, top 14% |
Closest competitors
The models ranked just above and below Qwen3.7 Max. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| DeepSeek V4.1 Flash | #38 | 52.8 | $0.26 | — | Compare |
| GPT-5.2 Pro | #39 | 52.3 | $57.75 | — | Compare |
| Gemini 3 Flash Preview | #40 | 52.3 | $1.13 | — | Compare |
| GLM-5.3-Flash | #41 | 51.8 | $0.24 | — | Compare |
| Qwen3.6 Max Preview | #43 | 51.5 | $2.92 | — | Compare |
| GLM-5.2 | #44 | 51.1 | $2.15 | 23 | Compare |
| GPT-5 | #45 | 50.9 | $3.44 | 2 | Compare |
| Muse Spark | #46 | 50.6 | — | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SWE-bench Verified | 77.3% | #7 of 32, top 22% | Epoch AI | 2026-06-18 | |
| LMArena WebDev | 1515 | #43 of 113, top 39% | LMArena | 2026-10-08 | |
| SciCode | 48.8% | #47 of 121, top 39% | Epoch AI | ||
| LMArena Coding | 1498 | #22 of 294, top 8% | LMArena | 2026-10-08 | |
| ALE-Bench | 1,189 | #31 of 105, top 30% | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GBAEval | 0.4% | #20 of 23, top 87% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SimpleBench | 70.4% | #11 of 77, top 15% | Epoch AI | ||
| NYT Connections (extended) | 85.1% | #29 of 91, top 32% | Lech Mazur benchmarks | ||
| CritPt | 13.4% | #40 of 134, top 30% | Epoch AI | ||
| Chess Puzzles | 19% | #63 of 129, top 49% | Epoch AI | 2026-08-07 | |
| EBR-Bench | 9.5% | #20 of 24, top 84% | Epoch AI | 2026-06-25 | |
| LMArena Hard Prompts | 1483 | #28 of 297, top 10% | LMArena | 2026-10-08 | |
| Mystery Game Puzzles | 32% | #25 of 74, top 34% | Epoch AI | 2026-07-28 | |
| Mystery Game Puzzles | 26% | none | Epoch AI | 2026-08-27 | |
| DTBench | 92.3% | #28 of 151, top 19% | Epoch AI | ||
| LMCA | 44% | #37 of 125, top 30% | Epoch AI | ||
| Epoch Capabilities Index | 153.68 | #39 of 213, top 19% | Epoch AI | 2026-05-19 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 64.6% | #33 of 81, top 41% | Epoch AI | 2026-06-13 | |
| FrontierMath Tier 4 | 34.1% | #25 of 63, top 40% | Epoch AI | 2026-06-13 | |
| OTIS Mock AIME 2024-2025 | 95.6% | #35 of 173, top 21% | Epoch AI | 2026-08-07 | |
| ProofBench | 26% | #44 of 77, top 58% | Epoch AI | ||
| LMArena Math | 1490 | #18 of 285, top 7% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 90.9% | #31 of 186, top 17% | Epoch AI | 2026-08-07 | |
| SimpleQA Verified | 55.8% | #19 of 77, top 25% | Epoch AI | 2026-08-27 | |
| LMArena Expert | 1488 | #34 of 273, top 13% | LMArena | 2026-10-08 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1474 | #15 of 297, top 6% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1530 | #14 of 285, top 5% | LMArena | 2026-10-08 | |
| LMArena Russian | 1484 | #16 of 283, top 6% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1460 | #34 of 298, top 12% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1482 | #19 of 291, top 7% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1476 | #19 of 297, top 7% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1449 | #29 of 295, top 10% | LMArena | 2026-10-08 | |
| EQ-Bench 4 | 1110 | #23 of 28, top 83% | EQ-Bench | ||
| LMArena Multi-Turn | 1481 | #18 of 295, top 7% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| alibaba | $2.50 | $7.50 | $0.50 | 2026-10-10 |
| deepinfra | $2.50 | $7.50 | $0.50 | 2026-10-10 |
| openrouter | $1.48 | $4.42 | $0.29 | 2026-10-10 |
| together | $1.25 | $3.75 | $0.13 | 2026-10-10 |
Compare Qwen3.7 Max
- Qwen3.7 Max vs Qwen3 Max
- Qwen3.7 Max vs GLM-5.3-Flash
- Qwen3.7 Max vs Qwen3.6 Max Preview
- Qwen3.7 Max vs Gemini 3 Flash Preview
- Qwen3.7 Max vs GLM-5.2
- Qwen3.7 Max vs GPT-5.2 Pro
- Qwen3.7 Max vs GPT-5
- Qwen3.7 Max vs GPT-6 Astra
- Qwen3.7 Max vs Claude Fable 5.1
- Qwen3.7 Max vs Gemini 3.8 Flash
- Qwen3.7 Max vs Kimi K3
- Qwen3.7 Max vs Grok 4.6
- Qwen3.7 Max vs GLM-5.3
- Qwen3.7 Max vs Muse Spark 1.3
Other Alibaba (Qwen) models
- Qwen3.8 Max56.8
- Qwen3.6 Max Preview51.5
- Qwen3.6 Plus47.5
- Qwen3.5 397B-A17B46.0
- Qwen3.8 27B46.0
- Qwen3.5 Max Preview45.3
- Qwen3.7 Plus45.3
- Qwen3 Max43.7
Frequently asked questions
How good is Qwen3.7 Max?
Qwen3.7 Max by Alibaba (Qwen) ranks 42nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 51.5. Its strongest category is multilingual, where it ranks 15th. API pricing starts at $2.50 per million input tokens and $7.50 per million output tokens, with a 1M-token context window.
How much does Qwen3.7 Max cost?
Qwen3.7 Max costs $2.50 per million input tokens and $7.50 per million output tokens on Alibaba (Qwen)'s own API, with cached input at $0.50.
What is Qwen3.7 Max's context window?
Qwen3.7 Max accepts up to 1M tokens of input and can write up to 131K tokens in one response.
Is Qwen3.7 Max open source?
No. Qwen3.7 Max is proprietary and available only through Alibaba (Qwen)'s API and partner platforms.
What are Qwen3.7 Max's strengths and weaknesses?
Relative to other ranked models, Qwen3.7 Max places best in multilingual, knowledge, math and lowest in agentic & tool use, writing & preference, long context.
What is Qwen3.7 Max best at?
Its best category is multilingual, where it ranks 15th on Noometry.