Alibaba (Qwen), open weights

# Qwen3 32B

> Qwen3 32B by Alibaba (Qwen), released April 2025. Ranked #172 of 354 with a Noometry Index of 39.2. API: $0.70 in / $2.80 out per M tokens. 131K context. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/qwen3-32b
- Last updated: 2026-10-10
- Title: Qwen3 32B Benchmarks, Price & Rank (October 2026) | Noometry

Qwen3 32B by Alibaba (Qwen) ranks 172nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.2. Its strongest category is agentic & tool use, where it ranks 62nd. API pricing starts at $0.70 per million input tokens and $2.80 per million output tokens, with a 131K-token context window.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #172 of 354
- **Index score:** 39.2
- **Evidence:** Confirmed 26 results
- **Provider:** [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba)
- **Released:** April 1, 2025
- **Weights:** Open weights
- **Reasoning:** Yes
- **Context window:** 131K
- **Max output:** 16K
- **Input price:** $0.70 / M
- **Output price:** $2.80 / M
- **Blended price:** $1.22 / M
- **Output speed:** 86 tokens/s [Kagi](https://help.kagi.com/kagi/ai/llm-benchmark.html)
- **Value:** #130 of 219
- **Knowledge cutoff:** April 2025
- **Input:** text
- **Hugging Face:** [Qwen/Qwen3-32B](https://huggingface.co/Qwen/Qwen3-32B)

## Category scores

Each category score combines every public result we have in that category.

Qwen3 32B category scores

1.  Coding 37.7
2.  Agentic & Tool Use 32.6
3.  Reasoning 20.2
4.  Math 39.7
5.  Knowledge 40.0
6.  Multilingual 45.6
7.  Instruction Following 68.9
8.  Long Context 43.8
9.  Writing & Preference 52.9
10.  020406080

Qwen3 32B category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 37.7 | #190 | 3 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 32.6 | #62 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 20.2 | #241 | 6 |
| [Math](https://noometry.com/best/math) | 39.7 | #99 | 2 |
| [Knowledge](https://noometry.com/best/knowledge) | 40.0 | #125 | 3 |
| [Multilingual](https://noometry.com/best/multilingual) | 45.6 | #167 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 68.9 | #179 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 43.8 | #87 | 2 |
| [Writing & Preference](https://noometry.com/best/writing) | 52.9 | #163 | 3 |

## Strengths and weaknesses

Categories where Qwen3 32B places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Qwen3 32B: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Long Context](https://noometry.com/best/long-context) | 43.8 | +2.8 | #87 of 296, top 30% |
| [Math](https://noometry.com/best/math) | 39.7 | +3.1 | #99 of 327, top 31% |
| [Knowledge](https://noometry.com/best/knowledge) | 40.0 | +2.7 | #125 of 314, top 40% |

### Weakest categories

Qwen3 32B: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 20.2 | −3.4 | #241 of 350, top 69% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 68.9 | −2.4 | #179 of 305, top 59% |
| [Multilingual](https://noometry.com/best/multilingual) | 45.6 | −1.8 | #167 of 297, top 57% |

## Closest competitors

The models ranked just above and below Qwen3 32B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen3 32B
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Olmo 3.1 32b Instruct](https://noometry.com/models/olmo-3-1-32b-instruct) | #168 | 39.4 | — | — | [Compare](https://noometry.com/compare/olmo-3-1-32b-instruct-vs-qwen3-32b) |
| [Granite 4.2 3b](https://noometry.com/models/granite-4-2-3b) | #169 | 39.4 | — | — | [Compare](https://noometry.com/compare/granite-4-2-3b-vs-qwen3-32b) |
| [Gemini 2.5 Flash](https://noometry.com/models/gemini-2-5-flash) | #170 | 39.3 | $0.85 | 152 | [Compare](https://noometry.com/compare/gemini-2-5-flash-vs-qwen3-32b) |
| [Step 2 16k Exp 202412](https://noometry.com/models/step-2-16k-exp-202412) | #171 | 39.2 | — | — | [Compare](https://noometry.com/compare/qwen3-32b-vs-step-2-16k-exp-202412) |
| [Gemini 2.0 Pro](https://noometry.com/models/gemini-2-0-pro) | #173 | 39.1 | — | — | [Compare](https://noometry.com/compare/gemini-2-0-pro-vs-qwen3-32b) |
| [Molmo 2 8b](https://noometry.com/models/molmo-2-8b) | #174 | 39.1 | — | — | [Compare](https://noometry.com/compare/molmo-2-8b-vs-qwen3-32b) |
| [Mercury 2](https://noometry.com/models/mercury-2) | #175 | 39.1 | $0.38 | — | [Compare](https://noometry.com/compare/mercury-2-vs-qwen3-32b) |
| [Mistral Large 3](https://noometry.com/models/mistral-large-3) | #176 | 39.1 | $0.38 | 7 | [Compare](https://noometry.com/compare/mistral-large-3-vs-qwen3-32b) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Qwen3 32B Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Aider Polyglot](https://noometry.com/benchmarks/aider-polyglot) | 40% | #28 of 44, top 64% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [SciCode](https://noometry.com/benchmarks/scicode) | 35.4% | #99 of 121, top 82% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1358 | #170 of 294, top 58% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Agentic & Tool Use

Qwen3 32B Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Berkeley Function Calling Leaderboard](https://noometry.com/benchmarks/bfcl) | 48.7% | #20 of 49, top 41% | fc | [Berkeley Function Calling Leaderboard](https://gorilla.cs.berkeley.edu/leaderboard.html) |  |

### Reasoning

Qwen3 32B Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Kagi LLM Benchmark](https://noometry.com/benchmarks/kagi-reasoning) | 54.9% | #53 of 99, top 54% |  | [Kagi LLM Benchmark](https://help.kagi.com/kagi/ai/llm-benchmark.html) |  |
| [Kagi LLM Benchmark](https://noometry.com/benchmarks/kagi-reasoning) | 48.7% |  |  | [Kagi LLM Benchmark](https://help.kagi.com/kagi/ai/llm-benchmark.html) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 0.3% | #98 of 134, top 74% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 5% | #94 of 129, top 73% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-28 |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 1% |  | none | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-28 |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1334 | #170 of 297, top 58% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 67.5% | #99 of 151, top 66% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMCA](https://noometry.com/benchmarks/lmca) | 17.3% | #98 of 125, top 79% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 138.51 | #113 of 213, top 54% |  | [Epoch AI](https://epoch.ai/eci) | 2025-04-29 |

### Math

Qwen3 32B Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 66.9% | #96 of 173, top 56% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-30 |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 23.1% |  | none | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-30 |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1399 | #126 of 285, top 45% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Qwen3 32B Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 65.7% | #105 of 186, top 57% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-28 |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 54.1% |  | none | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-30 |
| [Vectara Hallucination Rate](https://noometry.com/benchmarks/vectara-hallucination) (lower is better) | 5.9% | #21 of 96, top 22% |  | [Vectara Hallucination Leaderboard](https://github.com/vectara/hallucination-leaderboard) |  |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1362 | #146 of 273, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multilingual

Qwen3 32B Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1317 | #167 of 297, top 57% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1357 | #163 of 285, top 58% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1341 | #136 of 231, top 59% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1311 | #171 of 283, top 61% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Qwen3 32B Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1305 | #173 of 298, top 59% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Qwen3 32B Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Fiction.LiveBench](https://noometry.com/benchmarks/fiction-livebench) | 74.2% | #16 of 47, top 35% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1327 | #165 of 291, top 57% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Qwen3 32B Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1340 | #163 of 297, top 55% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1297 | #167 of 295, top 57% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1331 | #172 of 295, top 59% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## API pricing by provider

Qwen3 32B API prices
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
| --- | --- | --- | --- | --- |
| [alibaba](https://www.alibabacloud.com/help/en/model-studio/models) | $0.70 | $2.80 | — | 2026-10-10 |
| [bedrock](https://docs.aws.amazon.com/bedrock/latest/userguide/models-supported.html) | $0.15 | $0.60 | — | 2026-10-10 |
| [deepinfra](https://deepinfra.com/models) | $0.08 | $0.28 | — | 2026-10-10 |
| [openrouter](https://openrouter.ai/qwen/qwen3-32b) | $0.08 | $0.28 | — | 2026-10-10 |

[All Alibaba (Qwen) API prices →](https://noometry.com/llm-pricing/alibaba) [Estimate your cost →](https://noometry.com/tools/cost-calculator)

## Compare Qwen3 32B

-   [Qwen3 32B vs Qwen2-72B](https://noometry.com/compare/qwen2-72b-vs-qwen3-32b)
-   [Qwen3 32B vs Step 2 16k Exp 202412](https://noometry.com/compare/qwen3-32b-vs-step-2-16k-exp-202412)
-   [Qwen3 32B vs Gemini 2.0 Pro](https://noometry.com/compare/gemini-2-0-pro-vs-qwen3-32b)
-   [Qwen3 32B vs Gemini 2.5 Flash](https://noometry.com/compare/gemini-2-5-flash-vs-qwen3-32b)
-   [Qwen3 32B vs Molmo 2 8b](https://noometry.com/compare/molmo-2-8b-vs-qwen3-32b)
-   [Qwen3 32B vs Granite 4.2 3b](https://noometry.com/compare/granite-4-2-3b-vs-qwen3-32b)
-   [Qwen3 32B vs Mercury 2](https://noometry.com/compare/mercury-2-vs-qwen3-32b)
-   [Qwen3 32B vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-qwen3-32b)
-   [Qwen3 32B vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-qwen3-32b)
-   [Qwen3 32B vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-qwen3-32b)
-   [Qwen3 32B vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-qwen3-32b)
-   [Qwen3 32B vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-qwen3-32b)
-   [Qwen3 32B vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-qwen3-32b)
-   [Qwen3 32B vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-qwen3-32b)

## Other Alibaba (Qwen) models

-   [Qwen3.8 Max](https://noometry.com/models/qwen3-8-max)56.8
-   [Qwen3.7 Max](https://noometry.com/models/qwen3-7-max)51.5
-   [Qwen3.6 Max Preview](https://noometry.com/models/qwen3-6-max-preview)51.5
-   [Qwen3.6 Plus](https://noometry.com/models/qwen3-6-plus)47.5
-   [Qwen3.5 397B-A17B](https://noometry.com/models/qwen3-5-397b-a17b)46.0
-   [Qwen3.8 27B](https://noometry.com/models/qwen3-8-27b)46.0
-   [Qwen3.5 Max Preview](https://noometry.com/models/qwen3-5-max-preview)45.3
-   [Qwen3.7 Plus](https://noometry.com/models/qwen3-7-plus)45.3

## Frequently asked questions

### How good is Qwen3 32B?

Qwen3 32B by Alibaba (Qwen) ranks 172nd of 354 ranked models on the Noometry Index as of October 2026, with a score of 39.2. Its strongest category is agentic & tool use, where it ranks 62nd. API pricing starts at $0.70 per million input tokens and $2.80 per million output tokens, with a 131K-token context window.

### How much does Qwen3 32B cost?

Qwen3 32B costs $0.70 per million input tokens and $2.80 per million output tokens on Alibaba (Qwen)'s own API.

### What is Qwen3 32B's context window?

Qwen3 32B accepts up to 131K tokens of input and can write up to 16K tokens in one response.

### Is Qwen3 32B open source?

Yes. Qwen3 32B's weights are downloadable from Hugging Face (Qwen/Qwen3-32B); check the license for commercial terms.

### How fast is Qwen3 32B?

Qwen3 32B generated about 86 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

### What are Qwen3 32B's strengths and weaknesses?

Relative to other ranked models, Qwen3 32B places best in long context, math, knowledge and lowest in reasoning, instruction following, multilingual.

### What is Qwen3 32B best at?

Its best category is agentic & tool use, where it ranks 62nd on Noometry.

### Cite this page

Noometry. (2026). Qwen3 32B benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/qwen3-32b

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/qwen3-32b.md).
