Alibaba (Qwen), open weights

# Qwen3 8B

> Qwen3 8B by Alibaba (Qwen), released April 2025. Ranked #238 of 354 with a Noometry Index of 33.7. API: $0.18 in / $0.70 out per M tokens. 131K context. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/qwen3-8b
- Last updated: 2026-10-10
- Title: Qwen3 8B Benchmarks, Price & Rank (October 2026) | Noometry

Qwen3 8B by Alibaba (Qwen) ranks 238th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.7. Its strongest category is agentic & tool use, where it ranks 78th. API pricing starts at $0.18 per million input tokens and $0.70 per million output tokens, with a 131K-token context window.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #238 of 354
- **Index score:** 33.7
- **Evidence:** Confirmed 11 results
- **Provider:** [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba)
- **Released:** April 1, 2025
- **Weights:** Open weights
- **Reasoning:** Yes
- **Context window:** 131K
- **Max output:** 8K
- **Input price:** $0.18 / M
- **Output price:** $0.70 / M
- **Blended price:** $0.31 / M
- **Output speed:** Not measured
- **Value:** #55 of 219
- **Knowledge cutoff:** April 2025
- **Input:** text

## Category scores

Each category score combines every public result we have in that category.

Qwen3 8B category scores

1.  Coding 34.0
2.  Agentic & Tool Use 30.2
3.  Reasoning 16.6
4.  Math 34.9
5.  Knowledge 36.1
6.  Long Context 37.9
7.  010203040

Qwen3 8B category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 34.0 | #248 | 1 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 30.2 | #78 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 16.6 | #303 | 4 |
| [Math](https://noometry.com/best/math) | 34.9 | #191 | 1 |
| [Knowledge](https://noometry.com/best/knowledge) | 36.1 | #173 | 2 |
| [Long Context](https://noometry.com/best/long-context) | 37.9 | #210 | 1 |

## Strengths and weaknesses

Categories where Qwen3 8B places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Qwen3 8B: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 30.2 | −0.2 | #78 of 154, top 51% |
| [Knowledge](https://noometry.com/best/knowledge) | 36.1 | −1.2 | #173 of 314, top 56% |
| [Math](https://noometry.com/best/math) | 34.9 | −1.7 | #191 of 327, top 59% |

### Weakest categories

Qwen3 8B: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 16.6 | −7.0 | #303 of 350, top 87% |
| [Coding](https://noometry.com/best/coding) | 34.0 | −4.7 | #248 of 340, top 73% |
| [Long Context](https://noometry.com/best/long-context) | 37.9 | −3.0 | #210 of 296, top 71% |

## Closest competitors

The models ranked just above and below Qwen3 8B. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Qwen3 8B
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Qwen1.5-110B](https://noometry.com/models/qwen1-5-110b) | #234 | 34.2 | — | — | [Compare](https://noometry.com/compare/qwen1-5-110b-vs-qwen3-8b) |
| [o1-mini](https://noometry.com/models/o1-mini) | #235 | 34.0 | — | — | [Compare](https://noometry.com/compare/o1-mini-vs-qwen3-8b) |
| [Qwen3.5-9B](https://noometry.com/models/qwen3-5-9b) | #236 | 33.8 | $0.11 | — | [Compare](https://noometry.com/compare/qwen3-5-9b-vs-qwen3-8b) |
| [Codellama 70b Instruct](https://noometry.com/models/codellama-70b-instruct) | #237 | 33.7 | — | — | [Compare](https://noometry.com/compare/codellama-70b-instruct-vs-qwen3-8b) |
| [Grok-2 (Dec 2024)](https://noometry.com/models/grok-2) | #239 | 33.7 | — | — | [Compare](https://noometry.com/compare/grok-2-vs-qwen3-8b) |
| [GPT-4.1 mini](https://noometry.com/models/gpt-4-1-mini) | #240 | 33.6 | $0.70 | 86 | [Compare](https://noometry.com/compare/gpt-4-1-mini-vs-qwen3-8b) |
| [GPT-5 Nano](https://noometry.com/models/gpt-5-nano) | #241 | 33.5 | $0.14 | 4 | [Compare](https://noometry.com/compare/gpt-5-nano-vs-qwen3-8b) |
| [Mercury 2.5](https://noometry.com/models/mercury-2-5) | #242 | 33.5 | $0.0675 | — | [Compare](https://noometry.com/compare/mercury-2-5-vs-qwen3-8b) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Qwen3 8B Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [SciCode](https://noometry.com/benchmarks/scicode) | 22.6% | #116 of 121, top 96% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Agentic & Tool Use

Qwen3 8B Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Berkeley Function Calling Leaderboard](https://noometry.com/benchmarks/bfcl) | 42.6% | #21 of 49, top 43% | fc | [Berkeley Function Calling Leaderboard](https://gorilla.cs.berkeley.edu/leaderboard.html) |  |

### Reasoning

Qwen3 8B Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [CritPt](https://noometry.com/benchmarks/critpt) | 0% | #132 of 134, top 99% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 5% | #95 of 129, top 74% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 0% |  | none | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 59.7% | #117 of 151, top 78% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMCA](https://noometry.com/benchmarks/lmca) | 8.8% | #116 of 125, top 93% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 136.17 | #121 of 213, top 57% |  | [Epoch AI](https://epoch.ai/eci) | 2025-04-28 |

### Math

Qwen3 8B Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 56.1% | #107 of 173, top 62% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 22.2% |  | none | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |

### Knowledge

Qwen3 8B Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 56.8% | #117 of 186, top 63% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 47.5% |  | none | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [Vectara Hallucination Rate](https://noometry.com/benchmarks/vectara-hallucination) (lower is better) | 4.8% | #7 of 96, top 8% |  | [Vectara Hallucination Leaderboard](https://github.com/vectara/hallucination-leaderboard) |  |

### Long Context

Qwen3 8B Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Fiction.LiveBench](https://noometry.com/benchmarks/fiction-livebench) | 62.1% | #27 of 47, top 58% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

## API pricing by provider

Qwen3 8B API prices
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
| --- | --- | --- | --- | --- |
| [alibaba](https://www.alibabacloud.com/help/en/model-studio/models) | $0.18 | $0.70 | — | 2026-10-10 |

[All Alibaba (Qwen) API prices →](https://noometry.com/llm-pricing/alibaba) [Estimate your cost →](https://noometry.com/tools/cost-calculator)

## Compare Qwen3 8B

-   [Qwen3 8B vs Qwen2-72B](https://noometry.com/compare/qwen2-72b-vs-qwen3-8b)
-   [Qwen3 8B vs Codellama 70b Instruct](https://noometry.com/compare/codellama-70b-instruct-vs-qwen3-8b)
-   [Qwen3 8B vs Grok-2 (Dec 2024)](https://noometry.com/compare/grok-2-vs-qwen3-8b)
-   [Qwen3 8B vs Qwen3.5-9B](https://noometry.com/compare/qwen3-5-9b-vs-qwen3-8b)
-   [Qwen3 8B vs GPT-4.1 mini](https://noometry.com/compare/gpt-4-1-mini-vs-qwen3-8b)
-   [Qwen3 8B vs o1-mini](https://noometry.com/compare/o1-mini-vs-qwen3-8b)
-   [Qwen3 8B vs GPT-5 Nano](https://noometry.com/compare/gpt-5-nano-vs-qwen3-8b)
-   [Qwen3 8B vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-qwen3-8b)
-   [Qwen3 8B vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-qwen3-8b)
-   [Qwen3 8B vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-qwen3-8b)
-   [Qwen3 8B vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-qwen3-8b)
-   [Qwen3 8B vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-qwen3-8b)
-   [Qwen3 8B vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-qwen3-8b)
-   [Qwen3 8B vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-qwen3-8b)

## Other Alibaba (Qwen) models

-   [Qwen3.8 Max](https://noometry.com/models/qwen3-8-max)56.8
-   [Qwen3.7 Max](https://noometry.com/models/qwen3-7-max)51.5
-   [Qwen3.6 Max Preview](https://noometry.com/models/qwen3-6-max-preview)51.5
-   [Qwen3.6 Plus](https://noometry.com/models/qwen3-6-plus)47.5
-   [Qwen3.5 397B-A17B](https://noometry.com/models/qwen3-5-397b-a17b)46.0
-   [Qwen3.8 27B](https://noometry.com/models/qwen3-8-27b)46.0
-   [Qwen3.5 Max Preview](https://noometry.com/models/qwen3-5-max-preview)45.3
-   [Qwen3.7 Plus](https://noometry.com/models/qwen3-7-plus)45.3

## Frequently asked questions

### How good is Qwen3 8B?

Qwen3 8B by Alibaba (Qwen) ranks 238th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.7. Its strongest category is agentic & tool use, where it ranks 78th. API pricing starts at $0.18 per million input tokens and $0.70 per million output tokens, with a 131K-token context window.

### How much does Qwen3 8B cost?

Qwen3 8B costs $0.18 per million input tokens and $0.70 per million output tokens on Alibaba (Qwen)'s own API.

### What is Qwen3 8B's context window?

Qwen3 8B accepts up to 131K tokens of input and can write up to 8K tokens in one response.

### Is Qwen3 8B open source?

Yes. Qwen3 8B's weights are downloadable; check the license for commercial terms.

### What are Qwen3 8B's strengths and weaknesses?

Relative to other ranked models, Qwen3 8B places best in agentic & tool use, knowledge, math and lowest in reasoning, coding, long context.

### What is Qwen3 8B best at?

Its best category is agentic & tool use, where it ranks 78th on Noometry.

### Cite this page

Noometry. (2026). Qwen3 8B benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/qwen3-8b

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/qwen3-8b.md).
