StepFun, open weights

# Step 3

> Step 3 by StepFun. Ranked #149 of 354 with a Noometry Index of 40.5. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/step-3
- Last updated: 2026-10-10
- Title: Step 3 Benchmarks, Price & Rank (October 2026) | Noometry

Step 3 by StepFun ranks 149th of 354 ranked models on the Noometry Index as of October 2026, with a score of 40.5. Its strongest category is multimodal, where it ranks 86th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #149 of 354
- **Index score:** 40.5
- **Evidence:** Confirmed 17 results
- **Provider:** [![](/logos/stepfun.svg) StepFun](https://noometry.com/providers/stepfun)
- **Released:** Unknown
- **Weights:** Open weights
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** 7 tokens/s [Kagi](https://help.kagi.com/kagi/ai/llm-benchmark.html)
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Step 3 category scores

1.  Coding 40.1
2.  Reasoning 28.4
3.  Math 37.6
4.  Knowledge 36.8
5.  Multimodal 35.5
6.  Multilingual 46.3
7.  Instruction Following 70.4
8.  Long Context 40.3
9.  Writing & Preference 54.3
10.  020406080

Step 3 category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 40.1 | #147 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 28.4 | #105 | 2 |
| [Math](https://noometry.com/best/math) | 37.6 | #148 | 1 |
| [Knowledge](https://noometry.com/best/knowledge) | 36.8 | #164 | 1 |
| [Multimodal](https://noometry.com/best/multimodal) | 35.5 | #86 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 46.3 | #159 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 70.4 | #164 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 40.3 | #157 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 54.3 | #151 | 3 |

## Strengths and weaknesses

Categories where Step 3 places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Step 3: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 28.4 | +4.8 | #105 of 350, top 30% |
| [Coding](https://noometry.com/best/coding) | 40.1 | +1.4 | #147 of 340, top 44% |
| [Math](https://noometry.com/best/math) | 37.6 | +1.1 | #148 of 327, top 46% |

### Weakest categories

Step 3: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Multimodal](https://noometry.com/best/multimodal) | 35.5 | −3.0 | #86 of 128, top 68% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 70.4 | −0.9 | #164 of 305, top 54% |
| [Multilingual](https://noometry.com/best/multilingual) | 46.3 | −1.1 | #159 of 297, top 54% |

## Closest competitors

The models ranked just above and below Step 3. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Step 3
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Claude Sonnet 4](https://noometry.com/models/claude-sonnet-4) | #145 | 40.8 | $6 | 31 | [Compare](https://noometry.com/compare/claude-sonnet-4-vs-step-3) |
| [Qwen2.5-Max](https://noometry.com/models/qwen2-5-max) | #146 | 40.7 | — | — | [Compare](https://noometry.com/compare/qwen2-5-max-vs-step-3) |
| [Nemotron 3 Nano 30B A3B](https://noometry.com/models/nemotron-3-nano-30b-a3b) | #147 | 40.6 | $0.0875 | — | [Compare](https://noometry.com/compare/nemotron-3-nano-30b-a3b-vs-step-3) |
| [Granite 4.2 8B](https://noometry.com/models/granite-4-2-8b) | #148 | 40.5 | $0.11 | — | [Compare](https://noometry.com/compare/granite-4-2-8b-vs-step-3) |
| [MiniMax M1](https://noometry.com/models/minimax-m1) | #150 | 40.3 | $0.96 | — | [Compare](https://noometry.com/compare/minimax-m1-vs-step-3) |
| [Nvidia Llama 3.3 Nemotron Super 49b v1.5](https://noometry.com/models/nvidia-llama-3-3-nemotron-super-49b-v1-5) | #151 | 40.3 | $0.40 | — | [Compare](https://noometry.com/compare/nvidia-llama-3-3-nemotron-super-49b-v1-5-vs-step-3) |
| [Mistral Medium 3.5](https://noometry.com/models/mistral-medium-3-5) | #152 | 40.2 | $3 | 50 | [Compare](https://noometry.com/compare/mistral-medium-3-5-vs-step-3) |
| [Nemotron 3 Super](https://noometry.com/models/nemotron-3-super) | #153 | 40.1 | $0.17 | — | [Compare](https://noometry.com/compare/nemotron-3-super-vs-step-3) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Step 3 Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1367 | #161 of 294, top 55% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Reasoning

Step 3 Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Kagi LLM Benchmark](https://noometry.com/benchmarks/kagi-reasoning) | 62.3% | #38 of 99, top 39% |  | [Kagi LLM Benchmark](https://help.kagi.com/kagi/ai/llm-benchmark.html) |  |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1355 | #158 of 297, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Math

Step 3 Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1366 | #153 of 285, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Step 3 Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1333 | #164 of 273, top 61% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multimodal

Step 3 Multimodal benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Vision](https://noometry.com/benchmarks/arena-vision) | 1177 | #89 of 122, top 73% |  | [LMArena](https://lmarena.ai/leaderboard/vision) | 2026-10-09 |

### Multilingual

Step 3 Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1327 | #159 of 297, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1397 | #139 of 285, top 49% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1371 | #115 of 231, top 50% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1269 | #142 of 213, top 67% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1331 | #160 of 283, top 57% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1371 | #132 of 226, top 59% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Step 3 Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1332 | #157 of 298, top 53% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Step 3 Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1326 | #168 of 291, top 58% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Step 3 Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1350 | #159 of 297, top 54% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1321 | #151 of 295, top 52% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1341 | #160 of 295, top 55% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## Compare Step 3

-   [Step 3 vs Granite 4.2 8B](https://noometry.com/compare/granite-4-2-8b-vs-step-3)
-   [Step 3 vs MiniMax M1](https://noometry.com/compare/minimax-m1-vs-step-3)
-   [Step 3 vs Nemotron 3 Nano 30B A3B](https://noometry.com/compare/nemotron-3-nano-30b-a3b-vs-step-3)
-   [Step 3 vs Nvidia Llama 3.3 Nemotron Super 49b v1.5](https://noometry.com/compare/nvidia-llama-3-3-nemotron-super-49b-v1-5-vs-step-3)
-   [Step 3 vs Qwen2.5-Max](https://noometry.com/compare/qwen2-5-max-vs-step-3)
-   [Step 3 vs Mistral Medium 3.5](https://noometry.com/compare/mistral-medium-3-5-vs-step-3)
-   [Step 3 vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-step-3)
-   [Step 3 vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-step-3)
-   [Step 3 vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-step-3)
-   [Step 3 vs Kimi K3](https://noometry.com/compare/kimi-k3-vs-step-3)
-   [Step 3 vs Grok 4.6](https://noometry.com/compare/grok-4-6-vs-step-3)
-   [Step 3 vs Qwen3.8 Max](https://noometry.com/compare/qwen3-8-max-vs-step-3)
-   [Step 3 vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-step-3)
-   [Step 3 vs Muse Spark 1.3](https://noometry.com/compare/muse-spark-1-3-vs-step-3)

## Other StepFun models

-   [Step 5 Preview](https://noometry.com/models/step-5-preview)47.9
-   [Step 3.5 Flash](https://noometry.com/models/step-3-5-flash)42.3
-   [Step 1o Turbo 202506](https://noometry.com/models/step-1o-turbo-202506)39.7
-   [Step 2 16k Exp 202412](https://noometry.com/models/step-2-16k-exp-202412)39.2
-   [Step 3.7 Flash](https://noometry.com/models/step-3-7-flash)37.3
-   [Step 1 (32K)](https://noometry.com/models/step-1)
-   [Step 2 (16K)](https://noometry.com/models/step-2)

## Frequently asked questions

### How good is Step 3?

Step 3 by StepFun ranks 149th of 354 ranked models on the Noometry Index as of October 2026, with a score of 40.5. Its strongest category is multimodal, where it ranks 86th.

### Is Step 3 open source?

Yes. Step 3's weights are downloadable; check the license for commercial terms.

### How fast is Step 3?

Step 3 generated about 7 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

### What are Step 3's strengths and weaknesses?

Relative to other ranked models, Step 3 places best in reasoning, coding, math and lowest in multimodal, instruction following, multilingual.

### What is Step 3 best at?

Its best category is multimodal, where it ranks 86th on Noometry.

### Cite this page

Noometry. (2026). Step 3 benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/step-3

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/step-3.md).
