xAI, proprietary

# Grok 4.5

> Grok 4.5 by xAI, released July 2026. Ranked #25 of 354 with a Noometry Index of 55.0. API: $2 in / $6 out per M tokens. 500K context. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/grok-4-5
- Last updated: 2026-10-10
- Title: Grok 4.5 Benchmarks, Price & Rank (October 2026) | Noometry

Grok 4.5 by xAI ranks 25th of 354 ranked models on the Noometry Index as of October 2026, with a score of 55.0. Its strongest category is agentic & tool use, where it ranks 17th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 500K-token context window.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #25 of 354
- **Index score:** 55.0
- **Evidence:** Confirmed 52 results
- **Provider:** [xAI](https://noometry.com/providers/xai)
- **Released:** July 8, 2026
- **Weights:** Proprietary
- **Reasoning:** Yes
- **Context window:** 500K
- **Max output:** 500K
- **Input price:** $2 / M
- **Output price:** $6 / M
- **Blended price:** $3 / M
- **Output speed:** 4 tokens/s [Kagi](https://help.kagi.com/kagi/ai/llm-benchmark.html)
- **Value:** #155 of 219
- **Knowledge cutoff:** Unknown
- **Input:** text, image, pdf

## Category scores

Each category score combines every public result we have in that category.

Grok 4.5 category scores

1.  Coding 52.2
2.  Agentic & Tool Use 44.4
3.  Reasoning 56.1
4.  Math 60.9
5.  Knowledge 62.3
6.  Multimodal 37.6
7.  Multilingual 54.4
8.  Instruction Following 76.0
9.  Long Context 44.8
10.  Writing & Preference 65.8
11.  020406080

Grok 4.5 category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 52.2 | #35 | 6 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 44.4 | #17 | 5 |
| [Reasoning](https://noometry.com/best/reasoning) | 56.1 | #25 | 11 |
| [Math](https://noometry.com/best/math) | 60.9 | #35 | 5 |
| [Knowledge](https://noometry.com/best/knowledge) | 62.3 | #24 | 3 |
| [Multimodal](https://noometry.com/best/multimodal) | 37.6 | #72 | 3 |
| [Multilingual](https://noometry.com/best/multilingual) | 54.4 | #42 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 76.0 | #48 | 1 |
| [Long Context](https://noometry.com/best/long-context) | 44.8 | #56 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 65.8 | #42 | 4 |

## Strengths and weaknesses

Categories where Grok 4.5 places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Grok 4.5: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 56.1 | +32.5 | #25 of 350, top 8% |
| [Knowledge](https://noometry.com/best/knowledge) | 62.3 | +25.0 | #24 of 314, top 8% |
| [Coding](https://noometry.com/best/coding) | 52.2 | +13.5 | #35 of 340, top 11% |

### Weakest categories

Grok 4.5: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Multimodal](https://noometry.com/best/multimodal) | 37.6 | −0.9 | #72 of 128, top 57% |
| [Long Context](https://noometry.com/best/long-context) | 44.8 | +3.8 | #56 of 296, top 19% |
| [Instruction Following](https://noometry.com/best/instruction-following) | 76.0 | +4.8 | #48 of 305, top 16% |

## Closest competitors

The models ranked just above and below Grok 4.5. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Grok 4.5
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [Grok 4.6](https://noometry.com/models/grok-4-6) | #21 | 56.9 | $3 | — | [Compare](https://noometry.com/compare/grok-4-5-vs-grok-4-6) |
| [Qwen3.8 Max](https://noometry.com/models/qwen3-8-max) | #22 | 56.8 | $3 | — | [Compare](https://noometry.com/compare/grok-4-5-vs-qwen3-8-max) |
| [Gemini 3.1 Pro Preview](https://noometry.com/models/gemini-3-1-pro-preview) | #23 | 56.7 | $4.50 | — | [Compare](https://noometry.com/compare/gemini-3-1-pro-preview-vs-grok-4-5) |
| [Gemini 4 Argon](https://noometry.com/models/gemini-4-argon) | #24 | 56.5 | — | — | [Compare](https://noometry.com/compare/gemini-4-argon-vs-grok-4-5) |
| [GLM-5.3](https://noometry.com/models/glm-5-3) | #26 | 54.8 | $2.15 | — | [Compare](https://noometry.com/compare/glm-5-3-vs-grok-4-5) |
| [Muse Spark 1.3](https://noometry.com/models/muse-spark-1-3) | #27 | 54.8 | $2 | — | [Compare](https://noometry.com/compare/grok-4-5-vs-muse-spark-1-3) |
| [Gemini 3 Pro](https://noometry.com/models/gemini-3-pro) | #28 | 54.8 | — | 1 | [Compare](https://noometry.com/compare/gemini-3-pro-vs-grok-4-5) |
| [Claude Sonnet 5](https://noometry.com/models/claude-sonnet-5) | #29 | 54.6 | $4 | — | [Compare](https://noometry.com/compare/claude-sonnet-5-vs-grok-4-5) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Grok 4.5 Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [DeepSWE](https://noometry.com/benchmarks/deepswe) | 53.8% | #21 of 29, top 73% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [FrontierCode](https://noometry.com/benchmarks/frontiercode) | 42.4% | #17 of 37, top 46% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev) | 1553 | #33 of 113, top 30% |  | [LMArena](https://lmarena.ai/leaderboard/webdev) | 2026-10-08 |
| [SciCode](https://noometry.com/benchmarks/scicode) | 54.1% | #29 of 121, top 24% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [WeirdML](https://noometry.com/benchmarks/weirdml) | 46.4% | #56 of 119, top 48% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1474 | #52 of 294, top 18% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [ALE-Bench](https://noometry.com/benchmarks/ale-bench) | 1,309 | #24 of 105, top 23% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Agentic & Tool Use

Grok 4.5 Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [APEX-Agents](https://noometry.com/benchmarks/apex-agents) | 56.2% | #17 of 49, top 35% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [τ²-bench Banking](https://noometry.com/benchmarks/tau2-banking) | 47.9% | #3 of 26, top 12% | high | [τ²-bench](https://taubench.com/) | 2026-08-04 |
| [PostTrainBench](https://noometry.com/benchmarks/posttrainbench) | 23.4% | #9 of 11, top 82% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [GBAEval](https://noometry.com/benchmarks/gbaeval) | 65.4% | #4 of 23, top 18% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [GDP.pdf](https://noometry.com/benchmarks/gdp-pdf) | 14% | #31 of 36, top 87% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Search](https://noometry.com/benchmarks/arena-search) | 1213 | #8 of 32, top 25% |  | [LMArena](https://lmarena.ai/leaderboard/search) | 2026-08-24 |
| [Vending-Bench 2](https://noometry.com/benchmarks/vending-bench-2) | 3,887 | #36 of 60, top 60% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Reasoning

Grok 4.5 Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [ARC-AGI-2](https://noometry.com/benchmarks/arc-agi-2) | 52.6% | #34 of 83, top 41% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [ARC-AGI-2](https://noometry.com/benchmarks/arc-agi-2) | 33.1% |  | low | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [ARC-AGI-2](https://noometry.com/benchmarks/arc-agi-2) | 52.6% |  | medium | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [SimpleBench](https://noometry.com/benchmarks/simplebench) | 70% | #12 of 77, top 16% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Kagi LLM Benchmark](https://noometry.com/benchmarks/kagi-reasoning) | 83.5% | #5 of 99, top 6% |  | [Kagi LLM Benchmark](https://help.kagi.com/kagi/ai/llm-benchmark.html) |  |
| [NYT Connections (extended)](https://noometry.com/benchmarks/nyt-connections) | 79.9% | #36 of 91, top 40% | high reasoning | [Lech Mazur benchmarks](https://github.com/lechmazur/nyt-connections) |  |
| [ARC-AGI-1](https://noometry.com/benchmarks/arc-agi-1) | 85.7% |  | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [ARC-AGI-1](https://noometry.com/benchmarks/arc-agi-1) | 79.2% |  | low | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [ARC-AGI-1](https://noometry.com/benchmarks/arc-agi-1) | 87.2% | #32 of 83, top 39% | medium | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [CritPt](https://noometry.com/benchmarks/critpt) | 15.4% | #36 of 134, top 27% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 36% | #28 of 129, top 22% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-07-08 |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1462 | #43 of 297, top 15% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 96.5% | #10 of 151, top 7% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMCA](https://noometry.com/benchmarks/lmca) | 45.2% | #33 of 125, top 27% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Surface Evolver Bench](https://noometry.com/benchmarks/surface-evolver-bench) | 74.4% | #8 of 25, top 32% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 153.92 | #38 of 213, top 18% |  | [Epoch AI](https://epoch.ai/eci) | 2026-07-08 |

### Math

Grok 4.5 Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [FrontierMath (Tiers 1-3)](https://noometry.com/benchmarks/frontiermath) | 57.2% | #39 of 81, top 49% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-07-09 |
| [FrontierMath Tier 4](https://noometry.com/benchmarks/frontiermath-tier-4) | 24.4% | #39 of 63, top 62% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-07-09 |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 97.8% | #26 of 173, top 16% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-07-08 |
| [ProofBench](https://noometry.com/benchmarks/proofbench) | 31% | #42 of 77, top 55% | high | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1459 | #52 of 285, top 19% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Knowledge

Grok 4.5 Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 93.4% | #15 of 186, top 9% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-07-08 |
| [SimpleQA Verified](https://noometry.com/benchmarks/simpleqa-verified) | 48.3% | #29 of 77, top 38% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-27 |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1466 | #54 of 273, top 20% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multimodal

Grok 4.5 Multimodal benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Vision](https://noometry.com/benchmarks/arena-vision) | 1288 | #24 of 122, top 20% |  | [LMArena](https://lmarena.ai/leaderboard/vision) | 2026-10-09 |
| [Blueprint-Bench 2](https://noometry.com/benchmarks/blueprint-bench-2) | 27.3% | #18 of 31, top 59% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Furniture Assembly](https://noometry.com/benchmarks/furniture-assembly) | 22.5% | #28 of 31, top 91% | high | [Epoch AI](https://epoch.ai/benchmarks) | 2026-09-24 |
| [LMArena Document](https://noometry.com/benchmarks/arena-document) | 1452 | #21 of 38, top 56% |  | [LMArena](https://lmarena.ai/leaderboard/document) | 2026-09-13 |

### Multilingual

Grok 4.5 Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1440 | #41 of 297, top 14% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1496 | #48 of 285, top 17% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1456 | #57 of 223, top 26% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1446 | #49 of 231, top 22% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1428 | #34 of 211, top 17% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1404 | #49 of 213, top 24% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1448 | #46 of 283, top 17% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1450 | #50 of 226, top 23% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Grok 4.5 Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1446 | #44 of 298, top 15% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Grok 4.5 Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1463 | #39 of 291, top 14% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Grok 4.5 Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1448 | #47 of 297, top 16% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1442 | #34 of 295, top 12% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [EQ-Bench Creative Writing](https://noometry.com/benchmarks/eqbench-creative-writing) | 1579 | #45 of 115, top 40% |  | [EQ-Bench](https://eqbench.com/creative_writing.html) |  |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1456 | #44 of 295, top 15% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

## API pricing by provider

Grok 4.5 API prices
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
| --- | --- | --- | --- | --- |
| [openrouter](https://openrouter.ai/x-ai/grok-4.5) | $2 | $6 | $0.30 | 2026-10-10 |
| [xai](https://docs.x.ai/docs/models) | $2 | $6 | $0.30 | 2026-10-10 |

[All xAI API prices →](https://noometry.com/llm-pricing/xai) [Estimate your cost →](https://noometry.com/tools/cost-calculator)

## Compare Grok 4.5

-   [Grok 4.5 vs Grok 4.3](https://noometry.com/compare/grok-4-3-vs-grok-4-5)
-   [Grok 4.5 vs Gemini 4 Argon](https://noometry.com/compare/gemini-4-argon-vs-grok-4-5)
-   [Grok 4.5 vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-grok-4-5)
-   [Grok 4.5 vs Gemini 3.1 Pro Preview](https://noometry.com/compare/gemini-3-1-pro-preview-vs-grok-4-5)
-   [Grok 4.5 vs Muse Spark 1.3](https://noometry.com/compare/grok-4-5-vs-muse-spark-1-3)
-   [Grok 4.5 vs Qwen3.8 Max](https://noometry.com/compare/grok-4-5-vs-qwen3-8-max)
-   [Grok 4.5 vs Gemini 3 Pro](https://noometry.com/compare/gemini-3-pro-vs-grok-4-5)
-   [Grok 4.5 vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-grok-4-5)
-   [Grok 4.5 vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-grok-4-5)
-   [Grok 4.5 vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-grok-4-5)
-   [Grok 4.5 vs Kimi K3](https://noometry.com/compare/grok-4-5-vs-kimi-k3)
-   [Grok 4.5 vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-grok-4-5)
-   [Grok 4.5 vs MiMo-V2.6-Pro](https://noometry.com/compare/grok-4-5-vs-mimo-v2-6-pro)

## Other xAI models

-   [Grok 4.6](https://noometry.com/models/grok-4-6)56.9
-   [Grok 4.7](https://noometry.com/models/grok-4-7)53.1
-   [Grok 4.20 (Non-Reasoning)](https://noometry.com/models/grok-4-20)48.6
-   [Grok 4](https://noometry.com/models/grok-4)48.1
-   [Grok 4.20 Multi-Agent](https://noometry.com/models/grok-4-20-multi-agent)46.2
-   [Grok 4.3](https://noometry.com/models/grok-4-3)43.8
-   [Grok 4.1](https://noometry.com/models/grok-4-1)41.5
-   [Grok 4.1 Fast](https://noometry.com/models/grok-4-1-fast)41.4

## Frequently asked questions

### How good is Grok 4.5?

Grok 4.5 by xAI ranks 25th of 354 ranked models on the Noometry Index as of October 2026, with a score of 55.0. Its strongest category is agentic & tool use, where it ranks 17th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 500K-token context window.

### How much does Grok 4.5 cost?

Grok 4.5 costs $2 per million input tokens and $6 per million output tokens on xAI's own API, with cached input at $0.30.

### What is Grok 4.5's context window?

Grok 4.5 accepts up to 500K tokens of input and can write up to 500K tokens in one response.

### Is Grok 4.5 open source?

No. Grok 4.5 is proprietary and available only through xAI's API and partner platforms.

### How fast is Grok 4.5?

Grok 4.5 generated about 4 output tokens per second in the Kagi LLM Benchmark's timed runs. Speed varies by provider, load and reasoning effort.

### What are Grok 4.5's strengths and weaknesses?

Relative to other ranked models, Grok 4.5 places best in reasoning, knowledge, coding and lowest in multimodal, long context, instruction following.

### What is Grok 4.5 best at?

Its best category is agentic & tool use, where it ranks 17th on Noometry.

### Cite this page

Noometry. (2026). Grok 4.5 benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/grok-4-5

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/grok-4-5.md).
