xAI, proprietary

# Grok-2 (Dec 2024)

> Grok-2 (Dec 2024) by xAI, released August 2024. Ranked #239 of 354 with a Noometry Index of 33.7. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/grok-2
- Last updated: 2026-10-10
- Title: Grok-2 (Dec 2024) Benchmarks, Price & Rank (October 2026)

Grok-2 (Dec 2024) by xAI ranks 239th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.7. Its strongest category is multilingual, where it ranks 188th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #239 of 354
- **Index score:** 33.7
- **Evidence:** Confirmed 34 results
- **Provider:** [xAI](https://noometry.com/providers/xai)
- **Released:** August 13, 2024
- **Weights:** Proprietary
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Grok-2 (Dec 2024) category scores

1.  Coding 33.3
2.  Reasoning 16.9
3.  Math 20.8
4.  Knowledge 29.8
5.  Multilingual 43.1
6.  Instruction Following 66.9
7.  Long Context 38.8
8.  Writing & Preference 48.6
9.  020406080

Grok-2 (Dec 2024) category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 33.3 | #258 | 3 |
| [Reasoning](https://noometry.com/best/reasoning) | 16.9 | #299 | 5 |
| [Math](https://noometry.com/best/math) | 20.8 | #284 | 4 |
| [Knowledge](https://noometry.com/best/knowledge) | 29.8 | #233 | 3 |
| [Multilingual](https://noometry.com/best/multilingual) | 43.1 | #188 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 66.9 | #202 | 2 |
| [Long Context](https://noometry.com/best/long-context) | 38.8 | #190 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 48.6 | #198 | 5 |

## Strengths and weaknesses

Categories where Grok-2 (Dec 2024) places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Grok-2 (Dec 2024): strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Multilingual](https://noometry.com/best/multilingual) | 43.1 | −4.3 | #188 of 297, top 64% |
| [Writing & Preference](https://noometry.com/best/writing) | 48.6 | −5.1 | #198 of 312, top 64% |
| [Long Context](https://noometry.com/best/long-context) | 38.8 | −2.2 | #190 of 296, top 65% |

### Weakest categories

Grok-2 (Dec 2024): weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Math](https://noometry.com/best/math) | 20.8 | −15.8 | #284 of 327, top 87% |
| [Reasoning](https://noometry.com/best/reasoning) | 16.9 | −6.7 | #299 of 350, top 86% |
| [Coding](https://noometry.com/best/coding) | 33.3 | −5.4 | #258 of 340, top 76% |

## Closest competitors

The models ranked just above and below Grok-2 (Dec 2024). When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Grok-2 (Dec 2024)
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [o1-mini](https://noometry.com/models/o1-mini) | #235 | 34.0 | — | — | [Compare](https://noometry.com/compare/grok-2-vs-o1-mini) |
| [Qwen3.5-9B](https://noometry.com/models/qwen3-5-9b) | #236 | 33.8 | $0.11 | — | [Compare](https://noometry.com/compare/grok-2-vs-qwen3-5-9b) |
| [Codellama 70b Instruct](https://noometry.com/models/codellama-70b-instruct) | #237 | 33.7 | — | — | [Compare](https://noometry.com/compare/codellama-70b-instruct-vs-grok-2) |
| [Qwen3 8B](https://noometry.com/models/qwen3-8b) | #238 | 33.7 | $0.31 | — | [Compare](https://noometry.com/compare/grok-2-vs-qwen3-8b) |
| [GPT-4.1 mini](https://noometry.com/models/gpt-4-1-mini) | #240 | 33.6 | $0.70 | 86 | [Compare](https://noometry.com/compare/gpt-4-1-mini-vs-grok-2) |
| [GPT-5 Nano](https://noometry.com/models/gpt-5-nano) | #241 | 33.5 | $0.14 | 4 | [Compare](https://noometry.com/compare/gpt-5-nano-vs-grok-2) |
| [Mercury 2.5](https://noometry.com/models/mercury-2-5) | #242 | 33.5 | $0.0675 | — | [Compare](https://noometry.com/compare/grok-2-vs-mercury-2-5) |
| [Mistral Small](https://noometry.com/models/mistral-small) | #243 | 33.4 | $0.26 | 120 | [Compare](https://noometry.com/compare/grok-2-vs-mistral-small) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Grok-2 (Dec 2024) Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [WeirdML](https://noometry.com/benchmarks/weirdml) | 22.2% | #102 of 119, top 86% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Coding](https://noometry.com/benchmarks/livebench-coding) | 46.4% | #21 of 39, top 54% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1287 | #207 of 294, top 71% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Reasoning

Grok-2 (Dec 2024) Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [SimpleBench](https://noometry.com/benchmarks/simplebench) | 22.7% | #70 of 77, top 91% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Reasoning](https://noometry.com/benchmarks/livebench-reasoning) | 54.8% | #16 of 39, top 42% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1272 | #204 of 297, top 69% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 65.2% | #100 of 151, top 67% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Data Analysis](https://noometry.com/benchmarks/livebench-data-analysis) | 54.5% | #19 of 39, top 49% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 130.48 | #136 of 213, top 64% |  | [Epoch AI](https://epoch.ai/eci) | 2024-12-12 |
| [LiveBench](https://noometry.com/benchmarks/livebench) | 54.3% | #17 of 39, top 44% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Math

Grok-2 (Dec 2024) Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 11.5% | #132 of 173, top 77% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-02-25 |
| [LiveBench Math](https://noometry.com/benchmarks/livebench-math) | 54.9% | #18 of 39, top 47% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1283 | #190 of 285, top 67% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [MATH Level 5](https://noometry.com/benchmarks/math-level-5) | 63.5% | #36 of 79, top 46% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-02-17 |
| [FrontierMath (Feb 2025 set)](https://noometry.com/benchmarks/frontiermath-2025-02) | 0.7% | #61 of 68, top 90% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-03-06 |

### Knowledge

Grok-2 (Dec 2024) Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 53.8% | #123 of 186, top 67% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-01-27 |
| [Confabulations](https://noometry.com/benchmarks/confabulations) (lower is better) | 20.1% | #33 of 51, top 65% |  | [Lech Mazur benchmarks](https://github.com/lechmazur/confabulations) |  |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1254 | #193 of 273, top 71% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Multilingual

Grok-2 (Dec 2024) Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1282 | #188 of 297, top 64% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1289 | #191 of 285, top 68% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1318 | #156 of 223, top 70% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1287 | #155 of 231, top 68% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1244 | #143 of 211, top 68% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1237 | #148 of 213, top 70% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1286 | #185 of 283, top 66% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1281 | #168 of 226, top 75% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Grok-2 (Dec 2024) Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LiveBench Instruction Following](https://noometry.com/benchmarks/livebench-if) | 69.6% | #18 of 39, top 47% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1270 | #196 of 298, top 66% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Grok-2 (Dec 2024) Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1276 | #205 of 291, top 71% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Grok-2 (Dec 2024) Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1305 | #188 of 297, top 64% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1284 | #184 of 295, top 63% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [Short-Story Creative Writing](https://noometry.com/benchmarks/lech-mazur-writing) | 63.6% | #35 of 39, top 90% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1290 | #194 of 295, top 66% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LiveBench Language](https://noometry.com/benchmarks/livebench-language) | 45.6% | #15 of 39, top 39% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

## Compare Grok-2 (Dec 2024)

-   [Grok-2 (Dec 2024) vs Qwen3 8B](https://noometry.com/compare/grok-2-vs-qwen3-8b)
-   [Grok-2 (Dec 2024) vs GPT-4.1 mini](https://noometry.com/compare/gpt-4-1-mini-vs-grok-2)
-   [Grok-2 (Dec 2024) vs Codellama 70b Instruct](https://noometry.com/compare/codellama-70b-instruct-vs-grok-2)
-   [Grok-2 (Dec 2024) vs GPT-5 Nano](https://noometry.com/compare/gpt-5-nano-vs-grok-2)
-   [Grok-2 (Dec 2024) vs Qwen3.5-9B](https://noometry.com/compare/grok-2-vs-qwen3-5-9b)
-   [Grok-2 (Dec 2024) vs Mercury 2.5](https://noometry.com/compare/grok-2-vs-mercury-2-5)
-   [Grok-2 (Dec 2024) vs GPT-6 Astra](https://noometry.com/compare/gpt-6-astra-vs-grok-2)
-   [Grok-2 (Dec 2024) vs Claude Fable 5.1](https://noometry.com/compare/claude-fable-5-1-vs-grok-2)
-   [Grok-2 (Dec 2024) vs Gemini 3.8 Flash](https://noometry.com/compare/gemini-3-8-flash-vs-grok-2)
-   [Grok-2 (Dec 2024) vs Kimi K3](https://noometry.com/compare/grok-2-vs-kimi-k3)
-   [Grok-2 (Dec 2024) vs Qwen3.8 Max](https://noometry.com/compare/grok-2-vs-qwen3-8-max)
-   [Grok-2 (Dec 2024) vs GLM-5.3](https://noometry.com/compare/glm-5-3-vs-grok-2)
-   [Grok-2 (Dec 2024) vs Muse Spark 1.3](https://noometry.com/compare/grok-2-vs-muse-spark-1-3)
-   [Grok-2 (Dec 2024) vs DeepSeek V4 Pro](https://noometry.com/compare/deepseek-v4-pro-vs-grok-2)

## Other xAI models

-   [Grok 4.6](https://noometry.com/models/grok-4-6)56.9
-   [Grok 4.5](https://noometry.com/models/grok-4-5)55.0
-   [Grok 4.7](https://noometry.com/models/grok-4-7)53.1
-   [Grok 4.20 (Non-Reasoning)](https://noometry.com/models/grok-4-20)48.6
-   [Grok 4](https://noometry.com/models/grok-4)48.1
-   [Grok 4.20 Multi-Agent](https://noometry.com/models/grok-4-20-multi-agent)46.2
-   [Grok 4.3](https://noometry.com/models/grok-4-3)43.8
-   [Grok 4.1](https://noometry.com/models/grok-4-1)41.5

## Frequently asked questions

### How good is Grok-2 (Dec 2024)?

Grok-2 (Dec 2024) by xAI ranks 239th of 354 ranked models on the Noometry Index as of October 2026, with a score of 33.7. Its strongest category is multilingual, where it ranks 188th.

### Is Grok-2 (Dec 2024) open source?

No. Grok-2 (Dec 2024) is proprietary and available only through xAI's API and partner platforms.

### What are Grok-2 (Dec 2024)'s strengths and weaknesses?

Relative to other ranked models, Grok-2 (Dec 2024) places best in multilingual, writing & preference, long context and lowest in math, reasoning, coding.

### What is Grok-2 (Dec 2024) best at?

Its best category is multilingual, where it ranks 188th on Noometry.

### Cite this page

Noometry. (2026). Grok-2 (Dec 2024) benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/grok-2

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/grok-2.md).
