Model comparison

GPT-5.2 Pro vs Grok 4.7

GPT-5.2 Pro and Grok 4.7 score almost the same on the Noometry Index (52.3 vs 53.1), so choose on price, context window or the category you care about most.

Last verified . 4 shared benchmarks.

GPT-5.2 Pro OpenAI

52.3

Rank #39 Reported

Grok 4.7 xAI

53.1

Rank #37 Confirmed

Summary

  • They share 4 benchmarks with published results for both. GPT-5.2 Pro scores higher in 2 categories and Grok 4.7 in 0 categories; 2 gaps are clear of the uncertainty.
  • The widest gap is in math, where GPT-5.2 Pro leads 65.3 to 57.8.
  • The biggest single-benchmark swing is FrontierMath Tier 4: 46% for GPT-5.2 Pro and 17.1% for Grok 4.7.
  • Grok 4.7 is cheaper at $2 / $6 per million input/output tokens, against $21 / $168 for GPT-5.2 Pro.
  • Grok 4.7 accepts more context: 500K tokens versus 400K.

Side by side

GPT-5.2 Pro and Grok 4.7 specifications
GPT-5.2 ProGrok 4.7
ProviderOpenAIxAI
Noometry Index52.353.1
Released2025-12-112026-09-21
WeightsProprietaryProprietary
Context window400K500K
Max output128K500K
Input $ / M tokens$21$2
Output $ / M tokens$168$6
Results tracked839

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

GPT-5.2 Pro: —, Grok 4.7: 58.0 (#18)

Coding benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
FrontierCode—47.6%
CursorBench—46.3%
LMArena WebDev—1639
FrontierSWE—29.5%
SciCode—57.8%
LMArena Coding—1427

Agentic & Tool Use Not comparable

GPT-5.2 Pro: —, Grok 4.7: 36.7 (#37)

Agentic & Tool Use benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
APEX-Agents—54.6%
GDP.pdf—22.8%
Vending-Bench 2—10,537

Reasoning GPT-5.2 Pro leads

GPT-5.2 Pro: 51.5 (#33), Grok 4.7: 49.1 (#40)

Reasoning benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
NYT Connections (extended)79.3%76.8%
Epoch Capabilities Index155.4153.53
ARC-AGI-254.2%—
SimpleBench57.4%—
ARC-AGI-190.5%—
CritPt—18%
Chess Puzzles—38%
LMArena Hard Prompts—1413
Mystery Game Puzzles—29%
DTBench—96%
LMCA—49.4%

Math GPT-5.2 Pro leads

GPT-5.2 Pro: 65.3 (#29), Grok 4.7: 57.8 (#39)

Math benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
FrontierMath (Tiers 1-3)74%53%
FrontierMath Tier 446%17.1%
OTIS Mock AIME 2024-2025—98.1%
ProofBench—34%
LMArena Math—1407
FrontierMath Tier 4 (v1)31.3%—

Knowledge Not comparable

GPT-5.2 Pro: —, Grok 4.7: 62.8 (#22)

Knowledge benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
GPQA Diamond—92.7%
SimpleQA Verified—56%
LMArena Expert—1422

Multimodal Not comparable

GPT-5.2 Pro: —, Grok 4.7: 35.5 (#87)

Multimodal benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
LMArena Vision—1228
Blueprint-Bench 2—32.5%
Furniture Assembly—20.8%

Multilingual Not comparable

GPT-5.2 Pro: —, Grok 4.7: 50.8 (#116)

Multilingual benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
LMArena Non-English—1389
LMArena Chinese—1455
LMArena French—1455
LMArena Russian—1397
LMArena Spanish—1400

Instruction Following Not comparable

GPT-5.2 Pro: —, Grok 4.7: 74.1 (#105)

Instruction Following benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
LMArena Instruction Following—1404

Long Context Not comparable

GPT-5.2 Pro: —, Grok 4.7: 43.1 (#104)

Long Context benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
LMArena Longer Query—1413

Writing & Preference Not comparable

GPT-5.2 Pro: —, Grok 4.7: 70.0 (#24)

Writing & Preference benchmarks
BenchmarkGPT-5.2 ProGrok 4.7
LMArena Text—1399
LMArena Creative Writing—1391
EQ-Bench Creative Writing—2007
LMArena Multi-Turn—1393

Frequently asked questions

Is GPT-5.2 Pro better than Grok 4.7?

GPT-5.2 Pro and Grok 4.7 score almost the same on the Noometry Index (52.3 vs 53.1), so choose on price, context window or the category you care about most.

Which is cheaper, GPT-5.2 Pro or Grok 4.7?

Grok 4.7 is cheaper. It lists at $2 per million input tokens and $6 per million output tokens; GPT-5.2 Pro lists at $21 and $168.

Which has the bigger context window?

Grok 4.7 does, with 500K tokens against 400K.

How many benchmarks do GPT-5.2 Pro and Grok 4.7 share?

4 benchmarks have published results for both models. GPT-5.2 Pro has 8 scored results on Noometry and Grok 4.7 has 39.

Related comparisons

Go deeper