xAI, proprietary

Grok 4.7

Grok 4.7 by xAI ranks 37th of 354 ranked models on the Noometry Index as of October 2026, with a score of 53.1. Its strongest category is coding, where it ranks 18th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 500K-token context window.

Last verified

Specifications

Noometry rank
#37 of 354
Index score
53.1
Evidence
Confirmed 39 results
Provider
xAI
Released
September 21, 2026
Weights
Proprietary
Reasoning
Yes
Context window
500K
Max output
500K
Input price
$2 / M
Output price
$6 / M
Blended price
$3 / M
Output speed
Not measured
Value
#157 of 219
Knowledge cutoff
May 2026
Input
text, image, pdf

Category scores

Each category score combines every public result we have in that category.

Grok 4.7 category scores
  1. Coding 58.0
  2. Agentic & Tool Use 36.7
  3. Reasoning 49.1
  4. Math 57.8
  5. Knowledge 62.8
  6. Multimodal 35.5
  7. Multilingual 50.8
  8. Instruction Following 74.1
  9. Long Context 43.1
  10. Writing & Preference 70.0
Grok 4.7 category ranks
CategoryScoreRankResults
Coding58.0#186
Agentic & Tool Use36.7#372
Reasoning49.1#407
Math57.8#395
Knowledge62.8#223
Multimodal35.5#873
Multilingual50.8#1161
Instruction Following74.1#1051
Long Context43.1#1041
Writing & Preference70.0#244

Strengths and weaknesses

Categories where Grok 4.7 places highest and lowest among the models ranked in each, with its score against that category's median.

Strongest categories

Grok 4.7: strongest categories
CategoryScorevs medianRank
Coding58.0+19.3#18 of 340, top 6%
Knowledge62.8+25.4#22 of 314, top 8%
Writing & Preference70.0+16.2#24 of 312, top 8%

Weakest categories

Grok 4.7: weakest categories
CategoryScorevs medianRank
Multimodal35.5−3.0#87 of 128, top 68%
Multilingual50.8+3.3#116 of 297, top 40%
Long Context43.1+2.2#104 of 296, top 36%

Closest competitors

The models ranked just above and below Grok 4.7. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Grok 4.7
ModelRankScoreBlended $/MSpeed
Gemini 3.6 Flash#3354.1$1.50—Compare
GPT-5.2#3454.1$4.8115Compare
DeepSeek V4 Flash#3553.6$0.266Compare
GPT-6 Luna#3653.3$0.20—Compare
DeepSeek V4.1 Flash#3852.8$0.26—Compare
GPT-5.2 Pro#3952.3$57.75—Compare
Gemini 3 Flash Preview#4052.3$1.13—Compare
GLM-5.3-Flash#4151.8$0.24—Compare

Sponsored placements are available on pages like this one. Advertise on Noometry

Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

Coding

Grok 4.7 Coding benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierCode47.6%#10 of 37, top 28%Epoch AI
CursorBench43.9%highEpoch AI
CursorBench33.1%lowEpoch AI
CursorBench41.6%mediumEpoch AI
CursorBench46.3%#5 of 14, top 36%xhighEpoch AI
LMArena WebDev1639#12 of 113, top 11%xhighLMArena2026-10-08
FrontierSWE29.5%#10 of 18, top 56%xhighEpoch AI
SciCode57.8%#14 of 121, top 12%highEpoch AI
SciCode54.9%lowEpoch AI
SciCode57.4%xhighEpoch AI
LMArena Coding1427#113 of 294, top 39%xhighLMArena2026-10-08

Agentic & Tool Use

Grok 4.7 Agentic & Tool Use benchmark results
BenchmarkScorePositionSettingSourceDate
APEX-Agents54.6%#19 of 49, top 39%Epoch AI
GDP.pdf22.8%#17 of 36, top 48%xhighEpoch AI
Vending-Bench 210,537#6 of 60, top 10%Epoch AI

Reasoning

Grok 4.7 Reasoning benchmark results
BenchmarkScorePositionSettingSourceDate
NYT Connections (extended)76.8%#41 of 91, top 46%high reasoningLech Mazur benchmarks
CritPt18%#30 of 134, top 23%highEpoch AI
CritPt13.4%lowEpoch AI
CritPt17.7%xhighEpoch AI
Chess Puzzles38%#24 of 129, top 19%xhighEpoch AI2026-09-22
LMArena Hard Prompts1413#114 of 297, top 39%xhighLMArena2026-10-08
Mystery Game Puzzles29%#28 of 74, top 38%xhighEpoch AI2026-09-22
DTBench96%#15 of 151, top 10%xhighEpoch AI
LMCA49.4%#23 of 125, top 19%xhighEpoch AI
Epoch Capabilities Index153.53#40 of 213, top 19%Epoch AI2026-09-21

Math

Grok 4.7 Math benchmark results
BenchmarkScorePositionSettingSourceDate
FrontierMath (Tiers 1-3)53%#45 of 81, top 56%xhighEpoch AI2026-09-22
FrontierMath Tier 417.1%#47 of 63, top 75%xhighEpoch AI2026-09-22
OTIS Mock AIME 2024-202598.1%#23 of 173, top 14%xhighEpoch AI2026-09-22
ProofBench34%#40 of 77, top 52%Epoch AI
LMArena Math1407#117 of 285, top 42%xhighLMArena2026-10-08

Knowledge

Grok 4.7 Knowledge benchmark results
BenchmarkScorePositionSettingSourceDate
GPQA Diamond92.7%#20 of 186, top 11%xhighEpoch AI2026-09-22
SimpleQA Verified56%#18 of 77, top 24%xhighEpoch AI2026-09-22
LMArena Expert1422#104 of 273, top 39%xhighLMArena2026-10-08

Multimodal

Grok 4.7 Multimodal benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Vision1228#68 of 122, top 56%xhighLMArena2026-10-09
Blueprint-Bench 232.5%#12 of 31, top 39%Epoch AI
Furniture Assembly20.8%#30 of 31, top 97%xhighEpoch AI2026-09-24

Multilingual

Grok 4.7 Multilingual benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Non-English1389#116 of 297, top 40%xhighLMArena2026-10-08
LMArena Chinese1455#100 of 285, top 36%xhighLMArena2026-10-08
LMArena French1455#62 of 223, top 28%xhighLMArena2026-10-08
LMArena Russian1397#107 of 283, top 38%xhighLMArena2026-10-08
LMArena Spanish1400#111 of 226, top 50%xhighLMArena2026-10-08

Instruction Following

Grok 4.7 Instruction Following benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Instruction Following1404#97 of 298, top 33%xhighLMArena2026-10-08

Long Context

Grok 4.7 Long Context benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Longer Query1413#101 of 291, top 35%xhighLMArena2026-10-08

Writing & Preference

Grok 4.7 Writing & Preference benchmark results
BenchmarkScorePositionSettingSourceDate
LMArena Text1399#119 of 297, top 41%xhighLMArena2026-10-08
LMArena Creative Writing1391#93 of 295, top 32%xhighLMArena2026-10-08
EQ-Bench Creative Writing2007#8 of 115, top 7%EQ-Bench
LMArena Multi-Turn1393#126 of 295, top 43%xhighLMArena2026-10-08

API pricing by provider

Grok 4.7 API prices
RouteInput $/MOutput $/MCached input $/MChecked
bedrock$2$6$0.502026-10-10
openrouter$2$6$0.502026-10-10
vertex$2$6$0.502026-10-10
xai$2$6$0.502026-10-10

Compare Grok 4.7

Other xAI models

Frequently asked questions

How good is Grok 4.7?

Grok 4.7 by xAI ranks 37th of 354 ranked models on the Noometry Index as of October 2026, with a score of 53.1. Its strongest category is coding, where it ranks 18th. API pricing starts at $2 per million input tokens and $6 per million output tokens, with a 500K-token context window.

How much does Grok 4.7 cost?

Grok 4.7 costs $2 per million input tokens and $6 per million output tokens on xAI's own API, with cached input at $0.50.

What is Grok 4.7's context window?

Grok 4.7 accepts up to 500K tokens of input and can write up to 500K tokens in one response.

Is Grok 4.7 open source?

No. Grok 4.7 is proprietary and available only through xAI's API and partner platforms.

What are Grok 4.7's strengths and weaknesses?

Relative to other ranked models, Grok 4.7 places best in coding, knowledge, writing & preference and lowest in multimodal, multilingual, long context.

What is Grok 4.7 best at?

Its best category is coding, where it ranks 18th on Noometry.