xAI, proprietary
Grok 4.3
Grok 4.3 by xAI ranks 86th of 354 ranked models on the Noometry Index as of October 2026, with a score of 43.8. Its strongest category is knowledge, where it ranks 62nd. API pricing starts at $1.25 per million input tokens and $2.50 per million output tokens, with a 1M-token context window.
Last verified
Specifications
- Noometry rank
- #86 of 354
- Index score
- 43.8
- Evidence
- Confirmed 40 results
- Provider
- xAI
- Released
- April 17, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 1M
- Max output
- 30K
- Input price
- $1.25 / M
- Output price
- $2.50 / M
- Blended price
- $1.56 / M
- Output speed
- Not measured
- Value
- #137 of 219
- Knowledge cutoff
- Unknown
- Input
- text, image, pdf
Category scores
Each category score combines every public result we have in that category.
- Coding 41.6
- Agentic & Tool Use 27.7
- Reasoning 35.9
- Math 46.0
- Knowledge 52.5
- Multimodal 31.6
- Multilingual 50.5
- Instruction Following 72.1
- Long Context 42.5
- Writing & Preference 58.5
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 41.6 | #121 | 4 |
| Agentic & Tool Use | 27.7 | #99 | 1 |
| Reasoning | 35.9 | #68 | 6 |
| Math | 46.0 | #74 | 5 |
| Knowledge | 52.5 | #62 | 3 |
| Multimodal | 31.6 | #104 | 2 |
| Multilingual | 50.5 | #120 | 1 |
| Instruction Following | 72.1 | #140 | 1 |
| Long Context | 42.5 | #123 | 1 |
| Writing & Preference | 58.5 | #118 | 4 |
Strengths and weaknesses
Categories where Grok 4.3 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multimodal | 31.6 | −6.9 | #104 of 128, top 82% |
| Agentic & Tool Use | 27.7 | −2.7 | #99 of 154, top 65% |
| Instruction Following | 72.1 | +0.9 | #140 of 305, top 46% |
Closest competitors
The models ranked just above and below Grok 4.3. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Chatgpt 4o Latest 20250326 | #82 | 43.8 | — | 21 | Compare |
| ERNIE 5.1 | #83 | 43.8 | — | — | Compare |
| GLM-5V-Turbo | #84 | 43.8 | $1.90 | — | Compare |
| MiniMax-M3 | #85 | 43.8 | $0.52 | — | Compare |
| Qwen3 Max | #87 | 43.7 | $2.40 | 48 | Compare |
| MiMo-V2-Omni | #88 | 43.6 | $0.18 | — | Compare |
| Kimi K2.5 Instant | #89 | 43.6 | — | — | Compare |
| Gemma 4 31B IT | #90 | 43.5 | $0.15 | 3 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena WebDev | 1357 | #87 of 113, top 77% | LMArena | 2026-10-08 | |
| SciCode | 47.3% | #51 of 121, top 43% | high | Epoch AI | |
| WeirdML | 49.9% | #48 of 119, top 41% | Epoch AI | ||
| LMArena Coding | 1415 | #122 of 294, top 42% | LMArena | 2026-10-08 | |
| ALE-Bench | 944.17 | #45 of 105, top 43% | Epoch AI |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GDP.pdf | 8% | #35 of 36, top 98% | Epoch AI | ||
| GDP.pdf | 8% | high | Epoch AI | ||
| LMArena Search | 1165 | #22 of 32, top 69% | LMArena | 2026-08-24 | |
| Vending-Bench 2 | 35.26 | #55 of 60, top 92% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| NYT Connections (extended) | 55.2% | #59 of 91, top 65% | Lech Mazur benchmarks | ||
| CritPt | 0% | Epoch AI | |||
| CritPt | 8% | #53 of 134, top 40% | high | Epoch AI | |
| CritPt | 0% | none | Epoch AI | ||
| Chess Puzzles | 25% | #44 of 129, top 35% | high | Epoch AI | 2026-06-17 |
| LMArena Hard Prompts | 1396 | #131 of 297, top 45% | LMArena | 2026-10-08 | |
| DTBench | 90.7% | #36 of 151, top 24% | high | Epoch AI | |
| LMCA | 38.3% | #51 of 125, top 41% | high | Epoch AI | |
| Epoch Capabilities Index | 149.16 | #59 of 213, top 28% | Epoch AI | 2026-04-17 | |
| ForecastBench | 60.3 | #28 of 72, top 39% | Epoch AI |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 42.8% | #52 of 81, top 65% | high | Epoch AI | 2026-06-17 |
| FrontierMath Tier 4 | 14.6% | #49 of 63, top 78% | high | Epoch AI | 2026-06-17 |
| OTIS Mock AIME 2024-2025 | 93.3% | #42 of 173, top 25% | high | Epoch AI | 2026-06-17 |
| ProofBench | 11% | #62 of 77, top 81% | high | Epoch AI | |
| LMArena Math | 1388 | #141 of 285, top 50% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| GPQA Diamond | 88.8% | #45 of 186, top 25% | high | Epoch AI | 2026-06-17 |
| SimpleQA Verified | 33.2% | #54 of 77, top 71% | high | Epoch AI | 2026-08-27 |
| LMArena Expert | 1385 | #135 of 273, top 50% | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1229 | #67 of 122, top 55% | LMArena | 2026-10-09 | |
| Blueprint-Bench 2 | 0% | #31 of 31, top 100% | Epoch AI |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1385 | #120 of 297, top 41% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1422 | #126 of 285, top 45% | LMArena | 2026-10-08 | |
| LMArena French | 1412 | #109 of 223, top 49% | LMArena | 2026-10-08 | |
| LMArena German | 1395 | #97 of 231, top 42% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1379 | #78 of 211, top 37% | LMArena | 2026-10-08 | |
| LMArena Korean | 1356 | #100 of 213, top 47% | LMArena | 2026-10-08 | |
| LMArena Russian | 1399 | #106 of 283, top 38% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1398 | #113 of 226, top 50% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1366 | #135 of 298, top 46% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1393 | #123 of 291, top 43% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1397 | #121 of 297, top 41% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1380 | #104 of 295, top 36% | LMArena | 2026-10-08 | |
| EQ-Bench 4 | 1075 | #25 of 28, top 90% | EQ-Bench | ||
| LMArena Multi-Turn | 1406 | #115 of 295, top 39% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| bedrock | $1.25 | $2.50 | $0.20 | 2026-10-10 |
| openrouter | $1.25 | $2.50 | $0.20 | 2026-10-10 |
| vertex | $1.25 | $2.50 | $0.20 | 2026-10-10 |
| xai | $1.25 | $2.50 | $0.20 | 2026-10-10 |
Compare Grok 4.3
- Grok 4.3 vs Grok 4.1
- Grok 4.3 vs MiniMax-M3
- Grok 4.3 vs Qwen3 Max
- Grok 4.3 vs GLM-5V-Turbo
- Grok 4.3 vs MiMo-V2-Omni
- Grok 4.3 vs ERNIE 5.1
- Grok 4.3 vs Kimi K2.5 Instant
- Grok 4.3 vs GPT-6 Astra
- Grok 4.3 vs Claude Fable 5.1
- Grok 4.3 vs Gemini 3.8 Flash
- Grok 4.3 vs Kimi K3
- Grok 4.3 vs Qwen3.8 Max
- Grok 4.3 vs GLM-5.3
- Grok 4.3 vs Muse Spark 1.3
Other xAI models
- Grok 4.656.9
- Grok 4.555.0
- Grok 4.753.1
- Grok 4.20 (Non-Reasoning)48.6
- Grok 448.1
- Grok 4.20 Multi-Agent46.2
- Grok 4.141.5
- Grok 4.1 Fast41.4
Frequently asked questions
How good is Grok 4.3?
Grok 4.3 by xAI ranks 86th of 354 ranked models on the Noometry Index as of October 2026, with a score of 43.8. Its strongest category is knowledge, where it ranks 62nd. API pricing starts at $1.25 per million input tokens and $2.50 per million output tokens, with a 1M-token context window.
How much does Grok 4.3 cost?
Grok 4.3 costs $1.25 per million input tokens and $2.50 per million output tokens on xAI's own API, with cached input at $0.20.
What is Grok 4.3's context window?
Grok 4.3 accepts up to 1M tokens of input and can write up to 30K tokens in one response.
Is Grok 4.3 open source?
No. Grok 4.3 is proprietary and available only through xAI's API and partner platforms.
What are Grok 4.3's strengths and weaknesses?
Relative to other ranked models, Grok 4.3 places best in reasoning, knowledge, math and lowest in multimodal, agentic & tool use, instruction following.
What is Grok 4.3 best at?
Its best category is knowledge, where it ranks 62nd on Noometry.