Reasoning benchmark
Kagi LLM Benchmark leaderboard
As of October 2026, Claude Fable 5 has the highest published Kagi LLM Benchmark score on Noometry at 91.4%, out of 99 models with results.
Last verified
About Kagi LLM Benchmark
A private, frequently rotated set of reasoning, coding and instruction-following tasks from Kagi, designed so models cannot be trained on it.
- Category
- Reasoning
- Introduced
- 2024
- Format
- Mixed private tasks
- Unit
- Percent (random guessing ≈ 0%)
- Official site
- help.kagi.com
Top 15 models
- Claude Fable 5 91.4%
- Claude Opus 4.8 88.8%
- GPT-5.5 88.8%
- Claude Opus 4.6 83.6%
- Grok 4.5 83.5%
- Claude Opus 4.7 80.7%
- Claude Opus 4.5 80.2%
- Gemini 3 Pro 80.1%
- Kimi K2.5 78.5%
- GPT-5 Pro 76.8%
- Chatgpt 4o Latest 20250326 75%
- GLM-5 75%
- Grok 4.20 (Non-Reasoning) 75%
- Claude Opus 4 74.3%
- Qwen3.5 397B-A17B 73.7%
Sponsored placements are available on pages like this one. Advertise on Noometry
All results
Compare the leaders
Frequently asked questions
What does Kagi LLM Benchmark measure?
A private, frequently rotated set of reasoning, coding and instruction-following tasks from Kagi, designed so models cannot be trained on it.
Which model has the highest Kagi LLM Benchmark score?
As of October 2026, Claude Fable 5 has the highest published Kagi LLM Benchmark score on Noometry at 91.4%, out of 99 models with results.
What is the best open-weight model on Kagi LLM Benchmark?
Kimi K2.5 has the highest Kagi LLM Benchmark accuracy among open-weight models at 78.5%, ranking 9 of 99 overall.