Value ranking

Best Value LLMs for Math in October 2026

As of October 2026, gpt-oss-20b gives the most math performance per dollar: 39.4 points at a blended $0.04 per million tokens. gpt-oss-120b is next.

Last verified

Category score per dollar of blended API price, among models in the top half of the math ranking.

Best Value LLMs for Math: top 10
  1. gpt-oss-20b 1095.0 pts/$
  2. gpt-oss-120b 747.2 pts/$
  3. Qwen3.7 Flash 696.0 pts/$
  4. Gemma 4 26B A4B IT 445.0 pts/$
  5. Nemotron 3 Nano 30B A3B 431.4 pts/$
  6. Nemotron 3.5 Lightning 428.3 pts/$
  7. GPT-6 Luna 380.5 pts/$
  8. Claude Haiku 5.5 367.9 pts/$
  9. MiMo-V2.6-Flash 296.7 pts/$
  10. Step 3.5 Flash 284.2 pts/$

Full ranking

Best Value LLMs for Math
#ModelScoreInput $/MOutput $/MPoints per $
1gpt-oss-20b39.4$0.018$0.091095.0
2gpt-oss-120b52.5$0.037$0.17747.2
3Qwen3.7 Flash38.3$0.03$0.13696.0
4Gemma 4 26B A4B IT47.6$0.0675$0.23445.0
5Nemotron 3 Nano 30B A3B37.8$0.05$0.20431.4
6Nemotron 3.5 Lightning37.5$0.05$0.20428.3
7GPT-6 Luna76.1$0.10$0.50380.5
8Claude Haiku 5.573.6$0.10$0.50367.9
9MiMo-V2.6-Flash51.9$0.14$0.28296.7
10Step 3.5 Flash42.6$0.10$0.30284.2
11Gemma 4 31B IT43.2$0.09$0.34283.1
12Hy340.1$0.0825$0.33278.0
13DeepSeek V4.1 Flash66.7$0.15$0.60254.2
14DeepSeek V4 Flash60.3$0.15$0.60229.6
15Nemotron 3 Super39.6$0.08$0.45229.4
16GLM-5.3-Flash53.3$0.15$0.50224.2
17MiMo-V2-Omni39.1$0.14$0.28223.4
18MiMo-V2-Flash38.3$0.14$0.28219.0
19Gemini 2.5 Flash-Lite38.0$0.10$0.40216.9
20Qwen3.5-Flash37.4$0.10$0.40213.9
21MiMo-V2.536.8$0.14$0.28210.0
22Qwen3-30B-A3B37.4$0.12$0.50174.1
23GPT-5.6 Luna77.7$0.20$1.20172.6
24DeepSeek-V3.2-Exp41.7$0.26$0.38143.8
25Mistral Large 338.7$0.25$0.75103.3
26Step 3.7 Flash42.9$0.18$1.11103.1
27MiMo-V2.6-Pro54.5$0.43$0.87100.2
28Trinity Large Thinking37.6$0.25$0.8097.2
29Nvidia Llama 3.3 Nemotron Super 49b v1.538.2$0.40$0.4095.6
30Qwen3.6 Flash39.0$0.19$1.1392.5
31DeepSeek-V3.138.9$0.25$0.9591.4
32GPT-5.4 nano40.9$0.20$1.2588.4
33DeepSeek-V3.1-Terminus38.5$0.27$185.0
34MiniMax-M340.0$0.30$1.2076.3
35Solar Pro438.8$0.30$1.2073.8
36MiMo-V2.5-Pro40.0$0.43$0.8773.6
37MiniMax-M2.138.3$0.30$1.2073.0
38MiMo-V2-Pro39.5$0.43$0.8772.6
39Gemini 3.1 Flash Lite40.7$0.25$1.5072.3
40Qwen3.7 Plus50.5$0.40$1.6072.1
41MiniMax-M237.3$0.30$1.2071.1
42Inkling-Small45.1$0.45$1.2070.7
43Qwen3.6 35B-A3B38.9$0.25$1.4969.8
44GPT-5 Mini46.7$0.25$267.9
45DeepSeek V4 Pro64.8$0.66$1.9865.5
46Qwen3 14B38.6$0.35$1.4063.1
47Qwen3.5 35B-A3B39.9$0.25$258.0
48Kimi K2.551.8$0.45$2.2557.6
49Qwen3.5 Plus49.6$0.40$2.4055.1
50Hy4 preview55.7$0.75$2.2549.5
51DeepSeek-R143.8$0.50$2.1548.0
52Qwen3.5 27B38.8$0.30$2.4047.0
53Gemini 2.5 Flash39.9$0.30$2.5046.9
54Gemini 3.7 Flash69.6$0.75$3.7546.4
55Qwen3.6 Plus51.8$0.50$346.0
56Gemini 3 Flash Preview51.7$0.50$346.0
57Qwen3-Next 80B-A3B Instruct38.8$0.50$244.3
58Nova 2 Lite37.5$0.30$2.5044.1
59Gemini 3.8 Flash65.3$0.75$3.7543.6
60Kimi K2 (Jul 2025)42.7$0.57$2.3042.6
61GLM-4.5V37.4$0.60$1.8041.5
62Qwen3 235B-A22B50.4$0.70$2.8041.1
63Mistral Large 440.4$0.68$2.0939.2
64GLM-4.639.1$0.60$2.2039.1
65GLM-4.539.0$0.60$2.2039.0
66MiniMax M137.5$0.55$2.2039.0
67GLM-4.738.6$0.60$2.2038.6
68Gemini 3.6 Flash57.3$0.75$3.7538.2
69Muse Spark 1.373.1$1.25$4.2536.6
70Qwen3.6 27B48.5$0.60$3.6035.9
71Qwen3.5 122B-A10B39.1$0.40$3.2035.6
72Seed 2.0 Pro39.3$0.50$334.9
73Qwen3.5 397B-A17B46.1$0.60$3.6034.1
74Kimi K2.657.0$0.95$433.3
75Qwen3.8 27B37.1$0.99$1.4933.2
76Qwen3 32B39.7$0.70$2.8032.4
77Step 5 Preview46.1$1$2.7032.4
78Qwen3-VL 235B-A22B39.0$0.70$2.8031.8
79Kimi K2.7 Code52.9$0.95$430.9
80Grok 4.20 (Non-Reasoning)48.2$1.25$2.5030.8
81GLM-546.4$1$3.2029.9
82Grok 4.346.0$1.25$2.5029.4
83GLM-5.362.3$1.40$4.4029.0
84GPT-5.4 mini45.5$0.75$4.5027.0
85GLM-5.255.7$1.40$4.4025.9
86Grok 4.20 Multi-Agent39.4$1.25$2.5025.2
87Qwen3.8 Max73.2$2$624.4
88GPT-6.1 Sol93.7$2$1023.4
89Muse Spark 1.246.4$1.25$4.2523.2
90GLM-5.149.7$1.40$4.4023.1
91Muse Spark 1.145.5$1.25$4.2522.8
92Claude Haiku 4.544.9$1$522.4
93Grok 4.667.0$2$622.3
94Claude Sonnet 5.587.9$2$1022.0
95GPT-6 Sol87.2$2$1021.8
96o4-mini40.8$1.10$4.4021.2
97GLM-5V-Turbo39.4$1.20$420.7
98Grok 4.560.9$2$620.3
99Grok 4.757.8$2$619.3
100Qwen3.6 Max Preview54.1$1.30$7.8018.5
101GPT-5.6 Terra81.6$2$1218.1
102Gemini 3.5 Flash60.7$1.50$918.0
103Qwen3.7 Max62.4$2.50$7.5016.7
104Claude Sonnet 566.2$2$1016.6
105Qwen3 Max38.7$1.20$616.1
106GPT-555.0$1.25$1016.0
107GPT-5.152.2$1.25$1015.2
108o350.2$2$814.3
109Gemini 3.1 Pro Preview62.1$2$1213.8
110GPT-5.473.5$2.50$1513.1
111Mistral Medium 3.539.1$1.50$7.5013.0
112Qwen3-Coder 480B-A35B Instruct37.6$1.50$7.5012.5
113GPT-5.260.0$1.75$1412.5
114Kimi K374.2$3$1512.4
115Claude Opus 5.591.8$4$2011.5
116GPT-5.6 Sol85.6$4$2010.7
117Claude Sonnet 4.652.9$3$158.8
118Claude Opus 586.2$5$258.6
119GPT-5.3 Chat38.2$1.75$147.9
120Claude Opus 4.878.4$5$257.8
121GPT-5.581.7$5$307.3
122Claude Sonnet 443.3$3$157.2
123Claude Opus 4.766.7$5$256.7
124Claude Opus 4.663.0$5$256.3
125GPT-6 Astra93.5$10$504.7
126Claude Fable 5.189.6$10$504.5
127Claude Fable 588.5$10$504.4
128Claude Opus 4.538.6$5$253.9
129Claude Opus 442.0$15$751.4
130GPT-5.5 Pro84.0$30$1801.2
131GPT-5 Pro48.5$15$1201.2
132GPT-5.2 Pro65.3$21$1681.1
133GPT-5.4 Pro72.4$30$1801.1

Sponsored placements are available on pages like this one. Advertise on Noometry

Compare the leaders

Related rankings

Frequently asked questions

What is the most cost-effective AI model for math?

As of October 2026, gpt-oss-20b gives the most math performance per dollar: 39.4 points at a blended $0.04 per million tokens. gpt-oss-120b is next.

What are the top 5 in this ranking?

gpt-oss-20b, gpt-oss-120b, Qwen3.7 Flash, Gemma 4 26B A4B IT, Nemotron 3 Nano 30B A3B, in that order, as of October 2026.

How is this list ranked?

By score per dollar: the model’s score divided by its blended API price (three parts input to one part output), among models in the top half by score.