Model comparison

Amazon Nova Micro vs Kimi K3

Kimi K3 is the stronger model overall, scoring 59.5 to 30.4 on the Noometry Index. Amazon Nova Micro costs 98× less per token, which makes it the better buy when Kimi K3's lead doesn't matter for your workload.

Last verified . 17 shared benchmarks.

Amazon Nova Micro Amazon

30.4

Rank #294 Confirmed

Kimi K3 Moonshot AI

59.5

Rank #15 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Amazon Nova Micro scores higher in 0 categories and Kimi K3 in 9 categories; 9 gaps are clear of the uncertainty.
  • The widest gap is in math, where Kimi K3 leads 74.2 to 26.9.
  • Amazon Nova Micro is cheaper at $0.035 / $0.14 per million input/output tokens, against $3 / $15 for Kimi K3.
  • Kimi K3 accepts more context: 1.05M tokens versus 128K.
  • Kimi K3 has downloadable open weights; the other is API-only.

Side by side

Amazon Nova Micro and Kimi K3 specifications
Amazon Nova MicroKimi K3
ProviderAmazonMoonshot AI
Noometry Index30.459.5
Released2024-12-032026-07-16
WeightsProprietaryOpen
Context window128K1.05M
Max output10K1.05M
Input $ / M tokens$0.035$3
Output $ / M tokens$0.14$15
Results tracked3253

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Kimi K3 leads

Amazon Nova Micro: 30.5 (#295), Kimi K3: 61.0 (#10)

Coding benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Coding12181508
DeepSWE—68.5%
FrontierCode—44.2%
LMArena WebDev—1654
FrontierSWE—25.9%
SciCode—59.5%
WeirdML—82.6%
LiveBench Coding20.2%—
ALE-Bench—1,524

Agentic & Tool Use Kimi K3 leads

Amazon Nova Micro: 22.1 (#132), Kimi K3: 41.8 (#20)

Agentic & Tool Use benchmarks
BenchmarkAmazon Nova MicroKimi K3
APEX-Agents—50.6%
Berkeley Function Calling Leaderboard22.3%—
τ²-bench Banking—37.1%
PostTrainBench—32%
GBAEval—48.3%
GDP.pdf—19%
Vending-Bench 2—5,165

Reasoning Kimi K3 leads

Amazon Nova Micro: 17.4 (#294), Kimi K3: 63.0 (#17)

Reasoning benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Hard Prompts11911496
ARC-AGI-2—60.4%
SimpleBench—60.7%
NYT Connections (extended)—93.6%
ARC-AGI-1—94.5%
CritPt—23.4%
Chess Puzzles—39%
LiveBench Reasoning25.1%—
Mystery Game Puzzles—26%
DTBench—91.2%
LiveBench Data Analysis34%—
LMCA—52.7%
Surface Evolver Bench—95%
Epoch Capabilities Index—157.45
ForecastBench—61.1
LiveBench29.6%—

Math Kimi K3 leads

Amazon Nova Micro: 26.9 (#254), Kimi K3: 74.2 (#16)

Math benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Math12061491
FrontierMath (Tiers 1-3)—72.2%
FrontierMath Tier 4—39%
MathArena Final-Answer Competitions—87.8%
OTIS Mock AIME 2024-2025—97.2%
ProofBench—87%
Omni-MATH21.4%—
LiveBench Math34.5%—

Knowledge Kimi K3 leads

Amazon Nova Micro: 29.6 (#237), Kimi K3: 63.2 (#21)

Knowledge benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Expert11841521
GPQA Diamond—93.1%
SimpleQA Verified—50.6%
MMLU-Pro51.1%—
Vectara Hallucination Rate5.5%—
GPQA (HELM)38.3%—
MMLU70.8%—

Multimodal Not comparable

Amazon Nova Micro: —, Kimi K3: 37.8 (#70)

Multimodal benchmarks
BenchmarkAmazon Nova MicroKimi K3
Blueprint-Bench 2—29.5%
Furniture Assembly—34.2%

Multilingual Kimi K3 leads

Amazon Nova Micro: 36.5 (#239), Kimi K3: 56.3 (#21)

Multilingual benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Non-English11861466
LMArena Chinese12091529
LMArena French12381491
LMArena German11921488
LMArena Japanese11541487
LMArena Korean11501458
LMArena Russian11851482
LMArena Spanish12251472

Instruction Following Kimi K3 leads

Amazon Nova Micro: 56.3 (#272), Kimi K3: 77.7 (#14)

Instruction Following benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Instruction Following11741483
LiveBench Instruction Following48%—
IFEval76%—

Long Context Kimi K3 leads

Amazon Nova Micro: 36.5 (#229), Kimi K3: 45.8 (#29)

Long Context benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Longer Query12051494

Writing & Preference Kimi K3 leads

Amazon Nova Micro: 39.5 (#247), Kimi K3: 76.6 (#4)

Writing & Preference benchmarks
BenchmarkAmazon Nova MicroKimi K3
LMArena Text12081476
LMArena Creative Writing11721454
LMArena Multi-Turn11781488
EQ-Bench Creative Writing—2082
WildBench74.3%—
EQ-Bench 4—1339
LiveBench Language15.8%—

Frequently asked questions

Is Amazon Nova Micro better than Kimi K3?

Kimi K3 is the stronger model overall, scoring 59.5 to 30.4 on the Noometry Index. Amazon Nova Micro costs 98× less per token, which makes it the better buy when Kimi K3's lead doesn't matter for your workload.

Which is cheaper, Amazon Nova Micro or Kimi K3?

Amazon Nova Micro is cheaper. It lists at $0.035 per million input tokens and $0.14 per million output tokens; Kimi K3 lists at $3 and $15.

Is Amazon Nova Micro or Kimi K3 better for coding?

Kimi K3 scores higher on coding benchmarks: 61.0 versus 30.5 in the Noometry coding category.

Which has the bigger context window?

Kimi K3 does, with 1.05M tokens against 128K.

How many benchmarks do Amazon Nova Micro and Kimi K3 share?

17 benchmarks have published results for both models. Amazon Nova Micro has 32 scored results on Noometry and Kimi K3 has 53.

Related comparisons

Go deeper