Best model for your job

Best AI model for data extraction

For data extraction, GPT-5 scores highest on the current data (67.1), weighting instruction following 50%, long context 30%, multimodal 20%.

Last verified

Weights: Instruction Following 50%, Long Context 30%, Multimodal 20%. Extraction rewards following a schema exactly across long or scanned documents.

Best AI model for data extraction
#ModelProviderFitInstruction FollowingLong ContextMultimodalInput $/MOutput $/M
1GPT-5OpenAI67.173.869.546.8$1.25$10
2Claude Opus 5.5Anthropic65.780.047.157.8$4$20
3Grok 4xAI65.379.263.133.7——
4GPT-5.1OpenAI65.283.947.644.8$1.25$10
5Gemini 2.5 ProGoogle64.575.059.845.2$1.25$10
6Claude Fable 5.1Anthropic64.479.246.753.9$10$50
7Claude Opus 5Anthropic63.779.246.550.8$5$25
8Gemini 4 ArgonGoogle63.680.147.646.5——
9Claude Sonnet 5.5Anthropic63.278.345.951.5$2$10
10Gemini 3 ProGoogle62.876.344.057.6——
11GPT-5.5OpenAI62.677.548.346.9$5$30
12GPT-6.1 SolOpenAI62.577.044.952.7$2$10
13GPT-6 AstraOpenAI62.576.344.555.0$10$50
14GPT-5.4OpenAI62.477.150.343.7$2.50$15
15Claude Fable 5Anthropic62.278.646.345.3$10$50
16GPT-5.6 SolOpenAI62.277.745.448.6$4$20
17Claude Opus 4.6Anthropic61.679.548.137.3$5$25
18Kimi K2.5 (open weights)Moonshot AI61.575.352.141.1$0.45$2.25
19Claude Opus 4.7Anthropic61.378.446.241.2$5$25
20Gemini 3.5 FlashGoogle61.377.045.445.7$1.50$9
21Muse Spark 1.3Meta61.277.545.643.7$1.25$4.25
22MiMo-V2.6-Pro (open weights)Xiaomi61.178.246.040.8$0.43$0.87
23Gemini 3.8 FlashGoogle61.078.046.340.7$0.75$3.75
24GPT-5.6 TerraOpenAI61.076.444.447.3$2$12
25GLM-5.3-Flash (open weights)Z.ai (Zhipu)60.977.545.442.8$0.15$0.50

Sponsored placements are available on pages like this one. Advertise on Noometry

Top two head to head: GPT-5 vs Claude Opus 5.5

Frequently asked questions

What is the best ai model for data extraction?

For data extraction, GPT-5 scores highest on the current data (67.1), weighting instruction following 50%, long context 30%, multimodal 20%.

How is this shortlist built?

Extraction rewards following a schema exactly across long or scanned documents. Each model's category scores are blended with those weights; only ranked models with results in every needed category are listed.

Other jobs