Model comparison

Gemini 3 Flash Preview vs Qwen3.6 Max Preview

Gemini 3 Flash Preview and Qwen3.6 Max Preview score almost the same on the Noometry Index (52.3 vs 51.5), so choose on price, context window or the category you care about most.

Last verified . 29 shared benchmarks.

Gemini 3 Flash Preview Google

52.3

Rank #40 Confirmed

Qwen3.6 Max Preview Alibaba (Qwen)

51.5

Rank #43 Confirmed

Summary

  • They share 29 benchmarks with published results for both. Gemini 3 Flash Preview scores higher in 5 categories and Qwen3.6 Max Preview in 3 categories; 6 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Gemini 3 Flash Preview leads 49.2 to 41.7.
  • The biggest single-benchmark swing is Chess Puzzles: 40% for Gemini 3 Flash Preview and 20% for Qwen3.6 Max Preview.
  • Gemini 3 Flash Preview is cheaper at $0.50 / $3 per million input/output tokens, against $1.30 / $7.80 for Qwen3.6 Max Preview.
  • Gemini 3 Flash Preview accepts more context: 1.05M tokens versus 262K.

Side by side

Gemini 3 Flash Preview and Qwen3.6 Max Preview specifications
Gemini 3 Flash PreviewQwen3.6 Max Preview
ProviderGoogleAlibaba (Qwen)
Noometry Index52.351.5
Released2025-12-172026-04-20
WeightsProprietaryProprietary
Context window1.05M262K
Max output66K66K
Input $ / M tokens$0.50$1.30
Output $ / M tokens$3$7.80
Results tracked5929

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Gemini 3 Flash Preview leads

Gemini 3 Flash Preview: 50.9 (#42), Qwen3.6 Max Preview: 48.7 (#54)

Coding benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
SWE-bench Verified75.4%76.7%
LMArena WebDev14391482
LMArena Coding14601471
SWE-bench Verified (bash only)75.8%—
SWE-bench Multilingual72.7%—
GSO9.8%—
WeirdML61.6%—
ALE-Bench1,367—

Agentic & Tool Use Not comparable

Gemini 3 Flash Preview: 38.7 (#29), Qwen3.6 Max Preview: —

Agentic & Tool Use benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
Vending-Bench 23,6354,254
Terminal-Bench64.3%—
τ²-bench Airline82.5%—
τ²-bench Banking27.3%—
τ²-bench Retail76.8%—
τ²-bench Telecom91.2%—
DeepResearch Bench49.8%—
BALROG48.1%—
GDP.pdf10%—
LMArena Search1198—

Reasoning Gemini 3 Flash Preview leads

Gemini 3 Flash Preview: 49.2 (#37), Qwen3.6 Max Preview: 41.7 (#53)

Reasoning benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
SimpleBench61.1%63%
NYT Connections (extended)83.1%74.1%
Chess Puzzles40%20%
LMArena Hard Prompts14651457
Mystery Game Puzzles26%19%
DTBench89.1%87.2%
LMCA43.1%42.5%
Epoch Capabilities Index151.8149.24
ARC-AGI-233.6%—
ARC-AGI-184.7%—
ForecastBench58.5—

Math Qwen3.6 Max Preview leads

Gemini 3 Flash Preview: 51.7 (#55), Qwen3.6 Max Preview: 54.1 (#46)

Math benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
OTIS Mock AIME 2024-202595.6%91.1%
LMArena Math14731465
FrontierMath (Feb 2025 set)35.6%23.1%
FrontierMath Tier 4 (v1)4.2%4.2%
FrontierMath (Tiers 1-3)51.2%—
FrontierMath Tier 417.1%—
MathArena Final-Answer Competitions67.6%—
ProofBench15%—

Knowledge Gemini 3 Flash Preview leads

Gemini 3 Flash Preview: 58.8 (#33), Qwen3.6 Max Preview: 57.6 (#39)

Knowledge benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
GPQA Diamond89.4%87.4%
SimpleQA Verified66.8%52%
LMArena Expert14621478
Vectara Hallucination Rate13.5%—

Multimodal Not comparable

Gemini 3 Flash Preview: 45.5 (#16), Qwen3.6 Max Preview: —

Multimodal benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
LMArena Vision1285—
GeoBench88%—
VPCT72.6%—
Blueprint-Bench 20%—
LMArena Document1413—

Multilingual Gemini 3 Flash Preview leads

Gemini 3 Flash Preview: 55.7 (#27), Qwen3.6 Max Preview: 54.2 (#48)

Multilingual benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
LMArena Non-English14581437
LMArena Chinese15111487
LMArena French14771449
LMArena Russian14801445
LMArena Spanish14691454
LMArena German1497—
LMArena Japanese1489—
LMArena Korean1443—

Instruction Following Too close to call

Gemini 3 Flash Preview: 75.7 (#56), Qwen3.6 Max Preview: 75.7 (#55)

Instruction Following benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
LMArena Instruction Following14371438

Long Context Too close to call

Gemini 3 Flash Preview: 44.4 (#67), Qwen3.6 Max Preview: 44.6 (#61)

Long Context benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
LMArena Longer Query14521457

Writing & Preference Gemini 3 Flash Preview leads

Gemini 3 Flash Preview: 65.5 (#45), Qwen3.6 Max Preview: 63.8 (#60)

Writing & Preference benchmarks
BenchmarkGemini 3 Flash PreviewQwen3.6 Max Preview
LMArena Text14661447
LMArena Creative Writing14571435
LMArena Multi-Turn14711456

Frequently asked questions

Is Gemini 3 Flash Preview better than Qwen3.6 Max Preview?

Gemini 3 Flash Preview and Qwen3.6 Max Preview score almost the same on the Noometry Index (52.3 vs 51.5), so choose on price, context window or the category you care about most.

Which is cheaper, Gemini 3 Flash Preview or Qwen3.6 Max Preview?

Gemini 3 Flash Preview is cheaper. It lists at $0.50 per million input tokens and $3 per million output tokens; Qwen3.6 Max Preview lists at $1.30 and $7.80.

Is Gemini 3 Flash Preview or Qwen3.6 Max Preview better for coding?

Gemini 3 Flash Preview scores higher on coding benchmarks: 50.9 versus 48.7 in the Noometry coding category.

Which has the bigger context window?

Gemini 3 Flash Preview does, with 1.05M tokens against 262K.

How many benchmarks do Gemini 3 Flash Preview and Qwen3.6 Max Preview share?

29 benchmarks have published results for both models. Gemini 3 Flash Preview has 59 scored results on Noometry and Qwen3.6 Max Preview has 29.

Related comparisons

Go deeper