Model comparison

Longcat Flash Chat vs Qwen3.5 35B-A3B

Longcat Flash Chat and Qwen3.5 35B-A3B score almost the same on the Noometry Index (42.1 vs 42.0), so choose on price, context window or the category you care about most.

Last verified . 17 shared benchmarks.

Longcat Flash Chat Meituan

42.1

Rank #120 Confirmed

Qwen3.5 35B-A3B Alibaba (Qwen)

42.0

Rank #123 Confirmed

Summary

  • They share 17 benchmarks with published results for both. Longcat Flash Chat scores higher in 5 categories and Qwen3.5 35B-A3B in 3 categories; 7 gaps are clear of the uncertainty.
  • The widest gap is in coding, where Longcat Flash Chat leads 43.5 to 33.8.

Side by side

Longcat Flash Chat and Qwen3.5 35B-A3B specifications
Longcat Flash ChatQwen3.5 35B-A3B
ProviderMeituanAlibaba (Qwen)
Noometry Index42.142.0
Released—2026-02-01
WeightsOpenOpen
Context window—262K
Max output—66K
Input $ / M tokens—$0.25
Output $ / M tokens—$2
Results tracked1928

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Longcat Flash Chat leads

Longcat Flash Chat: 43.5 (#87), Qwen3.5 35B-A3B: 33.8 (#251)

Coding benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Coding14711410
LMArena WebDev—1254
SciCode—29.3%

Reasoning Qwen3.5 35B-A3B leads

Longcat Flash Chat: 19.0 (#272), Qwen3.5 35B-A3B: 24.6 (#161)

Reasoning benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Hard Prompts14401400
Kagi LLM Benchmark43.9%—
NYT Connections (extended)17.7%—
CritPt—0.6%
Chess Puzzles—10%
DTBench—80%
LMCA—29.5%
Epoch Capabilities Index—142.52

Math Too close to call

Longcat Flash Chat: 39.4 (#107), Qwen3.5 35B-A3B: 39.9 (#97)

Math benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Math14421404
MathArena Final-Answer Competitions—56%
OTIS Mock AIME 2024-2025—70%

Knowledge Qwen3.5 35B-A3B leads

Longcat Flash Chat: 40.6 (#116), Qwen3.5 35B-A3B: 47.8 (#79)

Knowledge benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Expert14541408
GPQA Diamond—83.5%
Vectara Hallucination Rate—10.5%

Multilingual Longcat Flash Chat leads

Longcat Flash Chat: 51.9 (#101), Qwen3.5 35B-A3B: 50.0 (#127)

Multilingual benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Non-English14041378
LMArena Chinese14651457
LMArena French14561412
LMArena German14081367
LMArena Japanese13731325
LMArena Korean13711356
LMArena Russian13951376
LMArena Spanish14451392

Instruction Following Longcat Flash Chat leads

Longcat Flash Chat: 74.4 (#96), Qwen3.5 35B-A3B: 72.8 (#128)

Instruction Following benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Instruction Following14111379

Long Context Longcat Flash Chat leads

Longcat Flash Chat: 43.5 (#93), Qwen3.5 35B-A3B: 42.4 (#127)

Long Context benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Longer Query14251389

Writing & Preference Longcat Flash Chat leads

Longcat Flash Chat: 61.0 (#91), Qwen3.5 35B-A3B: 57.9 (#124)

Writing & Preference benchmarks
BenchmarkLongcat Flash ChatQwen3.5 35B-A3B
LMArena Text14271395
LMArena Creative Writing13881346
LMArena Multi-Turn14181390

Frequently asked questions

Is Longcat Flash Chat better than Qwen3.5 35B-A3B?

Longcat Flash Chat and Qwen3.5 35B-A3B score almost the same on the Noometry Index (42.1 vs 42.0), so choose on price, context window or the category you care about most.

Is Longcat Flash Chat or Qwen3.5 35B-A3B better for coding?

Longcat Flash Chat scores higher on coding benchmarks: 43.5 versus 33.8 in the Noometry coding category.

How many benchmarks do Longcat Flash Chat and Qwen3.5 35B-A3B share?

17 benchmarks have published results for both models. Longcat Flash Chat has 19 scored results on Noometry and Qwen3.5 35B-A3B has 28.

Related comparisons

Go deeper