Model comparison

Longcat Flash Chat vs Qwen3.6 27B

Longcat Flash Chat and Qwen3.6 27B score almost the same on the Noometry Index (42.1 vs 42.2), so choose on price, context window or the category you care about most.

Last verified . 0 shared benchmarks.

Longcat Flash Chat Meituan

42.1

Rank #120 Confirmed

Qwen3.6 27B Alibaba (Qwen)

42.2

Rank #117 Confirmed

Summary

  • The widest gap is in knowledge, where Qwen3.6 27B leads 52.4 to 40.6.

Side by side

Longcat Flash Chat and Qwen3.6 27B specifications
Longcat Flash ChatQwen3.6 27B
ProviderMeituanAlibaba (Qwen)
Noometry Index42.142.2
Released—2026-04-22
WeightsOpenOpen
Context window—262K
Max output—66K
Input $ / M tokens—$0.60
Output $ / M tokens—$3.60
Results tracked1911

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Longcat Flash Chat leads

Longcat Flash Chat: 43.5 (#87), Qwen3.6 27B: 39.1 (#163)

Coding benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
SciCode—37.3%
LMArena Coding1471—

Reasoning Qwen3.6 27B leads

Longcat Flash Chat: 19.0 (#272), Qwen3.6 27B: 25.0 (#153)

Reasoning benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
Kagi LLM Benchmark43.9%—
NYT Connections (extended)17.7%—
CritPt—0.9%
Chess Puzzles—22%
LMArena Hard Prompts1440—
Mystery Game Puzzles—7%
DTBench—78.1%
LMCA—34.5%
Epoch Capabilities Index—146.5

Math Qwen3.6 27B leads

Longcat Flash Chat: 39.4 (#107), Qwen3.6 27B: 48.5 (#62)

Math benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
FrontierMath (Tiers 1-3)—35.1%
OTIS Mock AIME 2024-2025—91.1%
LMArena Math1442—

Knowledge Qwen3.6 27B leads

Longcat Flash Chat: 40.6 (#116), Qwen3.6 27B: 52.4 (#63)

Knowledge benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
GPQA Diamond—85.9%
LMArena Expert1454—

Multilingual Not comparable

Longcat Flash Chat: 51.9 (#101), Qwen3.6 27B: —

Multilingual benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
LMArena Non-English1404—
LMArena Chinese1465—
LMArena French1456—
LMArena German1408—
LMArena Japanese1373—
LMArena Korean1371—
LMArena Russian1395—
LMArena Spanish1445—

Instruction Following Not comparable

Longcat Flash Chat: 74.4 (#96), Qwen3.6 27B: —

Instruction Following benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
LMArena Instruction Following1411—

Long Context Not comparable

Longcat Flash Chat: 43.5 (#93), Qwen3.6 27B: —

Long Context benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
LMArena Longer Query1425—

Writing & Preference Longcat Flash Chat leads

Longcat Flash Chat: 61.0 (#91), Qwen3.6 27B: 50.3 (#181)

Writing & Preference benchmarks
BenchmarkLongcat Flash ChatQwen3.6 27B
LMArena Text1427—
LMArena Creative Writing1388—
EQ-Bench 4—1026
LMArena Multi-Turn1418—

Frequently asked questions

Is Longcat Flash Chat better than Qwen3.6 27B?

Longcat Flash Chat and Qwen3.6 27B score almost the same on the Noometry Index (42.1 vs 42.2), so choose on price, context window or the category you care about most.

Is Longcat Flash Chat or Qwen3.6 27B better for coding?

Longcat Flash Chat scores higher on coding benchmarks: 43.5 versus 39.1 in the Noometry coding category.

How many benchmarks do Longcat Flash Chat and Qwen3.6 27B share?

0 benchmarks have published results for both models. Longcat Flash Chat has 19 scored results on Noometry and Qwen3.6 27B has 11.

Related comparisons

Go deeper