Model comparison

Qwen3.6 Max Preview vs Yi-Large

Qwen3.6 Max Preview has enough public results to be ranked (#43); Yi-Large does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Qwen3.6 Max Preview Alibaba (Qwen)

51.5

Rank #43 Confirmed

Yi-Large 01.AI

38.1

Unranked Sparse

Summary

  • The widest gap is in coding, where Qwen3.6 Max Preview leads 48.7 to 36.7.

Side by side

Qwen3.6 Max Preview and Yi-Large specifications
Qwen3.6 Max PreviewYi-Large
ProviderAlibaba (Qwen)01.AI
Noometry Index51.538.1
Released2026-04-202024-05-13
WeightsProprietaryProprietary
Context window262K—
Max output66K—
Input $ / M tokens$1.30—
Output $ / M tokens$7.80—
Results tracked293

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Qwen3.6 Max Preview leads

Qwen3.6 Max Preview: 48.7 (#54), Yi-Large: 36.7 (#203)

Coding benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
SWE-bench Verified76.7%—
LMArena WebDev1482—
BigCodeBench Instruct—37.7%
LMArena Coding1471—
BigCodeBench Complete—47.2%

Agentic & Tool Use Not comparable

Qwen3.6 Max Preview: —, Yi-Large: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
Vending-Bench 24,254—

Reasoning Not comparable

Qwen3.6 Max Preview: 41.7 (#53), Yi-Large: —

Reasoning benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
SimpleBench63%—
NYT Connections (extended)74.1%—
Chess Puzzles20%—
LMArena Hard Prompts1457—
Mystery Game Puzzles19%—
DTBench87.2%—
LMCA42.5%—
Epoch Capabilities Index149.24—

Math Not comparable

Qwen3.6 Max Preview: 54.1 (#46), Yi-Large: —

Math benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
OTIS Mock AIME 2024-202591.1%—
LMArena Math1465—
FrontierMath (Feb 2025 set)23.1%—
FrontierMath Tier 4 (v1)4.2%—

Knowledge Not comparable

Qwen3.6 Max Preview: 57.6 (#39), Yi-Large: —

Knowledge benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
GPQA Diamond87.4%—
SimpleQA Verified52%—
LMArena Expert1478—
MMLU—79.3%

Multilingual Not comparable

Qwen3.6 Max Preview: 54.2 (#48), Yi-Large: —

Multilingual benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
LMArena Non-English1437—
LMArena Chinese1487—
LMArena French1449—
LMArena Russian1445—
LMArena Spanish1454—

Instruction Following Not comparable

Qwen3.6 Max Preview: 75.7 (#55), Yi-Large: —

Instruction Following benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
LMArena Instruction Following1438—

Long Context Not comparable

Qwen3.6 Max Preview: 44.6 (#61), Yi-Large: —

Long Context benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
LMArena Longer Query1457—

Writing & Preference Not comparable

Qwen3.6 Max Preview: 63.8 (#60), Yi-Large: —

Writing & Preference benchmarks
BenchmarkQwen3.6 Max PreviewYi-Large
LMArena Text1447—
LMArena Creative Writing1435—
LMArena Multi-Turn1456—

Frequently asked questions

Is Qwen3.6 Max Preview better than Yi-Large?

Qwen3.6 Max Preview has enough public results to be ranked (#43); Yi-Large does not yet, so treat this comparison as directional.

Is Qwen3.6 Max Preview or Yi-Large better for coding?

Qwen3.6 Max Preview scores higher on coding benchmarks: 48.7 versus 36.7 in the Noometry coding category.

How many benchmarks do Qwen3.6 Max Preview and Yi-Large share?

0 benchmarks have published results for both models. Qwen3.6 Max Preview has 29 scored results on Noometry and Yi-Large has 3.

Related comparisons

Go deeper