Model comparison

Qwen3.7 Plus vs Ring-2.6-1T

Qwen3.7 Plus has enough public results to be ranked (#72); Ring-2.6-1T does not yet, so treat this comparison as directional.

Last verified . 2 shared benchmarks.

Qwen3.7 Plus Alibaba (Qwen)

45.3

Rank #72 Confirmed

Summary

  • They share 2 benchmarks with published results for both. Qwen3.7 Plus scores higher in 1 category and Ring-2.6-1T in 1 category; 2 gaps are clear of the uncertainty.
  • The widest gap is in reasoning, where Qwen3.7 Plus leads 39.3 to 26.1.
  • The biggest single-benchmark swing is CritPt: 9.1% for Qwen3.7 Plus and 3.7% for Ring-2.6-1T.
  • Ring-2.6-1T has downloadable open weights; the other is API-only.

Side by side

Qwen3.7 Plus and Ring-2.6-1T specifications
Qwen3.7 PlusRing-2.6-1T
ProviderAlibaba (Qwen)Ant Group (inclusionAI)
Noometry Index45.336.9
Released2026-06-022026-05-14
WeightsProprietaryOpen
Context window1M—
Max output131K—
Input $ / M tokens$0.40—
Output $ / M tokens$1.60—
Results tracked323

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Ring-2.6-1T leads

Qwen3.7 Plus: 36.6 (#206), Ring-2.6-1T: 40.7

Coding benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
SciCode45.5%42.4%
FrontierCode10.2%—
LMArena Coding1473—
ALE-Bench—432.57

Agentic & Tool Use Not comparable

Qwen3.7 Plus: 21.4 (#138), Ring-2.6-1T: —

Agentic & Tool Use benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
OSWorld 2.02.8%—

Reasoning Qwen3.7 Plus leads

Qwen3.7 Plus: 39.3 (#59), Ring-2.6-1T: 26.1

Reasoning benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
CritPt9.1%3.7%
NYT Connections (extended)74.8%—
Chess Puzzles24%—
LMArena Hard Prompts1460—
Mystery Game Puzzles17%—
DTBench84%—
LMCA37.6%—
Epoch Capabilities Index147.37—

Math Not comparable

Qwen3.7 Plus: 50.5 (#56), Ring-2.6-1T: —

Math benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
FrontierMath (Tiers 1-3)34.4%—
OTIS Mock AIME 2024-202593.3%—
LMArena Math1466—

Knowledge Not comparable

Qwen3.7 Plus: 54.9 (#51), Ring-2.6-1T: —

Knowledge benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
GPQA Diamond87.9%—
LMArena Expert1467—

Multimodal Not comparable

Qwen3.7 Plus: 41.8 (#33), Ring-2.6-1T: —

Multimodal benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
LMArena Vision1279—
LMArena Document1444—

Multilingual Not comparable

Qwen3.7 Plus: 54.8 (#38), Ring-2.6-1T: —

Multilingual benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
LMArena Non-English1445—
LMArena Chinese1510—
LMArena French1473—
LMArena German1471—
LMArena Japanese1413—
LMArena Korean1415—
LMArena Russian1457—
LMArena Spanish1457—

Instruction Following Not comparable

Qwen3.7 Plus: 75.8 (#52), Ring-2.6-1T: —

Instruction Following benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
LMArena Instruction Following1440—

Long Context Not comparable

Qwen3.7 Plus: 44.5 (#65), Ring-2.6-1T: —

Long Context benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
LMArena Longer Query1455—

Writing & Preference Not comparable

Qwen3.7 Plus: 64.3 (#56), Ring-2.6-1T: —

Writing & Preference benchmarks
BenchmarkQwen3.7 PlusRing-2.6-1T
LMArena Text1455—
LMArena Creative Writing1439—
LMArena Multi-Turn1460—

Frequently asked questions

Is Qwen3.7 Plus better than Ring-2.6-1T?

Qwen3.7 Plus has enough public results to be ranked (#72); Ring-2.6-1T does not yet, so treat this comparison as directional.

Is Qwen3.7 Plus or Ring-2.6-1T better for coding?

Ring-2.6-1T scores higher on coding benchmarks: 40.7 versus 36.6 in the Noometry coding category.

How many benchmarks do Qwen3.7 Plus and Ring-2.6-1T share?

2 benchmarks have published results for both models. Qwen3.7 Plus has 32 scored results on Noometry and Ring-2.6-1T has 3.

Related comparisons

Go deeper