Model comparison

Nemotron 3.5 Lightning vs Qwen1.5 4b Chat

Nemotron 3.5 Lightning is the stronger model overall, scoring 40.0 to 28.8 on the Noometry Index.

Last verified . 13 shared benchmarks.

Nemotron 3.5 Lightning NVIDIA

40.0

Rank #155 Confirmed

Qwen1.5 4b Chat Alibaba (Qwen)

28.8

Rank #322 Confirmed

Summary

  • They share 13 benchmarks with published results for both. Nemotron 3.5 Lightning scores higher in 8 categories and Qwen1.5 4b Chat in 0 categories; 8 gaps are clear of the uncertainty.
  • The widest gap is in writing & preference, where Nemotron 3.5 Lightning leads 48.5 to 23.8.

Side by side

Nemotron 3.5 Lightning and Qwen1.5 4b Chat specifications
Nemotron 3.5 LightningQwen1.5 4b Chat
ProviderNVIDIAAlibaba (Qwen)
Noometry Index40.028.8
Released2026-08-11—
WeightsOpenOpen
Context window262K—
Max output262K—
Input $ / M tokens$0.05—
Output $ / M tokens$0.20—
Results tracked1813

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 40.4 (#141), Qwen1.5 4b Chat: 29.1 (#308)

Coding benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Coding1375999

Reasoning Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 26.8 (#127), Qwen1.5 4b Chat: 18.5 (#279)

Reasoning benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Hard Prompts1337976

Math Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 37.5 (#155), Qwen1.5 4b Chat: 30.4 (#234)

Math benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Math13591026

Knowledge Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 37.5 (#154), Qwen1.5 4b Chat: 26.7 (#255)

Knowledge benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Expert1356980

Multilingual Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 44.0 (#180), Qwen1.5 4b Chat: 24.1 (#290)

Multilingual benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Non-English1295979
LMArena Chinese13591024
LMArena German1282902
LMArena Russian1253952
LMArena French1366—
LMArena Japanese1206—
LMArena Korean1238—
LMArena Spanish1345—

Instruction Following Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 69.6 (#170), Qwen1.5 4b Chat: 49.0 (#300)

Instruction Following benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Instruction Following1318978

Long Context Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 39.9 (#165), Qwen1.5 4b Chat: 30.1 (#290)

Long Context benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Longer Query1314988

Writing & Preference Nemotron 3.5 Lightning leads

Nemotron 3.5 Lightning: 48.5 (#201), Qwen1.5 4b Chat: 23.8 (#309)

Writing & Preference benchmarks
BenchmarkNemotron 3.5 LightningQwen1.5 4b Chat
LMArena Text1327997
LMArena Creative Writing1254969
LMArena Multi-Turn1328977
EQ-Bench Creative Writing1280—

Frequently asked questions

Is Nemotron 3.5 Lightning better than Qwen1.5 4b Chat?

Nemotron 3.5 Lightning is the stronger model overall, scoring 40.0 to 28.8 on the Noometry Index.

Is Nemotron 3.5 Lightning or Qwen1.5 4b Chat better for coding?

Nemotron 3.5 Lightning scores higher on coding benchmarks: 40.4 versus 29.1 in the Noometry coding category.

How many benchmarks do Nemotron 3.5 Lightning and Qwen1.5 4b Chat share?

13 benchmarks have published results for both models. Nemotron 3.5 Lightning has 18 scored results on Noometry and Qwen1.5 4b Chat has 13.

Related comparisons

Go deeper