Model comparison
Nvidia Llama 3.3 Nemotron Super 49b v1.5 vs Qwen3.5 35B-A3B
Qwen3.5 35B-A3B is the stronger model overall, scoring 42.0 to 40.3 on the Noometry Index. Nvidia Llama 3.3 Nemotron Super 49b v1.5 costs 1.7× less per token, which makes it the better buy when Qwen3.5 35B-A3B's lead doesn't matter for your workload.
Last verified . 12 shared benchmarks.
Summary
- They share 12 benchmarks with published results for both. Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher in 2 categories and Qwen3.5 35B-A3B in 6 categories; 8 gaps are clear of the uncertainty.
- The widest gap is in knowledge, where Qwen3.5 35B-A3B leads 47.8 to 36.7.
- Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper at $0.40 / $0.40 per million input/output tokens, against $0.25 / $2 for Qwen3.5 35B-A3B.
- Qwen3.5 35B-A3B accepts more context: 262K tokens versus 131K.
Side by side
| Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B | |
|---|---|---|
| Provider | NVIDIA | Alibaba (Qwen) |
| Noometry Index | 40.3 | 42.0 |
| Released | 2025-07-25 | 2026-02-01 |
| Weights | Open | Open |
| Context window | 131K | 262K |
| Max output | 131K | 66K |
| Input $ / M tokens | $0.40 | $0.25 |
| Output $ / M tokens | $0.40 | $2 |
| Results tracked | 12 | 28 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 39.8 (#154), Qwen3.5 35B-A3B: 33.8 (#251)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Coding | 1355 | 1410 |
| LMArena WebDev | — | 1254 |
| SciCode | — | 29.3% |
Reasoning Nvidia Llama 3.3 Nemotron Super 49b v1.5 leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 26.8 (#128), Qwen3.5 35B-A3B: 24.6 (#161)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Hard Prompts | 1336 | 1400 |
| CritPt | — | 0.6% |
| Chess Puzzles | — | 10% |
| DTBench | — | 80% |
| LMCA | — | 29.5% |
| Epoch Capabilities Index | — | 142.52 |
Math Qwen3.5 35B-A3B leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 38.2 (#141), Qwen3.5 35B-A3B: 39.9 (#97)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Math | 1392 | 1404 |
| MathArena Final-Answer Competitions | — | 56% |
| OTIS Mock AIME 2024-2025 | — | 70% |
Knowledge Qwen3.5 35B-A3B leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 36.7 (#165), Qwen3.5 35B-A3B: 47.8 (#79)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Expert | 1330 | 1408 |
| GPQA Diamond | — | 83.5% |
| Vectara Hallucination Rate | — | 10.5% |
Multilingual Qwen3.5 35B-A3B leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 45.5 (#168), Qwen3.5 35B-A3B: 50.0 (#127)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Non-English | 1316 | 1378 |
| LMArena Japanese | 1300 | 1325 |
| LMArena Russian | 1332 | 1376 |
| LMArena Chinese | — | 1457 |
| LMArena French | — | 1412 |
| LMArena German | — | 1367 |
| LMArena Korean | — | 1356 |
| LMArena Spanish | — | 1392 |
Instruction Following Qwen3.5 35B-A3B leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 68.6 (#188), Qwen3.5 35B-A3B: 72.8 (#128)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Instruction Following | 1299 | 1379 |
Long Context Qwen3.5 35B-A3B leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 40.0 (#164), Qwen3.5 35B-A3B: 42.4 (#127)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Longer Query | 1315 | 1389 |
Writing & Preference Qwen3.5 35B-A3B leads
Nvidia Llama 3.3 Nemotron Super 49b v1.5: 53.1 (#159), Qwen3.5 35B-A3B: 57.9 (#124)
| Benchmark | Nvidia Llama 3.3 Nemotron Super 49b v1.5 | Qwen3.5 35B-A3B |
|---|---|---|
| LMArena Text | 1338 | 1395 |
| LMArena Creative Writing | 1307 | 1346 |
| LMArena Multi-Turn | 1334 | 1390 |
Frequently asked questions
Is Nvidia Llama 3.3 Nemotron Super 49b v1.5 better than Qwen3.5 35B-A3B?
Qwen3.5 35B-A3B is the stronger model overall, scoring 42.0 to 40.3 on the Noometry Index. Nvidia Llama 3.3 Nemotron Super 49b v1.5 costs 1.7× less per token, which makes it the better buy when Qwen3.5 35B-A3B's lead doesn't matter for your workload.
Which is cheaper, Nvidia Llama 3.3 Nemotron Super 49b v1.5 or Qwen3.5 35B-A3B?
Nvidia Llama 3.3 Nemotron Super 49b v1.5 is cheaper. It lists at $0.40 per million input tokens and $0.40 per million output tokens; Qwen3.5 35B-A3B lists at $0.25 and $2.
Is Nvidia Llama 3.3 Nemotron Super 49b v1.5 or Qwen3.5 35B-A3B better for coding?
Nvidia Llama 3.3 Nemotron Super 49b v1.5 scores higher on coding benchmarks: 39.8 versus 33.8 in the Noometry coding category.
Which has the bigger context window?
Qwen3.5 35B-A3B does, with 262K tokens against 131K.
How many benchmarks do Nvidia Llama 3.3 Nemotron Super 49b v1.5 and Qwen3.5 35B-A3B share?
12 benchmarks have published results for both models. Nvidia Llama 3.3 Nemotron Super 49b v1.5 has 12 scored results on Noometry and Qwen3.5 35B-A3B has 28.