Model comparison
Qwen Turbo vs Step 3.7 Flash
Step 3.7 Flash is the stronger model overall, scoring 37.3 to 27.1 on the Noometry Index. Qwen Turbo costs 4.8× less per token, which makes it the better buy when Step 3.7 Flash's lead doesn't matter for your workload.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in math, where Step 3.7 Flash leads 42.9 to 15.3.
- Qwen Turbo is cheaper at $0.05 / $0.20 per million input/output tokens, against $0.18 / $1.11 for Step 3.7 Flash.
- Qwen Turbo accepts more context: 1M tokens versus 256K.
- Step 3.7 Flash has downloadable open weights; the other is API-only.
Side by side
| Qwen Turbo | Step 3.7 Flash | |
|---|---|---|
| Provider | Alibaba (Qwen) | StepFun |
| Noometry Index | 27.1 | 37.3 |
| Released | 2024-11-01 | 2026-05-29 |
| Weights | Proprietary | Open |
| Context window | 1M | 256K |
| Max output | 16K | 256K |
| Input $ / M tokens | $0.05 | $0.18 |
| Output $ / M tokens | $0.20 | $1.11 |
| Results tracked | 3 | 5 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
Qwen Turbo: —, Step 3.7 Flash: 40.0 (#150)
Reasoning Not comparable
Qwen Turbo: —, Step 3.7 Flash: 21.6 (#219)
| Benchmark | Qwen Turbo | Step 3.7 Flash |
|---|---|---|
| NYT Connections (extended) | — | 39.7% |
| CritPt | — | 2.3% |
Math Step 3.7 Flash leads
Qwen Turbo: 15.3 (#297), Step 3.7 Flash: 42.9 (#82)
| Benchmark | Qwen Turbo | Step 3.7 Flash |
|---|---|---|
| MathArena Final-Answer Competitions | — | 68.5% |
| OTIS Mock AIME 2024-2025 | 6.1% | — |
| MATH Level 5 | 56.2% | — |
Knowledge Not comparable
Qwen Turbo: 22.2 (#272), Step 3.7 Flash: —
| Benchmark | Qwen Turbo | Step 3.7 Flash |
|---|---|---|
| GPQA Diamond | 41.8% | — |
Frequently asked questions
Is Qwen Turbo better than Step 3.7 Flash?
Step 3.7 Flash is the stronger model overall, scoring 37.3 to 27.1 on the Noometry Index. Qwen Turbo costs 4.8× less per token, which makes it the better buy when Step 3.7 Flash's lead doesn't matter for your workload.
Which is cheaper, Qwen Turbo or Step 3.7 Flash?
Qwen Turbo is cheaper. It lists at $0.05 per million input tokens and $0.20 per million output tokens; Step 3.7 Flash lists at $0.18 and $1.11.
Which has the bigger context window?
Qwen Turbo does, with 1M tokens against 256K.
How many benchmarks do Qwen Turbo and Step 3.7 Flash share?
0 benchmarks have published results for both models. Qwen Turbo has 3 scored results on Noometry and Step 3.7 Flash has 5.