Model comparison

DeepSeek Coder 33B vs Qwen2.5-VL 72B Instruct

Qwen2.5-VL 72B Instruct has enough public results to be ranked (#302); DeepSeek Coder 33B does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

DeepSeek Coder 33B DeepSeek

38.9

Unranked Sparse

Side by side

DeepSeek Coder 33B and Qwen2.5-VL 72B Instruct specifications
DeepSeek Coder 33BQwen2.5-VL 72B Instruct
ProviderDeepSeekAlibaba (Qwen)
Noometry Index38.929.9
Released2023-11-022024-09
WeightsOpenOpen
Context window—131K
Max output—8K
Input $ / M tokens—$2.80
Output $ / M tokens—$8.40
Results tracked96

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

DeepSeek Coder 33B: 38.0 (#184), Qwen2.5-VL 72B Instruct: —

Coding benchmarks
BenchmarkDeepSeek Coder 33BQwen2.5-VL 72B Instruct
BigCodeBench Instruct42%—
BigCodeBench Complete51.1%—
HumanEval+75%—
MBPP+70.1%—

Agentic & Tool Use Not comparable

DeepSeek Coder 33B: —, Qwen2.5-VL 72B Instruct: 18.6 (#144)

Agentic & Tool Use benchmarks
BenchmarkDeepSeek Coder 33BQwen2.5-VL 72B Instruct
OSWorld—5%

Reasoning Not comparable

DeepSeek Coder 33B: —, Qwen2.5-VL 72B Instruct: 20.7 (#233)

Reasoning benchmarks
BenchmarkDeepSeek Coder 33BQwen2.5-VL 72B Instruct
Kagi LLM Benchmark—36%
Epoch Capabilities Index96.32—
WinoGrande62%—

Math Not comparable

DeepSeek Coder 33B: —, Qwen2.5-VL 72B Instruct: —

Math benchmarks
BenchmarkDeepSeek Coder 33BQwen2.5-VL 72B Instruct
GSM8K35.4%—

Knowledge Not comparable

DeepSeek Coder 33B: —, Qwen2.5-VL 72B Instruct: —

Knowledge benchmarks
BenchmarkDeepSeek Coder 33BQwen2.5-VL 72B Instruct
ARC (AI2) Challenge42.2%—
MMLU39.4%—

Multimodal Not comparable

DeepSeek Coder 33B: —, Qwen2.5-VL 72B Instruct: 33.5 (#97)

Multimodal benchmarks
BenchmarkDeepSeek Coder 33BQwen2.5-VL 72B Instruct
LMArena Vision—1107
Video-MME—73.5%
GeoBench—62%
SpatialViz-Bench—33.3%

Frequently asked questions

Is DeepSeek Coder 33B better than Qwen2.5-VL 72B Instruct?

Qwen2.5-VL 72B Instruct has enough public results to be ranked (#302); DeepSeek Coder 33B does not yet, so treat this comparison as directional.

How many benchmarks do DeepSeek Coder 33B and Qwen2.5-VL 72B Instruct share?

0 benchmarks have published results for both models. DeepSeek Coder 33B has 9 scored results on Noometry and Qwen2.5-VL 72B Instruct has 6.

Related comparisons

Go deeper