Model comparison
DeepSeek Coder 33B vs Mercury
Mercury has enough public results to be ranked (#199); DeepSeek Coder 33B does not yet, so treat this comparison as directional.
Last verified . 0 shared benchmarks.
Summary
- DeepSeek Coder 33B has downloadable open weights; the other is API-only.
Side by side
| DeepSeek Coder 33B | Mercury | |
|---|---|---|
| Provider | DeepSeek | Inception |
| Noometry Index | 38.9 | 37.6 |
| Released | 2023-11-02 | — |
| Weights | Open | Proprietary |
| Context window | — | — |
| Max output | — | — |
| Input $ / M tokens | — | — |
| Output $ / M tokens | — | — |
| Results tracked | 9 | 9 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Too close to call
DeepSeek Coder 33B: 38.0 (#184), Mercury: 38.7 (#170)
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| BigCodeBench Instruct | 42% | — |
| LMArena Coding | — | 1322 |
| BigCodeBench Complete | 51.1% | — |
| HumanEval+ | 75% | — |
| MBPP+ | 70.1% | — |
Reasoning Not comparable
DeepSeek Coder 33B: —, Mercury: 17.5 (#293)
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| Kagi LLM Benchmark | — | 21.6% |
| LMArena Hard Prompts | — | 1285 |
| Epoch Capabilities Index | 96.32 | — |
| WinoGrande | 62% | — |
Math Not comparable
DeepSeek Coder 33B: —, Mercury: —
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| GSM8K | 35.4% | — |
Knowledge Not comparable
DeepSeek Coder 33B: —, Mercury: —
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| ARC (AI2) Challenge | 42.2% | — |
| MMLU | 39.4% | — |
Multilingual Not comparable
DeepSeek Coder 33B: —, Mercury: 41.6 (#206)
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| LMArena Non-English | — | 1260 |
Instruction Following Not comparable
DeepSeek Coder 33B: —, Mercury: 65.2 (#224)
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| LMArena Instruction Following | — | 1239 |
Long Context Not comparable
DeepSeek Coder 33B: —, Mercury: 38.4 (#198)
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| LMArena Longer Query | — | 1266 |
Writing & Preference Not comparable
DeepSeek Coder 33B: —, Mercury: 46.2 (#221)
| Benchmark | DeepSeek Coder 33B | Mercury |
|---|---|---|
| LMArena Text | — | 1282 |
| LMArena Creative Writing | — | 1191 |
| LMArena Multi-Turn | — | 1282 |
Frequently asked questions
Is DeepSeek Coder 33B better than Mercury?
Mercury has enough public results to be ranked (#199); DeepSeek Coder 33B does not yet, so treat this comparison as directional.
Is DeepSeek Coder 33B or Mercury better for coding?
They score almost the same on coding (38.0 vs 38.7); test both on your own repository before choosing.
How many benchmarks do DeepSeek Coder 33B and Mercury share?
0 benchmarks have published results for both models. DeepSeek Coder 33B has 9 scored results on Noometry and Mercury has 9.