Coding benchmark

# BigCodeBench Complete leaderboard

> BigCodeBench Complete results for 66 AI models, led by DeepSeek-V3 at 62.2%. What the benchmark measures, who runs it, and a source for every score.
- Canonical page: https://noometry.com/benchmarks/bigcodebench-complete
- Last updated: 2026-10-10
- Title: BigCodeBench Complete Leaderboard (October 2026): Scores by Model

As of October 2026, DeepSeek-V3 has the highest published BigCodeBench Complete score on Noometry at 62.2%, out of 66 models with results.

Last verified October 10, 2026

## About BigCodeBench Complete

The BigCodeBench tasks posed as docstring completion instead of instructions.

- **Category:** [Coding](https://noometry.com/best/coding)
- **Introduced:** 2024
- **Size:** 1,140 tasks
- **Format:** Code completion
- **Unit:** Percent (random guessing ≈ 0%)
- **Official site:** [bigcode-bench.github.io](https://bigcode-bench.github.io/)

## Top 15 models

Top models on BigCodeBench Complete

1.  DeepSeek-V3 62.2%
2.  Llama 4 Maverick 61.4%
3.  GPT-4o 61.1%
4.  Gemini 2.0 Flash (Feb 2025) 59.9%
5.  Deepseek Coder v2 59.7%
6.  DeepSeek-V2 (MoE-236B, May 2024) 59.4%
7.  Claude 3.5 Haiku 59%
8.  Claude 3.5 Sonnet 58.6%
9.  GPT-4 Turbo 58.2%
10.  Qwen2.5-Coder-32B 58%
11.  Gemini 1.5 Pro (May 2024) 57.5%
12.  Llama-3.3-70B-Instruct 57.5%
13.  Claude 3 Opus 57.4%
14.  GPT-4o mini 57.4%
15.  GPT-4 57.2%
16.  5658606264

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## All results

BigCodeBench Complete results by model
| # | Model | Provider | Score | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- | --- |
| 1 | [DeepSeek-V3](https://noometry.com/models/deepseek-v3) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 62.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-12-26 |
| 2 | [Llama 4 Maverick](https://noometry.com/models/llama-4-maverick) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 61.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-04-05 |
| 3 | [GPT-4o](https://noometry.com/models/gpt-4o) | [OpenAI](https://noometry.com/providers/openai) | 61.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-13 |
| 4 | [Gemini 2.0 Flash (Feb 2025)](https://noometry.com/models/gemini-2-0-flash) |  [![](/logos/google.svg) Google](https://noometry.com/providers/google) | 59.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-02-05 |
| 5 | [Deepseek Coder v2](https://noometry.com/models/deepseek-coder-v2) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 59.7% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-17 |
| 6 | [DeepSeek-V2 (MoE-236B, May 2024)](https://noometry.com/models/deepseek-v2) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 59.4% | 2024-06-28 | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-28 |
| 7 | [Claude 3.5 Haiku](https://noometry.com/models/claude-3-5-haiku) | [Anthropic](https://noometry.com/providers/anthropic) | 59% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-10-22 |
| 8 | [Claude 3.5 Sonnet](https://noometry.com/models/claude-3-5-sonnet) | [Anthropic](https://noometry.com/providers/anthropic) | 58.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-20 |
| 9 | [GPT-4 Turbo](https://noometry.com/models/gpt-4-turbo) | [OpenAI](https://noometry.com/providers/openai) | 58.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-09 |
| 10 | [Qwen2.5-Coder-32B](https://noometry.com/models/qwen2-5-coder-32b) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 58% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-19 |
| 11 | [Gemini 1.5 Pro (May 2024)](https://noometry.com/models/gemini-1-5-pro) |  [![](/logos/google.svg) Google](https://noometry.com/providers/google) | 57.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-14 |
| 12 | [Llama-3.3-70B-Instruct](https://noometry.com/models/llama-3-3-70b-instruct) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 57.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-12-19 |
| 13 | [Claude 3 Opus](https://noometry.com/models/claude-3-opus) | [Anthropic](https://noometry.com/providers/anthropic) | 57.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| 14 | [GPT-4o mini](https://noometry.com/models/gpt-4o-mini) | [OpenAI](https://noometry.com/providers/openai) | 57.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-07-18 |
| 15 | [GPT-4](https://noometry.com/models/gpt-4) | [OpenAI](https://noometry.com/providers/openai) | 57.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-13 |
| 16 | [Qwen2.5 72B Instruct](https://noometry.com/models/qwen2-5-72b-instruct) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 55.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-19 |
| 17 | [Phi-4](https://noometry.com/models/phi-4) |  [![](/logos/microsoft.svg) Microsoft](https://noometry.com/providers/microsoft) | 55.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-12-13 |
| 18 | [Gemini 1.5 Flash (May 2024)](https://noometry.com/models/gemini-1-5-flash) |  [![](/logos/google.svg) Google](https://noometry.com/providers/google) | 55.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-14 |
| 19 | [DeepSeek-R1-Distill-Qwen-32B](https://noometry.com/models/deepseek-r1-distill-qwen-32b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 54.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-01-20 |
| 20 | [Llama 3.1-70B](https://noometry.com/models/llama-3-1-70b) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 54.8% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-07-23 |
| 21 | [Llama 3-70B](https://noometry.com/models/llama-3-70b) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 54.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-18 |
| 22 | [QwQ-32B](https://noometry.com/models/qwq-32b) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 54.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-11-28 |
| 23 | [Qwen2-72B](https://noometry.com/models/qwen2-72b) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 54% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-07 |
| 24 | [Claude 3 Sonnet](https://noometry.com/models/claude-3-sonnet) | [Anthropic](https://noometry.com/providers/anthropic) | 53.8% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| 25 | [DeepSeek-V2.5 (Sep 2024)](https://noometry.com/models/deepseek-v2-5) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 53.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-12-10 |
| 26 | [Codestral](https://noometry.com/models/codestral) |  [![](/logos/mistral.svg) Mistral AI](https://noometry.com/providers/mistral) | 52.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-23 |
| 27 | [Gemma 2 27B](https://noometry.com/models/gemma-2-27b) |  [![](/logos/google.svg) Google](https://noometry.com/providers/google) | 52.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-19 |
| 28 | [Qwen2.5 32B Instruct](https://noometry.com/models/qwen2-5-32b-instruct) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 52.3% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-19 |
| 29 | [Qwen2.5 14B Instruct](https://noometry.com/models/qwen2-5-14b-instruct) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 52.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-19 |
| 30 | [DeepSeek Coder 33B](https://noometry.com/models/deepseek-coder-33b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 51.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2023-10-28 |
| 31 | [GPT-3.5-turbo](https://noometry.com/models/gpt-3-5-turbo) | [OpenAI](https://noometry.com/providers/openai) | 50.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-01-25 |
| 32 | [Mistral Small 3](https://noometry.com/models/mistral-small-3) |  [![](/logos/mistral.svg) Mistral AI](https://noometry.com/providers/mistral) | 50.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-01-31 |
| 33 | [Mixtral 8x22B](https://noometry.com/models/mixtral-8x22b) |  [![](/logos/mistral.svg) Mistral AI](https://noometry.com/providers/mistral) | 50.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-17 |
| 34 | [Claude 3 Haiku](https://noometry.com/models/claude-3-haiku) | [Anthropic](https://noometry.com/providers/anthropic) | 50.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-03-07 |
| 35 | [DeepSeek-R1-Distill-Llama-70B](https://noometry.com/models/deepseek-r1-distill-llama-70b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 49.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-01-20 |
| 36 | [Codellama 70b Instruct](https://noometry.com/models/codellama-70b-instruct) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 49.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2023-08-25 |
| 37 | [phi-3-medium 14B](https://noometry.com/models/phi-3-medium-14b) |  [![](/logos/microsoft.svg) Microsoft](https://noometry.com/providers/microsoft) | 48.7% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-21 |
| 38 | [DeepSeek-R1-Distill-Qwen-14B](https://noometry.com/models/deepseek-r1-distill-qwen-14b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 48.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-01-20 |
| 39 | [Llama 3.1 Nemotron 70b Instruct](https://noometry.com/models/llama-3-1-nemotron-70b-instruct) |  [![](/logos/nvidia.svg) NVIDIA](https://noometry.com/providers/nvidia) | 48.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-25 |
| 40 | [Yi-Large](https://noometry.com/models/yi-large) | [01.AI](https://noometry.com/providers/01-ai) | 47.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-13 |
| 41 | [Mistral Small](https://noometry.com/models/mistral-small) |  [![](/logos/mistral.svg) Mistral AI](https://noometry.com/providers/mistral) | 46.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-18 |
| 42 | [Qwen2.5 7B Instruct](https://noometry.com/models/qwen2-5-7b-instruct) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 46.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-19 |
| 43 | [Command R](https://noometry.com/models/command-r) |  [![](/logos/cohere.svg) Cohere](https://noometry.com/providers/cohere) | 45.2% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-08-30 |
| 44 | [Qwen1.5-110B](https://noometry.com/models/qwen1-5-110b) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 44.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-26 |
| 45 | [DeepSeek Coder 6.7B](https://noometry.com/models/deepseek-coder-6-7b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 43.8% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2023-10-28 |
| 46 | [Yi-1.5-34B](https://noometry.com/models/yi-1-5-34b) | [01.AI](https://noometry.com/providers/01-ai) | 43.8% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-20 |
| 47 | [Llama 4 Scout](https://noometry.com/models/llama-4-scout) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 43.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-04-05 |
| 48 | [Qwen1.5-32B](https://noometry.com/models/qwen1-5-32b) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 42% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-26 |
| 49 | [Command R+](https://noometry.com/models/command-r-plus) |  [![](/logos/cohere.svg) Cohere](https://noometry.com/providers/cohere) | 41.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-04 |
| 50 | [Gemma 2 9B](https://noometry.com/models/gemma-2-9b) |  [![](/logos/google.svg) Google](https://noometry.com/providers/google) | 40.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-06-19 |
| 51 | [Phi 3 Mini 128k Instruct](https://noometry.com/models/phi-3-mini-128k-instruct) |  [![](/logos/microsoft.svg) Microsoft](https://noometry.com/providers/microsoft) | 40.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-21 |
| 52 | [Llama 3.1-8B](https://noometry.com/models/llama-3-1-8b) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 40.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-07-23 |
| 53 | [Qwen1.5-72B](https://noometry.com/models/qwen1-5-72b) |  [![](/logos/alibaba.svg) Alibaba (Qwen)](https://noometry.com/providers/alibaba) | 40.3% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-26 |
| 54 | [Phi-3.5-mini](https://noometry.com/models/phi-3-5-mini) |  [![](/logos/microsoft.svg) Microsoft](https://noometry.com/providers/microsoft) | 38.5% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-08-21 |
| 55 | [StarCoder 2 15B](https://noometry.com/models/starcoder-2-15b) |  [![](/logos/nvidia.svg) NVIDIA](https://noometry.com/providers/nvidia) | 38.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| 56 | [Mistral Large](https://noometry.com/models/mistral-large) |  [![](/logos/mistral.svg) Mistral AI](https://noometry.com/providers/mistral) | 38.3% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-26 |
| 57 | [Codellama 34b Instruct](https://noometry.com/models/codellama-34b-instruct) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 37.1% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2023-08-25 |
| 58 | [Llama 3-8B](https://noometry.com/models/llama-3-8b) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 36.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-04-18 |
| 59 | [Granite 3.0 8b Instruct](https://noometry.com/models/granite-3-0-8b-instruct) | [IBM](https://noometry.com/providers/ibm) | 35.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-10-21 |
| 60 | [DeepSeek Coder 1.3B](https://noometry.com/models/deepseek-coder-1-3b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 29.6% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2023-10-28 |
| 61 | [Llama 3.2 3B](https://noometry.com/models/llama-3-2-3b) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 28.3% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-25 |
| 62 | [StarCoder 2 7B](https://noometry.com/models/starcoder-2-7b) |  [![](/logos/nvidia.svg) NVIDIA](https://noometry.com/providers/nvidia) | 27.7% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| 63 | [Mistral 7B](https://noometry.com/models/mistral-7b) |  [![](/logos/mistral.svg) Mistral AI](https://noometry.com/providers/mistral) | 27.3% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-05-22 |
| 64 | [StarCoder 2 3B](https://noometry.com/models/starcoder-2-3b) |  [![](/logos/nvidia.svg) NVIDIA](https://noometry.com/providers/nvidia) | 21.4% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| 65 | [Llama 3.2 1B](https://noometry.com/models/llama-3-2-1b) |  [![](/logos/meta.svg) Meta](https://noometry.com/providers/meta) | 11.3% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-09-25 |
| 66 | [DeepSeek-R1-Distill-Qwen-1.5B](https://noometry.com/models/deepseek-r1-distill-qwen-1-5b) |  [![](/logos/deepseek.svg) DeepSeek](https://noometry.com/providers/deepseek) | 7.9% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2025-01-20 |

## Compare the leaders

-   [DeepSeek-V3 vs Llama 4 Maverick](https://noometry.com/compare/deepseek-v3-vs-llama-4-maverick)
-   [DeepSeek-V3 vs GPT-4o](https://noometry.com/compare/deepseek-v3-vs-gpt-4o)
-   [DeepSeek-V3 vs Gemini 2.0 Flash (Feb 2025)](https://noometry.com/compare/deepseek-v3-vs-gemini-2-0-flash)
-   [DeepSeek-V3 vs Deepseek Coder v2](https://noometry.com/compare/deepseek-coder-v2-vs-deepseek-v3)
-   [Llama 4 Maverick vs GPT-4o](https://noometry.com/compare/gpt-4o-vs-llama-4-maverick)
-   [Llama 4 Maverick vs Gemini 2.0 Flash (Feb 2025)](https://noometry.com/compare/gemini-2-0-flash-vs-llama-4-maverick)

## Other coding benchmarks

-   [SWE-bench Verified](https://noometry.com/benchmarks/swe-bench-verified)
-   [DeepSWE](https://noometry.com/benchmarks/deepswe)
-   [FrontierCode](https://noometry.com/benchmarks/frontiercode)
-   [SWE-bench Verified (bash only)](https://noometry.com/benchmarks/swe-bench-bash-only)
-   [Aider Polyglot](https://noometry.com/benchmarks/aider-polyglot)
-   [LMArena WebDev](https://noometry.com/benchmarks/arena-webdev)
-   [CursorBench](https://noometry.com/benchmarks/cursorbench)
-   [SWE-bench Multilingual](https://noometry.com/benchmarks/swe-bench-multilingual)
-   [FrontierSWE](https://noometry.com/benchmarks/frontierswe)
-   [SciCode](https://noometry.com/benchmarks/scicode)
-   [GSO](https://noometry.com/benchmarks/gso-bench)
-   [WeirdML](https://noometry.com/benchmarks/weirdml)

## Frequently asked questions

### What does BigCodeBench Complete measure?

The BigCodeBench tasks posed as docstring completion instead of instructions.

### Which model has the highest BigCodeBench Complete score?

As of October 2026, DeepSeek-V3 has the highest published BigCodeBench Complete score on Noometry at 62.2%, out of 66 models with results.

### What is the best open-weight model on BigCodeBench Complete?

DeepSeek-V3 has the highest BigCodeBench Complete accuracy among open-weight models at 62.2%, ranking 1 of 66 overall.

### Cite this page

Noometry. (2026). BigCodeBench Complete leaderboard. Retrieved October 10, 2026, from https://noometry.com/benchmarks/bigcodebench-complete

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/benchmarks/bigcodebench-complete.md).
