Anthropic, proprietary

# Claude 3 Opus

> Claude 3 Opus by Anthropic, released February 2024. Ranked #310 of 354 with a Noometry Index of 29.5. Scores, sources and comparisons.
- Canonical page: https://noometry.com/models/claude-3-opus
- Last updated: 2026-10-10
- Title: Claude 3 Opus Benchmarks, Price & Rank (October 2026)

Claude 3 Opus by Anthropic ranks 310th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.5. Its strongest category is agentic & tool use, where it ranks 116th.

Last verified October 10, 2026

## Specifications

- **Noometry rank:** #310 of 354
- **Index score:** 29.5
- **Evidence:** Confirmed 46 results
- **Provider:** [Anthropic](https://noometry.com/providers/anthropic)
- **Released:** February 29, 2024
- **Weights:** Proprietary
- **Reasoning:** Unknown
- **Context window:** —
- **Max output:** —
- **Input price:** Not listed
- **Output price:** Not listed
- **Blended price:** Not listed
- **Output speed:** Not measured
- **Value:** Not ranked
- **Knowledge cutoff:** Unknown

## Category scores

Each category score combines every public result we have in that category.

Claude 3 Opus category scores

1.  Coding 32.9
2.  Agentic & Tool Use 24.6
3.  Reasoning 14.6
4.  Math 14.8
5.  Knowledge 24.5
6.  Multimodal 27.1
7.  Multilingual 41.4
8.  Instruction Following 64.1
9.  Long Context 38.2
10.  Writing & Preference 47.2
11.  020406080

Claude 3 Opus category ranks
| Category | Score | Rank | Results |
| --- | --- | --- | --- |
| [Coding](https://noometry.com/best/coding) | 32.9 | #267 | 5 |
| [Agentic & Tool Use](https://noometry.com/best/agentic) | 24.6 | #116 | 1 |
| [Reasoning](https://noometry.com/best/reasoning) | 14.6 | #324 | 8 |
| [Math](https://noometry.com/best/math) | 14.8 | #299 | 4 |
| [Knowledge](https://noometry.com/best/knowledge) | 24.5 | #267 | 4 |
| [Multimodal](https://noometry.com/best/multimodal) | 27.1 | #116 | 1 |
| [Multilingual](https://noometry.com/best/multilingual) | 41.4 | #207 | 1 |
| [Instruction Following](https://noometry.com/best/instruction-following) | 64.1 | #228 | 2 |
| [Long Context](https://noometry.com/best/long-context) | 38.2 | #202 | 1 |
| [Writing & Preference](https://noometry.com/best/writing) | 47.2 | #213 | 4 |

## Strengths and weaknesses

Categories where Claude 3 Opus places highest and lowest among the models ranked in each, with its score against that category's median.

### Strongest categories

Claude 3 Opus: strongest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Long Context](https://noometry.com/best/long-context) | 38.2 | −2.7 | #202 of 296, top 69% |
| [Writing & Preference](https://noometry.com/best/writing) | 47.2 | −6.6 | #213 of 312, top 69% |
| [Multilingual](https://noometry.com/best/multilingual) | 41.4 | −6.0 | #207 of 297, top 70% |

### Weakest categories

Claude 3 Opus: weakest categories
| Category | Score | vs median | Rank |
| --- | --- | --- | --- |
| [Reasoning](https://noometry.com/best/reasoning) | 14.6 | −9.0 | #324 of 350, top 93% |
| [Math](https://noometry.com/best/math) | 14.8 | −21.7 | #299 of 327, top 92% |
| [Multimodal](https://noometry.com/best/multimodal) | 27.1 | −11.4 | #116 of 128, top 91% |

## Closest competitors

The models ranked just above and below Claude 3 Opus. When scores are this close, price and speed are often the better way to choose.

Models ranked closest to Claude 3 Opus
| Model | Rank | Score | Blended $/M | Speed |  |
| --- | --- | --- | --- | --- | --- |
| [phi-3-medium 14B](https://noometry.com/models/phi-3-medium-14b) | #306 | 29.7 | — | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-phi-3-medium-14b) |
| [Gemma 2B](https://noometry.com/models/gemma-2b) | #307 | 29.6 | — | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-gemma-2b) |
| [Llama 3.1-70B](https://noometry.com/models/llama-3-1-70b) | #308 | 29.6 | $0.40 | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-llama-3-1-70b) |
| [Llama 2-13B](https://noometry.com/models/llama-2-13b) | #309 | 29.6 | — | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-llama-2-13b) |
| [DBRX](https://noometry.com/models/dbrx) | #311 | 29.4 | — | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-dbrx) |
| [Gemma 2 27B](https://noometry.com/models/gemma-2-27b) | #312 | 29.4 | $0.65 | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-gemma-2-27b) |
| [Gemma 1.1 2b IT](https://noometry.com/models/gemma-1-1-2b-it) | #313 | 29.3 | — | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-gemma-1-1-2b-it) |
| [Phi 3 Small 8k Instruct](https://noometry.com/models/phi-3-small-8k-instruct) | #314 | 29.3 | — | — | [Compare](https://noometry.com/compare/claude-3-opus-vs-phi-3-small-8k-instruct) |

Sponsored placements are available on pages like this one. [Advertise on Noometry](https://noometry.com/advertise)

## Benchmark results

Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.

### Coding

Claude 3 Opus Coding benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [WeirdML](https://noometry.com/benchmarks/weirdml) | 19.2% | #105 of 119, top 89% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [BigCodeBench Instruct](https://noometry.com/benchmarks/bigcodebench-instruct) | 45.5% | #18 of 64, top 29% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| [LiveBench Coding](https://noometry.com/benchmarks/livebench-coding) | 38.6% | #24 of 39, top 62% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Coding](https://noometry.com/benchmarks/arena-coding) | 1264 | #220 of 294, top 75% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [BigCodeBench Complete](https://noometry.com/benchmarks/bigcodebench-complete) | 57.4% | #13 of 66, top 20% |  | [BigCodeBench](https://bigcode-bench.github.io/) | 2024-02-29 |
| [HumanEval+](https://noometry.com/benchmarks/humaneval-plus) | 77.4% | #13 of 45, top 29% | mar 2024 | [EvalPlus](https://evalplus.github.io/leaderboard.html) |  |
| [MBPP+](https://noometry.com/benchmarks/mbpp-plus) | 73.3% | #8 of 38, top 22% | mar 2024 | [EvalPlus](https://evalplus.github.io/leaderboard.html) |  |

### Agentic & Tool Use

Claude 3 Opus Agentic & Tool Use benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [Cybench](https://noometry.com/benchmarks/cybench) | 10% | #15 of 21, top 72% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [METR Time Horizons](https://noometry.com/benchmarks/metr-time-horizons) | 29.5% | #31 of 32, top 97% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Reasoning

Claude 3 Opus Reasoning benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [SimpleBench](https://noometry.com/benchmarks/simplebench) | 23.5% | #67 of 77, top 88% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Chess Puzzles](https://noometry.com/benchmarks/chess-puzzles) | 5% | #92 of 129, top 72% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-07-15 |
| [EnigmaEval](https://noometry.com/benchmarks/enigmaeval) | 0.8% | #34 of 38, top 90% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Reasoning](https://noometry.com/benchmarks/livebench-reasoning) | 40.6% | #27 of 39, top 70% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Hard Prompts](https://noometry.com/benchmarks/arena-hard-prompts) | 1245 | #224 of 297, top 76% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [DTBench](https://noometry.com/benchmarks/dtbench) | 61.6% | #111 of 151, top 74% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench Data Analysis](https://noometry.com/benchmarks/livebench-data-analysis) | 57.9% | #16 of 39, top 42% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMCA](https://noometry.com/benchmarks/lmca) | 17% | #100 of 125, top 80% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [Epoch Capabilities Index](https://noometry.com/benchmarks/epoch-capabilities-index) | 126.91 | #152 of 213, top 72% |  | [Epoch AI](https://epoch.ai/eci) | 2024-02-29 |
| [ForecastBench](https://noometry.com/benchmarks/forecastbench) | 58.4 | #48 of 72, top 67% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LiveBench](https://noometry.com/benchmarks/livebench) | 49.2% | #21 of 39, top 54% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [WinoGrande](https://noometry.com/benchmarks/winogrande) | 88.5% | #2 of 43, top 5% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Math

Claude 3 Opus Math benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [OTIS Mock AIME 2024-2025](https://noometry.com/benchmarks/otis-mock-aime) | 4.7% | #148 of 173, top 86% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-02-25 |
| [LiveBench Math](https://noometry.com/benchmarks/livebench-math) | 43.6% | #22 of 39, top 57% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Math](https://noometry.com/benchmarks/arena-math) | 1273 | #198 of 285, top 70% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [MATH Level 5](https://noometry.com/benchmarks/math-level-5) | 37.5% | #54 of 79, top 69% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-01-27 |

### Knowledge

Claude 3 Opus Knowledge benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [GPQA Diamond](https://noometry.com/benchmarks/gpqa-diamond) | 47.2% | #138 of 186, top 75% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2025-01-27 |
| [SimpleQA Verified](https://noometry.com/benchmarks/simpleqa-verified) | 12.6% | #72 of 77, top 94% |  | [Epoch AI](https://epoch.ai/benchmarks) | 2026-08-31 |
| [Confabulations](https://noometry.com/benchmarks/confabulations) (lower is better) | 22.7% | #39 of 51, top 77% |  | [Lech Mazur benchmarks](https://github.com/lechmazur/confabulations) |  |
| [LMArena Expert](https://noometry.com/benchmarks/arena-expert) | 1223 | #213 of 273, top 79% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [MMLU](https://noometry.com/benchmarks/mmlu) | 84.6% | #9 of 81, top 12% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

### Multimodal

Claude 3 Opus Multimodal benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Vision](https://noometry.com/benchmarks/arena-vision) | 1023 | #115 of 122, top 95% |  | [LMArena](https://lmarena.ai/leaderboard/vision) | 2026-10-09 |

### Multilingual

Claude 3 Opus Multilingual benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Non-English](https://noometry.com/benchmarks/arena-non-english) | 1258 | #207 of 297, top 70% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Chinese](https://noometry.com/benchmarks/arena-chinese) | 1248 | #209 of 285, top 74% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena French](https://noometry.com/benchmarks/arena-french) | 1275 | #174 of 223, top 79% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena German](https://noometry.com/benchmarks/arena-german) | 1258 | #171 of 231, top 75% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Japanese](https://noometry.com/benchmarks/arena-japanese) | 1204 | #160 of 211, top 76% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Korean](https://noometry.com/benchmarks/arena-korean) | 1187 | #168 of 213, top 79% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Russian](https://noometry.com/benchmarks/arena-russian) | 1280 | #191 of 283, top 68% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Spanish](https://noometry.com/benchmarks/arena-spanish) | 1246 | #184 of 226, top 82% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Instruction Following

Claude 3 Opus Instruction Following benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LiveBench Instruction Following](https://noometry.com/benchmarks/livebench-if) | 63.9% | #24 of 39, top 62% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |
| [LMArena Instruction Following](https://noometry.com/benchmarks/arena-instruction-following) | 1248 | #214 of 298, top 72% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Long Context

Claude 3 Opus Long Context benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Longer Query](https://noometry.com/benchmarks/arena-longer-query) | 1259 | #216 of 291, top 75% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |

### Writing & Preference

Claude 3 Opus Writing & Preference benchmark results
| Benchmark | Score | Position | Setting | Source | Date |
| --- | --- | --- | --- | --- | --- |
| [LMArena Text](https://noometry.com/benchmarks/arena-text) | 1262 | #221 of 297, top 75% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Creative Writing](https://noometry.com/benchmarks/arena-creative-writing) | 1235 | #217 of 295, top 74% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LMArena Multi-Turn](https://noometry.com/benchmarks/arena-multi-turn) | 1275 | #205 of 295, top 70% |  | [LMArena](https://lmarena.ai/leaderboard/text) | 2026-10-08 |
| [LiveBench Language](https://noometry.com/benchmarks/livebench-language) | 50.4% | #11 of 39, top 29% |  | [Epoch AI](https://epoch.ai/benchmarks) |  |

## Compare Claude 3 Opus

-   [Claude 3 Opus vs Llama 2-13B](https://noometry.com/compare/claude-3-opus-vs-llama-2-13b)
-   [Claude 3 Opus vs DBRX](https://noometry.com/compare/claude-3-opus-vs-dbrx)
-   [Claude 3 Opus vs Llama 3.1-70B](https://noometry.com/compare/claude-3-opus-vs-llama-3-1-70b)
-   [Claude 3 Opus vs Gemma 2 27B](https://noometry.com/compare/claude-3-opus-vs-gemma-2-27b)
-   [Claude 3 Opus vs Gemma 2B](https://noometry.com/compare/claude-3-opus-vs-gemma-2b)
-   [Claude 3 Opus vs Gemma 1.1 2b IT](https://noometry.com/compare/claude-3-opus-vs-gemma-1-1-2b-it)
-   [Claude 3 Opus vs GPT-6 Astra](https://noometry.com/compare/claude-3-opus-vs-gpt-6-astra)
-   [Claude 3 Opus vs Gemini 3.8 Flash](https://noometry.com/compare/claude-3-opus-vs-gemini-3-8-flash)
-   [Claude 3 Opus vs Kimi K3](https://noometry.com/compare/claude-3-opus-vs-kimi-k3)
-   [Claude 3 Opus vs Grok 4.6](https://noometry.com/compare/claude-3-opus-vs-grok-4-6)
-   [Claude 3 Opus vs Qwen3.8 Max](https://noometry.com/compare/claude-3-opus-vs-qwen3-8-max)
-   [Claude 3 Opus vs GLM-5.3](https://noometry.com/compare/claude-3-opus-vs-glm-5-3)
-   [Claude 3 Opus vs Muse Spark 1.3](https://noometry.com/compare/claude-3-opus-vs-muse-spark-1-3)
-   [Claude 3 Opus vs DeepSeek V4 Pro](https://noometry.com/compare/claude-3-opus-vs-deepseek-v4-pro)

## Other Anthropic models

-   [Claude Fable 5.1](https://noometry.com/models/claude-fable-5-1)69.0
-   [Claude Opus 5.5](https://noometry.com/models/claude-opus-5-5)68.6
-   [Claude Opus 5](https://noometry.com/models/claude-opus-5)67.8
-   [Claude Fable 5](https://noometry.com/models/claude-fable-5)66.8
-   [Claude Sonnet 5.5](https://noometry.com/models/claude-sonnet-5-5)61.9
-   [Claude Opus 4.8](https://noometry.com/models/claude-opus-4-8)60.7
-   [Claude Opus 4.7](https://noometry.com/models/claude-opus-4-7)58.3
-   [Claude Opus 4.6](https://noometry.com/models/claude-opus-4-6)58.2

## Frequently asked questions

### How good is Claude 3 Opus?

Claude 3 Opus by Anthropic ranks 310th of 354 ranked models on the Noometry Index as of October 2026, with a score of 29.5. Its strongest category is agentic & tool use, where it ranks 116th.

### Is Claude 3 Opus open source?

No. Claude 3 Opus is proprietary and available only through Anthropic's API and partner platforms.

### What are Claude 3 Opus's strengths and weaknesses?

Relative to other ranked models, Claude 3 Opus places best in long context, writing & preference, multilingual and lowest in reasoning, math, multimodal.

### What is Claude 3 Opus best at?

Its best category is agentic & tool use, where it ranks 116th on Noometry.

### Cite this page

Noometry. (2026). Claude 3 Opus benchmarks and pricing. Retrieved October 10, 2026, from https://noometry.com/models/claude-3-opus

Quote Noometry with a link back to this page. It is also available in [Markdown](https://noometry.com/md/models/claude-3-opus.md).
