Model comparison

GPT-5.5 Pro vs Wizardlm 70b

GPT-5.5 Pro is the stronger model overall, scoring 64.3 to 33.0 on the Noometry Index.

Last verified . 0 shared benchmarks.

GPT-5.5 Pro OpenAI

64.3

Rank #8 Confirmed

Wizardlm 70b Microsoft

33.0

Rank #249 Confirmed

Summary

  • The widest gap is in reasoning, where GPT-5.5 Pro leads 73.3 to 20.7.
  • Wizardlm 70b has downloadable open weights; the other is API-only.

Side by side

GPT-5.5 Pro and Wizardlm 70b specifications
GPT-5.5 ProWizardlm 70b
ProviderOpenAIMicrosoft
Noometry Index64.333.0
Released2026-04-23—
WeightsProprietaryOpen
Context window1.05M—
Max output128K—
Input $ / M tokens$30—
Output $ / M tokens$180—
Results tracked1412

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

GPT-5.5 Pro: —, Wizardlm 70b: 31.4 (#285)

Coding benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
LMArena Coding—1081

Reasoning GPT-5.5 Pro leads

GPT-5.5 Pro: 73.3 (#10), Wizardlm 70b: 20.7 (#234)

Reasoning benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
ARC-AGI-284.6%—
SimpleBench76.9%—
ARC-AGI-196.5%—
CritPt30.6%—
Chess Puzzles64%—
LMArena Hard Prompts—1079
DTBench96%—
LMCA53.9%—
Epoch Capabilities Index162.07—

Math GPT-5.5 Pro leads

GPT-5.5 Pro: 84.0 (#10), Wizardlm 70b: 32.2 (#218)

Math benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
FrontierMath (Tiers 1-3)87.7%—
FrontierMath Tier 478%—
OTIS Mock AIME 2024-2025100%—
LMArena Math—1116
FrontierMath (Feb 2025 set)52.4%—
FrontierMath Tier 4 (v1)39.6%—

Knowledge Not comparable

GPT-5.5 Pro: 64.1 (#19), Wizardlm 70b: —

Knowledge benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
GPQA Diamond93.9%—

Multilingual Not comparable

GPT-5.5 Pro: —, Wizardlm 70b: 29.6 (#265)

Multilingual benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
LMArena Non-English—1078
LMArena Chinese—1052
LMArena German—1083
LMArena Russian—1155

Instruction Following Not comparable

GPT-5.5 Pro: —, Wizardlm 70b: 56.3 (#273)

Instruction Following benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
LMArena Instruction Following—1093

Long Context Not comparable

GPT-5.5 Pro: —, Wizardlm 70b: 33.3 (#263)

Long Context benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
LMArena Longer Query—1097

Writing & Preference Not comparable

GPT-5.5 Pro: —, Wizardlm 70b: 34.8 (#269)

Writing & Preference benchmarks
BenchmarkGPT-5.5 ProWizardlm 70b
LMArena Text—1120
LMArena Creative Writing—1149
LMArena Multi-Turn—1108

Frequently asked questions

Is GPT-5.5 Pro better than Wizardlm 70b?

GPT-5.5 Pro is the stronger model overall, scoring 64.3 to 33.0 on the Noometry Index.

How many benchmarks do GPT-5.5 Pro and Wizardlm 70b share?

0 benchmarks have published results for both models. GPT-5.5 Pro has 14 scored results on Noometry and Wizardlm 70b has 12.

Related comparisons

Go deeper