Model comparison
GPT-5.1-Codex vs Nova Premier 1.0
GPT-5.1-Codex and Nova Premier 1.0 score almost the same on the Noometry Index (38.6 vs 38.3), so choose on price, context window or the category you care about most.
Last verified . 0 shared benchmarks.
Summary
- The widest gap is in math, where Nova Premier 1.0 leads 33.8 to 30.3.
- GPT-5.1-Codex is cheaper at $1.25 / $10 per million input/output tokens, against $2.50 / $12.50 for Nova Premier 1.0.
- Nova Premier 1.0 accepts more context: 1M tokens versus 400K.
Side by side
| GPT-5.1-Codex | Nova Premier 1.0 | |
|---|---|---|
| Provider | OpenAI | Amazon |
| Noometry Index | 38.6 | 38.3 |
| Released | 2025-11-12 | 2025-04-30 |
| Weights | Proprietary | Proprietary |
| Context window | 400K | 1M |
| Max output | 128K | 10K |
| Input $ / M tokens | $1.25 | $2.50 |
| Output $ / M tokens | $10 | $12.50 |
| Results tracked | 6 | 6 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Not comparable
GPT-5.1-Codex: 41.9 (#116), Nova Premier 1.0: —
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| SWE-bench Verified (bash only) | 66% | — |
| LMArena WebDev | 1337 | — |
| ALE-Bench | 1,245 | — |
Agentic & Tool Use Not comparable
GPT-5.1-Codex: 38.0 (#33), Nova Premier 1.0: —
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| Terminal-Bench | 60.4% | — |
| METR Time Horizons | 70.8% | — |
Reasoning Not comparable
GPT-5.1-Codex: —, Nova Premier 1.0: 23.0 (#185)
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| Kagi LLM Benchmark | — | 44.8% |
Math Nova Premier 1.0 leads
GPT-5.1-Codex: 30.3 (#235), Nova Premier 1.0: 33.8 (#199)
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| ProofBench | 9% | — |
| Omni-MATH | — | 35% |
Knowledge Not comparable
GPT-5.1-Codex: —, Nova Premier 1.0: 35.8 (#180)
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| MMLU-Pro | — | 72.6% |
| GPQA (HELM) | — | 51.8% |
Instruction Following Not comparable
GPT-5.1-Codex: —, Nova Premier 1.0: 66.8 (#204)
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| IFEval | — | 80.3% |
Writing & Preference Not comparable
GPT-5.1-Codex: —, Nova Premier 1.0: 51.0 (#178)
| Benchmark | GPT-5.1-Codex | Nova Premier 1.0 |
|---|---|---|
| WildBench | — | 78.8% |
Frequently asked questions
Is GPT-5.1-Codex better than Nova Premier 1.0?
GPT-5.1-Codex and Nova Premier 1.0 score almost the same on the Noometry Index (38.6 vs 38.3), so choose on price, context window or the category you care about most.
Which is cheaper, GPT-5.1-Codex or Nova Premier 1.0?
GPT-5.1-Codex is cheaper. It lists at $1.25 per million input tokens and $10 per million output tokens; Nova Premier 1.0 lists at $2.50 and $12.50.
Which has the bigger context window?
Nova Premier 1.0 does, with 1M tokens against 400K.
How many benchmarks do GPT-5.1-Codex and Nova Premier 1.0 share?
0 benchmarks have published results for both models. GPT-5.1-Codex has 6 scored results on Noometry and Nova Premier 1.0 has 6.