Model comparison
Grok 4.1 vs Solar Pro4
Grok 4.1 and Solar Pro4 score almost the same on the Noometry Index (41.5 vs 42.1), so choose on price, context window or the category you care about most.
Last verified . 18 shared benchmarks.
Summary
- They share 18 benchmarks with published results for both. Grok 4.1 scores higher in 6 categories and Solar Pro4 in 2 categories; 6 gaps are clear of the uncertainty.
- The widest gap is in coding, where Solar Pro4 leads 40.1 to 33.7.
Side by side
| Grok 4.1 | Solar Pro4 | |
|---|---|---|
| Provider | xAI | Upstage |
| Noometry Index | 41.5 | 42.1 |
| Released | 2025-11-17 | 2026-08-06 |
| Weights | Proprietary | Proprietary |
| Context window | — | 524K |
| Max output | — | 131K |
| Input $ / M tokens | — | $0.30 |
| Output $ / M tokens | — | $1.20 |
| Results tracked | 19 | 18 |
Sponsored placements are available on pages like this one. Advertise on Noometry
Category by category
Coding Solar Pro4 leads
Grok 4.1: 33.7 (#253), Solar Pro4: 40.1 (#149)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena WebDev | 1214 | 1371 |
| LMArena Coding | 1445 | 1437 |
Agentic & Tool Use Not comparable
Grok 4.1: 34.1 (#49), Solar Pro4: —
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| Cybench | 39% | — |
Reasoning Grok 4.1 leads
Grok 4.1: 29.5 (#91), Solar Pro4: 28.5 (#104)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Hard Prompts | 1435 | 1399 |
Math Too close to call
Grok 4.1: 38.9 (#120), Solar Pro4: 38.8 (#128)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Math | 1422 | 1416 |
Knowledge Too close to call
Grok 4.1: 39.5 (#133), Solar Pro4: 39.8 (#129)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Expert | 1417 | 1427 |
Multilingual Grok 4.1 leads
Grok 4.1: 53.4 (#68), Solar Pro4: 48.7 (#139)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Non-English | 1425 | 1361 |
| LMArena Chinese | 1465 | 1415 |
| LMArena French | 1448 | 1397 |
| LMArena German | 1446 | 1364 |
| LMArena Japanese | 1397 | 1309 |
| LMArena Korean | 1407 | 1382 |
| LMArena Russian | 1434 | 1360 |
| LMArena Spanish | 1438 | 1401 |
Instruction Following Grok 4.1 leads
Grok 4.1: 73.8 (#111), Solar Pro4: 72.7 (#132)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Instruction Following | 1400 | 1377 |
Long Context Grok 4.1 leads
Grok 4.1: 43.2 (#100), Solar Pro4: 42.1 (#130)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Longer Query | 1416 | 1381 |
Writing & Preference Grok 4.1 leads
Grok 4.1: 62.4 (#75), Solar Pro4: 56.5 (#138)
| Benchmark | Grok 4.1 | Solar Pro4 |
|---|---|---|
| LMArena Text | 1437 | 1386 |
| LMArena Creative Writing | 1411 | 1316 |
| LMArena Multi-Turn | 1437 | 1385 |
Frequently asked questions
Is Grok 4.1 better than Solar Pro4?
Grok 4.1 and Solar Pro4 score almost the same on the Noometry Index (41.5 vs 42.1), so choose on price, context window or the category you care about most.
Is Grok 4.1 or Solar Pro4 better for coding?
Solar Pro4 scores higher on coding benchmarks: 40.1 versus 33.7 in the Noometry coding category.
How many benchmarks do Grok 4.1 and Solar Pro4 share?
18 benchmarks have published results for both models. Grok 4.1 has 19 scored results on Noometry and Solar Pro4 has 18.