Meta, proprietary
Muse Spark 1.1
Muse Spark 1.1 by Meta ranks 51st of 354 ranked models on the Noometry Index as of October 2026, with a score of 49.9. Its strongest category is writing & preference, where it ranks 11th. API pricing starts at $1.25 per million input tokens and $4.25 per million output tokens, with a 1.05M-token context window.
Last verified
Specifications
- Noometry rank
- #51 of 354
- Index score
- 49.9
- Evidence
- Confirmed 37 results
- Provider
Meta
- Released
- April 8, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 1.05M
- Max output
- 131K
- Input price
- $1.25 / M
- Output price
- $4.25 / M
- Blended price
- $2 / M
- Output speed
- Not measured
- Value
- #144 of 219
- Knowledge cutoff
- Unknown
- Input
- text, image, pdf, video
Category scores
Each category score combines every public result we have in that category.
- Coding 51.3
- Agentic & Tool Use 30.8
- Reasoning 47.1
- Math 45.5
- Knowledge 53.1
- Multimodal 42.6
- Multilingual 56.7
- Instruction Following 76.5
- Long Context 44.8
- Writing & Preference 73.4
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 51.3 | #40 | 4 |
| Agentic & Tool Use | 30.8 | #73 | 4 |
| Reasoning | 47.1 | #44 | 6 |
| Math | 45.5 | #76 | 2 |
| Knowledge | 53.1 | #59 | 2 |
| Multimodal | 42.6 | #29 | 1 |
| Multilingual | 56.7 | #17 | 1 |
| Instruction Following | 76.5 | #39 | 1 |
| Long Context | 44.8 | #58 | 1 |
| Writing & Preference | 73.4 | #11 | 5 |
Strengths and weaknesses
Categories where Muse Spark 1.1 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Writing & Preference | 73.4 | +19.7 | #11 of 312, top 4% |
| Multilingual | 56.7 | +9.3 | #17 of 297, top 6% |
| Coding | 51.3 | +12.6 | #40 of 340, top 12% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Agentic & Tool Use | 30.8 | +0.5 | #73 of 154, top 48% |
| Math | 45.5 | +8.9 | #76 of 327, top 24% |
| Multimodal | 42.6 | +4.1 | #29 of 128, top 23% |
Closest competitors
The models ranked just above and below Muse Spark 1.1. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Claude Opus 4.5 | #47 | 50.5 | $10 | 13 | Compare |
| Muse Spark 1.2 | #48 | 50.3 | $2 | — | Compare |
| MiMo-V2.6-Pro | #49 | 50.3 | $0.54 | — | Compare |
| Claude Sonnet 4.6 | #50 | 50.3 | $6 | — | Compare |
| Claude Haiku 5.5 | #52 | 49.5 | $0.20 | — | Compare |
| GPT-5.1 | #53 | 49.0 | $3.44 | — | Compare |
| Grok 4.20 (Non-Reasoning) | #54 | 48.6 | $1.56 | 61 | Compare |
| MiMo-V2.6-Flash | #55 | 48.5 | $0.18 | — | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| DeepSWE | 53.3% | #22 of 29, top 76% | Epoch AI | ||
| LMArena WebDev | 1542 | #35 of 113, top 31% | LMArena | 2026-10-08 | |
| SciCode | 58.8% | #13 of 121, top 11% | Epoch AI | ||
| SciCode | 58.8% | xhigh | Epoch AI | ||
| LMArena Coding | 1498 | #20 of 294, top 7% | LMArena | 2026-10-08 |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| APEX-Agents | 31.8% | #43 of 49, top 88% | Epoch AI | ||
| τ²-bench Banking | 40.5% | #6 of 26, top 24% | xhigh | τ²-bench | 2026-08-04 |
| GBAEval | 7.9% | #13 of 23, top 57% | Epoch AI | ||
| GDP.pdf | 15% | #27 of 36, top 75% | Epoch AI | ||
| GDP.pdf | 15% | medium | Epoch AI | ||
| Vending-Bench 2 | 6,520 | #16 of 60, top 27% | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| NYT Connections (extended) | 84.9% | #30 of 91, top 33% | high reasoning | Lech Mazur benchmarks | |
| CritPt | 15.1% | Epoch AI | |||
| CritPt | 15.1% | #37 of 134, top 28% | xhigh | Epoch AI | |
| LMArena Hard Prompts | 1486 | #21 of 297, top 8% | LMArena | 2026-10-08 | |
| DTBench | 94.4% | #23 of 151, top 16% | high | Epoch AI | |
| LMCA | 49.9% | #22 of 125, top 18% | high | Epoch AI | |
| Surface Evolver Bench | 52.5% | #15 of 25, top 60% | high | Epoch AI | |
| Epoch Capabilities Index | 154.21 | #37 of 213, top 18% | Epoch AI | 2026-07-09 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| ProofBench | 39% | #36 of 77, top 47% | Epoch AI | ||
| LMArena Math | 1483 | #25 of 285, top 9% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| SimpleQA Verified | 57.8% | #17 of 77, top 23% | Epoch AI | 2026-08-31 | |
| LMArena Expert | 1478 | #43 of 273, top 16% | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1293 | #21 of 122, top 18% | LMArena | 2026-10-09 | |
| LMArena Document | 1465 | #15 of 38, top 40% | LMArena | 2026-09-13 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1472 | #17 of 297, top 6% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1518 | #29 of 285, top 11% | LMArena | 2026-10-08 | |
| LMArena French | 1494 | #16 of 223, top 8% | LMArena | 2026-10-08 | |
| LMArena German | 1466 | #30 of 231, top 13% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1451 | #24 of 211, top 12% | LMArena | 2026-10-08 | |
| LMArena Korean | 1458 | #13 of 213, top 7% | LMArena | 2026-10-08 | |
| LMArena Russian | 1483 | #17 of 283, top 7% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1464 | #31 of 226, top 14% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1457 | #37 of 298, top 13% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1462 | #40 of 291, top 14% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1479 | #18 of 297, top 7% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1437 | #42 of 295, top 15% | LMArena | 2026-10-08 | |
| EQ-Bench Creative Writing | 1927 | #12 of 115, top 11% | EQ-Bench | ||
| EQ-Bench 4 | 1260 | #8 of 28, top 29% | EQ-Bench | ||
| LMArena Multi-Turn | 1485 | #14 of 295, top 5% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| meta | $1.25 | $4.25 | $0.15 | 2026-10-10 |
| openrouter | $1.25 | $4.25 | $0.15 | 2026-10-10 |
Compare Muse Spark 1.1
- Muse Spark 1.1 vs Claude Sonnet 4.6
- Muse Spark 1.1 vs Claude Haiku 5.5
- Muse Spark 1.1 vs MiMo-V2.6-Pro
- Muse Spark 1.1 vs GPT-5.1
- Muse Spark 1.1 vs Muse Spark 1.2
- Muse Spark 1.1 vs Grok 4.20 (Non-Reasoning)
- Muse Spark 1.1 vs GPT-6 Astra
- Muse Spark 1.1 vs Claude Fable 5.1
- Muse Spark 1.1 vs Gemini 3.8 Flash
- Muse Spark 1.1 vs Kimi K3
- Muse Spark 1.1 vs Grok 4.6
- Muse Spark 1.1 vs Qwen3.8 Max
- Muse Spark 1.1 vs GLM-5.3
- Muse Spark 1.1 vs DeepSeek V4 Pro
Other Meta models
- Muse Spark 1.354.8
- Muse Spark50.6
- Muse Spark 1.250.3
- Muse Glimmer41.7
- Codellama 70b Instruct33.7
- Llama 4 Maverick30.9
- Codellama 34b Instruct30.8
- Llama 3.1-405B30.7
Frequently asked questions
How good is Muse Spark 1.1?
Muse Spark 1.1 by Meta ranks 51st of 354 ranked models on the Noometry Index as of October 2026, with a score of 49.9. Its strongest category is writing & preference, where it ranks 11th. API pricing starts at $1.25 per million input tokens and $4.25 per million output tokens, with a 1.05M-token context window.
How much does Muse Spark 1.1 cost?
Muse Spark 1.1 costs $1.25 per million input tokens and $4.25 per million output tokens on Meta's own API, with cached input at $0.15.
What is Muse Spark 1.1's context window?
Muse Spark 1.1 accepts up to 1.05M tokens of input and can write up to 131K tokens in one response.
Is Muse Spark 1.1 open source?
No. Muse Spark 1.1 is proprietary and available only through Meta's API and partner platforms.
What are Muse Spark 1.1's strengths and weaknesses?
Relative to other ranked models, Muse Spark 1.1 places best in writing & preference, multilingual, coding and lowest in agentic & tool use, math, multimodal.
What is Muse Spark 1.1 best at?
Its best category is writing & preference, where it ranks 11th on Noometry.