Meta, proprietary
Muse Spark 1.3
Muse Spark 1.3 by Meta ranks 27th of 354 ranked models on the Noometry Index as of October 2026, with a score of 54.8. Its strongest category is multilingual, where it ranks 8th. API pricing starts at $1.25 per million input tokens and $4.25 per million output tokens, with a 1.05M-token context window.
Last verified
Specifications
- Noometry rank
- #27 of 354
- Index score
- 54.8
- Evidence
- Confirmed 37 results
- Provider
Meta
- Released
- September 2, 2026
- Weights
- Proprietary
- Reasoning
- Yes
- Context window
- 1.05M
- Max output
- 131K
- Input price
- $1.25 / M
- Output price
- $4.25 / M
- Blended price
- $2 / M
- Output speed
- Not measured
- Value
- #139 of 219
- Knowledge cutoff
- Unknown
- Input
- text, image, video, pdf, audio
Category scores
Each category score combines every public result we have in that category.
- Coding 56.6
- Agentic & Tool Use 38.6
- Reasoning 54.0
- Math 73.1
- Knowledge 42.6
- Multimodal 43.7
- Multilingual 57.4
- Instruction Following 77.5
- Long Context 45.6
- Writing & Preference 73.6
| Category | Score | Rank | Results |
|---|---|---|---|
| Coding | 56.6 | #21 | 4 |
| Agentic & Tool Use | 38.6 | #30 | 2 |
| Reasoning | 54.0 | #27 | 7 |
| Math | 73.1 | #21 | 5 |
| Knowledge | 42.6 | #95 | 1 |
| Multimodal | 43.7 | #22 | 1 |
| Multilingual | 57.4 | #8 | 1 |
| Instruction Following | 77.5 | #22 | 1 |
| Long Context | 45.6 | #32 | 1 |
| Writing & Preference | 73.6 | #9 | 4 |
Strengths and weaknesses
Categories where Muse Spark 1.3 places highest and lowest among the models ranked in each, with its score against that category's median.
Strongest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Multilingual | 57.4 | +10.0 | #8 of 297, top 3% |
| Writing & Preference | 73.6 | +19.8 | #9 of 312, top 3% |
| Coding | 56.6 | +17.9 | #21 of 340, top 7% |
Weakest categories
| Category | Score | vs median | Rank |
|---|---|---|---|
| Knowledge | 42.6 | +5.3 | #95 of 314, top 31% |
| Agentic & Tool Use | 38.6 | +8.3 | #30 of 154, top 20% |
| Multimodal | 43.7 | +5.1 | #22 of 128, top 18% |
Closest competitors
The models ranked just above and below Muse Spark 1.3. When scores are this close, price and speed are often the better way to choose.
| Model | Rank | Score | Blended $/M | Speed | |
|---|---|---|---|---|---|
| Gemini 3.1 Pro Preview | #23 | 56.7 | $4.50 | — | Compare |
| Gemini 4 Argon | #24 | 56.5 | — | — | Compare |
| Grok 4.5 | #25 | 55.0 | $3 | 4 | Compare |
| GLM-5.3 | #26 | 54.8 | $2.15 | — | Compare |
| Gemini 3 Pro | #28 | 54.8 | — | 1 | Compare |
| Claude Sonnet 5 | #29 | 54.6 | $4 | — | Compare |
| GPT-5.6 Luna | #30 | 54.6 | $0.45 | 12 | Compare |
| DeepSeek V4 Pro | #31 | 54.3 | $0.99 | 16 | Compare |
Sponsored placements are available on pages like this one. Advertise on Noometry
Benchmark results
Every published result we track, with its source. Bold rows are the ones used for ranking; where several exist we prefer independent runs over self-reported numbers.
Coding
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| CursorBench | 33.4% | high | Epoch AI | ||
| CursorBench | 29.3% | low | Epoch AI | ||
| CursorBench | 41.6% | #8 of 14, top 58% | max | Epoch AI | |
| CursorBench | 32.6% | medium | Epoch AI | ||
| CursorBench | 24.3% | minimal | Epoch AI | ||
| CursorBench | 37.5% | xhigh | Epoch AI | ||
| LMArena WebDev | 1657 | #10 of 113, top 9% | LMArena | 2026-10-08 | |
| LMArena WebDev | 1626 | muse-spark-1.3 (xHigh) | LMArena | 2026-10-08 | |
| SciCode | 58.8% | max | Epoch AI | ||
| SciCode | 59.7% | #8 of 121, top 7% | xhigh | Epoch AI | |
| LMArena Coding | 1514 | #9 of 294, top 4% | LMArena | 2026-10-08 |
Agentic & Tool Use
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| APEX-Agents | 57.8% | #15 of 49, top 31% | Epoch AI | ||
| GDP.pdf | 27.6% | #7 of 36, top 20% | Epoch AI | ||
| GDP.pdf | 27.6% | xhigh | Epoch AI |
Reasoning
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| NYT Connections (extended) | 85.1% | #28 of 91, top 31% | high reasoning | Lech Mazur benchmarks | |
| CritPt | 24.9% | max | Epoch AI | ||
| CritPt | 26% | #16 of 134, top 12% | xhigh | Epoch AI | |
| Chess Puzzles | 38% | #25 of 129, top 20% | max | Epoch AI | 2026-09-18 |
| Chess Puzzles | 35% | xhigh | Epoch AI | 2026-09-16 | |
| LMArena Hard Prompts | 1503 | #10 of 297, top 4% | LMArena | 2026-10-08 | |
| Mystery Game Puzzles | 25% | #34 of 74, top 46% | max | Epoch AI | 2026-09-18 |
| Mystery Game Puzzles | 14.1% | xhigh | Epoch AI | 2026-09-16 | |
| DTBench | 96.5% | #11 of 151, top 8% | Epoch AI | ||
| LMCA | 53.9% | #14 of 125, top 12% | Epoch AI | ||
| Bench to the Future 3 | 0.14 | #5 of 10, top 50% | xhigh | Epoch AI | |
| Epoch Capabilities Index | 156.75 | #19 of 213, top 9% | Epoch AI | 2026-09-02 |
Math
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| FrontierMath (Tiers 1-3) | 74% | max | Epoch AI | 2026-09-18 | |
| FrontierMath (Tiers 1-3) | 74.4% | #20 of 81, top 25% | xhigh | Epoch AI | 2026-09-16 |
| FrontierMath Tier 4 | 46.3% | #20 of 63, top 32% | max | Epoch AI | 2026-09-18 |
| FrontierMath Tier 4 | 41.5% | xhigh | Epoch AI | 2026-09-16 | |
| OTIS Mock AIME 2024-2025 | 99.2% | #14 of 173, top 9% | xhigh | Epoch AI | 2026-09-17 |
| ProofBench | 55% | Epoch AI | |||
| ProofBench | 58% | #21 of 77, top 28% | max | Epoch AI | |
| LMArena Math | 1494 | #16 of 285, top 6% | LMArena | 2026-10-08 |
Knowledge
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Expert | 1516 | #13 of 273, top 5% | LMArena | 2026-10-08 |
Multimodal
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Vision | 1309 | #11 of 122, top 10% | LMArena | 2026-10-09 | |
| LMArena Vision | 1301 | muse-spark-1.3 (xHigh) | LMArena | 2026-10-09 | |
| LMArena Document | 1471 | #12 of 38, top 32% | LMArena | 2026-09-13 |
Multilingual
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Non-English | 1481 | #8 of 297, top 3% | LMArena | 2026-10-08 | |
| LMArena Chinese | 1529 | #16 of 285, top 6% | LMArena | 2026-10-08 | |
| LMArena French | 1524 | #2 of 223, top 1% | LMArena | 2026-10-08 | |
| LMArena German | 1515 | #3 of 231, top 2% | LMArena | 2026-10-08 | |
| LMArena Japanese | 1474 | #16 of 211, top 8% | LMArena | 2026-10-08 | |
| LMArena Korean | 1501 | #3 of 213, top 2% | LMArena | 2026-10-08 | |
| LMArena Russian | 1490 | #13 of 283, top 5% | LMArena | 2026-10-08 | |
| LMArena Spanish | 1490 | #11 of 226, top 5% | LMArena | 2026-10-08 |
Instruction Following
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Instruction Following | 1477 | #19 of 298, top 7% | LMArena | 2026-10-08 |
Long Context
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Longer Query | 1488 | #14 of 291, top 5% | LMArena | 2026-10-08 |
Writing & Preference
| Benchmark | Score | Position | Setting | Source | Date |
|---|---|---|---|---|---|
| LMArena Text | 1490 | #10 of 297, top 4% | LMArena | 2026-10-08 | |
| LMArena Creative Writing | 1455 | #23 of 295, top 8% | LMArena | 2026-10-08 | |
| EQ-Bench Creative Writing | 1906 | #14 of 115, top 13% | EQ-Bench | ||
| LMArena Multi-Turn | 1482 | #16 of 295, top 6% | LMArena | 2026-10-08 |
API pricing by provider
| Route | Input $/M | Output $/M | Cached input $/M | Checked |
|---|---|---|---|---|
| meta | $1.25 | $4.25 | $0.15 | 2026-10-10 |
| openrouter | $1.25 | $4.25 | $0.15 | 2026-10-10 |
Compare Muse Spark 1.3
- Muse Spark 1.3 vs Muse Spark 1.2
- Muse Spark 1.3 vs GLM-5.3
- Muse Spark 1.3 vs Gemini 3 Pro
- Muse Spark 1.3 vs Grok 4.5
- Muse Spark 1.3 vs Claude Sonnet 5
- Muse Spark 1.3 vs Gemini 4 Argon
- Muse Spark 1.3 vs GPT-5.6 Luna
- Muse Spark 1.3 vs GPT-6 Astra
- Muse Spark 1.3 vs Claude Fable 5.1
- Muse Spark 1.3 vs Gemini 3.8 Flash
- Muse Spark 1.3 vs Kimi K3
- Muse Spark 1.3 vs Grok 4.6
- Muse Spark 1.3 vs Qwen3.8 Max
- Muse Spark 1.3 vs DeepSeek V4 Pro
Other Meta models
- Muse Spark50.6
- Muse Spark 1.250.3
- Muse Spark 1.149.9
- Muse Glimmer41.7
- Codellama 70b Instruct33.7
- Llama 4 Maverick30.9
- Codellama 34b Instruct30.8
- Llama 3.1-405B30.7
Frequently asked questions
How good is Muse Spark 1.3?
Muse Spark 1.3 by Meta ranks 27th of 354 ranked models on the Noometry Index as of October 2026, with a score of 54.8. Its strongest category is multilingual, where it ranks 8th. API pricing starts at $1.25 per million input tokens and $4.25 per million output tokens, with a 1.05M-token context window.
How much does Muse Spark 1.3 cost?
Muse Spark 1.3 costs $1.25 per million input tokens and $4.25 per million output tokens on Meta's own API, with cached input at $0.15.
What is Muse Spark 1.3's context window?
Muse Spark 1.3 accepts up to 1.05M tokens of input and can write up to 131K tokens in one response.
Is Muse Spark 1.3 open source?
No. Muse Spark 1.3 is proprietary and available only through Meta's API and partner platforms.
What are Muse Spark 1.3's strengths and weaknesses?
Relative to other ranked models, Muse Spark 1.3 places best in multilingual, writing & preference, coding and lowest in knowledge, agentic & tool use, multimodal.
What is Muse Spark 1.3 best at?
Its best category is multilingual, where it ranks 8th on Noometry.