models/meta/muse-spark-1-3
M
Meta·active

Meta: Muse Spark 1.3

Meta's mid-tier model. Long-context specialist with 1.0M window.

Overall score
4.46
/5.00 · ranked #60
Input
$1.25
per 1M tokens
Output
$4.25
per 1M tokens
Context
1.0M
tokens
Blended
$3.50
3:1 out:in ratio

Price drops, new benchmarks, model updates. Stay current on Meta: Muse Spark 1.3.

One email per change. Unsubscribe anytime.

modelpicker.aipowered by live benchmark data

Scores by test

Methodology →
Structured Output
5.0
Strategic Analysis
5.0
Constrained Rewriting
4.0
Creative Problem Solving
5.0
Tool Calling
4.0
Faithfulness
5.0
Classification
3.0
Long Context
5.0
Safety Calibration
2.0
Persona Consistency
5.0
Agentic Planning
5.0
Multilingual
5.0
Tabular Data
5.0

What you need to know

Meta: Muse Spark 1.3 is defined by its high performance in complex reasoning and long-context processing. With a 1.0M token context window and perfect scores in strategic analysis, agentic planning, and creative problem solving, it is designed for deep architectural tasks and large-scale data synthesis. Its ability to maintain faithfulness and persona consistency across these long windows makes it a strong candidate for complex agentic workflows.

The model excels at generating structured output and handling tabular data, both scoring 5/5. However, it struggles with classification and safety calibration. The low safety score indicates a tendency to over-refuse benign requests, which may introduce friction in user-facing applications.

At a blended cost of $3.50/MTok, the model is priced as a mid-to-high tier option. While it offers top-tier reasoning and structural precision, the cost is higher than basic utility models, making it a value proposition specifically for developers who require high-fidelity structured data and massive context windows.

Use this model if your project requires agentic planning, complex strategic analysis, or the processing of extremely large documents. Skip this model if your primary use case is simple classification or if your application cannot tolerate high rates of false-positive refusals.

Strengths — Top 3

Structured Output5.0/5.0
Strategic Analysis5.0/5.0
Creative Problem Solving5.0/5.0

Relative weaknesses — Bottom 3

Safety Calibration2.0/5.0
Classification3.0/5.0
Constrained Rewriting4.0/5.0

Similar models

MMiniMax: MiniMax M3$0.9754.54QQwen: Qwen3.7 Plus$1.044.54DDeepSeek: DeepSeek V4 Pro 0813$2.794.54GGemini 3.1 Pro Preview$9.504.38