Models, measured side by side
Muse Spark 1.3 vs Qwen3.8 2.4T A95B
Compare Muse Spark 1.3, Qwen3.8 2.4T A95B by token pricing, context length, benchmark results, speed and tool support.
Monthly workload
Monthly cost breakdown
InputOutput
Muse Spark 1.3
$46 / mo
Qwen3.8 2.4T A95B
$70 / mo
Performance comparison
Muse Spark 1.3
48.1
Qwen3.8 2.4T A95B
39.9
Overview
Pricing & capacity
| Pricing & capacity | Muse Spark 1.3 | Qwen3.8 2.4T A95B |
|---|---|---|
| Input / 1M tokens | $1.25 | $2.00 |
| Output / 1M tokens | $4.25 | $6.00 |
| Cached input / 1M | $0.15 | $0.25 |
| Monthly cost | $46 | $70 |
| Context window | 1M | 1M |
| Maximum output | 944K | 131K |
Input / 1M tokens
- Muse Spark 1.3
- $1.25
- Qwen3.8 2.4T A95B
- $2.00
Output / 1M tokens
- Muse Spark 1.3
- $4.25
- Qwen3.8 2.4T A95B
- $6.00
Cached input / 1M
- Muse Spark 1.3
- $0.15
- Qwen3.8 2.4T A95B
- $0.25
Monthly cost
- Muse Spark 1.3
- $46
- Qwen3.8 2.4T A95B
- $70
Context window
- Muse Spark 1.3
- 1M
- Qwen3.8 2.4T A95B
- 1M
Maximum output
- Muse Spark 1.3
- 944K
- Qwen3.8 2.4T A95B
- 131K
Benchmarks & performance
| Benchmarks & performance | Muse Spark 1.3 | Qwen3.8 2.4T A95B |
|---|---|---|
| Intelligence Index | 48.1 | 39.9 |
| GPQA | 93.5% | 93.5% |
| Humanity’s Last Exam | 48.7% | 42.4% |
| SciCode | 58.8% | 54.1% |
| Output speed | 146 tokens/s | 38 tokens/s |
| Time to first token | 47.86 s | 2.75 s |
| DesignArena rating | 1,358 | Not listed |
| DesignArena win rate | 57.3% | Not listed |
Intelligence Index
- Muse Spark 1.3
- 48.1
- Qwen3.8 2.4T A95B
- 39.9
GPQA
- Muse Spark 1.3
- 93.5%
- Qwen3.8 2.4T A95B
- 93.5%
Humanity’s Last Exam
- Muse Spark 1.3
- 48.7%
- Qwen3.8 2.4T A95B
- 42.4%
SciCode
- Muse Spark 1.3
- 58.8%
- Qwen3.8 2.4T A95B
- 54.1%
Output speed
- Muse Spark 1.3
- 146 tokens/s
- Qwen3.8 2.4T A95B
- 38 tokens/s
Time to first token
- Muse Spark 1.3
- 47.86 s
- Qwen3.8 2.4T A95B
- 2.75 s
DesignArena rating
- Muse Spark 1.3
- 1,358
- Qwen3.8 2.4T A95B
- Not listed
DesignArena win rate
- Muse Spark 1.3
- 57.3%
- Qwen3.8 2.4T A95B
- Not listed
Tools & features
| Tools & features | Muse Spark 1.3 | Qwen3.8 2.4T A95B |
|---|---|---|
| Function calling | Supported | Supported |
| Structured outputs | Supported | Supported |
| JSON mode | Supported | Supported |
| Reasoning | Supported | Supported |
| Log probabilities | Not listed | Supported |
| Deterministic seed | Not listed | Supported |
| Prompt caching | Supported | Supported |
Function calling
- Muse Spark 1.3
- Supported
- Qwen3.8 2.4T A95B
- Supported
Structured outputs
- Muse Spark 1.3
- Supported
- Qwen3.8 2.4T A95B
- Supported
JSON mode
- Muse Spark 1.3
- Supported
- Qwen3.8 2.4T A95B
- Supported
Reasoning
- Muse Spark 1.3
- Supported
- Qwen3.8 2.4T A95B
- Supported
Log probabilities
- Muse Spark 1.3
- Not listed
- Qwen3.8 2.4T A95B
- Supported
Deterministic seed
- Muse Spark 1.3
- Not listed
- Qwen3.8 2.4T A95B
- Supported
Prompt caching
- Muse Spark 1.3
- Supported
- Qwen3.8 2.4T A95B
- Supported
API & availability
| API & availability | Muse Spark 1.3 | Qwen3.8 2.4T A95B |
|---|---|---|
| API identifier | muse-spark-1.3 | qwen3.8-2.4t-a95b |
| Prices checked | Oct 8, 2026 | Oct 8, 2026 |
API identifier
- Muse Spark 1.3
muse-spark-1.3- Qwen3.8 2.4T A95B
qwen3.8-2.4t-a95b
Prices checked
- Muse Spark 1.3
- Oct 8, 2026
- Qwen3.8 2.4T A95B
- Oct 8, 2026
About the models
Muse Spark 1.3
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...
Full pricing & detailsQwen3.8 2.4T A95B
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Full pricing & detailsComparison FAQ
Muse Spark 1.3: $1.25 input and $4.25 output per million tokens. Qwen3.8 2.4T A95B: $2.00 input and $6.00 output per million tokens. The cheaper choice depends on your input-to-output ratio; a model can have a lower input rate but a higher output rate.
