Search Intelligence Score
66.1
+35.8 lift from search
050100
- Lift from search
- +35.8
- 30.3 without search
- Cost per 1K tasks
- $142
- #5 in Search Efficiency
- Time per task
- 274s
- Search plus inference time to complete one task
Accuracy by benchmark
Accuracy on each eval, with search (solid) and without (light).
- DSQA85.2 vs 48.9
- HLE46.0 vs 30.0
- WISER67.0 vs 12.0
Cost per 1K tasks
How the cost splits between model inference, search calls, and page extraction.
- Inference
- $129
- 91%
- Search
- $4.55
- 3%
- Extract
- $8.38
- 6%
Compared with
Other models from the same provider alongside models that score closest.
- Muse Spark 1.366.1Cost $142 · Lift +35.8 · 274s
- Sonnet 566.6Cost $688 · Lift +43.8 · 846s
- GPT-6 Sol66.6Cost $183 · Lift +28.7 · 260s
- Gemini 3.766.8Cost $130 · Lift +27.4 · 229s
- Gemini 3.865.1Cost $177 · Lift +26.1 · 317s
| Model | Score | Lift | Cost per 1K tasks | Time per task |
|---|---|---|---|---|
| Muse Spark 1.3 | 66.1 | +35.8 | $142 | 274s |
| Sonnet 5Claude Sonnet 5 | 66.6 | +43.8 | $688 | 846s |
| GPT-6 SolGPT-6 Sol | 66.6 | +28.7 | $183 | 260s |
| Gemini 3.7Gemini 3.7 Flash | 66.8 | +27.4 | $130 | 229s |
| Gemini 3.8Gemini 3.8 Flash | 65.1 | +26.1 | $177 | 317s |
Search Intelligence Score averages accuracy on DSQA, HLE, and WISER with equal weight; cost and time are averaged the same way, with cost shown per 1,000 tasks. Latest update September 25, 2026. Full methodology.
# Muse Spark 1.3 · Search Capability Leaderboard
- Search Intelligence Score: 66.1 (#9 in Search Intelligence)
- Without search: 30.3 · Lift from search +35.8
- Cost per 1K tasks: $142 (#5 in Search Efficiency)
- Time per task: 274s
- Lab: Meta · Model id: meta/muse-spark-1.3
## Accuracy by benchmark
| Suite | With search | Without search |
|---|---|---|
| DSQA | 85.2 | 48.9 |
| HLE | 46.0 | 30.0 |
| WISER | 67.0 | 12.0 |
## Compared with
| Model | Score | Lift | Cost per 1K tasks | Time per task |
|---|---|---|---|---|
| Muse Spark 1.3 | 66.1 | +35.8 | $142 | 274s |
| Claude Sonnet 5 | 66.6 | +43.8 | $688 | 846s |
| GPT-6 Sol | 66.6 | +28.7 | $183 | 260s |
| Gemini 3.7 Flash | 66.8 | +27.4 | $130 | 229s |
| Gemini 3.8 Flash | 65.1 | +26.1 | $177 | 317s |
Latest update September 25, 2026. Methodology: /leaderboard#methodology