Muse Spark 1.1
by Metatext+image+file+audio+video->text1 endpoint
Neura Intelligence Index
83.5/ 100
Rank
#27—
Confidence
mediumBenchmarks
5(13% coverage)
Overview
Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
Capabilities
Text generation
Image understanding
Audio processing
Video understanding
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
File input
Top-P sampling
Modalities
Input
TextImageVideoFileAudio
Output
Text
Technical Specifications
Context Window
1.0M tokens
Max Output
943.7K tokens
Tokenizer
Other
Content Moderation
Enabled
Uptime (24h)
10000.0%
Country
🇺🇸US
Release Date
July 9, 20262mo
Supported Parameters
include_reasoningmax_tokensreasoningreasoning_effortrepetition_penaltyresponse_formatstructured_outputstemperaturetool_choicetoolstop_ktop_p
Benchmark Performance
coding
| Benchmark | Score | Source |
|---|---|---|
| SWE-Bench Pro | 61.5 | verified |
reasoning
| Benchmark | Score | Source |
|---|---|---|
| Humanity's Last Exam | 62.1 | verified |
multimodal
| Benchmark | Score | Source |
|---|---|---|
| CharXiv Reasoning | 88.4 | verified |
agent
| Benchmark | Score | Source |
|---|---|---|
| MCP Atlas | 88.1 | verified |
| Toolathlon | 75.6 | verified |
Performance Over Time
Pricing
Per 1M tokens
Input$1.25
Output$4.25
Cache read$0.15
Added
July 16, 2026
Last synced 9/21/2026