Muse Spark 1.1
Meta's multimodal model for agentic and coding work, supporting text, image, video, and PDF input with a 1M+ token context window.
Muse Spark 1.1
**Muse Spark 1.1** is Meta's multimodal model designed for agentic and coding workloads. It accepts **text, images, video, and PDF** as input and generates text output, making it versatile for a wide range of real-world applications that require understanding across multiple content types. ### Key Capabilities - **Agentic tool calling**: Optimized for multi-step tool loops and workflows where the model needs to plan, call tools, evaluate results, and iterate - **Coding & technical assistance**: Strong performance on software engineering tasks, code generation, and technical problem-solving - **Multimodal understanding**: Native support for image understanding, video understanding, and PDF document processing - **Structured output**: Reliable generation of well-formed JSON and other structured formats for downstream automation - **Search grounding**: Can be grounded with search results for up-to-date, factual responses - **Long-context reasoning**: With a **1,048,576 token context window**, it can process massive amounts of information in a single request ### Best Use Cases Muse Spark 1.1 is ideal for building coding assistants, agentic sub-agent loops, document analysis pipelines, real-time chat experiences, and high-frequency workflows where both speed and multimodal capability matter.
Ready to build with Muse Spark 1.1?
Get Started FreeDocumentation & links
Parameters & options
Explore similar models
Start building with Muse Spark 1.1
No API keys required. Create AI-powered workflows with Muse Spark 1.1 in minutes — free.