Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation Model

Muse Spark 1.1

Meta's multimodal model for agentic and coding work, supporting text, image, video, and PDF input with a 1M+ token context window.

PublisherMeta
TypeText
Context Window1,048,576 tokens
Input$1.25/MTok
Output$4.25/MTok
LATESTLARGE CONTEXT

Muse Spark 1.1

**Muse Spark 1.1** is Meta's multimodal model designed for agentic and coding workloads. It accepts **text, images, video, and PDF** as input and generates text output, making it versatile for a wide range of real-world applications that require understanding across multiple content types. ### Key Capabilities - **Agentic tool calling**: Optimized for multi-step tool loops and workflows where the model needs to plan, call tools, evaluate results, and iterate - **Coding & technical assistance**: Strong performance on software engineering tasks, code generation, and technical problem-solving - **Multimodal understanding**: Native support for image understanding, video understanding, and PDF document processing - **Structured output**: Reliable generation of well-formed JSON and other structured formats for downstream automation - **Search grounding**: Can be grounded with search results for up-to-date, factual responses - **Long-context reasoning**: With a **1,048,576 token context window**, it can process massive amounts of information in a single request ### Best Use Cases Muse Spark 1.1 is ideal for building coding assistants, agentic sub-agent loops, document analysis pipelines, real-time chat experiences, and high-frequency workflows where both speed and multimodal capability matter.

Ready to build with Muse Spark 1.1?

Get Started Free

Parameters & options

Max Temperature1
Max Response Size128,000 tokens

Explore similar models

Start building with Muse Spark 1.1

No API keys required. Create AI-powered workflows with Muse Spark 1.1 in minutes — free.