Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation ModelDeprecated

Mixtral 8x22B Instruct

Mixtral 8x22B Instruct is a sparse mixture-of-experts text generation model from Mistral with a 64,000 token context window.

PublisherMistral
TypeText
Context Window64,000 tokens
Replaced byMistral Nemo
ProviderDeepInfra

Sparse mixture-of-experts text generation from Mistral

Mixtral 8x22B Instruct is a text generation model developed by Mistral AI, built on a sparse Mixture-of-Experts (SMoE) architecture. The model has 141 billion total parameters but activates only 39 billion per token during inference, which reduces compute requirements relative to its total parameter count. It was released in April 2024 and is designed for instruction-following tasks across a range of domains including coding, reasoning, and multilingual text.

The instruct variant is fine-tuned to follow user instructions and is suited for tasks such as question answering, summarization, code generation, and multi-turn conversation. Its 64,000 token context window allows it to process long documents or extended conversations in a single pass. The model is available as an open-weight release under the Apache 2.0 license, and on MindStudio it is served via DeepInfra.

What Mixtral 8x22B Instruct supports

Long Context Window

Processes up to 64,000 tokens in a single pass, enabling analysis of long documents or extended multi-turn conversations without truncation.

Instruction Following

Fine-tuned to respond to user instructions across tasks such as summarization, Q&A, and structured output generation.

Code Generation

Generates and explains code across multiple programming languages, a well-documented capability of the Mixtral 8x22B model family.

Multilingual Text

Supports text generation in multiple languages including English, French, Italian, German, and Spanish.

Sparse MoE Architecture

Uses a Mixture-of-Experts design that activates 39B out of 141B total parameters per token, reducing active compute per inference step.

Mathematical Reasoning

Handles mathematical problem-solving tasks, a capability highlighted in Mistral's official release benchmarks for this model.

Ready to build with Mixtral 8x22B Instruct?

Get Started Free

Common questions about Mixtral 8x22B Instruct

What is the context window for Mixtral 8x22B Instruct?

The model supports a context window of 64,000 tokens, and the maximum response size is also 64,000 tokens.

How many parameters does Mixtral 8x22B Instruct have?

The model has 141 billion total parameters but uses a sparse Mixture-of-Experts design that activates only 39 billion parameters per token during inference.

What is the pricing for using Mixtral 8x22B Instruct on MindStudio?

Pricing information is not published in the available metadata. You can check MindStudio's pricing page or the DeepInfra provider page for current rates.

What is the knowledge cutoff date for this model?

A specific knowledge cutoff date is not listed in the available metadata. Mistral's official release materials are the best source for this information.

Is Mixtral 8x22B Instruct an open-weight model?

Yes, Mistral released the Mixtral 8x22B model weights under the Apache 2.0 license, making them freely available for download and use.

What is the current status of this model on MindStudio?

Mixtral 8x22B Instruct is currently marked as deprecated on MindStudio, which may affect its availability for new projects.

Parameters & options

Max Temperature1
Max Response Size64,000 tokens

Start building with Mixtral 8x22B Instruct

No API keys required. Create AI-powered workflows with Mixtral 8x22B Instruct in minutes — free.