Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation Model

Qwen3.8 2.4T

Qwen3.8 2.4T is a mixture-of-experts text generation model from Qwen with a 262,144 token context window.

PublisherQwen
TypeText
Context Window262,144 tokens
ReleasedAugust 2026
Input$2.00/MTok
Output$6.00/MTok
ProviderDeepInfra
FLAGSHIPREASONING

Large-scale mixture-of-experts reasoning model

Qwen3.8 2.4T is a mixture-of-experts (MoE) language model developed by Qwen, with a total of 2.4 trillion parameters and 95 billion active parameters per forward pass. It is designed for text generation tasks and supports a context window of 262,144 tokens, making it suitable for processing very long documents and extended conversations within a single session.

The model carries a REASONING tag, indicating it is built to handle tasks that require multi-step logical inference, such as complex question answering, mathematical reasoning, and structured problem solving. Its MoE architecture allows a large total parameter count while keeping the number of active parameters per inference lower than a dense model of equivalent scale, which affects how the model is deployed and served.

What Qwen3.8 2.4T supports

Extended Context Window

Supports a context window of up to 262,144 tokens, enabling processing of very long documents or multi-turn conversations in a single pass.

Multi-Step Reasoning

Tagged as a reasoning model, it is designed for tasks requiring logical inference, such as math problems, structured analysis, and complex question answering.

Mixture-of-Experts Architecture

Uses a sparse MoE design with 2.4 trillion total parameters and approximately 95 billion active parameters per forward pass, distributing computation across expert subnetworks.

Long-Form Text Generation

Generates extended text outputs with a maximum response size matching the full 262,144 token context window.

Flagship-Scale Performance

Designated as a flagship model by Qwen, reflecting its position as one of the largest models in the Qwen3 series by total parameter count.

Ready to build with Qwen3.8 2.4T?

Get Started Free

Common questions about Qwen3.8 2.4T

What is the context window size for Qwen3.8 2.4T?

Qwen3.8 2.4T supports a context window of 262,144 tokens, and the maximum response size is also 262,144 tokens.

How many parameters does Qwen3.8 2.4T have?

The model has approximately 2.4 trillion total parameters. As a mixture-of-experts model, around 95 billion parameters are active during each forward pass.

What types of tasks is Qwen3.8 2.4T best suited for?

Based on its REASONING and FLAGSHIP tags, the model is designed for tasks requiring multi-step logical inference, complex question answering, mathematical reasoning, and processing of long documents.

Who developed Qwen3.8 2.4T and when was it released?

Qwen3.8 2.4T was developed by Qwen and has a listed release date of August 2026. It is served on MindStudio via the DeepInfra provider.

Does Qwen3.8 2.4T support image or video inputs?

Based on the available metadata, image and video analysis support is not confirmed for this model. It is listed as a text generation model with no declared multimodal input types.

Parameters & options

Max Temperature1
Max Response Size262,144 tokens

Start building with Qwen3.8 2.4T

No API keys required. Create AI-powered workflows with Qwen3.8 2.4T in minutes — free.