Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation Model

GLM 5.1

GLM 5.1 is a text generation model from Z.ai with a 200,000-token context window and configurable reasoning effort.

PublisherZ.ai
TypeText
Context Window200,000 tokens
ReleasedMarch 2026
Input$1.05/MTok
Output$3.50/MTok
ProviderDeepInfra

Long-context text generation with adjustable reasoning

GLM 5.1 is a large language model developed by Z.ai (formerly Zhipu AI) and made available through DeepInfra. It belongs to the GLM (General Language Model) family and supports a 200,000-token context window, making it suited for tasks that require processing long documents, extended conversations, or large codebases in a single pass. The model accepts a reasoning effort toggle as a configurable input, allowing users to adjust how much deliberation the model applies before generating a response.

GLM 5.1 is designed for text generation tasks including summarization, question answering, instruction following, and multi-turn dialogue. With a maximum response size of 16,384 tokens, it can produce detailed, long-form outputs. The model was released in March 2026 and is accessible on MindStudio without requiring users to manage separate API keys or provider accounts.

What GLM 5.1 supports

Long Context Window

Processes up to 200,000 tokens in a single request, enabling analysis of lengthy documents, codebases, or extended conversation histories without truncation.

Adjustable Reasoning

Exposes a reasoning effort toggle that lets users control how much deliberative processing the model applies before generating a response.

Long-Form Output

Supports responses of up to 16,384 tokens, making it suitable for generating detailed reports, summaries, or multi-section documents in one call.

Instruction Following

Handles multi-turn chat and single-turn instruction tasks, including summarization, question answering, and structured text generation.

Code Generation

Generates and explains code across common programming languages, consistent with the capabilities documented for the GLM model family.

Ready to build with GLM 5.1?

Get Started Free

Common questions about GLM 5.1

What is the context window size for GLM 5.1?

GLM 5.1 supports a context window of 200,000 tokens, meaning it can process up to that many tokens of combined input and conversation history in a single request.

What is the maximum response length GLM 5.1 can produce?

The model can generate responses of up to 16,384 tokens per call, which is suitable for long-form content such as detailed reports or extended code files.

What does the reasoning effort toggle do?

The reasoning effort input is a configurable toggle that adjusts how much deliberative processing the model performs before producing a response. Higher effort may improve accuracy on complex tasks at the cost of additional latency.

Who developed GLM 5.1 and when was it released?

GLM 5.1 was developed by Z.ai (published under the identifier zai-org) and released in March 2026. It is served through DeepInfra and available on MindStudio.

Is pricing information available for GLM 5.1 on MindStudio?

Published pricing for GLM 5.1 is not listed in the current metadata. You can check MindStudio's pricing page or the model's detail page for the most up-to-date cost information.

Does GLM 5.1 support image or video inputs?

Based on the available metadata, GLM 5.1 does not have confirmed support for image or video analysis inputs. It is classified as a text generation model.

Parameters & options

Max Temperature1
Max Response Size16,384 tokens
Reasoning EffortToggle Group
Default: medium
NoneLowMediumHigh

Start building with GLM 5.1

No API keys required. Create AI-powered workflows with GLM 5.1 in minutes — free.