Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation Model

Gemma 4 26B

Gemma 4 26B is a multimodal text generation model from Google with a 262,144 token context window.

PublisherGoogle
TypeText
Context Window262,144 tokens
ReleasedApril 2026
Input$0.07/MTok
Output$0.34/MTok
ProviderDeepInfra
MULTI-MODALCOST EFFECTIVE

Multimodal text generation from Google

Gemma 4 26B is an instruction-tuned language model developed by Google, released in April 2026 under the full model name google/gemma-4-26B-A4B-it. It is a mixture-of-experts architecture with 26 billion total parameters and 4 billion active parameters per forward pass, which allows it to handle a broad range of tasks while keeping inference costs relatively low. The model supports both text and image inputs, making it capable of multimodal reasoning across a single 262,144 token context window.

Gemma 4 26B is well suited for tasks that require long-context understanding, such as document analysis, extended conversations, and vision-language tasks involving images alongside text. Its active-parameter design means it can be served efficiently compared to dense models of similar total parameter counts, which contributes to its cost-effective positioning. The model is available through DeepInfra on MindStudio without requiring separate API key management.

What Gemma 4 26B supports

Long Context Window

Processes up to 262,144 tokens in a single context, enabling analysis of long documents, codebases, or extended conversations without truncation.

Image Understanding

Accepts image inputs alongside text, supporting tasks like visual question answering, image description, and multimodal reasoning.

Text Generation

Generates coherent, instruction-following text responses with a maximum output size of 16,384 tokens per response.

Mixture-of-Experts

Uses a sparse mixture-of-experts architecture with 26B total parameters but only 4B active per forward pass, reducing compute cost per inference.

Cost-Effective Inference

Designed for efficient deployment due to its active-parameter sparsity, making it accessible for high-volume or budget-conscious workloads.

Ready to build with Gemma 4 26B?

Get Started Free

Common questions about Gemma 4 26B

What is the context window size for Gemma 4 26B?

Gemma 4 26B supports a context window of 262,144 tokens, allowing it to process very long documents or conversations in a single pass.

Does Gemma 4 26B support image inputs?

Yes. The model accepts image inputs in addition to text, making it a multimodal model capable of vision-language tasks.

What is the maximum response size?

The model can generate up to 16,384 tokens per response.

Who developed Gemma 4 26B and when was it released?

Gemma 4 26B was developed by Google and released in April 2026. It is served on MindStudio via DeepInfra.

What does the '4B active parameters' mean for this model?

Gemma 4 26B uses a mixture-of-experts architecture with 26 billion total parameters, but only 4 billion parameters are active during each forward pass. This reduces the compute required per inference compared to a fully dense 26B model.

Is pricing information available for Gemma 4 26B on MindStudio?

Published pricing details are not listed in the current metadata. You can check MindStudio's pricing page or the DeepInfra provider page for current rate information.

Parameters & options

Max Temperature1
Max Response Size16,384 tokens

Start building with Gemma 4 26B

No API keys required. Create AI-powered workflows with Gemma 4 26B in minutes — free.