Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation Model

Gemini 3.1 Pro

Gemini 3.1 Pro is a multimodal text generation model from Google with a 1,048,576-token context window.

PublisherGoogle
TypeText
Context Window1,048,576 tokens
ReleasedFebruary 2026
Input$2.00/MTok
Output$12.00/MTok
LATESTLARGE CONTEXTREASONINGMULTI-MODALTOOLS

Large-context multimodal reasoning with tool use

Gemini 3.1 Pro is a large language model developed by Google, released in February 2026 under the full model identifier gemini-3.1-pro-preview. It supports text generation with multimodal inputs, accepting both images and video alongside text, and offers a context window of 1,048,576 tokens — allowing it to process very long documents or extended conversations in a single request. The model also exposes a configurable thinking budget, which lets developers control how much internal reasoning the model applies before producing a response.

Gemini 3.1 Pro is designed for tasks that benefit from deep reasoning, long-context comprehension, and tool integration, including document analysis, multi-step problem solving, and agentic workflows. It supports external tool calls natively and includes a priority mode setting that allows developers to tune the balance between response speed and output quality. With a maximum response size of 65,536 tokens, it is suited for use cases that require detailed, long-form outputs.

What Gemini 3.1 Pro supports

Large Context Window

Processes up to 1,048,576 tokens in a single request, enabling analysis of very long documents, codebases, or conversation histories without truncation.

Multimodal Input

Accepts text, images, and video as inputs within the same request, supporting tasks that require understanding across multiple content types.

Configurable Reasoning

Exposes a Thinking Budget input that lets developers set how much internal reasoning the model performs before generating a response, with a numeric limit for fine-grained control.

Tool Use

Supports native tool calling, allowing the model to invoke external functions or APIs as part of a response in agentic and workflow-driven applications.

Priority Mode

Includes a Priority Mode selector that lets developers adjust the trade-off between response latency and output quality depending on application requirements.

Long-Form Output

Generates responses of up to 65,536 tokens, making it suitable for producing detailed reports, long-form summaries, or extended code outputs.

Video Analysis

Analyzes video content passed as input, enabling tasks such as scene understanding, content summarization, and temporal reasoning across frames.

Ready to build with Gemini 3.1 Pro?

Get Started Free

Benchmark scores

Scores represent accuracy — the percentage of questions answered correctly on each test.

BenchmarkWhat it testsScore
GPQA DiamondPhD-level science questions (biology, physics, chemistry)94.1%
HLEQuestions that challenge frontier models across many domains44.7%
SciCodeScientific research coding and numerical methods58.9%
ARC-AGI-2Novel abstract reasoning and pattern recognition77.1%
SWE-bench VerifiedReal GitHub issues requiring multi-file code fixes80.6%
SWE-bench ProChallenging real-world software engineering tasks54.2%
Terminal-Bench 2.0Agentic coding and terminal command tasks68.5%
τ²-bench RetailAgentic tool use in retail scenarios90.8%
τ²-bench TelecomAgentic tool use in telecom scenarios99.3%
MCP-Atlas Tool UseStructured tool use via Model Context Protocol69.2%
BrowseCompComplex web browsing and information retrieval85.9%
MMMLUMultilingual and multimodal understanding92.6%

Common questions about Gemini 3.1 Pro

What is the context window size for Gemini 3.1 Pro?

Gemini 3.1 Pro has a context window of 1,048,576 tokens, which allows it to process very long inputs — such as large documents or extended conversations — in a single request.

What types of inputs does Gemini 3.1 Pro support?

The model supports text, images, and video as inputs, making it a multimodal model capable of handling tasks that involve more than one content type in the same request.

What is the Thinking Budget input used for?

The Thinking Budget is a configurable input that controls how much internal reasoning the model performs before generating a response. A numeric Thinking Budget Limit can also be set to cap the amount of reasoning applied.

What is the maximum response length for Gemini 3.1 Pro?

The model can generate responses of up to 65,536 tokens, which supports detailed, long-form outputs such as extended reports, summaries, or code.

When was Gemini 3.1 Pro released?

Gemini 3.1 Pro was released in February 2026 and is available on MindStudio under the full model identifier gemini-3.1-pro-preview.

What people think about Gemini 3.1 Pro

Community reception on r/singularity was largely positive at launch, with the benchmark announcement thread accumulating over 2,300 upvotes and 528 comments. Users frequently highlighted the ARC-AGI-2 score and the 1M-token context window as notable technical achievements.

Some community members raised questions about hallucination rates, with a dedicated thread asking whether Google had addressed accuracy issues seen in prior Gemini versions. Practical use cases discussed included coding assistance, long-document analysis, and agentic workflows.

View more discussions →

Parameters & options

Max Temperature2
Max Response Size65,536 tokens
Thinking BudgetSelect
Default: auto
ManualAuto
Thinking Budget LimitNumber

Must be less than Max Response Size

Range: 1–24576
ToolsTools
Google SearchGround responses in current events and facts from the web to reduce hallucinations.
Google MapsBuild location-aware assistants that can find places, get directions, and provide rich local context.
Code ExecutionAllow the model to write and run Python code to solve math problems or process data accurately.
URL ContextDirect the model to read and analyze content from specific web pages or documents.
Priority ModeSelect
Default: false
StandardPriority (1.8x cost, higher reliability)

Start building with Gemini 3.1 Pro

No API keys required. Create AI-powered workflows with Gemini 3.1 Pro in minutes — free.