Gemini 3.1 Pro
Gemini 3.1 Pro is a multimodal text generation model from Google with a 1,048,576-token context window.
Large-context multimodal reasoning with tool use
Gemini 3.1 Pro is a large language model developed by Google, released in February 2026 under the full model identifier gemini-3.1-pro-preview. It supports text generation with multimodal inputs, accepting both images and video alongside text, and offers a context window of 1,048,576 tokens — allowing it to process very long documents or extended conversations in a single request. The model also exposes a configurable thinking budget, which lets developers control how much internal reasoning the model applies before producing a response.
Gemini 3.1 Pro is designed for tasks that benefit from deep reasoning, long-context comprehension, and tool integration, including document analysis, multi-step problem solving, and agentic workflows. It supports external tool calls natively and includes a priority mode setting that allows developers to tune the balance between response speed and output quality. With a maximum response size of 65,536 tokens, it is suited for use cases that require detailed, long-form outputs.
What Gemini 3.1 Pro supports
Large Context Window
Processes up to 1,048,576 tokens in a single request, enabling analysis of very long documents, codebases, or conversation histories without truncation.
Multimodal Input
Accepts text, images, and video as inputs within the same request, supporting tasks that require understanding across multiple content types.
Configurable Reasoning
Exposes a Thinking Budget input that lets developers set how much internal reasoning the model performs before generating a response, with a numeric limit for fine-grained control.
Tool Use
Supports native tool calling, allowing the model to invoke external functions or APIs as part of a response in agentic and workflow-driven applications.
Priority Mode
Includes a Priority Mode selector that lets developers adjust the trade-off between response latency and output quality depending on application requirements.
Long-Form Output
Generates responses of up to 65,536 tokens, making it suitable for producing detailed reports, long-form summaries, or extended code outputs.
Video Analysis
Analyzes video content passed as input, enabling tasks such as scene understanding, content summarization, and temporal reasoning across frames.
Ready to build with Gemini 3.1 Pro?
Get Started FreeBenchmark scores
Scores represent accuracy — the percentage of questions answered correctly on each test.
| Benchmark | What it tests | Score |
|---|---|---|
| GPQA Diamond | PhD-level science questions (biology, physics, chemistry) | 94.1% |
| HLE | Questions that challenge frontier models across many domains | 44.7% |
| SciCode | Scientific research coding and numerical methods | 58.9% |
| ARC-AGI-2 | Novel abstract reasoning and pattern recognition | 77.1% |
| SWE-bench Verified | Real GitHub issues requiring multi-file code fixes | 80.6% |
| SWE-bench Pro | Challenging real-world software engineering tasks | 54.2% |
| Terminal-Bench 2.0 | Agentic coding and terminal command tasks | 68.5% |
| τ²-bench Retail | Agentic tool use in retail scenarios | 90.8% |
| τ²-bench Telecom | Agentic tool use in telecom scenarios | 99.3% |
| MCP-Atlas Tool Use | Structured tool use via Model Context Protocol | 69.2% |
| BrowseComp | Complex web browsing and information retrieval | 85.9% |
| MMMLU | Multilingual and multimodal understanding | 92.6% |
Common questions about Gemini 3.1 Pro
What is the context window size for Gemini 3.1 Pro?
Gemini 3.1 Pro has a context window of 1,048,576 tokens, which allows it to process very long inputs — such as large documents or extended conversations — in a single request.
What types of inputs does Gemini 3.1 Pro support?
The model supports text, images, and video as inputs, making it a multimodal model capable of handling tasks that involve more than one content type in the same request.
What is the Thinking Budget input used for?
The Thinking Budget is a configurable input that controls how much internal reasoning the model performs before generating a response. A numeric Thinking Budget Limit can also be set to cap the amount of reasoning applied.
What is the maximum response length for Gemini 3.1 Pro?
The model can generate responses of up to 65,536 tokens, which supports detailed, long-form outputs such as extended reports, summaries, or code.
When was Gemini 3.1 Pro released?
Gemini 3.1 Pro was released in February 2026 and is available on MindStudio under the full model identifier gemini-3.1-pro-preview.
What people think about Gemini 3.1 Pro
Community reception on r/singularity was largely positive at launch, with the benchmark announcement thread accumulating over 2,300 upvotes and 528 comments. Users frequently highlighted the ARC-AGI-2 score and the 1M-token context window as notable technical achievements.
Some community members raised questions about hallucination rates, with a dedicated thread asking whether Google had addressed accuracy issues seen in prior Gemini versions. Practical use cases discussed included coding assistance, long-document analysis, and agentic workflows.
Google releases Gemini 3.1 Pro with Benchmarks
Google just dropped Gemini 3.1 Pro. Mindblowing model.
Google Gemini 3.1 Pro Preview Soon?
Gemini 3.1 Pro Preview – Has Google finally fixed the hallucination problems they had?
Parameters & options
Must be less than Max Response Size
Explore similar models
Start building with Gemini 3.1 Pro
No API keys required. Create AI-powered workflows with Gemini 3.1 Pro in minutes — free.