Gemini 1.5 Pro
Gemini 1.5 Pro is a multimodal text generation model from Google with a 2,000,000 token context window.
Long-context multimodal understanding from Google
Gemini 1.5 Pro is a foundation model developed by Google, designed to handle a wide range of multimodal tasks including visual understanding, classification, summarization, and content generation from image, audio, and video inputs. It accepts photographs, documents, infographics, and screenshots alongside text, making it suited for workflows that require processing diverse input formats in a single session.
The model's most notable technical characteristic is its 2,000,000 token context window, which allows it to process very large volumes of text or multimodal content within a single request. This makes it particularly applicable to tasks involving long documents, extended conversations, or large codebases. The model was added to MindStudio in May 2024 and is currently listed as deprecated, meaning users should consider migrating to a successor model for production use.
What Gemini 1.5 Pro supports
Long Context Window
Supports up to 2,000,000 tokens in a single context, enabling processing of very large documents, codebases, or conversation histories without truncation.
Visual Understanding
Analyzes photographs, infographics, screenshots, and documents provided as image inputs, returning text-based descriptions, classifications, or summaries.
Audio Processing
Accepts audio inputs and can perform tasks such as transcription, summarization, and content extraction from spoken material.
Video Analysis
Processes video inputs to extract information, summarize content, or answer questions about visual and audio elements within the footage.
Text Generation
Generates text responses up to 8,192 tokens per reply, covering tasks such as drafting, summarization, classification, and question answering.
Document Comprehension
Reads and reasons over structured and unstructured documents including PDFs and infographics, extracting key information or answering specific questions.
Ready to build with Gemini 1.5 Pro?
Get Started FreeCommon questions about Gemini 1.5 Pro
What is the context window size for Gemini 1.5 Pro?
Gemini 1.5 Pro has a context window of 2,000,000 tokens, allowing very large amounts of text or multimodal content to be processed in a single request.
What is the maximum response size for Gemini 1.5 Pro?
The model supports a maximum response size of 8,192 tokens per output.
What input types does Gemini 1.5 Pro support?
Gemini 1.5 Pro is designed to handle text, image, audio, and video inputs, making it suitable for multimodal workflows.
Is Gemini 1.5 Pro still available for use?
Gemini 1.5 Pro is currently listed as deprecated on MindStudio. Users with active workflows using this model should consider migrating to a current successor model.
Who publishes Gemini 1.5 Pro?
Gemini 1.5 Pro is published by Google and is provided as a first-party model on MindStudio.
Parameters & options
Explore similar models
Start building with Gemini 1.5 Pro
No API keys required. Create AI-powered workflows with Gemini 1.5 Pro in minutes — free.