Gemini Omni Flash
Gemini Omni Flash is a video generation model from Google that supports image and video inputs with multimodal editing capabilities.
Multimodal video generation and editing from Google
Gemini Omni Flash (gemini-omni-flash-preview) is a video generation model developed by Google, released in June 2026. It accepts a range of input types including single images, arrays of reference images, and existing video files, making it suited for both generation and editing workflows. The model operates with a context window of 5,000 tokens and supports configurable aspect ratios and generation modes.
The model is tagged as multimodal and editing-capable, meaning it can work across image-to-video and video-to-video tasks within a single interface. Its support for reference image arrays suggests it can incorporate multiple visual references when generating or modifying video content. Gemini Omni Flash is available on MindStudio without requiring separate API key configuration, and it carries a preview status indicating it is an actively developing release.
What Gemini Omni Flash supports
Video Generation
Generates video output from text prompts or input media. Supports configurable aspect ratios via a toggle group input.
Image-to-Video
Accepts a single input image URL and converts it into video content. Useful for animating still images.
Reference Image Support
Accepts an array of reference images to guide video generation. Enables multi-image visual context for more controlled output.
Video Editing
Takes an existing video URL as input for editing or transformation tasks. Tagged explicitly as an editing-capable model.
Multimodal Input
Combines image, image array, and video inputs within a single model interface. Tagged as multimodal, supporting cross-media workflows.
Configurable Modes
Exposes a mode selector input allowing users to switch between different generation or editing behaviors at runtime.
Ready to build with Gemini Omni Flash?
Get Started FreeCommon questions about Gemini Omni Flash
What is the context window for Gemini Omni Flash?
Gemini Omni Flash has a context window of 5,000 tokens.
What input types does Gemini Omni Flash accept?
The model accepts a mode selector, aspect ratio toggle, a single input image URL, an array of reference image URLs, and an input video URL.
Is Gemini Omni Flash available for production use?
The model is listed with a 'preview' status in its full name (gemini-omni-flash-preview), indicating it is available but still in a preview stage as of its June 2026 release.
What is the pricing for Gemini Omni Flash on MindStudio?
Pricing information has not been published in the available metadata. Check MindStudio's pricing page or model catalog for current rates.
Who publishes Gemini Omni Flash?
Gemini Omni Flash is published by Google and provided as a first-party model on MindStudio.
Documentation & links
Parameters & options
An image to use as the starting frame or guide for the video.
Provide reference images of subjects, styles, or objects to include in the video.
Upload a video to edit.
Explore similar models
Start building with Gemini Omni Flash
No API keys required. Create AI-powered workflows with Gemini Omni Flash in minutes — free.