Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Video Generation Model

Gemini Omni Flash

Gemini Omni Flash is a video generation model from Google that supports image and video inputs with multimodal editing capabilities.

PublisherGoogle
TypeVideo
Context Window5,000 tokens
ReleasedJune 2026
Price<$0.0001/video
LATESTMULTIMODALEDITING

Multimodal video generation and editing from Google

Gemini Omni Flash (gemini-omni-flash-preview) is a video generation model developed by Google, released in June 2026. It accepts a range of input types including single images, arrays of reference images, and existing video files, making it suited for both generation and editing workflows. The model operates with a context window of 5,000 tokens and supports configurable aspect ratios and generation modes.

The model is tagged as multimodal and editing-capable, meaning it can work across image-to-video and video-to-video tasks within a single interface. Its support for reference image arrays suggests it can incorporate multiple visual references when generating or modifying video content. Gemini Omni Flash is available on MindStudio without requiring separate API key configuration, and it carries a preview status indicating it is an actively developing release.

What Gemini Omni Flash supports

Video Generation

Generates video output from text prompts or input media. Supports configurable aspect ratios via a toggle group input.

Image-to-Video

Accepts a single input image URL and converts it into video content. Useful for animating still images.

Reference Image Support

Accepts an array of reference images to guide video generation. Enables multi-image visual context for more controlled output.

Video Editing

Takes an existing video URL as input for editing or transformation tasks. Tagged explicitly as an editing-capable model.

Multimodal Input

Combines image, image array, and video inputs within a single model interface. Tagged as multimodal, supporting cross-media workflows.

Configurable Modes

Exposes a mode selector input allowing users to switch between different generation or editing behaviors at runtime.

Ready to build with Gemini Omni Flash?

Get Started Free

Common questions about Gemini Omni Flash

What is the context window for Gemini Omni Flash?

Gemini Omni Flash has a context window of 5,000 tokens.

What input types does Gemini Omni Flash accept?

The model accepts a mode selector, aspect ratio toggle, a single input image URL, an array of reference image URLs, and an input video URL.

Is Gemini Omni Flash available for production use?

The model is listed with a 'preview' status in its full name (gemini-omni-flash-preview), indicating it is available but still in a preview stage as of its June 2026 release.

What is the pricing for Gemini Omni Flash on MindStudio?

Pricing information has not been published in the available metadata. Check MindStudio's pricing page or model catalog for current rates.

Who publishes Gemini Omni Flash?

Gemini Omni Flash is published by Google and provided as a first-party model on MindStudio.

Parameters & options

ModeSelect
Default: text-to-video
Text to VideoImage to VideoReference to VideoEdit Video
Aspect RatioToggle Grouptext-to-video only
Default: 16:9
16:99:16
Input ImageImage URLimage-to-video only

An image to use as the starting frame or guide for the video.

Reference ImagesImage URL Arrayreference-to-video only

Provide reference images of subjects, styles, or objects to include in the video.

Input VideoVideo URLedit only

Upload a video to edit.

Start building with Gemini Omni Flash

No API keys required. Create AI-powered workflows with Gemini Omni Flash in minutes — free.