Gemini Omni
by Google DeepMind
Google's first any‑to‑any AI model.
Text, images, audio, and video in, a single video out.

Key features
Technical specifications
Omni Flash
First model in Google's Gemini Omni family
Video
Image and audio output planned in the Gemini Omni roadmap
Up to 10s
Flash clips capped at 10 seconds at launch to widen access
720p
Gemini Omni Flash outputs 720p video
Any mix
Text, image, audio, and video in one prompt
Native
Synchronized audio generated with every clip; voice via Avatars
SynthID
Imperceptible AI-provenance watermark on every clip
Google DeepMind
Successor positioning to Veo for any-to-any video creation
Use cases
Multi-input storyboarding
A character image, location photo, music cue, and beat go in; the model builds the shot and iterates.
Conversational video editing
Edit any clip in plain language: swap wardrobe, change a background, or retime a beat. The rest stays steady.
Marketing video
Ad cuts that respect brand colors, product shape, and on-screen text. One photo, one brief, one finished spot.
Educational explainers
Visualize science, history, and engineering with built-in physics. The science stays honest, the footage clean.
Spokesperson video
A portrait plus a voice reference gives the same on-camera presenter across shorts, courses, and walkthroughs.
Social shorts
10-second clips fit YouTube Shorts, Reels, and TikTok. Generate variations, then publish the one that lands.
Prompt examples


Product launch
Avant-garde sneaker mid-air over a titanium plinth, hard key light, launch mood
Edit prompt
Nature explainer
Droplet frozen as a crystalline crown on a dewy leaf, backlit sunrise macro
Edit prompt
Avatar spokesperson
Poised studio host addressing the lens, warm three-point light, 85mm bokeh
Edit prompt
Architectural walkthrough
Golden-hour light through a brutalist concrete villa, long shadows, drifting dust
Edit prompt
Simple pricing
Get started for free today, with the option to upgrade or cancel anytime.
Basic
1100 monthly credits
1 user only
All models
Workflows
Standard
3625 monthly credits
1 user only
All models
Workflows
Pro
6350 shared monthly credits
1 user
All models
Workflows
Pro Max
24650 shared monthly credits
1 user
All models
Workflows
Enterprise
For higher limits
Custom
pricing and billing terms

Free
For playing around
$0
forever free
FAQs
Gemini Omni Flash 1.1
Google DeepMind
Google's video model where the edit is a sentence. Now with 4K output and 40-second scenes.
MiniMax H3 Max
fal.ai
fal's post-trained MiniMax H3. Tuned for prompt adherence, rebuilt for speed.
Runway Ruby
Runway
Runway's SDR to HDR model. Lifts finished video into 16-bit EXR frames and ProRes masters.
Hyperion 2.5
Topaz Labs
Topaz Labs' SDR to HDR model. Turns 8-bit video into 10-bit ProRes and 16-bit EXR masters.