gemini-omni-flash-preview — Video Generation AI Model
Google Gemini Omni: text/image/video in, video out, with conversational editing. Which input ports apply depends on Task — Text to Video: Text only. Image to Video: + Image 1. Reference to Video: + Image 1–3. Edit: + Video (aspect ratio is ignored). Unused ports are dropped even if wired. Needs API key.
Details
- CategoryVideo Generation
- Year2026
- LicenseGoogle Cloud Terms
Compliance & Provenance
- ProviderGoogle · GPAI
- EU AI Act RiskLimited Risk
- Art. 50 TransparencyRequired — AI-generated outputs are marked
Inputs & Outputs
- TextInput · string
[all tasks] Prompt describing the video to generate, or the edit to apply.
- Image 1Input · image · optional
[Image to Video / Reference to Video / Edit] Source image, or reference image 1. Ignored by Text to Video.
- Image 2Input · image · optional
[Reference to Video / Edit] Reference image 2. Ignored by Text to Video and Image to Video.
- Image 3Input · image · optional
[Reference to Video / Edit] Reference image 3. Ignored by Text to Video and Image to Video.
- VideoInput · video · optional
[Edit only] The existing clip to modify. Ignored — and not uploaded — by every other task.
- Previous Interaction IDInput · string · optional
[any task] Interaction ID output by a previous Omni node. Continues editing that same video instead of generating a new one.
- VideoOutput · video
Generated or edited video.
- Interaction IDOutput · string
Wire into another Omni node's Previous Interaction ID port to keep editing this video.
Tags
Alternatives in Image/Video Models
- GPT Image
OpenAI GPT-Image text-to-image / edit. Needs API key.
- Nano Banana
Google Gemini-based image gen/edit with multimodal reasoning. Needs API key.
- Sora 2
OpenAI Sora 2 text-to-video / img2video. Needs API key.
- VEO
Google VEO 3.1 video gen (up to 4K, with native audio). Needs API key.
- Video Analysis
Gemini video analysis. Feed a video and get a transcript, timecodes, or highlight analysis as text. Needs API key.