Gemini Omni: Speak It, See It, Share It
Gemini Omni: Speak It, See It, Share It
Gemini Omni turns video creation into something as natural as a conversation. Think of it as Nano Banana, but for video. You can start from scratch, remix clips from your gallery, or begin with a premade template.
Create anything, from everything
Blend text, images, and video to bring ideas to life in motion.
From concept to clip
Gemini Omni is a creative partner for multimodal content. It keeps the soul of the shot intact: swap the background, change the wardrobe, or transfer styles while preserving the details that matter.
Instant inspiration
Rather than waiting for a spark, jump into curated styles to see what is possible in a single tap.
Easy editing
Just tell Gemini what to change. Swap characters, adjust lighting, stabilize the footage, or modify the background, all through chat.
Be the star of your own show
Add your AI avatar to make content that looks and sounds like you, without re-uploading your image every time.
Say hello to Gemini Omni
Gemini Omni will replace Veo in the Gemini app. It combines Gemini's core intelligence with advanced generative media, including image-to-video and video-to-video AI editing. It understands the world, blends different media types, and gives you more editing control for all your AI video generation and editing needs.
Gemini Omni Flash
A multimodal AI video generation and editing model that now replaces the previous Gemini Veo 3.1 model. Capabilities include:
- Create 10-second videos
- Native audio generation
- Turn up to 5 photos into a video
- Video-to-video editing
- Multi-turn editing
- AI avatar
A Google AI subscription is required. Features vary by tier and geography, and access is limited to users 18 and older.
Frequently asked questions
What is Gemini Omni?
Gemini Omni is a model that understands the world around you, so you can animate photos or create video from any input. Built on Gemini's world understanding and native multimodality, it produces outputs that follow real-world logic and lets you shape them step by step through natural conversation. With a single prompt, you effectively become an AI video editor. You can turn any combination of text, photos, or video into video, generate clips from up to five photo references, and edit footage easily.
Who can use it?
Gemini Omni is available to users 18 and older on a Google AI Plus, Pro, or Ultra plan, in every language and market where the Gemini app is available. Certain features, such as avatars and video-to-video editing, may be restricted in some countries; check the help center for details.
How do AI avatars work?
Creators have been making videos from photos and inserting themselves into their work, and avatars make that easier. An avatar is a digital version of yourself that lets you generate videos that look and sound like you, safely and securely. It is completely optional, and only you can use your own avatar. You can keep uploading photos or skip that step and use the avatar instead.
How is AI-generated content identified?
In line with Google's AI principles, every video generated in the Gemini app carries SynthID, an imperceptible watermark for identifying Google AI-generated content. Verification now extends to images, video, and audio: upload a file and ask whether it was generated using Google AI, and Gemini checks for SynthID and applies its own reasoning to respond.