VFX through conversation
Describe the transformation you want in everyday words and refine it without rebuilding the shot.
Official model guide plus a working online generator
Associez le raisonnement de Gemini à une génération vidéo rapide, au contrôle par références multimodales et à un montage vidéo conversationnel. The Gemini Omni Flash AI video generator just below lets you create from a prompt, an image, a pair of frames or supported references.

The full MuseGen video workspace is built in here with Gemini Omni Flash already selected. You can still try other variants or change models and keep the rest of your setup.
You need an account to generate, and credits depend on the model, variant, length, resolution and other settings you choose. Credits for failed tasks come back automatically.
Know where Gemini Omni Flash fits before you invest time in references, prompts and renders at full resolution.
Gemini Omni Flash links multimodal understanding straight to video creation. It reasons over text, pictures and video before it generates or edits, which makes it particularly handy when your instruction depends on how several references relate or on understanding footage you already have.
For creative and editing teams, product marketers, teachers, social creators and fast-moving prototypers, the real benefit is not one headline score. It is how Gemini Omni Flash pairs Text to video, multi-reference video and editing of existing footage with Text, several images and one source video. That pairing decides whether the model can hold on to a visual direction you have prepared, or has to imagine most of the scene from words.
Here, research and production happen in the same place. Read the official specs, look through the source material, use the prompt structure, then create in the built-in generator. It opens on Gemini Omni Flash, and everything else MuseGen offers — uploads, progress, history, reuse and downloads — is still there.
Plan your inputs, format, length and quality tier from these official capabilities before you spend a credit.
Begin with projects where this model's strongest controls give you a real edge, rather than choosing on top resolution alone.
Describe the transformation you want in everyday words and refine it without rebuilding the shot.
Feed in product and brand images so design details stay recognisable in a generated commercial moment.
Give an existing video a new look, location, visual joke or brand treatment for a short campaign.
Pull characters, places, props and art direction from different images into one request.
This official footage comes from the developer's own launch or product material and has been optimised to play quickly on this page.
From the official source: Footage released by Google for Gemini Omni Flash, downloaded and optimised for this guide.
Take an idea from first prompt to a configured Gemini Omni Flash render without leaving the model page.
Write a prompt covering the subject, action, setting, camera, look, pacing and audio. Got references? Upload them and say what each one is there for.
Leave Gemini Omni Flash selected, pick the variant, mode, length, frame shape, resolution and audio options that suit the shot, then look at the credit cost displayed.
Start the job, watch it progress beside the generator, check the finished clip, then reuse the same settings for a fresh attempt or download the file.
The core capabilities that decide how Gemini Omni Flash deals with direction, references, movement, audio and final delivery.
Gemini works out links, instructions, objects and surrounding context across every kind of input before it creates anything.
Change an existing clip in plain language — tweak an effect, swap visual elements or keep building on an idea.
Combine several images to steer subjects, products, locations, outfits, style and other parts of the scene.
A fixed video length suits social concepts, quick effects trials, product moments and fast iteration.
Transform footage you already have while keeping the motion or framing that made it worth using.
Get a full audiovisual result whenever sound, beat or effects belong to the change you asked for.
A good Gemini Omni Flash prompt reads like a short production brief: a subject, actions in order, a camera plan, an art direction and a soundtrack.
Subject + actions in order + setting + camera + lighting + look + timing + dialogue and sound + what must stay consistent
“Images 1–4 fix the exact bottle shape, label, materials and colours. In the source video, swap the grey studio for a sunlit desert of pink sand, leave the camera path and the timing of the hands as they are, and add shimmering heat haze with a soft wind-chime sound rising underneath.”
Attaching references is not enough; say which subject, style, object, place or behaviour should come from each.
When editing, separate what must stay untouched from the specific elements you want swapped out.
Stick to one readable idea with a clear start, a change and a final payoff that suits the fixed length.
Choose the tier and format that match your stage of the project. Draft settings help you find the shot; premium settings finish a direction that is already working.
| Feature | What Gemini Omni Flash offers |
|---|---|
| Tiers | Preview |
| Modes | Text to video, multi-reference video and editing of existing footage |
| Accepted inputs | Text, several images and one source video |
| Reference support | As many as 16 images plus one source video |
| Length | 8 seconds |
| Output resolution | 720p |
| Frame shapes | 16:9 widescreen and 9:16 vertical |
| Sound | Native audio generation |
AI video behaves best when each prompt carries one clear visual idea. Check the important details before you publish, and treat a first render as a directed take you can improve.
Busy interactions between several characters, fast objects crossing the frame, legible lettering, brand marks, fingers and exact counts of things can still change from take to take. Lean on clear references, keep crowded action simple and check continuity frame by frame.
More pixels are no substitute for art direction. Nail the story beat, framing, motion and sound on an inexpensive setting first, then move the best version up to a premium tier or a higher resolution.
Weigh up a different mix of motion, references, audio, speed, length and resolution without leaving the MuseGen model library.
Créez des plans cinématographiques maîtrisés, guidés par première et dernière image, avec audio natif, durée flexible et vraie sortie 4K.
View Kling v3Générez une vidéo fluide et cohérente à partir de texte, d’une image de départ ou de neuf images de référence, avec audio natif synchronisé.
View HappyHorse 1.1Générez de la vidéo 2K avec son stéréo natif à partir de texte, d’images clés et de quinze images, clips et pistes audio de référence à la fois.
View MiniMax H3Explorez la narration multi-plans native, l’audio et la vidéo synchronisés, la durée automatique, le Diffusion Fidelity Rendering et une sortie professionnelle 4K HDR du nouveau modèle open-weight de Lightricks.
View LTX 2.5Gemini Omni Flash is an AI video generation model from Google. Gemini Omni Flash links multimodal understanding straight to video creation. It reasons over text, pictures and video before it generates or edits, which makes it particularly handy when your instruction depends on how several references relate or on understanding footage you already have. On MuseGen the full generator sits right here, letting you go from reading about the model to creating with it without switching workspaces.
Gemini Omni Flash offers Text to video, multi-reference video and editing of existing footage. That means you can begin from a written idea, fix the first frame with a picture, or bring in extra references when the layout, the identity or the motion needs tighter control.
Clips run 8 seconds, with output at 720p. Use a lower resolution while you explore ideas, then switch to the highest sensible resolution once you are ready to judge fine detail or hand the shot over.
Gemini Omni Flash handles Native audio generation. Write speech, room tone, score and sound effects into the prompt on purpose so the soundtrack backs up the on-screen action and the emotional pace of the scene.
It takes Text, several images and one source video. For references it supports As many as 16 images plus one source video. Tell the prompt what each uploaded file is for rather than hoping the model guesses which image governs identity, style, composition or motion.
Gemini Omni Flash suits creative and editing teams, product marketers, teachers, social creators and fast-moving prototypers especially well. The right pick still comes down to the individual shot: use the official specs, features, examples and prompt tips on this page to judge whether its mix of control, pace, resolution, audio and references fits the job.
A dependable prompt covers the subject, the action, the place, the camera, the light, the look, the timing and the sound. List events in the order they happen, put exact dialogue in quotes, and say what has to stay consistent. If you upload references, name each one.
Yes. The complete Gemini Omni Flash generator is built into this page just below the hero. Pick text, an image, two frames or references as needed, set the available options, check the credit cost shown, and start generating without ever leaving this page.
Capabilities and media on this page were checked against the developer's official product pages, announcements and documentation.
Head back to the full Gemini Omni Flash AI video generator at the top, add a prompt or your references, and turn the next idea you have into a finished clip.
Generate nowSee every model