Gemini Omni Flash 1.14KDirect every frame

Start generating

REMIX REALITY

Change the aesthetic, action, or effect - all from your input video

Gemini Omni Flash example 1

Input video

Prompt

Move her onto a vast blooming flower field

Gemini Omni Flash example 2

Input video

Prompt

Cinematic close-up of her face

Gemini Omni Flash example 3

Input video

Prompt

Make the bicycle completely invisible

Tighter control, lower cost

Extend scenes for longer storytelling

1.1 reads up to ten seconds of what came before - not just the last second - and continues the motion, light and story in ten-second steps. One take, held as long as the story needs

Extend scenes for longer storytelling — Original

Direct it from both ends

Pin the opening frame and the closing frame - 1.1 builds the move between them: camera pushes, zoom transitions, seamless loops

Direct it from both ends
Start frame
Start frame
Start frame
Generated frame
Generated frame
Generated
End frame
End frame
End frame

References drive the shot

Pin the opening frame and the closing frame - 1.1 builds the move between them: camera pushes, zoom transitions, seamless loops

References drive the shot
Gemini Omni Flash feature 3 input image
Gemini Omni Flash feature 3 input image

Input image

Input video

Recut without reshooting

Load up to thirty seconds of real footage and direct changes in plain language - swap the subject, shift the mood, retime the action

Recut without reshooting — Spaceship

Physics that holds at 4K

Gravity, momentum, fluid dynamics - 1.1 keeps them honest through every ten-second extension, and the 4K master keeps every gear and jewel razor sharp

Physics that holds at 4K before
Physics that holds at 4K after

Transform your world

Change the aesthetic, action, or effect based on your input video

Gemini Omni Flash showcase 1
Gemini Omni Flash input image placeholder
Gemini Omni Flash input image placeholder

Input image

Input video

Input audio

Transform the entire scene into a stylized 2D cartoon animation aesthetic - hand-drawn illustration style with bold clean black outlines, cel-shaded coloring with flat shaded colors and soft gradients, painterly textures on the buildings, street, and sky, slightly exaggerated character proportions emphasizing his cool youthful vibe, vibrant saturated color palette reminiscent of modern animated feature films

Prompt

Combine multiple inputs. Feed Omni a mix of clips, stills, and references - and let it weave them into one coherent story instead of a collage

Gemini Omni Flash showcase 2
Gemini Omni Flash motion reference placeholder
Gemini Omni Flash motion reference placeholder

Input image

Input video

Transfer motion and styles. Lift the movement from one clip and the look from another - Omni applies both to your output in a single pass

Gemini Omni Flash showcase 3
Gemini Omni Flash character reference placeholder
Gemini Omni Flash character reference placeholder

Input image

Input video

Swap characters or objects with a reference image. Drop in a reference alongside your video, and the new character takes over the motion and dialogue without missing a beat

Gemini Omni Flash showcase 4
Gemini Omni Flash sketch reference placeholder
Gemini Omni Flash sketch reference placeholder

Input image

Translate drawings into video. Turn rough sketches into real footage - and use your doodles to direct exactly how each element should move

How to generate

Generate using the most advanced video model

  1. Input Image Reference
    1

    Input Image Reference

    Upload reference images to guide your vision

  2. WRITE THE PROMPT
    2

    WRITE THE PROMPT

    Use natural language to describe desired scenario and sounds

  3. Generate with Gemini Omni Flash
    3

    Generate with Gemini Omni Flash

    Click "Generate" button and Receive high-fidelity video in seconds.

Flow AI Video

CREATE FROM ANYTHING

Create anything from anything from any input - starting with video

Try it nowTry it now

What Creators Are Saying

From first projects to everyday creative work, these highlights from Flow AI Video creator reviews explore the tools, workflows and support that made a difference.

A
Alexandre

Strong creative tools and satisfying results stand out to Alexandre. But a question about credits made the biggest impression: Tim answered with clarity and care, turning a support request into a particularly positive experience.

S
Spock

Spock's first projects have already led to videos they are excited about. Helpful support rounds out that early experience, with more platform improvements something to look forward to.

R
Rha

After editing since 2008, Rha still finds the leap from a written prompt to cinematic video remarkable. The range of tools can feel like a lot to take in, but it also opens up plenty to explore.

AS
Akun Suliman

For Indonesian creator Akun Suliman, practical help with making content is what matters. Useful creative tools and encouragement from the team have made a real difference to the experience.

SJ
Shatanu Jachak

Shatanu Jachak appreciates a creative workflow that feels smooth and fast while delivering quality results. Support for the creative community and room to experiment make visual storytelling especially rewarding.

FM
Freissy Mediaart

An approachable interface, clearly organized tools and fast cloud rendering stand out to Freissy Mediaart. Across numerous image and video projects, the creator reports being pleased with the results.

LE
Ls estúdios

Regular improvements give Ls estúdios a reason to keep exploring. Useful tools and new features build on an experience the creator already enjoys, with a clear focus on helping creative work.

RR
Rian Rizky Ananta

Rian Rizky Ananta values competitive credit pricing alongside a steady flow of new features. Just as important is the impression of a team that listens and keeps moving the product forward.

DD
Deyo D

Deyo D puts creator-program support at the center of the experience. Professional, approachable help with questions makes the team feel invested in creators and their progress.

Trusted by 5.000+ people worldwide

Still have questions?

We’ve answered the most frequently asked questions

What is Gemini Omni Flash?
Gemini Omni Flash is Google DeepMind's video generation and editing model. It combines Gemini's reasoning and world knowledge with the ability to create and edit video from any combination of image, text, video, and audio inputs.
How does it work?
You give it a clip, an image, a sketch, or just text — and describe what to do in natural language. It generates or edits the video, maintaining scene consistency across multiple turns. Each instruction builds on the previous one, like directing a conversation with an editor.
How is it different from other video models?
Three things: conversational multi-turn editing with scene consistency, native multimodal inputs (image + text + video + audio in one prompt), and grounding in Gemini's world knowledge — physics, history, culture rendered accurately, not just plausibly.
What input and output formats are supported?
Inputs: image, video, audio, text, sketches. Image: PNG, JPG, WebP up to 20 MB. Video: MP4, MOV, WebM up to 60s. Audio: MP3, WAV up to 30s. Output: MP4 (H.264) in 16:9, 9:16, 1:1, or 4:5.
What resolution and duration can it generate?
Native output: 720p at 24fps. Default clip length is 8 seconds per generation, extendable to 60s via continuation. 1080p upscale available as a post-processing step.
Which languages does multilingual speech support?
English, Spanish, Mandarin, Japanese, French, German, Portuguese, Hindi, Korean, Russian, Arabic. Regional accents and dialects supported. Lip-sync re-renders to match the phonemes of the target language.
How many edit turns are supported per session?
Unlimited multi-turn editing within a session. Scene consistency holds across turns. Each turn produces a new generation; previous turns remain in session history and can be branched or restored.
What safety and provenance signals are included?
Every output carries SynthID — an imperceptible digital watermark — and C2PA Content Credentials embedded in the file metadata. Both persist through standard re-encoding and platform uploads.