Text to video
Describe a shot in plain language and get a moving scene back, with camera direction, lighting and pacing taken from the prompt.
Gemini Omni 1.1 Flash generates video from text, images and reference media, then keeps editing the result in a conversation. Write a prompt below to open the workspace.
This page is an independent product walkthrough. The prompt box is a preview and does not render video by itself.
Short clips generated from text, still images, reference media and first-to-last-frame interpolation. Every example below started as a one-line brief.






One model for text, image and reference-driven video generation, with an editing loop that keeps the same clip instead of restarting from scratch.
Describe a shot in plain language and get a moving scene back, with camera direction, lighting and pacing taken from the prompt.
Animate a still: keep its composition and subject, then add the motion, atmosphere and camera move you ask for.
Pass reference images or video clips so characters, products and art direction stay consistent across shots.
Supply a starting frame, an ending frame, or both, and the model interpolates the motion between them.
Keep editing through the Interactions API: ask for warmer light, a tighter frame or a different ending and the model revises the clip.
Generate 3-10 second clips, continue a scene past its original ending, and upscale drafts when you are happy with the cut.
Four steps that match how the model is used in Google AI Studio, the Gemini API and Gemini Enterprise Agent Platform.
Write a prompt, or upload an image, a reference clip or a first and last frame to anchor the shot.
Choose 9:16 or 16:9 and render a 360p draft, then move to 720p, 1080p or 4K when the edit is locked.
Ask for changes through the Interactions API: lighting, pacing, camera movement, ending, wardrobe.
Continue the scene, stitch the best takes into a sequence, and download the final render.
Gemini Omni 1.1 Flash accepts four kinds of input. Choosing the right one up front saves the most iterations.
From a single vertical ad to a previsualized scene, the same model covers drafts, revisions and finals.
Render 9:16 hooks in a draft resolution, then upscale the winners for paid placement.
Block a scene with first and last frames before a shoot day to test timing and camera moves.
Turn catalogue stills into looping product motion while the label and packaging stay accurate.
Extend a storyboard panel into a few seconds of motion to pitch pacing before animation.
Keep one look and produce alternate endings, hooks and framing for a launch.
Apply feedback as a conversation instead of rebuilding the shot from the prompt.
Monthly and annual subscriptions, plus a one-time credit booster when a launch needs extra renders. An active subscription is required to generate.
300 credits / month
1,000 credits / month
3,000 credits / month
Annual billing is available at $90 / $190 / $490. Need extra renders mid-cycle? A $29 booster adds 300 credits for 90 days.
Quick answers on inputs, editing, output formats, access paths and regional limits.
Gemini Omni 1.1 Flash (gemini-omni-1.1-flash) is a video generation model that turns text, images and reference media into short video, and then keeps editing that video in a conversation.
Yes. A written prompt can describe the subject, the action, the camera movement and the lighting, and the model renders the scene without any source media.
Yes. Upload a still and describe the motion you want. The model keeps the composition and subject of the image and animates it.
Reference media anchors things that must stay consistent between shots, such as a character design, a product label or an art direction.
You can provide a starting frame, an ending frame, or both. The model fills in the motion between them, which is useful when a shot has to land on an exact pose or product angle.
Through the Interactions API you keep working on the clip you already have. Asking for warmer light, a tighter frame or a different ending revises the video instead of starting over.
A single generation produces a 3-10 second clip. A scene can be continued past its original ending, and extensions can take a sequence up to about 40 seconds.
360p drafts render quickly for iteration, 720p is the standard output, and finished shots can be upscaled to 1080p or 4K.
9:16 for vertical and 16:9 for widescreen are both supported, so the same shot can serve social placement and long-form edits.
Inline uploads are limited to a few megabytes; larger files must be passed as a URI rather than inline data.
The model is available through Google AI Studio, the Gemini API and the Gemini Enterprise Agent Platform, depending on whether you are prototyping or shipping production volume.
Editing uploaded video is restricted in the EEA, the UK and Switzerland, so those accounts cannot use the video editing paths on uploaded material.
Yes. An active subscription is required, and accounts without one are sent to the pricing page. Sign-in is Google only.
No. geminiomni.lat is an independent third-party website and is not affiliated with, endorsed by, sponsored by or operated by Google.
Start with a prompt, a still or a pair of frames. Subscribe, generate a draft, and direct the shot from there.