AI video model

Gemini Omni Video AI Video Generator

Choose Gemini Omni Video when you need a Gemini AI video generator where a detailed written brief remains the source of truth even with optional images, and the delivery fits a fixed 4, 6, 8, or 10 second landscape or portrait clip.

Gemini Omni Video task directions

These local references help turn the current Gemini Omni Video input and control contract into a concrete task direction. They are not presented as output from this model.

Every card identifies whether it is a verified model-family example or a visual reference. Reference cards are not model benchmarks or before-and-after claims.

Input contract
Text prompt required; optional compatible reference media: image (up to 3)
Current controls
Prompt length: Up to 20,000 characters · Duration: 4s, 6s, 8s, 10s options · Output resolution: 720P / 1080P / 4K · Framing: 2 aspect-ratio choices
Selection boundary
A prompt remains required when images are attached. Compare another model when the task requires free-form duration, a square or other frame, driving audio, video references, or audio-generation controls.

Action beat

State one subject action and the moment the clip should begin from.

Task direction
State one subject action and the moment the clip should begin from.

Reference direction, not a model output

Local ImageStyle visual reference. It is not presented as a model-specific result.

Create with Gemini Omni Video

Camera direction

Name the camera behavior and the visual change it should reveal over time.

Task direction
Name the camera behavior and the visual change it should reveal over time.

Reference direction, not a model output

Local ImageStyle visual reference. It is not presented as a model-specific result.

Create with Gemini Omni Video

Ending state

Define the final visual state, without treating this reference as a result from the current model.

Task direction
Define the final visual state, without treating this reference as a result from the current model.

Reference direction, not a model output

Local ImageStyle visual reference. It is not presented as a model-specific result.

Create with Gemini Omni Video
Create
AI video
Input
Text prompt required; optional compatible reference media: image (up to 3)
Duration
4s, 6s, 8s, 10s options

Compare available settings with related models

Check whether the input, output, and framing controls fit the job before you open a workspace. This table does not treat the example media as a cross-model performance test.

What you can control with Gemini Omni Video

Keep the prompt as the source of truth

Gemini Omni Video currently supports a 20,000-character prompt with optional images, 720p through 4K choices, fixed 4, 6, 8, or 10 second clips, and 16:9 or 9:16 framing.

Add compatible media when it helps

Text prompt required; optional compatible reference media: image (up to 3). Only upload media you are allowed to use for the current project.

Confirm the clip settings

Current clip settings include 4s, 6s, 8s, 10s options, 720P / 1080P / 4K, and 2 aspect-ratio choices.

Current Gemini Omni Video settings

These settings come from the current ImageStyle model configuration and help you decide whether the workflow fits the job.

SettingCurrent availability
InputText prompt required; optional compatible reference media: image (up to 3)
Prompt lengthUp to 20,000 characters
Duration4s, 6s, 8s, 10s options
Output resolution720P / 1080P / 4K
Framing2 aspect-ratio choices
Additional controlsCurrent core settings

AI video model

When to choose Gemini Omni Video

Long prompt-led video briefs with optional images and fixed 4, 6, 8, or 10 seconds.

Current workflow fit

Choose Gemini Omni Video when you need a Gemini AI video generator where a detailed written brief remains the source of truth even with optional images, and the delivery fits a fixed 4, 6, 8, or 10 second landscape or portrait clip.

When to compare another model

A prompt remains required when images are attached. Compare another model when the task requires free-form duration, a square or other frame, driving audio, video references, or audio-generation controls.

How to create a video with Gemini Omni Video

Write the complete scene and movement brief first, then use optional images only to anchor details that the text needs to preserve.

  1. 1

    Write the complete timed brief

    Choose the fixed duration and orientation, then state the scene, action, camera, and required details before attaching optional images.

  2. 2

    Add reference media when needed

    Add approved compatible media only when an existing frame, motion, or sound should guide the result.

  3. 3

    Confirm duration and output controls

    Confirm the current 4s, 6s, 8s, 10s options, 720P / 1080P / 4K, and 2 aspect-ratio choices in the workspace before generation.

  4. 4

    Generate and review continuity

    Before public use, inspect subject motion, camera changes, visible text, sound, and the final frame.

Checks before you create and publish

Treat the workspace as current
Model availability and settings can change. Use the options shown in the current workspace before you submit.
Use approved source media
When you upload images, video, or audio, confirm that you are allowed to use and publish it for the current project.
Review the generated video
Before publishing, manually check motion continuity, text, sound, people, product details, and marks.

Gemini Omni Video FAQ

Is a prompt still required with images in Gemini Omni Video?
Yes. The current workflow requires a written prompt even when optional image references are attached.
Can I start Gemini Omni Video from reference media?
Text prompt required; optional compatible reference media: image (up to 3). Check the current media type and quantity limits before generation, and upload only approved source material.
How long can a Gemini Omni Video video be?
The current video-duration setup is 4s, 6s, 8s, 10s options.
What output controls does Gemini Omni Video currently offer?
Current output options include 720P / 1080P / 4K and 2 aspect-ratio choices. Other available controls include current core settings.
Can I publish a generated result without review?
No. Before public use, manually review motion continuity, text, sound, product details, and any other element that must be accurate.

Updated:

Capability source: Current ImageStyle video-model configuration.

Create with Gemini Omni Video