AI video model
Gemini Omni Video AI Video Generator
Choose Gemini Omni Video when you need a Gemini AI video generator where a detailed written brief remains the source of truth even with optional images, and the delivery fits a fixed 4, 6, 8, or 10 second landscape or portrait clip.
Gemini Omni Video task directions
These local references help turn the current Gemini Omni Video input and control contract into a concrete task direction. They are not presented as output from this model.
Every card identifies whether it is a verified model-family example or a visual reference. Reference cards are not model benchmarks or before-and-after claims.
- Input contract
- Text prompt required; optional compatible reference media: image (up to 3)
- Current controls
- Prompt length: Up to 20,000 characters · Duration: 4s, 6s, 8s, 10s options · Output resolution: 720P / 1080P / 4K · Framing: 2 aspect-ratio choices
- Selection boundary
- A prompt remains required when images are attached. Compare another model when the task requires free-form duration, a square or other frame, driving audio, video references, or audio-generation controls.
Action beat
State one subject action and the moment the clip should begin from.
- Task direction
- State one subject action and the moment the clip should begin from.
Reference direction, not a model output
Local ImageStyle visual reference. It is not presented as a model-specific result.
Camera direction
Name the camera behavior and the visual change it should reveal over time.
- Task direction
- Name the camera behavior and the visual change it should reveal over time.
Reference direction, not a model output
Local ImageStyle visual reference. It is not presented as a model-specific result.
Ending state
Define the final visual state, without treating this reference as a result from the current model.
- Task direction
- Define the final visual state, without treating this reference as a result from the current model.
Reference direction, not a model output
Local ImageStyle visual reference. It is not presented as a model-specific result.
- Create
- AI video
- Input
- Text prompt required; optional compatible reference media: image (up to 3)
- Duration
- 4s, 6s, 8s, 10s options
Compare available settings with related models
Check whether the input, output, and framing controls fit the job before you open a workspace. This table does not treat the example media as a cross-model performance test.
Gemini Omni Video
Current page
Long prompt-led video briefs with optional images and fixed 4, 6, 8, or 10 seconds.
- Input
- Text prompt required; optional compatible reference media: image (up to 3)
- Prompt length
- Up to 20,000 characters
- Duration
- 4s, 6s, 8s, 10s options
Seedance 2 Mini
Short multimodal clips with text, media references, audio generation, and web search.
- Input
- Text prompt or compatible reference media: image (up to 3), video (up to 3), audio (up to 3)
- Prompt length
- Up to 20,000 characters
- Duration
- 4–15 seconds
Seedance 2 Fast
The Fast-labelled multimodal workflow with a 15-second default clip.
- Input
- Text prompt or compatible reference media: image (up to 5), video (up to 3), audio (up to 3)
- Prompt length
- Up to 20,000 characters
- Duration
- 4–15 seconds
Seedance 2
Multimodal video work with explicit 480p through 4K output choices.
- Input
- Text prompt or compatible reference media: image (up to 5), video (up to 3), audio (up to 3)
- Prompt length
- Up to 20,000 characters
- Duration
- 4–15 seconds
What you can control with Gemini Omni Video
Keep the prompt as the source of truth
Gemini Omni Video currently supports a 20,000-character prompt with optional images, 720p through 4K choices, fixed 4, 6, 8, or 10 second clips, and 16:9 or 9:16 framing.
Add compatible media when it helps
Text prompt required; optional compatible reference media: image (up to 3). Only upload media you are allowed to use for the current project.
Confirm the clip settings
Current clip settings include 4s, 6s, 8s, 10s options, 720P / 1080P / 4K, and 2 aspect-ratio choices.
Current Gemini Omni Video settings
These settings come from the current ImageStyle model configuration and help you decide whether the workflow fits the job.
| Setting | Current availability |
|---|---|
| Input | Text prompt required; optional compatible reference media: image (up to 3) |
| Prompt length | Up to 20,000 characters |
| Duration | 4s, 6s, 8s, 10s options |
| Output resolution | 720P / 1080P / 4K |
| Framing | 2 aspect-ratio choices |
| Additional controls | Current core settings |
AI video model
When to choose Gemini Omni Video
Long prompt-led video briefs with optional images and fixed 4, 6, 8, or 10 seconds.
Current workflow fit
Choose Gemini Omni Video when you need a Gemini AI video generator where a detailed written brief remains the source of truth even with optional images, and the delivery fits a fixed 4, 6, 8, or 10 second landscape or portrait clip.
When to compare another model
A prompt remains required when images are attached. Compare another model when the task requires free-form duration, a square or other frame, driving audio, video references, or audio-generation controls.
How to create a video with Gemini Omni Video
Write the complete scene and movement brief first, then use optional images only to anchor details that the text needs to preserve.
- 1
Write the complete timed brief
Choose the fixed duration and orientation, then state the scene, action, camera, and required details before attaching optional images.
- 2
Add reference media when needed
Add approved compatible media only when an existing frame, motion, or sound should guide the result.
- 3
Confirm duration and output controls
Confirm the current 4s, 6s, 8s, 10s options, 720P / 1080P / 4K, and 2 aspect-ratio choices in the workspace before generation.
- 4
Generate and review continuity
Before public use, inspect subject motion, camera changes, visible text, sound, and the final frame.
Checks before you create and publish
- Treat the workspace as current
- Model availability and settings can change. Use the options shown in the current workspace before you submit.
- Use approved source media
- When you upload images, video, or audio, confirm that you are allowed to use and publish it for the current project.
- Review the generated video
- Before publishing, manually check motion continuity, text, sound, people, product details, and marks.
Gemini Omni Video FAQ
- Is a prompt still required with images in Gemini Omni Video?
- Yes. The current workflow requires a written prompt even when optional image references are attached.
- Can I start Gemini Omni Video from reference media?
- Text prompt required; optional compatible reference media: image (up to 3). Check the current media type and quantity limits before generation, and upload only approved source material.
- How long can a Gemini Omni Video video be?
- The current video-duration setup is 4s, 6s, 8s, 10s options.
- What output controls does Gemini Omni Video currently offer?
- Current output options include 720P / 1080P / 4K and 2 aspect-ratio choices. Other available controls include current core settings.
- Can I publish a generated result without review?
- No. Before public use, manually review motion continuity, text, sound, product details, and any other element that must be accurate.
Updated:
Capability source: Current ImageStyle video-model configuration.