Gemini Omni is a multimodal video model family. The current live generation workspace is Gemini Omni 1.1 Flash, with four input modes and synchronized native audio.

Example Image
Upload a sharp still — product, portrait, or scene
Upload a sharp still — product, portrait, or scene thumbnailVeo 3 animates motion while identity stays locked thumbnail

Generator examples

Current family status

Gemini Omni 1.1 Flash is live now

Gemini Omni is no longer a placeholder for a future model on Voor. The latest workspace is Gemini Omni 1.1 Flash, with separate text, image, mixed-reference, and existing-video edit modes. The player is a real text-to-video output from that same workspace: ten seconds, 720p, with synchronized native audio.

Latest model
Gemini Omni 1.1 Flash
Generation
3–10 seconds
Resolution
360p to 4K
Audio
Generated with picture
Actual Gemini Omni 1.1 Flash output · text to video · native audio

Version map

One family hub, one current generator

NameStatus on VoorUse now
Gemini Omni 1.1 FlashLive · currentText, image, reference, and edit workflows
Gemini Omni Flash previewPrevious family labelMove new work to 1.1 Flash

Existing briefs still transfer: subject, action, camera, sound, and delivery settings remain the useful structure. What changes is the model target. New work should open the versioned Gemini Omni 1.1 Flash page rather than relying on an older family label.

Four ways into one timeline

Let the source material choose the Gemini Omni mode

These are workflows, not quality tiers. Give the model the least material that still locks what cannot drift, then keep each source assigned to one job.

01

Text to video

Start with
A complete written shot brief
Get
3–10 seconds · 16:9 or 9:16 · native audio

Invent subject, set, camera, action, light, and sound together when no approved visual exists yet.

02

Image to video

Start with
One first frame, optional end frame
Get
Motion anchored to an approved still

Protect a packshot, portrait, illustration, or storyboard composition while directing what happens next.

03

Reference to video

Start with
Up to ten images and three short videos
Get
One clip reasoned across mixed evidence

Assign references to character, wardrobe, location, product, movement, or voice instead of blending a loose mood board.

04

Edit video

Start with
One source clip and one named change
Get
A revised clip at the source duration

Replace, restyle, or revise one visible element while naming every part of the source shot that must remain stable.

Picture and sound are one brief

Native audio changes how a Gemini Omni prompt is reviewed

Write dialogue word for word, then separate ambience, caused effects, and music into their own clauses. Watch once, listen once without looking, and scrub the beginning and end. A beautiful poster frame cannot compensate for a voice change, an effect with no visible cause, or ambience that jumps at the cut.

For a dedicated repair pass, use Lip Sync when dialogue needs replacement. Use AI Video Extender when the job is a merged continuation of an existing clip; the current Gemini Omni edit mode revises a source timeline rather than replacing Voor's dedicated extension workflow.

Choose with evidence

Place Gemini Omni beside the model that challenges it

Gemini Omni models FAQ

What are Gemini Omni models?

Gemini Omni models are multimodal video systems that can reason across text, images, short reference video, audio direction, and existing footage. On Voor, the current live family member is Gemini Omni 1.1 Flash.

What is the latest Gemini Omni model on Voor?

Gemini Omni 1.1 Flash is the latest live Gemini Omni model on Voor. It has separate text-to-video, image-to-video, mixed-reference, and existing-video edit modes with synchronized native audio.

Is Gemini Omni 1.1 Flash available now?

Yes. The Gemini Omni 1.1 Flash generator is live on Voor. Choose a mode, provide the required source material, set duration and resolution, review the credit estimate, and generate.

What can I upload to Gemini Omni models?

The current Gemini Omni 1.1 Flash workflows accept a written brief, a first frame with an optional end frame, up to ten reference images, up to three short reference videos, or one existing clip for a natural-language edit.

Do Gemini Omni models generate audio?

Gemini Omni 1.1 Flash generates synchronized native audio with the picture. Write dialogue exactly and separate ambience, effects, and music into distinct clauses, then review the finished file with sound on.

How long are Gemini Omni 1.1 Flash videos?

The three generation modes expose durations from three to ten seconds. Video edit keeps the source timeline rather than offering a separate duration control.

What resolution does Gemini Omni 1.1 Flash support?

Gemini Omni 1.1 Flash supports 360p, 720p, 1080p, and 4K tiers. Use a lower tier to approve subject, motion, camera, sound, and ending before moving a final brief to a delivery resolution.

How does Gemini Omni 1.1 Flash compare with H3 Max?

H3 Max is a narrower, speed-focused text and image video model. Gemini Omni 1.1 Flash adds mixed references, existing-video editing, and a 4K ceiling. Voor has a real same-prompt comparison page with observed files, time, and credits.

How does Gemini Omni 1.1 Flash compare with Wan 3.0?

Wan 3.0 reaches thirty seconds and accepts broad reference media, while Gemini Omni 1.1 Flash reaches 4K and includes a dedicated existing-video edit mode. Use the same-prompt comparison to choose by the actual workload.

Use the latest Gemini Omni model

Open Gemini Omni 1.1 Flash, choose the mode that matches your source material, and see the live credit estimate before generation.

Create with Gemini Omni 1.1 Flash