Skip to main content

Gemini Omni Flash. One prompt in, a whole scene out

By Steven Van ·

Gemini Omni Flash from Google DeepMind reads text, images, and video together and composes them into one coherent shot, so you build a scene from…

Higgsfield announced Gemini Omni Flash. One prompt in, a whole scene out on June 29, 2026.

What changed

Gemini Omni Flash from Google DeepMind reads text, images, and video together and composes them into one coherent shot, so you build a scene from references instead of describing it from scratch. Consistent characters. Face, outfit, and voice hold across every edit, so you can revise a shot without recasting it. Real-world physics. Gravity, weight, and collisions behave correctly, so motion looks filmed, not simulated. Edit by talking. Change one element of an existing clip in plain language instead of regenerating the whole scene. To get started: Video → model selector → Gemini Omni. Add a prompt, or bring an image, start and end frame, or a clip to edit. Availability: Preview (beta). 720p, clips of 4, 6, 8, or 10 seconds, 16:9 or 9:16, with native audio.

Higgsfield
Higgsfield
AI video and image generation platform — create videos, edit with motion control, swap faces, and collaborate in real time.
View Higgsfield →

Read the original announcement →

Read Gemini Omni Flash. One prompt in, a whole scene out on Creators Toolbox