Gemini Omni is the video AI game-changer everyone's been waiting for. I've been testing it the last 24 hours. It's not just another text-to-video model. It's a different category.
01 / What's newIt takes video as input
That's the headline. Most video AI takes a text prompt and renders a clip. Omni takes video, images, audio, and text as direct inputs. You literally talk to your footage and tell it what to change.
I dropped in a single talking-head clip of myself, ran a few different prompts, and watched the model rewrite the scene without re-rendering my face from scratch. The identity stays. The motion stays. The thing you wanted to change, changes.
02 / Why it's differentThe three things that make it special
After a day of pushing it, these are the unlocks that actually matter:
Older video models always glitch out, warp, or render text as unreadable mush. Omni puts text on a moving surface and keeps it sharp the whole way through. That alone changes what you can ship.
03 / The releaseFree for everyone, today
Yesterday this was gated behind a waitlist and a paid Google AI plan. Today Google flipped the switch globally.
You can use it inside Google Flow or directly in the native Gemini app. Same model, two surfaces.
Honestly? I'd try this today.
Open Google Flow, drop in a clip, talk to it. The creative ceiling just moved.
Open Google FlowOr read Google's announcement
Get the next one in your inbox.
I write about the AI tools I'm testing, what just shipped, and what's actually worth your time. Weekly, no fluff.
Subscribe to the newsletterOr join the community of AI builders and creators. DMs, AMAs, and behind-the-build.