Google’s Omni model has generated a lot of interest in the AI community. It can generate videos from anything and edits them too. What’s neat is you don’t need to over-explain everything. You can use Gemini’s understanding of the world to create outputs that look realistic. You can use a prompt like this to generate a fun video:
[The video shows items of the alphabet. An unusual item starting with each letter is shown sitting on a table (like a Capybara for C, disco globe for D and Lava Lamp for L). All 26 letters must be represented by 26 items with matching lower thirds displaying the letter. Only one item and lower third at a time. Each lower third must look like a black marker written on a slip of paper in the bottom left. Rapid fire, roughly 9 frames per item at 24FPS. Last frame is a slip of paper “THE END.” The whole video is accompanied by calm smooth music]

Omni is great at text rendering, so you can specify typography, placement, and animation styles. It is also possible to define shots, angles, and camera movements. Omni can edit your videos like a pro. You can change a background, swap a caption, and a lot more.
— Google AI (@GoogleAI) May 26, 2026
[HT]

