Gemini Omni
·2026.05.20 02:46
Key point
Gemini Omni introduced multimodal features that edit video and scenes using natural language.
1 / 2
Details
Gemini Omni unveiled multimodal editing features for video that combine Gemini's reasoning and generation capabilities.
- It enables modification and fine-tuning via natural language even in the middle of the generation process.
- Examples were shown of transforming a mirror scene into liquid-like reflections, mirror-textured arms, black-and-white line art, felt dolls, and voxel art.
- Cases were also included where the motion itself of a scene is redesigned, and sound and lighting are synchronized to match scene changes.
- Google DeepMind compared this to Nano Banana for video, emphasizing a flow where generation and editing continue through conversation.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.