Model ReleasesJun 3, 2026
Google introduces Gemini Omni for multimodal video generation and editing
Google DeepMind has launched Gemini Omni, a multimodal AI model that generates and edits video through text, image, and audio inputs. The release introduces natural-language video manipulation while enforcing SynthID watermarking, though general speech-editing features remain restricted pending safety tests.