HomeAIGoogle DeepMind Announces Gemini Omni
AI

Google DeepMind Announces Gemini Omni

Google DeepMind introduced Gemini Omni Flash, a multimodal model that generates and edits videos from text, image, audio, and video inputs.

WHAT YOU NEED TO KNOW
  • Google DeepMind launched Gemini Omni Flash, capable of generating video from text, image, video, and audio inputs.
  • The model supports conversational multi-turn video editing while preserving scene physics and character consistency.
  • Videos generated by the model include imperceptible SynthID digital watermarks for verification.
  • Gemini Omni Flash is available now for paid subscribers and is rolling out to YouTube Shorts this week.

Google DeepMind announced Gemini Omni, a new natively multimodal model family that generates and edits high-quality video using text, image, audio, and video inputs. The company introduced the first model in the lineup, Gemini Omni Flash, at Google I/O 2026.

According to reporting from Google DeepMind, Gemini Omni Flash allows video editing through natural language prompts across multiple turns. The system maintains character consistency, physics, and scene context as instructions build on previous steps.

Input capabilities and reasoning

The model combines visual rendering with reasoning about real-world physics, including gravity, kinetic energy, and fluid dynamics. It processes reference combinations to generate cohesive outputs and educational explainers. While initial audio inputs are restricted to voice, Google DeepMind plans to roll out additional audio input types alongside future output modalities for images and standalone audio.

Watermarking and distribution

All videos generated by Omni include an imperceptible SynthID digital watermark. Content origin can be verified through the Gemini app, Chrome, and Google Search. Google DeepMind restricted initial speech modification tools to personal digital avatars while testing continues on broader audio editing features.

Availability begins today for Google AI Plus, Pro, and Ultra subscribers globally through the Gemini app and Google Flow. Google DeepMind is also deploying the model at no cost on YouTube Shorts and the YouTube Create App this week, with developer and enterprise API access following in the coming weeks.

Xentir Media
Xentir Media NewsroomSource-backed AI and technology coverage, drafted by Xentir's automated editorial system under fixed human-set rules. See our editorial policy and AI usage policy.
J
Jomon · Founder & EditorFounder and editor of Xentir Media. Sets the editorial rules the newsroom system runs under, and is accountable for its corrections. About Jomon · [email protected]
The Xentir Brief
The developments worth knowing — one useful email.
Get the Brief →