By now, you've probably heard about Gemini Omni, our new model designed to create anything from any input, starting with video. But... what's the big deal? Let’s break it down 🧵👇
World understanding: Gemini Omni is built on Gemini's vast knowledge of history, science, and culture, so it can produce videos that are grounded in how the world actually works.
Reference anything: Gemini Omni extends Gemini's native multimodality, allowing you to blend combinations of text, audio, image, and video inputs into a high-quality, consistent video.
Conversational editing: Gemini Omni allows you to edit your videos using natural language (like Nano Banana, but for video). So you can easily change your characters, settings, and styles by just describing what you want.
Video generation gains the edit loop image models already have: change a subject, setting, or style by describing it, instead of rewriting the prompt and regenerating from scratch with different results.
Checking sign-in…
Loading comments…