At its I/O 2026 keynote, Google unveiled Gemini Omni, a multimodal flagship model capable of generating and editing high-quality videos from text, image, audio, and video inputs. Google also debuted Gemini 3.5 Flash, a model optimized for autonomous coding and complex long-horizon task execution. Gemini Omni Flash is rolling out globally to Google AI Plus, Pro, and Ultra subscribers while launching at no cost on YouTube Shorts.
Multimodal Video Generation and Gemini 3.5 Flash
Gemini Omni introduces conversational video creation and editing grounded in real-world knowledge. The Omni family debuts with Gemini Omni Flash, available through the Gemini app, Google Flow, and integrated directly into YouTube Create and YouTube Shorts.
Alongside Omni, Google introduced Gemini 3.5 Flash to power agentic workflows and software development. Gemini 3.5 Flash is generally available in Google Antigravity, the Gemini API via Google AI Studio and Android Studio, Gemini Enterprise, and AI Mode in Search. Google confirmed that a larger Gemini 3.5 Pro model is currently in internal testing with plans for release in the following month.
Autonomous Agents and Generative UI in Search
Google is expanding autonomous functionality across Search and its app ecosystem with dedicated agentic capabilities:
- Information Agents: Background agents in Search that monitor blogs, news, social platforms, and live financial or sports data to send users personalized updates.
- Google Antigravity in Search: Dynamic UI generation that codes mini-apps, trackers, and custom interactive tools on the fly within Search results.
- Daily Brief: An opt-in Gemini app feature that aggregates updates across Gmail and Google Calendar to present prioritized morning summaries and recommended next steps.
- Gemini Spark: A 24/7 cloud-based agent integrated across Google Workspace tools that performs background workflows and explicitly requests approval before executing high-stakes actions.
Content Watermarking and Hardware Announcements
Google announced expanded adoption of its SynthID watermarking technology, which has already tagged over 100 billion images and videos. Third-party platforms including OpenAI, Kakao, and ElevenLabs are implementing SynthID, while Google Cloud is launching an AI content detection API for enterprise customers. Native Content Credentials are also expanding from Pixel camera photo captures to video formats.
On the hardware and platform side, Google showcased Android XR intelligent eyewear—including audio glasses launching in the fall—and previewed a redesigned macOS Gemini app featuring screen-aware voice editing.