List layout
Understand the Gemma 4 model family
Learn about the Gemma 4 family of multimodal open source language models from Google. Google DeepMind’s Maarten Grootendorst walks through all five model sizes across four architectures, from the efficient E2B and E4B models…
🪄 Gemini Live API in action
See what’s new in the Gemini Live API! Thor Schaeff breaks down async function calling, Proactive Audio, real-time context injection with sendClientContent, frontier-level background reasoning in native audio. Products Mentioned: Gemini Live API, Gemini…
What's new in the Gemini Live API
Did you know that Gemini Live API enables low-latency, real-time voice and vision interactions with Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking? Watch along as Google DeepMind’s Thor Schaeff demonstrates: * *Async…
Manage your agents while you’re on the move with the Antigravity Remote Control
The Antigravity Remote Control lets you manage and monitor agents across different machines from a single interface. Launch your long-running agents, monitor them via browser or app, and step in exactly when needed. To…
Celebrating one billion Gemma downloads
Google DeepMind's Paige Bailey catches up with developers at the Gemma 1 billion downloads event to hear how they're using Gemma in their work. Watch along and hear about: * Blind and low vision…
Agentic approaches to processing long videos with Gemini
Save this for your next project with heavy video processing 📌 Agentic video understanding transforms how developers process long-form content across demanding applications, leading to massive token reductions, faster latency, and better accuracy. Products…
Agentic video understanding in Gemini
Processing long videos can use a lot of tokens. Learn how to do this more efficiently using our latest agentic video understanding capability available in Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite. In…
Koray Kavukcuoglu on frontier models, coding agents, and building AGI
Google DeepMind SVP and Chief AI Architect Koray Kavukcuoglu joins host Logan Kilpatrick to reflect on the journey from DeepMind’s early reinforcement learning milestones to Gemini and what it means to build AGI at…
Build voice-first apps with Gemini 3.5 Transcribe
See how Gemini 3.5 Transcribe allows you to build voice-first interfaces and transcribing multi-speaker recordings, delivering fast, contextually accurate transcriptions ready to be deployed in your apps. Subscribe to Google for Developers → https://goo.gle/developers…
How to build with Gemini 3.5 Transcribe
Learn how to build with Gemini 3.5 Transcribe, our latest transcription model, which is now available on the Interactions API and Live API. Google DeepMind’s Thor Schaeff gives a demo of real-time transcription of…