← Back to video archive

Airdroplet AI summary

The Most Important Google IO Announcements (SUPERCUT)

May 20, 2025Wes RothAI score 10038,739 views

Watch original on YouTube ↗

AI-generated summary

Okay, let's dive into the latest buzz from Google I/O!

This year's Google I/O was absolutely packed with major AI news, showcasing how Google is pushing the boundaries with its Gemini models, bringing advanced AI features to everything from search and creative tools to Android devices and even futuristic glasses and video communication. The key message is that AI is rapidly integrating into everyday products and research, aiming to make technology more helpful and accelerate scientific discovery.

Here are the key things you need to know from the event:

  • New Gemini Models (2.5 Pro & Flash): There are updated versions of the core Gemini models. Gemini 2.5 Pro is highlighted as their most intelligent and the best foundation model globally, especially strong in coding (topping WebDev Arena) and learning (leading on LM Arena) thanks to the integration of LearnLM models. Gemini Flash 2.5 is their efficient, fast, and low-cost workhorse model, which is super popular with developers. The new Flash is better across the board in reasoning, code, and handling long stuff. It's second only to 2.5 Pro on LM Arena, which is really impressive for a "workhorse." Flash is becoming generally available early June, with Pro following soon after.
  • Text-to-Speech Improvements: A cool new text-to-speech feature now supports multiple speakers for two voices using native audio output. This lets the model talk in more expressive ways, capturing subtle speech nuances, and even whispering! It works in over 24 languages and can switch between them seamlessly with the same voice. You can use this in the Gemini API right now, which is pretty awesome for building conversational apps.
  • Gemini Thinking Budgets: They launched Flash with "Thinking Budgets" to give developers control over cost and speed versus output quality, and folks loved it. Now they're bringing this control to 2.5 Pro soon. It lets you decide how many "tokens" the model uses to "think" before giving you an answer, or you can just turn it off. This is a great way to manage performance and cost depending on what you need.
  • Project Mariner (Web Agent): This is a research prototype of an agent that can actually use the web and get tasks done for you. It's seen as combining AI smarts with tools to take actions on your behalf. A key capability is "computer use," letting agents interact with browsers and software. Mariner was shown doing multitasking (up to 10 tasks at once) and learning tasks by being shown once ("Teach and Repeat"). They are bringing Mariner's computer use features to developers via the Gemini API this summer, aiming to build an "agent ecosystem."
  • Agent Ecosystem & Protocols: They're working on building tools for agents to work together. This includes an open agent-to-agent protocol launched with over 60 partners and compatibility with Anthropic's model context protocol (MCP), which helps agents access other services. This feels significant for the future of how AI agents will interact.
  • Jules (Coding Agent): This is an asynchronous coding agent you can give tasks to, and it works on its own to fix bugs or make updates. It hooks up with GitHub and can handle complex coding tasks in big codebases surprisingly quickly – things that used to take hours now take minutes. Jules is now in public beta, which is cool because anyone can sign up and try having an AI coding partner.
  • Gemini Diffusion (Text Diffusion): Google pioneered diffusion models for images and video, and now they're applying it to text with an experimental model called Gemini Diffusion. This model doesn't just generate text left-to-right; it can iterate and correct itself during the process, making it great for tasks like editing math or code. It's also incredibly fast, generating five times quicker than their previous fastest model while matching coding performance. Seeing it solve a math problem almost instantly was pretty wild.
  • Deep Think Mode for 2.5 Pro: Building on the idea that models get better with more thinking time (like AlphaGo), they're introducing a new "Deep Think" mode for 2.5 Pro. This mode pushes the model's performance to its limits using cutting-edge reasoning techniques, including parallel processing. It's showing impressive results on hard benchmarks like competitive math and coding (USAMO, LiveCodebench) and multimodal tasks (MMMU). It's being rolled out carefully to trusted testers first for frontier safety evaluations before wider availability, which makes sense given how powerful it is.
  • AI for Science: A huge area of focus is using AI to speed up scientific discovery. Google DeepMind has made big strides across math and life sciences. They mentioned AlphaProof (math problems), Co-scientists (research collaboration), and AlphaRevolve (discovering knowledge, speeding AI training). In life sciences, AMI (medical diagnosis help), AlphaFold3 (predicting molecular structures/interactions - seen as having massive impact already with millions using it), and Isomorphic Labs (using AlphaFold for drug discovery) were highlighted. The belief is that AGI, if built safely, could be the most beneficial technology ever for accelerating science and solving global diseases.
  • AI Mode for Search: Google Search is getting a total revamp with an "all-new AI mode." This uses more advanced reasoning to handle longer, more complex questions (users are already asking queries 2-3 times longer). You can also ask follow-up questions within the AI answer. It's available as a new tab in Search and is rolling out to everyone in the U.S. starting now. It's described as completely changing how Search is used.
  • Deep Research: For digging into complex topics, Deep Research is getting a key update: you can now upload your own files to guide the research agent. This was a top requested feature and soon you'll be able to pull info from Google Drive and Gmail too.
  • Canvas (Co-creation Space): This is an interactive space where Gemini can help you create things. You can now transform reports or detailed documents into various formats with one tap, like web pages, infographics, quizzes, or even custom podcasts in 45 languages. You can also "vibe code" and collaborate with Gemini to build interactive apps or simulations just by describing what you want. You can share and remix these creations, which seems great for collaboration.
  • Imagen 4 (Image Generation): The latest image generation model is coming to the Gemini app. Imagen 4 is described as a big leap, producing richer images with better colors, details, shadows, and water effects. The presenter feels it's gone from "good to great to stunning" and is also much better at generating text and typography within images, which has been a common challenge for these models.
  • Veo 3 (Video Generation with Audio): Veo 2 redefined video generation, and now Veo 3 is here and available today. The visual quality is even better with stronger understanding of physics. The major leap is native audio generation – Veo 3 can create sound effects, background sounds, and even dialogue based on your prompt, making generated videos incredibly realistic and immersive. Hearing characters speak adds a whole new dimension to video creation.
  • Lyria 2 (Music Generation): Lyria 2 generates high-fidelity music with vocals, solos, and choirs, creating expressive, rich music. It's available now for enterprises, YouTube creators, and musicians.
  • Flow (AI Filmmaking Tool): This new tool combines the best of Veo, Imagen, and Gemini into one place for creatives. It's built to help you get into a creative "zone." You can upload your own images or generate new ones using Imagen right in Flow. You can assemble clips, control camera movements with prompts, and crucially, extend clips or add new shots easily while maintaining character and scene consistency. If something isn't right, you can just trim or edit it like a normal video tool. Once done, you can download and edit in other software, adding music from Lyria. It looks like a powerful way to prototype and create video content quickly with AI help. Flow is launching today.
  • Google AI Subscription Plans: They're upgrading their AI subscription plans. There's Google AI Pro (global availability) with higher rate limits and special features, including the Gemini app version formerly known as Gemini Advanced. Then there's the all-new Google AI Ultra plan (U.S. today, global soon). This is presented as the VIP pass for cutting-edge AI, offering the highest rate limits, earliest access to new features (like Deep Think mode in the Gemini app when ready, and Flow with Veo 3 available today), plus YouTube Premium and lots of storage.
  • AI on Android Devices: Android is called the platform where you see the future first. Gemini is coming soon to the entire Android ecosystem, not just phones. You can already access Gemini via the power button on phones, but it's coming to your watch, car dashboard, and even your TV soon, putting a helpful AI assistant everywhere you are.
  • Android XR & AI Glasses: They're building Android XR, the first Android platform designed for the Gemini era, supporting various devices from headsets to lightweight glasses. They believe people will use different XR devices for different needs throughout the day (immersive headsets for media/work, lightweight glasses for timely info on the go). Android XR is being built with Samsung and optimized for Snapdragon with Qualcomm. They're reimagining Google apps for XR, and mobile/tablet apps also work.
  • Samsung's Project Wuhan (Android XR Headset): This is the first Android XR device from Samsung, available later this year. It offers an infinite screen for apps and integrates Gemini. You can "teleport" in Google Maps, ask Gemini about what you see, and watch things like sports with real-time stats chat.
  • Lightweight AI Glasses: Google has been working on glasses for over 10 years and hasn't stopped. New prototypes with Android XR are lightweight for all-day wear, packed with tech (camera, mics, speakers, optional in-lens display) to give Gemini the ability to see and hear the world. They work with your phone, keeping your hands free. A live demo showed sending texts, muting notifications, identifying objects (coffee shop, photo wall), accessing info (band details, cafe photos), getting directions, translating languages in real-time (Hindi and Farsi demo), and identifying people (Giannis!). They see glasses as a natural form factor for AI, putting Gemini right where you are.
  • Android XR Glasses Development: They're partnering with Samsung to extend Android XR beyond headsets to glasses, creating software and reference hardware. Prototypes are being used by testers, and developers can start building for glasses later this year.
  • Eyewear Partners (Gentle Monster, Warby Parker): To make sure the glasses are stylish and wearable all day, they've announced partnerships with Gentle Monster and Warby Parker as the first eyewear partners building glasses with Android XR. This feels important for getting widespread adoption beyond just tech early adopters.
  • Google Beam (AI-first Video Communication): Building on their Project Starline 3D video tech, Google Beam is a new platform that uses an AI model to turn regular 2D video streams into a realistic 3D experience. It uses six cameras and AI to merge streams and render you on a 3D light field display with near-perfect real-time head tracking. The goal is a much more natural and immersive conversation, making you feel like you're in the same room. Early devices will be available for customers later this year in collaboration with HP.
  • AI for Societal Impact (FireSat, Drone Deliveries): They shared examples of how AI is helping society now. FireSat is a constellation of satellites using AI and multispectral imagery to detect wildfires as small as a one-car garage in near real-time (imagery updated every 20 minutes instead of 12 hours). Wing (drone delivery partner) used AI during Hurricane Helene to deliver critical supplies based on real-time needs. These examples show the practical, life-saving potential of current AI.
  • Inspiration for the Future: The rapid progress towards things like next-gen robots, disease treatments, quantum computers, and fully autonomous vehicles is inspiring. They believe these aren't decades away but years, which is pretty amazing. A personal story about seeing an elderly father amazed by riding in a Waymo highlighted how powerful technology can be to inspire and move us forward.

Video transcript

Open transcript
That's a pretty nice convertible. I think you might have mistaken the garbage truck for a convertible. Is there anything else I can help you with? What's this skinny building doing in my neighborhood? It's a street light, not a building. Why are these palm trees so short? I'm worried about them. They're not short. They're actually pretty tall. Sick convertible. Garbage truck again. Anything else? Why do people keep delivering packages to my lawn? It's not a package. It's a utility box. Why is this person following me wherever I walk? No one's following you. That's just your shadow. Gemini is pretty good at telling you when you're wrong. We are rolling this out to everyone on Android and iOS starting today. Gemini 2.5 Pro is our most intelligent model ever and the best foundation model in the world. Just two weeks ago, we shipped a preview of an updated 2.5 Pro so you could get your hands on it and start building with it right away. You've been really impressed by what you've created, from turning sketches into interactive apps to simulating entire 3D cities. The new 2.5 Pro tops the popular coding leaderboard, WebDev Arena. And now that it incorporates LearnLM, our family of models built with educational experts, 2.5 Pro is also the leading model for learning. And it's number one across all the leaderboards on LM Arena. Gemini Flash is our most efficient workhorse model. It's been incredibly popular with developers who love its speed and low cost. Today, I'm thrilled to announce that we're releasing an updated version of 2.5 Flash. The new Flash is better in nearly every dimension, improving across key benchmarks for reasoning, code, and long context. In fact, it's second only to 2.5 Pro on the LM Arena leaderboard. I'm excited to say that Flash will be generally available in early June with Pro soon after. First, in addition to the new 2.5 Flash that Demis mentioned, we are also introducing new previews for text-to-speech. These now have a first-of-its-kind multi-speaker support for two voices built on native audio output. This means the model can converse in more expressive ways. It can capture the really subtle nuances of how we speak. It can even seamlessly switch to a whisper like this. This works in over 24 languages. And it can easily go between languages. So the model can begin speaking in English, but then... ...and switch back all with the same voice. That's pretty awesome, right? You can use this text-to-speech capability starting today in the Gemini API. Finally, we launched 2.5 Flash with Thinking Budgets to give you control over cost and latency versus quality. And the response was great. So we're bringing Thinking Budgets to 2.5 Pro, which will roll out in the coming weeks, along with our generally available model. With Thinking Budgets, you can have more control over how many tokens the model uses to think before it responds. Or you can simply turn it off. Next, we also have our research prototype, Project Mariner. It's an agent that can interact with the web and get stuff done. Stepping back, we think of agents as systems that combine the intelligence of advanced AI models with access to tools. They can take actions on your behalf and under your control. Computer use is an important agentic capability. It's what enables agents to interact with and operate browsers and other software. Project Mariner was an early step forward in testing computer use capabilities. We released it as an early research prototype in December, and we've made a lot of progress since. First, we are introducing multitasking, and it can now oversee up to 10 simultaneous tasks. Second, it's using a feature called Teach and Repeat. This is where you can show it a task once, and it learns a plan for similar tasks in the future. We are bringing Project Mariner's computer use capabilities to developers via the Gemini API. Trusted testers like Automation Anywhere and UiPath are already starting to build with it, and it will be available more broadly this summer. Computer use is part of a broader set of tools we will need to build for an agent ecosystem to flourish, like our open agent-to-agent protocol so that agents can talk to each other. We launched this at Cloud Next with the support of over 60 technology partners and hope to see that number grow. Then there is the model context protocol introduced by Anthropic so agents can access other services. And today we are excited to announce that our Gemini SDK is now compatible with MCP tools. And our asynchronous coding agent, Jules. Just submit a task, and Jules takes care of the rest, fixing bugs, making updates. It integrates with GitHub and works on its own. Jules can tackle complex tasks in large code bases that used to take hours, like updating an older version of Node.js. It can plan the steps, modify files, and more in minutes. So today I'm delighted to announce that Jules is now in public beta, so anyone can sign up at Jules.Google. And like Demis said, we're always innovating on new approaches to improve our models, including making them more efficient and performant. We first revolutionized image and video generation by pioneering diffusion techniques. A diffusion model learns to generate outputs by refining noise step by step. Today, we're bringing the power of diffusion to text with our newest research model. This helps it excel at tasks like editing, including in the context of math and code. Because it doesn't just generate left to right, it can iterate on a solution very quickly and error correct during the generation process. Gemini diffusion is a state-of-the-art experimental text diffusion model that leverages this parallel generation to achieve extremely low latency. For example, the version of Gemini diffusion we're releasing today generates five times faster than even 2.0 Flashlight, our fastest model so far, while matching its coding performance. So take this math example. Ready? Go. If you blinked, you missed it. Now earlier, we sped things up. But this time, we're going to slow it down a little bit. Pretty cool to see the process of how the model gets to the answer of 39. This model is currently testing with a small group. Thanks, Tulsi. We've been busy exploring the frontiers of thinking capabilities in Gemini 2.5. As we know from our experience with AlphaGo, responses improve when we give these models more time to think. Today, we're making 2.5 Pro even better by introducing a new mode we're calling Deep Think. Deep Think. It pushes model performance to its limits, delivering groundbreaking results. Deep Think uses our latest cutting-edge research in thinking and reasoning, including parallel techniques. So far, we've seen incredible performance. It gets an impressive score on USAMO 2025, currently one of the hardest math benchmarks. It leads on LiveCodebench, a difficult benchmark for competition-level coding. And since Gemini has been natively multimodal from the start, it's no surprise that it also excels on the main benchmark measuring this, MMMU. Because we're defining the frontier with 2.5 Pro Deep Think, we're taking a little bit of extra time to conduct more frontier safety evaluations and get further input from safety experts. As part of that, we're going to make it available to trusted testers via the Gemini API to get their feedback before making it widely available. My entire career at its core has been about using AI to advance knowledge and accelerate scientific discovery. At Google DeepMind, we've been applying AI across almost every branch of science for a long time. In just the past year, we've made some huge breakthroughs in a wide range of areas, from mathematics to life sciences. We've built AlphaProof that can solve math Olympiad problems at the silver medal level. Co-scientists that can collaborate with researchers, helping them develop and test novel hypotheses. And we've just released AlphaRevolve, which can discover new scientific knowledge and speed up AI training itself. In the life sciences, we've built AMI, a research system that could help clinicians with medical diagnoses. AlphaFold3, which can predict the structure and interactions of all of life's molecules. And Isomorphic Labs, which builds on our AlphaFold work to revolutionize the drug discovery process with AI. and will one day help to solve many global diseases. In just a few short years, AlphaFold has already had a massive impact in the scientific community. It's become a standard tool for biology and medical research, with over 2.5 million researchers worldwide using it in their critical work. As we continue to make progress towards AGI, I've always believed, if done safely and responsibly, it has the potential to accelerate scientific discovery and be the most beneficial technology ever invented. Taking a step back, it's amazing to me that even just a few years ago, the frontier technology you're seeing today would have seen nothing short of magic. It's exciting to see these technologies powering new experiences in products like Search and Gemini, and also coming together to help people in their daily lives. For example, we recently partnered with Aira, a company that assists people in the blind and low vision community to navigate the world by connecting them via video to human visual interpreters. Using Astra technology, we build a prototype to help more people have access to this type of assistance. We're getting ongoing feedback from users while Aira's interpreters are actively supervising for safety and reliability. We are introducing an all-new AI mode. It's a total reimagining of Search. With more advanced reasoning, you can ask AI mode longer and more complex queries like this. In fact, users have been asking much longer queries, two to three times the length of traditional searches. And you can go further with follow-up questions. All of this is available today as a new tab right in Search. I've been using it a lot, and it's completely changed how I use Search. And I'm excited to share that AI mode is coming to everyone in the U.S. starting today. Sometimes you need to go deep, unravel something complex. This is where deep research comes in. Starting today, deep research will now let you upload your own files to guide the research agent, which is one of the top requested features. And soon, we'll let you research across Google Drive and Gmail, so you can easily pull out information from there, too. So let's say you have this incredible detailed report. In this case, it's about the science of comets moving throughout space. How do you get all that brilliance to still down into something digestible, engaging, something you can share? To still down into something digestible, engaging, something you can share? This is where Canvas comes in. It's Gemini's interactive space for co-creation. Canvas will now let you transform that report with one tap into all kinds of new things, like a dynamic web page, an infographic, a helpful quiz, even a custom podcast in 45 languages. But if you want to go further, you can vibe code all sorts of amazing things in Canvas with as much back and forth as you want. You can get exactly the experience you're looking for. Check out this interactive comment simulation that one of our Googlers made just by describing what they wanted to build and collaborating with Gemini to get it just right. And as you can share apps like this, others can easily jump in and view it and modify it and remix it. Starting today, we're bringing our latest and most capable image generation model into the Gemini app. It's called Imagine 4, and it's a big leap forward. The images are richer with more nuanced colors and fine-grained details. The shadows in the different shots, the water droplets that come through in the photos. I've spent a lot of time around these models, and I can say, this model and the progression has gone from good to great to stunning. And Imagine 4 is so much better at text and topography. Images are incredible, but sometimes you need motion and sound to tell the whole story. Last December, VO2 came out, and it redefined video generation for the industry. And if you saw Demis' sizzling onions post yesterday, you know that we've been cooking something else. Today, I'm excited to announce our new state-of-the-art model, VO3. And like a lot of other things you've heard about from Stage Today, it's available today. The visual quality is even better. Its understanding of physics is stronger. But here's the leap forward. VO3 comes with native audio generation. That means, that means that VO3 can generate sound effects, background sounds, and dialogue. Now you prompt it, and your characters can speak. Here's a wise old owl and a nervous young badger in the forest. Take a listen. They left behind a ball today. It bounced higher than I can jump. What manner of magic is that? Pretty cool, right? VO added not just the sounds of the forest, but also the dialogue. We're entering a new era of creation with combined audio and video generation that's incredibly realistic. The quality is so good, it feels like you're there on the boat with this guy. This ocean, it's a force, a wild, untamed might. And she commands your awe with every breaking light. The photorealistic generation, the emotion, the movement of his mouth and the ocean in the background. It's incredible how fast VO continues to evolve as a powerful creative tool. We recently launched Lyria 2, which can generate high-fidelity music and professional-grade audio. The music is melodious with vocals and solos and choirs. As you hear, it makes expressive and rich music. Lyria 2 is available today for enterprises, YouTube creators and musicians. We've been building a new AI filmmaking tool for creatives, one that combines the best of VO, Imagine and Gemini, a tool built for creatives by creatives. It's inspired by that magical feeling you get when you get lost in the creative zone and time slows down. We're calling it Flow, and it's launching today. Let me show you how it works. Let's drop into a project I'm working on. Our hero, the grandpa, is building a flying car with help from a feathered friend. These are my ingredients, the old man and his car. We make it easy to upload your own images into the tool, or you can generate them on the fly using Imagine, which is built right in. We can create a custom gold gear shift just by describing it. There it is. Pretty cool. Next, you can start to assemble all of those clips together. With a single prompt, you can describe what you want, including very precise camera controls. Flow puts everything in place, and I can keep iterating in the scene builder. Now, here's where it gets really exciting. If I want to capture the next shot of the scene, I can just hit the plus icon to create the next shot. I can describe what I want to happen next, like adding a 10-foot-tall chicken in the back seat, and Flow will do the rest. The character consistency, the scene consistency, it just works. And if something isn't, oh, quite right, no problem. You can just go back in like any other video tool and trim it up if it's not working for you. But Flow works in the other direction as well. It lets you extend a clip too, so I can get the perfect ending that I've been working towards. Once I've got all the clips I need, I can download the files, I can bring them into my favorite editing software, add some music from Lyria, and now the old man finally has his flying car. So I'm excited to share that we're upgrading two AI subscription plans today. We will have Google AI Pro and the all-new Google AI Ultra. With the Pro plan, which is going to be available globally, you'll get a full suite of AI products with higher rate limits and special features compared to the free version. This includes the Pro version of the Gemini app that was formerly known as Gemini Advanced. Then there's the Ultra plan. It's for the trailblazers. The pioneers. Those of you who want cutting-edge AI from Google. The plan comes with the highest rate limits, the earliest access to new features and products from across Google. It's available in the U.S. today and will be rolling it out globally soon. You can think of this Ultra plan as your VIP pass for Google AI. So if you're an Ultra subscriber, you'll get huge rate limits and access to that 2.5 Pro deep think mode in the Gemini app when it's ready. You'll get access to Flow with VO3 available today. And it also comes with YouTube Premium and a massive amount of storage. There's so many exciting things happening in Android right now. It's the platform where you see the future first. Just last week at the Android show, we unveiled a bold new design and major updates to Android 16 and Wear OS 6. And of course, Android is the best place to experience AI. Many of the Gemini breakthroughs you saw today are coming soon to Android. You can already access Gemini instantly from the Power button. It understands your context and is ready to help. But Android is powering more than your phone. It's an entire ecosystem of devices. In the coming months, we're bringing Gemini to your watch, your car's dashboard, even your TV. So wherever you are, you have a helpful AI assistant to make your life easier. But what about emerging form factors that could let you experience an AI assistant in new ways? That's exactly why we're building Android XR. It's the first Android platform built in the Gemini era. And it supports a broad spectrum of devices for different use cases, from headsets to glasses and everything in between. We believe there's not a one-size-fits-all for XR and you'll use different devices throughout your day. For example, for watching movies, playing games, or getting work done, you'll want an immersive headset. But when you're on the go, you'll want lightweight glasses that can give you timely information without reaching for your phone. We built Android XR together as one team with Samsung and optimized it for Snapdragon with Qualcomm. Since releasing the Android XR developer preview last year, hundreds of developers are building for the platform. We're also reimagining all your favorite Google apps for XR. And it's Android, after all. So your mobile and tablet apps work too. This is Samsung's Project Wuhan, the first Android XR device. Wuhan gives you an infinite screen to explore your apps with Gemini by your side. With Google Maps in XR, you can teleport anywhere in the world simply by asking Gemini to take you there. you can talk with your AI assistant about anything you see and have it pull up videos and websites about what you're exploring. So many of us dream about sitting front row to watch our favorite team. Imagine watching them play in the MLB app as if you were right there in the stadium while chatting with Gemini about player and game stats. Samsung's Project Wuhan will be available for purchase later this year. Now, let's turn our attention to glasses. As you know, we've been building glasses for over 10 years and we've never stopped. Glasses with Android XR are lightweight and designed for all-day wear even though they're packed with technology. A camera and microphones give Gemini the ability to see and hear the world. Speakers let you listen to the AI, play music, or take calls. And an optional in-lens display privately shows you helpful information just when you need it. These glasses work with your phone giving you access to your apps while keeping your hands free. All this makes glasses a natural form factor for AI bringing the power of Gemini right to where you are. So, unlike Clark Kent you can get superpowers when you put your glasses on. Nishta is back there to show us how these glasses work for real. Let me send her a text now and let's get started. Hey everyone. Right now you should be seeing exactly what I'm seeing through the lens of my Android XR glasses. Like my delicious coffee over here and that text from Shuram that just came in. Let's see what he said. All right. It's definitely show time so I'm going to launch Gemini and get us going. Send Shuram a text that I'm getting started and silence my notifications please. Okay. I've sent that message to him and muted all your notifications. Oh, hey Nishta. Hey Dieter. I see the light's on on your glasses so I think it's safe to say that we're live right now. Yes. We're officially on with the IO crew. Hey everybody. It is pretty great to see IO from this angle. Nishta, you promised me I could get my own pair of Android XR glasses if I helped out back here. So what do you say? Of course. Let's get coffee after this and I'll bring you those glasses. Awesome. We'll see you then. Good luck. Thank you. As you all can see there's a ton going on backstage and is that pro basketball player Giannis wearing our glasses? I love them. It's placed on both of my hands for double high fives. Nice. Let me keep showing you guys what these glasses can do. I've been curious about this photo wall all day. Like what band is this and how are they connected to this place? Yeah. Shoreline amphitheater which are often seen as homecoming shows for the band. No way. Can you show me a photo of one of their performances here? Sure. Here's one. Want me to play one of their songs? I'd love that. I can listen while I make my way to the stage. Great. Here's Under the Aurora by Counting Crows. Welcome, mister. Hey everyone. Thanks for that star-studded behind the scenes look. By the way, do you want to book that coffee with Dieter now? Yes. The crew actually gave me some awesome coffee backstage so let me try something fun. Gemini, what was the name of the coffee shop on the cup I had earlier? Hmm, that might have been Bloomsgiving. From what I can tell, it's a vibrant coffee shop on Castro Street. Great memory. Can you show me the photos of that cafe? I want to check out the vibes. Definitely. Do these photos from Maps help? Oh, I know that spot. It's a flower shop as well as a coffee shop but it is downtown. Hmm, okay. Gemini, show me what it would take to walk here. Getting those directions now. It'll take you about an hour. Okay. I can get some steps in and these heads up directions and a full 3D map should make it super easy. Go ahead and send Dieter an invite for that cafe and get coffee at 3 p.m. today. I'll send out that invite now. This is a very risky demo but we're going to give it a shot. Nishter and I are going to speak to each other in our mother tongues. Nishter is going to speak Hindi. I'm going to speak Farsi very poorly and you'll see the feed from both of our glasses back here and so you can all follow along. We'll show an English translation in real time. Okay? Let's give it a shot. Fingers crossed. How can you actually perform «Zescalia or anything» that's funny. from Paris to English. Yes, absolutely. My whole world is open. And I can understand everything. It's easy. Can you talk to all the people of the world? And forgive me for my mom and dad? My French is very bad. Let's see. We said it's a risky demo. We're so excited about the possibilities when you have an incredibly helpful AI assistant by your side with these Android XR devices. But that's not all. We're taking our partnership with Samsung to the next level by extending Android XR beyond headsets to glasses. We're creating the software and reference hardware platform to enable the ecosystem to build great glasses alongside us. Our glasses prototypes are already being used by trusted testers. And you'll be able to start developing for glasses later this year. Now, we know that these need to be stylish glasses that you'll want to wear all day. That's why I'm excited to announce today that Gentle Monster and Warby Parker will be the first eyewear partners to build glasses with Android XR. The opportunity with AI is truly as big as it gets. And it will be up to this wave of developers, technology builders to make sure its benefits reach as many people as possible. I want to share three examples of how research is transforming our products today. Project Starline, Astra, and Mariner. We debuted Project Starline, our breakthrough 3D video technology at I.O. a few years back. The goal was to create a feeling of being in the same room as someone, even if you were far apart. We've continued to make technical advances. And today, we are ready to announce our next chapter, introducing Google Beam, a new AI-first video communications platform. Beam uses a new state-of-the-art video model to transform 2D video streams into a realistic 3D experience. Behind the scenes, an array of six cameras captures you from different angles. And with AI, we can merge these video streams together and render you on a 3D light field display. With near-perfect head tracking down to the millimeter and at 60 frames per second, all in real time. The result, a much more natural and deeply immersive conversational experience. We are so excited to bring this technology to others. In collaboration with HP, the first Google Beam devices will be available for early customers later this year. HP will have a lot more to share a few weeks from now. Stay tuned. I want to leave you with a few examples that inspire me. The first is top of mind for those who live here in California and so many places around the world. So many of us know someone who has been affected by wildfires. They can start suddenly and grow out of control in a matter of minutes. Speed and precision can make all the difference. Together, with an amazing group of partners, we are building something called FireSat. It's a constellation of satellites that use multispectral satellite imagery and AI aiming to provide near real-time insights. Just look at the resolution. It can detect fires as small as 270 square feet, about the size of a one-car garage. Our first satellite is in orbit now. When fully operational, imagery will be updated with a much greater frequency, down from every 12 hours today to every 20 minutes. Speed is also of the essence in other kinds of emergencies. During Hurricane Helene, Wing, in partnership with Walmart and the Red Cross, provided relief efforts with drone deliveries. Supported by AI, we were able to deliver critical items like food and medicine to a YMCA shelter in North Carolina based on real-time needs. We can imagine how this could be helpful in disaster relief in other communities, and we are actively working to scale up. These are examples of ways AI is helping society right now. It's especially inspiring to think about the research of today that will become reality in a few short years. Whether it's building the next generation of helpful robots, finding treatments for the world's deadliest diseases, advancing error-corrected quantum computers, or delivering fully autonomous vehicles that can safely bring you anywhere you want to go. All of this is very much possible within not decades, but years. It's amazing. This opportunity to improve lives is not something I take for granted. And the recent experience brought that home for me. I was in San Francisco with my parents. The first thing they wanted to do was ride in a Waymo, like a lot of other tourists. I had taken Waymos before, but watching my father, who's in his 80s, in the front seat, be totally amazed. I saw the progress in a whole new light. It was a reminder of how incredible the power of technology is to inspire, to awe, and to move us forward. And I can't wait to see what amazing things will build together next. Thank you.