
@GoogleAI
Making AI helpful for everyone. Show thinking ↓
It’s time for our end-of-week recap 👇 — Gemini 3.8 Live and 3.8 Live Extended Thinking, our most advanced live dialogue audio models yet — Dreambeans, an experiment from @GoogleLabs that curates a daily personalized collection of stories, is now GA — CC from @GoogleLabs has expanded from a personal productivity tool into a shared agent, designed to help families and households coordinate logistics, schedules, and daily tasks — Google Pics, a new @GoogleWorkspace tool that lets you generate, refine, and co-create images, is now GA — AlphaGenome Atlas, @GoogleDeepMind's new interactive platform for genomics discovery
Introducing our most advanced Gemini Audio models yet 🗣 Gemini 3.8 Live and 3.8 Live Extended Thinking let you speak, collaborate, and execute tasks seamlessly, meaning conversing with AI just got a lot more natural. So, what’s the difference between these two models? Let’s break it down: — Gemini 3.8 Live is built for scale, speed, and cost efficiency. It can handle mid-sentence interruptions, transitions across 97 languages on the fly, and understands visual context. Figure out how to fix a broken bike chain, or deal with a leaky pipe just by pointing your camera at the problem area in Search Live for step-by-step audio instructions. — Gemini 3.8 Live Extended Thinking goes one step further to bring increased intelligence to your most complex tasks. It reasons and speaks in parallel, even narrating its progress as it works. This lets it handle multi-step, behind-the-scenes projects, like planning an event, without ever losing the conversational flow. Watch how Gemini 3.8 Live combines real-time video and voice inputs in Search Live to tackle hands-on DIY plumbing tasks step by step 👇
Gemini 3.8 Live is rolling out to: — Consumers: in Search Live — Developers: in public preview in the Gemini API via @googleaistudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience) Gemini 3.8 Live Extended Thinking is rolling out to: — Consumers: in Gemini Live in the @GeminiApp, plus Google AI Pro and Ultra subscribers in @GoogleWorkspace in @GoogleDocs, and all Google AI subscribers in @gmail and Keep — Developers: in public preview in the Gemini API via @GoogleAIStudio — Enterprises: in private preview via Gemini Enterprise (coming soon to Gemini Enterprise Customer Experience and @GoogleWorkspace business customers) t.co/vVruOamq0A
How do you build personalized AI tech that truly understands what matters to you? With Dreambeans from @GoogleLabs, we're safely connecting the dots across your digital life. Instead of siloing pieces of data one app at a time, Dreambeans gathers details from connected sources like @Gmail, Calendar, Search, @GeminiApp, and through face grouping in your @googlephotos. It might notice your friend Beth's birthday is coming up, generate a one-of-a-kind illustrated story of the two of you, and surface the perfect gift ideas. Rather than doomscrolling an endless feed, you’ll get a daily in-app notification from Dreambeans when your customized “stories” are ready. And when a story clicks with you, you can tap the illustrated tile to find additional info and direct links to the next step, whether that’s watching a movie trailer or buying a suggested gift. This experience is opt-in, transparent, and protected by strict privacy filters. And because you are always in control, you can simply tap the "thumbs down" button on any story if a suggestion misses the mark, helping the system learn what is actually useful to you. Get started: t.co/XWeE7gbKvS
Oh, so this is why we mapped out all 166,000 of the male fruit fly's neurons. Check out the big community effort to show just how much these tiny fly brains are capable of 🪰🧵
First, the fruit fly found itself trapped inside a Minecraft box.
evnsnclr@evnsnclr·I’ve successfully run the full retained MaleCNS v1.0 fruit fly connectome, all 166,700 neurons, inside Minecraft, with its simulated neural activity driving a fly’s movement. V1 Currently in development. Built with the help of GPT-6 Astra. Props to the @OpenAI team and @thsottiaux for this release. Code and mod coming soon!
Bring your imagination to life with Google Pics. Built with Nano Banana, Pics makes it easy to generate, refine, and co-create images with unparalleled precision. You can: — Isolate and edit specific objects without altering the rest of your image — Modify or translate text directly inside an image — Share and collaborate with friends and teammates on the same creation — Generate multiple options from a single prompt Google Pics is now available to Google AI Pro and Ultra subscribers, and most @GoogleWorkspace business customers. Google Pics is also integrated into @googledocs and Slides, and will roll out to @googledrive in the coming weeks. Try it now at t.co/4ZLfftKTty
Learn more: blog.google/products-and-p…
Our @GoogleResearch Connectomics team, in collaboration with @HHMIJanelia, has released the complete wiring diagram of a male fruit fly’s brain and central nervous system — the largest brain map by number of proofread neurons to date. So, why do we care so much about a tiny fly's brain? 🧵👇
Because it’s a stepping stone to understanding our own. Nervous systems across species share surprising similarities. By using AI to map these smaller organisms (like flies and fish), we are paving the way for research that could one day help treat and repair human neural and cognitive ailments.
Check out this week’s shipping recap: — Gemini 3.8 Flash, our most intelligent workhorse model yet, delivers upgrades across coding, agentic workflows, and critical multi-step reasoning. — Gemini 3.8 Flash Cyber, our most capable cybersecurity model, features frontier-level performance in vulnerability detection and automated patching. — Lyria 3.5, our newest music generation model, is now available via the Gemini API, and across @GoogleAIStudio, @GeminiApp, @FlowbyGoogle, and Google Vids. — WeatherNext 3, our most advanced and accurate global weather AI model, is here from @GoogleDeepMind and @GoogleResearch. — Agentic Video Understanding, our new video analysis feature available via the Gemini API in @GoogleAIStudio and the Gemini Enterprise Agent Platform, improves accuracy while dramatically reducing token usage and costs.
Introducing WeatherNext 3️⃣— our most advanced global weather AI model yet from @GoogleDeepmind and @GoogleResearch With prediction capabilities that are up to 5x sharper than WeatherNext 2, the model generates a forecast with high spatial resolution in order to catch fast-evolving rainstorms, map local temperature shifts, and even help wind farms predict their power output. So, how does it do that? While traditional weather models rely on massive, physics-based supercomputer simulations that can carry a 6-hour forecast lag, WeatherNext 3 leverages live geostationary satellite observations as inputs and trains directly on real-world surface and atmospheric observations. By pulling this raw satellite data, it’s able to update the global forecast every single hour. And because weather develops at lightning speed, these quick, detailed insights can help bring more localized forecasting to billions of people and local businesses, especially in regions that are historically underserved due to the high costs of traditional weather forecasting models.
WeatherNext 3 will begin enhancing weather experiences within Google Search, @GeminiApp, @googlemaps, @GMapsPlatform Weather API and @googleearth Engine starting today Dive into the tech: blog.google/innovation-and…
Identifying security flaws is only half the battle; generating automated, reliable fixes in real time is where defense gets real. Enter Gemini 3.8 Flash Cyber ⚡️🛡 The new release is a major improvement from our 3.5 generation and among our most capable defensive models to date. Built specifically for defenders (like security teams), it acts as an expert partner to help write new code to fix flaws at scale, a process known as patch generation. We’re already using it to secure @Google’s own code, helping our @GoogleChrome team produce 2.6X more correct patches than top commercial models. Even better? It brings a cost advantage by delivering frontier-level performance for a fraction of the price. Given these capabilities, we're providing prioritized access to government authorities, critical infrastructure operators, and software maintainers through our FairwindProgram.
Apply for access today: deepmind.google/fairwind-progr…
We’re introducing Gemini 3.8 Flash ⚡️ built to tackle complex agentic and multi-step tasks with even greater diligence. Our most intelligent workhorse model yet delivers significant improvements in reasoning, evolving to an AI partner that doesn’t just write code, but can also navigate complex projects. While solving ambiguous and high-friction tasks is immensely helpful, it can also be expensive. Fortunately, 3.8 Flash features the usual effort controls, ensuring that the amount of thinking required to accomplish the task at hand is proportional to the tokens spent. Watch Gemini 3.8 Flash combine native video understanding with advanced coding to autonomously build this 3D game, play it to find errors, and execute code changes in a seamless agentic loop in @antigravity.
— Available to Google AI Pro and Ultra subscribers across the @GeminiApp, AI Mode in Google Search, and Google Sheets — Build in the Gemini API via @GoogleAIStudio and @AndroidStudio, explore agent-first workflows in @antigravity, and generate UIs with 3.8 Flash in @stitchbygoogle — Gemini Enterprise users by picking it from the drop down model menu or in Gemini Enterprise Agent Platform. t.co/OzPm3Et10i
Here’s what launched this week: — Gemini 3.5 Transcribe, our most precise speech-to-text model yet, designed to deliver intelligent transcriptions — Gemini Omni 1.1 Flash, bringing expanded creative capabilities and controls for video generation and editing — @GeminiApp’s Live experience, moving beyond conversation to complex tasks with new features like Daily Brief, Gemini Spark, Personal Intelligence, and @Gmail inbox management — Expert Intelligence, a new cross-@Google initiative that lets you engage with and combine insights from trusted sources, starting with eligible @GooglePlay ebooks in @Gemini_Notebook
Meet Gemini Omni 1.1 Flash ⚡️ Our newest multimodal model for video generation and editing. It now features your favorite creative controls from Veo, plus brand new capabilities. Enjoy features like 4K upscaling, first / last frame control, and fast 360p drafting. But, the biggest upgrade? Next-level scene extension. With Omni 1.1 you can extend scenes based on 10 seconds of context from your original video, a big jump from Veo's 1 second! That means tighter consistency, deeper control, and longer, more cohesive storytelling. See it in action ↓
— Upgrades coming to @FlowbyGoogle (scene extension is coming soon!) — Extend scenes in the @GeminiApp (globally rolling out for all Google AI Plus, Pro and Ultra subscribers) — Build directly in @GoogleAIStudio — Deploy on the Gemini Enterprise Agent Platform t.co/NiUxxCh0Mj
— Upgrades coming to @FlowbyGoogle — Extend scenes in the @GeminiApp (globally rolling out for all Google AI Plus, Pro and Ultra subscribers) — Build directly in @GoogleAIStudio — Deploy on the Gemini Enterprise Agent Platform blog.google/innovation-and…
Today we’re introducing Gemini 3.5 Transcribe, our latest transcription model built for incredibly precise, smart dictation across your favorite apps and devices. Remember when traditional speech-to-text meant shouting over background noise, constantly hitting backspace to fix misspelled words, and manually deleting every "um" and "uh"? Those days are over. Gemini 3.5 Transcribe isn't just dictation — it’s active intelligence with precise, context-aware speech-to-text support in 85+ languages. The model automatically filters out filler words, formats unstructured speech, and even pairs with your screen context to execute voice commands. Watch as Gemini 3.5 Transcribe removes filler words and uses multimodal capabilities to seamlessly turn messy voice input and local files into a polished email draft.
— Try it out in the @Geminiapp on macOS and Gboard on @Android — Build in the Gemini API via @googleaistudio, @antigravity, and the Gemini Enterprise Agent Platform (public preview) — Coming soon to @googlechrome and Gemini Enterprise for Customer Experience Learn more: t.co/mGPWSdloR8
It’s (finally) Friday 🎉 Here’s our end-of-week recap: — This year’s @madebygoogle lineup (Pixel 11 series, Pixel Watch 5, and Pixel Tag) brings new AI integrations across devices. A few of the key announcements were Magic Capture for simultaneous video and photo capture, Rambler’s AI voice typing and text transformation, and expanded Live Transcribe for real-time ASL-to-text translation using the Pixel camera. Tying it all together is Gemini Intelligence, our proactive, agentic AI layer to anticipate user needs. — Gemini 3.7 Flash, our most intelligent workhorse model yet for coding and agents, is now available in the Gemini API via @googleaistudio, @antigravity, Spark in the @geminiapp, and the Gemini Enterprise Agent Platform. It’s also rolling out to paid users on @googleworkspace, Gemini App, and Search. — @GoogleDeepMind's WeatherNext 2, an AI forecasting tool that can give meteorologists an extra day of accuracy when predicting a cyclone's track, intensity, and wind structure, is now open source. — The upgraded @Gemini_Notebook experience has been fully rolled out to all Pro users and the ability to copy a notebook has been rolled out to all users.
Our most intelligent workhorse model yet for coding and agents has arrived ⚡ Meet Gemini 3.7 Flash. — Crush that seemingly endless to-do list. Gemini Spark in the @geminiapp now uses 3.7 Flash. The new model can equip your personal AI agent to work even smarter for you by seamlessly handling complex, multi-step tasks across your @GoogleWorkspace apps like @gmail, @googlecalendar and @googledocs — Enjoy a smoother build experience. The model thinks more diligently, putting more effort into multi-step planning and tool calls. A more disciplined execution means less manual oversight and fewer retries across engineering workflows — Build more, spend less. 3.7 Flash is available through the end of the year with an introductory price of half the original 3.6 Flash cost per million tokens ($0.75/1M input tokens and $3.75/1M output tokens)
— Google AI Pro and Ultra subscribers can experience 3.7 Flash today via Spark in the @GeminiApp — Access the model in the Gemini Enterprise Agent Platform and Gemini Enterprise app — Build in the Gemini API via @googleaistudio and @androidstudio, and explore agent-first workflows in @antigravity — Learn more in the blog ↓ t.co/s7Sdj99H32
Tomorrow is the total solar eclipse 🌑☀️ Whether you’re in the path of totality across Greenland, Iceland, and Spain, or watching the partial phases online, we’ve got you covered. The interactive 3D visualizer built in @GoogleAIStudio lets you follow the Moon's shadow across the globe in real-time, simulate local sky views from stations like Reykjavík or Valencia, and fast-forward or rewind to catch key eclipse milestones: t.co/cJyF3EcKxG
You can also check @NASA’s map for your local eclipse timing and exact path details. science.nasa.gov/eclipses/futur…
It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robots — Gemini 3.5 Flash-Lite is our fastest, most cost-effective model optimized for high-speed agentic workflows — Gemini 3.6 Flash delivers faster, more accurate performance across tasks while using significantly fewer tokens — Gemini 3.5 Flash Cyber is our new, cost-effective model optimized for finding and fixing software vulnerabilities at scale, available exclusively to governments and trusted partners — Nano Banana 2 in @GoogleEarth lets you reimagine places and generate custom images using satellite, aerial, and 3D imagery — Lyria 3.5 is the newest music model from Google Deepmind, now powering @googleflowmusic — @Gemini_Notebook (formerly NotebookLM) launched Collections, a new way to organize your notebooks
It’s been a busy couple of weeks! ICYMI, here’s the recap ⬇️ — Gemini Robotics 2 from @GoogleDeepmind brings whole-body intelligence to robots — Gemini 3.5 Flash-Lite is our fastest, most cost-effective model optimized for high-speed agentic workflows — Gemini 3.6 Flash delivers faster, more accurate performance across tasks while using significantly fewer tokens — Gemini 3.5 Flash Cyber is our new, cost-effective model optimized for finding and fixing software vulnerabilities at scale, available exclusively to governments and trusted partners — Nano Banana 2 in @GoogleEarth lets you reimagine places and generate custom images using satellite, aerial, and 3D imagery — Lyria 3.5 is the newest music model from Google Deepmind, now powering @googleflowmusic — @Gemini_Notebook (formerly NotebookLM) launched Collections, a new way to organize your notebooks
What if you could see a city as it looked 100 years ago, or visualize a new basketball court at your local community center? Starting today, we're bringing our latest image generation capabilities to @GoogleEarth on the web. Powered by Nano Banana 2, you can combine rich satellite and 3D imagery with text prompts to reimagine your favorite place around the globe. Just zoom in, tap “create image,” and type whatever you want to see! Available now for everyone to start creating: t.co/FEtpYC43SO
For decades, we’ve dreamed of robots that can seamlessly step into our world and lend a hand. Today, we take a major stride toward making that dream a reality: Introducing Gemini Robotics 2 from @GoogleDeepMind, the intelligence layer powering the next generation of truly adaptable robots. This major advance unlocks intelligent whole-body control, advanced dexterity, and even multi-robot collaboration 🤯. Ok but... how does a robot actually "think"? Real-world tasks take time and planning. To manage that complexity, our new embodied reasoning model, Gemini Robotics ER 2, acts as the robot’s high-level brain, enhancing the robot’s capabilities to: — Observe the environment — Reason about the actions needed to complete the task — Coordinate with the vision-language-action model to carry out actions — Track progress until the job is done This setup allows robots to execute complex multi-step workflows, self-correct if a step fails, and adapt to completely novel situations. Learn more about Gemini Robotics ER 2 (and our two other brand new models) here: t.co/1YEpoYAhww
@GoogleDeepMind Humanoid legs or wheeled rovers? Should robots be cracking eggs? Watch as the @GoogleDeepMind team behind Gemini Robotics 2 shares their behind-the-scenes insights and stories ↓
As AI models are now finding vulnerabilities faster than we can fix them, our approach to securing software must be built on highly efficient and capable models. Which brings us to our third (!) model launch of the day: Gemini 3.5 Flash Cyber ⚡🛡️ Built on top of 3.5 Flash, in CodeMender (our AI agent for code security) it delivers competitive performance at the frontier. on benchmarks like CyberGym and is optimized for finding and fixing cybersecurity vulnerabilities at scale at a lower cost. Given the dual-use nature of this technology, we have taken an intentional approach to its deployment. The model will be available exclusively to governments and trusted partners via CodeMender soon as part of a limited-access pilot program.
Today, we're introducing not one but TWO new models, striking the balance between efficiency and quality to enable you to build production AI agents. — Gemini 3.6 Flash: Addresses efficiency feedback we received from Gemini 3.5 Flash with upgrades in coding, knowledge work, and multimodal tasks faster, more accurately, and with substantially fewer tokens per task — Gemini 3.5 Flash-Lite: Our fastest, most cost-effective 3.5-class model yet built for agentic workflows, hitting ~350 output tokens/sec with improved coding and overall quality Start building with these today via the Gemini API in @GoogleAIStudio or try them out in the @GeminiApp
Step into the map with the Street View grounding feature in Project Genie from @GoogleDeepmind and @GoogleLabs. Announced at I/O, this research prototype uses locations from @GoogleMaps Street View as a foundation, letting you generate and explore interactive, 360-degree virtual environments from just a text prompt or real-world starting place. Cool, right? But… How does it actually work? 🤔 As an experimental tool, Project Genie tackles the "blank space" problem (showing both what’s in front of the camera and behind it) by utilizing Street View data to realistically generate a 360-degree view of the location you selected as the starting point for your world to generate from. Worlds generated by Genie are far more dynamic and rich because they’re created frame-by-frame based on the world description and user actions. By predicting each subsequent frame, Genie is able to simulate what it looks like to swim across an ocean, or hike to the top of a peak, marking a massive shift in interactive media and simulation pipelines. What real-world place would you want to step into and explore?
As generative AI tools continue to evolve, we believe it's more important than ever to know what's AI-generated and what isn't. That’s why @GoogleDeepMind launched SynthID in 2023—a technology that adds a hidden digital watermark to AI content. Here’s a summary of SynthID’s journey and where the provenance technology (the documented history and origin of digital content) is today: — SynthID watermarking was originally built for images, but now supports video, audio, and text. — The technology has watermarked over 100 billion images and videos, alongside 60,000 years of audio. — You can now verify content with SynthID directly in Google Search, Gemini in Chrome, and the @GeminiApp, where it has been utilized over 50 million times. — We’ve also adopted C2PA Content Credentials across a growing number of our generative AI tools. This includes the images and videos created within the Gemini app. So now, in addition to the SynthID watermark, you can also see where an image or video originated and how it’s been altered. — We have open-sourced our text watermarking technology, and we are working with companies like @OpenAI, @NVIDIA, and @Apple to apply SynthID to generative media. Let us know what you think of the tool so far!
We’re shipping two major updates to streamline your creative workflow, allowing you to generate high-speed images with one model and then instantly animate them with the other—all at a fraction of the cost 🍌⚡️ 1️⃣ Introducing Nano Banana 2 Lite: Our fastest and most cost-efficient Gemini Image model yet delivers text-to-image outputs in under 4 seconds. Now available via the Gemini API and Google AI Studio, and rolling out soon across @NotebookLM, @FlowbyGoogle, @geminiapp, @stitchbygoogle, Google Search and @GooglePhotos. 2️⃣ Gemini Omni Flash in Public Preview: Our natively multimodal model for cost-efficient video generation and conversational editing. Now available via the Gemini API, @googleaistudio, and Gemini Enterprise Agent Platform so you can integrate the model into your workflow. While exciting on their own, the real magic happens when you build using these models together. Watch how our interior design demo integrates Nano Banana 2 Lite and Omni to instantly reimagine any space. Upload a photo, swipe through tailored design concepts, and see Omni bring the details to life in cinematic motion. Try out the demo app in AI Studio: t.co/EjYC2oHIDG
Explore ideas, scale visual concepts, and start creating: goo.gle/4bcThNt
Here’s what launched this week: — Gemini 3.5 Live Translate our latest audio model for live speech-to-speech translation — @NotebookLM got a major upgrade including agentic capabilities in chat, more advanced reasoning, and a suite of new output formats — Project Genie from
Today, we released Gemini 3.5 Live Translate, our latest audio model for live speech-to-speech translation. It supports over 70 languages and starts translating as soon as you start talking, streaming translations while listening to what you say next. No awkward pauses or choppy
Read the blog to learn more: blog.google/innovation-and…
Here’s this week’s shipping recap 👇 — Nano Banana 2 & Nano Banana Pro are now GA and available via the Gemini Enterprise Agent Platform, Gemini API, and in @GoogleAIStudio —Co-Scientist, our new multi-agent system for structured scientific thinking, generates and refines novel
Read more about all the fun ways we used AI to bring I/O to life this year: blog.google/innovation-and…
Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer sit down on camera together to share a behind-the-scenes look at the people behind the model,
Tune in for the full story ↓ youtube.com/watch?v=8hfpLa…
Watch your avatar explore extreme jobs in extreme locations x.com/ZefredAi/statu…
Watch your avatar speak Spanish, English and Japanese x.com/DotCSV/status/…