
@JeffDean
Chief Scientist, Google DeepMind & Google Research. Gemini Lead. Opinions stated here are my own, not those of Google. TensorFlow, MapReduce, Bigtable, ...
A nice aspect of our new Gemini 3.6 Flash model is that it is much more token efficient than our 3.5 Flash model. Here's a side-by-side demonstration of that. Nice work by everyone who worked on this model and release! x.com/GoogleDeepMind…
Yesterday I was fortunate enough to go to my first-ever World Cup game, with my long-time colleagues/friends @OriolVinyalsML and @quocleix and their spouses). We were sitting in the corner area and had a quite good view of the goal! I present to you a 43 second multi-part drama filled with emotion: The initial promising-looking cross coming in, but looking overhit Nico Williams cleverly knocking it down at the back post into a dangerous area Ferran Torres striking it cleanly into the roof of the net The crowd rising as one (forcing me to stand up as well) The elation of the Spanish players racing off the bench to celebrate The dejection of the Argentinian players, their defense having finally been breached in extra time The elation of my Spanish colleague Oriol and his wife Meire next to me (he and I are both Barça fans, so it was nice to see a Barça player score the winning goal) The entire stadium reacting Whew!
@giffmana Maybe if it was a TPU smack instead people would talk smack about their chip allocations?
The video from the commencement speech I gave a couple of weeks ago at @uwcse is now up. Congratulations again to all the graduates, and thanks for having me, @MBalazinska! youtu.be/Pvkmwh7Mito?si… (video is of whole ceremony: my part starts around the 3 minute mark) x.com/JeffDean/statu…
My @Google colleagues @NormJouppi, Sridhar Lakshmanamurthy, Cliff Young, and David Patterson recently wrote a paper that will appear in the July/August 2026 edition of @ieeemicro titled "Google's Training Supercomputers from TPU v2 to Ironwood: Architectural Stability, Scale, Resilience, Power Efficiency, and Sustainability Across Five Generations". It's chock full of interesting data about the evolution of TPU chip generations, as well as how workloads at Google have transformed over time (hint: lots more transformer-based models!), and how the generations have gotten ~30X more energy efficient per flop. Lots of changes over these generations: Air cooling in TPUv2 to water cooling in TPUv3 onwards 2D to 3D torus-based interconnects 30X improvement TFLOPS/Watt 256 chips (TPUv2) to 9216 chips (Ironwood) per pod Read the full paper: t.co/D5NFYFv19V
I really enjoyed this game today. Vozinha and the whole Cape Verde team were amazing against Spain! x.com/minwooshi8/sta…
A good essay by @pgasawa and @profjoeyg on a more nuanced view of AI advances. x.com/pgasawa/status…