
@drfeifei
Cofounder/CEO @theworldlabs, Prof (CS @Stanford), Co-Director @StanfordHAI, #AI #SpatialIntelligence #GenAI #computervision #robotics #AI-healthcare
Agents can be AI, but agency is human. It is every individual’s responsibility to empower our own agency, with the help of tools. But the North Star should always remain human centered.
Incredible opportunity for robotic learning researchers/engineers to join @theworldlabs! ❤️🔥
Yunzhu Li@YunzhuLiYZ·We're hiring in robot learning at @theworldlabs! Join me, @drfeifei, and the team to define and scale the next generation of world models for robot learning! Atlas for Robotics: worldlabs.ai/blog/atlas#rob… Real-to-Sim-to-Real: worldlabs.ai/blog/real-to-s… Apply: jobs.ashbyhq.com/worldlabs/8599…
@eccvconf Oh what an honor! Thank you! And congrats to all the fellow awardees! 🤩
The world of R&D is forking into two paths: the token-abundant research, and the token-starved research. The future is in the former - evidentially, the progress by today's top AI industry teams and neolabs is breathtaking, where researchers’ human brilliance is super charged by AI’s assistance. Every research university president should be reading this report and reflecting on what the future of higher education research should be. t.co/y74Ai3mIhk
Atlas runs on real-time!!🚀
World Labs@theworldlabs·One more thing…
Next view prediction is the key to Atlas, enabling us to unify pixel-level generation and reconstruction. @jcjohnss @BenMildenhall @martin_casado and I had a deeper discussion on some of the most exciting technical innovations of Atlas, our newly released world model for spatial intelligence!
a16z@a16z·World Labs co-founders Fei-Fei Li, Justin Johnson, Ben Mildenhall, and a16z's Martin Casado on Atlas, a world model for spatial intelligence: LLMs are built on next token prediction. Video models are built on next frame prediction. Atlas is built on new view prediction, and it's the first model to unify pixel generation and pixel reconstruction, two problems computer vision has kept in separate tracks for half a century. The practical result is a 50 to 100x reduction in what it takes to digitally capture a 3D representation of a space. Previously, you needed 100 to 300 photos of a single room. Atlas can work from just three. In this conversation, they get into the slow motion shot from The Matrix that took hundreds of cameras and now takes three iPhones, the overnight Slack message that made them bet the company in five seconds, why robotics is bottlenecked on data rather than chips, and the case that new view prediction is AI-complete. 00:00 Intro 01:50 The Matrix slow motion scene now takes three iPhones 02:48 Why new view prediction is the primitive 07:10 Unifying generation and reconstruction 11:15 Gaussian splats became the bottleneck 14:17 Dense capture used to mean 300 photos 17:30 Why reconstruction needs generation to fill the gaps 18:44 The LLM lesson image models missed 23:39 The video that made them go all in 28:04 3D design is 95% revisions 30:50 The problem in robotics is data, not chips 32:48 Why a robot policy can't be trained like an image model 34:44 When the simulator becomes the planner 36:45 Frozen time required footage full of movement 40:57 Why new view prediction is AI-complete 42:43 Nature gave animals eyes but not trees YouTube: t.co/AvR59efen0 @drfeifei @jcjohnss @BenMildenhall @theworldlabs @martin_casado
@jcjohnss @BenMildenhall @martin_casado More on Atlas!
World Labs@theworldlabs·Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D. Model the world, move the camera, and simulate space & time.
I'm so excited that our @theworldlabs team has achieved a major milestone today! Introducing Atlas - a first of its kind multimodal world model trained from scratch! 🚀 Atlas is capable of generating frames with pixel-perfect camera control, reconstructing large scenes from as few as one single input image, simulating space-time by reframing videos, natively outputting 3D spaces from one or more input images, composing multiple posed images into a consistent 3d world, and more! This is the best camera conditioned world model ever, opening doors to many possible use cases from VFX to robotics. I'm so so so proud of our team!♥️
World Labs@theworldlabs·Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D. Model the world, move the camera, and simulate space & time.
On this random Sunday afternoon, I just noticed that I have 1M followers on X now. Feels like a big responsibility 🤔❤️
Indeed all tools should be about augmenting human agency, including AI! I had a fun chat with @hubermanlab . x.com/hubermanlab/st…
When SceniX joined World Labs, we said spatial intelligence was never only about perceiving and generating virtual and physical worlds, but also interacting with them. Today, we’re sharing early results from that vision: building worlds that train robots. 🌎🤖↓
With the help of generative world models, a real-to-sim-to-real (R2S2R) simulation engine turns one physical task into many controllable, reusable worlds, helping robotics teams train policy models and test changes faster, uncover failures earlier, and reduce costly experimentation on hardware.
The world is not just made of words, and spatial intelligence was never just about perceiving and generating worlds. It's about interacting with them. Today, SceniX is joining World Labs. 🌎🤖👇
@YunzhuLiYZ, @fast_sploosh, and @xhsonny have built a remarkable team that is training and evaluating robots in high-fidelity simulation – proven not only in lab demos, but in live deployments on real hardware. Last month, we wrote that the boundaries between rendering, simulation, and planning are beginning to blur. Bringing our world models together with SceniX's simulation and robotics expertise is an important step in our quest for spatial intelligence. Welcome to the team. Read the full announcement: t.co/IUSawjTwz5
I’m very excited by this test time training work for robotic learning! It’s an awesome collaboration between @StanfordSVL and @NVIDIARobotics ! x.com/drjimfan/statu…
AI’s next chapter will be defined not only by technical progress, but by how responsibly and thoughtfully we bring it into the world. I’m looking forward to be speaking on stage at Ai4 2026, Aug 4–6 in Vegas. Register here at ai4.io/register/ #Ai42026
1/N Long horizon, complex tasks that truly matter in everyday life are not solved problems by today’s robotics, requiring planning, object detection, object manipulation, and failure recovery. That's why Stanford's BEHAVIOR Challenge is back for year 2! Last year, the winning solution reached only 12.4% full task success. This year, the BEHAVIOR challenge has more tasks, better evaluation, and is easier to use. 🚨 ⏰ Submission deadline: 10/16/2026 📣 Winners announced: 11/04/2026 🏆 Prize pool: $11,000
2/N Real-world robot evaluation is essential but hard to scale: experiments are difficult to control, reproduce, and compare. Simulation is a powerful testbed for scalable, controlled, reproducible evaluation. BEHAVIOR-1K is an open-source simulation benchmark of 1,000 everyday household activities requiring long-horizon reasoning, navigation, and bimanual manipulation, giving us a scalable way to measure how well robot AI models generalize. t.co/lV8BelxBCq
Your thoughtful reflection is so inspiring and encouraging @smallfly ! As everyone talks about AI and automation, human creativity, story telling and productivity are even more important and essential to our society. @theworldlabs is founded on the premise of empowering human ingenuity and productivity. We are very grateful to be able to work with people like you! 🙏🌐
Scientific research is fundamental to advancing civilization and helping people globally to solve the most critical problems, from medicine to materials, from brain science to physics, and much beyond. This is only possible when scientists have access to the best tools of the time to conduct scientific research, including having access to AI-based tools.
The creativity and imagination is out of the world! So grateful that @theworldlabs got to partner with the amazing talents @withloreco to translate their incredible ideas into an interactive experiences for users to enjoy!🤩 x.com/theworldlabs/s…