李飞飞宣布其创办的空间智能公司 World Labs 加入 AMD,她本人将出任 AMD 执行副总裁兼首席科学家,直接与 CEO Lisa Su 合作。
推荐理由:作者本人说明收购动机与此前技术积累,读者可以了解World Labs为何选择并入硬件公司这条路径。
@drfeifei · X
李飞飞宣布其创办的空间智能公司 World Labs 加入 AMD,她本人将出任 AMD 执行副总裁兼首席科学家,直接与 CEO Lisa Su 合作。
推荐理由:作者本人说明收购动机与此前技术积累,读者可以了解World Labs为何选择并入硬件公司这条路径。
构建任何技术——包括 AI——的目标都应是改善人类生活和社会。
“Any threat to human society, including existential, is within ourselves,” says World Labs Technologies CEO Fei-Fei Li as she discusses the risks surrounding AI, and the responsibility humans have in shaping how the technology is developed and used. Listen to our full interview here: https://bloom.bg/3V74HgP
机器人学习的研究人员/工程师加入 @theworldlabs 的绝佳机会!❤️🔥
We're hiring in robot learning at @theworldlabs! Join me, @drfeifei, and the team to define and scale the next generation of world models for robot learning! Atlas for Robotics: https://www.worldlabs.ai/blog/atlas#robotics-simulation Real-to-Sim-to-Real: https://www.worldlabs.ai/blog/real-to-sim-to-real Apply: https://jobs.ashbyhq.com/worldlabs/85994fa5-c44c-48b4-84fc-aece7934c2cb
One more thing…
@jcjohnss @BenMildenhall @martin_casado 关于 Atlas 的更多内容!https://t.co/sy8pabeojI
World Labs co-founders Fei-Fei Li, Justin Johnson, Ben Mildenhall, and a16z's Martin Casado on Atlas, a world model for spatial intelligence: LLMs are built on next token prediction. Video models are built on next frame prediction. Atlas is built on new view prediction, and it's the first model to unify pixel generation and pixel reconstruction, two problems computer vision has kept in separate tracks for half a century. The practical result is a 50 to 100x reduction in what it takes to digitally capture a 3D representation of a space. Previously, you needed 100 to 300 photos of a single room. Atlas can work from just three. In this conversation, they get into the slow motion shot from The Matrix that took hundreds of cameras and now takes three iPhones, the overnight Slack message that made them bet the company in five seconds, why robotics is bottlenecked on data rather than chips, and the case that new view prediction is AI-complete. 00:00 Intro 01:50 The Matrix slow motion scene now takes three iPhones 02:48 Why new view prediction is the primitive 07:10 Unifying generation and reconstruction 11:15 Gaussian splats became the bottleneck 14:17 Dense capture used to mean 300 photos 17:30 Why reconstruction needs generation to fill the gaps 18:44 The LLM lesson image models missed 23:39 The video that made them go all in 28:04 3D design is 95% revisions 30:50 The problem in robotics is data, not chips 32:48 Why a robot policy can't be trained like an image model 34:44 When the simulator becomes the planner 36:45 Frozen time required footage full of movement 40:57 Why new view prediction is AI-complete 42:43 Nature gave animals eyes but not trees YouTube: https://www.youtube.com/watch?v=qn1QDDBnTA0 @drfeifei @jcjohnss @BenMildenhall @theworldlabs @martin_casado
Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D. Model the world, move the camera, and simulate space & time.