跳到正文
原文
Fei-Fei Li· @drfeifei · X·· 2026-09-02AI 评分57
AI 导读

Fei-Fei Li 宣布 World Labs 团队发布从零训练的多模态世界模型 Atlas。该模型支持像素级精准的相机控制生成帧、从单张输入图像重建大场景、通过重排视频帧模拟时空、从一或多张图像原生输出 3D 空间,以及将多张带位姿图像组合成一致的 3D 世界,潜在应用覆盖 VFX 到机器人。

正文

I'm so excited that our @theworldlabs team has achieved a major milestone today! Introducing Atlas - a first of its kind multimodal world model trained from scratch! 🚀

Atlas is capable of generating frames with pixel-perfect camera control, reconstructing large scenes from as few as one single input image, simulating space-time by reframing videos, natively outputting 3D spaces from one or more input images, composing multiple posed images into a consistent 3d world, and more! This is the best camera conditioned world model ever, opening doors to many possible use cases from VFX to robotics. I'm so so so proud of our team!♥️

引用World Labs@theworldlabs
Introducing Atlas: The world's first multimodal world model that generates image and video frames with pixel-perfect camera control and reconstructs them in 3D. Model the world, move the camera, and simulate space & time.
在 X 查看被引用的帖子

来源:Fei-Fei Li · x.com