Odyssey 发布基础世界模型 Odyssey-3,其 Pro 版报告了 Physics-IQ Verified 视频-视频最高分 66.1。模型是自回归扩散 Transformer,基于先前帧和输入动作预测下一帧;蒸馏版 Flash 可将文本提示转为实时响应动作的可探索世界,免费开放体验。
Odyssey released Odyssey-3, a world model whose Pro version posts the highest reported Physics-IQ Verified video-to-video score, 66.1.
It's an autoregressive diffusion transformer that predicts next frames from earlier ones plus incoming actions, letting prompted scenes react to each move.
> Odyssey-3 Flash, a distilled few-step version, turns a text prompt into an explorable world that reacts to movement and events in real time, and anyone can try it free today.
> With Odyssey-3 frozen, a policy trained on 20 hours of Indian road data drove a real car, and a robot arm recovered from missed grasps its demos never showed, though neither result reports success rates.
> on image-to-video Odyssey-3 Pro roughly ties FLUX 3 [large]
Today we're launching Odyssey-3, the most powerful foundation world model yet. It sets a new state of the art on Physics-IQ, and powers robots, trains AIs, and generates interactive experiences. It's really cool. Experience the model today, all for free!在 X 查看被引用的帖子
来源:Rohan Paul · x.com