跳到正文
Rohan Paul· @rohanpaul_ai · X·· 1 小时前AI 评分62
AI 导读

Odyssey 发布基础世界模型 Odyssey-3,其 Pro 版报告了 Physics-IQ Verified 视频-视频最高分 66.1。模型是自回归扩散 Transformer,基于先前帧和输入动作预测下一帧;蒸馏版 Flash 可将文本提示转为实时响应动作的可探索世界,免费开放体验。

正文

Odyssey released Odyssey-3, a world model whose Pro version posts the highest reported Physics-IQ Verified video-to-video score, 66.1.

It's an autoregressive diffusion transformer that predicts next frames from earlier ones plus incoming actions, letting prompted scenes react to each move.

> Odyssey-3 Flash, a distilled few-step version, turns a text prompt into an explorable world that reacts to movement and events in real time, and anyone can try it free today.

> With Odyssey-3 frozen, a policy trained on 20 hours of Indian road data drove a real car, and a robot arm recovered from missed grasps its demos never showed, though neither result reports success rates.

> on image-to-video Odyssey-3 Pro roughly ties FLUX 3 [large]

引用Odyssey@odysseyml
Today we're launching Odyssey-3, the most powerful foundation world model yet. It sets a new state of the art on Physics-IQ, and powers robots, trains AIs, and generates interactive experiences. It's really cool. Experience the model today, all for free!
在 X 查看被引用的帖子

来源:Rohan Paul · x.com