跳到正文
Rohan Paul· @rohanpaul_ai · X·· 2 小时前AI 评分66
AI 导读

Odyssey 发布基础世界模型 Odyssey-3,Pro 版在 Physics-IQ Verified 视频到视频任务上报告最高分 66.1。它是自回归扩散 Transformer,根据先前帧和输入动作预测下一帧,使提示场景随每个动作实时反应;蒸馏的少步版本 Odyssey-3 Flash 可从文本提示生成可探索的实时交互世界,今日免费开放。

正文

Odyssey released Odyssey-3, a world model whose Pro version posts the highest reported Physics-IQ Verified video-to-video score, 66.1.

It's an autoregressive diffusion transformer that predicts next frames from earlier ones plus incoming actions, letting prompted scenes react to each move.

Odyssey-3 Flash, a distilled few-step version, turns a text prompt into an explorable world that reacts to movement and events in real time, and anyone can try it free today.

With Odyssey-3 frozen, a policy trained on 20 hours of Indian road data drove a real car, and a robot arm recovered from missed grasps its demos never showed, though neither result reports success rates.

on image-to-video Odyssey-3 Pro roughly ties FLUX 3 [large]

引用Odyssey@odysseyml
Today we're launching Odyssey-3, the most powerful foundation world model yet. It sets a new state of the art on Physics-IQ, and powers robots, trains AIs, and generates interactive experiences. It's really cool. Experience the model today, all for free!
在 X 查看被引用的帖子

来源:Rohan Paul · x.com