登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

alphaXiv
@askalphaxiv
High fidelity research
参加 November 2023
101 フォロー中    56.9K ファン
“World in World: Explore the World with World Models” Video world models can generate long rollouts, but controlling them from new viewpoints usually needs task-specific training or adapters. This paper instead turns source frames, geometry, and past generated states into visual evidence that a frozen world model can directly read through self-attention. This then gives training-free camera-controlled rerendering with better long-horizon consistency, unseen-view completion, and lower camera error than prior methods.
もっと見る