가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

alphaXiv
@askalphaxiv
High fidelity research
가입 November 2023
101 팔로잉 중    56.9K 팬
“World in World: Explore the World with World Models” Video world models can generate long rollouts, but controlling them from new viewpoints usually needs task-specific training or adapters. This paper instead turns source frames, geometry, and past generated states into visual evidence that a frozen world model can directly read through self-attention. This then gives training-free camera-controlled rerendering with better long-horizon consistency, unseen-view completion, and lower camera error than prior methods.
더 보기