Skip to content
arXiv cs.CV · Papers

Video Models as Native 4D Renderers: World-Grounded Conditioning from Animated Mesh

arXiv:2608.00094v2 Announce Type: replace Abstract: Pretrained video diffusion models can act as renderers when the desired scene state is already specified by an animated mesh, a camera trajectory, and a reference image. This 4D generative rendering setting raises a representation question: what image-format condition