Joint work by our awesome research intern Qihang Zhang, together with colleagues Shuangfei, Miguel, Kevin, Alex and Josh at Apple MLR!
Want to dive deeper? Check out our paper for full details
ArXiv: arxiv.org/abs/2412.01821
Project page: zqh0253.github.io/wvd/ (9/n, n=9)
arxiv.org
World-consistent Video Diffusion with Explicit 3D Modeling
Recent advancements in diffusion models have set new benchmarks in image and video generation, enabling realistic visual synthesis across single- and multi-frame contexts. However, these models still ...