🤔Image-to-3D, monocular depth estimation, camera pose estimation, …, can we achieve all of this with just ONE model easily? 🚀Our answer is Yes -- Excited to introduce our latest work: World-consistent Video Diffusion (WVD) with Explicit 3D Modeling! arxiv.org/abs/2412.01821
Jiatao Gu
@jgu32.bsky.social
Machine Learning Researcher @Apple MLR Incoming Assistant Professor @Penn CIS See more details https://jiataogu.me
I am seeking multiple PhD students passionate about Generative Intelligence and its applications in empowering AI agents to interact with the physical world to join us at UPenn CIS for the 2024-2025 academic cycle. You can find more information at www.cis.upenn.edu/graduate/pro...
Doctoral Program
Doctoral Program
cis.upenn.edu