Abstract
In this paper, we present a method for generating consistent novel views from a single source image. Our approach focuses on maximizing the reuse of visible pixels from the source image. To achieve this, we use a monocular depth estimator that transfers visible pixels from the source view to the target view. Starting from a pre-trained 2D inpainting diffusion model, we train our method on the large-scale Objaverse dataset to learn 3D object priors. While training we use a novel masking mechanism based on epipolar lines to further improve the quality of our approach. This allows our framework to perform zero-shot novel view synthesis on a variety of objects. We evaluate the zero-shot abilities of our framework on three challenging datasets: Google Scanned Objects, Ray Traced Multiview, and Common Objects in 3D.
Author supplied keywords
Cite
CITATION STYLE
Kant, Y., Siarohin, A., Vasilkovsky, M., Guler, R. A., Ren, J., Tulyakov, S., & Gilitschenski, I. (2023). Repurposing Diffusion Inpainters for Novel View Synthesis. In Proceedings - SIGGRAPH Asia 2023 Conference Papers, SA 2023. Association for Computing Machinery, Inc. https://doi.org/10.1145/3610548.3618149
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.