TensorX
返回文献探索

Paper · arXiv 2312.04560

NeRFiller: Completing Scenes via Generative 3D Inpainting

Ethan Weber, Aleksander Hołyński, Varun Jampani, Saurabh Saxena, Noah Snavely, Abhishek Kar, Angjoo Kanazawa

13 upvotesDecember 7, 2023arXiv 预印本
AI 摘要

NeRFiller utilizes a 2D inpainting diffusion model to generate consistent 3D scene completions from sparse observations.

generative 3D inpaintingdiffusion models2D inpainting diffusion modelscene completion3D scenescene reconstruction

Abstract

We propose NeRFiller, an approach that completes missing portions of a 3D capture via generative 3D inpainting using off-the-shelf 2D visual generative models. Often parts of a captured 3D scene or object are missing due to mesh reconstruction failures or a lack of observations (e.g., contact regions, such as the bottom of objects, or hard-to-reach areas). We approach this challenging 3D inpainting problem by leveraging a 2D inpainting diffusion model. We identify a surprising behavior of these models, where they generate more 3D consistent inpaints when images form a 2times2 grid, and show how to generalize this behavior to more than four images. We then present an iterative framework to distill these inpainted regions into a single consistent 3D scene. In contrast to related works, we focus on completing scenes rather than deleting foreground objects, and our approach does not require tight 2D object masks or text. We compare our approach to relevant baselines adapted to our setting on a variety of scenes, where NeRFiller creates the most 3D consistent and plausible scene completions. Our project page is at https://ethanweber.me/nerfiller.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
NeRFiller: Completing Scenes via Generative 3D Inpainting | TensorX