TensorX
返回文献探索

Paper · arXiv 2404.09995

Taming Latent Diffusion Model for Neural Radiance Field Inpainting

Chieh Hubert Lin, Changil Kim, Jia-Bin Huang, Qinbo Li, Chih-Yao Ma, Johannes Kopf, Ming-Hsuan Yang, Hung-Yu Tseng

7 upvotesApril 15, 2024arXiv 预印本
AI 摘要

The proposed framework addresses limitations in NeRF editing by tempering diffusion model stochasticity and mitigating textural shifts, leading to state-of-the-art NeRF inpainting results.

Neural Radiance FieldNeRFdiffusion priorlatent diffusion modelsradiance fieldtextural shiftauto-encoding errorspixel-distance lossesperceptual lossesmasked adversarial trainingNeRF inpainting

Abstract

Neural Radiance Field (NeRF) is a representation for 3D reconstruction from multi-view images. Despite some recent work showing preliminary success in editing a reconstructed NeRF with diffusion prior, they remain struggling to synthesize reasonable geometry in completely uncovered regions. One major reason is the high diversity of synthetic contents from the diffusion model, which hinders the radiance field from converging to a crisp and deterministic geometry. Moreover, applying latent diffusion models on real data often yields a textural shift incoherent to the image condition due to auto-encoding errors. These two problems are further reinforced with the use of pixel-distance losses. To address these issues, we propose tempering the diffusion model's stochasticity with per-scene customization and mitigating the textural shift with masked adversarial training. During the analyses, we also found the commonly used pixel and perceptual losses are harmful in the NeRF inpainting task. Through rigorous experiments, our framework yields state-of-the-art NeRF inpainting results on various real-world scenes. Project page: https://hubert0527.github.io/MALD-NeRF

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Taming Latent Diffusion Model for Neural Radiance Field Inpainting | TensorX