TensorX
返回文献探索

Paper · arXiv 2402.18842

ViewFusion: Towards Multi-View Consistency via Interpolated Denoising

Xianghui Yang, Yan Zuo, Sameera Ramasinghe, Loris Bazzani, Gil Avraham, Anton van den Hengel

14 upvotesFebruary 29, 2024arXiv 预印本
AI 摘要

ViewFusion, an auto-regressive, training-free algorithm, integrates with pre-trained diffusion models to synthesize consistent novel views by leveraging interpolated denoising from known views.

diffusion modelsnovel-view synthesismulti-view consistencyauto-regressive methodinterpolated denoisingsingle-view conditioned modelsmultiple-view conditional settings

Abstract

Novel-view synthesis through diffusion models has demonstrated remarkable potential for generating diverse and high-quality images. Yet, the independent process of image generation in these prevailing methods leads to challenges in maintaining multiple-view consistency. To address this, we introduce ViewFusion, a novel, training-free algorithm that can be seamlessly integrated into existing pre-trained diffusion models. Our approach adopts an auto-regressive method that implicitly leverages previously generated views as context for the next view generation, ensuring robust multi-view consistency during the novel-view generation process. Through a diffusion process that fuses known-view information via interpolated denoising, our framework successfully extends single-view conditioned models to work in multiple-view conditional settings without any additional fine-tuning. Extensive experimental results demonstrate the effectiveness of ViewFusion in generating consistent and detailed novel views.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
ViewFusion: Towards Multi-View Consistency via Interpolated Denoising | TensorX