TensorX
返回文献探索

Paper · arXiv 2312.02201

ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation

Peng Wang, Yichun Shi

35 upvotesDecember 2, 2023arXiv 预印本
AI 摘要

ImageDream is a multi-view diffusion model for 3D object generation that uses image prompts to produce high-quality 3D models with improved visual geometry accuracy.

image-promptmulti-view diffusion model3D object generationcanonical camera coordinationglobal controllocal control

Abstract

We introduce "ImageDream," an innovative image-prompt, multi-view diffusion model for 3D object generation. ImageDream stands out for its ability to produce 3D models of higher quality compared to existing state-of-the-art, image-conditioned methods. Our approach utilizes a canonical camera coordination for the objects in images, improving visual geometry accuracy. The model is designed with various levels of control at each block inside the diffusion model based on the input image, where global control shapes the overall object layout and local control fine-tunes the image details. The effectiveness of ImageDream is demonstrated through extensive evaluations using a standard prompt list. For more information, visit our project page at https://Image-Dream.github.io.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation | TensorX