TensorX
返回文献探索

Paper · arXiv 2409.17280

Disco4D: Disentangled 4D Human Generation and Animation from a Single Image

Hui En Pang, Shuai Liu, Zhongang Cai, Lei Yang, Tianwei Zhang, Ziwei Liu

10 upvotesSeptember 25, 2024arXiv 预印本
AI 摘要

Disco4D is a Gaussian Splatting framework for 4D human generation and animation from a single image, which disentangles clothing and body models, uses diffusion models for enhanced generation, and supports dynamic human animation.

Gaussian SplattingSMPL-X modeldiffusion models

Abstract

We present Disco4D, a novel Gaussian Splatting framework for 4D human generation and animation from a single image. Different from existing methods, Disco4D distinctively disentangles clothings (with Gaussian models) from the human body (with SMPL-X model), significantly enhancing the generation details and flexibility. It has the following technical innovations. 1) Disco4D learns to efficiently fit the clothing Gaussians over the SMPL-X Gaussians. 2) It adopts diffusion models to enhance the 3D generation process, e.g., modeling occluded parts not visible in the input image. 3) It learns an identity encoding for each clothing Gaussian to facilitate the separation and extraction of clothing assets. Furthermore, Disco4D naturally supports 4D human animation with vivid dynamics. Extensive experiments demonstrate the superiority of Disco4D on 4D human generation and animation tasks. Our visualizations can be found in https://disco-4d.github.io/.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Disco4D: Disentangled 4D Human Generation and Animation from a Single Image | TensorX