TensorX
返回文献探索

Paper · arXiv 2403.13535

IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image Models

Siying Cui, Jiankang Deng, Jia Guo, Xiang An, Yongle Zhao, Xinyu Wei, Ziyong Feng

23 upvotesMarch 20, 2024arXiv 预印本
AI 摘要

IDAdapter, a tuning-free method, enhances the diversity and identity preservation in personalized image generation using a single face image by integrating personalized textual and visual injections and a face identity loss.

Stable Diffusionpersonalized portraitshigh-fidelitycharacter avatarstest-time fine-tuninginput imagesidentity preservationdiversityIDAdaptertextual injectionsvisual injectionsface identity lossmixed featuresreference imagesidentity-related content detailsimage generationdiversityidentity fidelity

Abstract

Leveraging Stable Diffusion for the generation of personalized portraits has emerged as a powerful and noteworthy tool, enabling users to create high-fidelity, custom character avatars based on their specific prompts. However, existing personalization methods face challenges, including test-time fine-tuning, the requirement of multiple input images, low preservation of identity, and limited diversity in generated outcomes. To overcome these challenges, we introduce IDAdapter, a tuning-free approach that enhances the diversity and identity preservation in personalized image generation from a single face image. IDAdapter integrates a personalized concept into the generation process through a combination of textual and visual injections and a face identity loss. During the training phase, we incorporate mixed features from multiple reference images of a specific identity to enrich identity-related content details, guiding the model to generate images with more diverse styles, expressions, and angles compared to previous works. Extensive evaluations demonstrate the effectiveness of our method, achieving both diversity and identity fidelity in generated images.

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image Models | TensorX