PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 5, 20241 citationsOpen Access

Identity Decoupling for Multi-Subject Personalization of Text-to-Image Models

View Full Paper
SJSangwon JangJJJaehyeong JoKLKimin Lee

Key Points

Key points are not available for this paper at this time.

Abstract

Text-to-image diffusion models have shown remarkable success in generating a personalized subject based on a few reference images. However, current methods struggle with handling multiple subjects simultaneously, often resulting in mixed identities with combined attributes from different subjects. In this work, we present MuDI, a novel framework that enables multi-subject personalization by effectively decoupling identities from multiple subjects. Our main idea is to utilize segmented subjects generated by the Segment Anything Model for both training and inference, as a form of data augmentation for training and initialization for the generation process. Our experiments demonstrate that MuDI can produce high-quality personalized images without identity mixing, even for highly similar subjects as shown in Figure 1. In human evaluation, MuDI shows twice as many successes for personalizing multiple subjects without identity mixing over existing baselines and is preferred over 70% compared to the strongest baseline. More results are available at https://mudi-t2i.github.io/.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Jang et al. (2024) studied this question.

synapsesocial.com/papers/68e70459b6db64358767e373https://doi.org/10.48550/arxiv.2404.04243
Ask AI
Helpful
Bookmark
Share
View Full Paper