PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 24, 20260 citations

Unifying Multi-modal Hair Editing via Proxy Feature Blending.

View Full Paper
TWTianyi WeiDCDongdong ChenWZWenbo Zhou

Key Points

  • The aim is to create a unified hair editing framework that supports multiple interaction modes while maintaining attribute preservation.
  • Developed a proxy-based hair transfer method using StyleGAN
  • Utilized dense latent space for precise editing
  • Converted editing conditions into transfer proxies
  • Extended approach to 3D-aware settings via EG3D and PanoHead
  • Implemented multi-view boosted hair feature localization strategies
  • Outperformed previous methods in hair editing effects and visual naturalness
  • Achieved consistent attribute preservation across diverse editing modes
  • Demonstrated improved multi-view consistency in outputs
  • Enabled unprecedented support for multimodal interactions

Abstract

Hair editing is a long-standing problem in computer vision that demands both fine-grained local control and intuitive user interactions across diverse modalities. Despite the remarkable progress of GANs and diffusion models, existing methods still lack a unified framework that simultaneously supports arbitrary interaction modes (e.g., text, sketch, mask, and reference image) while ensuring precise editing and faithful preservation of irrelevant attributes. In this work, we introduce a novel paradigm that reformulates hair editing as proxy-based hair transfer. Specifically, we leverage the dense and semantically disentangled latent space of StyleGAN for precise manipulation and exploit its feature space for disentangled attribute preservation, thereby decoupling the objectives of editing and preservation. Our framework unifies different modalities by converting editing conditions into distinct transfer proxies, whose features are seamlessly blended to achieve global or local edits. Beyond 2D, we extend our paradigm to 3D-aware settings by incorporating EG3D and PanoHead, where we propose a multi-view boosted hair feature localization strategy together with 3D-tailored proxy generation methods that exploit the inherent properties of 3D-aware generative models. Extensive experiments demonstrate that our method consistently outperforms prior approaches in editing effects, attribute preservation, visual naturalness, and multi-view consistency, while offering unprecedented support for multimodal and mixed-modal interactions.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Wei et al. (2026) studied this question.

synapsesocial.com/papers/697461a8bb9d90c67120b7a5https://doi.org/10.1109/tpami.2026.3656763
Ask AI
Helpful
Bookmark
Share
View Full Paper