PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 12, 20247 citations

Sketch-Guided Text-to-Image Generation with Spatial Control

View Full Paper
TZTianyu ZhangHXHaoran Xie

Key Points

Key points are not available for this paper at this time.

Abstract

Recent text-to-image generation models can produce high-quality images from textual prompts. However, it is difficult to correctly interpret instructions specifying the complex images with multiple objects using only texts. To solve this issue, we propose a sketch-guided spatial control for text-to-image diffusion models. In the feature extraction stage of the proposed framework, sketch inputs are segmented into individual objects using the image segmentation approach. The obtained bounding boxes and labels are used as spatial-guided inputs into the attention layers of the diffusion model. For the image generation stage, the proposed model utilizes a pretrained text-to-image diffusion model as the image generator. We assess the proposed method through both quantitative and qualitative evaluations, demonstrating its versatility in spatial control based on user sketches.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhang et al. (2024) studied this question.

synapsesocial.com/papers/6a12d64545487b7639a74016https://doi.org/10.1109/cgip62525.2024.00035
Ask AI
Helpful
Bookmark
Share
View Full Paper