PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 29, 20231 citations

Story-to-Images Translation: Leveraging Diffusion Models and Large Language Models for Sequence Image Generation

View Full Paper
HKHaruka KumagaiRYRyosuke YamakiHNHiroki Naganuma

Key Points

Key points are not available for this paper at this time.

Abstract

Diffusion models are catalyzing breakthroughs in creative fields, with a notable impact on text-to-image generation. This study centers on the transformation of textual narratives into coherent sequences of images - a process currently hampered by issues of consistency and contextual fidelity. To address these challenges, we propose a method utilizing a large language model, with an emphasis on context and character information. Empirical evaluations, carried out using Hollywood movie scripts, clearly indicate that our approach improves both the consistency and contextual fidelity of the resulting image sequences.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Kumagai et al. (2023) studied this question.

synapsesocial.com/papers/6a0f559b8f3ca410b09bcb99https://doi.org/10.1145/3607540.3617144
Ask AI
Helpful
Bookmark
Share
View Full Paper