Generative AI has advanced rapidly in architectural design; however, existing building footprint generation models tend to emphasize stylistic exploration while insufficiently integrating site context as a fundamental physical constraint that facilitates alignment with the surrounding urban fabric. To address this limitation, this study proposes a context-responsive methodology for generating building footprints using a multi-layered four-channel representation of site conditions—including roads, sidewalks, adjacent buildings, and site boundaries—within a Latent Diffusion Model framework. The proposed approach encodes these physical conditions into a structured tensor and concatenates them directly to the U-Net input, enabling site context to function as an explicit spatial control variable during generation. An ablation study evaluated the effectiveness of the proposed contextual configuration. Compared with a single-channel model, the four-channel model achieved an 18.08% reduction in average pixel-wise information entropy, indicating a measurable decrease in generative uncertainty. Qualitative analyses further demonstrated that the enriched contextual input promotes geometrically coherent footprint configurations, such as context-responsive setbacks and spatial alignment with surrounding built forms. These findings suggest that structured multi-channel site information enhances contextual grounding in generative design processes and may contribute to more environmentally integrated and spatially coherent architectural outcomes.
장은석 et al. (Fri,) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: