SSGLD: Spatial-Semantic Guided Latent Diffusion for Person Image Synthesis
Existing latent diffusion models (LDMs) for pose-guided person image synthesis (PGPIS) face an inherent trade-off between semantic consistency and texture fidelity. This dilemma stems from their reliance on a single conditioning pathway: semantic guidance is robust but prone to over-smoothing high-frequency details, wh...