DIFFCZSL: Compositional Zero-Shot Learning Regularized by Diffusion Representations
DIFFCZSL is proposed, a diffusion-augmented framework that injects generative priors from pre-trained diffusion models into CLIP-based CZSL pipelines and highlights the complementary strengths of generative diffusion representations and discriminative vision-language models for compositional generalization.