Direct, Parallel, or Sequential? A Comparative Study of Training-Free Multi-Subject Image-to-Video Generation
A systematic study of three representative paradigms for training-free multi-subject I2V generation: direct, parallel, and sequential generation, revealing the strengths and limitations of each paradigm and offering practical insights for designing controllable multi-subject video generation systems.