Jul 2026
Mgsc: multimodal generation and self-supervised contrast learning for mitigating language bias in visual question answering
This work proposes Multimodal Generation and Self-Supervised Contrast Learning (MGSC), which first leverages a generative adversarial network to train a bias model, guiding the target model to capture complementary information between vision and language.
Xinyu Jiang, Zhenfang Zhu, Qiang Lu et al.
· Multimedia Systems · 0 citations