Recent large audio language models (LALMs) have achieved impressive progress in audio understanding. However, existing evaluations remain largely constrained to English and narrow audio domains. Prior benchmarks typically focus on a single audio modality, i.e., speech, sound, or music, limiting the systematic investiga...
Jia-Wen Wang, Xiao-Xue Gao, Ziliang Pang et al.· 0 citations
This paper comprehensively studies effective prompt injection attacks against 14 widely used open-source and three closed-source LLMs on five attack benchmarks and proposes a straightforward and effective hypnotism attack, showing that this attack causes aligned language models to generate objectionable behaviors.
Jia-Wen Wang, Pritha Gupta, E. Hüllermeier et al.· 9 citations
The causes of modal divergence are probed, offering insights into fostering culturally robust MLLMs, and a Multilingual, Multimodal Alignment framework for Cultural grounding evaluation is proposed.
Weihua Zheng, Zhengyuan Liu, Tanmoy Chakraborty et al.· Annual Meeting of the Associ...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.