A Methodology to Compare Real and LLM-Simulated Multimodal Interactions: The Case of Gender Bias
Large Language Models (LLMs) are increasingly used to generate multimodal data for building corpora or simulating synthetic behaviours. A key challenge remains the evaluation of generated data. Comparison with real multimodal human interactions may help address this issue, but unified representations and evaluation fra...