MentalHospital, a virtual evaluation environment for LLM-based psychiatric clinical encounters, and MentalEval, five domain-specific evaluators covering communication empathy, interviewing professionalism, clinical-note quality, diagnostic rigor, and treatment appropriateness, trained with rubric-grounded SFT and expert-guided DPO are introduced.
Yu-Ming Yang, Xiao Sun, Yuanwei Zou et al.· arXiv.org· 0 citations
A new multimodal benchmark dataset, CoMMET, a comprehensive mental states and moral evaluation task inspired by the Theory of Mind Booklet Task is proposed, which is the first psychology-grounded benchmark to evaluate MLLMs across multiple mental states in a multimodal, open-ended, and multi-turn setting.
Rui-Rui Chen, Wei-Feng Jiang, Cheng-Wei Qin et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.