Skip to content

Author

Jia-Hong Huang

We have 3 of 61 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

MMArt: A Multi-Perspective Multimodal Dataset for Visual Art Understanding

MMArt is introduced, a large-scale dataset of 74,234 WikiArt paintings, each annotated with four independently annotated perspectives plus a harmonized unified caption, produced by specialized vision-language models or human annotation and validated through complementary quality evaluations.

Shuai Wang, Wang-Yuan Ding, Yixian Shen et al. · 0 citations
Preprint Aug 2026

Can We Read the Mind of an Audio LLM? A Verbalizable, Multilingual Middle-Layer Workspace

Reading a base Qwen3-Omni with a logit lens at the audio-token positions, it is found that the answer to a spoken question becomes legible - in words - in the model's middle layers, before it emits any token.

Jiajun Fan, Jing-Yuan Li, Prashanth Gurunath Shivakumar et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.