Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Oct 2026

KVE-KD: Key Visual Evidence-Guided Knowledge Distillation for Vision-Language Models

Knowledge distillation is crucial for deploying vision-language models on resource-constrained devices. However, existing methods typically impose uniform supervision across visual tokens or rely on static token selection, which confuses task-relevant cues with background noise and degrades cross-modal reasoning. To ad...

Jian-Bing Zhang, Xin Sun, Shan-Wen Wang et al. · 0 citations
Preprint Sep 2026

Binaural Audio-Visual Instance Segmentation

Audio-visual segmentation (AVS) aims to segment sounding objects at the pixel level by integrating auditory and visual cues. However, existing methods are predominantly developed under the monaural setting and primarily rely on cross-modal semantic correspondence, which limits their ability to distinguish visually simi...

Sai-Jun Wang, Guan-Feng Tang, Hong-Bo Zhao et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.