COSMO replaces expert-to-expert guidance with co-adaptation through an anchored shared consensus and achieves state-of-the-art performance under matched VLM backbones, indicating that it better balances the retention of valid source-derived evidence with the absorption of complementary VLM evidence.
Abstract
Source-free domain adaptation (SFDA) adapts a source-trained model to an unlabeled target domain without source data, a practical setting under privacy or storage constraints. Yet its self-generated supervision can reinforce source bias under substantial domain shifts. Pretrained vision-language models (VLMs) offer complementary semantic knowledge, but the relative reliability of the source model and VLM varies across target samples. Existing cross-model guidance does not explicitly account for this variation and may overwrite valid source-derived evidence under conflict, a failure we term source-derived evidence forgetting. We formulate VLM-guided SFDA as a sample-wise reliability-allocation problem and propose Consensus-Driven Shift Modulation (COSMO). COSMO replaces expert-to-expert guidance with co-adaptation through an anchored shared consensus. It first forms a sample-specific initial consensus that favors the more concentrated prediction. During adaptation, COSMO re-aggregates both branches'evolving evidence and regulates how far the resulting consensus moves from its initial anchor based on consensus uncertainty and training progress. This keeps the shared supervision anchored yet adaptive. Across four benchmarks, COSMO achieves state-of-the-art performance under matched VLM backbones. Further analyses indicate that it better balances the retention of valid source-derived evidence with the absorption of complementary VLM evidence.
A Peer-level Heterogeneous Perception Framework is proposed that departs from such paradigms by enabling balanced collaboration between heterogeneous models by introducing an auxiliary domain that is significantly different from the target domain and employ an auxiliary model with the same architecture as the source mo...
Zhi-Ze Wu, Yu-Tao Fu, Huan-Xin Zou et al.· Multimedia Systems· 0 citations
Source-Free Domain Adaptation (SFDA) adapts a pre-trained source model to an unlabeled target domain without accessing source data, alleviating the need for direct access to source data during adaptation, which reduces the risk of data transmission. Existing methods leverage pre-trained Vision-Language (ViL) models to...
Bing-Tao Zhou, Mian Xiang, Qian Ning· Journal of King Saud Univers...· 0 citations
This paper proposes ADA-CS, a plug-and-play module compatible with any ADA or ASFDA framework, and introduces a CSS metric to quantify the Concept Shift Severity across domains, revealing that non-negligible concept shift exists in many transfer tasks.
Motivated by privacy concerns and the high cost of measured data transmission, source-free domain adaptation (SFDA) has attracted increasing attention for intelligent fault diagnosis. Instead of accessing raw measured source-domain data, SFDA transfers knowledge from pre-trained source models to target domains. However...
Yi-Ming Yuan, Kang Wu, Xing-Xing Jiang et al.· Measurement science and tech...· 0 citations
MASA (Multimodal-LLM-Anchored Semantic Adaptation), which complements model-internal evidence with structured semantic descriptions from a frozen multimodal large language model (MLLM) to limit inference cost.
Zhen-Bin Wang, Lei Zhang, Li-Tuan Wang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.