COSMO replaces expert-to-expert guidance with co-adaptation through an anchored shared consensus and achieves state-of-the-art performance under matched VLM backbones, indicating that it better balances the retention of valid source-derived evidence with the absorption of complementary VLM evidence.
Abstract
Source-free domain adaptation (SFDA) adapts a source-trained model to an unlabeled target domain without source data, a practical setting under privacy or storage constraints. Yet its self-generated supervision can reinforce source bias under substantial domain shifts. Pretrained vision-language models (VLMs) offer complementary semantic knowledge, but the relative reliability of the source model and VLM varies across target samples. Existing cross-model guidance does not explicitly account for this variation and may overwrite valid source-derived evidence under conflict, a failure we term source-derived evidence forgetting. We formulate VLM-guided SFDA as a sample-wise reliability-allocation problem and propose Consensus-Driven Shift Modulation (COSMO). COSMO replaces expert-to-expert guidance with co-adaptation through an anchored shared consensus. It first forms a sample-specific initial consensus that favors the more concentrated prediction. During adaptation, COSMO re-aggregates both branches'evolving evidence and regulates how far the resulting consensus moves from its initial anchor based on consensus uncertainty and training progress. This keeps the shared supervision anchored yet adaptive. Across four benchmarks, COSMO achieves state-of-the-art performance under matched VLM backbones. Further analyses indicate that it better balances the retention of valid source-derived evidence with the absorption of complementary VLM evidence.
A Peer-level Heterogeneous Perception Framework is proposed that departs from such paradigms by enabling balanced collaboration between heterogeneous models by introducing an auxiliary domain that is significantly different from the target domain and employ an auxiliary model with the same architecture as the source mo...
Zhi-Ze Wu, Yu-Tao Fu, Huan-Xin Zou et al.· Multimedia Systems· 0 citations
This work proposes a novel criterion, termed Maximum Refinement Decision score (MRD-score), which replaces softmax with median centering to normalize source models’ predictions along both positive and negative axes, thereby harnessing both affirmative and complementary guidance.
Bing-Tao Zhou, Mian Xiang, Qian Ning· Journal of King Saud Univers...· 0 citations
This paper proposes ADA-CS, a plug-and-play module compatible with any ADA or ASFDA framework, and introduces a CSS metric to quantify the Concept Shift Severity across domains, revealing that non-negligible concept shift exists in many transfer tasks.
This paper proposes a novel SFDA with high-confidence sample selection and feature disentanglement for machinery fault diagnosis, which effectively alleviates the adverse influence of noisy pseudo-labels during the stage of adaptation.
Yi-Ming Yuan, Kang Wu, Xing-Xing Jiang et al.· Measurement science and tech...· 0 citations
MASA (Multimodal-LLM-Anchored Semantic Adaptation), which complements model-internal evidence with structured semantic descriptions from a frozen multimodal large language model (MLLM) to limit inference cost.
Zhen-Bin Wang, Lei Zhang, Li-Tuan Wang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.