Jul 2026· IEEE transactions on multimedia· Vol abs/2607.10087, pp. 1-13· 0 citations· 53 references
Computer Science
TL;DR
CVKD-UDA is proposed, which revisits voxel size as a core design factor to construct domain-similar representations and leverages cross-view complementary cues to balance transferability and discriminability of the warm-up model.
Abstract
3D unsupervised domain adaptive (UDA) segmentation mitigates the high cost of manual annotations of the new domain data. Self-training has emerged as the dominant approach in this area, where its success heavily depends on a well-initialized warm-up model to generate reliable pseudo labels. However, existing methods often depend on source supervision or output-level adversarial alignment to obtain the warm-up model, which suffer from limited generalization and training instability due to the large domain gap between domains. Constructing domain-similar representations is an effective way to bridge this gap. In this work, we propose CVKD-UDA, which revisits voxel size as a core design factor to construct domain-similar representations and leverages cross-view complementary cues to balance transferability and discriminability of the warm-up model. First, we generate two complementary views by varying voxel sizes and introduce a cross-view knowledge distillation (CVKD) to enhance generalization and target perception of the model. Second, to balance transferability and discriminability, we design a lightweight Decouple-Adapter and an auxiliary imitation classifier to decouple cross-view knowledge transfer. Extensive experiments on two benchmarks demonstrate that CVKD-UDA effectively improves the performance of self-training methods and provides a new perspective for 3D UDA segmentation. Our code will be available at GitHub.
Efficient Unsupervised Domain Adaptation (EUDA) is proposed, a parameter-efficient framework that leverages a frozen DINOv2 backbone as a feature extractor and updates only a lightweight bottleneck and classification head to promote both discriminative learning and cross-domain alignment.
Ali Abedi, Q. M. Jonathan Wu, Ning Zhang et al.· International Journal of Mac...· 9 citations
Experimental results show that TransPileSiam consistently improves downstream recognition under limited-label conditions and improves robustness, label efficiency, and cross-domain generalization for practical SEI tasks.
Junwei Peng, Siyang Xu, Jiao Wang et al.· International Conference on...· 0 citations
PolyMem, an exemplar-free approach that implicitly models rich high-order statistics of the feature distribution to enhance cross-domain robustness, is introduced that effectively alleviates the performance discrepancy while improving the model's performance across domains.
Jin-Ge Ma, Gautham Vinod, Bruce Coburn et al.· 0 citations
Single domain generalization (SDG) aims to learn a model from one labeled source domain that generalizes to unseen target domains. A common strategy is to enrich the source distribution with augmented or generated samples, and recent text-to-image (T2I) diffusion models provide a strong generative prior for this purpos...
Zhi-Peng Xu, De Cheng, Xinyang Jiang et al.· 0 citations
The development of deep learning over the past decade has revolutionized medical imaging segmentation, allowing the extraction of precise descriptors from large volumes to characterize pathologies. Data augmentation is a technique widely regarded as a way to improve model training. It includes simple transformations li...
Robin Trombetta, Carole Lartizien· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.