Sep 2026· Proceedings of the Thirty-Fifth International Joint Conference on Artificial Intelligence· 0 citations· 37 references
TL;DR
Reciprocal Distillation with a Frozen Foundation Model (ReDi-FM), a novel framework with two core components: using structure-aware prompts from the source model to guide the frozen VFM to generate target-adapted supervision, and distilling the resulting knowledge back into the adapting model.
Abstract
Continual test-time adaptation (CTTA) adapts a pre-trained medical segmentation model online to an unlabeled target stream whose distribution changes over time. However, most existing CTTA methods rely on pseudo-labeling and self-supervised objectives, which inevitably yield noisy supervision under domain shifts. To mitigate this limitation, we introduce off-the-shelf Vision Foundation Models (VFMs) as external knowledge sources. Zero-shot VFMs are insufficient for medical segmentation because they lack medical semantics, yet they contain rich and heterogeneous generic knowledge. To exploit such external knowledge, we propose Reciprocal Distillation with a Frozen Foundation Model (ReDi-FM), a novel framework with two core components: using structure-aware prompts from the source model to guide the frozen VFM to generate target-adapted supervision, and distilling the resulting knowledge back into the adapting model. For robust distillation, we introduce two complementary objectives: uncertainty-driven hard distillation for precise guidance in ambiguous regions, and hard class-balanced soft distillation for richer supervision of under-represented and challenging structures. A consensus-aware gate further stabilizes adaptation when the two teachers disagree. Extensive experiments on multi-domain medical segmentation benchmarks demonstrate that ReDi-FM outperforms state-of-the-art CTTA methods. Code is available at https://github.com/M4cheal/ReDi-FM.
Medical image segmenters often get worse when sites, scanner vendors, or protocols change. Continual test-time adaptation (CTTA) addresses this problem without target labels, but it can be impossible to update a model on a non-stationary stream and can lead to a lot of errors. We examine a more reasonable and meaningfu...
DABAL is proposed, a semi-supervised framework designed to improve both supervision reliability and contour localization and introduces a Dynamic-static Domain Adaptive Adapter (DDAA) into the Segment Anything Model (SAM) encoder to preserve stable structural priors while providing input-dependent feature compensation...
Wei-Yan Zeng, Zhi-Ming Cheng, Bin Lin et al.· Computerized Medical Imaging...· 0 citations
Few-shot medical image segmentation relies on dense, boundary-sensitive prototype matching, yet common pre-training objectives mainly optimize global alignment or reconstruction, creating an objective gap that hurts boundary delineation and increases adaptation cost. This raises the question: how to pre-train represent...
Shou-Peng Chen, Yi-Ming Miao, Li-Mei Peng et al.· Proceedings of the Thirty-Fi...· 0 citations
Medical image segmentation plays a pivotal role in computer-aided diagnosis. However, the scarcity of annotated data severely hinders the deployment of deep learning models. Few-shot learning (FSL) is designed to achieve rapid adaptation to unseen classes using limited labeled samples, among which prototype-based metho...
Wen-Jie Meng, Kai Liu, Minghui Wang· IEEE Transactions on Medical...· 0 citations
Semi-supervised medical image segmentation methods have drawn wide attention as they reduce reliance on heavily annotated data. However, existing models suffer from confirmation bias with limited annotations, and structural or parameter coupling hinders self-correction, especially for medical images with ambiguous boun...
Dong-Sheng Wang, Xiao-Han Lang· Biomedical engineering and p...· 0 citations
Medical image segmentation remains challenging in practical deployment, as models often struggle to generalize beyond the distributions covered by their training data and high-quality pixel-level annotations are typically unavailable for adaptation. Inspired by the cross-task transferability of large language models, w...
Xiao-Ye Liang, Ye Yan, Ming-Ze Yin et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.