Aug 2026· Medical Image Analysis· Vol 115, pp.
104293
· 0 citations· 291 references
Medicine
TL;DR
This review provides a comprehensive and structured synthesis of FMs in medical image analysis by systematically organizing studies into two primary categories: vision-only foundation models (VFMs) and vision-language foundation models (VLFMs), based on their architectural foundations, training strategies, and downstream clinical tasks.
Abstract
Recent advancements in foundation models (FMs) have catalyzed a paradigm shift in medical image analysis. Unlike traditional task-specific artificial intelligence (AI) models, FMs leverage large-scale datasets to learn generalized representations that can be adapted to downstream clinical applications. Despite the rapid proliferation of FM research in medical imaging, there is a lack of unified synthesis that systematically maps the evolution of architectures, training paradigms, and clinical applications across modalities. To address this gap, this review provides a comprehensive and structured synthesis of FMs in medical image analysis by systematically organizing studies into two primary categories: vision-only foundation models (VFMs) and vision-language foundation models (VLFMs), based on their architectural foundations, training strategies, and downstream clinical tasks. A quantitative analysis was conducted on both VFMs and VLFMs to characterize temporal trends in dataset utilization and application domains, along with pooled performance and subgroup analyses. We also critically discuss persistent challenges, including cross-domain generalization, computational scalability, FM evaluation, fairness, and deployment. Finally, we identify key future research directions aimed at enhancing the robustness, interpretability, and clinical integration of FMs, thereby accelerating their translation into real-world medical practice.
This systematic review evaluated forty-two studies published between 2024 and 2026 across eleven databases using PRISMA 2020 guidelines and AI-specific quality frameworks and demonstrated high organ segmentation accuracy with Dice coefficients exceeding 0.90, whereas performance declined for infiltrative tumours, and e...
The retina provides a unique, non-invasive window into the human microvascular and central nervous systems. Recent advancements in deep learning have catalyzed the emergence of “Oculomics” transitioning automated retinal image analysis from localized ophthalmic diagnostics to holistic systemic health assessment. This s...
Hítalo Silva, Arlington Rodrigues, Rafael Albuquerque et al.· Research on Biomedical Engin...· 0 citations
This narrative review examines the evolution of artificial intelligence (AI) in healthcare, with a focus on the transition from early rule-based systems to modern deep learning architectures and their integration into clinical practice. We examine foundational technologies, including convolutional neural networks for i...
Abdulkadir Yıldırım, Ö. Özdemi̇r· Artificial Intelligence in M...· 0 citations
This review examines retinal image analysis from the perspective of clinical translation rather than benchmark-oriented model comparison, arguing that the next stage of retinal AI requires transferable representations, robust external evaluation, trustworthy uncertainty handling, and demonstrated value within real-worl...
Jia-Yao Chen, Xiao Feng, Hao Hu et al.· Artificial Intelligence in M...· 0 citations
Large language models (LLMs) have demonstrated strong capabilities across diverse domains, showing considerable potential in medicine. However, their application in medical settings remains limited by the scarcity of visual question answering (VQA) datasets that capture clinical reasoning and explicit image-text alignm...
Ling-Xuan Hou, Yu-Hua Xie, Yue Hu et al.· 0 citations
Satellite imagery is proposed as a novel pretraining domain for MedVFM development and benchmarking, motivated by its closer visual alignment with medical data and its freedom from the privacy constraints that limit medical datasets.
Lovre Antonio Budimir, Ming Gong, Alyssa Foong Quinney et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.