Skip to content
Preprint

Domain-Specific Self-Supervised Representation Learning for Retinal Fundus Classification

Aug 2026 · 0 citations · 22 references
Computer Science

TL;DR

The results show that tailoring augmentation strategies to the characteristics of retinal images plays a critical role in improving performance, and even under constrained settings, lightweight SSL frameworks can learn transferable representations that reduce dependence on large annotated datasets and achieve competitive results.

Abstract

Despite the growing number of public datasets, annotated medical images remain scarce. Supervised learning methods achieve strong performance on many benchmarks, however require large amounts of labeled data, which are costly and time-consuming to obtain in the medical domain. To address this limitation, contrastive self-supervised learning (SSL) has emerged as a promising alternative for learning useful representations from unlabeled data. In this work, we investigate two SSL frameworks, SimSiam and SimCLR, for retinal disease classification from fundus images. We focus on understanding how augmentation strategies and training parameters influence representation learning under resource-constrained settings. Given limited data and computational capacity, we explore the feasibility of training SSL models with small batch sizes incorporated with retinal-specific augmentation techniques. Through a series of experiments, we assess the quality of learned representations via linear evaluation and fine-tuning across downstream tasks, including multi-disease classification and diabetic retinopathy grading. Our results show that tailoring augmentation strategies to the characteristics of retinal images plays a critical role in improving performance. Even under constrained settings, lightweight SSL frameworks can learn transferable representations that reduce dependence on large annotated datasets and achieve competitive results.

View source

Similar papers

Aug 2026

Self-supervised vision transformers for intelligent OCT-based retinal disease classification and severity assessment

Vision Transformers (ViTs) have demonstrated strong performance in medical image analysis due to their ability to model long-range dependencies through self-attention mechanisms. However, training such models typically requires large annotated datasets, which are often difficult to obtain in medical imaging. To address...

Abdulkarem Almshnanah, Rehab M. Duwairi · 0 citations
Open access Sep 2026

Improving Generalization of Deep Learning for Glaucoma Classification Under Real-World Domain Shift via Unsupervised Domain Adaptation

Purpose To develop and evaluate an unsupervised domain adaptation (UDA) framework for glaucoma classification from fundus images that improves the generalizability of deep learning (DL) models across heterogeneous imaging characteristics and clinical settings. Methods We developed an adversarial UDA framework that adap...

Homa Rashidisabet, R. V. Paul Chan, T. Vajaranant et al. · 0 citations
Review Open access Aug 2026

A trustworthy cross-domain AI framework for fundus disease classification using hybrid CNN fusion and supervised domain adaptation

FusionEye-Net combines strong internal performance with substantially improved cross-domain generalization through lightweight adaptation, underscoring the critical role of external validation and domain-aware fine-tuning in formulating trustworthy, adaptive AI-assisted decision-support tools that can safely augment au...

Ali M. Duhaim, A. M. Al-Bakry · 0 citations
Conference Aug 2026

Tiling to transfer: patch-level contrastive learning framework for cross-dataset generalization on downstream diabetic retinopathy grading tasks

This study presents a patch-based contrastive learning framework (Patch-SimCLR) designed to improve the generalizability and calibration of diabetic retinopathy classification across heterogeneous fundus datasets. By extracting overlapping, and non-overlapping patches and leveraging contrastive pretraining, the model l...

Usman Ali, A. A. Imam, R. Apong · 0 citations
Open access Sep 2026

Selective Confidence-Guided Projection-Based Encoding for Medical Image Classification

Selective Confidence-guided Projection-based Encoding (SCOPE) is proposed, a conflict-aware KD framework comprising Selective Relation Alignment (SRA) and Gradient Conflict Resolution (GCR), which demonstrates competitive predictive performance, improved training stability, and low computational overhead.

Tao Chen, Chuan Zhou, Yi-Fan Wang et al. · 0 citations
Conference Aug 2026

Optimizing multi-dimensional retinal segmentation: a performance-practicality analysis

This study introduces a multi-dimensional framework to systematically evaluate the generalization and computational efficiency of deep learning models for retinal vessel segmentation. Using the DRIVE and STARE datasets, this research evaluates the performance of SegNet, U-Net, and DeepLab across in-domain, cross-domain...

Yu-Lung Chang, Wei-Chun Chi, Yen-Ching Chang · 1 citation · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.