Class balanced diabetic retinopathy image synthesis using a latent diffusion framework

Diabetic Retinopathy (DR) is one of the major causes of preventable blindness globally. Automated screening for DR is critical but is severely hindered by data scarcity and class imbalance. Real-world datasets, such as APTOS 2019, exhibit extreme class imbalance, where sight-threatening classes are statistically rare. Traditional augmentation fails to capture complex pathological features, while Generative Adversarial Networks (GANs) often suffer from mode collapse. Existing diffusion approaches typically operate in pixel space, limiting image resolution and requiring extreme computational resources. To address this, we introduce the Diabetic Retinopathy Latent Diffusion Synthesizer (DR-LDS), a class-conditional framework that synthesizes high-fidelity, $$512 \\times 512$$ fundus images natively deployable on consumer-grade hardware. By leveraging a fine-tuned Variational Autoencoder (VAE) for domain-adapted latent space compression, alongside an optimized U-Net, our method achieves superior anatomical realism with convergence in just 150 epochs (at 17 min 22 s per epoch). Extensive benchmarking demonstrates that DR-LDS outperforms state-of-the-art baselines, achieving a Fréchet Inception Distance (FID) of 8.05. When used to augment minority classes to a uniform target for the APTOS 2019 dataset across 11 deep learning architectures, our synthetic data significantly improved diagnostic accuracy by up to 15.25% ( $$p < 0.001$$ ), enabling a lightweight SqueezeNet model to reach 95.60% validation accuracy and an F1 Score of 0.96. Moreover, rigorous data leakage analyses and Explainable AI (XAI) verify that DR-LDS learns genuine pathological biomarkers. This study establishes DR-LDS as a highly resource-efficient solution to the medical data bottleneck. The source code is available at: https://github.com/Touhid-Alam/DR-LDS

Authors

Institutions

Publication Details

Journal
Discover Artificial Intelligence
Published
2026-09-21
DOI
https://doi.org/10.1007/s44163-026-02184-1
Primary Topic
Retinal Imaging and Analysis
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Class balanced diabetic retinopathy image synthesis using a latent diffusion framework

Md. Kishor Morol, Dip Nandi, Touhid Alam, Mashiour Rahman et al.
Discover Artificial Intelligence
Retinal Imaging and Analysis
article

Class balanced diabetic retinopathy image synthesis using a latent diffusion framework

Md. Kishor Morol, Dip Nandi, Touhid Alam, Mashiour Rahman, Md. Abdullah-Al Jubair, Tze Hui Liew
article en

Abstract

Diabetic Retinopathy (DR) is one of the major causes of preventable blindness globally. Automated screening for DR is critical but is severely hindered by data scarcity and class imbalance. Real-world datasets, such as APTOS 2019, exhibit extreme class imbalance, where sight-threatening classes are statistically rare. Traditional augmentation fails to capture complex pathological features, while Generative Adversarial Networks (GANs) often suffer from mode collapse. Existing diffusion approaches typically operate in pixel space, limiting image resolution and requiring extreme computational resources. To address this, we introduce the Diabetic Retinopathy Latent Diffusion Synthesizer (DR-LDS), a class-conditional framework that synthesizes high-fidelity, $$512 \times 512$$ fundus images natively deployable on consumer-grade hardware. By leveraging a fine-tuned Variational Autoencoder (VAE) for domain-adapted latent space compression, alongside an optimized U-Net, our method achieves superior anatomical realism with convergence in just 150 epochs (at 17 min 22 s per epoch). Extensive benchmarking demonstrates that DR-LDS outperforms state-of-the-art baselines, achieving a Fréchet Inception Distance (FID) of 8.05. When used to augment minority classes to a uniform target for the APTOS 2019 dataset across 11 deep learning architectures, our synthetic data significantly improved diagnostic accuracy by up to 15.25% ( $$p < 0.001$$ ), enabling a lightweight SqueezeNet model to reach 95.60% validation accuracy and an F1 Score of 0.96. Moreover, rigorous data leakage analyses and Explainable AI (XAI) verify that DR-LDS learns genuine pathological biomarkers. This study establishes DR-LDS as a highly resource-efficient solution to the medical data bottleneck. The source code is available at: https://github.com/Touhid-Alam/DR-LDS

Discover Artificial IntelligenceVol. 6(1)
American International University-Bangladesh (BD), Multimedia University (MY), Cloud Computing Center (CN)
Decent work and economic growth
Openalex Percentile: Top 11%
Retinal Imaging and Analysis
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.