Single-cell transcriptomic integrated with machine learning reveals retinal cell-specific biomarkers in diabetic retinopathy

preprint OA: closed CC-BY-NC-ND-4.0
⚙ AI-generated summary by claude@2026-06, 2026-06-08 ⓘ

This study integrated single-cell transcriptomics with machine learning to identify novel cell-specific biomarkers associated with diabetic retinopathy.

One-sentence paraphrase of the abstract; not a substitute for reading it. No clinical advice. How this works

Abstract

Diabetic retinopathy (DR) remains a principal cause of vision impairment worldwide, involved complex retinal cellular pathophysiology that remains incompletely understood. To elucidate cell-type-specific molecular signatures underlying DR, we generated a high-resolution single-cell transcriptomic atlas of 297,121 retinal cells from 20 Chinese donors, including non-diabetic controls (26.4%), diabetic without retinopathy (23.4%) and DR (50.2%). Following rigorous quality control, batch-effect correction, and clustering and annotation, 10 major retinal cell populations were delineated. Differential expression analyses across disease states within each cell type yielded candidate gene sets, which were further refined via a multi-stage machine-learning pipeline combining L1-regularized logistic regression and recursive feature elimination with cross-validation, alongside bootstrap stability selection. Resulting cell-type-specific classifiers achieved high accuracy (79–95%) and AUCs (0.85–0.99) in distinguishing DR disease states. Enrichment analyses implicated immune activation, oxidative stress, neurodegeneration and synaptic dysfunction pathways across multiple cell types in retina. Integrating 567 unique marker genes from all cell types, a general multilayer perceptron classifier achieved 95.31% overall accuracy on held-out test data, demonstrating the translational potential of these signatures for non-diabetic controls, diabetic without retinopathy and DR classification. This high-resolution atlas and the accompanying analytic framework provide a robust computational framework for biomarker discovery, mechanistic insight and targeted intervention strategies in diabetic retinal diseases.
Full text 1,802 characters · extracted from oa-doi-fallback · click to expand
Abstract Diabetic retinopathy (DR) remains a principal cause of vision impairment worldwide, involved complex retinal cellular pathophysiology that remains incompletely understood. To elucidate cell-type-specific molecular signatures underlying DR, we generated a high-resolution single-cell transcriptomic atlas of 297,121 retinal cells from 20 Chinese donors, including non-diabetic controls (26.4%), diabetic without retinopathy (23.4%) and DR (50.2%). Following rigorous quality control, batch-effect correction, and clustering and annotation, 10 major retinal cell populations were delineated. Differential expression analyses across disease states within each cell type yielded candidate gene sets, which were further refined via a multi-stage machine-learning pipeline combining L1-regularized logistic regression and recursive feature elimination with cross-validation, alongside bootstrap stability selection. Resulting cell-type-specific classifiers achieved high accuracy (79–95%) and AUCs (0.85–0.99) in distinguishing DR disease states. Enrichment analyses implicated immune activation, oxidative stress, neurodegeneration and synaptic dysfunction pathways across multiple cell types in retina. Integrating 567 unique marker genes from all cell types, a general multilayer perceptron classifier achieved 95.31% overall accuracy on held-out test data, demonstrating the translational potential of these signatures for non-diabetic controls, diabetic without retinopathy and DR classification. This high-resolution atlas and the accompanying analytic framework provide a robust computational framework for biomarker discovery, mechanistic insight and targeted intervention strategies in diabetic retinal diseases. Competing Interest Statement The authors have declared no competing interest.

Text is read by the "Ask this paper" AI Q&A widget below. Extraction quality varies by source — PMC NXML preserves structure cleanly, OA-HTML may include some navigation residue, and OA-PDF can have broken hyphenation. The publisher copy (via DOI) is the canonical version.

My notes (saved in your browser only)

⚙ Ask this paper AI returns verbatim quotes from the full text · source: oa-doi-fallback ⓘ

Answers must be backed by verbatim quotes from this paper's full text. Hallucinated quotes are dropped automatically; if no verbatim passage answers the question, we say so. How this works

Citation neighborhood (sparse)

Too few in-corpus citations on either side for a chart; here are the lists.

Cites (1)

References (63)

Source provenance

crossref
last seen: 2026-06-04T01:00:35.597411+00:00
europepmc
last seen: 2026-05-20T01:45:00.602351+00:00
unpaywall
last seen: 2026-05-22T02:00:06.705733+00:00
License: CC-BY-NC-ND-4.0