Contrastive Learning for Robust Cell Annotation and Representation from Single-Cell Transcriptomics
preprint
OA: closed
CC-BY-4.0
Abstract
Batch effects are a significant concern in single-cell RNA sequencing (scRNA-Seq) data analysis, where variations in the data can be attributed to factors unrelated to cell types. This can make downstream analysis a challenging task. In this study, we present a novel DL approach using contrastive learning and a carefully designed loss function for learning a generalizable embedding space from scRNA-Seq data. We call this model CELLULAR: CELLUlar contrastive Learning for Annotation and Representation. When benchmarked against multiple established methods for scRNA-Seq integration, CELLULAR outperforms existing methods in learning a generalizable embedding space on multiple datasets. Cell annotation was also explored as a downstream application for the learned embedding space. When compared against multiple well-established methods, CELLULAR demonstrates competitive performance with top cell classification methods in terms of accuracy, balanced accuracy, and F1 score. CELLULAR is also capable of performing novel cell type detection. These findings assess the biological relevance of the model’s learned embedding space by demonstrating the robustness of its cell representations across various applications. The model has been structured into an opensource Python package, specifically designed to simplify and streamline its usage for bioinformaticians and other scientists interested in cell representation learning.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2024) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- europepmc
- last seen: 2026-05-20T01:45:00.602351+00:00
- unpaywall
- last seen: 2026-06-05T02:00:03.366016+00:00
License: CC-BY-4.0