CholBindNet: Interpretable Neural Networks for Cholesterol Binding Site Prediction

preprint OA: closed
Full text JSON View at publisher

Abstract

Cholesterol is a key modulator of membrane protein structure and function, yet predicting cholesterol binding sites remains challenging due to its undrug-like physicochemical properties. Here, we curated more than 800 high-resolution transmembrane protein structures containing cholesterol and developed an interpretable, atom-based deep-learning framework, CholBindNet, comprising four model architectures: a 3D convolutional neural network, a graph neural network, a graph attention network, and a graph convolutional network. A Positive–Unlabeled (PU) training strategy was employed to address the scarcity of true negative samples resulting from the promiscuous nature of cholesterol binding. We show that CholBindNet substantially outperforms existing deep-learning models trained on general ligand-binding datasets. The performance and generalizability of the model were further demonstrated by rapidly assessing strong, median, and weak cholesterol-binding sites in the PIEZO2 ion channel in excellent agreement with computationally expensive all-atom molecular dynamics (MD) simulations. Additionally, strong model interpretability was achieved for CholBindNet through atom-level feature encoding, Grad-CAM visualization, and attention-based scoring analysis. Overall, CholBindNet provides an efficient and scalable approach for predicting cholesterol binding sites on membrane proteins, achieving performance comparable to MD simulations while offering mechanistic biophysical insights beyond amino-acid sequence. This work hence lays the foundation for future development of deep-learning models targeting membrane protein drug-binding sites and cholesterol-modulated therapeutics. Significance Statement Deep-learning models for ligand-binding prediction have advanced rapidly, yet those trained on soluble proteins perform poorly for membrane proteins, particularly for cholesterol binding. We introduce CholBindNet, a set of neural network models specifically designed to identify cholesterol-binding sites in transmembrane proteins. CholBindNet substantially outperforms existing deep-learning approaches and accurately ranks strong, intermediate, and weak cholesterol-binding sites in close agreement with computationally intensive all-atom molecular dynamics simulations. This work provides a practical and scalable alternative to long-timescale simulations for studying cholesterol–protein interactions. The curated cholesterol benchmark and open-source models will enable broader adoption of deep learning for investigating lipid regulation of membrane proteins and for guiding the design of drugs targeting membrane-embedded binding sites.
Full text 2,745 characters · extracted from oa-doi-fallback · click to expand
Abstract Cholesterol is a key modulator of membrane protein structure and function, yet predicting cholesterol binding sites remains challenging due to its undrug-like physicochemical properties. Here, we curated more than 800 high-resolution transmembrane protein structures containing cholesterol and developed an interpretable, atom-based deep-learning framework, CholBindNet, comprising four model architectures: a 3D convolutional neural network, a graph neural network, a graph attention network, and a graph convolutional network. A Positive–Unlabeled (PU) training strategy was employed to address the scarcity of true negative samples resulting from the promiscuous nature of cholesterol binding. We show that CholBindNet substantially outperforms existing deep-learning models trained on general ligand-binding datasets. The performance and generalizability of the model were further demonstrated by rapidly assessing strong, median, and weak cholesterol-binding sites in the PIEZO2 ion channel in excellent agreement with computationally expensive all-atom molecular dynamics (MD) simulations. Additionally, strong model interpretability was achieved for CholBindNet through atom-level feature encoding, Grad-CAM visualization, and attention-based scoring analysis. Overall, CholBindNet provides an efficient and scalable approach for predicting cholesterol binding sites on membrane proteins, achieving performance comparable to MD simulations while offering mechanistic biophysical insights beyond amino-acid sequence. This work hence lays the foundation for future development of deep-learning models targeting membrane protein drug-binding sites and cholesterol-modulated therapeutics. Significance Statement Deep-learning models for ligand-binding prediction have advanced rapidly, yet those trained on soluble proteins perform poorly for membrane proteins, particularly for cholesterol binding. We introduce CholBindNet, a set of neural network models specifically designed to identify cholesterol-binding sites in transmembrane proteins. CholBindNet substantially outperforms existing deep-learning approaches and accurately ranks strong, intermediate, and weak cholesterol-binding sites in close agreement with computationally intensive all-atom molecular dynamics simulations. This work provides a practical and scalable alternative to long-timescale simulations for studying cholesterol–protein interactions. The curated cholesterol benchmark and open-source models will enable broader adoption of deep learning for investigating lipid regulation of membrane proteins and for guiding the design of drugs targeting membrane-embedded binding sites. Competing Interest Statement The authors have declared no competing interest.

Text is read by the "Ask this paper" AI Q&A widget below. Extraction quality varies by source — PMC NXML preserves structure cleanly, OA-HTML may include some navigation residue, and OA-PDF can have broken hyphenation. The publisher copy (via DOI) is the canonical version.

My notes (saved in your browser only)

Ask this paper AI returns verbatim quotes from the full text · source: oa-doi-fallback

Answers must be backed by verbatim quotes from this paper's full text. Hallucinated quotes are dropped automatically; if no verbatim passage answers the question, we say so. How this works

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2026) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.

Source provenance

europepmc
last seen: 2026-05-20T01:45:00.602351+00:00