ASSESSMENT OF ALPHAFOLD PROTEIN MODELS FOR SMALL-MOLECULE LIGAND DOCKING

preprint OA: closed
Full text JSON View at publisher

Abstract

ABSTRACT Molecular docking is a powerful computational tool for predicting protein-ligand interactions, widely employed in drug discovery. However, its effectiveness is often constrained by the availability of experimentally resolved X-ray protein structures, a process that is both time consuming and resource-intensive. AlphaFold (AF), a deep learning method, offers an efficient alternative by predicting high-accuracy 3D protein structures directly from amino acid sequences. This study assesses the utility of AF-generated protein models for fragment and larger ligand docking with Glide, a widely used docking approach. The docking workflow is evaluated in an unbiased manner by carrying out binding site identification with FTMap, a binding hot spot prediction software. We show that fragment docking to AF models outperforms docking to the respective unbound protein crystal structures, and performs comparably to docking to the corresponding ligand bound structures when using an unbiased approach. Leveraging computational efficiency of AF model generation, we also employ ensembles of AF models to incorporate protein flexibility. Results show that docking to AF ensembles improves larger-ligand docking compared to docking to singular AF models and outperforms docking to unbound structures. The results provide insights into the effectiveness of integrating AF protein models into docking procedures, highlighting the potential for streamlining computational drug discovery processes. STATEMENT OF SIGNIFICANCE This work addresses a critical bottleneck in computational drug discovery by demonstrating that AlphaFold (AF) models can serve as an alternative or complement to experimental structures for molecular docking. Specifically, a systematic study assessing Glide docking to rigid AF protein models and ensembles of models compared to experimentally determined ligand-bound and unbound protein X-ray structures was performed. The evaluation employs an unbiased methodology using FTMap-identified binding sites, eliminating the need for prior knowledge of the native ligand binding location. Additionally, protein flexibility is incorporated through a multiseed ensemble approach that generates a conformational ensembles of AF models at minimal computational cost, improving the docking accuracy without the need for ligand-bound templates.
Full text 2,437 characters · extracted from oa-doi-fallback · click to expand
ABSTRACT Molecular docking is a powerful computational tool for predicting protein-ligand interactions, widely employed in drug discovery. However, its effectiveness is often constrained by the availability of experimentally resolved X-ray protein structures, a process that is both time consuming and resource-intensive. AlphaFold (AF), a deep learning method, offers an efficient alternative by predicting high-accuracy 3D protein structures directly from amino acid sequences. This study assesses the utility of AF-generated protein models for fragment and larger ligand docking with Glide, a widely used docking approach. The docking workflow is evaluated in an unbiased manner by carrying out binding site identification with FTMap, a binding hot spot prediction software. We show that fragment docking to AF models outperforms docking to the respective unbound protein crystal structures, and performs comparably to docking to the corresponding ligand bound structures when using an unbiased approach. Leveraging computational efficiency of AF model generation, we also employ ensembles of AF models to incorporate protein flexibility. Results show that docking to AF ensembles improves larger-ligand docking compared to docking to singular AF models and outperforms docking to unbound structures. The results provide insights into the effectiveness of integrating AF protein models into docking procedures, highlighting the potential for streamlining computational drug discovery processes. STATEMENT OF SIGNIFICANCE This work addresses a critical bottleneck in computational drug discovery by demonstrating that AlphaFold (AF) models can serve as an alternative or complement to experimental structures for molecular docking. Specifically, a systematic study assessing Glide docking to rigid AF protein models and ensembles of models compared to experimentally determined ligand-bound and unbound protein X-ray structures was performed. The evaluation employs an unbiased methodology using FTMap-identified binding sites, eliminating the need for prior knowledge of the native ligand binding location. Additionally, protein flexibility is incorporated through a multiseed ensemble approach that generates a conformational ensembles of AF models at minimal computational cost, improving the docking accuracy without the need for ligand-bound templates. Competing Interest Statement The authors have declared no competing interest.

Text is read by the "Ask this paper" AI Q&A widget below. Extraction quality varies by source — PMC NXML preserves structure cleanly, OA-HTML may include some navigation residue, and OA-PDF can have broken hyphenation. The publisher copy (via DOI) is the canonical version.

My notes (saved in your browser only)

Ask this paper AI returns verbatim quotes from the full text · source: oa-doi-fallback

Answers must be backed by verbatim quotes from this paper's full text. Hallucinated quotes are dropped automatically; if no verbatim passage answers the question, we say so. How this works

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2026) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.

Source provenance

europepmc
last seen: 2026-05-20T01:45:00.602351+00:00