Mapping the diverse topologies of protein-protein interaction fitness landscapes

preprint OA: closed
📄 Open PDF View at publisher

Abstract

De novo binder discovery is unpredictable and inefficient due to a lack of quantitative understanding of protein-protein interaction (PPI) sequence-function landscapes. Here, we use our PANCS-Binder technology to perform >1,300 independent selections of various library sizes and compositions of a randomized small protein to identify binders to a panel of 96 distinct target proteins. For successful selections, we discovered reproducible fitness landscapes that group into a few, target-specific, clusters. Each cluster defines a minimal binding motif whose frequency is inversely proportional to the number of specified amino acids (∼2–8) and determines selection success, which is quantifiable by the density of binders to the target within a theoretical sequence space. We leverage these data to develop a supervised contrastive learning approach that discriminates binders from non-binders and demonstrates generalization beyond a threshold amount of data. Together, this framework renders PPI landscapes measurable and predictive, accelerating de novo binder discovery and optimization.

My notes (saved in your browser only)

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2025) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.

Source provenance

europepmc
last seen: 2026-05-20T01:45:00.602351+00:00
unpaywall
last seen: 2026-08-06T06:41:17.185923+00:00