Published 2024 | Version v2
Journal article Open

De Novo Design of Peptide Binders to Conformationally Diverse Targets with Contrastive Language Modeling

Description

Designing binders to target undruggable proteins presents a formidable challenge in drug discovery, requiring innovative approaches to overcome the lack of putative binding sites. Recently, generative models have been trained to design binding proteins via three-dimensional structures of target proteins, but as a result, struggle to design binders to disordered or conformationally unstable targets. In this work, we provide a generalizable algorithmic framework to design short, target-binding linear peptides, requiring only the amino acid sequence of the target protein. To do this, we propose a process to generate naturalistic peptide candidates through Gaussian perturbation of the peptidic latent space of the ESM-2 protein language model, and subsequently screen these novel linear sequences for target-selective interaction activity via a CLIP-based contrastive learning architecture. By integrating these generative and discriminative steps, we create a Peptide Prioritization via CLIP (PepPrCLIP) pipeline and validate highly-ranked, target-specific peptides experimentally, both as inhibitory peptides and as fusions to E3 ubiquitin ligase domains, demonstrating functionally potent binding and degradation of conformationally diverse protein targets in vitro.  Overall, our design strategy provides a modular toolkit for designing short binding linear peptides to any target protein without the reliance on stable and ordered tertiary structure, enabling generation of programmable modulators to undruggable and disordered proteins such as transcription factors and fusion oncoproteins.

Files

PCOligos-FACSFiles-20240414T170407Z-001.zip

Files (4.8 GB)

Name Size Download all
md5:e5af66ce53e36f046b8387236742cdc4
706.1 kB Download
md5:5908abb32ffabd63a5a74852703a8438
213.5 kB Download
md5:7f95e3a06c376faaca8d47743c2254c7
10.0 MB Preview Download
md5:ea32d4203bfc8b14dc8a1223ca3d6ada
1.2 MB Download
md5:16cf34c3ff13458983f9fa4a9c77dbf8
4.8 GB Preview Download