
==== Front
eLife
Elife
eLife
eLife
2050-084X
eLife Sciences Publications, Ltd

39240197
99702
10.7554/eLife.99702
version of record
Research Article
Structural Biology and Molecular Biophysics
Empowering AlphaFold2 for protein conformation selective drug discovery with AlphaFold2-RAVE
Gu Xinyu 12
Aranganathan Akashnathan https://orcid.org/0000-0001-7938-2141
13
Tiwary Pratyush https://orcid.org/0000-0002-2412-6922
pratyush.tiwary@gmail.com
124
1 https://ror.org/047s2c258 Institute for Physical Science and Technology, University of Maryland College Park United States
2 https://ror.org/047s2c258 University of Maryland Institute for Health Computing Bethesda United States
3 https://ror.org/047s2c258 Biophysics Program, University of Maryland College Park United States
4 https://ror.org/047s2c258 Department of Chemistry and Biochemistry, University of Maryland College Park United States
Panchenko Anna Reviewing Editor https://ror.org/02y72wh86 Queen's University Canada

Cui Qiang Senior Editor https://ror.org/05qwgg493 Boston University United States

06 9 2024
2024
13 RP9970227 5 2024
This manuscript was published as a preprint.12 5 2024

This manuscript was published as a reviewed preprint.01 7 2024

The reviewed preprint was revised.14 8 2024

© 2024, Gu et al
2024
Gu et al
https://creativecommons.org/licenses/by/4.0/ This article is distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use and redistribution provided that the original author and source are credited.

Small-molecule drug design hinges on obtaining co-crystallized ligand-protein structures. Despite AlphaFold2’s strides in protein native structure prediction, its focus on apo structures overlooks ligands and associated holo structures. Moreover, designing selective drugs often benefits from the targeting of diverse metastable conformations. Therefore, direct application of AlphaFold2 models in virtual screening and drug discovery remains tentative. Here, we demonstrate an AlphaFold2-based framework combined with all-atom enhanced sampling molecular dynamics and Induced Fit docking, named AF2RAVE-Glide, to conduct computational model-based small-molecule binding of metastable protein kinase conformations, initiated from protein sequences. We demonstrate the AF2RAVE-Glide workflow on three different mammalian protein kinases and their type I and II inhibitors, with special emphasis on binding of known type II kinase inhibitors which target the metastable classical DFG-out state. These states are not easy to sample from AlphaFold2. Here, we demonstrate how with AF2RAVE these metastable conformations can be sampled for different kinases with high enough accuracy to enable subsequent docking of known type II kinase inhibitors with more than 50% success rates across docking calculations. We believe the protocol should be deployable for other kinases and more proteins generally.

artificial intelligence
molecular simulation
AlphaFold
docking
enhanced sampling
Research organism

Human
http://dx.doi.org/10.13039/100000057 National Institute of General Medical Sciences R35GM142719 Tiwary Pratyush The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.Author impact statementAlphaFold2's inability to generate non-native conformations for docking is addressed by AI-molecular dynamics method AlphaFold2-RAVE, which produces Boltzmann-ranked conformations, enabling the retrospective discovery of type II kinase inhibitors.
publishing-routeprc
==== Body
pmcIntroduction

Despite the groundbreaking impact of AlphaFold2 (AF2) (Jumper et al., 2021; Mirdita et al., 2022) on the computational prediction of the ligand-free protein native apo structures, it appears that the determination of high-quality crystal or cryo-EM ligands-bound holo structures remains irreplaceable in the field of structure-based drug design. When ligands bind, residues within the protein pockets may adjust their side-chain rotamer configurations to optimize contacts with ligands, a phenomenon known as the induced fit effect. Furthermore, thermodynamic fluctuations induce protein dynamics and structural flexibility, leading to rearrangements of side-chains and even large-scale backbone movements that can reveal cryptic pockets (Amaro, 2019). These metastable conformations and associated cryptic pockets can be stabilized upon binding to specific ligands, a phenomenon known as conformational selection. Moreover, due to the similarity of native structures among protein homologs, it has been widely believed that ligands targeting highly diverse metastable conformations should result in better selectivity (Davis et al., 2011). It is thus highly desirable to account for metastable protein conformations or states, instead of investigating native states only. Several AF2-based techniques, including reduced multiple sequence alignment (rMSA) AF2 (or MSA subsampling AF2) (Del Alamo et al., 2022; Monteiro da Silva et al., 2024; Porter et al., 2023), AF2-cluster (Wayment-Steele et al., 2024), and AlphaFlow (Jing et al., 2024), have been devised to generate distinct decoy structures from native states. However, the suitability of these decoys for subsequent docking and virtual screening remains uncertain. Moreover, accurately assigning Boltzmann weights to decoys produced by those methods lacking a direct physical interpretation is challenging. Such a Boltzmann ranking is critical simply because of the explosion in number of decoys that can be hallucinated from AF2 or future generative AI methods, now including AlphaFold3 (Abramson et al., 2024).

The AF2RAVE protocol integrates rMSA AF2 and the machine learning-based reweighted autoencoded variational Bayes for enhanced sampling (RAVE) method (Wang et al., 2019; Wang and Tiwary, 2021; Mehdi et al., 2024) to systematically explore metastable states and accurately rank structures using Boltzmann weights (Vani et al., 2023). Subsequently, traditional grid-based docking methods or recent generative diffusion models, like DiffDock (Corso et al., 2022) and DynamicBind (Lu et al., 2024), can be employed to dock ligands with a few top-ranked structures in metastable states, enabling further virtual screening on large ligand libraries. In this work, we chose the docking method Glide XP (Friesner et al., 2004; Halgren et al., 2004; Friesner et al., 2006) and Induced Fit docking (IFD) (Sherman et al., 2006a; Sherman et al., 2006b) from the Schrödinger suite (Maestro, 2023) as our primary docking method. By combining these two steps, we propose the AF2RAVE-Glide workflow (Figure 1) as an innovative approach for small-molecule drug design, initiated from protein sequences. In this workflow, the conformational selection effect is addressed by probing various metastable states using AF2RAVE. Glide-IFD is then applied to account for the induced fit effect, refine the holo-like pockets further, and predict the ligand-bound holo structures.

Figure 1. A schematic of the AF2RAVE-Glide workflow.

A schematic of the AF2RAVE-Glide workflow: (i) decoy structures generated by reduced multiple sequence alignment (MSA) AlphaFold2 (AF2), (ii) regular space clustering and unbiased molecular dynamics (MD) simulations starting from cluster centers, (iii) state predictive information bottleneck model (SPIB, a reweighted autoencoded variational Bayes for enhanced sampling [RAVE] variant) to learn reaction coordinates from unbiased MD, (iv) enhanced sampling runs to calculate free energy landscape, (v) distinguish holo-like structures from decoys in metastable states based on Boltzmann rank and conduct Glide XP or Induced Fit docking (IFD) on holo-like structures for ligands targeting metastable states. (iv’) and (v’) The decoy structure set and the learnt SPIB coordinates are transferable to homologous systems.

In this work, we demonstrate this AF2RAVE-Glide workflow retrospectively on three different protein kinases and their type I and type II inhibitors. Protein kinases are involved in the regulation of various cellular pathways by catalyzing the hydrolyzing of ATP and transferring the phosphate group to substrate peptides/proteins. Dysfunctions of various kinases are known to cause human pathologies and cancers. The human genome contains about 500 protein kinases which share highly conserved structures in their catalytic ATP-binding pocket, due to the selection pressure toward functional catalysis. This poses significant challenges in developing selective small-molecule ATP-competitive kinase inhibitors, as they must effectively target the intended kinase while avoiding off-target interactions and associated side effects. Research efforts aimed at achieving selectivity have led to the development of four distinct types of kinase inhibitors (Müller et al., 2015), including two types of ATP-competitive kinase inhibitors: type I inhibitors bind to the ATP-binding site adopting the active catalytic state, while type II inhibitors target the binding pockets adjacent to the ATP-binding site adopting the inactive state. These two states primarily differ in the configurations of their activation loop (A-loop), which is a flexible loop of approximately 20 residues. In the active state, the A-loop is ‘extended’ to create a cleft for substrate peptides to bind, while in the inactive state, it is collapsed or ‘folded’ onto the protein surface, blocking substrate binding. Additionally, in the active state, the three residues ‘Asp-Phe-Gly’ (DFG motif) at the N-terminus of the A-loop bind to an ATP-binding Mg2+ ion, with the Asp side-chain pointing inward to the ATP-binding pocket, while in the inactive state, the Asp side-chain is flipped outward, and the DFG motif adopts the DFG-out conformation (Modi and Dunbrack, 2019). Henceforth, we will refer to the type I inhibitor binding state as the active DFG-in state and designate the type II inhibitor binding state as the classical DFG-out state, as previously proposed in the literature (Gizzio et al., 2022; Thakur et al., 2024; Gizzio and Thakur, 2024) (demonstrated in Figure 2A and B using Abl1 kinase as an example).

Figure 2. DFG-in and DFG-out conformation adopted by DDR1 kinase and their relation with Type I and Type II inhibitors.

(A) The NMR (nuclear magnetic resonance) structures of Abl1 kinase overlay, comparing the activation loop (A-loop) in the active DFG-in state (red, Protein Data Bank [PDB]: 6XR6) with the classical DFG-out state (blue, PDB: 6XRG). Type I inhibitors target the active DFG-in state, where the DFG motif adopting the DFG-in configuration and the A-loop adopting the ‘extended’ configuration, while type II inhibitors target the classical DFG-out state, where the DFG motif adopting the DFG-out configuration and the A-loop is ‘folded’. The distance between CB atoms of residue N98 (gray bead) and residue R162 (red/blue bead) in Abl1 kinase serves as an order parameter here to illustrate the location of the A-loop. The dashed black block emphasizes the different configurations of the DFG motif in these two states. (B) The Dunbrack definition for DFG motif configuration is employed here. The Dunbrack space is delineated by two-order parameters: D1=dist(F158-CZ, M66-CA),D2=dist(F158-CZ, K47-CA). (Cor D) The docking poses with the lowest ligand RMSD for four known kinase inhibitors targeting the Abl1 or DDR1 kinase AlphaFold2 (AF2) structure, generated by Glide XP. Co-crystallized structures are shown as light-cyan cartoons (proteins), green sticks (ligand), and magenta sticks (DFG motif). Docking poses are shown as light-gray cartoons (proteins), gray sticks (ligand), and blue sticks (DFG motif). Comparing with type I inhibitors, AF2 structures of protein kinases fail to dock with type II inhibitors.

Figure 2—figure supplement 1. The distributions of ligand RMSDs for Glide XP docking poses of DDR1 and type I/type II inhibitors (upper/lower panel).

Results from cross-docking against four crystal holo structures, docking against the AF structure, and docking against 15 classical DFG-out structure in reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) ensemble are shown as green, blue, and red, separately.

In this work, we investigated three kinases: (i) Abl1, which is targeted by the first clinically approved small-molecule kinase inhibitor imatinib for cancer therapy; (ii) DDR1, a structurally more flexible kinase which is identified as a promiscuous kinase targeted by chemically diverse inhibitors (Hanson et al., 2019); and (iii) Src kinase, another crucially important member of the tyrosine kinase family. In the following section, we applied AF2RAVE to enrich holo-like structures adopting the classical DFG-out state from the AF2-generated ensembles of DDR1 and Abl1 kinase, and further validated those holo-like structures by docking them with known type II kinase inhibitors. We initiated our investigation by conducting docking experiments involving both type I and type II inhibitors with the AF2 structures of these two kinases. We observed the incapacity of AF2 structures to effectively dock with type II inhibitors targeting the metastable classical DFG-out state. Subsequently, we employed rMSA AF2, an associative memory-like process that generates diverse structure ensembles for both kinases (Porter et al., 2023; Roney and Ovchinnikov, 2022), exploring their relatively improved but still limited potential to contain holo-like structures suitable for type II inhibitor binding. Ultimately, our study showcases AF2RAVE’s effectiveness in enhancing the generation and selection of holo-like structures in metastable states by integrating AF2-based ensemble generation with physics-based methods.

Results and discussion

AF2 structures fail to dock with type II kinase inhibitors

The utility of AF2-generated structures for structure-based drug discovery and virtual screening campaigns has been a subject of controversy and skeptical enthusiasm (Lyu et al., 2024; Holcomb et al., 2023; Scardino et al., 2023; Díaz-Rovira et al., 2023). For proteins lacking crystal and cryo-EM structures but with homologous structures available in PDB, such as CDK20 kinase, AF2 demonstrates effectiveness as a homology modeling method for generating initial structures suitable for subsequent virtual screening (Ren et al., 2023). Other AF2-based homology modeling method can bias AF2-generated structures toward user-selected template structures with specific druggable conformations (Sala et al., 2023). However, a significant drop in the hits enrichment factor during virtual screening has been reported when employing AF2 structures as rigid receptors for docking, compared to using holo PDB structures. This occurs even in cases where the binding pockets of AF2 structures differ only slightly at two to three residues, from those of holo PDBs (Scardino et al., 2023; Díaz-Rovira et al., 2023). Given that ligands can induce slight relocation and side-chain rotation of pocket residues upon binding, it is important to note that AF2 does not account for this induced fit effect, as it does not encode co-factors like ligands. Therefore, it appears necessary to perform ligand induced fit modeling or relaxation on AF2 structures before engaging in any further structure-based drug design. Molecular dynamics (MD) simulations biased toward adjacent holo-template structure have proved effective in refining apo structures and improving their early enrichment performance (Guterres et al., 2021). It has also been demonstrated that AF2 structures can achieve comparable accuracy to crystal holo structures in free energy perturbation (FEP) calculations, by superposing AF2 structures with crystal structures, grafting the co-crystallized ligands onto the AF2 structure, and optimizing the AF2 structure/ligand complex to account for subtle induced fit effects (Beuming et al., 2022). AF2 structures, decorated with ligands from template ligand grafting method or known-hits docking method, and further refined by Schrödinger IFD-MD protocol, exhibit promising performance in early enrichment (Zhang et al., 2023) and the prediction of novel ligand/protein complex structures (Coskun et al., 2024). Therefore, if large backbone motion is not required, it appears feasible to refine the apo pockets in AF2 structures into holo-like pockets for structure-based drug design. However, when AF2 structures exhibit significant steric clashes with holo ligands, especially in cases ligands targeting metastable states, direct application of out-of-the-box AF2 structures in docking methods for virtual screening and early enrichment may pose more challenges. The AF2-predicted kinase structures predominantly exhibit the DFG-in state, with over 95% of human kinases predicted in this conformation (Modi and Dunbrack, 2022; Al-Masri et al., 2023). Significant A-loop motion and backbone flipping of the DFG motif are necessary to transition from the AF2-predicted DFG-in state to holo-like states, for type II ligands targeting the classical DFG-out state. As a result, AF2 structures of Abl1/DDR1 kinases exhibit superior performance when docking with type I inhibitors (achieving a minimum ligand RMSD of 2.14 Å) compared to docking with type II inhibitors using Glide XP (Figure 2C and D). Even with Glide’s IFD (Figure 3—figure supplement 4) and the highly side-chain clashes forgiving docking method DiffDock (Figure 3—figure supplement 6), AF2 structures struggle to dock with these metastable state-targeting ligands (type II inhibitors), with ligand RMSDs above 8 Å across all docking poses from all three docking methods.

Holo-like metastable structures may be present among decoys generated from rMSA AF2

AF2-based methods can achieve structural diversity by introducing dropouts in MSA inputs stochastically (rMSA AF2) or in a clustering manner (AF2-cluster). Additionally, models modified from the AF2 framework, such as the flow-match generative model AlphaFlow (Jing et al., 2024), have been developed to explore the diversity of conformational space. Similar protocols to rMSA AF2 have demonstrated the potential to address the induced fit effect by generating diverse structures at the binding pocket, ranging from closed apo pockets to opened holo-like pockets (Meller et al., 2023). Other investigations have also indicated that larger backbone motions, such as DFG motif backbone flipping and A-loop movement in at least some protein kinases, can be captured by the rMSA AF2 ensemble, although the distributions of conformations deviate significantly from the correct Boltzmann distributions (Vani et al., 2024; Monteiro da Silva et al., 2024).

In this section, we utilized rMSA AF2 to generate 1280 diverse structures for Abl1 kinase or DDR1 kinase: 640 for MSAs of depth 16 and 32. See Supplementary Material for information regarding Src kinase. Following a filtering step based on RMSD from corresponding AF2 structures, 1198 and 1147 structures remain for Abl1 and DDR1, respectively (Figure 3—figure supplement 2). As shown in Figure 3A and B, for Abl1 kinase, only 4 out of 1198 structures have a folded A-loop, when using a distance cutoff of 15 Å between CB atoms of N98 and R162 in Abl1. However, for DDR1 kinase, 124 out of 1147 rMSA AF2 structures exhibit a folded A-loop, when using a salt bridge distance cutoff of 10 Å between the aligned residue pairs in DDR1 (E110 and R191). We then clustered the A-loop folded structures in the Dunbrack space. For Abl1, only two clusters were identified: one adopts the DFG-in state, and the other adopts the DFG-inter state, referring to an intermediate conformation during the backbone flipping from DFG-in to DFG-out, according to the Dunbrack definition. For DDR1, structures were divided into five clusters, with one cluster of size 15 being the closest to the classical DFG-out state.

Figure 3. Reduced MSA AF2 is capable of generating crystal-like DFG-out state for DDR1 kinase.

(A) Upper panel: distribution of A-loop location for the reduced multiple sequence alignment (MSA) AlphaFold2 (AF2) structures of Abl1 kinase. 4 out of 1198 structures are A-loop folded. Lower panel: k-means clustering of four A-loop folded Abl1 structures in Dunbrack space, with the number of clusters (n cluster) set to 2. The structures in the cluster closest to the classical DFG-out state (red circles) fail to dock with type II inhibitors using Induced Fit docking (IFD). (B) Upper panel: distribution of A-loop location for the reduced MSA AF2 structures of DDR1 kinase. 124 out of 1147 structures are A-loop folded. Lower panel: k-means clustering of 124 A-loop folded DDR1 structures in Dunbrack space, with n cluster set to 5. Among the 15 structures in the cluster closest to the classical DFG-out state (red circles), one structure (‘holo-model’, highlighted by a red circle filled with green) demonstrates successful docking with type II inhibitors, showcasing a ligand RMSD<2 Å, utilizing IFD or an extended-sampling version of IFD, IFD-trim. (C) The docking poses with the lowest ligand RMSD for two type II kinase inhibitors targeting the DDR1 kinase ‘holo-model’ structure, generated by IFD or IFD-trim. The color code is the same as Figure 2.

Figure 3—figure supplement 1. The plot illustrates the number of gaps in the multiple sequence alignment (MSA) generated by mmseq2 (using ColabFold; Mirdita et al., 2022) for different kinases.

The non-gap count describes the coverage of each position in the MSA. The presence of residue positions with gap counts higher than 40% of the total sequence in DDR1 implies that it has fewer conserved regions than abl1 kinase and src kinase. This characteristic of DDR1 MSA enables the reduced MSA (rMSA) AlphaFold2 (AF2) protocol to generate multiple conformations for DDR1, including the classical DFG-out conformation, by initializing it at various states. However, the highly conserved nature of abl1 and src makes it challenging for the rMSA AF2 to initialize at a state that can lead to a classical DFG-out conformation. Therefore, we used the AlphaFold template protocol to overcome this initialization issue with rMSA AF2.

Figure 3—figure supplement 2. The AlphaFold2 (AF2) pLDDT rank is plotted against the CA RMSDs from the AF2 structure (the one with the highest pLDDT) for each structure in the reduced multiple sequence alignment (rMSA) AF2 ensemble for Abl1, DDR1, or Src kinase.

An RMSD cutoff of 7 Å (dashed black line) is applied to filter out unphysical structures with large RMSD from the native structure. Each rMSA AF2 ensemble consists of 1280 structures, 640 for MSAs of depth 8:16 (red) and 16:32 (blue), separately.

Figure 3—figure supplement 3. The AlphaFold2 (AF2) pLDDT rank is plotted against the CA RMSDs from the AF2 structure for each structure in the AF2-cluster ensemble for Abl1, DDR1, or SrcK.

An RMSD cutoff of 10 Å (dashed black line) is applied to filter out unphysical structures with large RMSD from the native structure. After the RMSD filter, 197 out of 362 structures remain for Abl1, 134 out of 251 structures remain for DDR1, and 93 out of 355 structures remain for SrcK.

Figure 3—figure supplement 4. Ligand RMSDs are plotted against the docking scores for the Induced Fit docking (IFD) poses of type II inhibitors (ponatinib and imatinib) against AlphaFold2 (AF2) structure (blue) or classical DFG-out structures in reduced multiple sequence alignment (rMSA) AF2 ensembles (red).

(A) IFD results for Abl1. (B) IFD results for DDR1. The pose with the lowest ligand RMSD from each input structure is marked by hexagon.

Figure 3—figure supplement 5. Docking studies on holo-model from rMSA AF2 for DDR1 kinase.

(A) Comparison of the DFG motif for DDR1 in its co-crystalized structure with imatinib (Protein Data Bank [PDB] 4BKJ), its ‘holo-model’ structure, and its AlphaFold2 (AF2) structure. (B and C) In the ‘holo-model’ structure, the Phe residue in the DFG motif requires rotation to prevent steric clashes with imatinib. Proteins from crystal structure are shown as cyan cartoon, while all the other proteins are shown as gray cartoon. (D) Ligand RMSDs are plotted against the docking scores for the IFD-trim docking poses of type II inhibitors (ponatinib and imatinib) against the ‘holo-model’ structure in DDR1 reduced multiple sequence alignment (rMSA) AF2 ensembles. The pose with the lowest ligand RMSD is marked by hexagon.

Figure 3—figure supplement 6. Ligand RMSDs are plotted against the DiffDock confidence scores for the DiffDock poses of type II inhibitors (ponatinib and imatinib) against DDR1 AlphaFold2 (AF2) structure (blue) or the classical DFG-out structures in DDR1 reduced multiple sequence alignment (rMSA) AF2 ensemble (red).

The pose with the lowest ligand RMSD from each input structure is marked by hexagon.

We subsequently docked type II inhibitors, ponatinib and imatinib, to the most ‘classical DFG-out’-like cluster from rMSA AF2 ensembles (indicated as red circles in Figure 3A and B), using IFD. Despite the A-loop being folded, the DFG motif in structures from the Abl1 cluster is not strictly DFG-out. As expected, type II ligands fail to dock with those structures, with all IFD poses exhibiting ligand RMSDs above 9 Å (Figure 3—figure supplement 4A, Figure 5B). While AF2-based methods can indeed generate decoy structures that deviate from the native state, it remains uncertain whether these decoys correspond to metastable basins. Additionally, it’s unclear whether these decoys include structures that can represent the specific metastable states required for the intended types of drug design.

In contrast to Abl1, rMSA AF2 ensemble for DDR1 contains holo-like structure for type II inhibitors. One structure (which we label the ‘holo-model’, shown as a red circle filled with green in Figure 3B), out of the 15 in the DDR1 classical DFG-out cluster docked ponatinib with a remarkably low RMSD of 0.89 Å using IFD (Figure 3C). However, IFD poses from all the other structures exhibit large ligand RMSDs above 6 Å (Figure 3—figure supplement 4B). Hence, an enrichment process to select holo-like structures from decoys becomes essential to ensure a practical number of pocket structures for ensemble docking and virtual screening.

We have also docked the ‘holo-model’ structure with another type II ligand, imatinib. In this case, the steric clashes are more significant, rendering the refinement for the induced fit effect more challenging compared to the ponatinib scenario. We thus introduced an extra trimming step (see Methods section for details) for the DFG-Phe residue which exhibit significant clashes with the holo-imatinib in the ‘holo-model’ structure (Figure 3—figure supplement 5B). The IFD-trim protocol successfully achieves a minimum ligand RMSD of 1.04 Å for docking poses of the ‘holo-model’ structure with imatinib (Figure 3C, Figure 3—figure supplement 5C and D).

AF2RAVE on DDR1 enriches holo-like classical DFG-out structures in rMSA AF2 decoys

To assess the utility of physics-based methods in selecting holo-like structures from AF2-generated ensembles, we utilized a physics-based protocol, AF2RAVE (Vani et al., 2023), to explore the energy landscape of DDR1 kinase. We employed the identical set of collective variables (CVs) as in our previous study (Vani et al., 2024) to perform regular space clustering for the DDR1 rMSA AF2 ensemble. Two structures were selected from the clustering centers for each combination of DFG type (in, inter, or out) and A-loop position (folded or extended), resulting in a total of 12 initial structures (Figure 4—figure supplement 1). Subsequently, 50 ns unbiased MD simulations were conducted starting from each initial structure. The CVs extracted from all MD trajectories were input into the SPIB model with a linear encoder, to learn the reaction coordinates indicative of the slow motions linked to transitions between metastable states. The learnt SPIB reaction coordinates possess clear physical interpretations, with the first coordinate signifying the A-loop position and the second indicating the DFG type (Figure 4A and B). This alignment with physical features is expected as we deliberately selected diverse and representative initial structures for unbiased MD simulations. Additionally, it’s noteworthy that these reaction coordinates are transferable across various kinases and can effectively discern different metastable states.

Figure 4. Ranking the DDR1 structures using AF2RAVE protocol on the learnt latent space.

(A or B) The unbiased molecular dynamics (MD) trajectories of DDR1 are projected onto the learnt state predictive information bottleneck model (SPIB) latent space. In plot (A), the colors of sample points represent the A-loop location, while in plot (B), they depict the Dunbrack DFG state. The first SPIB coordinate, σ1, correlates with the A-loop location, and the second SPIB coordinate, σ2, correlates with configuration of the DFG motif. (C) The reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) structures of DDR1 are projected onto the latent space. Sample points are color-coded based on the A-loop location. Light green stars highlight the 15 classical DFG-out structures selected based on prior information in Figure 3. (D) Free energy profile in the A-loop folded region of the latent space, calculated from umbrella sampling simulations. The 15 classical DFG-out structures from rMSA AF2 are shown as red cross and circles (structures with free energy less than 1 kJ/mol). The ‘holo-model’ structure is emphasized using a red circle filled with red. The embedding table shows the lowest ligand RMSD in IFD poses of the rMSA AF2 structure with ponatinib. The ‘holo-model’ is among the two structures selected by AF2RAVE (potential of mean force [PMF]<1 kJ/mol).

Figure 4—figure supplement 1. Reduced MSA AF2 generated DDR1 kinase structures on the Dunbrack space.

(A) The reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) ensemble for DDR1 is projected in the Dunbrack space. Sample points are color-coded based on the CA RMSD from the AF2 structure with the highest pLDDT. Regular space cluster centers are marked by blue hexagons. For each DFG type (in, inter, or out), top two cluster centers with the lowest CA RMSD are selected as AF2RAVE initial structures. (B) To take account of the underrepresented A-loop folded configurations, an extra regular space clustering is conducted only for the A-loop folded structures in the rMSA AF2 ensemble. The color code, notation, and the way to select initial structures are the same as plot A. Combining AF2RAVE initial structures from both plots (A and B), there are 12 initial structures in total.

Figure 4—figure supplement 2. Umbrella sampling for DDR1 kinase.

(A) Distributions from different umbrella sampling windows in the latent space. (B) The distribution overlap graph for all the umbrella sampling windows. The mean value of each distribution is shown as blue dots. Each distribution’s 2D histogram is flattened into 1D vectors, and the cosine similarity between two distributions is then indicated by the width and color of the edge connecting the respective dots. Windows from the A-loop folded region are not overlapped well with the windows from the A-loop extended region, while windows inside the A-loop folded region (the left part of the graph) are well connected and are used for the local potential of mean force (PMF) calculation in Figure 4D.

Figure 4—figure supplement 3. DDR1 potential of mean force (PMF) calculated with all the umbrella sampling windows.

Hanson et al., 2019, found the A-loop folded DFG-out state to be more stable than the A-loop folded DFG-in/inter state for DDR1; Vani et al., 2024, reported that the A-loop extended DFG-out state is more stable than the A-loop extended DFG-in/inter state for DDR1. Although our umbrella sampling setup is not sufficient to sample the A-loop movement, the observed relative stability corresponds with the findings of Hanson et al. and Vani et al.

Figure 4—figure supplement 4. Potential of mean force (PMF) values and Boltzmann ranks of candidate structures fluctuate with the selection of the umbrella sampling windows and the simulation length of umbrella sampling trajectories, demonstrated with the DDR1 system.

Figure 4—figure supplement 5. Free energy profile for DDR1 in the latent space, calculated from unbiased molecular dynamics (MD) simulations.

The 15 DDR1 classical DFG-out structures in reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) are shown as red cross and circles (top 5 structures ranked by free energy values). The ‘holo-model’ structure is emphasized using a red circle filled with red.

Figure 4—figure supplement 6. Unbiased simulation coverage study on DDR1 and Abl1 kinase.

(A) The distribution overlap graph for all the unbiased molecular dynamics (MD) trajectories starting from 15 classical DFG-out structures in DDR1 reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) ensemble. (B) The distribution overlap graph for all the unbiased MD trajectories starting from 30 Abl1 tAF2 structures in classical DFG-out state. The color code is the same as Figure 4—figure supplement 4.

Interestingly, the 15 classical DFG-out structures within the DDR1 rMSA AF2 ensemble are situated in a region of the latent space that were not thoroughly explored by the 12 unbiased MD trajectories. To address this gap, we employed enhanced sampling to sample along the SPIB-approximated reaction coordinates and compute the free energy profile inside the classical DFG-out basin. Considering both the flipping of the DFG motif and the overall motion of the large flexible A-loop, it might be impractical to sample direct back-and-forth transitions between various states using metadynamics. Therefore, we opted for umbrella sampling for its simplicity. The reliability of umbrella sampling hinges on two issues, first whether the latent space adequately represents the conformational space and second, whether there is sufficient overlap between different windows for efficient reweighting. Addressing the first challenge remains an ongoing endeavor in the dimensionality reduction research field, and we anticipate that our SPIB latent space is sufficient enough for our current purpose. The second challenge can be managed through careful setup of umbrella sampling windows and bias strength. Given our current setup, sampling the extensive motion of A-loop relocation remains challenging, showing insufficient overlap between the A-loop folded and extended regions (Figure 4—figure supplement 2). Consequently, the quantitative reliability of the absolute ΔG values between different states is limited, allowing us only to qualitatively assess the relative thermodynamic stability of the DFG-in versus the DFG-out basins. Nevertheless, the qualitative relative stability observed from umbrella sampling aligns with previous studies (Vani et al., 2024; Hanson et al., 2019; Figure 4—figure supplement 3). Furthermore, the local potential of mean force (PMF) surrounding each basin should provide quantitatively reliable insights, given their thorough sampling through umbrella sampling. Reassuringly, when we ranked the classical DFG-out structures within the DDR1 rMSA AF2 ensemble, using the local PMF values in the latent space, the ‘holo-model’ structure emerged among the top 2 structures with free energy relative to the minimum smaller than 1 kJ/mol, as illustrated in Figure 4D.

Additionally, we used DiffDock to conduct docking experiments with type II ligands on the 15 classical DFG-out structures within the DDR1 rMSA AF2 ensemble. Despite poses from most structures surprisingly demonstrating very low ligand RMSD, typically below 2 Å (Figure 3—figure supplement 6), they exhibited significant steric clashes and low DiffDock confidence score due to DiffDock’s disregard for side-chain configurations and clashes. Notably, upon ranking structures based on their AF2RAVE PMF and examining the corresponding poses with the lowest ligand RMSD, we observed that only the poses derived from the top 2 structures selected by AF2RAVE had DiffDock confidence scores higher than –1.5, surpassing the default threshold of the DiffDock model (Supplementary file 1). This implies the general advantage of structures selected by physics-based methods across various docking methods.

In summary, Boltzmann ranking from AF2RAVE effectively distinguishes holo-like structures from other decoys. Beginning with the DDR1 rMSA AF2 ensemble, AF2RAVE substantially enhances the likelihood of identifying the holo-like structure from 1 out of 15 to at least 1 out of 2 when using a PMF cutoff of 1 kJ/mol. We noted that although the PMF values and Boltzmann ranks may fluctuate with the setup of umbrella sampling, the enrichment effect of holo-like structures remains (Figure 4—figure supplement 4). Besides, we applied an alternative protocol to compute the PMF profile within the DDR1 classical DFG-out basin, by running unbiased MD simulations starting from the 15 classical DFG-out decoys in the DDR1 rMSA AF2 ensemble. However, this unbiased protocol poses a risk of failing to sample rare events if barriers exit within the region of interest. In this case, the mini-barriers inside the classical DFG-out basin are low enough to enable efficient sampling using 50 ns unbiased MD simulations. Consequently, the Boltzmann ranks derived from unbiased MD simulations also successfully enriched the single holo-like structure in the DDR1 rMSA AF2 ensemble to the top 5 of the structure set (Figure 4—figure supplement 5).

Transferable learning of holo-like structure for Abl1 and Src kinases with AF2RAVE from DDR1 templates

As mentioned earlier, our current Abl1 rMSA AF2 ensemble lacks any decoy structure in the classical DFG-out state. This is further illustrated in Figure 5A, where the Abl1 rMSA AF2 ensemble is projected onto the same latent space learnt during the DDR1 AF2RAVE protocol. Hence, it’s necessary to prepare the Abl1 decoy set adopting the classical DFG-out state before we can rank them using Boltzmann weights derived from physics-based methods and conduct further docking for selected holo-like structures.

Figure 5. Transferrable learning of holo-like structure for Abl1 kinase.

(A) The reduced multiple sequence alignment (MSA) AlphaFold2 (AF2) structures of Abl1 are projected onto the latent space. Sample points are color-coded based on the A-loop location. Light green stars highlight the 30 AF2-template Abl1 structures modeled from the 15 DDR1 classical DFG-out structures. (B) Upper panel: the distribution of ligand RMSD for the Induced Fit docking (IFD) poses of Abl1 structures and two type II ligands. Lower panel: IFD poses with the lowest ligand RMSD for Abl1 AF2-template structures and two type II ligands. The color code is the same as Figure 2. (C) Free energy profile in the latent space, calculated from unbiased molecular dynamics (MD) simulations. The 30 Abl1 classical DFG-out structures from AF2-template are shown as red cross and circles (structures with free energy less than 1 kJ/mol). The ‘holo-model’ structures are emphasized using red circles filled with red. (D) The table shows the lowest ligand RMSD in IFD poses of the AF2-template structures with two type II inhibitors. All the four ‘holo-models’ are among the eight structures selected by AF2RAVE (potential of mean force [PMF]<1 kJ/mol).

Figure 5—figure supplement 1. Ligand RMSDs are plotted against the docking scores for the Induced Fit docking (IFD) poses of type II inhibitors (ponatinib and imatinib) against Abl1 tAF2 structures.

The pose with the lowest ligand RMSD from each input structure is marked by hexagon.

Figure 5—figure supplement 2. The projection of A-loop folded structures from the reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) ensemble or the AF2-cluster ensemble (cAF2) on the AF2RAVE potential of mean force (PMF) for Abl1 or DDR1.

Figure 5—figure supplement 3. Structural visualization of tAF2 generated classical DFG-out Src kinase.

(A) The AlphaFold2 (AF2)-template structure for Src kinase is superimposed with its template structure (classical DFG-out in DDR1 reduced multiple sequence alignment [rMSA] AF2 ensemble, ‘holo-model’). The tAF2 structure of Src is shown as light-orange cartoon (protein) and yellow sticks (DFG motif), while DDR1 template is shown as light-gray cartoon (protein) and blue sticks (DFG motif). (B) The AF2-template structure for Src kinase is again superimposed with Src/imatinib co-crystallized structure (Protein Data Bank [PDB] 2OIQ). Crystal structure is shown as light-cyan cartoon (protein), green sticks (ligand), and magenta sticks (DFG motif).

Figure 5—figure supplement 4. Comparison between performance of rMSA AF2 and AF2-cluster on Src kinase.

The projection of (A) the reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) ensemble or (B) the AF2-cluster ensemble on the AF2RAVE latent space for SrcK. The classical DFG-out SrcK structure generated from AF2-template in Figure 5—figure supplement 5 is shown as the green star. The color code shows the A-loop location.

Figure 5—figure supplement 5. Ligand RMSDs are plotted against the docking scores for the Induced Fit docking (IFD)/IFD-trim docking poses of type II inhibitors (ponatinib and imatinib) against the SrcK tAF2 structure.

The pose with the lowest ligand RMSD from each input structure is marked by hexagon.

Figure 5—figure supplement 6. Accounting for broken αC helix during umbrella sampling of DDR1 and Abl1 kinase.

(A) One representative frame with αC helix broken in Abl1 umbrella sampling trajectories. The backbone of the αC helix is shown with cyan sticks, while the DFG motif is shown as orange sticks. (B or C) The distribution of the ratios of frames with αC helix broken in each umbrella sampling window for Abl1 or DDR1.

Figure 5—figure supplement 7. Abl1 potential of mean force (PMF) calculated from umbrella sampling after discarding windows with αC helix broken.

The four holo-like structures (‘holo-models’) are enriched to the top six based on PMF values.

There are several approaches to generate the Abl1 decoy sets. One can conduct enhanced sampling on the latent space starting from Abl1 rMSA AF2 structures to reach the classical DFG-out basin. Subsequently, MD structures from this basin can serve as templates for asking AF2 to generate crystal-like structures in classical DFG-out state. However, for simplicity, we opted to use the 15 classical DFG-out structures from the DDR1 rMSA AF2 ensemble directly as templates and employed an AF2-based homology modeling protocol, referred to as AF2-template (tAF2 for short, detailed protocol can be found in the Methods section), to generate a decoy set comprising 30 Abl1 structures, as illustrated by the green stars in Figure 5A.

Compared to the AF2 and rMSA AF2 structures, the performance of IFD on type II inhibitors shows a significant improvement when using the 30 tAF2 decoy structures of Abl1. The lowest ligand RMSD achieved is 2.74 Å for imatinib and 0.78 Å for ponatinib (Figure 5B). However, only four structures (labeled as ‘holo-model’ structures hereafter) out of the 30 decoys are capable of docking with type II inhibitors with ligand RMSD <3 Å, and all other structures produce IFD poses with ligand RMSD>6 Å (Figure 5—figure supplement 1).

We then investigated whether physics-based methods could enrich the 4 ‘holo-model’ structures from the 30 tAF2 decoys to a practical number of holo-like candidates. To explore the local classical DFG-out basin, we conducted 50 ns unbiased MD simulations starting from each tAF2 decoy structure and combined the trajectories to obtain the final Boltzmann distribution. Remarkably, the PMF for the Abl1 classical DFG-out basin calculated from the unbiased protocol (Figure 5C) showed the enrichment of all 4 ‘holo-model’ structures among the top 8 tAF2 structures with PMF value <1 kJ/mol. The PMF values of the top 8 tAF2 structures and the lowest ligand RMSD from the corresponding structures are presented in Figure 5D. We also employed umbrella sampling for the Abl1 kinase by transferring setups and initial structures using the AF2-template from DDR1 umbrella sampling. However, we noted a significant occurrence of αC helix breakage in the Abl1 umbrella sampling trajectories compared to DDR1 (Figure 5—figure supplement 6). After excluding windows with broken αC helix, the Abl1 PMF derived from umbrella sampling ultimately enriched all 4 ‘holo-model’ structures to the top 6 among the 30 tAF2 structures (Figure 5—figure supplement 7). In Supplementary Information, we also provide results for Src kinase which are in the same quality as for Abl1 kinase.

Discussion

Through our retrospective analysis, we have thus demonstrated that the default AF2 models are ineffective for docking ligands targeting metastable protein kinase conformations. While AF2-based methods can be coaxed into generating diverse structures, they still struggle to produce reliable accuracy decoys for metastable conformations since the AF2 ensembles do not follow Boltzmann distribution. This failure is evident in the inability to generate an Abl1 AF2 ensemble containing holo-like structures for type II inhibitors. To further investigate whether this limitation is common among AF2-based methods, including the rMSA AF2 method we employed earlier, we tested another AF2-based approach, AF2-cluster (Wayment-Steele et al., 2024). The AF2-cluster ensemble of Abl1 comprises more A-loop folded structures (21 out of 197) compared to our Abl1 rMSA AF2 ensemble (4 out of 1198). However, similar to rMSA AF2, there are still no decoys in the classical DFG-out state, and all A-loop folded structures are located far from the PMF basin of the classical DFG-out state in the latent space (Figure 5—figure supplement 2C). This indicates that AF2-cluster, like rMSA AF2, also fails to generate Abl1 metastable states effectively. Interestingly, we observed comparable structural diversity in sampling the A-loop folded configurations within the AF2-cluster ensembles for DDR1 and Abl1 (Figure 5—figure supplement 2C and D). While for the rMSA AF2 method, the promiscuous kinase DDR1 ensemble exhibits superior structural diversity compared to the Abl1 kinase. This enhanced diversity leads to the identification of one dockable structure for type II inhibitors among decoys in DDR1 rMSA AF2 ensemble. Through the application of a homology modeling method, AF2-template, we demonstrated that the classical DFG-out decoys in the DDR1 rMSA AF2 ensemble can be transferred to Abl1 kinase. Furthermore, we tested an additional kinase, Src, for which non-native state decoys are reported to be even more challenging to produce using AF2 subsampling methods than Abl1 kinase (Monteiro da Silva et al., 2024). We have also verified that rMSA AF2 and AF2-cluster struggle in producing distinct decoys from native structure of Src kinase (Figure 3—figure supplement 2, Figure 3—figure supplement 3). Besides, neither the rMSA AF2 nor the AF2-cluster ensembles for Src kinase adequately sample the classical DFG-out basin in the latent space (Figure 5—figure supplement 4). Remarkably, AF2-template, as a homology modeling method, can easily produce the classical DFG-out structure of Src kinase, using the top 2 AF2RAVE-picked DDR1 classical DFG-out structures as templates (Figure 5—figure supplement 3). The IFD poses from the tAF2 structure of Src kinase shown a minimal ligand RMSD of 2.82 Å with imatinib (Figure 5—figure supplement 5).

With the rapid expansion of chemical space, virtual screening on libraries containing billions of diverse molecules becomes enticing for novel drug discovery (Lyu et al., 2019). Therefore, the enrichment of candidate holo-like structures emerges as a necessary step, offering significant benefits in terms of computational efficiency and feasibility. As summarized in Table 1, unlike the non-dockable AF2 structures for type II inhibitors, the diverse rMSA AF2 ensemble (referred to as rAF2 in Table 1 for brevity) shows potential in generating holo-like structures within a large set of decoys. However, it’s only upon AF2RAVE ranking and selection that the ratio of holo-like structures in selected structure set increases to a plausible value of 50%, facilitating further virtual screening on computational models of protein pockets.

Table 1. Comparing the Induced Fit docking (IFD) performance of various structure generation methods for docking type II kinase inhibitors.

Source-protein (# of structures)	Lowest imatinib ligand RMSD (Å)	Lowest ponatinib ligand RMSD (Å)	Ratio of structs. w. ligand RMSD <3 Å	
AF2-Abl1 (1)	9.22	9.4	0/1	
AF2-DDR1 (1)	9.24	9.33	0/1	
rMSA AF2-Abl1 (2)	10.14	9.11	0/2	
rMSA AF2-DDR1 (15)	1.04*	0.89	1/15	
tAF2-Abl1 (30)	2.74	0.78	4/30	
AF2RAVE-DDR1 (2)	1.04*	0.89	1/2	
AF2RAVE-Abl1 (8)	2.74	0.78	4/8	
* Result from IFD-trim.

Conclusion

AF2 has arguably revolutionized protein structure prediction, but it remains to be constructively demonstrated if it can be reliably used for drug discovery purposes, especially involving non-native protein conformations. In this work we have demonstrated through retrospective studies on kinase inhibitors that a combination of AF2, statistical mechanics-based enhanced sampling, and IFD can be deployed for such calculations. Specifically, we have utilized the AF2RAVE protocol by inputting the sequences of the DDR1 kinases, along with two additional pieces of prior information: a pairwise distance cutoff for evaluating A-loop positions and the Dunbrack definition of DFG type. We then employed Glide IFD to assess our AF2RAVE-generated computational models for type II kinase inhibitor binding pockets. This AF2RAVE-Glide workflow yielded holo-like structure candidates with a 50% successful docking rate for known type II inhibitors. Notably, the holo-like structures in metastable state and the latent space constructed from AF2RAVE of DDR1 are transferable to other kinases. This includes the challenging cases (Monteiro da Silva et al., 2024) of Abl1 and Src kinases, wherein we showed that SPIB and sampling performed for DDR1 allowed generating classical DFG-out structures for both Abl1 and Src kinases. This severely reduces the computational cost for retraining SPIB to learn low-dimensional latent space for different kinases.

This demonstration of AF2RAVE-Glide on kinase inhibitors shows its promising application for discovering drugs targeting general proteins in addition to kinases, such as G-protein-coupled receptors, which are the targets for over one-third of Food and Drug Administration (FDA)-approved drugs (Hauser et al., 2018). For the design of novel drugs targeting general proteins, developing a protocol that does not require prior information about the system is left for future exploration. Besides, in this study we only investigate the classical DFG-out metastable state for kinases. For a comprehensive protocol, all top-ranked metastable states identified by AF2RAVE should be explored in subsequent docking experiments. Integration of algorithms capable of predicting ligand binding sites on protein surfaces, such as the Graph Attention Site Prediction (GrASP) model (Smith et al., 2024), is then essential before utilizing AF2RAVE-selected structures in docking, thus expanding the workflow to AF2RAVE-GrASP-Glide. Additionally, the inclusion of FEP calculations for front-runner ligands to evaluate the actual ligand binding affinity can further enhance this workflow.

The integration of AF2-based and physics-based methods presents a promising approach toward the development of a mature workflow for computer-aided drug design. AF2-based methods are capable of producing ensembles with structural diversity, which aids physics-based methods in better sampling and exploring the energy landscape of proteins. Additionally, AF2-generated structures can serve as crystal-like decoys, free from distortion that may occur in biased simulations. Physics-based methods play a crucial role in accurately assigning Boltzmann weights and ranking decoy structures to guide the enrichment of holo-like structures, essential for virtual screening on large libraries. This collaborative approach leverages the strengths of both methodologies, leading to enhanced efficiency and efficacy in drug discovery efforts.

We conclude this manuscript by making a final comment on the horizons opened by generative AI methods, including those involving diffusion models (Abramson et al., 2024; Wang et al., 2022) and any such future frameworks. These approaches make it possible to easily hypothesize regions of the conformational space underlying arbitrarily complex molecules of life, that then could serve as a starting point to launch more careful investigations. However, these predictions without careful in situ or a posteriori testing through advanced simulations or experiments are only predictions and could just be hallucinations. Having the capability to quickly generate numerous - thousands or more - such structural hypotheses is what made AF2 so crucial to this current work. We believe that a deep integration of the hypothesis creation possibilities of generative AI with careful MD (Herron et al., 2023; Zheng et al., 2024) experiments or other forms of rapid testing is one way these methods will truly facilitate new and reliable scientific discoveries.

Methods

Summary of systems tested and tools used

In this work we tested AF2RAVE-Glide protocol on active and inactive conformations of DDR1, Abl1, and Src kinases against type I inhibitors VX-680 and dasatinib, and type II inhibitors imatinib and ponatinib. The top-ranked structures obtained from AF2RAVE-Glide were compared against publicly available crystal structures deposited in the Protein Data Bank (PDB), with PDB codes provided in the main text.

To generate diverse structural ensembles for DDR1, Abl1, and Src kinases, we primarily used the rMSA AF2 method. For comparison, we also produced AF2-cluster ensembles for these three kinases.

For MD simulations, all the systems in this paper were parameterized with the Amber99SB*-ILDN force field (Case et al., 2005; Lindorff-Larsen et al., 2010) with the TIP3P water model (Jorgensen et al., 1983) and neutralized with Na+ ions or Cl- ions. The simulations are performed at 300 K with the LangevinMiddleIntegrator (Zhang et al., 2019) in OpenMM (Eastman et al., 2017) with the step size of 2 fs. Particle mesh Ewald (Darden et al., 1993) is used for calculating electrostatics and the lengths of bonds to hydrogen atoms are constrained using LINCS (Hess et al., 1997) throughout all simulations. Before performing MD simulations for analysis, energy minimization is conducted for all initial structures, followed by equilibrium runs under NVT and NPT for 500 ps and 1 ns respectively.

To account for the induced fit effect, IFD from the Schrödinger suite is the primary docking method used in this work. For comparison, we also tested the Glide XP docking from the Schrödinger suite and DiffDock to dock the two known type II inhibitors (imatinib and ponatinib) against DDR1 AF2 structure and the 15 classical DFG-out structures in the DDR1 rMSA ensemble.

Code and data for this paper are available in https://github.com/tiwarylab/AF2RAVE_Glide-kinase.

Generation of rMSA AF2 ensemble, AF2-cluster ensemble, and AF2-template structures

In this work, ColabFold (Mirdita et al., 2022) was employed to generate all the AF2-based ensembles and structures. The MSAs were produced using mmseq2. For rMSA AF2 ensembles, the MSA depth was reduced to either 16 or 32 for each kinase. For each depth, 128 random seeds were utilized, and each seed produced 5 structural models via ColabFold, resulting in a total of 1280 rMSA AF2 structures per kinase. The structure with the highest pLDDT score among all 1280 was identified as the AF2 structure (native structure). Subsequently, any unphysical structures with an RMSD greater than 7 Å from the AF2 structure were discarded, culminating in the final rMSA AF2 ensembles.

AF2-cluster for DDR1, Abl1, and SrcK were run with default setups as provided in the ColabFold notebook in the original paper of AF2-cluster (Wayment-Steele et al., 2024).

The AF2-based homology modeling protocol, AF2-template (tAF2) method, is implemented using ColabFold. For a given query sequence, tAF2 structures are generated by ColabFold upon uploading desired template structure and deactivating the Evoformer module. For each template, five AF2-template models are generated, and the last three structures exhibiting lower pLDDT will be discarded. The AF2-template (tAF2) method is employed here to transfer structure sets between homologous systems. To generate tAF2 structures for Abl1 in the classical DFG-out state, each of the 15 classical DFG-out structures from the DDR1 rMSA AF2 ensemble is used as a template, resulting in 30 tAF2 Abl1 structures in total. For Src kinase, a single representative tAF2 structure is generated using the ‘holo-model’ DDR1 structure as the template.

AF2RAVE protocol

Regular space clustering on the rMSA AF2 ensemble

We used the same 14 CVs for regular space clustering as in the previous AF2RAVE work on kinases (Vani et al., 2024). These CVs are pairwise distances, selected to describe the kinase conformations around the ATP-binding pocket and the A-loop.

Unbiased MD and SPIB

50 ns unbiased MD simulation was run for each AF2RAVE initial structure of DDR1 from Figure 4—figure supplement 1. The standard deviations of the 14 CVs are calculated after concatenating all the 12 unbiased MD trajectories. 8 CVs with standard deviations larger than 0.25 of the maximum standard deviation remain as the input features of the SPIB model. We conducted a parameter screening of the SPIB time lag, ranging from 1 ns to 40 ns with intervals of 1 ns. Eventually, we selected a time lag of 16 ns based on the performance of SPIB coordinates in representing physical features, including DFG type and A-loop position.

PMF calculations from umbrella sampling

2D umbrella sampling is conducted along the two learnt SPIB coordinates, employing 11×11 windows. The bias potential equilibrium points are uniformly distributed in the SPIB latent space, with σ1 ranging from –0.8 to –0.1 and σ2 ranging from 0 to 0.8. The strength of the bias potential is set to 1000kJ/mol/nm2. Each window originates from the structure closest to the window’s equilibrium point in Euclidean distance within the latent space among the 12 AF2RAVE initial structures and lasts for 100 ns (with the first 10 ns discarded in PMF calculation). For Abl1 kinase, we noticed a substantial proportion of the αC helix breaking in the umbrella sampling trajectories (Figure 5—figure supplement 6). This finding aligns with earlier enhanced sampling investigations on DDR1 using metadynamics (Vani et al., 2024), where the authors imposed restraints to prevent αC helix breakage. In our study, we opted to exclude all umbrella sampling windows where the ratio of broken αC helix exceeded 20% for the Abl1 PMF calculation (Figure 5—figure supplement 7). The WHAM algorithm is applied to bin and reweight the biased trajectories and compute the final PMF.

PMF calculations from unbiased MD

50 ns unbiased MD simulation was run starting from each structure in the 15 classical DFG-out structures in DDR1 rMSA AF2 ensemble. Upon discarding first 10 ns, all the unbiased trajectories are simply concatenated to calculate the Boltzmann distribution and PMF for each bin around the classical DFG-out basin in the latent space. For Abl1, unbiased simulations start from 30 AF2-template structures to calculate the PMF around the classical DFG-out basin.

Boltzmann ranks assignment for structures in AF2-based ensembles

After calculating the PMF value for each bin in the latent space, we projected AF2-generated structures into latent space. We then directly assign the PMF values of the corresponding bins to these AF2-generated structures.

We must acknowledge the limitations of the way we assigned PMF values to AF-generated candidate holo structures. First of all, the free energy profiles are derived from MD simulations, and the PMF values directly correspond to the MD structures. Here, we assumed that the latent space adequately represents the conformational changes of protein pocket within specific metastable states. Additionally, while the enrichment of holo structures in the top Boltzmann-ranked structures persists, the absolute PMF values and Boltzmann ranks of proper holo structures may fluctuate with the umbrella sampling setups, as depicted in Figure 4—figure supplement 4. Theoretically, the number of umbrella sampling windows and the simulation length should be sufficiently large for PMF convergence. However, there is always a trade-off between PMF accuracy and computational costs, so we opted to stick with the current setups.

Docking details

All the input structure for our docking experiments were first relaxed in solution with an MD energy minimization step. This work investigates two type I inhibitors (VX-680 and dasatinib) and two type II inhibitors (imatinib and ponatinib) by docking. For Glide XP docking or IFD, we used Ligprep in Maestro to prepare the ligand inputs from SMILES files. For DiffDock, the ligand inputs were directly provided as SMILES files.

Glide XP docking

Glide XP docking experiments in this work were run with default setups in the Maestro software. Glide XP docking was performed for all four ligands on the AF2 structures of DDR1 or Abl1, as well as on the 15 classical DFG-out conformations of DDR1 in the rMSA AF2 ensemble.

Induced Fit docking

For IFD, we used the Glide XP for initial docking, followed by Prime relaxation and final Glide XP docking. Parameters remain default in Maestro IFD. IFD was performed only for the type II ligands on the AF2 structures of DDR1 or Abl1, on the classical DFG-out conformations of DDR1 or Abl1 in the rMSA AF2 ensembles, as well as tAF2 structures of Abl1 and SrcK.

Given ponatinib’s backbone features, notably its lengthy and slender carbon-carbon triple bond, it exhibits reduced sensitivity to steric clashes, resulting in successful docking with the DDR1 ‘holo-model’ structure at a ligand RMSD of 0.89 Å. Conversely, the ‘holo-model’ structure struggles to accurately dock with imatinib using IFD (Figure 3—figure supplement 4B). We then employed an extended-sampling IFD approach by initially trimming the DFG-Phe residue from the ‘holo-model’ structure. The trimmed residue is temporarily mutated to alanine during initial docking and later restored in subsequent Prime relaxation and final docking steps. Here, in the ‘holo-model’ structure, we manually chose the DFG-Phe residue which is significantly hindered by holo-imatinib (Figure 3—figure supplement 5B). For generic systems lacking groundtruth information, broader screening of single residue trimming for the protein pocket may be necessary for this extended-sampling IFD method. In essence, it’s a trade-off between the quality of holo-like structure to dock with and the accuracy/complexity of the docking method.

DiffDock performance on DDR1 classical DFG-out in rMSA AF2 ensemble

DiffDock docking experiments in this work were run in the webserver with default setups in the version before March 8, 2024 (https://huggingface.co/spaces/simonduerr/diffdock; Corso et al., 2022). DiffDock was performed only for type II ligands on the AF2 structures of DDR1, as well as on the 15 classical DFG-out conformations of DDR1 in the rMSA AF2 ensemble.

Funding Information

This paper was supported by the following grant:

http://dx.doi.org/10.13039/100000057 National Institute of General Medical Sciences R35GM142719 to Pratyush Tiwary.

Acknowledgements

Research in this publication was supported by the National Institute of General Medical Sciences of the National Institutes of Health under Award Number R35GM142719. The content is solely the responsibility of the authors and does not represent the official views of the National Institutes of Health. We thank UMD HPC’s Zaratan and NSF ACCESS (project CHE180027P) for computational resources. AA was supported by NCI-UMD Partnership for Integrative Cancer Research. PT is an investigator at the University of Maryland-Institute for Health Computing, which is supported by funding from Montgomery County, Maryland and The University of Maryland Strategic Partnership: MPowering the State, a formal collaboration between the University of Maryland, College Park, and the University of Maryland, Baltimore. We thank UMD HPC’s Zaratan and NSF ACCESS (project CHE180027P) for computational resources. We thank Bodhi Vani, Dedi Wang, Zachary Smith, and Anjali Verma for helpful discussions.

Additional information

Competing interests

Author contributions

Additional files

MDAR checklist

Supplementary file 1. Comparison between AF2RAVE ranks and DiffDock confidence scores.

Confidence score for the DiffDock pose aligns with AF2RAVE potential of mean force (PMF) values. The DiffDock confidence score of the pose with the lowest ligand RMSD (marked in red/bold) from each classical DFG-out structure in DDR1 reduced multiple sequence alignment (rMSA) AlphaFold2 (AF2) ensemble is compared with the AF2RAVE PMF value for corresponding structure (marked in red/bold).

Data availability

Code and data for this paper are available in https://github.com/tiwarylab/AF2RAVE_Glide-kinase (copy archived at Gu, 2024).

10.7554/eLife.99702.3.sa0
eLife assessment
Panchenko Anna Reviewing Editor Queen's University Canada

Compelling
Important
This important study demonstrates that combining AlphaFold2 with the author's sampling method AF2-RAVE improves protein-ligand docking for three protein kinases and their inhibitors. The evidence is compelling and the results will be of interest to researchers who work on computer-aided drug design.

10.7554/eLife.99702.3.sa1
Reviewer #1 (Public Review):
Reviewer
The development of effective computational methods for protein-ligand binding remains an outstanding challenge to the field of drug design. This impressive computational study combines a variety of structure prediction (AlphaFold2) and sampling (RAVE) tools to generate holo-like protein structures of three kinases (DDR1, Abl1, and Src kinases) for binding to type I and type II inhibitors. Of central importance to the work is the conformational state of the Asp-Phy-Gly "DFG motif" where the Asp points inward (DFG-in) in the active state and outward (DFG-out) in the inactive state. The kinases bind to type I or type II inhibitors when in the DFG-in or DFG-out states, respectively.

It is noted that while AlphaFold2 can be effective in generating ligand-free apo protein structures, it is ineffective at generating holo structures appropriate for ligand binding. Starting from the native apo structure, structural fluctuations are necessary to access holo-like structures appropriate for ligand-binding. A variety of methods, including reduced multiple sequence alignment (rMSA), AF2-cluster, and AlphaFlow may be used to create decoy structures. However, those methods can be limited in the diversity of structures generated and lack a physics-based analysis of Boltzmann weight critical to their relative evaluation.

To address this need, the authors combine AlphaFold2 with the Reweighted Autoencoded Variational Bayes for Enhanced Sampling (RAVE) method, to explore metastable states and create a Boltzmann ranking. With that variety of structures in hand, grid-based docking methods Glide and Induced-Fit Docking (IFD) were used to generate protein-ligand (kinase-inhibitor) complexes.

The authors demonstrate that using AlphaFold2 alone, there is a failure to generate DFG-out structures needed for binding to type II inhibitors. By applying the AlphaFold2 with rMSA followed by RAVE (using short MD trajectories, SPIB-based collective variable analysis, and enhanced sampling using umbrella sampling), metastable DFG-out structures with Boltzmann weighting are generated enabling protein-ligand binding. Moreover, the authors found that the successful sampling of DFG-out states for one kinase (DDR1) could be used to model similar states for other proteins (Abl1 and Src kinase). The AF2RAVE approach is shown to result in a set of holo-like protein structures with a 50% rate of docking type II inhibitors.

Overall, this is excellent work and a valuable contribution to the field that demonstrates the strengths and weaknesses of state-of-the-art computational methods for protein-ligand binding. The authors also suggest promising directions for future study, noting that potential enhancements in the workflow may result from the use of binding site prediction models and free energy perturbation calculations.

10.7554/eLife.99702.3.sa2
Reviewer #2 (Public Review):
Reviewer
This manuscript explores the utility of AlphaFold2 (AF2) and the author's own AF2-RAVE method for drug discovery. As has been observed elsewhere, the predictive power of docking against AF2 structures is quite limited, particularly for proteins like kinases that have non-trivial conformational dynamics. However, using enhanced sampling methods like RAVE to explore beyond AF2 starting structures leads to a significant improvement.

Comments on revised version:

I'm happy with the changes made.

10.7554/eLife.99702.3.sa3
Reviewer #3 (Public Review):
Reviewer
In this manuscript, the authors aim to enhance AlphaFold2 for protein conformation-selective drug discovery through the integration of AlphaFold2 and physics-based methods, focusing on improving the accuracy of predicting protein structures ensemble and small molecule binding of metastable protein conformations to facilitate targeted drug design.

The major strength of the paper lies in the methodology, which includes the innovative integration of AlphaFold2 with all-atom enhanced sampling molecular dynamics and induced fit docking to produce protein ensembles with structural diversity. Moreover, the generated structures can be used as reliable crystal-like decoys to enrich metastable conformations of holo-like structures. The authors demonstrate the effectiveness of the proposed approach in producing metastable structures of three different protein kinases and perform docking with their type I and II inhibitors. The paper provides strong evidence supporting the potential impact of this technology in drug discovery. However, limitations may exist in the generalizability of the approach across other structures, especially complex structures such as protein-protein or DNA-protein complexes.

The authors largely achieved their aims by demonstrating that the AF2RAVE-Glide workflow can generate holo-like structure candidates with a 50% successful docking rate for known type II inhibitors. This work is likely to have a significant impact on the field by offering a more precise and efficient method for predicting protein structure ensemble, which is essential for designing targeted drugs. The utility of the integrated AF2RAVE-Glide approach may streamline the drug discovery process, potentially leading to the development of more effective and specific medications for various diseases.

Comments on revised version:

The revised manuscript looks great to me. I have no further comments.

10.7554/eLife.99702.3.sa4
Author response
Gu Xinyu Author University of Maryland, College Park College Park United States

Aranganathan Akashnathan Author University of Maryland, College Park College Park United States

Tiwary Pratyush Author University of Maryland, College Park College Park United States

The following is the authors’ response to the original reviews.

Public Reviews:

Reviewer #1 (Public Review):

The development of effective computational methods for protein-ligand binding remains an outstanding challenge to the field of drug design. This impressive computational study combines a variety of structure prediction (AlphaFold2) and sampling (RAVE) tools to generate holo-like protein structures of three kinases (DDR1, Abl1, and Src kinases) for binding to type I and type II inhibitors. Of central importance to the work is the conformational state of the Asp-Phy-Gly "DFG motif" where the Asp points inward (DFG-in) in the active state and outward (DFG-out) in the inactive state. The kinases bind to type I or type II inhibitors when in the DFG-in or DFG-out states, respectively.

It is noted that while AlphaFold2 can be effective in generating ligand-free apo protein structures, it is ineffective at generating holo-structures appropriate for ligand binding. Starting from the native apo structure, structural fluctuations are necessary to access holo-like structures appropriate for ligand binding. A variety of methods, including reduced multiple sequence alignment (rMSA), AF2-cluster, and AlphaFlow may be used to create decoy structures. However, those methods can be limited in the diversity of structures generated and lack a physics-based analysis of Boltzmann weight critical to their relative evaluation.

To address this need, the authors combine AlphaFold2 with the Reweighted Autoencoded Variational Bayes for Enhanced Sampling (RAVE) method, to explore metastable states and create a Boltzmann ranking. With that variety of structures in hand, grid-based docking methods Glide and Induced-Fit Docking (IFD) were used to generate protein-ligand (kinase-inhibitor) complexes.

The authors demonstrate that using AlphaFold2 alone, there is a failure to generate DFG-out structures needed for binding to type II inhibitors. By applying the AlphaFold2 with rMSA followed by RAVE (using short MD trajectories, SPIB-based collective variable analysis, and enhanced sampling using umbrella sampling), metastable DFG-out structures with Boltzmann weighting are generated enabling protein-ligand binding. Moreover, the authors found that the successful sampling of DFG-out states for one kinase (DDR1) could be used to model similar states for other proteins (Abl1 and Src kinase). The AF2RAVE approach is shown to result in a set of holo-like protein structures with a 50% rate of docking type II inhibitors.

Overall, this is excellent work and a valuable contribution to the field that demonstrates the strengths and weaknesses of state-of-the-art computational methods for protein-ligand binding. The authors also suggest promising directions for future study, noting that potential enhancements in the workflow may result from the use of binding site prediction models and free energy perturbation calculations.

Reviewer #2 (Public Review):

Summary:

This manuscript explores the utility of AlphaFold2 (AF2) and the author's own AF2-RAVE method for drug discovery. As has been observed elsewhere, the predictive power of docking against AF2 structures is quite limited, particularly for proteins like kinases that have non-trivial conformational dynamics. However, using enhanced sampling methods like RAVE to explore beyond AF2 starting structures leads to a significant improvement.

Strengths:

This is a nice demonstration of the utility of the authors' previously published RAVE method.

Weaknesses:

My only concern is the authors' discussion of induced fit. I'm quite confident the structures discussed are present in the absence of ligand binding, consistent with conformational selection. It seems the author's own data also argues for an important role in conformational selection. It would be nice to acknowledge this instead of going along with the common practice in drug discovery of attributing any conformational changes to induced fit without thoughtful consideration of conformational selection.

The reviewer is correct. We aim to highlight the significant role of conformational selection. To clarify this, we have expanded the discussion on conformational selection in the introduction.

Reviewer #3 (Public Review):

In this manuscript, the authors aim to enhance AlphaFold2 for protein conformation-selective drug discovery through the integration of AlphaFold2 and physics-based methods, focusing on improving the accuracy of predicting protein structures ensemble and small molecule binding of metastable protein conformations to facilitate targeted drug design.

The major strength of the paper lies in the methodology, which includes the innovative integration of AlphaFold2 with all-atom enhanced sampling molecular dynamics and induced fit docking to produce protein ensembles with structural diversity. Moreover, the generated structures can be used as reliable crystal-like decoys to enrich metastable conformations of holo-like structures. The authors demonstrate the effectiveness of the proposed approach in producing metastable structures of three different protein kinases and perform docking with their type I and II inhibitors. The paper provides strong evidence supporting the potential impact of this technology in drug discovery. However, limitations may exist in the generalizability of the approach across other structures, especially complex structures such as protein-protein or DNA-protein complexes.

Proteins undergo thermodynamic fluctuations and can occasionally reach metastable configurations. It can be assumed that other biomolecules, such as proteins and DNA, stabilize these metastable states when forming protein-protein or protein-DNA complexes. Since our method has the potential to identify these metastable states, it shows promise for designing drugs targeting proteins in allosteric configurations induced by other biomolecules.

The authors largely achieved their aims by demonstrating that the AF2RAVE-Glide workflow can generate holo-like structure candidates with a 50% successful docking rate for known type II inhibitors. This work is likely to have a significant impact on the field by offering a more precise and efficient method for predicting protein structure ensemble, which is essential for designing targeted drugs. The utility of the integrated AF2RAVE-Glide approach may streamline the drug discovery process, potentially leading to the development of more effective and specific medications for various diseases.

Recommendations for the authors:

Reviewer #1 (Recommendations For The Authors):

Suggestions

(1) The computational protocol is found to be insufficient to generate precise values of the relative free energies between structures generated. The authors note in the Conclusion that an enhancement in the workflow might result from the addition of free energy calculations. Can the authors comment on the prospects for generating more accurate estimates of the free energy that might be used to qualitatively evaluate poses and the free energy landscape surrounding putative metastable states? What are the principal challenges and what might help overcome them? What would the most effective computational protocol be?

More accurate estimates of the free energy can theoretically be achieved by increasing the number of umbrella sampling windows and extending the simulation length until the PMF converges. However, there is always a trade-off between PMF accuracy and computational costs, so we have chosen to stick with the current setup. Metadynamics is another method to obtain a more accurate free energy profile, which we have used in previous versions of AlphaFold2-RAVE, but for the specific systems we investigated, it had issues in achieving back and forth movement given the high entropic nature of the activation loop. Research in enhanced sampling methods and dimensionality reduction techniques for reaction coordinates is continually evolving and will play a critical role in alleviating this problem.

(2) I was surprised that there was not more correlation of a funnel-like shape in Figures S16 and S18, showing a stronger correlation between low RMSD and better docking score. This is true for both the ponatinib and imatinib applications in DDR1 and Abl1. That also seems true for the trimmed results for Src kinase in Figure S19. I was also surprised that there are structures with very large RMSD but docking scores comparable to the best structures of the lowest RMSD. Might something be done to make the docking score a more effective discriminator?

The docking algorithm and docking score are used to filter out highly improbable docking poses. False positives in predicted docking poses are a common issue across all docking methods as described for instance in:

Fan, Jiyu, Ailing Fu, and Le Zhang. "Progress in molecular docking." Quantitative Biology 7 (2019): 83-89.

Ferreira, R.S., Simeonov, A., Jadhav, A., Eidam, O., Mott, B.T., Keiser, M.J., McKerrow, J.H., Maloney, D.J., Irwin, J.J. and Shoichet, B.K., 2010. "Complementarity between a docking and a high-throughput screen in discovering new cruzain inhibitors." Journal of medicinal chemistry, 53(13), pp.4891-4905.

Moreover, there is always a trade-off between docking accuracy and computational cost. While employing more accurate docking methods may decrease false positives, it can also be resource-intensive. In such scenarios, our approach to enriching holo-structures can be impactful by reducing the number of pocket structures in the input ensembles and significantly enhancing docking efficiency.

(3) I think that it is fine to identify one structure as "IFD winner" but also feel that its significance is overstressed, especially given that it can be identified only in a retrospective analysis rather than through de novo prediction.

We agree with the reviewer. We did not intend to emphasize the specific structure "IFD winner". Rather, we aimed to demonstrate that our method can enrich promising candidates for holo-structures. We verified this by showing that our holo-structure candidates performed well in retrospective docking using IFD, which we previously referred to as "IFD winner". We have now revised this term to "holo-model".

Minor Points

p. 3 "DymanicBind" should be "DynamicBind"

p. 3 Change "We chosen" to "We have chosen" or "we chose."

p. 3 In identifying the Schrödinger software Glide and IFD, I recommend removing the subjective modifier "industry-leading."

Modifications done.

Reviewer #2 (Recommendations For The Authors):

In the view of this reviewer, the writing is 'choppy'.

We have tried to improve the writing.

Reviewer #3 (Recommendations For The Authors):

(1) In Figure 1, the workflow labels (i) to (iv) are not shown on the figures, making it difficult for readers to follow. Consider adding these labels to the figures.

Modifications done.

(2) Explain how Boltzmann ranks were calculated based on unbiased MD simulations to guide the enrichment of holo-like structures in metastable states.

The Methods section is now updated for clarification.

(3) The authors could clarify how the classical DFG-out decoys in the DDR1 rMSA AF2 ensemble are transferred to Abl1 kinase in the Methods section.

The Methods section is now updated for clarification.

(4) The authors can clarify the methodology section by providing more detailed explanations about how the unbiased MD simulations are performed, including which MD simulation software was used and whether energy minimization and equilibrium steps were needed as in conventional MD simulations, and other setup details.

The Methods section is now updated for clarification.

(5) The validation of the proposed approach in this work used three kinase proteins. The authors can enhance the discussion section by addressing other types of protein structure prediction that can use the proposed approach in drug discovery, beyond the three kinase proteins tested.

The proposed approach is theoretically applicable to other types of proteins, such as GPCRs, where both conformational selection and the induced-fit effect are crucial. We have expanded the discussion on the generalization of our protocol in the Conclusion section.

(6) The authors should add appropriate citations for the software and tools used in the manuscript. For example, a reference should be added for the Glide XP docking experiments that utilized the Maestro software. Double-check all related software citations.

We have now updated the citations for docking experiments based on the instruction of the Maestro Glide User manual and IFD User manual.

(7) The authors should consider offering a comprehensive list of software tools and databases utilized in the study to assist in replicating the experiments and further validating the results.

We have now added a summary of tools used in the Methods section.

No competing interests declared.

P.T. is a consultant to Schrodinger, Inc and is on their Scientific Advisory Board.

Conceptualization, Resources, Data curation, Software, Formal analysis, Supervision, Validation, Investigation, Visualization, Methodology, Writing - original draft, Project administration, Writing – review and editing.

Conceptualization, Resources, Data curation, Software, Formal analysis, Validation, Investigation, Visualization, Methodology, Writing - original draft, Writing – review and editing.

Conceptualization, Formal analysis, Supervision, Funding acquisition, Project administration, Writing – review and editing.
==== Refs
References

Abramson J Adler J Dunger J Evans R Green T Pritzel A Ronneberger O Willmore L Ballard AJ Bambrick J Bodenstein SW Evans DA Hung C-C O’Neill M Reiman D Tunyasuvunakool K Wu Z Žemgulytė A Arvaniti E Beattie C Bertolli O Bridgland A Cherepanov A Congreve M Cowen-Rivers AI Cowie A Figurnov M Fuchs FB Gladman H Jain R Khan YA Low CMR Perlin K Potapenko A Savy P Singh S Stecula A Thillaisundaram A Tong C Yakneen S Zhong ED Zielinski M Žídek A Bapst V Kohli P Jaderberg M Hassabis D Jumper JM 2024 Accurate structure prediction of biomolecular interactions with AlphaFold 3 Nature 630 493 500 10.1038/s41586-024-07487-w 38718835
Al-Masri C Trozzi F Lin S-H Tran O Sahni N Patek M Cichonska A Ravikumar B Rahman R 2023 Investigating the conformational landscape of AlphaFold2-predicted protein kinase structures Bioinformatics Advances 3 vbad129 10.1093/bioadv/vbad129 37786533
Amaro RE 2019 Will the real cryptic pocket please stand out? Biophysical Journal 116 753 754 10.1016/j.bpj.2019.01.018 30739726
Beuming T Martín H Díaz-Rovira AM Díaz L Guallar V Ray SS 2022 Are Deep Learning Structural Models Sufficiently Accurate for Free-Energy Calculations? Application of FEP+ to AlphaFold2-Predicted Structures Journal of Chemical Information and Modeling 62 4351 4360 10.1021/acs.jcim.2c00796 36099477
Case DA Cheatham TE Darden T Gohlke H Luo R Merz KM Onufriev A Simmerling C Wang B Woods RJ 2005 The Amber biomolecular simulation programs Journal of Computational Chemistry 26 1668 1688 10.1002/jcc.20290 16200636
Corso G Stärk H Jing B Barzilay R Jaakkola T 2022 Diffdock: Diffusion Steps, Twists, and Turns for Molecular Docking arXiv https://arxiv.org/abs/2210.01776
Coskun D Lihan M Rodrigues J Vass M Robinson D Friesner RA Miller EB 2024 Using AlphaFold and Experimental Structures for the Prediction of the Structure and Binding Affinities of GPCR Complexes via Induced Fit Docking and Free Energy Perturbation Journal of Chemical Theory and Computation 20 477 489 10.1021/acs.jctc.3c00839 38100422
Darden T York D Pedersen L 1993 Particle mesh ewald: An n⋅ log (n) method for ewald sums in large systems The Journal of Chemical Physics 98 10089 10092 10.1063/1.464397
Davis MI Hunt JP Herrgard S Ciceri P Wodicka LM Pallares G Hocker M Treiber DK Zarrinkar PP 2011 Comprehensive analysis of kinase inhibitor selectivity Nature Biotechnology 29 1046 1051 10.1038/nbt.1990 22037378
Del Alamo D Sala D Mchaourab HS Meiler J 2022 Sampling alternative conformational states of transporters and receptors with AlphaFold2 eLife 11 e75751 10.7554/eLife.75751 35238773
Díaz-Rovira AM Martín H Beuming T Díaz L Guallar V Ray SS 2023 Are deep learning structural models sufficiently accurate for virtual screening? Application of docking algorithms to alphafold2 predicted structures Journal of Chemical Information and Modeling 63 1668 1674 10.1021/acs.jcim.2c01270 36892986
Eastman P Swails J Chodera JD McGibbon RT Zhao Y Beauchamp KA Wang L-P Simmonett AC Harrigan MP Stern CD Wiewiora RP Brooks BR Pande VS 2017 OpenMM 7: Rapid development of high performance algorithms for molecular dynamics PLOS Computational Biology 13 e1005659 10.1371/journal.pcbi.1005659 28746339
Friesner RA Banks JL Murphy RB Halgren TA Klicic JJ Mainz DT Repasky MP Knoll EH Shelley M Perry JK Shaw DE Francis P Shenkin PS 2004 Glide: a new approach for rapid, accurate docking and scoring. 1. Method and assessment of docking accuracy Journal of Medicinal Chemistry 47 1739 1749 10.1021/jm0306430 15027865
Friesner RA Murphy RB Repasky MP Frye LL Greenwood JR Halgren TA Sanschagrin PC Mainz DT 2006 Extra precision glide: docking and scoring incorporating a model of hydrophobic enclosure for protein-ligand complexes Journal of Medicinal Chemistry 49 6177 6196 10.1021/jm051256o 17034125
Gizzio J Thakur A Haldane A Levy RM 2022 Evolutionary divergence in the conformational landscapes of tyrosine vs serine/threonine kinases eLife 11 e83368 10.7554/eLife.83368 36562610
Gizzio J Thakur A 2024 Evolutionary Sequence and Structural Basis for the Distinct Conformational Landscapes of Tyr and Ser/Thr Kinases bioRxiv 10.1101/2024.03.08.584161
Gu X 2024 AF2RAVE_Glide-kinase swh:1:rev:f9722a58d97ce0e94f8777c9ba02087a7e78644c Software Heritage https://archive.softwareheritage.org/swh:1:dir:f665bc3113e2b011c6a8f54054f1831fa744e2f3;origin=https://github.com/tiwarylab/AF2RAVE_Glide-kinase;visit=swh:1:snp:6199883ebbe7be2ac6a6269b705e96e477ba35c8;anchor=swh:1:rev:f9722a58d97ce0e94f8777c9ba02087a7e78644c
Guterres H Park S-J Jiang W Im W 2021 Ligand-binding-site refinement to generate reliable holo protein structure conformations from apo structures Journal of Chemical Information and Modeling 61 535 546 10.1021/acs.jcim.0c01354 33337877
Halgren TA Murphy RB Friesner RA Beard HS Frye LL Pollard WT Banks JL 2004 Glide: a new approach for rapid, accurate docking and scoring. 2. Enrichment factors in database screening Journal of Medicinal Chemistry 47 1750 1759 10.1021/jm030644s 15027866
Hanson SM Georghiou G Thakur MK Miller WT Rest JS Chodera JD Seeliger MA 2019 What makes a kinase promiscuous for inhibitors? Cell Chemical Biology 26 390 399 10.1016/j.chembiol.2018.11.005 30612951
Hauser AS Chavali S Masuho I Jahn LJ Martemyanov KA Gloriam DE Babu MM 2018 Pharmacogenomics of GPCR Drug Targets Cell 172 41 54 10.1016/j.cell.2017.11.033 29249361
Herron L Mondal K Schneekloth JS Tiwary P 2023 Inferring Phase Transitions and Critical Exponents from Limited Observations with Thermodynamic Maps arXiv https://arxiv.org/abs/2308.14885
Hess B Bekker H Berendsen HJC Fraaije JGEM 1997 LINCS: A linear constraint solver for molecular simulations Journal of Computational Chemistry 18 1463 1472 10.1002/(SICI)1096-987X(199709)18:12<1463::AID-JCC4>3.3.CO;2-L
Holcomb M Chang Y-T Goodsell DS Forli S 2023 Evaluation of AlphaFold2 structures as docking targets Protein Science 32 e4530 10.1002/pro.4530 36479776
Jing B Berger B Jaakkola T 2024 Alphafold Meets Flow Matching for Generating Protein Ensembles arXiv https://arxiv.org/abs/2402.04845
Jorgensen WL Chandrasekhar J Madura JD Impey RW Klein ML 1983 Comparison of simple potential functions for simulating liquid water The Journal of Chemical Physics 79 926 935 10.1063/1.445869
Jumper J Evans R Pritzel A Green T Figurnov M Ronneberger O Tunyasuvunakool K Bates R Žídek A Potapenko A Bridgland A Meyer C Kohl SAA Ballard AJ Cowie A Romera-Paredes B Nikolov S Jain R Adler J Back T Petersen S Reiman D Clancy E Zielinski M Steinegger M Pacholska M Berghammer T Bodenstein S Silver D Vinyals O Senior AW Kavukcuoglu K Kohli P Hassabis D 2021 Highly accurate protein structure prediction with AlphaFold Nature 596 583 589 10.1038/s41586-021-03819-2 34265844
Lindorff-Larsen K Piana S Palmo K Maragakis P Klepeis JL Dror RO Shaw DE 2010 Improved side-chain torsion potentials for the Amber ff99SB protein force field Proteins 78 1950 1958 10.1002/prot.22711 20408171
Lu W Zhang J Huang W Zhang Z Jia X Wang Z Shi L Li C Wolynes PG Zheng S 2024 DynamicBind: predicting ligand-specific protein-ligand complex structure with a deep equivariant generative model Nature Communications 15 1071 10.1038/s41467-024-45461-2 38316797
Lyu J Wang S Balius TE Singh I Levit A Moroz YS O’Meara MJ Che T Algaa E Tolmachova K Tolmachev AA Shoichet BK Roth BL Irwin JJ 2019 Ultra-large library docking for discovering new chemotypes Nature 566 224 229 10.1038/s41586-019-0917-9 30728502
Lyu J Kapolka N Gumpper R Alon A Wang L Jain MK Barros-Álvarez X Sakamoto K Kim Y DiBerto J Kim K Glenn IS Tummino TA Huang S Irwin JJ Tarkhanova OO Moroz Y Skiniotis G Kruse AC Shoichet BK Roth BL 2024 AlphaFold2 structures guide prospective ligand discovery Science 384 eadn6354 10.1126/science.adn6354 38753765
Maestro 2023 Schrödinger Release 2023-3: Glide; Induced Fit Docking Protocol; Prime New York Schrödinger, LLC
Mehdi S Smith Z Herron L Zou Z Tiwary P 2024 Enhanced sampling with machine learning Annual Review of Physical Chemistry 75 347 370 10.1146/annurev-physchem-083122-125941 38382572
Meller A Bhakat S Solieva S Bowman GR 2023 Accelerating cryptic pocket discovery using alphafold Journal of Chemical Theory and Computation 19 4355 4363 10.1021/acs.jctc.2c01189 36948209
Mirdita M Schütze K Moriwaki Y Heo L Ovchinnikov S Steinegger M 2022 ColabFold: making protein folding accessible to all Nature Methods 19 679 682 10.1038/s41592-022-01488-1 35637307
Modi V Dunbrack RL 2019 Defining a new nomenclature for the structures of active and inactive kinases PNAS 116 6818 6827 10.1073/pnas.1814279116 30867294
Modi V Dunbrack RL 2022 Kincore: a web resource for structural classification of protein kinases and their inhibitors Nucleic Acids Research 50 D654 D664 10.1093/nar/gkab920 34643709
Monteiro da Silva G Cui JY Dalgarno DC Lisi GP Rubenstein BM 2024 High-throughput prediction of protein conformational distributions with subsampled AlphaFold2 Nature Communications 15 2464 10.1038/s41467-024-46715-9 38538622
Müller S Chaikuad A Gray NS Knapp S 2015 The ins and outs of selective kinase inhibitor development Nature Chemical Biology 11 818 821 10.1038/nchembio.1938 26485069
Porter LL Chakravarty D Schafer JW Chen EA 2023 Colabfold Predicts Alternative Protein Structures from Single Sequences, Coevolution Unnecessary for Af-Cluster bioRxiv 10.1101/2023.11.21.567977
Ren F Ding X Zheng M Korzinkin M Cai X Zhu W Mantsyzov A Aliper A Aladinskiy V Cao Z Kong S Long X Man Liu BH Liu Y Naumov V Shneyderman A Ozerov IV Wang J Pun FW Polykovskiy DA Sun C Levitt M Aspuru-Guzik A Zhavoronkov A 2023 AlphaFold accelerates artificial intelligence powered drug discovery: efficient discovery of a novel CDK20 small molecule inhibitor Chemical Science 14 1443 1452 10.1039/d2sc05709c 36794205
Roney JP Ovchinnikov S 2022 State-of-the-art estimation of protein model accuracy using alphafold Physical Review Letters 129 238101 10.1103/PhysRevLett.129.238101 36563190
Sala D Hildebrand PW Meiler J 2023 Biasing AlphaFold2 to predict GPCRs and kinases with user-defined functional or structural properties Frontiers in Molecular Biosciences 10 1121962 10.3389/fmolb.2023.1121962 36876042
Scardino V Di Filippo JI Cavasotto CN 2023 How good are AlphaFold models for docking-based virtual screening? iScience 26 105920 10.1016/j.isci.2022.105920 36686396
Sherman W Beard HS Farid R 2006a Use of an induced fit receptor structure in virtual screening Chemical Biology & Drug Design 67 83 84 10.1111/j.1747-0285.2005.00327.x 16492153
Sherman W Day T Jacobson MP Friesner RA Farid R 2006b Novel procedure for modeling ligand/receptor induced fit effects Journal of Medicinal Chemistry 49 534 553 10.1021/jm050540c 16420040
Smith Z Strobel M Vani BP Tiwary P 2024 Graph attention site prediction (GrASP): Identifying druggable binding sites using graph neural networks with attention Journal of Chemical Information and Modeling 64 2637 2644 10.1021/acs.jcim.3c01698 38453912
Thakur A Gizzio J Levy RM 2024 Potts hamiltonian models and molecular dynamics free energy simulations for predicting the impact of mutations on protein kinase stability The Journal of Physical Chemistry. B 128 1656 1667 10.1021/acs.jpcb.3c08097 38350894
Vani BP Aranganathan A Wang D Tiwary P 2023 AlphaFold2-RAVE: From Sequence to Boltzmann Ranking Journal of Chemical Theory and Computation 19 4351 4354 10.1021/acs.jctc.3c00290 37171364
Vani BP Aranganathan A Tiwary P 2024 Exploring Kinase Asp-Phe-Gly (DFG) Loop Conformational Stability with AlphaFold2-RAVE Journal of Chemical Information and Modeling 64 2789 2797 10.1021/acs.jcim.3c01436 37981824
Wang Y Ribeiro JML Tiwary P 2019 Past-future information bottleneck for sampling molecular reaction coordinate simultaneously with thermodynamics and kinetics Nature Communications 10 3573 10.1038/s41467-019-11405-4 31395868
Wang D Tiwary P 2021 State predictive information bottleneck The Journal of Chemical Physics 154 134111 10.1063/5.0038198 33832235
Wang Y Herron L Tiwary P 2022 From data to noise to data for mixing physics across temperatures with generative artificial intelligence PNAS 119 e2203656119 10.1073/pnas.2203656119 35925885
Wayment-Steele HK Ojoawo A Otten R Apitz JM Pitsawong W Hömberger M Ovchinnikov S Colwell L Kern D 2024 Predicting multiple conformations via sequence clustering and AlphaFold2 Nature 625 832 839 10.1038/s41586-023-06832-9 37956700
Zhang Z Liu X Yan K Tuckerman ME Liu J 2019 Unified efficient thermostat scheme for the canonical ensemble with holonomic or isokinetic constraints via molecular dynamics The Journal of Physical Chemistry. A 123 6056 6079 10.1021/acs.jpca.9b02771 31117592
Zhang Y Vass M Shi D Abualrous E Chambers JM Chopra N Higgs C Kasavajhala K Li H Nandekar P Sato H Miller EB Repasky MP Jerome SV 2023 Benchmarking refined and unrefined alphafold2 structures for hit discovery Journal of Chemical Information and Modeling 63 1656 1667 10.1021/acs.jcim.2c01219 36897766
Zheng S He J Liu C Shi Y Lu Z Feng W Ju F Wang J Zhu J Min Y Zhang H Tang S Hao H Jin P Chen C Noé F Liu H Liu T-Y 2024 Predicting equilibrium distributions for molecular systems with deep learning Nature Machine Intelligence 6 558 567 10.1038/s42256-024-00837-3
