bioRxiv ScienceSearch

bioRxiv · 10.1101/790949

Novel, provable algorithms for efficient ensemble-based computational protein design and their application to the redesign of the c-Raf-RBD:KRas protein-protein interface

Abstract

The K* algorithm provably approximates partition functions for a set of states (e.g., protein, ligand, and protein-ligand complex) to a user-specified accuracy{varepsilon} . Often, reaching an{varepsilon} -approximation for a particular set of partition functions takes a prohibitive amount of time and space. To alleviate some of this cost, we introduce two algorithms into the osprey suite for protein design: O_SCPLOWFRIESC_SCPLOW, a Fast Removal of Inadequately Energied Sequences, and EWAK*, an Energy Window Approximation to K*. In combination, these algorithms provably retain calculational accuracy while limiting the input sequence space and the conformations included in each partition function calculation to only the most energetically favorable. This combined approach leads to significant speed-ups compared to the previous state-of-the-art multi-sequence algorithm, BBK*. As a proof of concept, we used these new algorithms to redesign the protein-protein interface (PPI) of the c-Raf-RBD:KRas complex. The Ras-binding domain of the protein kinase c-Raf (c-Raf-RBD) is the tightest known binder of KRas, a historically \"undruggable\" protein implicated in difficult-to-treat cancers including pancreatic ductal adenocarcinoma (PDAC). O_SCPLOWFRIESC_SCPLOW/EWAK* accurately retrospectively predicted the effect of 38 out of 41 different sets of mutations in the PPI of the c-Raf-RBD:KRas complex. Notably, these mutations include mutations whose effect had previously been incorrectly predicted using other computational methods. Next, we used O_SCPLOWFRIESC_SCPLOW/EWAK* for prospective design and discovered a novel point mutation that improves binding of c-Raf-RBD to KRas in its active, GTP-bound state (KRasGTP). We combined this new mutation with two previously reported mutations (which were also highly-ranked by O_SCPLOWOSPREYC_SCPLOW) to create a new variant of c-Raf-RBD, c-Raf-RBD(RKY). O_SCPLOWFRIESC_SCPLOW/EWAK* in O_SCPLOWOSPREYC_SCPLOW computationally predicted that this new variant would bind even more tightly than the previous best-binding variant, c-Raf-RBD(RK). We measured the binding affinity of c-Raf-RBD(RKY) using a bio-layer interferometry (BLI) assay and found that this new variant exhibits single-digit nanomolar affinity for KRasGTP, confirming the computational predictions made with O_SCPLOWFRIESC_SCPLOW/EWAK*. This study steps through the advancement and development of computational protein design by presenting theory, new algorithms, accurate retrospective designs, new prospective designs, and biochemical validation.\n\nAuthor summaryComputational structure-based protein design is an innovative tool for redesigning proteins to introduce a particular or novel function. One such possible function is improving the binding of one protein to another, which can increase our understanding of biomedically important protein systems toward the improvement or development of novel therapeutics. Herein we introduce two novel, provable algorithms, O_SCPLOWFRIESC_SCPLOW and EWAK*, for more efficient computational structure-based protein design as well as their application to the redesign of the c-Raf-RBD:KRas protein-protein interface. These new algorithms speed up computational structure-based protein design while maintaining accurate calculations, allowing for larger, previously infeasible protein designs. Using O_SCPLOWFRIESC_SCPLOW and EWAK* within the O_SCPLOWOSPREYC_SCPLOW suite, we designed the tightest known binder of KRas, an \"undruggable\" cancer target. This new variant of a KRas-binding domain, c-Raf-RBD, should serve as an important tool to probe the protein-protein interface between KRas and its effectors as work continues toward an effective therapeutic targeting KRas.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Lowegard, A. U., Frenkel, M., Jou, J., Ojewole, A., Holt, G., Donald, B.. 2019-10-02. Novel, provable algorithms for efficient ensemble-based computational protein design and their application to the redesign of the c-Raf-RBD:KRas protein-protein interface. https://doi.org/10.1101/790949

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Hierarchical cysteine oxidation controls reversible amyloid formation in an ankyrin repeat protein

The formation of amyloids, including functional amyloids, is observed for an increasing number of proteins but the molecular mechanisms that control this structural transition remain poorly understood. Here we report that the kinase inhibitor protein P18 (drP18) from Danio rerio (zebrafish), which contains two cysteine residues, undergoes a complex and hierarchical redox switch that strictly governs reversible amyloid formation. We identify cysteine 50 (C50) acting as a regulatory residue. Upon oxidation, C50 forms an intramolecular disulfide bond with the executioner cysteine 128 (C128), thereby blocking it. C50 can become S-glutathionylated, and upon oxidation, C128 then forms intermolecular disulfides that lead to rapid transition into amyloid fibrils. S-glutathionylation of C50 therefore enables amyloid formation of drP18 and the outcome is oxidant-dependent with diamide, hydrogen peroxide, peroxymonocarbonate and hypothiocyanous acid each leading to amyloid assembly with distinct kinetics and morphologies. These amyloids are fully reversible, where disulfide reduction is leading to disassembly. Whereas monomeric drP18 inhibits CDK4-mediated retinoblastoma phosphorylation, the amyloid conformation abolishes this inhibition, and reduction restores both structure and function. Expression of drP18 in zebrafish embryos yields Congo red-positive, oxidation-dependent aggregates in vivo. Together, our findings show that a regulatory cysteine controls an executioner cysteine to induce reversible, functional amyloid formation, revealing that proteins can encode sophisticated mechanisms to control amyloid assembly.

biochemistry

Snapshots from the Catalytic Landscape of Chalcone Isomerase

Chalcone isomerase (CHI) catalyzes the cyclization of 3-ring scaffolds of flavonoids, a class of plant-based natural products important for nutrition and disease prevention. A persistent question has been whether the enzyme uses dynamics to facilitate conformational rearrangements of substrates within the active site. To help resolve this question, CHI was crystallized with phloretin, a flexible substrate analogue that cannot undergo cyclization. The crystal structure possesses eight protein molecules per asymmetric unit, revealing different active site conformations that accommodate different bound conformers of phloretin. Together, the structural snapshots depict a series of coordinated, dynamic chemical interactions that lower barriers to substrate rearrangements approaching bond formation. Differential scanning fluorimetry combined with mutational analysis and enzyme kinetics further confirm that phloretin binds to the enzyme active site and that it acts as a competitive inhibitor of CHI. Together these findings answer outstanding questions about the flexibility and dynamics of CHI catalysis, information that may be useful for future biosynthetic design and enzyme engineering goals. Overall, this work supports a catalytic model in which the CHI enzyme operates as a dynamic ensemble of structures necessary to facilitate catalytic substrate rearrangements.

biochemistry

Structures of pUG-fold RNA bound to DNMT1 reveal a mechanism for RNA-mediated epigenetic regulation

Many chromatin-associated proteins have been found to bind RNA as a means of epigenetic regulation. Specifically, DNA methyltransferase 1 (DNMT1), which maintains cytosine methylation at CpG dinucleotides, is inhibited by RNA at transcribed DNA loci in cells. However, the mechanisms by which RNA binds DNMT1 and inhibits its activity remain unknown. Here, we determine a series of cryogenic electron microscopy (cryo-EM) structures of human DNMT1 bound to pUG-fold RNA, a non-canonical G-quadruplex previously observed to inhibit activity, revealing two distinct RNA-binding modes. The pUG-fold RNA binds the surface of DNMT1 in its autoinhibited conformation across a positively charged surface between the methyltransferase domain and the CXXC domain, and it binds directly in the active site of an open DNMT1 conformation. RNA binding is sterically incompatible with substrate DNA engagement in both states. Our 2.5 [A] structure captures the intricate network of hydrogen bonds and electrostatic interactions between amino acids in the methyltransferase domain and the tetrad layers of pUG-fold RNA. Metadynamics molecular dynamics simulations provide an orthogonal view of the conformational landscape of DNMT1, revealing the two distinct RNA-binding modes. Furthermore, our analysis of published DNMT1 RIP-seq and eCLIP-seq data confirms that DNMT1-interacting RNAs in cells exhibit a strong propensity to form non-canonical G-quadruplex RNA structures. Collectively, our study provides the first structural basis for pUG-fold RNA recognition by a protein and illustrates how cryo-EM and AI-based methods for protein and RNA structure prediction synergize to inform the mechanism of RNA-mediated regulation of DNMT1.

biochemistry