The aim of de novo protein design is to find the amino acid sequences that will fold into a desired 3-dimensional structure with improvements in specific properties, such as binding affinity, agonist or antagonist behavior, or stability, relative to the native sequence. Protein design lies at the center of current advances drug design and discovery. Not only does protein design provide predictions for potentially useful drug targets, but it also enhances our understanding of the protein folding process and protein-protein interactions. Experimental methods such as directed evolution have shown success in protein design. However, such methods are restricted by the limited sequence space that can be searched tractably. In contrast, computational design strategies allow for the screening of a much larger set of sequences covering a wide variety of properties and functionality. We have developed a range of computational de novo protein design methods capable of tackling several important areas of protein design. These include the design of monomeric proteins for increased stability and complexes for increased binding affinity.
To disseminate these methods for broader use we present Protein WISDOM (https://www.proteinwisdom.org), a tool that provides automated methods for a variety of protein design problems. Structural templates are submitted to initialize the design process. The first stage of design is an optimization sequence selection stage that aims at improving stability through minimization of potential energy in the sequence space. Selected sequences are then run through a fold specificity stage and a binding affinity stage. A rank-ordered list of the sequences for each step of the process, along with relevant designed structures, provides the user with a comprehensive quantitative assessment of the design. Here we provide the details of each design method, as well as several notable experimental successes attained through the use of the methods.
16 Related JoVE Articles!
A Protocol for Computer-Based Protein Structure and Function Prediction
Institutions: University of Michigan , University of Kansas.
Genome sequencing projects have ciphered millions of protein sequence, which require knowledge of their structure and function to improve the understanding of their biological role. Although experimental methods can provide detailed information for a small fraction of these proteins, computational modeling is needed for the majority of protein molecules which are experimentally uncharacterized. The I-TASSER server is an on-line workbench for high-resolution modeling of protein structure and function. Given a protein sequence, a typical output from the I-TASSER server includes secondary structure prediction, predicted solvent accessibility of each residue, homologous template proteins detected by threading and structure alignments, up to five full-length tertiary structural models, and structure-based functional annotations for enzyme classification, Gene Ontology terms and protein-ligand binding sites. All the predictions are tagged with a confidence score which tells how accurate the predictions are without knowing the experimental data. To facilitate the special requests of end users, the server provides channels to accept user-specified inter-residue distance and contact maps to interactively change the I-TASSER modeling; it also allows users to specify any proteins as template, or to exclude any template proteins during the structure assembly simulations. The structural information could be collected by the users based on experimental evidences or biological insights with the purpose of improving the quality of I-TASSER predictions. The server was evaluated as the best programs for protein structure and function predictions in the recent community-wide CASP experiments. There are currently >20,000 registered scientists from over 100 countries who are using the on-line I-TASSER server.
Biochemistry, Issue 57, On-line server, I-TASSER, protein structure prediction, function prediction
Determination of Protein-ligand Interactions Using Differential Scanning Fluorimetry
Institutions: University of Exeter.
A wide range of methods are currently available for determining the dissociation constant between a protein and interacting small molecules. However, most of these require access to specialist equipment, and often require a degree of expertise to effectively establish reliable experiments and analyze data. Differential scanning fluorimetry (DSF) is being increasingly used as a robust method for initial screening of proteins for interacting small molecules, either for identifying physiological partners or for hit discovery. This technique has the advantage that it requires only a PCR machine suitable for quantitative PCR, and so suitable instrumentation is available in most institutions; an excellent range of protocols are already available; and there are strong precedents in the literature for multiple uses of the method. Past work has proposed several means of calculating dissociation constants from DSF data, but these are mathematically demanding. Here, we demonstrate a method for estimating dissociation constants from a moderate amount of DSF experimental data. These data can typically be collected and analyzed within a single day. We demonstrate how different models can be used to fit data collected from simple binding events, and where cooperative binding or independent binding sites are present. Finally, we present an example of data analysis in a case where standard models do not apply. These methods are illustrated with data collected on commercially available control proteins, and two proteins from our research program. Overall, our method provides a straightforward way for researchers to rapidly gain further insight into protein-ligand interactions using DSF.
Biophysics, Issue 91, differential scanning fluorimetry, dissociation constant, protein-ligand interactions, StepOne, cooperativity, WcbI.
Real Time Measurements of Membrane Protein:Receptor Interactions Using Surface Plasmon Resonance (SPR)
Institutions: The Technion-Israel Institute of Technology.
Protein-protein interactions are pivotal to most, if not all, physiological processes, and understanding the nature of such interactions is a central step in biological research. Surface Plasmon Resonance (SPR) is a sensitive detection technique for label-free study of bio-molecular interactions in real time. In a typical SPR experiment, one component (usually a protein, termed 'ligand') is immobilized onto a sensor chip surface, while the other (the 'analyte') is free in solution and is injected over the surface. Association and dissociation of the analyte from the ligand are measured and plotted in real time on a graph called a sensogram, from which pre-equilibrium and equilibrium data is derived. Being label-free, consuming low amounts of material, and providing pre-equilibrium kinetic data, often makes SPR the method of choice when studying dynamics of protein interactions. However, one has to keep in mind that due to the method's high sensitivity, the data obtained needs to be carefully analyzed, and supported by other biochemical methods.
SPR is particularly suitable for studying membrane proteins since it consumes small amounts of purified material, and is compatible with lipids and detergents. This protocol describes an SPR experiment characterizing the kinetic properties of the interaction between a membrane protein (an ABC transporter) and a soluble protein (the transporter's cognate substrate binding protein).
Structural Biology, Issue 93, ABC transporter, substrate binding protein, bio-molecular interaction kinetics, label-free, protein-protein interaction, Surface plasmon resonance (SPR), Biacore
Bio-layer Interferometry for Measuring Kinetics of Protein-protein Interactions and Allosteric Ligand Effects
Institutions: SUNY Upstate Medical University.
We describe the use of Bio-layer Interferometry to study inhibitory interactions of subunit ε with the catalytic complex of Escherichia coli
ATP synthase. Bacterial F-type ATP synthase
is the target of a new, FDA-approved antibiotic to combat drug-resistant tuberculosis. Understanding bacteria-specific auto-inhibition of ATP synthase by the C-terminal domain of subunit ε could provide a new means to target the enzyme for discovery of antibacterial drugs. The C-terminal domain of ε undergoes a dramatic conformational change when the enzyme transitions between the active and inactive states, and catalytic-site ligands can influence which of ε's conformations is predominant. The assay measures kinetics of ε's binding/dissociation with the catalytic complex, and indirectly measures the shift of enzyme-bound ε to and from the apparently nondissociable inhibitory conformation. The Bio-layer Interferometry signal is not overly sensitive to solution composition, so it can also be used to monitor allosteric effects of catalytic-site ligands on ε's conformational changes.
Chemistry, Issue 84, ATP synthase, Bio-Layer Interferometry, Ligand-induced conformational change, Biomolecular Interaction Analysis, Allosteric regulation, Enzyme inhibition
Nanomechanics of Drug-target Interactions and Antibacterial Resistance Detection
Institutions: University College London.
The cantilever sensor, which acts as a transducer of reactions between model bacterial cell wall matrix immobilized on its surface and antibiotic drugs in solution, has shown considerable potential in biochemical sensing applications with unprecedented sensitivity and specificity1-5
. The drug-target interactions generate surface stress, causing the cantilever to bend, and the signal can be analyzed optically when it is illuminated by a laser. The change in surface stress measured with nano-scale precision allows disruptions of the biomechanics of model bacterial cell wall targets to be tracked in real time. Despite offering considerable advantages, multiple cantilever sensor arrays have never been applied in quantifying drug-target binding interactions.
Here, we report on the use of silicon multiple cantilever arrays coated with alkanethiol self-assembled monolayers mimicking bacterial cell wall matrix to quantitatively study antibiotic binding interactions. To understand the impact of vancomycin on the mechanics of bacterial cell wall structures1,6,7
. We developed a new model1
which proposes that cantilever bending can be described by two independent factors; i) namely a chemical factor, which is given by a classical Langmuir adsorption isotherm, from which we calculate the thermodynamic equilibrium dissociation constant (Kd
) and ii) a geometrical factor, essentially a measure of how bacterial peptide receptors are distributed on the cantilever surface. The surface distribution of peptide receptors (p
) is used to investigate the dependence of geometry and ligand loading. It is shown that a threshold value of p ~
10% is critical to sensing applications. Below which there is no detectable bending signal while above this value, the bending signal increases almost linearly, revealing that stress is a product of a local chemical binding factor and a geometrical factor combined by the mechanical connectivity of reacted regions and provides a new paradigm for design of powerful agents to combat superbug infections.
Immunology, Issue 80, Engineering, Technology, Diagnostic Techniques and Procedures, Early Diagnosis, Bacterial Infections and Mycoses, Lipids, Amino Acids, Peptides, and Proteins, Chemical Actions and Uses, Diagnosis, Therapeutics, Surface stress, vancomycin, mucopeptides, cantilever sensor
Reconstitution of a Kv Channel into Lipid Membranes for Structural and Functional Studies
Institutions: University of Texas Southwestern Medical Center at Dallas.
To study the lipid-protein interaction in a reductionistic fashion, it is necessary to incorporate the membrane proteins into membranes of well-defined lipid composition. We are studying the lipid-dependent gating effects in a prototype voltage-gated potassium (Kv) channel, and have worked out detailed procedures to reconstitute the channels into different membrane systems. Our reconstitution procedures take consideration of both detergent-induced fusion of vesicles and the fusion of protein/detergent micelles with the lipid/detergent mixed micelles as well as the importance of reaching an equilibrium distribution of lipids among the protein/detergent/lipid and the detergent/lipid mixed micelles. Our data suggested that the insertion of the channels in the lipid vesicles is relatively random in orientations, and the reconstitution efficiency is so high that no detectable protein aggregates were seen in fractionation experiments. We have utilized the reconstituted channels to determine the conformational states of the channels in different lipids, record electrical activities of a small number of channels incorporated in planar lipid bilayers, screen for conformation-specific ligands from a phage-displayed peptide library, and support the growth of 2D crystals of the channels in membranes. The reconstitution procedures described here may be adapted for studying other membrane proteins in lipid bilayers, especially for the investigation of the lipid effects on the eukaryotic voltage-gated ion channels.
Molecular Biology, Issue 77, Biochemistry, Genetics, Cellular Biology, Structural Biology, Biophysics, Membrane Lipids, Phospholipids, Carrier Proteins, Membrane Proteins, Micelles, Molecular Motor Proteins, life sciences, biochemistry, Amino Acids, Peptides, and Proteins, lipid-protein interaction, channel reconstitution, lipid-dependent gating, voltage-gated ion channel, conformation-specific ligands, lipids
Collecting Variable-concentration Isothermal Titration Calorimetry Datasets in Order to Determine Binding Mechanisms
Institutions: McGill University.
Isothermal titration calorimetry (ITC) is commonly used to determine the thermodynamic parameters associated with the binding of a ligand to a host macromolecule. ITC has some advantages over common spectroscopic approaches for studying host/ligand interactions. For example, the heat released or absorbed when the two components interact is directly measured and does not require any exogenous reporters. Thus the binding enthalpy and the association constant (Ka) are directly obtained from ITC data, and can be used to compute the entropic contribution. Moreover, the shape of the isotherm is dependent on the c-value and the mechanistic model involved. The c-value is defined as c = n[P]tKa, where [P]t is the protein concentration, and n is the number of ligand binding sites within the host. In many cases, multiple binding sites for a given ligand are non-equivalent and ITC allows the characterization of the thermodynamic binding parameters for each individual binding site. This however requires that the correct binding model be used. This choice can be problematic if different models can fit the same experimental data. We have previously shown that this problem can be circumvented by performing experiments at several c-values. The multiple isotherms obtained at different c-values are fit simultaneously to separate models. The correct model is next identified based on the goodness of fit across the entire variable-c dataset. This process is applied here to the aminoglycoside resistance-causing enzyme aminoglycoside N-6'-acetyltransferase-Ii (AAC(6')-Ii). Although our methodology is applicable to any system, the necessity of this strategy is better demonstrated with a macromolecule-ligand system showing allostery or cooperativity, and when different binding models provide essentially identical fits to the same data. To our knowledge, there are no such systems commercially available. AAC(6')-Ii, is a homo-dimer containing two active sites, showing cooperativity between the two subunits. However ITC data obtained at a single c-value can be fit equally well to at least two different models a two-sets-of-sites independent model and a two-site sequential (cooperative) model. Through varying the c-value as explained above, it was established that the correct binding model for AAC(6')-Ii is a two-site sequential binding model. Herein, we describe the steps that must be taken when performing ITC experiments in order to obtain datasets suitable for variable-c analyses.
Biochemistry, Issue 50, ITC, global fitting, cooperativity, binding model, ligand
A High Throughput MHC II Binding Assay for Quantitative Analysis of Peptide Epitopes
Institutions: Dartmouth College, University of Rhode Island, Dartmouth College.
Biochemical assays with recombinant human MHC II molecules can provide rapid, quantitative insights into immunogenic epitope identification, deletion, or design1,2
. Here, a peptide-MHC II binding assay is scaled to 384-well format. The scaled down protocol reduces reagent costs by 75% and is higher throughput than previously described 96-well protocols1,3-5
. Specifically, the experimental design permits robust and reproducible analysis of up to 15 peptides against one MHC II allele per 384-well ELISA plate. Using a single liquid handling robot, this method allows one researcher to analyze approximately ninety test peptides in triplicate over a range of eight concentrations and four MHC II allele types in less than 48 hr. Others working in the fields of protein deimmunization or vaccine design and development may find the protocol to be useful in facilitating their own work. In particular, the step-by-step instructions and the visual format of JoVE should allow other users to quickly and easily establish this methodology in their own labs.
Biochemistry, Issue 85, Immunoassay, Protein Immunogenicity, MHC II, T cell epitope, High Throughput Screen, Deimmunization, Vaccine Design
Structure and Coordination Determination of Peptide-metal Complexes Using 1D and 2D 1H NMR
Institutions: The Hebrew University of Jerusalem, The Hebrew University of Jerusalem.
Copper (I) binding by metallochaperone transport proteins prevents copper oxidation and release of the toxic ions that may participate in harmful redox reactions. The Cu (I) complex of the peptide model of a Cu (I) binding metallochaperone protein, which includes the sequence MTCSGCSRPG (underlined is conserved), was determined in solution under inert conditions by NMR spectroscopy.
NMR is a widely accepted technique for the determination of solution structures of proteins and peptides. Due to difficulty in crystallization to provide single crystals suitable for X-ray crystallography, the NMR technique is extremely valuable, especially as it provides information on the solution state rather than the solid state. Herein we describe all steps that are required for full three-dimensional structure determinations by NMR. The protocol includes sample preparation in an NMR tube, 1D and 2D data collection and processing, peak assignment and integration, molecular mechanics calculations, and structure analysis. Importantly, the analysis was first conducted without any preset metal-ligand bonds, to assure a reliable structure determination in an unbiased manner.
Chemistry, Issue 82, solution structure determination, NMR, peptide models, copper-binding proteins, copper complexes
Analyzing Protein Dynamics Using Hydrogen Exchange Mass Spectrometry
Institutions: University of Heidelberg.
All cellular processes depend on the functionality of proteins. Although the functionality of a given protein is the direct consequence of its unique amino acid sequence, it is only realized by the folding of the polypeptide chain into a single defined three-dimensional arrangement or more commonly into an ensemble of interconverting conformations. Investigating the connection between protein conformation and its function is therefore essential for a complete understanding of how proteins are able to fulfill their great variety of tasks. One possibility to study conformational changes a protein undergoes while progressing through its functional cycle is hydrogen-1
H-exchange in combination with high-resolution mass spectrometry (HX-MS). HX-MS is a versatile and robust method that adds a new dimension to structural information obtained by e.g.
crystallography. It is used to study protein folding and unfolding, binding of small molecule ligands, protein-protein interactions, conformational changes linked to enzyme catalysis, and allostery. In addition, HX-MS is often used when the amount of protein is very limited or crystallization of the protein is not feasible. Here we provide a general protocol for studying protein dynamics with HX-MS and describe as an example how to reveal the interaction interface of two proteins in a complex.
Chemistry, Issue 81, Molecular Chaperones, mass spectrometers, Amino Acids, Peptides, Proteins, Enzymes, Coenzymes, Protein dynamics, conformational changes, allostery, protein folding, secondary structure, mass spectrometry
Characterization of Complex Systems Using the Design of Experiments Approach: Transient Protein Expression in Tobacco as a Case Study
Institutions: RWTH Aachen University, Fraunhofer Gesellschaft.
Plants provide multiple benefits for the production of biopharmaceuticals including low costs, scalability, and safety. Transient expression offers the additional advantage of short development and production times, but expression levels can vary significantly between batches thus giving rise to regulatory concerns in the context of good manufacturing practice. We used a design of experiments (DoE) approach to determine the impact of major factors such as regulatory elements in the expression construct, plant growth and development parameters, and the incubation conditions during expression, on the variability of expression between batches. We tested plants expressing a model anti-HIV monoclonal antibody (2G12) and a fluorescent marker protein (DsRed). We discuss the rationale for selecting certain properties of the model and identify its potential limitations. The general approach can easily be transferred to other problems because the principles of the model are broadly applicable: knowledge-based parameter selection, complexity reduction by splitting the initial problem into smaller modules, software-guided setup of optimal experiment combinations and step-wise design augmentation. Therefore, the methodology is not only useful for characterizing protein expression in plants but also for the investigation of other complex systems lacking a mechanistic description. The predictive equations describing the interconnectivity between parameters can be used to establish mechanistic models for other complex systems.
Bioengineering, Issue 83, design of experiments (DoE), transient protein expression, plant-derived biopharmaceuticals, promoter, 5'UTR, fluorescent reporter protein, model building, incubation conditions, monoclonal antibody
Protein Purification-free Method of Binding Affinity Determination by Microscale Thermophoresis
Institutions: National Cancer Institute, SAIC-Frederick, Inc., Georgetown University Medical Center, National Cancer Institute.
Quantitative characterization of protein interactions is essential in practically any field of life sciences, particularly drug discovery. Most of currently available methods of KD
determination require access to purified protein of interest, generation of which can be time-consuming and expensive. We have developed a protocol that allows for determination of binding affinity by microscale thermophoresis (MST) without purification of the target protein from cell lysates. The method involves overexpression of the GFP-fused protein and cell lysis in non-denaturing conditions. Application of the method to STAT3-GFP transiently expressed in HEK293 cells allowed to determine for the first time the affinity of the well-studied transcription factor to oligonucleotides with different sequences. The protocol is straightforward and can have a variety of application for studying interactions of proteins with small molecules, peptides, DNA, RNA, and proteins.
Molecular Biology, Issue 78, Biochemistry, Cellular Biology, Genetics, Chemistry, Pharmacology, Intracellular Signaling Peptides and Proteins, Proteins, protein-inhibitor interaction, KD, transcription factor, ligand binding, binding affinity, thermophoresis, fluorescence, microscopy
The Importance of Correct Protein Concentration for Kinetics and Affinity Determination in Structure-function Analysis
Institutions: GE Healthcare Bio-Sciences AB.
In this study, we explore the interaction between the bovine cysteine protease inhibitor cystatin B and a catalytically inactive form of papain (Fig. 1), a plant cysteine protease, by real-time label-free analysis using Biacore X100. Several cystatin B variants with point mutations in areas of interaction with papain, are produced. For each cystatin B variant we determine its specific binding concentration using calibration-free concentration analysis (CFCA) and compare the values obtained with total protein concentration as determined by A280
. After that, the kinetics of each cystatin B variant binding to papain is measured using single-cycle kinetics (SCK). We show that one of the four cystatin B variants we examine is only partially active for binding. This partial activity, revealed by CFCA, translates to a significant difference in the association rate constant (ka
) and affinity (KD
), compared to the values calculated using total protein concentration. Using CFCA in combination with kinetic analysis in a structure-function study contributes to obtaining reliable results, and helps to make the right interpretation of the interaction mechanism.
Cellular Biology, Issue 37, Protein interaction, Surface Plasmon Resonance, Biacore X100, CFCA, Cystatin B, Papain
Isothermal Titration Calorimetry for Measuring Macromolecule-Ligand Affinity
Institutions: University of Tennessee .
Isothermal titration calorimetry (ITC) is a useful tool for understanding the complete thermodynamic picture of a binding reaction. In biological sciences, macromolecular interactions are essential in understanding the machinery of the cell. Experimental conditions, such as buffer and temperature, can be tailored to the particular binding system being studied. However, careful planning is needed since certain ligand and macromolecule concentration ranges are necessary to obtain useful data. Concentrations of the macromolecule and ligand need to be accurately determined for reliable results. Care also needs to be taken when preparing the samples as impurities can significantly affect the experiment. When ITC experiments, along with controls, are performed properly, useful binding information, such as the stoichiometry, affinity and enthalpy, are obtained. By running additional experiments under different buffer or temperature conditions, more detailed information can be obtained about the system. A protocol for the basic setup of an ITC experiment is given.
Molecular Biology, Issue 55, Isothermal titration calorimetry, thermodynamics, binding affinity, enthalpy, entropy, free energy
Discovering Protein Interactions and Characterizing Protein Function Using HaloTag Technology
Institutions: Promega Corporation, MS Bioworks LLC.
Research in proteomics has exploded in recent years with advances in mass spectrometry capabilities that have led to the characterization of numerous proteomes, including those from viruses, bacteria, and yeast. In comparison, analysis of the human proteome lags behind, partially due to the sheer number of proteins which must be studied, but also the complexity of networks and interactions these present. To specifically address the challenges of understanding the human proteome, we have developed HaloTag technology for protein isolation, particularly strong for isolation of multiprotein complexes and allowing more efficient capture of weak or transient interactions and/or proteins in low abundance. HaloTag is a genetically encoded protein fusion tag, designed for covalent, specific, and rapid immobilization or labelling of proteins with various ligands. Leveraging these properties, numerous applications for mammalian cells were developed to characterize protein function and here we present methodologies including: protein pull-downs used for discovery of novel interactions or functional assays, and cellular localization. We find significant advantages in the speed, specificity, and covalent capture of fusion proteins to surfaces for proteomic analysis as compared to other traditional non-covalent approaches. We demonstrate these and the broad utility of the technology using two important epigenetic proteins as examples, the human bromodomain protein BRD4, and histone deacetylase HDAC1. These examples demonstrate the power of this technology in enabling the discovery of novel interactions and characterizing cellular localization in eukaryotes, which will together further understanding of human functional proteomics.
Cellular Biology, Issue 89, proteomics, HaloTag, protein interactions, mass spectrometry, bromodomain proteins, BRD4, histone deacetylase (HDAC), HDAC cellular assays, and confocal imaging
Actin Co-Sedimentation Assay; for the Analysis of Protein Binding to F-Actin
Institutions: University of California, San Francisco - UCSF.
The actin cytoskeleton within the cell is a network of actin filaments that allows the movement of cells and cellular processes, and that generates tension and helps maintains cellular shape. Although the actin cytoskeleton is a rigid structure, it is a dynamic structure that is constantly remodeling. A number of proteins can bind to the actin cytoskeleton. The binding of a particular protein to F-actin is often desired to support cell biological observations or to further understand dynamic processes due to remodeling of the actin cytoskeleton. The actin co-sedimentation assay is an in vitro assay routinely used to analyze the binding of specific proteins or protein domains with F-actin. The basic principles of the assay involve an incubation of the protein of interest (full length or domain of) with F-actin, ultracentrifugation step to pellet F-actin and analysis of the protein co-sedimenting with F-actin. Actin co-sedimentation assays can be designed accordingly to measure actin binding affinities and in competition assays.
Biochemistry, Issue 13, F-actin, protein, in vitro binding, ultracentrifugation