Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “protein interaction”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Phycobilisome protein ApcG interacts with PSII and regulates energy transfer in Synechocystis

Photosynthetic organisms harvest light using pigment–protein complexes. In cyanobacteria, these are water-soluble antennae known as phycobilisomes (PBSs). The light absorbed by PBS is transferred to the photosystems in the thylakoid membrane to drive photosynthesis. The energy transfer between these complexes implies that protein–protein interactions allow the association of PBS with the photosystems. However, the specific proteins involved in the interaction of PBS with the photosystems are not fully characterized. Here, we show in Synechocystis sp. PCC 6803 that the recently discovered PBS linker protein ApcG (sll1873) interacts specifically with PSII through its N-terminal region. Growth of cyanobacteria is impaired in apcG deletion strains under light-limiting conditions. Furthermore, complementation of these strains using a phospho-mimicking version of ApcG causes reduced growth under normal growth conditions. Interestingly, the interaction of ApcG with PSII is affected when a phospho-mimicking version of ApcG is used, targeting the positively charged residues interacting with the thylakoid membrane, suggesting a regulatory role mediated by phosphorylation of ApcG. Low-temperature fluorescence measurements showed decreased PSI fluorescence in apcG deletion and complementation strains. The PSI fluorescence was the lowest in the phospho-mimicking complementation strain, while the pull-down experiment showed no interaction of ApcG with PSI under any tested condition. In conclusion, our results highlight the importance of ApcG for selectively directing energy harvested by the PBS and imply that the phosphorylation status of ApcG plays a role in regulating energy transfer from PSII to PSI.

59 BASIC BIOLOGICAL SCIENCES↗

Examining polymer‐protein biophysical interactions with small‐angle x‐ray scattering and quartz crystal microbalance with dissipation

Abstract Polymer‐protein hybrids can be deployed to improve protein solubility and stability in denaturing environments. While previous work used robotics and active machine learning to inform new designs, further biophysical information is required to ascertain structure–function behavior. Here, we show the value of tandem small‐angle x‐ray scattering (SAXS) and quartz crystal microbalance with dissipation (QCMD) experiments to reveal detailed polymer‐protein interactions with horseradish peroxidase (HRP) as a test case. Of particular interest was the process of polymer‐protein complex formation under thermal stress whereby SAXS monitors formation in solution while QCMD follows these dynamics at an interface. The radius of gyration ( R g ) of the protein as measured by SAXS does not change significantly in the presence of polymer under denaturing conditions, but thickness and dissipation changes were observed in QCMD data. SAXS data with and without thermal stress were utilized to create bead models of the potential complexes and denatured enzyme, and each model fit provided insight into the degree of interactions. Additionally, QCMD data demonstrated that HRP deforms by spreading upon surface adsorption at low concentration as shown by longer adsorption times and smaller frequency shifts. In contrast, thermally stressed and highly inactive HRP had faster adsorption kinetics. The combination of SAXS and QCMD serves as a framework for biophysical characterization of interactions between proteins and polymers which could be useful in designing polymer‐protein hybrids.

60 APPLIED LIFE SCIENCES↗

Deep Learning Prediction of Protein Complex Structures

Proteins interact to form protein complex to carry out biological functions such as catalytic chemical reaction. Therefore, it is important to develop computational methods to predict protein-protein interaction and the structures of protein complexes to study and enhance protein function. In this project, we successfully developed several deep learning methods to predict inter-protein contacts and the reinforcement learning and optimization methods to reconstruct protein complex structures from predicted inter-chain contacts. The methods were integrated with the MULTICOM protein complex structure prediction system and applied to predict the complex structures of biomass production-related proteins of green algae. During the two and a half years of research and development, all the specific milestones of the project were achieved successfully. 16 publications/manuscripts were produced. 10 software tools were developed. A patent application was submitted. Our MULTICOM predictors leveraging some tools developed in this project were ranked among the top predictors in the 15th Critical Assessment of Techniques for Protein Structure Prediction (CASP15) in 2022.

59 BASIC BIOLOGICAL SCIENCES↗

Improving Protein–Ligand Interaction Modeling with cryo-EM Data, Templates, and Deep Learning in 2021 Ligand Model Challenge

Elucidating protein–ligand interaction is crucial for studying the function of proteins and compounds in an organism and critical for drug discovery and design. The problem of protein–ligand interaction is traditionally tackled by molecular docking and simulation, which is based on physical forces and statistical potentials and cannot effectively leverage cryo-EM data and existing protein structural information in the protein–ligand modeling process. In this work, we developed a deep learning bioinformatics pipeline (DeepProLigand) to predict protein–ligand interactions from cryo-EM density maps of proteins and ligands. DeepProLigand first uses a deep learning method to predict the structure of proteins from cryo-EM maps, which is averaged with a reference (template) structure of the proteins to produce a combined structure to add ligands. The ligands are then identified and added into the structure to generate a protein–ligand complex structure, which is further refined. The method based on the deep learning prediction and template-based modeling was blindly tested in the 2021 EMDataResource Ligand Challenge and was ranked first in fitting ligands to cryo-EM density maps. These results demonstrate that the deep learning bioinformatics approach is a promising direction for modeling protein–ligand interactions on cryo-EM data using prior structural information.

59 BASIC BIOLOGICAL SCIENCES↗

Decoding the protein–ligand interactions using parallel graph neural networks

Abstract Protein–ligand interactions (PLIs) are essential for biochemical functionality and their identification is crucial for estimating biophysical properties for rational therapeutic design. Currently, experimental characterization of these properties is the most accurate method, however, this is very time-consuming and labor-intensive. A number of computational methods have been developed in this context but most of the existing PLI prediction heavily depends on 2D protein sequence data. Here, we present a novel parallel graph neural network (GNN) to integrate knowledge representation and reasoning for PLI prediction to perform deep learning guided by expert knowledge and informed by 3D structural data. We develop two distinct GNN architectures: $$\hbox {GNN}_{\mathrm{F}}$$ GNN F is the base implementation that employs distinct featurization to enhance domain-awareness, while $$\hbox {GNN}_{\mathrm{P}}$$ GNN P is a novel implementation that can predict with no prior knowledge of the intermolecular interactions. The comprehensive evaluation demonstrated that GNN can successfully capture the binary interactions between ligand and protein’s 3D structure with 0.979 test accuracy for $$\hbox {GNN}_{\mathrm{F}}$$ GNN F and 0.958 for $$\hbox {GNN}_{\mathrm{P}}$$ GNN P for predicting activity of a protein–ligand complex. These models are further adapted for regression tasks to predict experimental binding affinities and $$\hbox {pIC}_{\mathrm{50}}$$ pIC 50 crucial for compound’s potency and efficacy. We achieve a Pearson correlation coefficient of 0.66 and 0.65 on experimental affinity and 0.50 and 0.51 on $$\hbox {pIC}_{\mathrm{50}}$$ pIC 50 with $$\hbox {GNN}_{\mathrm{F}}$$ GNN F and $$\hbox {GNN}_{\mathrm{P}}$$ GNN P , respectively, outperforming similar 2D sequence based models. Our method can serve as an interpretable and explainable artificial intelligence (AI) tool for predicted activity, potency, and biophysical properties of lead candidates. To this end, we show the utility of $$\hbox {GNN}_{\mathrm{P}}$$ GNN P on SARS-Cov-2 protein targets by screening a large compound library and comparing the prediction with the experimentally measured data.

59 BASIC BIOLOGICAL SCIENCES↗

Energy-Screened Many-Body Expansion for Protein–Ligand Interactions: Examining Convergence for Metalloenzymes Through Seven–Body Interactions

Fragment-based quantum chemistry is a powerful strategy for calculating protein−ligand interaction energies using quantum chemistry methods. Rigorous convergence often requires hundreds of atoms in the protein binding-site model, especially if that model is constructed using distance-based criteria to select amino acid residues, while three- and four-body calculations exhibit instability related to combinatorial proliferation in the number of subsystem calculations. Here, we report an energy-based screening protocol for the many-body expansion applied to protein−ligand interactions, implemented in the open-source FRAGME∩T code. Using a combination of aggressive screening based on semiempirical quantum chemistry, with an improved graph-theoretical algorithm to eliminate unimportant subsystems, we are able to perform n-body calculations up to n = 7 using density functional theory in triple-ζ basis sets. Distance cutoffs further reduce the cost without compromising accuracy. Rapid and stable convergence of the many-body expansion is obtained by n = 4, for a pair of metalloenzymes in which a divalent ion coordinates directly to the ligand. As compared to previous results that relied solely on distance cutoffs, oscillations in the n-body corrections are reduced or eliminated, although residual errors remain in one case. This work demonstrates that benchmark-quality protein−ligand interaction energies can be systematically converged using a method with excellent parallel efficiency and scalability.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Identification of mosquito proteins that differentially interact with alphavirus nonstructural protein 3, a determinant of vector specificity

Chikungunya virus (CHIKV) and the closely related onyong-nyong virus (ONNV) are arthritogenic arboviruses that have caused significant, often debilitating, disease in millions of people. However, despite their kinship, they are vectored by different mosquito subfamilies that diverged 180 million years ago (anopheline versus culicine subfamilies). Previous work indicated that the nonstructural protein 3 (nsP3) of these alphaviruses was partially responsible for this vector specificity. To better understand the cellular components controlling alphavirus vector specificity, a cell culture model system of the anopheline restriction of CHIKV was developed along with a protein expression strategy. Mosquito proteins that differentially interacted with CHIKV nsP3 or ONNV nsP3 were identified. Six proteins were identified that specifically bound ONNV nsP3, ten that bound CHIKV nsP3 and eight that interacted with both. In addition to identifying novel factors that may play a role in virus/vector processing, these lists included host proteins that have been previously implicated as contributing to alphavirus replication.

Byers, Nathaniel M. (ORCID:0000000157725940)↗

Behavior of Water Near Multimodal Chromatography Ligands and Its Consequences for Modulating Protein–Ligand Interactions

Multimodal chromatography is a powerful approach for purifying proteins that uses ligands containing multiple modes of interaction. Recent studies have shown that selectivity in multimodal chromatographic separations is a function of the ligand structure and geometry. Here, we performed molecular dynamics simulations to explore how the ligand structure and geometry affect ligand–water interactions and how these differences in solution affect the nature of protein–ligand interactions. Our investigation focused on three chromatography ligands: Capto MMC, Nuvia cPrime, and Prototype 4, a structural variant of Nuvia cPrime. First, the solvation characteristics of each ligand were quantified via three metrics: average water density, fluctuations, and residence time. We then explored how solvation was perturbed when the ligand was bound to the protein surface and found that the probability of the phenyl ring dewetting followed the order: Capto MMC > Prototype 4 > Nuvia cPrime. To explore how these differences in dewetting affect protein–ligand interactions, we calculated the probability of each ligand binding to different types of residues on the protein surface and found that the probability of binding to a hydrophobic residue followed the same order as the dewetting behavior. This study illustrates the role that wetting and dewetting play in modulating protein–ligand interactions.

59 BASIC BIOLOGICAL SCIENCES↗

Protein-Nanoparticle Interactions Govern the Interfacial Behavior of Polymeric Nanogels: Study of Protein Corona Formation at the Air/Water Interface

Biomedical applications of nanoparticles require a fundamental understanding of their interactions and behavior with biological interfaces. Protein corona formation can alter the morphology and properties of nanomaterials, and knowledge of the interfacial behavior of the complexes, using in situ analytical techniques, will impact the development of nanocarriers to maximize uptake and permeability at cellular interfaces. In this study we evaluate the interactions of acrylamide-based nanogels, with neutral, positive, and negative charges, with serum-abundant proteins albumin, fibrinogen, and immunoglobulin G. The formation of a protein corona complex between positively charged nanoparticles and albumin is characterized by dynamic light scattering, circular dichroism, and surface tensiometry; we use neutron reflectometry to resolve the complex structure at the air/water interface and demonstrate the effect of increased protein concentration on the interface. Surface tensiometry data suggest that the structure of the proteins can impact the interfacial properties of the complex formed. These results contribute to the understanding of the factors that influence the bio-nano interface, which will help to design nanomaterials with improved properties for applications in drug delivery.

Traldi, Federico (ORCID:0000000161251012)↗

The gammaherpesviral TATA-box-binding protein directly interacts with the CTD of host RNA Pol II to direct late gene transcription

β- and γ-herpesviruses include the oncogenic human viruses Kaposi’s sarcoma-associated virus (KSHV) and Epstein-Barr virus (EBV), and human cytomegalovirus (HCMV), which is a significant cause of congenital disease. Near the end of their replication cycle, these viruses transcribe their late genes in a manner distinct from host transcription. Late gene transcription requires six virally encoded proteins, one of which is a functional mimic of host TATA-box-binding protein (TBP) that is also involved in recruitment of RNA polymerase II (Pol II) via unknown mechanisms. Here, we applied biochemical protein interaction studies together with electron microscopy-based imaging of a reconstituted human preinitiation complex to define the mechanism underlying Pol II recruitment. These data revealed that the herpesviral TBP, encoded by ORF24 in KSHV, makes a direct protein-protein contact with the C-terminal domain of host RNA polymerase II (Pol II), which is a unique feature that functionally distinguishes viral from cellular TBP. The interaction is mediated by the N-terminal domain (NTD) of ORF24 through a conserved motif that is shared in its β- and γ-herpesvirus homologs. Thus, these herpesviruses employ an unprecedented strategy in eukaryotic transcription, wherein promoter recognition and polymerase recruitment are facilitated by a single transcriptional activator with functionally distinct domains.

59 BASIC BIOLOGICAL SCIENCES↗

LDIP cooperates with SEIPIN and LDAP to facilitate lipid droplet biogenesis in Arabidopsis

Cytoplasmic lipid droplets (LDs) are evolutionarily conserved organelles that store neutral lipids and play critical roles in plant growth, development, and stress responses. However, the molecular mechanisms underlying their biogenesis at the endoplasmic reticulum (ER) remain obscure. Here we show that a recently identified protein termed LD-associated protein [LDAP]-interacting protein (LDIP) works together with both endoplasmic reticulum-localized SEIPIN and the LD-coat protein LDAP to facilitate LD formation in Arabidopsis thaliana. Heterologous expression in insect cells demonstrated that LDAP is required for the targeting of LDIP to the LD surface, and both proteins are required for the production of normal numbers and sizes of LDs in plant cells. LDIP also interacts with SEIPIN via a conserved hydrophobic helix in SEIPIN and LDIP functions together with SEIPIN to modulate LD numbers and sizes in plants. Further, the co-expression of both proteins is required to restore normal LD production in SEIPIN-deficient yeast cells. These data, combined with the analogous function of LDIP to a mammalian protein called LD Assembly Factor 1, are discussed in the context of a new model for LD biogenesis in plant cells with evolutionary connections to LD biogenesis in other eukaryotes.

59 BASIC BIOLOGICAL SCIENCES↗

Artificial intelligence in the prediction of protein–ligand interactions: recent advances and future directions

Abstract New drug production, from target identification to marketing approval, takes over 12 years and can cost around $2.6 billion. Furthermore, the COVID-19 pandemic has unveiled the urgent need for more powerful computational methods for drug discovery. Here, we review the computational approaches to predicting protein–ligand interactions in the context of drug discovery, focusing on methods using artificial intelligence (AI). We begin with a brief introduction to proteins (targets), ligands (e.g. drugs) and their interactions for nonexperts. Next, we review databases that are commonly used in the domain of protein–ligand interactions. Finally, we survey and analyze the machine learning (ML) approaches implemented to predict protein–ligand binding sites, ligand-binding affinity and binding pose (conformation) including both classical ML algorithms and recent deep learning methods. After exploring the correlation between these three aspects of protein–ligand interaction, it has been proposed that they should be studied in unison. We anticipate that our review will aid exploration and development of more accurate ML-based prediction strategies for studying protein–ligand interactions.

59 BASIC BIOLOGICAL SCIENCES↗

Engineering an efficient and bright split Corynactis californica green fluorescent protein

Split green fluorescent protein (GFP) has been used in a panoply of cellular biology applications to study protein translocation, monitor protein solubility and aggregation, detect protein–protein interactions, enhance protein crystallization, and even map neuron contacts. Recent work shows the utility of split fluorescent proteins for large scale labeling of proteins in cells using CRISPR, but sets of efficient split fluorescent proteins that do not cross-react are needed for multiplexing experiments. We present a new monomeric split green fluorescent protein (ccGFP) engineered from a tetrameric GFP found in Corynactis californica, a bright red colonial anthozoan similar to sea anemones and scleractinian stony corals. Split ccGFP from C. californica complements up to threefold faster compared to the original Aequorea victoria split GFP and enable multiplexed labeling with existing A. victoria split YFP and CFP.

59 BASIC BIOLOGICAL SCIENCES↗

Microscale Thermophoresis (MST) as a Tool to Study Binding Interactions of Oxygen-Sensitive Biohybrids

Microscale thermophoresis (MST) is a technique used to measure the strength of molecular interactions. MST is a thermophoretic-based technique that monitors the change in fluorescence associated with the movement of fluorescent-labeled molecules in response to a temperature gradient triggered by an IR LASER. MST has advantages over other approaches for examining molecular interactions, such as isothermal titration calorimetry, nuclear magnetic resonance, biolayer interferometry, and surface plasmon resonance, requiring a small sample size that does not need to be immobilized and a high-sensitivity fluorescence detection. In addition, since the approach involves the loading of samples into capillaries that can be easily sealed, it can be adapted to analyze oxygen-sensitive samples. In this Bio-protocol, we describe the troubleshooting and optimization we have done to enable the use of MST to examine protein–protein interactions, protein–ligand interactions, and protein–nanocrystal interactions. The salient elements in the developed procedures include 1) loading and sealing capabilities in an anaerobic chamber for analysis using a NanoTemper MST located on the benchtop in air, 2) identification of the optimal reducing agents compatible with data acquisition with effective protection against trace oxygen, and 3) the optimization of data acquisition and analysis procedures. The procedures lay the groundwork to define the determinants of molecular interactions in these technically demanding systems.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Revealing functional insights into ER proteostasis through proteomics and interactomics

The endoplasmic reticulum (ER), responsible for processing approximately one-third of the human proteome including most secreted and membrane proteins, plays a pivotal role in protein homeostasis (proteostasis). Dysregulation of ER proteostasis has been implicated in a number of disease states. As such, continued efforts are directed at elucidating mechanisms of ER protein quality control which are mediated by transient and dynamic protein-protein interactions with molecular chaperones, co-chaperones, protein folding and trafficking factors that take place in and around the ER. Technological advances in mass spectrometry have played a pivotal role in characterizing and understanding these protein-protein interactions that dictate protein quality control mechanisms. Here, we highlight the recent progress from mass spectrometry-based investigation of ER protein quality control in revealing the topological arrangement of the proteostasis network, stress response mechanisms that adjust the ER proteostasis capacity, and disease specific changes in proteostasis network engagement. We close by providing a brief outlook on underexplored areas of ER proteostasis where mass spectrometry is a tool uniquely primed to further expand our understanding of the regulation and coordination of protein quality control processes in diverse diseases.

60 APPLIED LIFE SCIENCES↗

Protein–chromophore interactions controlling photoisomerization in red/green cyanobacteriochromes

Abstract Photoreceptors in the phytochrome superfamily use 15,16-photoisomerization of a linear tetrapyrrole (bilin) chromophore to photoconvert between two states with distinct spectral and biochemical properties. Canonical phytochromes include master regulators of plant growth and development in which light signals trigger interconversion between a red-absorbing 15Z dark-adapted state and a metastable, far-red-absorbing 15E photoproduct state. Distantly related cyanobacteriochromes (CBCRs) carry out a diverse range of photoregulatory functions in cyanobacteria and exhibit considerable spectral diversity. One widespread CBCR subfamily typically exhibits a red-absorbing 15Z dark-adapted state similar to that of phytochrome that gives rise to a distinct green-absorbing 15E photoproduct. This red/green CBCR subfamily also includes red-inactive examples that fail to undergo photoconversion, providing an opportunity to study protein–chromophore interactions that either promote photoisomerization or block it. In this work, we identified a conserved lineage of red-inactive CBCRs. This enabled us to identify three substitutions sufficient to block photoisomerization in photoactive red/green CBCRs. The resulting red-inactive variants faithfully replicated the fluorescence and circular dichroism properties of naturally occurring examples. Converse substitutions restored photoconversion in naturally red-inactive CBCRs. This work thus identifies protein–chromophore interactions that control the fate of the excited-state population in red/green cyanobacteriochromes.

59 BASIC BIOLOGICAL SCIENCES↗