Engineering PapersSearch

SEARCH · Engineering Papers

Results for “protein design”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Computational design of potent and selective binders of BAK and BAX

Potent and selective binders of the key proapoptotic proteins BAK and BAX have not been described. We use computational protein design to generate high affinity binders of BAK and BAX with greater than 100-fold specificity for their target. Both binders activate their targets when at low concentration, driving pore formation, but inhibit membrane permeabilization when in excess. Crystallography shows that the BAK binder induces BAK unfolding, exposing the α6 helix and BH3 domain. Together, these data suggest that upon binding, BAK or BAX unfold; at high binder concentrations, self-association of the partially folded BAK or BAX proteins is blocked and the membrane remains intact, whereas at low concentrations, dimers form, and the membrane ruptures. Our designed binders modulate apoptosis via direct, specific interactions with BAK and BAX and reveal that for therapeutic strategies targeting BAK and BAX, inhibition requires saturating binder concentrations at the site of action.

Berger, Stephanie

LDBT instead of DBTL: combining machine learning and rapid cell-free testing

Synthetic biology is defined by Design-Build-Test-Learn cycles. Machine learning has yielded significant improvements in protein design; thus, we propose that “Learning” can precede “Design” to optimize engineering workflows. Additionally, shifting to cell-free platforms will further accelerate “Build” and “Test” capabilities by allowing the rapid expression of proteins and functional testing without the need for cellular transformation or isolation.

59 BASIC BIOLOGICAL SCIENCES

NCAP: Noncanonical Amino Acid Parameterization Software for CHARMM Potentials

Noncanonical Amino Acids (NCAAs) provide numerous avenues for introduction of novel functionality to peptides and proteins. NCAAs can be incorporated through solid phase synthesis or genetic code expansion in conjugation with heterologous expression of the encoded protein modification. Due to the difficulty of synthesis, wide chemical space and lack of empirically resolved structures modeling the effects of NCAA mutation is critical for rational protein design. To evaluate the structural and functional perturbations NCAAs introduce we utilize molecular potentials that describe the forces in protein structure. Most potentials such as CHARMM are designed to model canonical residues but can be parameterized in include novel NCAAs. Here, in this work, we introduce NCAP a software package to generate CHARMM compatible parameters from quantum chemical calculation. Unlike currently available tools NCAP is designed to recognize NCAA structure and automatically bridge the gap between DFT calculations and potential parameters. For our software we discuss workflow, validation against canonical parameter sets and comparison to published NCAA-protein structures.

59 BASIC BIOLOGICAL SCIENCES

Designing Peptide Fossils That Model the Evolution of the Bacterial Ferredoxin Fold

Electron transfer coupled to redox chemistry is at the heart of metabolism. The proteins responsible for moving electrons (protein electron carriers) must have emerged at the origin of life. The small iron–sulfur-binding bacterial ferredoxins were likely among these first proteins. Embedded within the ferredoxin sequence and structure is a symmetry that points to an ancient gene duplication event. Little is understood about the nature of ferredoxins prior to this duplication event or what environmental factors may have driven the selection for more complex forms. The deep-time molecular history of ferredoxins goes back billions of years and cannot be reconstructed by phylogenetic analyses based on amino acid sequences. Here, we use structure-guided protein design to model a fossil half-ferredoxin stage in the evolution of this fold, the semidoxins, and their symmetric full-length counterparts, the symdoxins. Semidoxin designs homodimerize, exhibiting structural, thermodynamic, and electrochemical behaviors in most cases identical to cognate symdoxins. However, the semi- and symdoxin fossil stages behave differently when incorporated into an in vivo electron transfer complementation assay. Both can support bacterial growth dependent on protein expression. Growth rates of bacteria expressing the semidoxins are much more sensitive to oxygen than those of bacteria expressing symdoxins. Motivated by the in vivo functionality of designed semidoxins, we identified putative naturally occurring semidoxins in extant anaerobic microorganisms. This is consistent with the observed in vivo oxygen sensitivity of the semidoxin designs. One natural semidoxin is shown to be folded and redox active. However, it exists as a mixture of monomers and dimers, suggesting a potential connection between semidoxins and even simpler single iron–sulfur cluster-binding peptides.

59 BASIC BIOLOGICAL SCIENCES

Flow matching meets biology and life science: a survey

Over the past decade, advances in generative modeling, such as generative adversarial networks, masked autoencoders, and diffusion models, have significantly transformed biological research and discovery, enabling breakthroughs in molecule design, protein generation, catalysis discovery, drug discovery, and beyond. At the same time, biological applications have served as valuable testbeds for evaluating the capabilities of generative models. Recently, flow matching has emerged as a powerful and efficient alternative to diffusion-based generative modeling, with growing interest in its application to problems in biology and life sciences. This paper presents the first comprehensive survey of recent developments in flow matching and its applications in biological domains. We begin by systematically reviewing the foundations and variants of flow matching, and then categorize its applications into three major areas: biological sequence modeling, molecule generation and design, and peptide and protein generation. For each, we provide an in-depth review of recent progress. We also summarize commonly used datasets and software tools, and conclude with a discussion of potential future directions.

59 BASIC BIOLOGICAL SCIENCES

Sequence and structural implications of a bovine corneal keratan sulfate proteoglycan core protein. Protein 37B represents bovine lumican and proteins 37A and 25 are unique

Amino acid sequence from tryptic peptides of three different bovine corneal keratan sulfate proteoglycan (KSPG) core proteins (designated 37A, 37B, and 25) showed similarities to the sequence of a chicken KSPG core protein lumican. Bovine lumican cDNA was isolated from a bovine corneal expression library by screening with chicken lumican cDNA. The bovine cDNA codes for a 342-amino acid protein, M(r) 38,712, containing amino acid sequences identified in the 37B KSPG core protein. The bovine lumican is 68% identical to chicken lumican, with an 83% identity excluding the N-terminal 40 amino acids. Location of 6 cysteine and 4 consensus N-glycosylation sites in the bovine sequence were identical to those in chicken lumican. Bovine lumican had about 50% identity to bovine fibromodulin and 20% identity to bovine decorin and biglycan. About two-thirds of the lumican protein consists of a series of 10 amino acid leucine-rich repeats that occur in regions of calculated high beta-hydrophobic moment, suggesting that the leucine-rich repeats contribute to beta-sheet formation in these proteins. Sequences obtained from 37A and 25 core proteins were absent in bovine lumican, thus predicting a unique primary structure and separate mRNA for each of the three bovine KSPG core proteins.

NASA Discipline Cell Biology

VPS26 Moonlights as a β-Arrestin-like Adapter for a 7-Transmembrane RGS Protein in Arabidopsis thaliana

Extracellular signals perceived by 7-transmembrane (7TM)-spanning receptors initiate desensitization that involves the removal of these receptors from the plasma membrane. Agonist binding often evokes phosphorylation in the flexible C-terminal region and/or intracellular loop 3 of many 7TM G-protein-coupled receptors in animal cells, which consequently recruits a cytoplasmic intermediate adaptor, β-arrestin, resulting in clathrin-mediated endocytosis (CME) and downstream signaling such as transcriptional changes. Some 7TM receptors undergo CME without recruiting β-arrestin, but it is not clear how. Arrestins are not encoded in the Arabidopsis thaliana genome, yet Arabidopsis cells have a well-characterized signal-induced CME of a 7TM protein, designated Regulator of G Signaling 1 (AtRGS1). Here we show that a component of the retromer complex, Vacuolar Protein Sorting-Associated 26 (VPS26), binds the phosphorylated C-terminal region of AtRGS1 as a VPS26A/B heterodimer to form a complex that is required for downstream signaling. We propose that VPS26 moonlights as an arrestin-like adaptor in the CME of AtRGS1.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Artificial intelligence methods for protein structure and interaction prediction: Recent advances and challenges

Recent advances in artificial intelligence have introduced novel methods for high-accuracy prediction of protein tertiary structures, protein complex structures, and interactions between proteins and other biomolecules, such as small molecules and nucleic acids. Such advancements are accelerating biomedical research and the development of new protein design and bioengineering methods among many other important biotechnology applications. Here, in this review, we outline the recent advances in protein-centric biomolecular structure and interaction prediction, highlight some major challenges in the field, and discuss potential directions to address them.

Morehead, Alex [Lawrence Berkeley National Laborat

Proteins with Novel Structure, Function and Dynamics

Recently, a small enzyme that ligates two RNA fragments with the rate of 10(exp 6) above background was evolved in vitro (Seelig and Szostak, Nature 448:828‐831, 2007). This enzyme does not resemble any contemporary protein (Chao et al., Nature Chem. Biol. 9:81‐83, 2013). It consists of a dynamic, catalytic loop, a small, rigid core containing two zinc ions coordinated by neighboring amino acids, and two highly flexible tails that might be unimportant for protein function. In contrast to other proteins, this enzyme does not contain ordered secondary structure elements, such as alpha‐helix or beta‐sheet. The loop is kept together by just two interactions of a charged residue and a histidine with a zinc ion, which they coordinate on the opposite side of the loop. Such structure appears to be very fragile. Surprisingly, computer simulations indicate otherwise. As the coordinating, charged residue is mutated to alanine, another, nearby charged residue takes its place, thus keeping the structure nearly intact. If this residue is also substituted by alanine a salt bridge involving two other, charged residues on the opposite sides of the loop keeps the loop in place. These adjustments are facilitated by high flexibility of the protein. Computational predictions have been confirmed experimentally, as both mutants retain full activity and overall structure. These results challenge our notions about what is required for protein activity and about the relationship between protein dynamics, stability and robustness. We hypothesize that small, highly dynamic proteins could be both active and fault tolerant in ways that many other proteins are not, i.e. they can adjust to retain their structure and activity even if subjected to mutations in structurally critical regions. This opens the doors for designing proteins with novel functions, structures and dynamics that have not been yet considered.

Proteins

Isolation, characterization, and amino acid sequences of auracyanins, blue copper proteins from the green photosynthetic bacterium Chloroflexus aurantiacus

Three small blue copper proteins designated auracyanin A, auracyanin B-1, and auracyanin B-2 have been isolated from the thermophilic green gliding photosynthetic bacterium Chloroflexus aurantiacus. All three auracyanins are peripheral membrane proteins. Auracyanin A was described previously (Trost, J. T., McManus, J. D., Freeman, J. C., Ramakrishna, B. L., and Blankenship, R. E. (1988) Biochemistry 27, 7858-7863) and is not glycosylated. The two B forms are glycoproteins and have almost identical properties to each other, but are distinct from the A form. The sodium dodecyl sulfate-polyacrylamide gel electrophoresis apparent monomer molecular masses are 14 (A), 18 (B-2), and 22 (B-1) kDa. The amino acid sequences of the B forms are presented. All three proteins have similar absorbance, circular dichroism, and resonance Raman spectra, but the electron spin resonance signals are quite different. Laser flash photolysis kinetic analysis of the reactions of the three forms of auracyanin with lumiflavin and flavin mononucleotide semiquinones indicates that the site of electron transfer is negatively charged and has an accessibility similar to that found in other blue copper proteins. Copper analysis indicates that all three proteins contain 1 mol of copper per mol of protein. All three auracyanins exhibit a midpoint redox potential of +240 mV. Light-induced absorbance changes and electron spin resonance signals suggest that auracyanin A may play a role in photosynthetic electron transfer. Kinetic data indicate that all three proteins can donate electrons to cytochrome c-554, the electron donor to the photosynthetic reaction center.

NASA Discipline Exobiology

Nanomolar Sensitivity Chirality Transfer from Designed Helical Repeat Proteins to Achiral CdS Nanorods

Bridging chirality across length scales with inorganic− organic hybrid materials is a rapidly expanding area of research. Here, we establish asymmetry at CdS nanorod (NR) interfaces using a designed helical repeat protein bearing four cysteine residues (DHR- 4Cys). Hydrophobic NRs are transferred into water with glycine, and then glycine is displaced by DHR-4Cys, leveraging the thiophilicity of cadmium. Circular dichroism (CD) in the visible, coincident with CdS electronic transitions, reveals a chiral DHR-4Cys:CdS interface. Here, the dissymmetry factor [g-factor = 4.5 × 10 −4 (short NRs) and 5.0 × 10 −4 (long NRs)] is weakly dependent on the NR length, and CD persists at nanomolar protein loadings. Additionally, control experiments demonstrate that DHR-4Cys:CdS NR chirality is dictated by the local coordination of Cys with no significant contribution from the chiral secondary structure of the protein (g-factors of short and long Cys:CdS NRs are 4.8 × 10 −4 and 4.0 × 10 −4 , respectively). Together with far-UV CD and transmission electron microscopy, which provide evidence of preserved protein structure, these results provide the first demonstration that a structurally defined protein can induce chirality in CdS nanocrystals while maintaining protein structure at biologically relevant concentrations.

Cadmium sulfide

Architector 2.0: Expanded Capabilities for Metal Complex Engineering

Automated three-dimensional molecular construction from two-dimensional graph representations is critical to high-throughput discovery eIorts. Software capabilities in this area have accelerated research across fields ranging from protein design and drug discovery to transition metal catalyst development. When Architector was first introduced, it uniquely enabled high-throughput, chemically relevant three-dimensional construction of f-element complexes. Since its introduction, Architector has been applied in large-scale computational campaigns, targeted studies in critical mineral extraction, and artificial intelligence-driven discovery eIorts.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Size and Structure of the Sequence Space of Repeat Proteins

The coding space of protein sequences is shaped by evolutionary constraints set by requirements of function and stability. We show that the coding space of a given protein family— the total number of sequences in that family—can be estimated using models of maximum entropy trained on multiple sequence alignments of naturally occurring amino acid sequences. We analyzed and calculated the size of three abundant repeat proteins families, whose members are large proteins made of many repetitions of conserved portions of *30 amino acids. While amino acid conservation at each position of the alignment explains most of the reduction of diversity relative to completely random sequences, we found that correlations between amino acid usage at different positions significantly impact that diversity. We quantified the impact of different types of correlations, functional and evolutionary, on sequence diversity. Analysis of the detailed structure of the coding space of the families revealed a rugged landscape, with many local energy minima of varying sizes with a hierarchical structure, reminiscent of frustrated energy landscapes of spin glass in physics. This clustered structure indicates a multiplicity of subtypes within each family and suggests new strategies for protein design.

Jacopo Marchi

Investigation of design principles for metal-binding and conductive protein assemblies

Throughout the lifetime of this initiative, including renewals, we focused on understanding the fundamental principles of protein-protein interface design that enable predictable and modular spatial and kinetic control of multi-component protein self-assembly in 1D, 2D, and 3D, including the interface with inorganic materials, small molecules, and metal ions. We designed individual protein components that bind specific metal ions, including REEs and transport ions across lipid membranes. We created helical 1D filaments of repeating units with programmed periodicity, pitch, and multi-component environmentally responsive self-assembling protein fibers. We showed that these filaments reversibly assemble and disassemble under specific pH conditions and created end-specific caps that independently tune the balance of attachment and detachment rates at each terminus of the filament. Using similar filaments, we succeeded in binding arrays of heme and chlorophyll molecules and assembling patterned helical coatings around carbon nanotubes in efforts to create de novo conductive nanowires. By arraying REE binding sites in a large circular tandem array with a repeat protein-based cyclic oligomer, we created a molecular scaffold for superradiance and paramagnetic quantum sensing. We created a range of one-component and two-component self-assembling 2D arrays and showed that when designed to engage cell receptors, these arrays can control cell behavior from outside the cell signal to inside the cell. We designed helical repeat proteins with variable lengths displaying charged residues in a pattern matched to the cation lattice of mica. achieved a range of ordered states with an epitaxial match to the underlying crystal lattice. We further applied the learned principles of protein-induced biomineralization to design proteins with an interface lattice matching CaCO 3 and guide the formation of specific crystal forms of CaCO 3 from solution, a significant advance toward the global need to manage carbon. In all cases of mineral lattice matching and biomineralization, we followed assembly using molecularly resolved in situ AFM imaging and extracted information about assembly pathways and energetics, applying deep learning to quantify the dynamics of protein self-organization. We developed techniques for using dynamic metal-dependent interfaces on protein nanopores for discriminatively sensing dilute REEs in solution and demonstrated the use of strong metal-binding interfaces to drive nanocage disassembly for conditional nanocompartmentalization applications. This grant supported 11 people, including Asim Bera, Evans Brackenbrough, Andrew Borst, Nikita Hanikel, Timothy Huddy, Emily Joyce, Alex Young-Seug Kang, Ryan Kibler, Joshua Morris Lubner, Harley Pyles, and Shuai Zhang. The research effort culminated in the production of published papers and theses. Electronic Thesis/Dissertation are distributed by ProQuest/UMI Dissertation Publishing and made available on an open access basis through UW Libraries ResearchWorks Service.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Design of intrinsically disordered protein variants with diverse structural properties

Intrinsically disordered proteins (IDPs) perform a broad range of functions in biology, suggesting that the ability to design IDPs could help expand the repertoire of proteins with novel functions. Computational design of IDPs with specific conformational properties has, however, been difficult because of their substantial dynamics and structural complexity. We describe a general algorithm for designing IDPs with specific structural properties. We demonstrate the power of the algorithm by generating variants of naturally occurring IDPs that differ in compaction, long-range contacts, and propensity to phase separate. We experimentally tested and validated our designs and analyzed the sequence features that determine conformations. We show how our results are captured by a machine learning model, enabling us to speed up the algorithm. Our work expands the toolbox for computational protein design and will facilitate the design of proteins whose functions exploit the many properties afforded by protein disorder.

Science & Technology - Other Topics

Science Issues Associated with the Use of a Microfluidic Chip Designed Specifically for Protein Crystallization

The Iterative Biological Crystallization team in partnership with Caliper Technologies has produced a prototype microfluidic chip for batch crystallization that has been designed and tested. The chip is designed for the mixing and dispensing of up to five solutions with possible variation of the recipe being delivered to two growth wells. Developments that have led to the successful on-chip crystallization of a few model proteins have required investigative insight into many different areas, including fluid mixing dynamics, surface treatments, quantification and fidelity of reagent delivery. This presentation will encompass the ongoing studies and data accumulated toward these efforts.

Holmes, Anna M.