Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “structured RNA”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Room-temperature structural studies of SARS-CoV-2 protein NendoU with an X-ray free-electron laser

NendoU from SARS-CoV-2 is responsible for the virus’s ability to evade the innate immune system by cleaving the polyuridine leader sequence of antisense viral RNA. Here we report the room-temperature structure of NendoU, solved by serial femtosecond crystallography at an X-ray free-electron laser to 2.6 Å resolution. The room-temperature structure provides insight into the flexibility, dynamics, and other intrinsic properties of NendoU, with indications that the enzyme functions as an allosteric switch. Functional studies examining cleavage specificity in solution and in crystals support the uridine-purine cleavage preference, and we demonstrate that enzyme activity is fully maintained in crystal form. Optimizing the purification of NendoU and identifying suitable crystallization conditions set the benchmark for future time-resolved serial femtosecond crystallography studies. This could advance the design of antivirals with higher efficacy in treating coronaviral infections, since drugs that block allosteric conformational changes are less prone to drug resistance.

59 BASIC BIOLOGICAL SCIENCES↗

Sarecycline interferes with tRNA accommodation and tethers mRNA to the 70S ribosome

Sarecycline is a new narrow-spectrum tetracycline-class antibiotic approved for the treatment of acne vulgaris. Tetracyclines share a common four-ring naphthacene core and inhibit protein synthesis by interacting with the 70S bacterial ribosome. Sarecycline is distinguished chemically from other tetracyclines because it has a 7-[[methoxy(methyl)amino]methyl] group attached at the C7 position of ring D. To investigate the functional role of this C7 moiety, we determined the X-ray crystal structure of sarecycline bound to the Thermus thermophilus 70S ribosome. Our 2.8-Å resolution structure revealed that sarecycline binds at the canonical tetracycline binding site located in the decoding center of the small ribosomal subunit. Importantly, unlike other tetracyclines, the unique C7 extension of sarecycline extends into the messenger RNA (mRNA) channel to form a direct interaction with the A-site codon to possibly interfere with mRNA movement through the channel and/or disrupt A-site codon–anticodon interaction. Based on our biochemical studies, sarecycline appears to be a more potent initiation inhibitor compared to other tetracyclines, possibly due to drug interactions with the mRNA, thereby blocking accommodation of the first aminoacyl transfer RNA (tRNA) into the A site. Overall, our structural and biochemical findings rationalize the role of the unique C7 moiety of sarecycline in antibiotic action.

59 BASIC BIOLOGICAL SCIENCES↗

P finder: genomic and metagenomic annotation of RNase P RNA gene (rnpB)

Abstract Background The rnpB gene encodes for an essential catalytic RNA (RNase P). Like other essential RNAs, RNase P’s sequence is highly variable. However, unlike other essential RNAs (i.e. tRNA, 16 S, 6 S,...) its structure is also variable with at least 5 distinct structure types observed in prokaryotes. This structural variability makes it labor intensive and challenging to create and maintain covariance models for the detection of RNase P RNA in genomic and metagenomic sequences. The lack of a facile and rapid annotation algorithm has led to the rnpB gene being the most grossly under annotated essential gene in completed prokaryotic genomes with only a 24% annotation rate. Here we describe the coupling of the largest RNase P RNA database with the local alignment scoring algorithm to create the most sensitive and rapid prokaryote rnpB gene identification and annotation algorithm to date. Results Of the 2772 completed microbial genomes downloaded from GenBank only 665 genomes had an annotated rnpB gene. We applied P Finder to these genomes and were able to identify 2733 or nearly 99% of the 2772 microbial genomes examined. From these results four new rnpB genes that encode the minimal T-type P RNase P RNAs were identified computationally for the first time. In addition, only the second C-type RNase P RNA was identified in Sphaerobacter thermophilus . Of special note, no RNase P RNAs were detected in several obligate endosymbionts of sap sucking insects suggesting a novel evolutionary adaptation. Conclusions The coupling of the largest RNase P RNA database and associated structure class identification with the P Finder algorithm is both sensitive and rapid, yielding high quality results to aid researchers annotating either genomic or metagenomic data. It is the only algorithm to date that can identify challenging RNAse P classes such as C-type and the minimal T-type RNase P RNAs. P Finder is written in C# and has a user-friendly GUI that can run on multiple 64-bit windows platforms (Windows Vista/7/8/10). P Finder is free available for download at https://github.com/JChristopherEllis/P-Finder as well as a small sample RNase P RNA file for testing.

59 BASIC BIOLOGICAL SCIENCES↗

Novel viruses of the family Partitiviridae discovered in Saccharomyces cerevisiae

It has been 49 years since the last discovery of a new virus family in the model yeast Saccharomyces cerevisiae . A large-scale screen to determine the diversity of double-stranded RNA (dsRNA) viruses in S . cerevisiae has identified multiple novel viruses from the family Partitiviridae that have been previously shown to infect plants, fungi, protozoans, and insects. Most S . cerevisiae partitiviruses (ScPVs) are associated with strains of yeasts isolated from coffee and cacao beans. The presence of partitiviruses was confirmed by sequencing the viral dsRNAs and purifying and visualizing isometric, non-enveloped viral particles. ScPVs have a typical bipartite genome encoding an RNA-dependent RNA polymerase (RdRP) and a coat protein (CP). Phylogenetic analysis of ScPVs identified three species of ScPV, which are most closely related to viruses of the genus Cryspovirus from the mammalian pathogenic protozoan Cryptosporidium parvum . Molecular modeling of the ScPV RdRP revealed a conserved tertiary structure and catalytic site organization when compared to the RdRPs of the Picornaviridae . The ScPV CP is the smallest so far identified in the Partitiviridae and has structural homology with the CP of other partitiviruses but likely lacks a protrusion domain that is a conspicuous feature of other partitivirus particles. ScPVs were stably maintained during laboratory growth and were successfully transferred to haploid progeny after sporulation, which provides future opportunities to study partitivirus-host interactions using the powerful genetic tools available for the model organism S . cerevisiae .

59 BASIC BIOLOGICAL SCIENCES↗

Equally parsimonious pathways through an RNA sequence space are not equally likely

An experimental system for determining the potential ability of sequences resembling 5S ribosomal RNA (rRNA) to perform as functional 5S rRNAs in vivo in the Escherichia coli cellular environment was devised previously. Presumably, the only 5S rRNA sequences that would have been fixed by ancestral populations are ones that were functionally valid, and hence the actual historical paths taken through RNA sequence space during 5S rRNA evolution would have most likely utilized valid sequences. Herein, we examine the potential validity of all sequence intermediates along alternative equally parsimonious trajectories through RNA sequence space which connect two pairs of sequences that had previously been shown to behave as valid 5S rRNAs in E. coli. The first trajectory requires a total of four changes. The 14 sequence intermediates provide 24 apparently equally parsimonious paths by which the transition could occur. The second trajectory involves three changes, six intermediate sequences, and six potentially equally parsimonious paths. In total, only eight of the 20 sequence intermediates were found to be clearly invalid. As a consequence of the position of these invalid intermediates in the sequence space, seven of the 30 possible paths consisted of exclusively valid sequences. In several cases, the apparent validity/invalidity of the intermediate sequences could not be anticipated on the basis of current knowledge of the 5S rRNA structure. This suggests that the interdependencies in RNA sequence space may be more complex than currently appreciated. If ancestral sequences predicted by parsimony are to be regarded as actual historical sequences, then the present results would suggest that they should also satisfy a validity requirement and that, in at least limited cases, this conjecture can be tested experimentally.

NASA Discipline Exobiology↗

Cryo-EM structure of the diapause chaperone artemin

The protein artemin acts as both an RNA and protein chaperone and constitutes over 10% of all protein in Artemia cysts during diapause. However, its mechanistic details remain elusive since no high-resolution structure of artemin exists. Here we report the full-length structure of artemin at 2.04 Å resolution. The cryo-EM map contains density for an intramolecular disulfide bond between Cys22-Cys61 and resolves the entire C-terminus extending into the core of the assembled protein cage but in a different configuration than previously hypothesized with molecular modeling. We also provide data supporting the role of C-terminal helix F towards stabilizing the dimer form that is believed to be important for its chaperoning activity. We were able to destabilize this effect by placing a tag at the C-terminus to fully pack the internal cavity and cause limited steric hindrance.

59 BASIC BIOLOGICAL SCIENCES↗

Structural mechanisms for binding and activation of a contact-quenched fluorophore by RhoBAST

The fluorescent light-up aptamer RhoBAST, which binds and activates the fluorophore–quencher conjugate tetramethylrhodamine-dinitroaniline with high affinity, super high brightness, remarkable photostability, and fast exchange kinetics, exhibits excellent performance in super-resolution RNA imaging. Here we determine the co-crystal structure of RhoBAST in complex with tetramethylrhodamine-dinitroaniline to elucidate the molecular basis for ligand binding and fluorescence activation. The structure exhibits an asymmetric “A”-like architecture for RhoBAST with a semi-open binding pocket harboring the xanthene of tetramethylrhodamine at the tip, while the dinitroaniline quencher stacks over the phenyl of tetramethylrhodamine instead of being fully released. Molecular dynamics simulations show highly heterogeneous conformational ensembles with the contact-but-unstacked fluorophore–quencher conformation for both free and bound tetramethylrhodamine-dinitroaniline being predominant. The simulations also show that, upon RNA binding, the fraction of xanthene-dinitroaniline stacked conformation significantly decreases in free tetramethylrhodamine-dinitroaniline. This highlights the importance of releasing dinitroaniline from xanthene tetramethylrhodamine to unquench the RhoBAST–tetramethylrhodamine-dinitroaniline complex. Using SAXS and ITC, we characterized the magnesium dependency of the folding and binding mode of RhoBAST in solution and indicated its strong structural robustness. The structures and binding modes of relevant fluorescent light-up aptamers are compared, providing mechanistic insights for rational design and optimization of this important fluorescent light-up aptamer-ligand system.

59 BASIC BIOLOGICAL SCIENCES↗

Magnesium ions mitigate metastable states in the regulatory landscape of mRNA elements

Residing in the 5' untranslated region of the mRNA, the 2'-deoxyguanosine (2'-dG) riboswitch mRNA element adopts an alternative structure upon binding of the 2'-dG molecule, which terminates transcription. RNA conformations are generally strongly affected by positively charged metal ions (especially Mg 2+ ). We have quantitatively explored the combined effect of ligand (2'-dG) and Mg 2+ binding on the energy landscape of the aptamer domain of the 2'-dG riboswitch with both explicit solvent all-atom molecular dynamics simulations (99 μsec aggregate sampling for the study) and selective 2'-hydroxyl acylation analyzed by primer extension (SHAPE) experiments. We show that both ligand and Mg 2+ are required for the stabilization of the aptamer domain; however, the two factors act with different modalities. The addition of Mg 2+ remodels the energy landscape and reduces its frustration by the formation of additional contacts. In contrast, the binding of 2'-dG eliminates the metastable states by nucleating a compact core for the aptamer domain. Mg 2+ ions and ligand binding are required to stabilize the least stable helix, P1 (which needs to unfold to activate the transcription platform), and the riboswitch core formed by the backbone of the P2 and P3 helices. Mg 2+ and ligand also facilitate a more compact structure in the three-way junction region.

59 BASIC BIOLOGICAL SCIENCES↗

Exploring Connectivity in Sequence Space of Functional RNA

Emergence of replicable genetic molecules was one of the marking points in the origin of life, evolution of which can be conceptualized as a walk through the space of all possible sequences. A theoretical concept of fitness landscape helps to understand evolutionary processes through assigning a value of fitness to each genotype. Then, evolution of a phenotype is viewed as a series of consecutive, single-point mutations. Natural selection biases evolution toward peaks of high fitness and away from valleys of low fitness. whereas neutral drift occurs in the sequence space without direction as mutations are introduced at random. Large networks of neutral or near-neutral mutations on a fitness landscape, especially for sufficiently long genomes, are possible or even inevitable. Their detection in experiments, however, has been elusive. Although a few near-neutral evolutionary pathways have been found, recent experimental evidence indicates landscapes consist of largely isolated islands. The generality of these results, however, is not clear, as the genome length or the fraction of functional molecules in the genotypic space might have been insufficient for the emergence of large, neutral networks. Thorough investigation on the structure of the fitness landscape is essential to understand the mechanisms of evolution of early genomes. RNA molecules are commonly assumed to play the pivotal role in the origin of genetic systems. They are widely believed to be early, if not the earliest, genetic and catalytic molecules, with abundant biochemical activities as aptamers and ribozymes, i.e. RNA molecules capable, respectively, to bind small molecules or catalyze chemical reactions. Here, we present results of our recent studies on the structure of the sequence space of RNA ligase ribozymes selected through in vitro evolution. Several hundred thousands of sequences active to a different degree were obtained by way of deep sequencing. Analysis of these sequences revealed several large clusters defined such that every sequence in a cluster can be reached from any other sequence in the same cluster through a series of single point mutations. Sequences in a single cluster appear to adopt more than one secondary structure. The mechanism of refolding within a single cluster was examined. To shed light on possible evolutionary paths in the space of ribozymes, the connectivity between clusters was investigated. The effect of length of RNA molecules on the structure of the fitness landscape and possible evolutionary paths was examined by way of comparing functional sequences of 20 and 80 nucleobases in length. It was found that sequences of different lengths shared secondary structure motifs that were presumed responsible for catalytic activity, with increasing complexity and global structural rearrangements emerging in longer molecules.

Wei, Chenyu↗

Template and target-site recognition by human LINE-1 in retrotransposition

The long interspersed element-1 (LINE-1, hereafter L1) retrotransposon has generated nearly one-third of the human genome and serves as an active source of genetic diversity and human disease. L1 spreads through a mechanism termed target-primed reverse transcription, in which the encoded enzyme (ORF2p) nicks the target DNA to prime reverse transcription of its own or non-self RNAs. Here we purified full-length L1 ORF2p and biochemically reconstituted robust target-primed reverse transcription with template RNA and target-site DNA. We report cryo-electron microscopy structures of the complete human L1 ORF2p bound to structured template RNAs and initiating cDNA synthesis. The template polyadenosine tract is recognized in a sequence-specific manner by five distinct domains. Among them, an RNA-binding domain bends the template backbone to allow engagement of an RNA hairpin stem with the L1 ORF2p C-terminal segment. Moreover, structure and biochemical reconstitutions demonstrate an unexpected target-site requirement: L1 ORF2p relies on upstream single-stranded DNA to position the adjacent duplex in the endonuclease active site for nicking of the longer DNA strand, with a single nick generating a staggered DNA break. Our research provides insights into the mechanism of ongoing transposition in the human genome and informs the engineering of retrotransposon proteins for gene therapy.

59 BASIC BIOLOGICAL SCIENCES↗

A DNA enzyme with N-glycosylase activity

In vitro evolution was used to develop a DNA enzyme that catalyzes the site-specific depurination of DNA with a catalytic rate enhancement of about 10(6)-fold. The reaction involves hydrolysis of the N-glycosidic bond of a particular deoxyguanosine residue, leading to DNA strand scission at the apurinic site. The DNA enzyme contains 93 nucleotides and is structurally complex. It has an absolute requirement for a divalent metal cation and exhibits optimal activity at about pH 5. The mechanism of the reaction was confirmed by analysis of the cleavage products by using HPLC and mass spectrometry. The isolation and characterization of an N-glycosylase DNA enzyme demonstrates that single-stranded DNA, like RNA and proteins, can form a complex tertiary structure and catalyze a difficult biochemical transformation. This DNA enzyme provides a new approach for the site-specific cleavage of DNA molecules.

Non-NASA Center↗

Proteins with Novel Structure, Function and Dynamics

Recently, a small enzyme that ligates two RNA fragments with the rate of 10(exp 6) above background was evolved in vitro (Seelig and Szostak, Nature 448:828‐831, 2007). This enzyme does not resemble any contemporary protein (Chao et al., Nature Chem. Biol. 9:81‐83, 2013). It consists of a dynamic, catalytic loop, a small, rigid core containing two zinc ions coordinated by neighboring amino acids, and two highly flexible tails that might be unimportant for protein function. In contrast to other proteins, this enzyme does not contain ordered secondary structure elements, such as alpha‐helix or beta‐sheet. The loop is kept together by just two interactions of a charged residue and a histidine with a zinc ion, which they coordinate on the opposite side of the loop. Such structure appears to be very fragile. Surprisingly, computer simulations indicate otherwise. As the coordinating, charged residue is mutated to alanine, another, nearby charged residue takes its place, thus keeping the structure nearly intact. If this residue is also substituted by alanine a salt bridge involving two other, charged residues on the opposite sides of the loop keeps the loop in place. These adjustments are facilitated by high flexibility of the protein. Computational predictions have been confirmed experimentally, as both mutants retain full activity and overall structure. These results challenge our notions about what is required for protein activity and about the relationship between protein dynamics, stability and robustness. We hypothesize that small, highly dynamic proteins could be both active and fault tolerant in ways that many other proteins are not, i.e. they can adjust to retain their structure and activity even if subjected to mutations in structurally critical regions. This opens the doors for designing proteins with novel functions, structures and dynamics that have not been yet considered.

Proteins↗

RolyPoly (rp) v0.1.0

The Rolypoly pipeline is designed to process raw RNA-seq data and identify potential RNA viral sequences. It is split into several self contained steps: 1. input data filtering and QC, 2. Genome assembly and refinement, 3. Assembly filtering, 4. Mapping to known RNA viral genomes, 5. Searching for RNA viral marker genes. 6. Genome functional and structural annotation. 6. Report preparation and potential downstream analysis The last module, may include taxonomic assignment, host range estimation, and phenotypic prediction. There are many similar software, but they focus on human related viruses, and lack the downstream applications or differ in their sensitivity. The initial user base are non-computational microbial ecologists who wish to better understand the potential RNA viruses in their own generated samples.

Neri, Uri↗

Universally Accessible Structural Data on Macromolecular Conformation, Assembly, and Dynamics by Small Angle X-Ray Scattering for DNA Repair Insights.

Structures provide a critical breakthrough step for biological analyses, and small angle X-ray scattering (SAXS) is a powerful structural technique to study dynamic DNA repair proteins. As toxic and mutagenic repair intermediates need to be prevented from inadvertently harming the cell, DNA repair proteins often chaperone these intermediates through dynamic conformations, coordinated assemblies, and allosteric regulation. By measuring structural conformations in solution for both proteins, DNA, RNA, and their complexes, SAXS provides insight into initial DNA damage recognition, mechanisms for validation of their substrate, and pathway regulation. Here, we describe exemplary SAXS analyses of a DNA damage response protein spanning from what can be derived directly from the data to obtaining super resolution through the use of SAXS selection of atomic models. We outline strategies and tactics for practical SAXS data collection and analysis. Making these structural experiments in reach of any basic and clinical researchers who have protein, SAXS data can readily be collected at government-funded synchrotrons, typically at no cost for academic researchers. In addition to discussing how SAXS complements and enhances cryo-electron microscopy, X-ray crystallography, NMR, and computational modeling, we furthermore discuss taking advantage of recent advances in protein structure prediction in combination with SAXS analysis.

Chinnam, Naga Babu↗

ADAR activation by inducing a syn conformation at guanosine adjacent to an editing site

Abstract ADARs (adenosine deaminases acting on RNA) can be directed to sites in the transcriptome by complementary guide strands allowing for the correction of disease-causing mutations at the RNA level. However, ADARs show bias against editing adenosines with a guanosine 5′ nearest neighbor (5′-GA sites), limiting the scope of this approach. Earlier studies suggested this effect arises from a clash in the RNA minor groove involving the 2-amino group of the guanosine adjacent to an editing site. Here we show that nucleosides capable of pairing with guanosine in a syn conformation enhance editing for 5′-GA sites. We describe the crystal structure of a fragment of human ADAR2 bound to RNA bearing a G:G pair adjacent to an editing site. The two guanosines form a Gsyn:Ganti pair solving the steric problem by flipping the 2-amino group of the guanosine adjacent to the editing site into the major groove. Also, duplexes with 2′-deoxyadenosine and 3-deaza-2′-deoxyadenosine displayed increased editing efficiency, suggesting the formation of a Gsyn:AH+anti pair. This was supported by X-ray crystallography of an ADAR complex with RNA bearing a G:3-deaza dA pair. This study shows how non-Watson–Crick pairing in duplex RNA can facilitate ADAR editing enabling the design of next generation guide strands for therapeutic RNA editing.

Doherty, Erin E.↗

Multi-omics analysis reveals the dynamic interplay between Vero host chromatin structure and function during vaccinia virus infection

The genome folds into complex configurations and structures thought to profoundly impact its function. The intricacies of this dynamic structure-function relationship are not well understood particularly in the context of viral infection. To unravel this interplay, here we provide a comprehensive investigation of simultaneous host chromatin structural (via Hi-C and ATAC-seq) and functional changes (via RNA-seq) in response to vaccinia virus infection. Over time, infection significantly impacts global and local chromatin structure by increasing long-range intra-chromosomal interactions and B compartmentalization and by decreasing chromatin accessibility and inter-chromosomal interactions. Local accessibility changes are independent of broad-scale chromatin compartment exchange (~12% of the genome), underscoring potential independent mechanisms for global and local chromatin reorganization. While infection structurally condenses the host genome, there is nearly equal bidirectional differential gene expression. Despite global weakening of intra-TAD interactions, functional changes including downregulated immunity genes are associated with alterations in local accessibility and loop domain restructuring. Therefore, chromatin accessibility and local structure profiling provide impactful predictions for host responses and may improve development of efficacious anti-viral counter measures including the optimization of vaccine design.

59 BASIC BIOLOGICAL SCIENCES↗

Methods for Identifying Ligands that Target Nucleic Acid Molecules and Nucleic Acid Structural Motifs

Disclosed are methods for identifying a nucleic acid (e.g., RNA, DNA, etc.) motif which interacts with a ligand. The method includes providing a plurality of ligands immobilized on a support, wherein each particular ligand is immobilized at a discrete location on the support; contacting the plurality of immobilized ligands with a nucleic acid motif library under conditions effective for one or more members of the nucleic acid motif library to bind with the immobilized ligands; and identifying members of the nucleic acid motif library that are bound to a particular immobilized ligand. Also disclosed are methods for selecting, from a plurality of candidate ligands, one or more ligands that have increased likelihood of binding to a nucleic acid molecule comprising a particular nucleic acid motif, as well as methods for identifying a nucleic acid which interacts with a ligand.

Disney, Matthew D.↗

RNA oligomers at atomic resolution containing 1‐methylpseudouridine, an essential building block of mRNA vaccines

Abstract All widely used mRNA vaccines against COVID‐19 contain in their sequence 1‐methylpseudouridine ( m1Ψ ) instead of uridine. In this publication, we report two high resolution crystal structures (at up to 1.01 and 1.32 Å, respectively) of one such double‐stranded 12‐mer RNA sequence crystallized in two crystal forms. The structures are compared with similar structures which do not contain this modification. Additionally, the X‐ray structure of 1‐methyl‐pseudouridine itself was determined.

Nievergelt, Philipp↗