Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “consensus sequence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Binding profiles for 961 Drosophila and C. elegans transcription factors reveal tissue-specific regulatory relationships

A catalog of transcription factor (TF) binding sites in the genome is critical for deciphering regulatory relationships. Here, we present the culmination of the efforts of the modENCODE (model organism Encyclopedia of DNA Elements) and modERN (model organism Encyclopedia of Regulatory Networks) consortia to systematically assay TF binding events in vivo in two major model organisms,Drosophila melanogaster(fly) andCaenorhabditis elegans(worm). These data sets comprise 605 TFs identifying 3.6 M sites in the fly and 356 TFs identifying 0.9 M sites in the worm, and represent the majority of the regulatory space in each genome. We demonstrate that TFs associate with chromatin in clusters termed “metapeaks,” that larger metapeaks have characteristics of high-occupancy target (HOT) regions, and that the importance of consensus sequence motifs bound by TFs depends on metapeak size and complexity. Combining ChIP-seq data with single-cell RNA-seq data in a machine-learning model identifies TFs with a prominent role in promoting target gene expression in specific cell types, even differentiating between parent–daughter cells during embryogenesis. These data are a rich resource for the community that should fuel and guide future investigations into TF function. To facilitate data accessibility and utility, all strains expressing green fluorescent protein (GFP)-tagged TFs are available at the stock centers for each organism. The chromatin immunoprecipitation sequencing data are available through the ENCODE Data Coordinating Center, GEO, and through a direct interface that provides rapid access to processed data sets and summary analyses, as well as widgets to probe the cell-type-specific TF–target relationships.

Biochemistry & Molecular Biology↗

Regulation of L - and D -Aspartate Transport and Metabolism in Acinetobacter baylyi ADP1

Here, the regulated uptake and consumption of d-amino acids by bacteria remain largely unexplored, despite the physiological importance of these compounds. Unlike other characterized bacteria, such as Escherichia coli, which utilizes only l-Asp, Acinetobacter baylyi ADP1 can consume both d-Asp and l-Asp as the sole carbon or nitrogen source. As described here, two LysR-type transcriptional regulators (LTTRs), DarR and AalR, control d- and l-Asp metabolism in strain ADP1. Heterologous expression of A. baylyi proteins enabled E. coli to use d-Asp as the carbon source when either of two transporters (AspT or AspY) and a racemase (RacD) were coexpressed. A third transporter, designated AspS, was also discovered to transport Asp in ADP1. DarR and/or AalR controlled the transcription of aspT, aspY, racD, and aspA (which encodes aspartate ammonia lyase). Conserved residues in the N-terminal DNA-binding domains of both regulators likely enable them to recognize the same DNA consensus sequence (ATGC-N7-GCAT) in several operator-promoter regions. In strains lacking AalR, suppressor mutations revealed a role for the ClpAP protease in Asp metabolism. In the absence of the ClpA component of this protease, DarR can compensate for the loss of AalR. ADP1 consumed l- and d-Asn and l-Glu, but not d-Glu, as the sole carbon or nitrogen source using interrelated pathways.

59 BASIC BIOLOGICAL SCIENCES↗

EDGE COVID-19: a web platform to generate submission-ready genomes from SARS-CoV-2 sequencing efforts

Abstract Summary Genomics has become an essential technology for surveilling emerging infectious disease outbreaks. A range of technologies and strategies for pathogen genome enrichment and sequencing are being used by laboratories worldwide, together with different and sometimes ad hoc, analytical procedures for generating genome sequences. A fully integrated analytical process for raw sequence to consensus genome determination, suited to outbreaks such as the ongoing COVID-19 pandemic, is critical to provide a solid genomic basis for epidemiological analyses and well-informed decision making. We have developed a web-based platform and integrated bioinformatic workflows that help to provide consistent high-quality analysis of SARS-CoV-2 sequencing data generated with either the Illumina or Oxford Nanopore Technologies (ONT). Using an intuitive web-based interface, this workflow automates data quality control, SARS-CoV-2 reference-based genome variant and consensus calling, lineage determination and provides the ability to submit the consensus sequence and necessary metadata to GenBank, GISAID and INSDC raw data repositories. We tested workflow usability using real world data and validated the accuracy of variant and lineage analysis using several test datasets, and further performed detailed comparisons with results from the COVID-19 Galaxy Project workflow. Our analyses indicate that EC-19 workflows generate high-quality SARS-CoV-2 genomes. Finally, we share a perspective on patterns and impact observed with Illumina versus ONT technologies on workflow congruence and differences. Availability and implementation https://edge-covid19.edgebioinformatics.org, and https://github.com/LANL-Bioinformatics/EDGE/tree/SARS-CoV2. Supplementary information Supplementary data are available at Bioinformatics online.

59 BASIC BIOLOGICAL SCIENCES↗

Evaluation of Potential In Vitro Recombination Events in Codon Deoptimized FMDV Strains

Codon deoptimization (CD) has been recently used as a possible strategy to derive foot-and-mouth disease (FMD) live-attenuated vaccine (LAV) candidates containing DIVA markers. However, reversion to virulence, or loss of DIVA, from possible recombination with wild-type (WT) strains has yet to be analyzed. An in vitro assay was developed to quantitate the levels of recombination between WT and a prospective A24-P2P3 partially deoptimized LAV candidate. By using two genetically engineered non-infectious RNA templates, we demonstrate that recombination can occur within non-deoptimized viral genomic regions (i.e., 3'end of P3 region). The sequencing of single plaque recombinants revealed a variety of genome compositions, including full-length WT sequences at the consensus level and deoptimized sequences at the sub-consensus/consensus level within the 3'end of the P3 region. Notably, after further passage, two recombinants that contained deoptimized sequences evolved to WT. Overall, recombinants featuring large stretches of CD or DIVA markers were less fit than WT viruses. Our results indicate that the developed assay is a powerful tool to evaluate the recombination of FMDV genomes in vitro and should contribute to the improved design of FMDV codon deoptimized LAV candidates.

59 BASIC BIOLOGICAL SCIENCES↗

Cloning the promoter for transforming growth factor-beta type III receptor. Basal and conditional expression in fetal rat osteoblasts

Transforming growth factor-beta binds to three high affinity cell surface molecules that directly or indirectly regulate its biological effects. The type III receptor (TRIII) is a proteoglycan that lacks significant intracellular signaling or enzymatic motifs but may facilitate transforming growth factor-beta binding to other receptors, stabilize multimeric receptor complexes, or segregate growth factor from activating receptors. Because various agents or events that regulate osteoblast function rapidly modulate TRIII expression, we cloned the 5' region of the rat TRIII gene to assess possible control elements. DNA fragments from this region directed high reporter gene expression in osteoblasts. Sequencing showed no consensus TATA or CCAAT boxes, whereas several nuclear factors binding sequences within the 3' region of the promoter co-mapped with multiple transcription initiation sites, DNase I footprints, gel mobility shift analysis, or loss of activity by deletion or mutation. An upstream enhancer was evident 5' proximal to nucleotide -979, and a silencer region occurred between nucleotides -2014 and -2194. Glucocorticoid sensitivity mapped between nucleotides -687 and -253, whereas bone morphogenetic protein 2 sensitivity co-mapped within the silencer region. Thus, the TRIII promoter contains cooperative basal elements and dispersed growth factor- and hormone-sensitive regulatory regions that can control TRIII expression by osteoblasts.

Non-NASA Center↗

Continuous in vitro evolution of bacteriophage RNA polymerase promoters

Rapid in vitro evolution of bacteriophage T7, T3, and SP6 RNA polymerase promoters was achieved by a method that allows continuous enrichment of DNAs that contain functional promoter elements. This method exploits the ability of a special class of nucleic acid molecules to replicate continuously in the presence of both a reverse transcriptase and a DNA-dependent RNA polymerase. Replication involves the synthesis of both RNA and cDNA intermediates. The cDNA strand contains an embedded promoter sequence, which becomes converted to a functional double-stranded promoter element, leading to the production of RNA transcripts. Synthetic cDNAs, including those that contain randomized promoter sequences, can be used to initiate the amplification cycle. However, only those cDNAs that contain functional promoter sequences are able to produce RNA transcripts. Furthermore, each RNA transcript encodes the RNA polymerase promoter sequence that was responsible for initiation of its own transcription. Thus, the population of amplifying molecules quickly becomes enriched for those templates that encode functional promoters. Optimal promoter sequences for phage T7, T3, and SP6 RNA polymerase were identified after a 2-h amplification reaction, initiated in each case with a pool of synthetic cDNAs encoding greater than 10(10) promoter sequence variants.

Non-NASA Center↗

HiFiAdapterFilt, a memory efficient read processing pipeline, prevents occurrence of adapter sequence in PacBio HiFi reads and their negative impacts on genome assembly

Abstract Background Pacific Biosciences HiFi read technology is currently the industry standard for high accuracy long-read sequencing that has been widely adopted by large sequencing and assembly initiatives for generation of de novo assemblies in non-model organisms. Though adapter contamination filtering is routine in traditional short-read analysis pipelines, it has not been widely adopted for HiFi workflows. Results Analysis of 55 publicly available HiFi datasets revealed that a read-sanitation step to remove sequence artifacts derived from PacBio library preparation from read pools is necessary as adapter sequences can be erroneously integrated into assemblies. Conclusions Here we describe the nature of adapter contaminated reads, their consequences in assembly, and present HiFiAdapterFilt, a simple and memory efficient solution for removing adapter contaminated reads prior to assembly.

59 BASIC BIOLOGICAL SCIENCES↗

Role of hypoxia-inducible factor-1 in transcriptional activation of ceruloplasmin by iron deficiency

A role of the copper protein ceruloplasmin (Cp) in iron metabolism is suggested by its ferroxidase activity and by the tissue iron overload in hereditary Cp deficiency patients. In addition, plasma Cp increases markedly in several conditions of anemia, e.g. iron deficiency, hemorrhage, renal failure, sickle cell disease, pregnancy, and inflammation. However, little is known about the cellular and molecular mechanism(s) involved. We have reported that iron chelators increase Cp mRNA expression and protein synthesis in human hepatocarcinoma HepG2 cells. Furthermore, we have shown that the increase in Cp mRNA is due to increased rate of transcription. We here report the results of new studies designed to elucidate the molecular mechanism underlying transcriptional activation of Cp by iron deficiency. The 5'-flanking region of the Cp gene was cloned from a human genomic library. A 4774-base pair segment of the Cp promoter/enhancer driving a luciferase reporter was transfected into HepG2 or Hep3B cells. Iron deficiency or hypoxia increased luciferase activity by 5-10-fold compared with untreated cells. Examination of the sequence showed three pairs of consensus hypoxia-responsive elements (HREs). Deletion and mutation analysis showed that a single HRE was necessary and sufficient for gene activation. The involvement of hypoxia-inducible factor-1 (HIF-1) was shown by gel-shift and supershift experiments that showed HIF-1alpha and HIF-1beta binding to a radiolabeled oligonucleotide containing the Cp promoter HRE. Furthermore, iron deficiency (and hypoxia) did not activate Cp gene expression in Hepa c4 hepatoma cells deficient in HIF-1beta, as shown functionally by the inactivity of a transfected Cp promoter-luciferase construct and by the failure of HIF-1 to bind the Cp HRE in nuclear extracts from these cells. These results are consistent with in vivo findings that iron deficiency increases plasma Cp and provides a molecular mechanism that may help to understand these observations.

NASA Discipline Cardiopulmonary↗

Molecular relationships between closely related strains and species of nematodes

Electrophoretic comparisons have been made for 24 enzymes in the Bergerac and Bristol strains of Caenorhabditis elegans and the related species, Caenorhabditis briggsae. No variation was detected between the two strains of C. elegans. In contrast, the two species, C. elegans and C. briggsae exhibited electrophoretic differences in 22 of 24 enzymes. A consensus 5S rRNA sequence was determined for C. elegans and found to be identical to that from C. briggsae. By analogy with other species with relatively well established fossil records it can be inferred that the time of divergence between the two nematode species is probably in the tens of millions of years. The limited anatomical evolution during a time period in which proteins undergo extensive changes supports the hypothesis that anatomical evolution is not dependent on overall protein changes.

Butler, M. H.↗

Sequence-defined structural transitions by calcium-responsive proteins

Biopolymer sequences dictate their functions, and protein-based polymers are a promising platform to establish sequence–function relationships for novel biopolymers. To efficiently explore vast sequence spaces of natural proteins, sequence repetition is a common strategy to tune and amplify specific functions. This strategy is applied to repeats-in-toxin (RTX) proteins with calcium-responsive folding behavior, which stems from tandem repeats of the nonapeptide GGXGXDXUX in which X can be any amino acid and U is a hydrophobic amino acid. To determine the functional range of this nonapeptide, we modified a naturally occurring RTX protein that forms β-roll structures in the presence of calcium. Sequence modifications focused on calcium-binding turns within the repetitive region, including either global substitution of nonconserved residues or complete replacement with tandem repeats of a consensus nonapeptide GGAGXDTLY. Some sequence modifications disrupted the typical transition from intrinsically disordered random coils to folded β rolls, despite conservation of the underlying nonapeptide sequence. Proteins enriched with smaller, hydrophobic amino acids adopted secondary structures in the absence of calcium and underwent structural rearrangements in calcium-rich environments. In contrast, proteins with bulkier, hydrophilic amino acids maintained intrinsic disorder in the absence of calcium. In conclusion, these results indicate a significant role of nonconserved amino acids in calcium-responsive folding, thereby revealing a strategy to leverage sequences in the design of tunable, calcium-responsive biopolymers.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Matrix Metalloproteinases as Candidate Antigenic Determinants for Anti‐Tumor Autoantibodies in Human Ovarian Cancer: A Post Hoc Analysis

Circulating antibodies in patients with cancer can facilitate the identification of accessible epitopes on autoantigens expressed by tumors. To identify previously unrecognized protein targets in ovarian cancer, we computationally assessed a heptapeptide consensus motif (VPELGHE, flanked by two cysteine residues yielding a cyclic nonapeptide under oxidizing conditions) previously discovered via phage display-based epitope mapping of autoantibodies in patients. Eight proteins associated with ovarian cancer encompass amino acid sequences similar to the consensus motif and were, therefore, considered as candidate native autoantigens. Among these candidate targets, however, matrix metalloproteinase 14 (MMP14) demonstrates gene expression that is both high and negatively correlated with survival in ovarian cancer patient cohorts. MMP14 protein levels are also stable in tumor versus non-tumor tissues. Moreover, the corresponding heptapeptide mimic in MMP14 occurs within an α-helical secondary structural element observed in its catalytic domain. These findings demonstrate that a subset of patient-derived autoantibodies may interact with a previously unknown antigenic epitope found in MMP14 and other MMPs, thereby providing opportunities for the development of new targeted agents.

Biochemistry & Molecular Biology↗

De novo design of symmetric ferredoxins that shuttle electrons in vivo

A symmetric origin for bacterial ferredoxins was first proposed over 50 y ago, yet, to date, no functional symmetric molecule has been constructed. It is hypothesized that extant proteins have drifted from their symmetric roots via gene duplication followed by mutations. Phylogenetic analyses of extant ferredoxins support the independent evolution of N- and C-terminal sequences, thereby allowing consensus-based design of symmetric 4Fe-4S molecules. All designs bind two [4Fe-4S] clusters and exhibit strongly reducing midpoint potentials ranging from −405 to −515 mV. One of these constructs efficiently shuttles electrons through a designed metabolic pathway in Escherichia coli. These finding establish that ferredoxins consisting of a symmetric core can be used as a platform to design novel electron transfer carriers for in vivo applications. Outer-shell asymmetry increases sequence space without compromising electron transfer functionality.

Andrew C. Mutter↗

CONSTAX2: improved taxonomic classification of environmental DNA markers

Abstract Summary CONSTAX—the CONSensus TAXonomy classifier—was developed for accurate and reproducible taxonomic annotation of fungal rDNA amplicon sequences and is based upon a consensus approach of RDP, SINTAX and UTAX algorithms. CONSTAX2 extends these features to classify prokaryotes as well as eukaryotes and incorporates BLAST-based classifiers to reduce classification errors. Additionally, CONSTAX2 implements a conda-installable command-line tool with improved classification metrics, faster training, multithreading support, capacity to incorporate external taxonomic databases and new isolate matching and high-level taxonomy tools, replete with documentation and example tutorials. Availability and implementation CONSTAX2 is available at https://github.com/liberjul/CONSTAXv2, and is packaged for Linux and MacOS from Bioconda with use under the MIT License. A tutorial and documentation are available at https://constax.readthedocs.io/en/latest/. Data and scripts associated with the manuscript are available at https://github.com/liberjul/CONSTAXv2_ms_code. Supplementary information Supplementary data are available at Bioinformatics online.

59 BASIC BIOLOGICAL SCIENCES↗

Recombinant And Mix-Infection Finder for SARS-CoV-2 sample

The scientific and public health communities responded to the COVID-19 pandemic with sample acquisition and genome sequencing on a scale that eclipsed all prior sequencing efforts. While this can only be characterized as a resounding success story that has cemented the use of genomics for epidemiological investigations for any future infectious disease outbreak, several retrospective studies are cataloging an array of lessons learned and issues that have yet to be addressed in order to realize the full potential of genomics as a routine biosurveillance tool. We have been both developing methods to accurately assess SARS-CoV-2 genomes from complex samples, and analyzing the large volumes of international data, both at the consensus level and the raw sequencing data. During the course of our investigations and similar to other groups, we have examined COVID-19 samples with signatures from multiple lineages of SARS-CoV-2 and will describe some of our findings during the development of a novel workflow that incorporates detection and reporting of potential co-infection within samples and also highlights any evidence of within-host recombination.

Lo, Chien-Chi↗

U.S. Efforts in Support of Examinations at Fukushima Daiichi - November 2022 Meeting Notes and Information Request Status

Information obtained from Fukushima Daiichi Nuclear Power Station (Daiichi) is required to inform future Decontamination and Decommissioning (D&D) activities, improving the ability of the Tokyo Electric Power Company Holdings, Incorporated (TEPCO Holdings) to characterize potential hazards and to ensure the safety of workers involved with cleanup activities. This information also has important implications for the safety and operation of U.S. commercial nuclear power plants. This document summarizes results from the Fiscal Year 2023 (FY2023) U.S. effort to review Daiichi information and extract insights to enhance the safety of existing and future nuclear power plant designs. This U.S. effort, which was initiated in 2014 by the Department of Energy Office of Nuclear Energy, is completed by a group of experts in reactor safety and plant operations that identify examination needs and evaluate recent Daiichi examination data to address these needs. Fukushima-related information and associated discussions during these meetings benefit operating, new, and advanced reactors. Significant safety insights have been and are continuing to be obtained in several areas: system and component performance, radionuclide surveys and sampling, debris end-state location, combustible gas effects, and plant operations and maintenance. In addition to reducing uncertainties related to severe accident modeling progression, these insights have and continue to be used to update guidance for severe accident prevention, mitigation, and emergency planning. Furthermore, Daiichi-related activities, such as code modeling improvements and analysis, testing, and new technology deployment efforts, have the potential to offer additional benefits to the operating fleet and new LWR and non-LWR designs. U.S. evaluations of obtained examination information and input regarding future Daiichi examinations are of interest to several organizations within Japan. Since its inception, the U.S. has provided consensus input for high priority time-sequenced examination tasks and supporting research activities. In their Mid-to-Long-term Examination Plan for 1F investigations, TEPCO included all remaining U.S. consensus information requests and additional information requests they identified. TEPCO periodically provides reports on the status of these requests (reflecting D&D priorities, new insights from investigations, and new technologies that become available). Hence, U.S. experts agreed that it was appropriate for TEPCO to track and prioritize these information requests as D&D progresses. U.S. experts will continue to review and comment on the information obtained from examinations and, as needed, provide additional details and relevant background material to support future examinations. As documented in this report, several other items, such as additional details on information requests pertaining to ex-vessel examinations, relevant references from prior research, additional documents to provide insights regarding recent investigation findings, and reviews of recently released documents, were agreed to during the FY2023 meeting.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Prevalence of gp160 polymorphisms known to be related to decreased susceptibility to temsavir in different subtypes of HIV-1 in the Los Alamos National Laboratory HIV Sequence Database

Fostemsavir, a prodrug of the gp120-directed attachment inhibitor temsavir, is indicated for use in heavily treatment-experienced individuals with MDR HIV-1. Reduced susceptibility to temsavir in the clinic maps to discrete changes at amino acid positions in gp160: S375, M426, M434 and M475.To query the Los Alamos National Laboratory (LANL) HIV Sequence Database for the prevalence of polymorphisms at gp160 positions of interest. Full-length gp160 sequences (N = 7560) were queried for amino acid polymorphisms relative to the subtype B consensus at positions of interest; frequencies were reported for all sequences and among subtypes/circulating recombinant forms (CRFs) with ≥10 isolates in the database. Among 239 subtypes in the database, the 5 most prevalent were B (n = 2651, 35.1%), C (n = 1626, 21.5%), CRF01_AE (n = 674, 8.9%), A1 (n = 273, 3.6%) and CRF02_AG (n = 199, 2.6%). Among all 7560 sequences, the most prevalent amino acids at positions of interest (S 375 , 73.5%; M 426 , 82.1%; M 434 , 88.2%; M 475 , 89.9%) were the same as the subtype B consensus. Specific polymorphisms with the potential to decrease temsavir susceptibility (S 375 H/I/M/N/T/Y, M 426 L/P, M 434 I/K and M475I) were found in <10% of isolates of subtypes D, G, A6, BC, F1, CRF07_BC, CRF08_BC, 02A, CRF06_cpx, F2, 02G and 02B. S 375 H and M 475 I were predominant among CRF01_AE (S375H, 99.3%; M 475 I, 76.3%; consistent with previously reported low temsavir susceptibility of this CRF) and 01B (S 375 H, 71.7%; M 475 I, 49.5%). Analysis of the LANL HIV Sequence Database found a low prevalence of gp160 amino acid polymorphisms with the potential to reduce temsavir susceptibility overall and among most of the common subtypes.

59 BASIC BIOLOGICAL SCIENCES↗

17beta-estradiol potently suppresses cAMP-induced insulin-like growth factor-I gene activation in primary rat osteoblast cultures

Insulin-like growth factor-I (IGF-I) is a key factor in bone remodeling. In osteoblasts, IGF-I synthesis is enhanced by parathyroid hormone and prostaglandin E2 (PGE2) through cAMP-activated protein kinase. In rats, estrogen loss after ovariectomy leads to a rise in serum IGF-I and an increase in bone remodeling, both of which are reversed by estrogen treatment. To examine estrogen-dependent regulation of IGF-I expression at the molecular level, primary fetal rat osteoblasts were co-transfected with the estrogen receptor (hER, to ensure active ER expression), and luciferase reporter plasmids controlled by promoter 1 of the rat IGF-I gene (IGF-I P1), used exclusively in these cells. As reported, 1 microM PGE2 increased IGF-I P1 activity by 5-fold. 17beta-Estradiol alone had no effect, but dose-dependently suppressed the stimulatory effect of PGE2 by up to 90% (ED50 approximately 0.1 nM). This occurred within 3 h, persisted for at least 16 h, required ER, and appeared specific, since 17alpha-estradiol was 100-300-fold less effective. By contrast, 17beta-estradiol stimulated estrogen response element (ERE)-dependent reporter expression by up to 10-fold. 17beta-Estradiol also suppressed an IGF-I P1 construct retaining only minimal promoter sequence required for cAMP-dependent gene activation, but did not affect the 60-fold increase in cAMP induced by PGE2. There is no consensus ERE in rat IGF-I P1, suggesting novel downstream interactions in the cAMP pathway that normally enhances IGF-I expression in skeletal cells. To explore this, nuclear extract from osteoblasts expressing hER were examined by electrophoretic mobility shift assay using the atypical cAMP response element in IGF-I P1. Estrogen alone did not cause DNA-protein binding, while PGE2 induced a characteristic gel shift complex. Co-treatment with both hormones caused a gel shift greatly diminished in intensity, consistent with their combined effects on IGF-I promoter activity. Nonetheless, hER did not bind IGF-I cAMP response element or any adjacent sequences. These results provide new molecular evidence that estrogen may temper the biological effects of hormones acting through cAMP to regulate skeletal IGF-I expression and activity.

Non-NASA Center↗