Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Direct Sequence”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Evaluation of the Impact of Concentration and Extraction Methods on the Targeted Sequencing of Human Viruses from Wastewater

Sequencing human viruses in wastewater is challenging due to their low abundance compared to the total microbial background. This study compared the impact of four virus concentration/extraction methods (Innovaprep, Nanotrap, Promega, and Solids extraction) on probe-capture enrichment for human viruses followed by sequencing. Different concentration/extraction methods yielded distinct virus profiles. Innovaprep ultrafiltration (following solids removal) had the highest sequencing sensitivity and richness, resulting in the successful assembly of several near-complete human virus genomes. However, it was less sensitive in detecting SARS-CoV-2 by digital polymerase chain reaction (dPCR) compared to Promega and Nanotrap. Across all preparation methods, astroviruses and polyomaviruses were the most highly abundant human viruses, and SARS-CoV-2 was rare. These findings suggest that sequencing success can be increased using methods that reduce nontarget nucleic acids in the extract, though the absolute concentration of total extracted nucleic acid, as indicated by Qubit, and targeted viruses, as indicated by dPCR, may not be directly related to targeted sequencing performance. Further, using broadly targeted sequencing panels may capture viral diversity but risks losing signals for specific low-abundance viruses. Overall, this study highlights the importance of aligning wet lab and bioinformatic methods with specific goals when employing probe-capture enrichment for human virus sequencing from wastewater.

59 BASIC BIOLOGICAL SCIENCES↗

Nanopore Sequencing in Space: The Advancement of In Situ Microbiome Analysis for the International Space Station and Beyond

Introduction: Routine assessment of the International Space Station (ISS) microbiome has been performed through in-flight culture and Earth-based analysis. Through utilization of the ISS, significant progress toward in situ microbial identification has been achieved. In 2016, the molecular space age began when DNA-based, Earth-prepared samples were amplified within miniPCR (Genes in Space-1), and subsequent samples were sequenced with the MinION (Biomolecule Sequencer). The following year, these platforms were used collectively to yield the first off-Earth identification of unknown bacteria collected and cultured onboard the ISS (Genes in Space-3). Further advancement occurred in 2018 when a culture-independent, direct swab-to-sequencer method revealed a more thorough depiction of the ISS surface microbiome (BEST). As NASA looks towards Artemis and extended exploration missions, it is critical to continue to harness the ISS to expand in situ nanopore sequencing-based analysis in support of microbial-related crew health, planetary protection, and space research initiatives.

Christian G. Mena↗

A comparison of RNA with DNA in template-directed synthesis

Nonenzymatic template-directed copying of RNA sequences rich in cytidylic acid using nucleoside 5'-(2-methylimidazol-1-yl phosphates) as substrates is substantially more efficient than the copying of corresponding DNA sequences. However, many sequences cannot be copied, and the prospect of replication in this system is remote, even for RNA. Surprisingly, wobble-pairing leads to much more efficient incorporation of G opposite U on RNA templates than of G opposite T on DNA templates.

Non-NASA Center↗

A Chemoselective and Stereodivergent Platform of Heme‐Nitrene Transferases to Access Chiral Aryl‐β‐Amino Esters and An Investigation of the Sequence‐Activity Landscape

Engineered biocatalysts can utilize nitrene precursors to access enantioenriched amination products, yet they have not been applied to produce valuable, enantiomerically enriched noncanonical β-amino esters. Current approaches to synthesizing β-amino acids rely on pre-oxidized precursors and multistep synthetic approaches involving various protecting groups. We engineered a platform of heme enzymes for stereoselective C–H bond amination of readily available carboxylic ester derivatives to install primary amines. A directed evolution campaign coupled with sequencing of over 1000 variants enabled us to develop engineered variants that use either O-pivaloylhydroxylamine triflic acid (PONT) or hydroxylamine hydrochloride (H 2 NOH∙HCl) as aminating reagents. An analysis of the resulting sequence–activity dataset revealed additional improvements that could be made to the final variant, highlighting the utility of sequencing data to guide future steps in directed evolution campaigns. Furthermore, the evolved nitrene transferases expand the scope of accessible chiral β-amino acid building blocks for peptidomimetic applications and provide new starting points for the design and synthesis of enantioenriched β-amino acid motifs.

amino ester building blocks↗

Use of T7 RNA polymerase to direct expression of outer Surface Protein A (OspA) from the Lyme disease Spirochete, Borrelia burgdorferi

The OspA gene from a North American strain of the Lyme disease Spirochete, Borrelia burgdorferi, was cloned under the control of transciption and translation signals from bacteriophage T7. Full-length OspA protein, a 273 amino acid (31kD) lipoprotein, is expressed poorly in Escherichia coli and is associated with the insoluble membrane fraction. In contrast, a truncated form of OspA lacking the amino-terminal signal sequence which normally would direct localization of the protein to the outer membrane is expressed at very high levels (less than or equal to 100 mg/liter) and is soluble. The truncated protein was purified to homogeneity and is being tested to see if it will be useful as an immunogen in a vaccine against Lyme disease. Circular dichroism and fluorescence spectroscopy was used to characterize the secondary structure and study conformational changes in the protein. Studies underway with other surface proteins from B burgdorferi and a related spirochete, B. hermsii, which causes relapsing fever, leads us to conclude that a strategy similar to that used to express the truncated OspA can provide a facile method for producing variations of Borrelia lipoproteins which are highly expressed in E. coli and soluble without exposure to detergents.

Dunn, John J.↗

Biases in genome reconstruction from metagenomic data

Background Advances in sequencing, assembly, and assortment of contigs into species-specific bins has enabled the reconstruction of genomes from metagenomic data (MAGs). Though a powerful technique, it is difficult to determine whether assembly and binning techniques are accurate when applied to environmental metagenomes due to a lack of complete reference genome sequences against which to check the resulting MAGs. Methods We compared MAGs derived from an enrichment culture containing ~20 organisms to complete genome sequences of 10 organisms isolated from the enrichment culture. Factors commonly considered in binning software—nucleotide composition and sequence repetitiveness—were calculated for both the correctly binned and not-binned regions. This direct comparison revealed biases in sequence characteristics and gene content in the not-binned regions. Additionally, the composition of three public data sets representing MAGs reconstructed from the Tara Oceans metagenomic data was compared to a set of representative genomes available through NCBI RefSeq to verify that the biases identified were observable in more complex data sets and using three contemporary binning software packages. Results Repeat sequences were frequently not binned in the genome reconstruction processes, as were sequence regions with variant nucleotide composition. Genes encoded on the not-binned regions were strongly biased towards ribosomal RNAs, transfer RNAs, mobile element functions and genes of unknown function. Our results support genome reconstruction as a robust process and suggest that reconstructions determined to be >90% complete are likely to effectively represent organismal function; however, population-level genotypic heterogeneity in natural populations, such as uneven distribution of plasmids, can lead to incorrect inferences.

54 ENVIRONMENTAL SCIENCES↗

Genome dependent Cas9/gRNA search time underlies sequence dependent gRNA activity

Abstract CRISPR-Cas9 is a powerful DNA editing tool. A gRNA directs Cas9 to cleave any DNA sequence with a PAM. However, some gRNA sequences mediate cleavage at higher efficiencies than others. To understand this, numerous studies have screened large gRNA libraries and developed algorithms to predict gRNA sequence dependent activity. These algorithms do not predict other datasets as well as their training dataset and do not predict well between species. Here, to better understand these discrepancies, we retrospectively examine sequence features that impact gRNA activity in 44 published data sets. We find strong evidence that gRNA sequence dependent activity is largely influenced by the ability of the Cas9/gRNA complex to find the target site rather than activity at the target site and that this drives sequence dependent differences in gRNA activity between different species. This understanding will help guide future work to understand Cas9 activity as well as efforts to identify optimal gRNAs and improve Cas9 variants.

59 BASIC BIOLOGICAL SCIENCES↗

Differential laboratory passaging of SARS-CoV-2 viral stocks impacts the in vitro assessment of neutralizing antibodies

Viral populations in natural infections can have a high degree of sequence diversity, which can directly impact immune escape. However, antibody potency is often tested in vitro with a relatively clonal viral populations, such as laboratory virus or pseudotyped virus stocks, which may not accurately represent the genetic diversity of circulating viral genotypes. This can affect the validity of viral phenotype assays, such as antibody neutralization assays. To address this issue, we tested whether recombinant virus carrying SARS-CoV-2 spike (VSV-SARS-CoV-2-S) stocks could be made more genetically diverse by passage, and if a stock passaged under selective pressure was more capable of escaping monoclonal antibody (mAb) neutralization than unpassaged stock or than viral stock passaged without selective pressures. We passaged VSV-SARS-CoV-2-S four times concurrently in three cell lines and then six times with or without polyclonal antiserum selection pressure. All three of the monoclonal antibodies tested neutralized the viral population present in the unpassaged stock. The viral inoculum derived from serial passage without antiserum selection pressure was neutralized by two of the three mAbs. However, the viral inoculum derived from serial passage under antiserum selection pressure escaped neutralization by all three mAbs. Deep sequencing revealed the rapid acquisition of multiple mutations associated with antibody escape in the VSV-SARS-CoV-2-S that had been passaged in the presence of antiserum, including key mutations present in currently circulating Omicron subvariants. These data indicate that viral stock that was generated under polyclonal antiserum selection pressure better reflects the natural environment of the circulating virus and may yield more biologically relevant outcomes in phenotypic assays. Thus, mAb assessment assays that utilize a more genetically diverse, biologically relevant, virus stock may yield data that are relevant for prediction of mAb efficacy and for enhancing biosurveillance.

60 APPLIED LIFE SCIENCES↗

EHR-BERT: A BERT-based model for effective anomaly detection in electronic health records

Objective: Physicians and clinicians rely on data contained in electronic health records (EHRs), as recorded by health information technology (HIT), to make informed decisions about their patients. The reliability of HIT systems in this regard is critical to patient safety. Consequently, better tools are needed to monitor the performance of HIT systems for potential hazards that could compromise the collected EHRs, which in turn could affect patient safety. In this paper, we propose a new framework for detecting anomalies in EHRs using sequence of clinical events. This new framework, EHR-Bidirectional Encoder Representations from Transformers (BERT), is motivated by the gaps in the existing deep-learning related methods, including high false negatives, sub-optimal accuracy, higher computational cost, and the risk of information loss. EHR-BERT is an innovative framework rooted in the BERT architecture, meticulously tailored to navigate the hurdles in the contemporary BERT method; thus, enhancing anomaly detection in EHRs for healthcare applications.Methods: The EHR-BERT framework was designed using the Sequential Masked Token Prediction (SMTP) method. This approach treats EHRs as natural language sentences and iteratively masks input tokens during both training and prediction stages. This method facilitates the learning of EHR sequence patterns in both directions for each event and identifies anomalies based on deviations from the normal execution models trained on EHR sequences.Results: Extensive experiments on large EHR datasets across various medical domains demonstrate that EHR-BERT markedly improves upon existing models. It significantly reduces the number of false positives and enhances the detection rate, thus bolstering the reliability of anomaly detection in electronic health records. This improvement is attributed to the model’s ability to minimize information loss and maximize data utilization effectively.Conclusion: EHR-BERT showcases immense potential in decreasing medical errors related to anomalous clinical events, positioning itself as an indispensable asset for enhancing patient safety and the overall standard of healthcare services. The framework effectively overcomes the drawbacks of earlier models, making it a promising solution for healthcare professionals to ensure the reliability and quality of health data.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

A Multiplexed Quantitative Analysis of Germline Single Amino Acid Variants by Targeted Proteomics in Nondepleted Human Plasma

Single amino acid variants (SAAVs) in protein sequences are often a direct result of single-nucleotide polymorphisms (SNPs). Certain germline SAAVs have shown biological relevance in different disease conditions but lack precise quantification in circulation, which could hinder functional investigations and progress in biomarker development. Here, we have developed a multiplexed liquid chromatography-selected reaction monitoring (LC-SRM) assay that monitors 5 wild-type and variant peptide pairs (Complement Factor B: CFB-R32Q/R32W, Clusterin: CLU-N317H, Fetuin B: FETUB-K360R, and Kininogen: KNG1-L212P) in nondepleted human plasma. The assay was optimized for imprecision, linearity, stability, and calibration assessments with CVs of under 20%. The wild-type and variant peptide pairs were characterized in a set of healthy individual plasma samples. These target identifications were also validated by SNP genotyping with more than 99% accuracy. For all protein targets, we observed significantly lower concentrations of WT species in the presence variant peptides. In CFB, the concentration of R32Q was significantly lower than its counterpart R32W variant and WT species. Furthermore, our results distinguished phenotypes of homozygosity and heterozygosity of the SAAV presence through direct concentration level characterization. These findings provide some insights into how SAAVs affect quantitative assessments of target peptides. The assay demonstrates a platform for proteogenomic analyses with potential applications in both research and clinical settings.

genetics↗

Decoding substrate specificity determining factors in glycosyltransferase-B enzymes – insights from machine learning models

Substrate specificity is an essential characteristic of any enzyme's function and an understanding of the factors that determine this specificity is crucial for enzyme engineering. Unlike the structure of an enzyme which is directly impacted by its sequence, substrate specificity as an enzyme attribute involves a rather indirect relationship with sequence as it also depends on structural aspects that dictate substrate accessibility and active site dynamics. In this study, we explore the performance of classifier-based machine learning models trained on curated sequence and structural data for a class of glycosyltransferases (GTs), namely GT-Bs, to understand their substrate specificity determining factors. GTs enable the transfer of sugar moieties to other biomolecules such as oligosaccharides or proteins and are found in all kingdoms of life. In plants, GTs participate in the biosynthesis of plant cell wall biopolymers (e.g.: hemicelluloses and pectins) and are an integral part of the enzymatic machinery that enables the storage of carbon and energy as plant biomass. To elucidate the substrate specificity of uncharacterized GT-Bs, we constructed multi-label machine learning models (Support Vector Classifier, K-Nearest Neighbors, Gaussian Naïve-Bayes, Random Forest) that incorporate both sequence and structural features. These models achieve good predictive accuracies on test datasets. However, despite our use of structural information, we highlight that there is further scope for improvement in training these models to draw interpretable relationships between sequence, structure and substrate specificity determining motifs in GT-Bs.

97 MATHEMATICS AND COMPUTING↗

Detecting Spin-Bath Polarization with Quantum Quench Phase Shifts of Single Spins in Diamond

Single-qubit sensing protocols can be used to measure qubit-bath coupling parameters. However, for sufficiently large coupling, the sensing protocol itself perturbs the bath, which is predicted to result in a characteristic response in the sensing measurements. Here, we observe this bath perturbation, also known as a quantum quench, by preparing the nuclear spin bath of a nitrogen-vacancy (NV) center in polarized initial states and performing phase-resolved spin-echo measurements on the NV electron spin. These measurements reveal a time-dependent phase determined by the initial state of the bath. We derive the relationship between the sensor phase and the Gaussian spin-bath polarization and apply it to reconstruct both the axial and transverse polarization components. Using this insight, we optimize the transfer efficiency of our dynamic nuclear polarization sequence. This technique for directly measuring bath polarization may assist in preparing high-fidelity quantum memory states, improving nanoscale NMR methods, and investigating non-Gaussian quantum baths.

36 MATERIALS SCIENCE↗

Roles for epigenetics in wood formation and stress response intrees–from basic biology to forest management

Annual model and crop species have been the subject of most epigenetic studies for plants. In contrast to annuals, forest trees persist on natural landscapes and experience environmental variation within and across seasons, years, and decades or even centuries. Most forest trees species are undomesticated and typically grown on variable landscapes with no irrigation or application of agricultural chemicals. Forest trees must thus rely on their inherent ability to alter growth and physiology to mitigate the effects of changing abiotic and biotic stressors. Like other plants, trees have mechanisms encoded in their genomic DNA sequence that can respond directly to stress events such as drought or heat. Hypothetically, it would be highly advantageous to join these mechanisms with a dynamic “memory” of past exposure to stress. It is now well established that annual model and crop plants can establish epigenetic-based memory of stress events that support more rapid and robust response to stress in the future. Here, evidence is discussed for epigenetic regulation and “memory” in two fundamental biological processes in trees, wood formation and abiotic stress response. Wood formation is an ideal trait for epigenetic research in trees, as wood formation is highly responsive to environmental conditions and includes multiple rapid developmental changes as cells adopt distinct fates within complex tissues. This is followed by a discussion of research needs that would provide the foundation for new epigenetic applications for forestry.

Groover, Andrew↗

General verification description

A brief general description of the ASTP flight program verification was presented. The total program verification effort assures the accuracy and adequacy of the LVDC flight program and verifies that the final program meets mission requirements and conforms to program documentation. The flight program's functional requirements to integrate the guidance and control system with the launch vehicle sequencing system are verified directly by analysis of many special logic checks designed for this purpose and indirectly by the correct overall program response to nominal and numerous perturbed conditions. Verification of the interaction of function requirements is accomplished on every case run during the verification effort.

Source record↗

Dynamic delamination buckling in composite laminates under impact loading: Computational simulation

A unique dynamic delamination buckling and delamination propagation analysis capability has been developed and incorporated into a finite element computer program. This capability consists of the following: (1) a modification of the direct time integration solution sequence which provides a new analysis algorithm that can be used to predict delamination buckling in a laminate subjected to dynamic loading, and (2) a new method of modeling the composite laminate using plate bending elements and multipoint constraints. This computer program is used to predict both impact induced buckling in composite laminates with initial delaminations and the strain energy release rate due to extension of the delamination. It is shown that delaminations near the outer surface of a laminate are susceptible to local buckling and buckling-induced delamination propagation when the laminate is subjected to transverse impact loading. The capability now exists to predict the time at which the onset of dynamic delamination buckling occurs, the dynamic buckling mode shape, and the dynamic delamination strain energy release rate.

Grady, Joseph E.↗

Dynamic delamination buckling in composite laminates under impact loading - Computational simulation

A unique dynamic delamination buckling and delamination propagation analysis capability has been developed and incorporated into a finite element computer program. This capability consists of the following: (1) a modification of the direct time integration solution sequence which provides a new analysis algorithm that can be used to predict delamination buckling in a laminate subjected to dynamic loading, and (2) a new method of modeling the composite laminate using plate bending elements and multipoint constraints. This computer program is used to predict both impact induced buckling in composite laminates with initial delaminations and the strain energy release rate due to extension of the delamination. It is shown that delaminations near the outer surface of a laminate are susceptible to local buckling and buckling-induced delamination propagation when the laminate is subjected to transverse impact loading. The capability now exists to predict the time at which the onset of dynamic delamination buckling occurs, the dynamic buckling mode shape, and the dynamic delamination strain energy release rate.

Grady, Joseph E.↗

Removing spurious reflections from CFD solutions by using the Complex Cepstrum

The Complex Cepstrum is shown to remove spurious reflections from artificial boundaries in computational fluid dynamic (CFD) solutions. First, the Complex Cepstrum theory is presented. A model time sequence consisting of a direct signal and reflections is analyzed theoretically with the Complex Cepstrum, and it is shown that the direct signal uncontaminated by reflections may be recovered in the time domain. Next, the Complex Cepstrum is applied to one- and three-dimensional CFD solutions, and spurious reflections from the boundary conditions are removed. By eliminating spurious reflections introduced by artificial boundary conditions, the applicability of CFD methods to aeroacoustic problems is greatly enhanced.

Meadows, Kristine R.↗