Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Sequencing data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Weight distributions for turbo codes using random and nonrandom permutations

This article takes a preliminary look at the weight distributions achievable for turbo codes using random, nonrandom, and semirandom permutations. Due to the recursiveness of the encoders, it is important to distinguish between self-terminating and non-self-terminating input sequences. The non-self-terminating sequences have little effect on decoder performance, because they accumulate high encoded weight until they are artificially terminated at the end of the block. From probabilistic arguments based on selecting the permutations randomly, it is concluded that the self-terminating weight-2 data sequences are the most important consideration in the design of constituent codes; higher-weight self-terminating sequences have successively decreasing importance. Also, increasing the number of codes and, correspondingly, the number of permutations makes it more and more likely that the bad input sequences will be broken up by one or more of the permuters. It is possible to design nonrandom permutations that ensure that the minimum distance due to weight-2 input sequences grows roughly as the square root of (2N), where N is the block length. However, these nonrandom permutations amplify the bad effects of higher-weight inputs, and as a result they are inferior in performance to randomly selected permutations. But there are 'semirandom' permutations that perform nearly as well as the designed nonrandom permutations with respect to weight-2 input sequences and are not as susceptible to being foiled by higher-weight inputs.

Dolinar, S.↗

Transcriptomic data sets for Novosphingobium aromaticivorans DSM12444 and a ΔSARO_RS14285 mutant grown in the presence of glucose and either protocatechuic, vanillic, syringic, or 4-coumaric acid

The SARO_RS14285 gene, encoding a transcription factor, was deleted in Novosphingobium aromaticivorans DSM12444. The transcriptomes of the parent and ΔSARO_RS14285 strains were determined when grown in medium containing glucose with or without protocatechuic, vanillic, syringic, or 4-coumaric acid. We present the raw RNA sequencing data obtained from these cultures.

Novosphingobium aromaticivorans↗

Atomic Data and Spectral Line Intensities for Be-like Ions

Atomic data and collision rates are needed to model the spectrum of optically thin astrophysical sources. Recent observations from solar instrumentation such as SOH0 and Hinode have revealed the presence of hosts of lines emitted by high-energy configurations from ions belonging to the Be-like to the 0-like isoelectronic sequences. Data for such configurations are often unavailable in the literature. We have started a program to calculate the atomic parameters and rates for the high-energy configurations of Be-like ions of the type ls2.21.nl' where n=3,4,5. We report on the results of this project and on the diagnostic application of the predicted spectral lines.

Bhatia, Anand↗

Transcriptomic data sets for Novosphingobium aromaticivorans grown with the β-5-linked aromatic dimer dehydrodiconiferyl alcohol and the related G-aromatic monomers vanillin and ferulic acid

ABSTRACT The transcriptomes of a 2-pyrone-4,6-dicarboxylic acid-producing strain of Novosphingobium aromaticivorans DSM12444 were determined when grown in minimal medium containing glucose alone or glucose plus vanillin, ferulic acid, or the β-5-linked aromatic dimer dehydrodiconiferyl alcohol as carbon sources. Here, we present the RNA-sequencing data we obtained.

Metz, Fletcher↗

Measurements of Cumulonimbus Clouds using quantitative satellite and radar data

Results are reported for a preliminary study of SMS-2 digital brightness and IR data obtained at frequent 5-7.5 min intervals. The clouds studied were over the Central and Great Plains in midlatitudes and thus were typical of an environment much different from that of the tropical oceans. The satellite data are compared to radar data for both a severe weather event and weak thundershower activity of the type which might be a target for weather modification efforts. The relative importance of short time interval satellite data is shown for both cases, and possible relationships between the two types of data are presented. It is concluded that (1) using a threshold technique for visible reflected brightness, precipitating vs. nonprecipitating clouds can be discriminated; (2) brightness is well related to cloud size and shape; and (3) satellite-derived growth rates may be a significant parameter to be used in determining storm severity, especially if rapid time sequence data are used during the development phase of the storm.

Negri, A. J.↗

Interplanetary approach optical navigation with applications

The use of optical data from onboard television cameras for the navigation of interplanetary spacecraft during the planet approach phase is investigated. Three optical data types were studied: the planet limb with auxiliary celestial references, the satellite-star, and the planet-star two-camera methods. Analysis and modelling issues related to the nature and information content of the optical methods were examined. Dynamic and measurement system modelling, data sequence design, measurement extraction, model estimation and orbit determination, as relating optical navigation, are discussed, and the various error sources were analyzed. The methodology developed was applied to the Mariner 9 and the Viking Mars missions. Navigation accuracies were evaluated at the control and knowledge points, with particular emphasis devoted to the combined use of radio and optical data. A parametric probability analysis technique was developed to evaluate navigation performance as a function of system reliabilities.

Jerath, N.↗

Pioneer Venus spacecraft design and operation

The Pioneer Venus Orbiter and Multiprobe spacecraft design and operation enabled both remote and in-situ measurements of the Venusian environment from the outermost fringes of the atmosphere all the way to the surface. Both spacecraft were spin-stabilized and solar-cell powered from launch to Venus. Since orbit insertion, the Orbiter has been transmitting measurements from a highly elliptical 24-h orbit with periapsis altitudes down to about 150 km. Data rates up to 2048 bits/s have been utilized through a despun high-gain antenna transmitting at S-band frequency. Spacecraft attitudes, orbit periods, and periapsis altitudes are being maintained as required with a hydrazine propulsion system. The Multiprobe spacecraft (Bus with all four Probes attached) performed the necessary Probe checkouts and deployed the Probes to achieve the desired Probe and Bus targeting. Silver-zinc batteries provided the necessary power on each of the four Probes from separation from the Bus through the entry/descent sequence. Data rates of 256 and 128 bits/s on the Large Probe were maintained with 40-W radiated power, and 64 and 16 bits/s on the Small Probes were maintained with 10-W radiated power, through omni antennas directly to Earth-based stations. Each Probe's entry/descent sequence was controlled with a hardwired entry sequence programmer to achieve the desired scientific and spacecraft operations.

Nothwang, G. J.↗

From microbial diversity to functional potential using dimensionality reduction

The high dimensionality of microbial diversity data from ‘omics observations can be reduced using Machine Learning, with many recent studies showcasing ML utility for exploratory ecological feature finding and process prediction. Here, we compare the Self Organizing Map (SOM) dimensionality reduction method to the well-documented sample-based Principal Coordinate Analysis (PCoA) and taxa-based Weighted Gene Correlation Network Analysis (WGCNA) using near daily 16S rRNA gene amplicon sequencing data from the 2019 to 2020 MOSAiC International Arctic Drift Expedition. We then map k-means clustering outputs from each method to available metagenomes, extracting functionally distinct seasonal microbial ecotypes in the surface Arctic Ocean. Our results indicate the SOM method better represented expected seasonal transitions and identified a greater number of metabolically distinct functional groups than the more traditional PCoA ordination. Ultimately, we identified four community ecotypes with distinct taxonomic and functional cut-offs driven by seasonality, water mass, and substrate turnover, highlighting the importance of succession in functional diversity for the central Arctic Ocean. These results reinforce ML dimensionality reduction as a meaningful translator in the mining of historical amplicon datasets to address modern mechanistic questions and potentially provide ’omics informed ecotype diversity to leverage in mechanistic biogeochemical models.

Arctic Ocean↗

Using Fractional Clock-Period Delays in Telemetry Arraying

A set of special digital all-pass finite-impulse- response (FIR) filters produces phase shifts equivalent to delays that equal fractions of the sampling or clock period of a telemetry-data-processing system. These filters have been used to enhance the arraying of telemetry signals that have been received at multiple ground stations from spacecraft (see figure). Somewhat more specifically, these filters have been used to align, in the time domain, the telemetry-data sequences received by the various antennas, in order to maximize the signal-to-noise ratio of the composite telemetric signal obtained by summing the signals received by the antennas. The term arraying in this context denotes a method of enhanced reception of telemetry signals in which several antennas are used to track a single spacecraft. Each antenna receives a signal that comprises a sum of telemetry data plus noise, and these sum data are sent to an arraying combiner for processing. Correlation is the means used to align the set of data from one antenna with that from another antenna. After the data from all the antennas have been aligned in the time domain, they are all added together, sample by sample.

Fort, David↗

Multi-omics data resource: Data package 24 (Pck024)

The data package consists of isolated pancreatic islets from 3 human donors treated with IL-1β, IFNγ or IL-1β + IFNγ for 6 h and IL-1β, IFNγ, IL-1β + IFNγ, IL-1β + IFNγ + NMMA or NMMA for 18 h and submitted for scRNA-seq. This study examines cytokine-stimulated changes in gene expression in human islets using single-cell RNA sequencing. Data contributors: Jennifer S Stancill & John A Corbett: Department of Biochemistry, Medical College of Wisconsin, Milwaukee, WI, USA Data repository: GSE251730 Publication: 10.1093/function/zqae015

Sarkar, Soumyadeep [Pacific Northwest National Lab↗

Genomics Study of Effect of Redox-Active Metalloporphyrin on Murine Retina During Spaceflight

Astronauts returning from spaceflight have experienced eye problems, which may decrease retinal performance and lead to long-term effects on visual acuity. This study leverages the collected data from spaceflown murine retinas that were treated with redox-active metalloporphyrin (BuOE) to mitigate spaceflight-induced changes. 10-week-old adult C57BL/6 male mice (n=5 in each of BuOE treated and saline control groups) were flown on Space-X 24 to the ISS national lab, kept in low earth orbit for 35 days and returned to Earth alive. Our analysis of RNA-sequencing data generated from subsequent murine retina tissues uncovered genes, pathways, and epigenetic modifications consistent with therapeutic potential of BuOE. For spaceflown murine samples, the treatment group show differentially expressed genes relative to saline controls that reached significance (adjusted p-value < 0.05) and included genes Gpx3 and Crhbp, which are related to protection against cell oxidative damage and cellular response to organonitrogen compounds. Ranked fold-changes from the same contrast were used for gene set enrichment analysis, which showed biological processes reaching significance (adjusted p-value < 0.05) including glutathione metabolic processes and cellular response to xenobiotic stimulus. The findings from this investigation have the potential to provide valuable insights into the molecular mechanisms underlying conditions like spaceflight associated neuro-ocular syndrome and assess the effectiveness of BuOE as a countermeasure for astronauts experiencing neuro-ophthalmic abnormalities, which can lead to long-term effects on visual acuity.

Machine Learning↗

Genomics Study of Effect of Redox-Active Metalloporphyrin on Murine Retina During Spaceflight

Astronauts returning from spaceflight have experienced eye problems, which may decrease retinal performance and lead to long-term effects on visual acuity. This study leverages the collected data from spaceflown murine retinas that were treated with redox-active metalloporphyrin (BuOE) to mitigate spaceflight-induced changes. 10-week-old adult C57BL/6 male mice (n=5 in each of BuOE treated and saline control groups) were flown on Space-X 24 to the ISS national lab, kept in low earth orbit for 35 days and returned to Earth alive. Our analysis of RNA-sequencing data generated from subsequent murine retina tissues uncovered genes, pathways, and epigenetic modifications consistent with therapeutic potential of BuOE. For spaceflown murine samples, the treatment group show differentially expressed genes relative to saline controls that reached significance (adjusted p-value < 0.05) and included genes Gpx3 and Crhbp, which are related to protection against cell oxidative damage and cellular response to organonitrogen compounds. Ranked fold-changes from the same contrast were used for gene set enrichment analysis, which showed biological processes reaching significance (adjusted p-value < 0.05) including glutathione metabolic processes and cellular response to xenobiotic stimulus. The findings from this investigation have the potential to provide valuable insights into the molecular mechanisms underlying conditions like spaceflight associated neuro-ocular syndrome and assess the effectiveness of BuOE as a countermeasure for astronauts experiencing neuro-ophthalmic abnormalities, which can lead to long-term effects on visual acuity.

Machine Learning↗

GL4U: Training the next generation of bioinformaticians, one omics datatype at a time

Spaceflight modifies gene expression in every organism examined to date, including humans. Understanding how these gene expression changes affect physiology is crucial for the development of countermeasures to enable long-duration manned missions. NASA’s GeneLab project provides researchers open access to multi-omics data, including genetic and gene expression data, from spaceflight experiments that can be mined to understand the effects of spaceflight on biological systems. To ensure new knowledge generation through data re-use, it is important to maximize the number of scientists who utilize GeneLab data. Training students on the GeneLab platform is the best way to create long-term adopters of this NASA database and its tools. Turning students into future instructors and advocates will also accelerate the dissemination of these data and tools to the broader scientific community. Therefore, in collaboration with the GeneLab Educational Working Group (EWG), GeneLab has created GeneLab for Colleges and Universities (GL4U). GL4U provides space biology-relevant training in bioinformatics to the next generation of scientists through direct and indirect approaches. The GeneLab team plans to host two annual data processing bootcamps, one for college-level students (direct) and one for college educators (indirect – training of trainers), in which participants learn to analyze GeneLab’s space-relevant omics data. During the bootcamp, educators will receive materials and training to enable them to run the bootcamp at their home institutions or alternatively to adapt the content to implement within existing courses, thereby extending the reach of this initiative. The GL4U direct training pilot program was conducted in June 2021 in collaboration with USRA and San Jose State University (SJSU). During the pilot, SJSU students participated in a week-long bootcamp consisting of space biology-specific lectures and hands-on instruction using Jupyter Notebooks to analyze RNA sequence data. This pilot demonstrates the capacity of GL4U for training young scientists and encouraging data re-use.

Jonathan Matthew Galazka↗

LinkFinder: An expert system that constructs phylogenic trees

An expert system has been developed using the C Language Integrated Production System (CLIPS) that automates the process of constructing DNA sequence based phylogenies (trees or lineages) that indicate evolutionary relationships. LinkFinder takes as input homologous DNA sequences from distinct individual organisms. It measures variations between the sequences, selects appropriate proportionality constants, and estimates the time that has passed since each pair of organisms diverged from a common ancestor. It then designs and outputs a phylogenic map summarizing these results. LinkFinder can find genetic relationships between different species, and between individuals of the same species, including humans. It was designed to take advantage of the vast amount of sequence data being produced by the Genome Project, and should be of value to evolution theorists who wish to utilize this data, but who have no formal training in molecular genetics. Evolutionary theory holds that distinct organisms carrying a common gene inherited that gene from a common ancestor. Homologous genes vary from individual to individual and species to species, and the amount of variation is now believed to be directly proportional to the time that has passed since divergence from a common ancestor. The proportionality constant must be determined experimentally; it varies considerably with the types of organisms and DNA molecules under study. Given an appropriate constant, and the variation between two DNA sequences, a simple linear equation gives the divergence time.

Inglehart, James↗

Emerging anomaly detection techniques for electronic health records: A survey

Background Anomaly detection in electronic health records (EHRs) is a cornerstone of biomedical informatics, with direct implications for patient safety, clinical decision-making, and the prevention of healthcare fraud. Once guided primarily by simple rule-based methods, the field has advanced rapidly, driven by increased computing power, richer and more detailed health data, and the rise of machine learning and deep learning techniques. The objective of this paper is to provide a comprehensive overview of modern approaches to detecting anomalies in EHRs, outlining their strengths, limitations, and relevance to key healthcare challenges. We review traditional statistical methods alongside newer ML- and DL-based strategies and hybrid models, with particular attention to how these techniques support transparency and build clinical trust. Methods This paper presents a thorough and critical survey through systematic review (PRISMA-based) of the latest anomaly detection strategies in time-sequence data domains within electronic health record systems. Results We explore a broad spectrum of methodologies, including statistical models, supervised and unsupervised learning approaches, hybrid frameworks, and state-of-the-art ML-based techniques that collectively advance the precision and scalability of detecting anomalies in complex clinical datasets. In addition to mapping current capabilities, we address the enduring challenges that hinder widespread implementation and provide a forward-looking perspective on the future of anomaly detection in the data-rich landscape of modern healthcare. Summary The advancement in AI-based approaches is reported along with the basic principles of the individual approaches and their applicability. The increased availability of high-quality data, advancements in DL approaches, and enhanced computation power are leading to more frequent adaptation of DL-based approaches. Emerging DL-based approaches that have been adapted in other domains or recently applied in the EHR domain are also discussed in detail. Although DL-based approaches can improve model predictions by incorporating comorbidities, their application is limited in low-frequency data domains (e.g., when the total available data remains in the single digits). Therefore, the user must carefully consider the application based on data availability.

Anomaly detection↗

Exploring Saccharomycotina Yeast Ecology Through an Ecological Ontology Framework

Yeasts in the subphylum Saccharomycotina are found across the globe in disparate ecosystems. A major aim of yeast research is to understand the diversity and evolution of ecological traits, such as carbon metabolic breadth, insect association, and cactophily. This includes studying aspects of ecological traits like genetic architecture or association with other phenotypic traits. Genomic resources in the Saccharomycotina have grown rapidly. Ecological data, however, are still limited for many species, especially those only known from species descriptions where usually only a limited number of strains are studied. Moreover, ecological information is recorded in natural language format limiting high throughput computational analysis. To address these limitations, we developed an ontological framework for the analysis of yeast ecology. A total of 1,088 yeast strains were added to the Ontology of Yeast Environments (OYE) and analyzed in a machine-learning framework to connect genotype to ecology. This framework is flexible and can be extended to additional isolates, species, or environmental sequencing data. Widespread adoption of OYE would greatly aid the study of macroecology in the Saccharomycotina subphylum.

59 BASIC BIOLOGICAL SCIENCES↗

Post-Flight Microbial Analysis of Samples from the International Space Station Water Recovery System and Oxygen Generation System

The Regenerative, Environmental Control and Life Support System (ECLSS) on the International Space Station (ISS) includes the the Water Recovery System (WRS) and the Oxygen Generation System (OGS). The WRS consists of a Urine Processor Assembly (UPA) and Water Processor Assembly (WPA). This report describes microbial characterization of wastewater and surface samples collected from the WRS and OGS subsystems, returned to KSC, JSC, and MSFC on consecutive shuttle flights (STS-129 and STS-130) in 2009-10. STS-129 returned two filters that contained fluid samples from the WPA Waste Tank Orbital Recovery Unit (ORU), one from the waste tank and the other from the ISS humidity condensate. Direct count by microscopic enumeration revealed 8.38 x 104 cells per mL in the humidity condensate sample, but none of those cells were recoverable on solid agar media. In contrast, 3.32 x lOs cells per mL were measured from a surface swab of the WRS waste tank, including viable bacteria and fungi recovered after S12 days of incubation on solid agar media. Based on rDNA sequencing and phenotypic characterization, a fungus recovered from the filter was determined to be Lecythophora mutabilis. The bacterial isolate was identified by rDNA sequence data to be Methylobacterium radiotolerans. Additional UPA subsystem samples were returned on STS-130 for analysis. Both liquid and solid samples were collected from the Russian urine container (EDV), Distillation Assembly (DA) and Recycle Filter Tank Assembly (RFTA) for post-flight analysis. The bacterium Pseudomonas aeruginosa and fungus Chaetomium brasiliense were isolated from the EDV samples. No viable bacteria or fungi were recovered from RFTA brine samples (N= 6), but multiple samples (N = 11) from the DA and RFTA were found to contain fungal and bacterial cells. Many recovered cells have been identified to genus by rDNA sequencing and carbon source utilization profiling (BiOLOG Gen III). The presence of viable bacteria and fungi from WRS and OGS subsystems demonstrates the need for continued monitoring of ECLSS during future ISS operations and investigation of advanced antimicrobial controls.

Birmele, Michele N.↗

Signatures of Mollicutes-related endobacteria in publicly available Mucoromycota genomes

ABSTRACT Mucoromycota fungi and their Mollicutes-related endobacteria (MRE) are an ideal system for studying bacterial–fungal interactions and evolution due to the long-term and intimate nature of their interactions. However, methods for detecting MRE face specific challenges due to the poor representation of MRE in sequencing databases coupled with the high sequence divergence of their genomes, making traditional similarity searches unreliable. This has precluded estimations on the diversity of MRE associated with Mucoromycota. To determine the prevalence of previously undetected MRE in fungal genome sequences, we scanned 389 Mucoromycota genome assemblies available from the National Center for Biotechnology Information for the presence of MRE sequences using publicly available tools to map contigs from fungal assemblies to publicly available MRE genomes. We demonstrate a higher diversity of MRE genomes than previously described in Mucoromycota and a lack of cophylogeny between MRE and the majority of their fungal hosts. This supports the late invasion hypothesis regarding MRE acquisition across most of the examined fungal families. In contrast with other Mucoromycota lineages, MRE from the Gigasporaceae displayed some degree of cophylogeny with their hosts, which may indicate that horizontal transmission is restricted between members of this family or that transmission is strictly vertical. These results underscore the need for a refined process to capture sequencing data from potential fungal endosymbionts to discern their evolution and transmission. Screens of fungal genomes for MRE can help improve the quality of fungal genome assemblies while identifying new MRE lineages to further test hypotheses on their origin and evolution. IMPORTANCE Mollicutes-related endobacteria (MRE) are obligate intracellular bacteria found within Mucoromycota fungi. Despite their frequent detection, MRE roles in host functioning are still unknown. Comparative genomic investigations can improve our understanding of the impact of MRE on their fungal hosts by identifying similarities and differences in MRE genome evolution. However, MRE genomes have only been assembled from a small fraction of Mucoromycota hosts. Here, we demonstrate that MRE can be present yet undetected in publicly available Mucoromycota genome assemblies. We use these newfound sequences to assess the broader diversity of MRE and their phylogenetic relationships with respect to their hosts. We demonstrate that publicly available tools can be used to extract novel MRE sequences from assembled fungal genomes leading to insights on MRE evolution. This work contributes to a greater understanding of the fungal microbiome, which is crucial to improving knowledge on the dynamics and impacts of fungi in microbial ecosystems.

59 BASIC BIOLOGICAL SCIENCES↗