Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Sequence analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Mutational analysis of photosystem I polypeptides in the cyanobacterium Synechocystis sp. PCC 6803. Targeted inactivation of psaI reveals the function of psaI in the structural organization of psaL

We cloned, characterized, and inactivated the psaI gene encoding a 4-kDa hydrophobic subunit of photosystem I from the cyanobacterium Synechocystis sp. PCC 6803. The psaI gene is located 90 base pairs downstream from psaL, and is transcribed on 0.94- and 0.32-kilobase transcripts. To identify the function of PsaI, we generated a cyanobacterial strain in which psaI has been interrupted by a gene for chloramphenicol resistance. The wild-type and the mutant cells showed comparable rates of photoautotrophic growth at 25 degrees C. However, the mutant cells grew slower and contained less chlorophyll than the wild-type cells, when grown at 40 degrees C. The PsaI-less membranes from cells grown at either temperature showed a small decrease in NADP+ photoreduction rate when compared to the wild-type membranes. Inactivation of psaI led to an 80% decrease in the PsaL level in the photosynthetic membranes and to a complete loss of PsaL in the purified photosystem I preparations, but had little effect on the accumulation of other photosystem I subunits. Upon solubilization with nonionic detergents, photosystem I trimers could be obtained from the wild-type, but not from the PsaI-less membranes. The PsaI-less photosystem I monomers did not contain detectable levels of PsaL. Therefore, a structural interaction between PsaL and PsaI may stabilize the association of PsaL with the photosystem I core. PsaL in the wild-type and PsaI-less membranes showed equal resistance to removal by chaotropic agents. However, PsaL in the PsaI-less strain exhibited an increased susceptibility to proteolysis. From these data, we conclude that PsaI has a crucial role in aiding normal structural organization of PsaL within the photosystem I complex and the absence of PsaI alters PsaL organization, leading to a small, but physiologically significant, defect in photosystem I function.

NASA Discipline Cell Biology↗

The Arabidopsis SKU5 gene encodes an extracellular glycosyl phosphatidylinositol-anchored glycoprotein involved in directional root growth

To investigate how roots respond to directional cues, we characterized a T-DNA-tagged Arabidopsis mutant named sku5 in which the roots skewed and looped away from the normal downward direction of growth on inclined agar surfaces. sku5 roots and etiolated hypocotyls were slightly shorter than normal and exhibited a counterclockwise (left-handed) axial rotation bias. The surface-dependent skewing phenotype disappeared when the roots penetrated the agar surface, but the axial rotation defect persisted, revealing that these two directional growth processes are separable. The SKU5 gene belongs to a 19-member gene family designated SKS (SKU5 Similar) that is related structurally to the multiple-copper oxidases ascorbate oxidase and laccase. However, the SKS proteins lack several of the conserved copper binding motifs characteristic of copper oxidases, and no enzymatic function could be assigned to the SKU5 protein. Analysis of plants expressing SKU5 reporter constructs and protein gel blot analysis showed that SKU5 was expressed most strongly in expanding tissues. SKU5 was glycosylated and modified by glycosyl phosphatidylinositol and localized to both the plasma membrane and the cell wall. Our observations suggest that SKU5 affects two directional growth processes, possibly by participating in cell wall expansion.

NASA Discipline Plant Biology↗

Mass analysis for the Space Station ECLSS using the balance spreadsheet method

The balance spreadsheet method is applied to mass analysis of the Environmental Control and Life Support System (ECLSS). The spreadsheet layout reduces the complexity of the ECLSS analysis by concisely defining the sources, sinks, and net changes in mass for each fluid. The analysis method is illustrated by using information from the latest Space Station ECLSS Architectural Control Documents and a given Space Station assembly sequence. The analysis results are plotted and discussed.

Chu, Wen-Ho↗

Implied alignment: a synapomorphy-based multiple-sequence alignment method and its use in cladogram search

A method to align sequence data based on parsimonious synapomorphy schemes generated by direct optimization (DO; earlier termed optimization alignment) is proposed. DO directly diagnoses sequence data on cladograms without an intervening multiple-alignment step, thereby creating topology-specific, dynamic homology statements. Hence, no multiple-alignment is required to generate cladograms. Unlike general and globally optimal multiple-alignment procedures, the method described here, implied alignment (IA), takes these dynamic homologies and traces them back through a single cladogram, linking the unaligned sequence positions in the terminal taxa via DO transformation series. These "lines of correspondence" link ancestor-descendent states and, when displayed as linearly arrayed columns without hypothetical ancestors, are largely indistinguishable from standard multiple alignment. Since this method is based on synapomorphy, the treatment of certain classes of insertion-deletion (indel) events may be different from that of other alignment procedures. As with all alignment methods, results are dependent on parameter assumptions such as indel cost and transversion:transition ratios. Such an IA could be used as a basis for phylogenetic search, but this would be questionable since the homologies derived from the implied alignment depend on its natal cladogram and any variance, between DO and IA + Search, due to heuristic approach. The utility of this procedure in heuristic cladogram searches using DO and the improvement of heuristic cladogram cost calculations are discussed. c2003 The Willi Hennig Society. Published by Elsevier Science (USA). All rights reserved.

Non-NASA Center↗

Generalized Levy-walk model for DNA nucleotide sequences

We propose a generalized Levy walk to model fractal landscapes observed in noncoding DNA sequences. We find that this model provides a very close approximation to the empirical data and explains a number of statistical properties of genomic DNA sequences such as the distribution of strand-biased regions (those with an excess of one type of nucleotide) as well as local changes in the slope of the correlation exponent alpha. The generalized Levy-walk model simultaneously accounts for the long-range correlations in noncoding DNA sequences and for the apparently paradoxical finding of long subregions of biased random walks (length lj) within these correlated sequences. In the generalized Levy-walk model, the lj are chosen from a power-law distribution P(lj) varies as lj(-mu). The correlation exponent alpha is related to mu through alpha = 2-mu/2 if 2 < mu < 3. The model is consistent with the finding of "repetitive elements" of variable length interspersed within noncoding DNA.

NASA Discipline Number 14-10↗

A NASTRAN DMAP alter for linear buckling analysis under dynamic loading

A modification to the NASTRAN solution sequence for transient analysis with direct time integration (COSMIC NASTRAN rigid format 9) was developed and incorporated into a DMAP alter. This DMAP alter calculates the buckling stability of a dynamically loaded structure, and is used to predict the onset of structural buckling under stress-wave loading conditions. The modified solution sequence incorporates the linear buckling analysis capability (rigid format 5) of NASTRAN into the existing Transient solution rigid format in such a way as to provide a time dependent eigensolution which is used to assess the buckling stability of the structure as it responds to the impulsive load. As a demonstration of the validity of this modified solution procedure, the dynamic buckling of a prismatic bar subjected to an impulsive longitudinal compression is analyzed and compared to the known theoretical solution. In addition, a dynamic buckling analysis is performed for the analytically less tractable problem of the localized dynamic buckling of an initially flawed composite laminate under transverse impact loading. The addition of this DMAP alter to the transient solution sequence in NASTRAN facilitates the computational prediction of both the time at which the onset of dynamic buckling occurs in an impulsively loaded structure, and the dynamic buckling mode shapes of that structure.

Aiello, Robert A.↗

Cyote-attack Chain Estimator

Attack Chain Estimator (ACE) Application Overview The Attack Chain Estimator (ACE) Application is a sophisticated tool designed for the ingestion, classification, sequencing, and enrichment of cybersecurity threat reports. This application leverages advanced machine learning models and extensive historical data to provide comprehensive insights into cyber threats, specifically targeting Industrial Control Systems (ICS). Purpose The primary functions of the ACE Application include: Ingestion of Cybersecurity Threat Reporting: Capable of ingesting text-based threat reports in markdown or text file format. Supports ingestion of structured data from other sources in STIX/JSON format. Classification of Report’s Text-Based Events: Utilizes a DeBERTa classifier, specifically trained on cybersecurity data, to map the events to MITRE ATT&CK for ICS Tactics and Techniques. Classification is performed using multiple Jupyter notebooks and machine learning workflows hosted as FastAPI microservices: regex_data deberta_base_35_train_hft_classifier_mlflow.ipynb hft_regex_classifier_mlflow.ipynb param_train_hft_classifier_mlflow.ipynb regex_tactic_tech.ipynb Ordering of Tactics, Techniques, and Observable Events: Sequences the identified tactics, techniques, and events to form a coherent attack chain. Enrichment with Historical Attack Chain Details: Enhances the attack chain with details from historical attacks using a Markov model developed from CyOTE Precursor Analysis Report data. The Markov model is available as a FastAPI endpoint for seamless integration. Enrichment with Adversary Emulation Capabilities Data: Integrates adversary emulation capabilities data using MITRE Caldera for OT adversary abilities UUIDs. Export of Output Files: Provides options to export the enriched attack chain in JSON or CSV formats. Routing of Output to Other Applications: Facilitates routing of output to various platforms and applications, including: Threat Intelligence Platforms COREII Scout for Threat Intelligence Analysis COREII Modeling and Simulation for Adversary Emulation Technical Description The ACE Application is an advanced cybersecurity tool designed to provide detailed threat analysis and sequence generation. It is built on a robust architecture that integrates natural language processing, machine learning, and historical data modeling. Key Components: Data Ingestion Module: Handles the input of threat reports and data from various formats, ensuring flexibility in data sources. Classification Engine: Employs DeBERTa-based classifiers hosted as FastAPI microservices to analyze and classify threat report events in accordance with the MITRE ATT&CK framework for ICS. Sequence Generator: Orders the classified events into a logical attack chain, providing clear insight into the sequence of tactics and techniques used in the threat. Enrichment Engine: Integrates historical data and adversary emulation capabilities to enhance the attack chain with valuable context and additional details. The historical data enrichment is powered by a Markov model, which is available as a FastAPI endpoint. Export and Routing Module: Facilitates the export of the enriched attack chain in multiple formats and routes the output to designated applications for further analysis or emulation.

Paul, Tony [Idaho National Laboratory (INL), Idaho↗

Behavior of Single-Line-Ground Faults in Inverter-Based Resource Dominated Grids Explained

It has been observed by protection engineers that it is difficult for a protective relay to identify the faulted phase during a single-line-ground (SLG) fault in a power system with a high ingression of inverter-based resources (IBR) using currents (phase or sequence). Further studies using electromagnetic transient (EMT) simulation show that the initial operating conditions of the IBRs influence the response of phase currents during an SLG fault. In this letter, we conduct a quantitative analysis using sequence components. We find that the pre-fault condition determines the relative position of the current contributed by the grid versus that from the IBR, and further dictates which phase has the largest magnitude during an SLG condition. Finally, this finding is further verified by the EMT simulation results.

electromagnetic transient simulation↗

VIZARD: analysis of Affymetrix Arabidopsis GeneChip data

SUMMARY: The Affymetrix GeneChip Arabidopsis genome array has proved to be a very powerful tool for the analysis of gene expression in Arabidopsis thaliana, the most commonly studied plant model organism. VIZARD is a Java program created at the University of California, Berkeley, to facilitate analysis of Arabidopsis GeneChip data. It includes several integrated tools for filtering, sorting, clustering and visualization of gene expression data as well as tools for the discovery of regulatory motifs in upstream sequences. VIZARD also includes annotation and upstream sequence databases for the majority of genes represented on the Affymetrix Arabidopsis GeneChip array. AVAILABILITY: VIZARD is available free of charge for educational, research, and not-for-profit purposes, and can be downloaded at http://www.anm.f2s.com/research/vizard/ CONTACT: moseyko@uclink4.berkeley.edu.

Non-NASA Center↗

Signature lipids and stable carbon isotope analyses of Octopus Spring hyperthermophilic communities compared with those of Aquificales representatives

The molecular and isotopic compositions of lipid biomarkers of cultured Aquificales genera have been used to study the community and trophic structure of the hyperthermophilic pink streamers and vent biofilm from Octopus Spring. Thermocrinis ruber, Thermocrinis sp. strain HI 11/12, Hydrogenobacter thermophilus TK-6, Aquifex pyrophilus, and Aquifex aeolicus all contained glycerol-ether phospholipids as well as acyl glycerides. The n-C(20:1) and cy-C(21) fatty acids dominated all of the Aquificales, while the alkyl glycerol ethers were mainly C(18:0). These Aquificales biomarkers were major constituents of the lipid extracts of two Octopus Spring samples, a biofilm associated with the siliceous vent walls, and the well-known pink streamer community (PSC). Both the biofilm and the PSC contained mono- and dialkyl glycerol ethers in which C(18) and C(20) alkyl groups were prevalent. Phospholipid fatty acids included both the Aquificales n-C(20:1) and cy-C(21), plus a series of iso-branched fatty acids (i-C(15:0) to i-C(21:0)), indicating an additional bacterial component. Biomass and lipids from the PSC were depleted in (13)C relative to source water CO(2) by 10.9 and 17.2 per thousand, respectively. The C(20-21) fatty acids of the PSC were less depleted than the iso-branched fatty acids, 18.4 and 22.6 per thousand, respectively. The biomass of T. ruber grown on CO(2) was depleted in (13)C by only 3.3 per thousand relative to C source. In contrast, biomass was depleted by 19.7 per thousand when formate was the C source. Independent of carbon source, T. ruber lipids were heavier than biomass (+1.3 per thousand). The depletion in the C(20-21) fatty acids from the PSC indicates that Thermocrinis biomass must be similarly depleted and too light to be explained by growth on CO(2). Accordingly, Thermocrinis in the PSC is likely to have utilized formate, presumably generated in the spring source region.

Carbon Isotopes/analysis↗

Sequence, overproduction and purification of Vibrio proteolyticus ribosomal protein L18 for in vitro and in vivo studies

A strategy suggested by comparative genomic studies was used to amplify the entire Vibrio proteolyticus (Vp) gene for ribosomal protein L18. Vp L18 and its flanking regions were sequenced and compared with the deduced amino acid (aa) sequences of other known L18 proteins. A 26-aa residue segment at the carboxy terminus contains many strongly conserved residues and may be critical for the L18 interaction with 5S rRNA. This approach should allow rapid characterization of L18 from large numbers of bacteria. Both Vp L18 and Escherichia coli (Ec) L18 were overproduced and purified using a T7 expression vector which fuses an N-terminal peptide segment (His-tag) containing 6 histidine residues to the recombinant protein. The purified fusion proteins, Vp His::L18 and Ec His::L18, were both found to bind to either the Vp 5S or Ec 5S rRNAs in vitro. Vp His::L18 protein was also shown to incorporate into Ec ribosomes in vivo. This His-tag strategy likely will have general applicability for the study of ribosomal proteins in vitro and in vivo.

Non-NASA Center↗

Atrogin-1, a muscle-specific F-box protein highly expressed during muscle atrophy

Muscle wasting is a debilitating consequence of fasting, inactivity, cancer, and other systemic diseases that results primarily from accelerated protein degradation by the ubiquitin-proteasome pathway. To identify key factors in this process, we have used cDNA microarrays to compare normal and atrophying muscles and found a unique gene fragment that is induced more than ninefold in muscles of fasted mice. We cloned this gene, which is expressed specifically in striated muscles. Because this mRNA also markedly increases in muscles atrophying because of diabetes, cancer, and renal failure, we named it atrogin-1. It contains a functional F-box domain that binds to Skp1 and thereby to Roc1 and Cul1, the other components of SCF-type Ub-protein ligases (E3s), as well as a nuclear localization sequence and PDZ-binding domain. On fasting, atrogin-1 mRNA levels increase specifically in skeletal muscle and before atrophy occurs. Atrogin-1 is one of the few examples of an F-box protein or Ub-protein ligase (E3) expressed in a tissue-specific manner and appears to be a critical component in the enhanced proteolysis leading to muscle atrophy in diverse diseases.

NASA Discipline Musculoskeletal↗

Performance analysis of multiple PRF technique for ambiguity resolution

For short wavelength spaceborne synthetic aperture radar (SAR), ambiguity in Doppler centroid estimation occurs when the azimuth squint angle uncertainty is larger than the azimuth antenna beamwidth. Multiple pulse recurrence frequency (PRF) hopping is a technique developed to resolve the ambiguity by operating the radar in different PRF's in the pre-imaging sequence. Performance analysis results of the multiple PRF technique are presented, given the constraints of the attitude bound, the drift rate uncertainty, and the arbitrary numerical values of PRF's. The algorithm performance is derived in terms of the probability of correct ambiguity resolution. Examples, using the Shuttle Imaging Radar-C (SIR-C) and X-SAR parameters, demonstrate that the probability of correct ambiguity resolution obtained by the multiple PRF technique is greater than 95 percent and 80 percent for the SIR-C and X-SAR applications, respectively. The success rate is significantly higher than that achieved by the range cross correlation technique.

Chang, C. Y.↗

Fractal landscapes in biological systems: long-range correlations in DNA and interbeat heart intervals

Here we discuss recent advances in applying ideas of fractals and disordered systems to two topics of biological interest, both topics having common the appearance of scale-free phenomena, i.e., correlations that have no characteristic length scale, typically exhibited by physical systems near a critical point and dynamical systems far from equilibrium. (i) DNA nucleotide sequences have traditionally been analyzed using models which incorporate the possibility of short-range nucleotide correlations. We found, instead, a remarkably long-range power law correlation. We found such long-range correlations in intron-containing genes and in non-transcribed regulatory DNA sequences as well as intragenomic DNA, but not in cDNA sequences or intron-less genes. We also found that the myosin heavy chain family gene evolution increases the fractal complexity of the DNA landscapes, consistent with the intron-late hypothesis of gene evolution. (ii) The healthy heartbeat is traditionally thought to be regulated according to the classical principle of homeostasis, whereby physiologic systems operate to reduce variability and achieve an equilibrium-like state. We found, however, that under normal conditions, beat-to-beat fluctuations in heart rate display long-range power law correlations.

Non-NASA Center↗

A NASTRAN DMAP alter for linear buckling analysis under dynamic loading

A unique modification to the NASTRAN solution sequence for transient analysis with direct time integration (COSMIC NASTRAN rigid format 9) was developed and incorporated into a DMAP alter. This DMAP alter calculates the buckling stability of a dynamically loaded structure, and is used to predict the onset of structural buckling under stress wave loading conditions. The modified solution sequence incorporates the linear buckling analysis capability (rigid format 5) of NASTRAN into the existing Transient solution rigid format in such a way as to provide a time dependent eigensolution which is used to assess the buckling stability of the structure as it responds to the impulsive load. As a demonstration of the validity of this modified solution procedure, the dynamic buckling of a prismatic bar subjected to an impulsive longitudinal compression is analyzed and compared to the known theoretical solution. In addition, a dynamic buckling analysis is performed for the analytically less tractable problem of the localized dynamic buckling of an initially flawed composite laminate under transverse impact loading. The addition of this DMAP alter to the transient solution sequence in NASTRAN facilitates the prediction of both time and mode of buckling.

Aiello, Robert A.↗

Scaling features of noncoding DNA

We review evidence supporting the idea that the DNA sequence in genes containing noncoding regions is correlated, and that the correlation is remarkably long range--indeed, base pairs thousands of base pairs distant are correlated. We do not find such a long-range correlation in the coding regions of the gene, and utilize this fact to build a Coding Sequence Finder Algorithm, which uses statistical ideas to locate the coding regions of an unknown DNA sequence. Finally, we describe briefly some recent work adapting to DNA the Zipf approach to analyzing linguistic texts, and the Shannon approach to quantifying the "redundancy" of a linguistic text in terms of a measurable entropy function, and reporting that noncoding regions in eukaryotes display a larger redundancy than coding regions. Specifically, we consider the possibility that this result is solely a consequence of nucleotide concentration differences as first noted by Bonhoeffer and his collaborators. We find that cytosine-guanine (CG) concentration does have a strong "background" effect on redundancy. However, we find that for the purine-pyrimidine binary mapping rule, which is not affected by the difference in CG concentration, the Shannon redundancy for the set of analyzed sequences is larger for noncoding regions compared to coding regions.

Non-NASA Center↗

Direct profiling of environmental microbial populations by thermal dissociation analysis of native rRNAs hybridized to oligonucleotide microarrays

Oligonucleotide microarrays were used to profile directly extracted rRNA from environmental microbial populations without PCR amplification. In our initial inspection of two distinct estuarine study sites, the hybridization patterns were reproducible and varied between estuarine sediments of differing salinities. The determination of a thermal dissociation curve (i.e., melting profile) for each probe-target duplex provided information on hybridization specificity, which is essential for confirming adequate discrimination between target and nontarget sequences.

Non-NASA Center↗

Dynamic stress analysis of smooth and notched fiber composite flexural specimens

A detailed analysis of the dynamic stress field in smooth and notched fiber composite (Charpy-type) specimens is reported in this paper. The analysis is performed with the aid of the direct transient response analysis solution sequence of MSC/NASTRAN. Three unidirectional composites were chosen for the study. They are S-Glass/Epoxy, Kevlar/Epoxy and T-300/Epoxy composite systems. The specimens are subjected to an impact load which is modeled as a triangular impulse with a maximum of 2000 lb and a duration of 1 ms. The results are compared with those of static analysis of the specimens subjected to a peak load of 2000 lb. For the geometry and type of materials studied, the static analysis results gave close conservative estimates for the dynamic stresses. Another interesting inference from the study is that the impact induced effects are felt by S-Glass/Epoxy specimens sooner than Kevlar/Epoxy or T-300/Epoxy specimens.

Murthy, P. L. N.↗