Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “sequence alignment”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Testing putative hemichordate homologues of the chordate dorsal nervous system and endostyle: expression of NK2.1 (TTF-1) in the acorn worm Ptychodera flava (Hemichordata, Ptychoderidae)

Recent phylogenetic investigations have confirmed that hemichordates and echinoderms are sister taxa. However, hemichordates share several cardinal characterstics with chordates and are thus an important taxon for testing hypotheses of homology between key chordate characters and their putative hemichordate antecedents. The chordate dorsal nervous system (DNS) and endostyle are intriguing characters because both hemichordate larval and adult structures have been hypothesized as homologues. This study attempts to test these purported homologies through examination of the expression pattem of a Ptychodera flava NK2 gene, PfNK2.1, because this gene is expressed both in the DNS and endostyle/thyroid in a wide range of chordate taxa. We found that PfNK2.1 is expressed in both neuronal and pharyngeal structures, but its expression pattem is broken up into distinct embryonic and juvenile phases. During embryogenesis, PfNK2.1 is expressed in the apical ectoderm, with transcripts later detected in presumable neuronal structures, including the apical organ and ciliated feeding band. In the developing juvenile we detected PfNK2.1 signal throughout the pharynx, including the stomochord, and later in the hindgut. We conclude that the similar utilization of NK2.1 in apical organ development and chordate DNS is probably due to a more general role for NK2.1 in neurogenesis and that hemichordates do not possess a homologue of the chordate DNS. In addition, we conclude that P. flava most likely does not possess a true endostyle; rather during the evolution of the endostyle NK2.1 was recruited from its more general role in pharynx development.

NASA Discipline Evolutionary Biology↗

A calmodulin-binding/CGCG box DNA-binding protein family involved in multiple signaling pathways in plants

We reported earlier that the tobacco early ethylene-responsive gene NtER1 encodes a calmodulin-binding protein (Yang, T., and Poovaiah, B. W. (2000) J. Biol. Chem. 275, 38467-38473). Here we demonstrate that there is one NtER1 homolog as well as five related genes in Arabidopsis. These six genes are rapidly and differentially induced by environmental signals such as temperature extremes, UVB, salt, and wounding; hormones such as ethylene and abscisic acid; and signal molecules such as methyl jasmonate, H(2)O(2), and salicylic acid. Hence, they were designated as AtSR1-6 (Arabidopsis thaliana signal-responsive genes). Ca(2+)/calmodulin binds to all AtSRs, and their calmodulin-binding regions are located on a conserved basic amphiphilic alpha-helical motif in the C terminus. AtSR1 targets the nucleus and specifically recognizes a novel 6-bp CGCG box (A/C/G)CGCG(G/T/C). The multiple CGCG cis-elements are found in promoters of genes such as those involved in ethylene signaling, abscisic acid signaling, and light signal perception. The DNA-binding domain in AtSR1 is located on the N-terminal 146 bp where all AtSR1-related proteins share high similarity but have no similarity to other known DNA-binding proteins. The calmodulin-binding nuclear proteins isolated from wounded leaves exhibit specific CGCG box DNA binding activities. These results suggest that the AtSR gene family encodes a family of calmodulin-binding/DNA-binding proteins involved in multiple signal transduction pathways in plants.

Non-NASA Center↗

Hydrogen peroxide homeostasis: activation of plant catalase by calcium/calmodulin

Environmental stimuli such as UV, pathogen attack, and gravity can induce rapid changes in hydrogen peroxide (H(2)O(2)) levels, leading to a variety of physiological responses in plants. Catalase, which is involved in the degradation of H(2)O(2) into water and oxygen, is the major H(2)O(2)-scavenging enzyme in all aerobic organisms. A close interaction exists between intracellular H(2)O(2) and cytosolic calcium in response to biotic and abiotic stresses. Studies indicate that an increase in cytosolic calcium boosts the generation of H(2)O(2). Here we report that calmodulin (CaM), a ubiquitous calcium-binding protein, binds to and activates some plant catalases in the presence of calcium, but calcium/CaM does not have any effect on bacterial, fungal, bovine, or human catalase. These results document that calcium/CaM can down-regulate H(2)O(2) levels in plants by stimulating the catalytic activity of plant catalase. Furthermore, these results provide evidence indicating that calcium has dual functions in regulating H(2)O(2) homeostasis, which in turn influences redox signaling in response to environmental signals in plants.

NASA Discipline Plant Biology↗

Phylogenetic diversity in the genus Bacillus as seen by 16S rRNA sequencing studies

Comparative sequence analysis of 16S ribosomal (r)RNAs or DNAs of Bacillus alvei, B. laterosporus, B. macerans, B. macquariensis, B. polymyxa and B. stearothermophilus revealed the phylogenetic diversity of the genus Bacillus. Based on the presently available data set of 16S rRNA sequences from bacilli and relatives at least four major "Bacillus clusters" can be defined: a "Bacillus subtilis cluster" including B. stearothermophilus, a "B. brevis cluster" including B. laterosporus, a "B. alvei cluster" including B. macerans, B. maquariensis and B. polymyxa and a "B. cycloheptanicus branch".

NASA Discipline Number 52-30↗

Differential assembly of alpha- and gamma-filagenins into thick filaments in Caenorhabditis elegans

Muscle thick filaments are highly organized supramolecular assemblies of myosin and associated proteins with lengths, diameters and flexural rigidities characteristic of their source. The cores of body wall muscle thick filaments of the nematode Caenorhabditis elegans are tubular structures of paramyosin sub-filaments coupled by filagenins and have been proposed to serve as templates for the assembly of native thick filaments. We have characterized alpha- and gamma-filagenins, two novel proteins of the cores with calculated molecular masses of 30,043 and 19,601 and isoelectric points of 10.52 and 11.49, respectively. Western blot and immunoelectron microscopy using affinity-purified antibodies confirmed that the two proteins are core components. Immunoelectron microscopy of the cores revealed that they assemble with different periodicities. Immunofluorescence microscopy showed that alpha-filagenin is localized in the medial regions of the A-bands of body wall muscle cells whereas gamma-filagenin is localized in the flanking regions, and that alpha-filagenin is expressed in 1.5-twofold embryos while gamma-filagenin becomes detectable only in late vermiform embryos. The expression of both proteins continues throughout later stages of development. C. elegans body wall muscle thick filaments of these developmental stages have distinct lengths. Our results suggest that the differential assembly of alpha- and gamma-filagenins into thick filaments of distinct lengths may be developmentally regulated.

NASA Discipline Musculoskeletal↗

A calcium-dependent protein kinase can inhibit a calmodulin-stimulated Ca2+ pump (ACA2) located in the endoplasmic reticulum of Arabidopsis

The magnitude and duration of a cytosolic Ca(2+) release can potentially be altered by changing the rate of Ca(2+) efflux. In plant cells, Ca(2+) efflux from the cytoplasm is mediated by H(+)/Ca(2+)-antiporters and two types of Ca(2+)-ATPases. ACA2 was recently identified as a calmodulin-regulated Ca(2+)-pump located in the endoplasmic reticulum. Here, we show that phosphorylation of its N-terminal regulatory domain by a Ca(2+)-dependent protein kinase (CDPK isoform CPK1), inhibits both basal activity ( approximately 10%) and calmodulin stimulation ( approximately 75%), as shown by Ca(2+)-transport assays with recombinant enzyme expressed in yeast. A CDPK phosphorylation site was mapped to Ser(45) near a calmodulin binding site, using a fusion protein containing the N-terminal domain as an in vitro substrate for a recombinant CPK1. In a full-length enzyme, an Ala substitution for Ser(45) (S45/A) completely blocked the observed CDPK inhibition of both basal and calmodulin-stimulated activities. An Asp substitution (S45/D) mimicked phosphoinhibition, indicating that a negative charge at this position is sufficient to account for phosphoinhibition. Interestingly, prior binding of calmodulin blocked phosphorylation. This suggests that, once ACA2 binds calmodulin, its activation state becomes resistant to phosphoinhibition. These results support the hypothesis that ACA2 activity is regulated as the balance between the initial kinetics of calmodulin stimulation and CDPK inhibition, providing an example in plants for a potential point of crosstalk between two different Ca(2+)-signaling pathways.

NASA Discipline Plant Biology↗

Intramolecular activation of a Ca(2+)-dependent protein kinase is disrupted by insertions in the tether that connects the calmodulin-like domain to the kinase

Ca(2+)-dependent protein kinases (CDPK) have a calmodulin-like domain (CaM-LD) tethered to the C-terminal end of the kinase. Activation is proposed to involve intramolecular binding of the CaM-LD to a junction sequence that connects the CaM-LD to the kinase domain. Consistent with this model, a truncated CDPK (DeltaNC) in which the CaM-LD has been deleted can be activated in a bimolecular interaction with an isolated CaM-LD or calmodulin, similar to the activation of a calmodulin-dependent protein kinase (CaMK) by calmodulin. Here we provide genetic evidence that this bimolecular activation requires a nine-residue binding segment from F436 to I444 (numbers correspond to CPK-1 accession number L14771). Two mutations at either end of this core segment (F436/A and VI444/AA) severely disrupted bimolecular activation, whereas flanking mutations had only minor effects. Intramolecular activation of a full-length kinase was also disrupted by a VI444/AA mutation, but surprisingly not by a F436/A mutation (at the N-terminal end of the binding site). Interestingly, intramolecular but not bimolecular activation was disrupted by insertion mutations placed immediately downstream of I444. To show that mutant enzymes were not misfolded, latent kinase activity was stimulated through binding of an antijunction antibody. Results here support a model of intramolecular activation in which the tether (A445 to G455) that connects the CaM-LD to the kinase provides an important structural constraint and is not just a simple flexible connection.

NASA Discipline Plant Biology↗

ARG1 (altered response to gravity) encodes a DnaJ-like protein that potentially interacts with the cytoskeleton

Gravitropism allows plant organs to direct their growth at a specific angle from the gravity vector, promoting upward growth for shoots and downward growth for roots. Little is known about the mechanisms underlying gravitropic signal transduction. We found that mutations in the ARG1 locus of Arabidopsis thaliana alter root and hypocotyl gravitropism without affecting phototropism, root growth responses to phytohormones or inhibitors of auxin transport, or starch accumulation. The positional cloning of ARG1 revealed a DnaJ-like protein containing a coiled-coil region homologous to coiled coils found in cytoskeleton-interacting proteins. These data suggest that ARG1 participates in a gravity-signaling process involving the cytoskeleton. A combination of Northern blot studies and analysis of ARG1-GUS fusion-reporter expression in transgenic plants demonstrated that ARG1 is expressed in all organs. Ubiquitous ARG1 expression in Arabidopsis and the identification of an ortholog in Caenorhabditis elegans suggest that ARG1 is involved in other essential processes.

Non-NASA Center↗

Characterization of the proteins comprising the integral matrix of Strongylocentrotus purpuratus embryonic spicules

In the present study, we enumerate and characterize the proteins that comprise the integral spicule matrix of the Strongylocentrotus purpuratus embryo. Two-dimensional gel electrophoresis of [35S]methionine radiolabeled spicule matrix proteins reveals that there are 12 strongly radiolabeled spicule matrix proteins and approximately three dozen less strongly radiolabeled spicule matrix proteins. The majority of the proteins have acidic isoelectric points; however, there are several spicule matrix proteins that have more alkaline isoelectric points. Western blotting analysis indicates that SM50 is the spicule matrix protein with the most alkaline isoelectric point. In addition, two distinct SM30 proteins are identified in embryonic spicules, and they have apparent molecular masses of approximately 43 and 46 kDa. Comparisons between embryonic spicule matrix proteins and adult spine integral matrix proteins suggest that the embryonic 43-kDa SM30 protein is an embryonic isoform of SM30. An adult 49-kDa spine matrix protein is also identified as a possible adult isoform of SM30. Analysis of the SM30 amino acid sequences indicates that a portion of SM30 proteins is very similar to the carbohydrate recognition domain of C-type lectin proteins.

Non-NASA Center↗

Three-dimensional structure of Schistosoma japonicum glutathione S-transferase fused with a six-amino acid conserved neutralizing epitope of gp41 from HIV

The 3-dimensional crystal structure of glutathione S-transferase (GST) of Schistosoma japonicum (Sj) fused with a conserved neutralizing epitope on gp41 (glycoprotein, 41 kDa) of human immunodeficiency virus type 1 (HIV-1) (Muster T et al., 1993, J Virol 67:6642-6647) was determined at 2.5 A resolution. The structure of the 3-3 isozyme rat GST of the mu gene class (Ji X, Zhang P, Armstrong RN, Gilliland GL, 1992, Biochemistry 31:10169-10184) was used as a molecular replacement model. The structure consists of a 4-stranded beta-sheet and 3 alpha-helices in domain 1 and 5 alpha-helices in domain 2. The space group of the Sj GST crystal is P4(3)2(1)2, with unit cell dimensions of a = b = 94.7 A, and c = 58.1 A. The crystal has 1 GST monomer per asymmetric unit, and 2 monomers that form an active dimer are related by crystallographic 2-fold symmetry. In the binding site, the ordered structure of reduced glutathione is observed. The gp41 peptide (Glu-Leu-Asp-Lys-Trp-Ala) fused to the C-terminus of Sj GST forms a loop stabilized by symmetry-related GSTs. The Sj GST structure is compared with previously determined GST structures of mammalian gene classes mu, alpha, and pi. Conserved amino acid residues among the 4 GSTs that are important for hydrophobic and hydrophilic interactions for dimer association and glutathione binding are discussed.

Schistosoma japonicum/enzymology/genetics/immunolo↗

A full-coordinate model of the polymerase domain of HIV-1 reverse transcriptase and its interaction with a nucleic acid substrate

We present a full-coordinate model of residues 1-319 of the polymerase domain of HIV-I reverse transcriptase. This model was constructed from the x-ray crystallographic structure of Jacobo-Molina et al. (Jacobo-Molina et al., P.N.A.S. USA 90, 6320-6324 (1993)) which is currently available to the degree of C-coordinates. The backbone and side-chain atoms were constructed using the MAXSPROUT suite of programs (L. Holm and C. Sander, J. Mol. Biol. 218, 183-194 (1991)) and refined through molecular modeling. A seven base pair A-form dsDNA was positioned in the nucleic acid binding cleft to represent the template-primer complex. The orientation of the template-primer complex in the nucleic acid binding cleft was guided by the positions of phosphorus atoms in the crystal structure.

Non-NASA Center↗

The sequence, and its evolutionary implications, of a Thermococcus celer protein associated with transcription

Through random search, a gene from Thermococcus celer has been identified and sequenced that appears to encode a transcription-associated protein (110 amino acid residues). The sequence has clear homology to approximately the last half of an open reading frame reported previously for Sulfolobus acidocaldarius [Langer, D. & Zillig, W. (1993) Nucleic Acids Res. 21, 2251]. The protein translations of these two archaeal genes in turn are homologs of a small subunit found in eukaryotic RNA polymerase I (A12.2) and the counterpart of this from RNA polymerase II (B12.6). Homology is also seen with the eukaryotic transcription factor TFIIS, but it involves only the terminal 45 amino acids of the archaeal proteins. Evolutionary implications of these homologies are discussed.

Non-NASA Center↗

OrthoPhylo

This software builds on PHAME developed at LANL to generate phylogenetic trees of bacterial whole genome sequences. Where PHAME uses whole genome alignments to generate informative sites to base tress on, PHAME-OuS annotates bacterial genes, identifies orthologous sequences, aligns related proteins, uses those alignments to inform transcript alignments, then builds trees with several methods. The first is a conventional gene concatenation and ML tree estemation method. The second attempts to reconcile gene tree with a unified species tree using quartets (ASTRAL). Both methods allow filtering of gene lists on number of species represented, length, and gappiness in order to tune noise-to-signal for tree estimation

Middlebrook, Earl↗

Biological Information Signal Processor

Biological Information Signal Processor (BISP) is computing system analyzing data on deoxyribonucleic acid (DNA) sequences for molecular genetic analysis. Includes coprocessors, specialized microprocessors complementing present and future computers by performing rapidly most-time-consuming DNA-sequence-analyzing functions, establishing relationships (alignments) between both global sequences and defining patterns in multiple sequences. Also includes state-of-art software and data-base systems on both conventional and parallel computer systems to augment analytical abilities of developmental coprocessors.

Chow, Edward T.↗

5S ribosomal ribonucleic acid sequences in Bacteroides and Fusobacterium: evolutionary relationships within these genera and among eubacteria in general

The 5S ribosomal ribonucleic acid (rRNA) sequences were determined for Bacteroides fragilis, Bacteroides thetaiotaomicron, Bacteroides capillosus, Bacteroides veroralis, Porphyromonas gingivalis, Anaerorhabdus furcosus, Fusobacterium nucleatum, Fusobacterium mortiferum, and Fusobacterium varium. A dendrogram constructed by a clustering algorithm from these sequences, which were aligned with all other hitherto known eubacterial 5S rRNA sequences, showed differences as well as similarities with respect to results derived from 16S rRNA analyses. In the 5S rRNA dendrogram, Bacteroides clustered together with Cytophaga and Fusobacterium, as in 16S rRNA analyses. Intraphylum relationships deduced from 5S rRNAs suggested that Bacteroides is specifically related to Cytophaga rather than to Fusobacterium, as was suggested by 16S rRNA analyses. Previous taxonomic considerations concerning the genus Bacteroides, based on biochemical and physiological data, were confirmed by the 5S rRNA sequence analysis.

NASA Discipline Exobiology↗

Partial gene sequences for the A subunit of methyl-coenzyme M reductase (mcrI) as a phylogenetic tool for the family Methanosarcinaceae

Representatives of the family Methanosarcinaceae were analyzed phylogenetically by comparing partial sequences of their methyl-coenzyme M reductase (mcrI) genes. A 490-bp fragment from the A subunit of the gene was selected, amplified by the PCR, cloned, and sequenced for each of 25 strains belonging to the Methanosarcinaceae. The sequences obtained were aligned with the corresponding portions of five previously published sequences, and all of the sequences were compared to determine phylogenetic distances by Fitch distance matrix methods. We prepared analogous trees based on 16S rRNA sequences; these trees corresponded closely to the mcrI trees, although the mcrI sequences of pairs of organisms had 3.01 +/- 0.541 times more changes than the respective pairs of 16S rRNA sequences, suggesting that the mcrI fragment evolved about three times more rapidly than the 16S rRNA gene. The qualitative similarity of the mcrI and 16S rRNA trees suggests that transfer of genetic information between dissimilar organisms has not significantly affected these sequences, although we found inconsistencies between some mcrI distances that we measured and and previously published DNA reassociation data. It is unlikely that multiple mcrI isogenes were present in the organisms that we examined, because we found no major discrepancies in multiple determinations of mcrI sequences from the same organism. Our primers for the PCR also match analogous sites in the previously published mcrII sequences, but all of the sequences that we obtained from members of the Methanosarcinaceae were more closely related to mcrI sequences than to mcrII sequences, suggesting that members of the Methanosarcinaceae do not have distinct mcrII genes.

NASA Discipline Number 52-30↗

Genome-Wide Transcription Factor DNA Binding Sites and Gene Regulatory Networks in Clostridium thermocellum

Clostridium thermocellum is a thermophilic bacterium recognized for its natural ability to effectively deconstruct cellulosic biomass. While there is a large body of studies on the genetic engineering of this bacterium and its physiology to-date, there is limited knowledge in the transcriptional regulation in this organism and thermophilic bacteria in general. The study herein is the first report of a large-scale application of DNA-affinity purification sequencing (DAP-seq) to transcription factors (TFs) from a bacterium. We applied DAP-seq to > 90 TFs in C. thermocellum and detected genome-wide binding sites for 11 of them. We then compiled and aligned DNA binding sequences from these TFs to deduce the primary DNA-binding sequence motifs for each TF. These binding motifs are further validated with electrophoretic mobility shift assay (EMSA) and are used to identify individual TFs’ regulatory targets in C. thermocellum . Our results led to the discovery of novel, uncharacterized TFs as well as homologues of previously studied TFs including RexA-, LexA-, and LacI-type TFs. We then used these data to reconstruct gene regulatory networks for the 11 TFs individually, which resulted in a global network encompassing the TFs with some interconnections. As gene regulation governs and constrains how bacteria behave, our findings shed light on the roles of TFs delineated by their regulons, and potentially provides a means to enable rational, advanced genetic engineering of C. thermocellum and other organisms alike toward a desired phenotype.

59 BASIC BIOLOGICAL SCIENCES↗

Genome-wide Transcription Factor DNA Binding Sites and Gene Regulatory Networks in Clostridium thermocellum

Clostridium thermocellum is a thermophilic bacterium recognized for its natural ability to effectively deconstruct cellulosic biomass. While there is a large body of studies on the genetic engineering of this bacterium and its physiology to-date, there is limited knowledge in the transcriptional regulation in this organism and thermophilic bacteria in general. The study herein is the first report of a high-throughput application of DNA-affinity purification sequencing (DAP-seq) to transcription factors (TFs) from a thermophile. We applied DAP-seq to >90 TFs in C. thermocellum and detected genome-wide binding sites for 11 of them. We then compiled and aligned DNA binding sequences from these TFs to deduce the primary DNA-binding sequence motifs for each TF. These binding motifs are further validated with electrophoretic mobility shift assay (EMSA) and are used to identify individual TFs’ regulatory targets in C. thermocellum. Our results led to the discovery of novel, uncharacterized TFs as well as homologues of previously studied TFs including RexA-, LexA- and LacI-type TFs. We then used these data to reconstruct gene regulatory networks for the 11 TFs individually, which resulted in a global network encompassing the TFs with some interconnections. As gene regulation governs and constrains how bacteria behave, our findings shed light on the roles of TFs delineated by their regulons, and potentially provides a means to enable rational, advanced genetic engineering of C. thermocellum and other organisms alike towards a desired phenotype.

09 BIOMASS FUELS↗