Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Genome, Bacterial”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Identification of shared viral sequences in peat moss metagenomes reveals elements of a possible Sphagnum core virome

Viruses are an understudied component of plant microbiomes. Identifying viruses that are shared between individual plants, or members of the “core virome”, could reveal stable viral populations with the potential to modulate the composition and function of the microbiome. Here, we examined the virome associated with Sphagnum mosses, a keystone species that has direct influence over the fate of peatland carbon stores. We analyzed bulk metagenomes and metatranscriptomes generated from Sphagnum field samples collected over a ten-month period to identify virus-like sequences shared among plants. Individual Sphagnum samples harbored distinct DNA and RNA viromes where only a small percentage (< 1%) of the total number of identified viral contigs were shared among all samples. Based on taxonomic classification, the shared viral contigs represent bacterial viruses, or phage (Caudoviricetes), as well as viruses of eukaryotes, namely nucleocytoplasmic large DNA viruses (Nucleocytoviricota) and RNA viruses (Riboviria). We linked the shared phage-like contigs to viral regions within sequenced genomes of bacterial taxa that are members of the Sphagnum core microbiome, suggesting that these contigs represent temperate phage or degraded prophage. The putative nucleocytoplasmic large DNA viruses and RNA viruses were phylogenetically diverse and showed sequence similarity to viruses associated with a broad range of hosts and environmental sources. The identification of shared viral contigs suggested that, despite the compositional heterogeneity between samples, Sphagnum mosses may harbor a core virome. Future work validating the presence of the core virome is warranted as it may aid in understanding how persistent viruses impact microbiome ecology and symbiont evolution within this climatically relevant keystone species.

Metagenomics↗

Comparative Genomic Analysis of Labrenzia aggregata (Alphaproteobacteria) Strains Isolated From the Mariana Trench: Insights Into the Metabolic Potentials and Biogeochemical Functions

Hadal zones are marine environments deeper than 6,000 m, most of which comprise oceanic trenches. Microbes thriving at such depth experience high hydrostatic pressure and low temperature. The genomic potentials of these microbes to such extreme environments are largely unknown. Here, we compare five complete genomes of bacterial strains belonging to Labrenzia aggregata ( Alphaproteobacteria ), including four from the Mariana Trench at depths up to 9,600 m and one reference from surface seawater of the East China Sea, to uncover the genomic potentials of this species. Genomic investigation suggests all the five strains of L. aggregata as participants in nitrogen and sulfur cycles, including denitrification, dissimilatory nitrate reduction to ammonium (DNRA), thiosulfate oxidation, and dimethylsulfoniopropionate (DMSP) biosynthesis and degradation. Further comparisons show that, among the five strains, 85% gene functions are similar with 96.7% of them encoded on the chromosomes, whereas the numbers of functional specific genes related to osmoregulation, antibiotic resistance, viral infection, and secondary metabolite biosynthesis are majorly contributed by the differential plasmids. A following analysis suggests the plasmidic gene numbers increase along with isolation depth and most plasmids are dissimilar among the five strains. These findings provide a better understanding of genomic potentials in the same species throughout a deep-sea water column and address the importance of externally originated plasmidic genes putatively shaped by deep-sea environment.

Zhong, Haohui↗

A fast comparative genome browser for diverse bacteria and archaea

Genome sequencing has revealed an incredible diversity of bacteria and archaea, but there are no fast and convenient tools for browsing across these genomes. It is cumbersome to view the prevalence of homologs for a protein of interest, or the gene neighborhoods of those homologs, across the diversity of the prokaryotes. We developed a web-based tool, fast . genomics , that uses two strategies to support fast browsing across the diversity of prokaryotes. First, the database of genomes is split up. The main database contains one representative from each of the 6,377 genera that have a high-quality genome, and additional databases for each taxonomic order contain up to 10 representatives of each species. Second, homologs of proteins of interest are identified quickly by using accelerated searches, usually in a few seconds. Once homologs are identified, fast . genomics can quickly show their prevalence across taxa, view their neighboring genes, or compare the prevalence of two different proteins. Fast . genomics is available at https://fast.genomics.lbl.gov .

59 BASIC BIOLOGICAL SCIENCES↗

Multicellular magnetotactic bacteria are genetically heterogeneous consortia with metabolically differentiated cells

Consortia of multicellular magnetotactic bacteria (MMB) are currently the only known example of bacteria without a unicellular stage in their life cycle. Because of their recalcitrance to cultivation, most previous studies of MMB have been limited to microscopic observations. To study the biology of these unique organisms in more detail, we use multiple culture-independent approaches to analyze the genomics and physiology of MMB consortia at single-cell resolution. We separately sequenced the metagenomes of 22 individual MMB consortia, representing 8 new species, and quantified the genetic diversity within each MMB consortium. This revealed that, counter to conventional views, cells within MMB consortia are not clonal. Single consortia metagenomes were then used to reconstruct the species-specific metabolic potential and infer the physiological capabilities of MMB. To validate genomic predictions, we performed stable isotope probing (SIP) experiments and interrogated MMB consortia using fluorescence in situ hybridization (FISH) combined with nanoscale secondary ion mass spectrometry (NanoSIMS). By coupling FISH with bioorthogonal noncanonical amino acid tagging (BONCAT), we explored their in situ activity as well as variation of protein synthesis within cells. We demonstrate that MMB consortia are mixotrophic sulfate reducers and that they exhibit metabolic differentiation between individual cells, suggesting that MMB consortia are more complex than previously thought. These findings expand our understanding of MMB diversity, ecology, genomics, and physiology, as well as offer insights into the mechanisms underpinning the multicellular nature of their unique lifestyle.

59 BASIC BIOLOGICAL SCIENCES↗

Genome expansion by a CRISPR trimmer-integrase

CRISPR–Cas adaptive immune systems capture DNA fragments from invading mobile genetic elements and integrate them into the host genome to provide a template for RNA-guided immunity. CRISPR systems maintain genome integrity and avoid autoimmunity by distinguishing between self and non-self, a process for which the CRISPR/Cas1–Cas2 integrase is necessary but not sufficient. In some microorganisms, the Cas4 endonuclease assists CRISPR adaptation, but many CRISPR–Cas systems lack Cas4. Here we show here that an elegant alternative pathway in a type I-E system uses an internal DnaQ-like exonuclease (DEDDh) to select and process DNA for integration using the protospacer adjacent motif (PAM). The natural Cas1–Cas2/exonuclease fusion (trimmer-integrase) catalyses coordinated DNA capture, trimming and integration. Five cryo-electron microscopy structures of the CRISPR trimmer-integrase, visualized both before and during DNA integration, show how asymmetric processing generates size-defined, PAM-containing substrates. Before genome integration, the PAM sequence is released by Cas1 and cleaved by the exonuclease, marking inserted DNA as self and preventing aberrant CRISPR targeting of the host. Together, these data support a model in which CRISPR systems lacking Cas4 use fused or recruited exonucleases for faithful acquisition of new CRISPR immune sequences.

59 BASIC BIOLOGICAL SCIENCES↗

Comparative and pangenomic analysis of the genus Streptomyces

Abstract Streptomycetes are highly metabolically gifted bacteria with the abilities to produce bioproducts that have profound economic and societal importance. These bioproducts are produced by metabolic pathways including those for the biosynthesis of secondary metabolites and catabolism of plant biomass constituents. Advancements in genome sequencing technologies have revealed a wealth of untapped metabolic potential from Streptomyces genomes. Here, we report the largest Streptomyces pangenome generated by using 205 complete genomes. Metabolic potentials of the pangenome and individual genomes were analyzed, revealing degrees of conservation of individual metabolic pathways and strains potentially suitable for metabolic engineering. Of them, Streptomyces bingchenggensis was identified as a potent degrader of plant biomass. Polyketide, non-ribosomal peptide, and gamma-butyrolactone biosynthetic enzymes are primarily strain specific while ectoine and some terpene biosynthetic pathways are highly conserved. A large number of transcription factors associated with secondary metabolism are strain-specific while those controlling basic biological processes are highly conserved. Although the majority of genes involved in morphological development are highly conserved, there are strain-specific varieties which may contribute to fine tuning the timing of cellular differentiation. Overall, these results provide insights into the metabolic potential, regulation and physiology of streptomycetes, which will facilitate further exploitation of these important bacteria.

59 BASIC BIOLOGICAL SCIENCES↗

Aeromonas in South Asia: genomic insights into an environmental pathogen and reservoir of antimicrobial resistance

Aeromonads are an ecologically versatile group of bacteria that cause infections in aquatic animals and are recognised as emerging human pathogens. Despite this, our understanding of Aeromonas diversity, especially the relationship between clinical and environmental strains, remains limited. Here, we present a genomic analysis of the Aeromonas genus, comprising 1853 genomes, and a detailed comparison of clinical and environmental strains from South Asia, including 996 newly sequenced genomes from Bangladesh and India. Phylogenetic analyses revealed that Aeromonas is a highly diverse genus, with no distinct clade separating clinical and environmental isolates. We identified 28 Aeromonas species and 905 novel sequence types, comprising 72.5% of the genomes. Notably, we show a high incidence of antimicrobial resistance (AMR) genes across all isolates, including against front and last-line antibiotics. Finally, we highlight frequent misidentification of Aeromonas as Vibrio cholerae, which is relevant to cholera-endemic regions where both genera co-exist and are associated with diarrhoeal disease. Our study underscores Aeromonas as an important environmental AMR reservoir and emerging multi-species pathogen capable of spilling over into human populations.

59 BASIC BIOLOGICAL SCIENCES↗

Mixed heavy metal stress induces global iron starvation response

Abstract Multiple heavy metal contamination is an increasingly common global problem. Heavy metals have the potential to disrupt microbially mediated biogeochemical cycling. However, systems-level studies on the effects of combinations of heavy metals on bacteria are lacking. For this study, we focused on the Oak Ridge Reservation (ORR; Oak Ridge, TN, USA) subsurface which is contaminated with several heavy metals and high concentrations of nitrate. Using a native Bacillus cereus isolate that represents a dominant species at this site, we assessed the combined impact of eight metal contaminants, all at site-relevant concentrations, on cell processes through an integrated multi-omics approach that included discovery proteomics, targeted metabolomics, and targeted gene-expression profiling. The combination of eight metals impacted cell physiology in a manner that could not have been predicted from summing phenotypic responses to the individual metals. Exposure to the metal mixture elicited a global iron starvation response not observed during individual metal exposures. This disruption of iron homeostasis resulted in decreased activity of the iron-cofactor-containing nitrate and nitrite reductases, both of which are important in biological nitrate removal at the site. We propose that the combinatorial effects of simultaneous exposure to multiple heavy metals is an underappreciated yet significant form of cell stress in the environment with the potential to disrupt global nutrient cycles and to impede bioremediation efforts at mixed waste sites. Our work underscores the need to shift from single- to multi-metal studies for assessing and predicting the impacts of complex contaminants on microbial systems.

59 BASIC BIOLOGICAL SCIENCES↗

CRISPR-COPIES: an in silico platform for discovery of neutral integration sites for CRISPR/Cas-facilitated gene integration

Abstract The CRISPR/Cas system has emerged as a powerful tool for genome editing in metabolic engineering and human gene therapy. However, locating the optimal site on the chromosome to integrate heterologous genes using the CRISPR/Cas system remains an open question. Selecting a suitable site for gene integration involves considering multiple complex criteria, including factors related to CRISPR/Cas-mediated integration, genetic stability, and gene expression. Consequently, identifying such sites on specific or different chromosomal locations typically requires extensive characterization efforts. To address these challenges, we have developed CRISPR-COPIES, a COmputational Pipeline for the Identification of CRISPR/Cas-facilitated intEgration Sites. This tool leverages ScaNN, a state-of-the-art model on the embedding-based nearest neighbor search for fast and accurate off-target search, and can identify genome-wide intergenic sites for most bacterial and fungal genomes within minutes. As a proof of concept, we utilized CRISPR-COPIES to characterize neutral integration sites in three diverse species: Saccharomyces cerevisiae, Cupriavidus necator, and HEK293T cells. In addition, we developed a user-friendly web interface for CRISPR-COPIES (https://biofoundry.web.illinois.edu/copies/). We anticipate that CRISPR-COPIES will serve as a valuable tool for targeted DNA integration and aid in the characterization of synthetic biology toolkits, enable rapid strain construction to produce valuable biochemicals, and support human gene and cell therapy applications.

59 BASIC BIOLOGICAL SCIENCES↗

A stable 15-member bacterial SynCom promotes Brachypodium growth under drought stress

Introduction: Rhizosphere microbiomes are known to drive soil nutrient cycling and influence plant fitness during adverse environmental conditions. Field-derived robust Synthetic Communities (SynComs) of microbes mimicking the diversity of rhizosphere microbiomes can greatly advance a deeper understanding of such processes. However, assembling stable, genetically tractable, reproducible, and scalable SynComs remains challenging. Methods: Here, we present a systematic approach using a combination of network analysis and cultivation-guided methods to construct a 15-member SynCom from the rhizobiome of Brachypodium distachyon. This SynCom incorporates diverse strains from five bacterial phyla. Genomic analysis of the individual strains was performed to reveal encoded plant growth-promoting traits, including genes for the synthesis of osmoprotectants (trehalose and betaine) and Na+/K+ transporters, and some predicted traits were validated by laboratory phenotypic assays. Results: The SynCom demonstrates strong stability both in vitro and in planta. Most strains encoded multiple plant growth-promoting functions, and several of these were confirmed experimentally. The presence of osmoprotectant and ion transporter genes likely contributed to the observed resilience of Brachypodium to drought stress, where plants amended with the SynCom recovered better than those without. We further observed preferential colonization of SynCom strains around root tips under stress, likely due to active interactions between plant root metabolites and bacteria. Discussion: Our results demonstrate that trait-informed construction of synthetic communities can yield stable, functionally diverse consortia that enhance plant resilience under drought. Preferential colonization near root tips points to active, localized plant-microbe signaling as a component of stress-responsive recruitment. This stable SynCom provides a scalable platform for probing mechanisms of plant-microbe interaction and for developing microbiome-based strategies to improve soil and crop performance in variable environments.

Yadav, Archana↗

CRISPR-COPIES: Web Tool

CRISPR/Cas system has emerged as a powerful genome-editing tool for metabolic engineering and human gene therapy. However, the conundrum of where to integrate heterologous genes on the chromosome using the CRISPR/Cas system remains an open question. Selecting a site for gene integration requires incorporation of complex criteria such as factors involved in CRISPR/Cas-mediated integration, genetic stability, and gene expression and therefore, usually requires strenuous characterization of sites on particular or different chromosomal locations. To address these issues, we developed CRISPR-COPIES, a COmputational Pipeline for the Identification of CRISPR/Cas-facilitated intEgration Sites. The tool applies ScaNN, a state-of-the-art model on the embedding-based nearest neighbor search for fast and accurate off-target search and can identify genome-wide intergenic sites for most bacterial and fungal genomes within minutes. This submission contains the code we developed to create a user-friendly web interface for CRISPR-COPIES (https://biofoundry.web.illinois.edu/copies/). We anticipate CRISPR-COPIES will serve as a useful tool for targeted DNA integration and aid in the characterization of synthetic biology toolkits, rapid strain construction to produce valuable biochemicals, and human gene and cell therapy.

Bioinformatics↗

Structural models of the MscL gating mechanism

Three-dimensional structural models of the mechanosensitive channel of large conductance, MscL, from the bacteria Mycobacterium tuberculosis and Escherichia coli were developed for closed, intermediate, and open conformations. The modeling began with the crystal structure of M. tuberculosis MscL, a homopentamer with two transmembrane alpha-helices, M1 and M2, per subunit. The first 12 N-terminal residues, not resolved in the crystal structure, were modeled as an amphipathic alpha-helix, called S1. A bundle of five parallel S1 helices are postulated to form a cytoplasmic gate. As membrane tension induces expansion, the tilts of M1 and M2 are postulated to increase as they move away from the axis of the pore. Substantial expansion is postulated to occur before the increased stress in the S1 to M1 linkers pulls the S1 bundle apart. During the opening transition, the S1 helices and C-terminus amphipathic alpha-helices, S3, are postulated to dock parallel to the membrane surface on the perimeter of the complex. The proposed gating mechanism reveals critical spatial relationships between the expandable transmembrane barrel formed by M1 and M2, the gate formed by S1 helices, and "strings" that link S1s to M1s. These models are consistent with numerous experimental results and modeling criteria.

NASA Discipline Cell Biology↗

Transcription factor IID in the Archaea: sequences in the Thermococcus celer genome would encode a product closely related to the TATA-binding protein of eukaryotes

The first step in transcription initiation in eukaryotes is mediated by the TATA-binding protein, a subunit of the transcription factor IID complex. We have cloned and sequenced the gene for a presumptive homolog of this eukaryotic protein from Thermococcus celer, a member of the Archaea (formerly archaebacteria). The protein encoded by the archaeal gene is a tandem repeat of a conserved domain, corresponding to the repeated domain in its eukaryotic counterparts. Molecular phylogenetic analyses of the two halves of the repeat are consistent with the duplication occurring before the divergence of the archael and eukaryotic domains. In conjunction with previous observations of similarity in RNA polymerase subunit composition and sequences and the finding of a transcription factor IIB-like sequence in Pyrococcus woesei (a relative of T. celer) it appears that major features of the eukaryotic transcription apparatus were well-established before the origin of eukaryotic cellular organization. The divergence between the two halves of the archael protein is less than that between the halves of the individual eukaryotic sequences, indicating that the average rate of sequence change in the archael protein has been less than in its eukaryotic counterparts. To the extent that this lower rate applies to the genome as a whole, a clearer picture of the early genes (and gene families) that gave rise to present-day genomes is more apt to emerge from the study of sequences from the Archaea than from the corresponding sequences from eukaryotes.

NASA Discipline Exobiology↗

Genome-wide transcriptional analysis of flagellar regeneration in Chlamydomonas reinhardtii identifies orthologs of ciliary disease genes

The important role that cilia and flagella play in human disease creates an urgent need to identify genes involved in ciliary assembly and function. The strong and specific induction of flagellar-coding genes during flagellar regeneration in Chlamydomonas reinhardtii suggests that transcriptional profiling of such cells would reveal new flagella-related genes. We have conducted a genome-wide analysis of RNA transcript levels during flagellar regeneration in Chlamydomonas by using maskless photolithography method-produced DNA oligonucleotide microarrays with unique probe sequences for all exons of the 19,803 predicted genes. This analysis represents previously uncharacterized whole-genome transcriptional activity profiling study in this important model organism. Analysis of strongly induced genes reveals a large set of known flagellar components and also identifies a number of important disease-related proteins as being involved with cilia and flagella, including the zebrafish polycystic kidney genes Qilin, Reptin, and Pontin, as well as the testis-expressed tubby-like protein TULP2.

Polycystic Kidney Diseases/genetics↗

Filling gaps in bacterial catabolic pathways with computation and high-throughput genetics

To discover novel catabolic enzymes and transporters, we combined high-throughput genetic data from 29 bacteria with an automated tool to find gaps in their catabolic pathways. GapMind for carbon sources automatically annotates the uptake and catabolism of 62 compounds in bacterial and archaeal genomes. For the compounds that are utilized by the 29 bacteria, we systematically examined the gaps in GapMind’s predicted pathways, and we used the mutant fitness data to find additional genes that were involved in their utilization. We identified novel pathways or enzymes for the utilization of glucosamine, citrulline, myo-inositol, lactose, and phenylacetate, and we annotated 299 diverged enzymes and transporters. We also curated 125 proteins from published reports. For the 29 bacteria with genetic data, GapMind finds high-confidence paths for 85% of utilized carbon sources. In diverse bacteria and archaea, 38% of utilized carbon sources have high-confidence paths, which was improved from 27% by incorporating the fitness-based annotations and our curation. GapMind for carbon sources is available as a web server ( http://papers.genomics.lbl.gov/carbon ) and takes just 30 seconds for the typical genome.

59 BASIC BIOLOGICAL SCIENCES↗