Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “genome analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

The Role of Stress Proteins in Cell Stabilization: A Perspective from an Extremophile

The existence of organisms that live at near boiling temperatures is living proof that all of the complex biochemical machinery of life can be adapted to function under these harsh conditions. The purpose of our research is to elucidate the role of a group of proteins known as heat shock proteins or HSP60s in this adaptation to high temperatures. HSP60s are found in all organisms and they are among the most highly conserved proteins known. We are investigating HSP60s in an organism growing at 80 C and pH 2.0 (Sulfolobus shibatae). This organism produces three closely-related HSP60 proteins, referred to as HSP60 alpha, beta, and gamma. Our DOE-funded research during the last two years has focused on clarifying the role of FiSP60 alpha and beta. These are among the two most abundant proteins in S. shibatae grown at high temperatures and significantly increase in abundance when the cells are exposed to near-lethal temperatures. We have demonstrated that these proteins protect the cells from lethal temperatures by stabilizing their membranes. During this last year we have been studying gamma, which was discovered by genome sequence analysis but nothing was known about its function. We have determined that gamma is only expressed at low temperatures. that it interacts with alpha and beta, and that it influences their ability to form higher-order structures critical to their function. We propose that gamma modulates HSP60 function at low temperatures.

Trent, Jonathan↗

The origin and evolution of model organisms

The phylogeny and timescale of life are becoming better understood as the analysis of genomic data from model organisms continues to grow. As a result, discoveries are being made about the early history of life and the origin and development of complex multicellular life. This emerging comparative framework and the emphasis on historical patterns is helping to bridge barriers among organism-based research communities.

Review, Academic↗

Somatic Mutation Analysis in Spaceflight: NASA Twins Genome Study

The NASA Twins Genome Study investigates the effects of spaceflight on somatic mutation accumulation by comparing genome-wide sequence data from a spaceflight astronaut and his Earth-bound twin. Utilizing advanced computational software on high performance computers, this study identifies and maps somatic mutations, with implications for understanding spaceflight-associated health risks, including cancer, neurodegeneration, and cardiovascular disease. The findings aim to bridge rodent and human space research, offering insights into tissue-specific pathophysiology, risk models, and potential therapeutic interventions.

somatic mutation↗

Quantifying microbial roles in environmental iron oxidation via an integrated kinetics, `omics and metabolic modeling study (Final Report)

Iron oxyhydroxides are extremely reactive components of environmental systems, and therefore exert a strong influence on biogeochemical cycles. These oxyhydroxides strongly adsorb many biologically-relevant elements, including organic carbon and phosphate, as well as a wide range of metals including uranium and actinide species. Thus, the formation mechanism of iron oxyhydroxides are key to understanding both nutrient and contaminant cycling. Microorganisms can catalyze iron oxidation and promote the formation of Fe biominerals and thus are increasingly recognized as important players in biogeochemical cycling. However, it is completely unknown how much of environmental iron oxidation is biologically mediated versus abiotic, and various challenges in studying microbial iron oxidation have hindered accurate incorporation into hydrobiogeochemical models. The overarching goal of our work was to quantify and constrain microbial iron oxidation rates and use ‘omics to gain insight into the controls on this process, while developing tools to enable integration of biotic iron oxidation into hydrobiogeochemical models. Our work focused on the Savannah River Site (SRS) in South Carolina, where extensive microbial iron oxidation has been observed. At Tims Branch, part of the Argonne National Laboratory Wetland Hydrobiogeochemistry Science Focus Area (Argonne SFA), where groundwater discharges into a stream, iron-oxidizing microbial mats form and appear to be a major sink of uranium. In the wetlands that surround Tims Branch, there are wide swaths of iron microbial mats and flocs (mobilized mat). We measured biotic and abiotic iron oxidation rates using mats and water sampled from these sites and found that iron oxidation is primarily carried out by chemolithotrophic microorganisms. The resulting rate constants can be incorporated into models. These mats were characterized by metagenomics and metatranscriptomics, which showed that aerobic chemolithotrophs were the dominant iron-oxidizing bacteria (FeOB), and these included Gallionellaceae and Leptothrix, and possibly Rhodoferax, which is known as an Fe-reducer but may also oxidize Fe(II). This demonstrated that diverse FeOB can coexist and suggests that there are a range of niches and therefore drivers of chemolithotrophic iron oxidation. Analysis of reconstructed genomes strongly suggests that a major factor in diversity is carbon source, as genomes contained varied pathways for autotrophy and heterotrophy. We performed an in-depth analysis of Leptothrix ochracea genomes, since this sheath-former is one of the primary mat builders, yet its physiology remained unresolved. A combination of genomics, transcriptomics, and metabolic modeling suggest that L. ochracea grows mixotrophically using a combination of Fe(II) and organics for energy and both inorganic and organic carbon to create biomass. This contrasts with the largely autotrophic Gallionellaceae (Gallionella, Sideroxydans, and Ferriphaselus) also present in the mats and flocs. Remarkably, multiple FeOB, both Leptothrix and Gallionellaceae, showed activity in response to Fe(II) in live mat incubations. We tracked the gene expression of individual MAGs to Fe(II) and found that various autotrophic and heterotrophic FeOB responded to Fe(II), increasing expression of both carbon fixation and organic utilization genes. The results of the integrated field, kinetics, and omics studies give detailed insight into 1) the taxa that oxidize Fe, and 2) how they connect Fe, C, and N cycles. Towards the goal of connecting omics data to hydrobiogeochemical models, we worked with the KBase team to create a template metabolic model for chemolithotrophic iron oxidation. We initially modeled the well-characterized isolate Gallionellaceae Sideroxydans lithotrophicus, and also applied the model to the mixotroph L. ochracea. In all, we have characterized diverse FeOB in a representative wetland system and solved key problems that enable better incorporation of iron-oxidizing microbes into hydrobiogeochemical models.

54 ENVIRONMENTAL SCIENCES↗

Expanded genome and proteome reallocation in a novel, robust Bacillus coagulans strain capable of utilizing pentose and hexose sugars

Bacillus coagulans, a Gram-positive thermophilic bacterium, is recognized for its probiotic properties and recent development as a microbial cell factory. Despite its importance for biotechnological applications, the current understanding of B. coagulans’ robustness is limited, especially for undomesticated strains. To fill this knowledge gap, we characterized the metabolic capability and performed functional genomics and systems analysis of a novel, robust strain, B. coagulans B-768. Genome sequencing revealed that B-768 has the largest B. coagulans genome known to date (3.94 Mbp), about 0.63 Mbp larger than the average genome of sequenced B. coagulans strains, with expanded carbohydrate metabolism and mobilome. Functional genomics identified a well-equipped genetic portfolio for utilizing a wide range of C5 (xylose, arabinose), C6 (glucose, mannose, galactose), and C12 (cellobiose) sugars present in biomass hydrolysates, which was validated experimentally. For growth on individual xylose and glucose, the dominant sugars in biomass hydrolysates, B-768 exhibited distinct phenotypes and proteome profiles. Faster growth and glucose uptake rates resulted in lactate overflow metabolism, which makes B. coagulans a lactate overproducer; however, slower growth and xylose uptake diminished overflow metabolism due to the high energy demand for sugar assimilation. Carbohydrate Transport and Metabolism (COG-G), Translation (COG-J), and Energy Conversion and Production (COG-C) made up 60%–65% of the measured proteomes but were allocated differently when growing on xylose and glucose. The trade-off in proteome reallocation, with high investment in COG-C over COG-G, explains the xylose growth phenotype with significant upregulation of xylose metabolism, pyruvate metabolism, and tricarboxylic acid (TCA) cycle. Strain B-768 tolerates and effectively utilizes inhibitory biomass hydrolysates containing mixed sugars and exhibits hierarchical sugar utilization with glucose as the preferential substrate.

carbohydrate metabolism↗

Diversity of Sordariales Fungi: Identification of Seven New Species of Naviculisporaceae Through Morphological Analyses and Genome Sequencing

Thanks to next-generation sequencing (NGS) technologies, the diversity of fungi can now be investigated through the analysis of their genome sequences. Naviculisporaceae is a family within the Sordariales, whose diversity is not well-known, with only one genome sequence published for this family. Here, we report on the isolation and cultivation of 20 new strains of Naviculisporaceae. Their genome sequences, as well as those of the five commercially available strains, were determined, thus providing complete genome sequences for 25 new Naviculisporaceae strains. Species delimitation was conducted using a combination of (1) ITS + LSU phylogenetic analysis of the new isolates along with other known species of the family, (2) comparisons between DNA barcode sequences of the new strains with those of the known species, and (3) average genome-wide nucleotide identity calculation. We built a phylogenomic tree and studied the organization of the mating-type locus. In vitro fruiting was obtained for 16 strains, enabling the definition of seven new species, namely Pseudorhypophila gallica, Pseudorhypophila guyanensis Rhypophila alpibus, Rhypophila brasiliensis, Rhypophila camarguensis, Rhypophila reunionensis and Rhypophila thailandica, as well as two new combinations, namely Pseudorhypophila latipes and Pseudorhypophila oryzae. Eight strains for which in vitro fruiting was not obtained may belong to additional new species. These results expand the known diversity of the Naviculisporaceae and greatly enlarge the genomic data available for the family.

Naviculisporaceae↗

GenomeDepot: data management system for microbial comparative genomics

Summary GenomeDepot is an open-source web-based platform for annotation, management, and comparative analysis of microbial genomic sequences and associated data including ortholog families, protein domains, operons, regulatory interactions, strain taxonomy, and sample metadata. GenomeDepot supports rapid creation of websites for user-defined genome collections that include bioinformatic tools for interactive genome browsing, Basic Local Alignment Search Tool (BLAST) search, annotation search, comparative genomic neighborhood visualization, and sequence download. Gene function annotations are generated by a customizable annotation pipeline. The pipeline runs annotation tools in Conda environments and can be easily extended with additional user-specified tools. Availability and implementation GenomeDepot is open source and distributed under the GNU General Public License via GitHub (https://github.com/aekazakov/genome-depot). GenomeDepot is implemented in Python and was tested in Ubuntu Linux. Full installation instructions and documentation are available at https://aekazakov.github.io/genome-depot/. GenomeDepot demo server is freely accessible at https://iseq.lbl.gov/demogd/.

Kazakov, Alexey [Lawrence Berkeley National Labora↗

Modification and analysis of context-specific genome-scale metabolic models: methane-utilizing microbial chassis as a case study

ABSTRACT Context-specific genome-scale model (CS-GSM) reconstruction is becoming an efficient strategy for integrating and cross-comparing experimental multi-scale data to explore the relationship between cellular genotypes, facilitating fundamental or applied research discoveries. However, the application of CS modeling for non-conventional microbes is still challenging. Here, we present a graphical user interface that integrates COBRApy, EscherPy, and RIPTiDe, Python-based tools within the BioUML platform, and streamlines the reconstruction and interrogation of the CS genome-scale metabolic frameworks via Jupyter Notebook. The approach was tested using -omics data collected for Methylotuvimicrobium alcaliphilum 20Z R , a prominent microbial chassis for methane capturing and valorization. We optimized the previously reconstructed whole genome-scale metabolic network by adjusting the flux distribution using gene expression data. The outputs of the automatically reconstructed CS metabolic network were comparable to manually optimized i IA409 models for Ca-growth conditions. However, the CS model questions the reversibility of the phosphoketolase pathway and suggests higher flux via primary oxidation pathways. The model also highlighted unresolved carbon partitioning between assimilatory and catabolic pathways at the formaldehyde-formate node. Only a very few genes and only one enzyme with a predicted function in C1 metabolism, a homolog of the formaldehyde oxidation enzyme ( fae1-2 ), showed a significant change in expression in La-growth conditions. The CS-GSM predictions agreed with the experimental measurements under the assumption that the Fae1-2 is a part of the tetrahydrofolate-linked pathway. The cellular roles of the tungsten (W)-dependent formate dehydrogenase ( fdhAB ) and fae homologs ( fae1-2 and fae3 ) were investigated via mutagenesis. The phenotype of the f dhAB mutant followed the model prediction. Furthermore, a more significant reduction of the biomass yield was observed during growth in La-supplemented media, confirming a higher flux through formate. M. alcaliphilum 20Z R mutants lacking fae1-2 did not display any significant defects in methane or methanol-dependent growth. However, contrary to fae1, the fae1-2 homolog failed to restore the formaldehyde-activating enzyme function in complementation tests. Overall, the presented data suggest that the developed computational workflow supports the reconstruction and validation of CS-GSM networks of non-model microbes. IMPORTANCE The interrogation of various types of data is a routine strategy to explore the relationship between genotype and phenotype. An efficient approach for integrating and cross-comparing experimental multi-scale data in the context of whole-genome-based metabolic network reconstruction becomes a powerful tool that facilitates fundamental and applied research discoveries. The present study describes the reconstruction of a context-specific (CS) model for the methane-utilizing bacterium, Methylotuvimicrobium alcaliphilum 20Z R . M. alcaliphilum 20Z R is becoming an attractive microbial platform for the production of biofuels, chemicals, pharmaceuticals, and bio-sorbents for capturing atmospheric methane. We demonstrate that this pipeline can help reconstruct metabolic models that are similar to manually curated networks. Furthermore, the model is able to highlight previously overlooked pathways, thus advancing fundamental knowledge of non-model microbial systems or promoting their development toward biotechnological or environmental implementations.

Kulyashov, M. A.↗

Genomic and morphological characterization of Knufia obscura isolated from the Mars 2020 spacecraft assembly facility

Members of the family Trichomeriaceae, belonging to the Chaetothyriales order and the Ascomycota phylum, are known for their capability to inhabit hostile environments characterized by extreme temperatures, oligotrophic conditions, drought, or presence of toxic compounds. The genus Knufia encompasses many polyextremophilic species. In this report, the genomic and morphological features of the strain FJI-L2-BK-P2 presented, which was isolated from the Mars 2020 mission spacecraft assembly facility located at the Jet Propulsion Laboratory in Pasadena, California. The identification is based on sequence alignment for marker genes, multi-locus sequence analysis, and whole genome sequence phylogeny. The morphological features were studied using a diverse range of microscopic techniques (bright field, phase contrast, differential interference contrast and scanning electron microscopy). The phylogenetic marker genes of the strain FJI-L2-BK-P2 exhibited highest similarities with type strain of Knufia obscura (CBS 148926 T ) that was isolated from the gas tank of a car in Italy. To validate the species identity, whole genomes of both strains (FJI-L2-BK-P2 and CBS 148926 T ) were sequenced, annotated, and strain FJI-L2-BK-P2 was confirmed as K. obscura. The morphological analysis and description of the genomic characteristics of K. obscura FJI-L2-BK-P2 may contribute to refining the taxonomy of Knufia species. Key morphological features are reported in this K. obscura strain, resembling microsclerotia and chlamydospore-like propagules. These features known to be characteristic features in black fungi which could potentially facilitate their adaptation to harsh environments.

59 BASIC BIOLOGICAL SCIENCES↗

OrthoPhyl—streamlining large-scale, orthology-based phylogenomic studies of bacteria at broad evolutionary scales

Abstract There are a staggering number of publicly available bacterial genome sequences (at writing, 2.0 million assemblies in NCBI's GenBank alone), and the deposition rate continues to increase. This wealth of data begs for phylogenetic analyses to place these sequences within an evolutionary context. A phylogenetic placement not only aids in taxonomic classification but informs the evolution of novel phenotypes, targets of selection, and horizontal gene transfer. Building trees from multi-gene codon alignments is a laborious task that requires bioinformatic expertise, rigorous curation of orthologs, and heavy computation. Compounding the problem is the lack of tools that can streamline these processes for building trees from large-scale genomic data. Here we present OrthoPhyl, which takes bacterial genome assemblies and reconstructs trees from whole genome codon alignments. The analysis pipeline can analyze an arbitrarily large number of input genomes (>1200 tested here) by identifying a diversity-spanning subset of assemblies and using these genomes to build gene models to infer orthologs in the full dataset. To illustrate the versatility of OrthoPhyl, we show three use cases: E. coli/Shigella, Brucella/Ochrobactrum and the order Rickettsiales. We compare trees generated with OrthoPhyl to trees generated with kSNP3 and GToTree along with published trees using alternative methods. We show that OrthoPhyl trees are consistent with other methods while incorporating more data, allowing for greater numbers of input genomes, and more flexibility of analysis.

59 BASIC BIOLOGICAL SCIENCES↗

epicsuite(EAS)

Software, Documentation, Tutorials, Testing data, Example data for processing, analysis, filtering, querying, visualization and otherwise transforming genomic data for scientific analysis and discovery.

Rogers, David H.↗

AlloSHP: deconvoluting single homeologous polymorphism for phylogenetic analysis of allopolyploids

Background The genomic and evolutionary study of allopolyploid organisms involves multiple copies of homeologous chromosomes, making their assembly, annotation, and phylogenetic analysis challenging. Bioinformatics tools and protocols have been developed to study polyploid genomes, but sometimes require the assembly of their genomes, or at least the genes, limiting their use. Results We have developed AlloSHP, a command-line tool for detecting and extracting single homeologous polymorphisms (SHPs) from the subgenomes of allopolyploid species. This tool integrates three main algorithms, WGA, VCF2ALIGNMENT and VCF2SYNTENY, and allows the detection of SHPs for the study of diploid-polyploid complexes with available diploid progenitor genomes, without assembling and annotating the genomes of the allopolyploids under study. AlloSHP has been validated on three diploid-polyploid plant complexes, Brachypodium, Brassica, and Triticum-Aegilops, and a set of synthetic hybrid yeasts and their progenitors of the genus Saccharomyces. The results and congruent phylogenies obtained from the four datasets demonstrate the potential of AlloSHP for the evolutionary analysis of allopolyploids with a wide range of ploidy and genome sizes. Conclusions AlloSHP combines the strategies of simultaneous mapping against multiple reference genomes and syntenic alignment of these genomes to call SHPs, using as input data a single VCF file and the reference genomes of the known or closest extant diploid progenitor species. This novel approach provides a valuable tool for the evolutionary study of allopolyploid species, both at the interspecific and intraspecific levels, allowing the simultaneous analysis of a large number of accessions and avoiding the complex process of assembling polyploid genomes.

Allopolyploids↗

Supporting Information for manuscript: “A latitudinal gradient in S/G lignin monomer ratio driven by laccase in natural poplar variants”

Lignin composition plays a crucial role in plant structural integrity and environmental adaptation. However, the genetic and molecular mechanisms underlying natural variation in lignin composition remain poorly understood. This study investigates the syringyl-to-guaiacyl (S/G) lignin monomer ratio across a natural population of Populus trichocarpa spanning a latitudinal gradient along the Northwest coast of North America. By integrating biochemical, genomic, and geographic analysis, we identify key gene variants associated with S/G ratio differences. These datasets provide valuable insights into the evolutionary and functional genomics of lignin composition and serve as a resource for developing poplar variants optimized for forestry and bioenergy applications.

Poplar, lignin composition, laccases, latitude, ad↗

An archaeal genomic signature

Comparisons of complete genome sequences allow the most objective and comprehensive descriptions possible of a lineage's evolution. This communication uses the completed genomes from four major euryarchaeal taxa to define a genomic signature for the Euryarchaeota and, by extension, the Archaea as a whole. The signature is defined in terms of the set of protein-encoding genes found in at least two diverse members of the euryarchaeal taxa that function uniquely within the Archaea; most signature proteins have no recognizable bacterial or eukaryal homologs. By this definition, 351 clusters of signature proteins have been identified. Functions of most proteins in this signature set are currently unknown. At least 70% of the clusters that contain proteins from all the euryarchaeal genomes also have crenarchaeal homologs. This conservative set, which appears refractory to horizontal gene transfer to the Bacteria or the Eukarya, would seem to reflect the significant innovations that were unique and fundamental to the archaeal "design fabric." Genomic protein signature analysis methods may be extended to characterize the evolution of any phylogenetically defined lineage. The complete set of protein clusters for the archaeal genomic signature is presented as supplementary material (see the PNAS web site, www.pnas.org).

Non-NASA Center↗

Seed coat transcriptomic profiling of 5-593, a genotype important for genetic studies of seed coat color and patterning in common bean ( Phaseolus vulgaris L.)

Common bean (Phaseolus vulgaris L.) market classes have distinct seed coat colors, which are directly related to the diverse flavonoids found in the mature seed coat. To understand and elucidate the molecular mechanisms underlying the regulation of seed coat color, RNA-Seq data was collected from the black bean 5-593 and used for a differential gene expression and enrichment analysis from four different seed coat color development stages. 5-593 carries dominant alleles for 10 of the 11 major genes that control seed coat color and expression and has historically been used to develop introgression lines used for seed coat genetic analysis. Pairwise comparison among the four stages identified 6,294 differentially expressed genes (DEGs) varying from 508 to 5,780 DEGs depending on the compared stages. Kyoto Encyclopedia of Genes and Genomes (KEGG) enrichment analysis revealed that phenylpropanoid biosynthesis, flavonoid biosynthesis, and plant hormone signal transduction comprised the principal pathways expressed during bean seed coat pigment development. Transcriptome analysis suggested that most structural genes for flavonoid biosynthesis and some potential regulatory genes were significantly differentially expressed. Further studies detected 29 DEGs as important candidate genes governing the key enzymatic flavonoid biosynthetic pathways for common bean seed coat color development. Additionally, four gene models, Pv5-593.02G016100, 593.02G078700, Pv5-593.02G090900, and Pv5-593.06G121300, encode MYB-like transcription factor family protein were identified as strong candidate regulatory genes in anthocyanin biosynthesis which could regulate the expression levels of some important structural genes in flavonoid biosynthesis pathway. These findings provide a framework to draw new insights into the molecular networks underlying common bean seed coat pigment development.

60 APPLIED LIFE SCIENCES↗

Identification of a QTL region for tomato brown rugose fruit virus resistance in Solanum pimpinellifolium

Abstract Tomato (Solanum lycopersicumL.), one of the most widely grown vegetables in the world, has been seriously impacted in the past decade by the emerging tomato brown rugose fruit virus (ToBRFV). ToBRFV is a seed-borne tobamovirus, with ability to overcome the commonly usedTm-2 2 resistance gene in tomato. The objective of this study was to conduct quantitative trait locus (QTL) mapping and identify single-nucleotide polymorphism (SNP) markers associated with ToBRFV resistance in tomato. Two F 2 populations were used for QTL mapping: One derived from a cross betweenS. pimpinellifoliumUSVL333 (PI 390718) × USVL332 (PI 390717) and another from ‘Moneymaker’ × USVL332 (PI 390717), with population sizes of 195 and 79 plants, respectively. The resistance trait was derived from theS. pimpinellifoliumaccession USVL332 (PI 390717). A major QTL for ToBRFV resistance was identified on chromosome 11 (SL4.0ch11), with the peak located at approximately 46.84 Mbp. This QTL spans a 22-kb interval between 46,825,788 bp and 46,847,421 bp, as determined through both genome-wide association study (GWAS) and QTL linkage mapping. Three SNP markers, SL4.0ch11_46825788, SL4.0ch11_46847421, and SL4.0ch11_46850215, demonstrated the most significant association with high LOD values (LOD = 13 in the Blink model) in GWAS analysis. In this genomic region, two disease resistance gene analogs, Solyc11g062150 (TIR-NBS-LRR resistance protein, Toll-Interleukin receptor) and Solyc11g062180 (disease resistance protein, leucine-rich repeat), were identified, which may serve as candidates for ToBRFV resistance. The QTL identified in this study could be valuable for plant breeders in facilitating tomato breeding with ToBRFV resistance.

Agriculture↗

The reference genome for the northeastern Pacific bull kelp, Nereocystis luetkeana

Bull kelp, Nereocystis luetkeana, is a northeastern Pacific kelp with broad distribution from Alaska to central California. Its population declines have caused severe concerns in northern California, the Salish Sea in Washington, and recently in some populations in Oregon. Despite bull kelp's accumulated ecological and physiological studies, an assembled and annotated genomic reference was still unavailable. Here, we report the complete and annotated genome of Nereocystis luetkeana, produced by the California Conservation Genomics Project (CCGP), which aims to reveal genomic diversity patterns across California by sequencing the complete genomes of approximately 150 carefully selected species. The genome was assembled into 1562 scaffolds with 449.82 Mb, 80x of coverage and 22 952 gene models. BUSCO assembly showed a completeness score of 72% for the stramenopiles gene set. The mitochondria and chloroplast genome sequences have 37 Kb and 131 Mb, respectively. The orthology analysis between 10 Phaeophycean genomes showed 1065 expanded and 286 unique orthogroups for this species. Pairwise comparisons showed 542 orthogroups present only in N. luetkeana and M. pyrifera, another large-body kelp. The enrichment analysis of these orthogroups showed important functions related to central metabolism and signaling due to ATPases enrichment in these two species. This genome assembly will provide an essential resource for the ecology, evolution, conservation, and breeding of bull kelp.

California Conservation Genomics Project—CCGP↗