Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “sequence alignment”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Analyses of transcriptomes and the first complete genome of Leucocalocybe mongolica provide new insights into phylogenetic relationships and conservation

In this study, we report a de novo assembly of the first high-quality genome for a wild mushroom species Leucocalocybe mongolica (LM). We performed high-throughput transcriptome sequencing to analyze the genetic basis for the life history of LM. Our results show that the genome size of LM is 46.0 Mb, including 26 contigs with a contig N50 size of 3.6 Mb. In total, we predicted 11,599 protein-coding genes, of which 65.7% (7630) could be aligned with high confidence to annotated homologous genes in other species. We performed phylogenetic analyses using genes form 3269 single-copy gene families and showed support for distinguishing LM from the genus Tricholoma (L.) P.Kumm., in which it is sometimes circumscribed. We believe that one reason for limited wild occurrences of LM may be the loss of key metabolic genes, especially carbohydrate-active enzymes (CAZymes), based on comparisons with other closely related species. The results of our transcriptome analyses between vegetative (mycelia) and reproductive (fruiting bodies) organs indicated that changes in gene expression among some key CAZyme genes may help to determine the switch from asexual to sexual reproduction. Taken together, our genomic and transcriptome data for LM comprise a valuable resource for both understanding the evolutionary and life history of this species.

54 ENVIRONMENTAL SCIENCES↗

High-throughput Single-Cell Proteomics and Transcriptomics from the Same Cells with a Nanoliter-Scale Spin-Transfer Approach

Single-cell multiomic platforms provide a comprehensive snapshot of cellular states and cell types by offering critical insights into the spatiotemporal regulation of biomolecular networks at a systems level, thereby defining the basis of multicellularity. Here, we introduce nanoSPINS, an advanced platform that enables high-throughput profiling and integrative analysis of the transcriptome and proteome from the same single cells using RNA sequencing and isobaric labeling LC-MS-based proteomics, respectively. NanoSPINS can efficiently transfer mRNA-containing droplets across two microarrays via a centrifugation-based approach, while proteins are retained on the initial platform. Benchmarking of nanoSPINS on two cell lines demonstrates its ability to generate global proteomic and transcriptomic profiles that align well with previously established methodologies/platforms. The incorporation of isobaric TMTpro labeling into this single-cell multiomics platform significantly enhances the throughput of single-cell proteomic analyses. Through the high-throughput quantification of the proteome and transcriptome, nanoSPINS not only facilitates the identification of molecular features at both mRNA and protein level but also provides larger sample sizes for improved statistical power in clustering and differential abundance. Given the broad applicability of single-cell multiomics in biological research and clinical settings, we believe nanoSPINS represents a powerful platform for the characterization of heterogeneous cell populations.

multi 'omics↗

Building a framework to genetically characterize “feather spots” and understand demographic impacts of solar energy sites on migratory bird populations

The lack of data on the impact of utility-scale solar facilities on avian species and populations adds to the cost of siting and operation. As much as 32 percent of the avian biological material (feathers and carcasses) recovered from solar facilities remain unidentified, because they often take the form of “feather spots”. Feather spots are remains of impacted animals that can be separated into two broad categories: 1) those remains that may be visually identified to a species, or 2) those that cannot be visually identified to a species due to degradation from the environment and/or scavenger activity (listed as “unknown”). Even when feather spots can be identified to species, they cannot be visually assigned to particular breeding populations. In some cases, it is unknown whether multiple feather spots represent single or multiple individuals. This project’s objectives were to: 1. Use a developed, genetic-based technique to identify and determine the species, population of origin, and number of individuals found in feather spots recovered from solar facilities. 2. Implement collected data and resulting analyses to develop a publicly accessible web-based decision-making tool that can be used by the solar industry, regulators and other stakeholders to inform siting, mitigation, and conservation management efforts. 3. Establish a not-for-profit fee-for-service center at UCLA to ensure collection and identification of feather spots continue after the project period of performance. During the Project Period, we proposed to establish a pipeline for collecting, transporting, and storing of avian biological material collected at solar facilities and the collection and identification of feather spots to species and individual. We proposed the development of a genetic-based framework that would recover viable DNA from feather spots, amplify this DNA (i.e., make millions of copies of the original DNA), and use it to match the resulting sequences to a national database of known species of birds. The result would be the identification of feathers spots that were previously unidentified, and the incorporation of these samples into a larger database that included all samples recovered from solar facilities. The resulting report (below) details the result of this work and its alignment with proposed activities. We proposed the use of the data collected to assess the comparative risk to specific species or populations of species from solar facilities. For some species, we have already identified genomic markers of specific breeding populations and developed “genoscapes,” maps of unique genetic variation across the full breeding range of a species. We used these (previously and newly developed) genoscapes to probabilistically link a feather spot to the specific breeding populations from which it originated (assignment probabilities range from 75%-100% depending on species and population groups). For those species without genoscapes, we developed a vulnerability and susceptibility estimate that determines the relative local and regional risk to populations that are in geographic proximity to solar facilities, using citizen science data (Breeding Bird Survey (BBS) and eBird). These two feather spot processing pipelines (see Figure 1 below) provide quantitative estimates as to the numbers of individuals from a given population of origin that are affected by solar facilities, and ultimately can reduce costs to the consumer by reducing the industry costs associated with mitigation and siting strategies for future solar energy development.

14 SOLAR ENERGY↗

Two major chromosome evolution events with unrivaled conserved gene content in pomegranate

Pomegranate has a unique evolutionary history given that different cultivars have eight or nine bivalent chromosomes with possible crossability between the two classes. Therefore, it is important to study chromosome evolution in pomegranate to understand the dynamics of its population. Here, we de novo assembled the Azerbaijani cultivar “Azerbaijan guloyshasi” (AG2017; 2n = 16) and re-sequenced six cultivars to track the evolution of pomegranate and to compare it with previously published de novo assembled and re-sequenced cultivars. High synteny was observed between AG2017, Bhagawa (2n = 16), Tunisia (2n = 16), and Dabenzi (2n = 18), but these four cultivars diverged from the cultivar Taishanhong (2n = 18) with several rearrangements indicating the presence of two major chromosome evolution events. Major presence/absence variations were not observed as >99% of the five genomes aligned across the cultivars, while >99% of the pan-genic content was represented by Tunisia and Taishanhong only. We also revisited the divergence between soft- and hard-seeded cultivars with less structured population genomic data, compared to previous studies, to refine the selected genomic regions and detect global migration routes for pomegranate. We reported a unique admixture between soft- and hard-seeded cultivars that can be exploited to improve the diversity, quality, and adaptability of local pomegranate varieties around the world. Our study adds body knowledge to understanding the evolution of the pomegranate genome and its implications for the population structure of global pomegranate diversity, as well as planning breeding programs aiming to develop improved cultivars.

59 BASIC BIOLOGICAL SCIENCES↗

Spontaneous Valley Polarization of Interacting Carriers in a Monolayer Semiconductor

Here, we report magnetoabsorption spectroscopy of gated WSe 2 monolayers in high magnetic fields up to 60 T. When doped with a 2D Fermi sea of mobile holes, well-resolved sequences of optical transitions are observed in both σ ± circular polarizations, which unambiguously and separately indicate the number of filled Landau levels (LLs) in both K and K ' valleys. This reveals the interaction-enhanced valley Zeeman energy, which is found to be highly tunable with hole density p . We exploit this tunability to align the LLs in K and K ' , and find that the 2D hole gas becomes unstable against small changes in LL filling and can spontaneously valley polarize. These results cannot be understood within a single-particle picture, highlighting the importance of exchange interactions in determining the ground state of 2D carriers in monolayer semiconductors.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Decay of the proton-unbound superradiant state in 13 N

Here, the 12 C ( 3 He, 𝑑)⁢ 13 N* reaction is studied in an experiment with a high-resolution magnetic spectrograph, in coincidence with protons detected in silicon detectors near the target. This allows for the observation of angular correlation patterns between the proton transfer and proton decays from populated unbound resonances. A formalism describing the spin polarization of direct reactions is developed to analyze these correlations, and is verified on the known directionally asymmetric decay distributions arising from parity mixing in the 13 N ⁡($\frac{3}{2}$ − , $\frac{5}{2}$ + ) doublet. The same formalism is used to study the decays from the continuum-aligned, broad 3/2 + resonance at 7.9 MeV excitation energy, which arises from superradiant coupling. The observed asymmetric angular correlation patterns are approximately reproduced by adding an “artificial” 3/2 − resonance with strength equal to that of the reaction formalism. This parity-mixing approach serves as a first approximation to a more advanced reaction model of rapid reaction and decay sequences.

13N↗

Craspase is a CRISPR RNA-guided, RNA-activated protease

The CRISPR-Cas type III-E RNA-targeting effector complex gRAMP/Cas7-11 is associated with a caspase-like protein (TPR-CHAT/Csx29) to form Craspase (CRISPR-guided caspase). Here, we use cryo–electron microscopy snapshots of Craspase to explain its target RNA cleavage and protease activation mechanisms. Target-guide pairing extending into the 5' region of the guide RNA displaces a gating loop in gRAMP, which triggers an extensive conformational relay that allosterically aligns the protease catalytic dyad and opens an amino acid side-chain–binding pocket. We further define Csx30 as the endogenous protein substrate that is site-specifically proteolyzed by RNA-activated Craspase. This protease activity is switched off by target RNA cleavage by gRAMP and is not activated by RNA targets containing a matching protospacer flanking sequence. We thus conclude that Craspase is a target RNA–activated protease with self-regulatory capacity.

Science & Technology - Other Topics↗

Understanding the Reactions Between Fe and Se Binary Diffusion Couples

Spurred by recent discoveries of high-temperature superconductivity in Fe-Se based materials, the magnetic, electronic, and catalytic properties of iron-chalcogenides have drawn significant attention. Furthermore, much remains to be understood about the sequence of phase formation in these systems. In this work, we shed light on this issue by preparing a series of binary Fe-Se ultrathin diffusion couples via designed thin film precursors and investigating their structural evolution as a function of composition and annealing temperature. Two previously unreported Fe-Se phases crystallized during the deposition process on a nominally room-temperature Si substrate in the 27-33% and 37-47% Fe (atomic percent) composition regimes. Both phases completely decompose after annealing to 200°C in a nitrogen glovebox. At higher temperatures, the sequence of phase formation is governed by Se loss in the annealing process, consistent with what would be expected from the phase diagram. Films rich in Fe (53-59% Fe) crystalized during deposition as β-FeSe (P4/nmm) with preferred c-axis orientation to the amorphous SiO 2 substrate surface, providing a means to non-epitaxial self-assembly of crystallographically aligned, iron-rich β-FeSe for future research. Our findings suggest the crystallization of binary Fe-Se compounds at room temperature via near diffusionless transformations should be a significant consideration in future attempts to prepare metastable ternary and higher order compounds containing Fe and Se.

36 MATERIALS SCIENCE↗

CAPG: comprehensive allopolyploid genotyper

Genotyping by sequencing is a powerful tool for investigating genetic variation in plants, but many economically important plants are allopolyploids, where homoeologous similarity obscures the subgenomic origin of reads and confounds allelic and homoeologous SNPs. Recent polyploid genotyping methods use allelic frequencies, rate of heterozygosity, parental cross or other information to resolve read assignment, but good subgenomic references offer the most direct information. The typical strategy aligns reads to the joint reference, performs diploid genotyping within each subgenome, and filters the results, but persistent read misassignment results in an excess of false heterozygous calls. We introduce the Comprehensive Allopolyploid Genotyper (CAPG), which formulates an explicit likelihood to weight read alignments against both subgenomic references and genotype individual allopolyploids from whole-genome resequencing data. We demonstrate CAPG in allotetraploids, where it performs better than Genome Analysis Toolkit’s HaplotypeCaller applied to reads aligned to the combined subgenomic references.

59 BASIC BIOLOGICAL SCIENCES↗

Layer-engineered interlayer charge transfer in WSe 2 /WS 2 heterostructures

The layer thickness determines the electronic structure of two-dimensional (2D) materials, leading to different band alignments, which are crucial for the transition metal dichalcogenides heterostructures. Here, we investigated the heterostructure of WSe 2 /WS 2 with different layer thicknesses by steady-state and transient absorption spectroscopy. We observed different ultrafast charge transfer behaviors in 1L-WSe 2 /2L-WS 2 and 2L-WSe 2 /2L-WS 2 few-layer heterostructures. We demonstrate that the layer thickness determines the sequence of intralayer exciton relaxation and interlayer charge transfer. The valley transfer of the band edge induced by the layer thickness can effectively mediate the hot carrier transfer time and interlayer exciton lifetime. Furthermore, these provide us a deeper understanding of carrier dynamics in 2D indirect bandgap semiconductor heterostructures.

77 NANOSCIENCE AND NANOTECHNOLOGY↗

An investigation of the failure modes in U-10Mo monolithic fuel irradiated to high burnup

Here, this study investigated possible failure mode sequences in U-10Mo monolithic fuel irradiated to very-high-burnup. The U-Mo fuel plate used in this study was characterized using scanning electron microscopy (SEM), wavelength dispersive spectroscopy (WDS), and image analysis techniques to investigate how microstructural features may initiate crack formation and propagation through the microstructure, potentially leading to blistering. Distinctly large fission gas pores (FGPs) were observed to preferentially align 5-15µm away from the U-10Mo/Zr interface, parallel to the Zr diffusion barrier. Preferential growth, alignment, and interconnection of FGP could initiate blistering in the fuel plate by creating a fission gas channel, parallel to the U-Mo/Zr interface. FGP alignment in the fuel phase near the U-Mo/Zr interface is believed to be one of the precursors to crack and Type 2 blister formation in monolithic U-Mo fuels. The chemical maps revealed differences in fission product behavior near the U-Mo/Zr interface such that Nd remains in the fuel matrix at the FGP sites. On the other hand, other fission products like Xe and Cs can get trapped in FGPs by Nd in the fuel phase and/or diffuse toward the Zr diffusion barrier.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Maast: genotyping thousands of microbial strains efficiently

Existing single nucleotide polymorphism (SNP) genotyping algorithms do not scale for species with thousands of sequenced strains, nor do they account for conspecific redundancy. Here we present a bioinformatics tool, Maast, which empowers population genetic meta-analysis of microbes at an unrivaled scale. Maast implements a novel algorithm to heuristically identify a minimal set of diverse conspecific genomes, then constructs a reliable SNP panel for each species, and enables rapid and accurate genotyping using a hybrid of whole-genome alignment and k-mer exact matching. We demonstrate Maast’s utility by genotyping thousands of Helicobacter pylori strains and tracking SARS-CoV-2 diversification.

59 BASIC BIOLOGICAL SCIENCES↗

Single nucleotide variants drive evolutionary phage-host arms race in anaerobic carbon dioxide-converting microbiome

Microbial bioconversions are shaped by environmental perturbations and the adaptation of resident microbiomes. Prokaryotes coexist with bacteriophages, yet their coevolutionary trajectories remain underexplored. Here, we investigate the effects of a cultivation vessel leak on an anaerobic consortium performing carbon dioxide reduction. Using time-series shotgun metagenomic sequencing, we reconstruct microbial and viral genomes to track community shifts. We further apply single-nucleotide variant profiling and CRISPR array analysis to monitor viral microdiversity and host defense mechanisms. After bioaugmentation restores bioconversion efficiency, the consortium undergoes pronounced restructuring, with new dominant taxa emerging from the rare biosphere. We identify patterns consistent with phage predation selectively removing certain species, while others exhibit resilience to infection. This shift aligns with a widespread viral outbreak and a transient increased frequency of single nucleotide variants in bacterial CRISPR–Cas defense genes. Expansion of CRISPR spacers further supports that CRISPR-mediated processes influence microbial resilience. Concurrently, phages infecting resilient hosts exhibited adaptive evolution, marked by high genetic heterogeneity. Selective pressure varies across their genomes, targeting infectivity genes and protospacer-adjacent motifs. These findings highlight a dynamic evolutionary arms race driven by the selection of beneficial genetic variants, providing a mechanistic framework for multi-omics investigations, and informing biotechnological applications, including phage-based microbiome manipulation.

Ghiotto, G↗

Observational constraints of an anisotropic boost due to the projection effects using redMaPPer clusters

Optical clusters identified from red-sequence galaxies suffer from projection effects, where interloper galaxies along the line of sight to a cluster are mistaken as genuine members of the cluster. In the previous study, we found that the projection effects cause the boost on the amplitudes of clustering and lensing on large scale compared to the expected amplitudes in the absence of any projection effects. These boosts are caused by preferential selections of filamentary structure aligned to the line of sight due to distance uncertainties in photometric surveys. We model the projection effects with two simple assumptions and develop a novel method to quantify the size of the boost using cluster-galaxy cross-correlation functions. We validate our method using mock cluster catalogues built from cosmological N-body simulations and find that we can obtain unbiased constraints on the boost parameter with our model. We then apply our analysis on the SDSS redMaPPer clusters and find that the size of the boost is roughly 20 per cent for all the richness bins except the cluster sample with the richness bin λ ∈ [30, 40]. This is the first study to constrain the boost parameter independent from cluster cosmology studies and provides a self-consistency test for the projection effects.

79 ASTRONOMY AND ASTROPHYSICS↗

Galba: genome annotation with miniprot and AUGUSTUS

The Earth Biogenome Project has rapidly increased the number of available eukaryotic genomes, but most released genomes continue to lack annotation of protein-coding genes. In addition, no transcriptome data is available for some genomes. Various gene annotation tools have been developed but each has its limitations. Here, we introduce GALBA, a fully automated pipeline that utilizes miniprot, a rapid protein-to-genome aligner, in combination with AUGUSTUS to predict genes with high accuracy. Accuracy results indicate that GALBA is particularly strong in the annotation of large vertebrate genomes. We also present use cases in insects, vertebrates, and a land plant. GALBA is fully open source and available as a docker image for easy execution with Singularity in high-performance computing environments. Our pipeline addresses the critical need for accurate gene annotation in newly sequenced genomes, and we believe that GALBA will greatly facilitate genome annotation for diverse organisms.

59 BASIC BIOLOGICAL SCIENCES↗

Simulations of n -dodecane/oxygen/nitrogen cellular detonations

In this work, two-dimensional n-dodecane/air/nitrogen cellular detonations are simulated with various equivalence ratios (ERs). A skeletal mechanism consisting of 54 species and 269 reactions is used. The lower and upper equivalence ratio boundaries for self-sustained detonation are 0.3 and 2.2, respectively. Detonation with different regimes characterized by the detonation cell patterns is observed, which aligns well with the category based on the stability parameter, i.e., weakly and highly unstable detonations, and extinction. Further, in terms of the frontal structure, non-negligible effect of diffusion on the cellular detonation is revealed, especially in the vicinity of the leading shock front. In highly unstable and quenched detonations, the alternation in reaction pathway within the induction zone accounts for the changes of detonation dynamics, such as the absence or extended sequence of important radical formation, e.g., OH. In addition, the composition of the unburned pockets depends on both pocket location from the leading shock front and the ER in the fresh mixture, because the former determines the residence time, whilst the latter affects the pocket reaction rate.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Aligning Standards Communities for Omics Biodiversity Data: Sustainable Darwin Core-MIxS Interoperability

The standardization of data, encompassing both primary and contextual information (metadata), plays a pivotal role in facilitating data (re-)use, integration, and knowledge generation. However, the biodiversity and omics communities, converging on omics biodiversity data, have historically developed and adopted their own distinct standards, hindering effective (meta)data integration and collaboration. In response to this challenge, the Task Group (TG) for Sustainable DwC-MIxS Interoperability was established. Convening experts from the Biodiversity Information Standards (TDWG) and the Genomic Standards Consortium (GSC) alongside external stakeholders, the TG aimed to promote sustainable interoperability between the Minimum Information about any (x) Sequence (MIxS) and Darwin Core (DwC) specifications. To achieve this goal, the TG utilized the Simple Standard for Sharing Ontology Mappings (SSSOM) to create a comprehensive mapping of DwC keys to MIxS keys. This mapping, combined with the development of the MIxS-DwC extension, enables the incorporation of MIxS core terms into DwC-compliant metadata records, facilitating seamless data exchange between MIxS and DwC user communities. Through the implementation of this translation layer, data produced in either MIxS- or DwC-compliant formats can now be efficiently brokered, breaking down silos and fostering closer collaboration between the biodiversity and omics communities. To ensure its sustainability and lasting impact, TDWG and GSC have both signed a Memorandum of Understanding (MoU) on creating a continuous model to synchronize their standards. These achievements mark a significant step forward in enhancing data sharing and utilization across domains, thereby unlocking new opportunities for scientific discovery and advancement.

59 BASIC BIOLOGICAL SCIENCES↗

Exploring the Evolution of Stellar Rotation Using Galactic Kinematics

The rotational evolution of cool dwarfs is poorly constrained after ∼1–2 Gyr due to a lack of precise ages and rotation periods for old main-sequence stars. In this work, we use velocity dispersion as an age proxy to reveal the temperature-dependent rotational evolution of low-mass Kepler dwarfs and demonstrate that kinematic ages could be a useful tool for calibrating gyrochronology in the future. We find that a linear gyrochronology model, calibrated to fit the period–T{sub eff} relationship of the Praesepe cluster, does not apply to stars older than around 1 Gyr. Although late K dwarfs spin more slowly than early-K dwarfs when they are young, at old ages, we find that late K dwarfs rotate at the same rate or faster than early-K dwarfs of the same age. This result agrees qualitatively with semiempirical models that vary the rate of surface-to-core angular momentum transport as a function of time and mass. It also aligns with recent observations of stars in the NGC 6811 cluster, which indicate that the surface rotation rates of K dwarfs go through an epoch of inhibited evolution. We find that the oldest Kepler stars with measured rotation periods are late K and early M dwarfs, indicating that these stars maintain spotted surfaces and stay magnetically active longer than more massive stars. Finally, based on their kinematics, we confirm that many rapidly rotating GKM dwarfs are likely to be synchronized binaries.

79 ASTRONOMY AND ASTROPHYSICS↗