Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Functional data analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Non-universal stellar initial mass functions: large uncertainties in star formation rates at z ≈ 2–4 and other astrophysical probes

ABSTRACT We explore the assumption, widely used in many astrophysical calculations, that the stellar initial mass function (IMF) is universal across all galaxies. By considering both a canonical broken-power-law IMF and a non-universal IMF, we are able to compare the effect of different IMFs on multiple observables and derived quantities in astrophysics. Specifically, we consider a non-universal IMF that varies as a function of the local star formation rate, and explore the effects on the star formation rate density (SFRD), the extragalactic background light, the supernova (both core-collapse and thermonuclear) rates, and the diffuse supernova neutrino background. Our most interesting result is that our adopted varying IMF leads to much greater uncertainty on the SFRD at $z \approx 2-4$ than is usually assumed. Indeed, we find an SFRD (inferred using observed galaxy luminosity distributions) that is a factor of $\gtrsim 3$ lower than canonical results obtained using a universal IMF. Secondly, the non-universal IMF we explore implies a reduction in the supernova core-collapse rate of a factor of $\sim 2$, compared against a universal IMF. The other potential tracers are only slightly affected by changes to the properties of the IMF. We find that currently available data do not provide a clear preference for universal or non-universal IMF. However, improvements to measurements of the star formation rate and core-collapse supernova rate at redshifts $z \gtrsim 2$ may offer the best prospects for discernment.

79 ASTRONOMY AND ASTROPHYSICS↗

On single-crystal total scattering data reduction and correction protocols for analysis in direct space

Data reduction and correction steps and processed data reproducibility in the emerging single-crystal total-scattering-based technique of three-dimensional differential atomic pair distribution function (3D-ΔPDF) analysis are explored. All steps from sample measurement to data processing are outlined using a crystal of CuIr 2 S 4 as an example, studied in a setup equipped with a high-energy X-ray beam and a flat-panel area detector. Computational overhead as pertains to data sampling and the associated data-processing steps is also discussed. Various aspects of the final 3D-ΔPDF reproducibility are explicitly tested by varying the data-processing order and included steps, and by carrying out a crystal-to-crystal data comparison. Situations in which the 3D-ΔPDF is robust are identified, and caution against a few particular cases which can lead to inconsistent 3D-ΔPDFs is noted. Although not all the approaches applied herein will be valid across all systems, and a more in-depth analysis of some of the effects of the data-processing steps may still needed, the methods collected herein represent the start of a more systematic discussion about data processing and corrections in this field.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Optical Variability of ICRF3 Quasars in the Pan-STARRS 3Pi Survey with Functional Principal Component Analysis

We make use of individual (epoch) detection data from the Pan-STARRS “3π” survey for 2863 optical ICRF3 counterparts in the five wavelength bands g, r, i, z, and y, published as part of the Data Release 2. A dedicated method based on the Functional Principal Component Analysis is developed for these sparse and irregularly sampled data. With certain regularization and normalization constraints, it allows us to obtain uniform and compatible estimates of the variability amplitudes and average magnitudes between the passbands and objects. We find that the starting assumption of affinity of the light curves for a given object at different wavelengths is violated for several percent of the sample. The distributions of rms variability amplitudes are strongly skewed toward small values, peaking at ∼0.1 mag with tails stretching to 2 mag. Statistically, the lowest variability is found for the r band and the largest for the reddest y band. A small “brighter-redder” effect is present, with amplitudes in y greater than amplitudes in g in 57% of the sample. The variability versus redshift dependence shows a strong decline with z toward redshift 3, which we interpret as the time dilation of the dominant time frequencies. The colors of radio-loud ICRF3 quasars are correlated with redshift in a complicated, wavy pattern governed by the emergence of brightest emission lines within the five passbands.

79 ASTRONOMY AND ASTROPHYSICS↗

The Completed SDSS-IV Extended Baryon Oscillation Spectroscopic Survey: N-body Mock Challenge for Galaxy Clustering Measurements

We develop a series of N-body data challenges, functional to the final analysis of the extended Baryon Oscillation Spectroscopic Survey (eBOSS) Data Release 16 (DR16) galaxy sample. The challenges are primarily based on high-fidelity catalogues constructed from the Outer Rim simulation - a large box size realization (3h(-1) Gpc) characterized by an unprecedented combination of volume and mass resolution, down to 1.85 x 10(9) h(-1)M(circle dot). We generate synthetic galaxy mocks by populating Outer Rim haloes with a variety of halo occupation distribution (HOD) schemes of increasing complexity, spanning different redshift intervals. We then assess the performance of three complementary redshift space distortion (RSD) models in configuration and Fourier space, adopted for the analysis of the complete DR16 eBOSS sample of Luminous Red Galaxies (LRG5). We find all the methods mutually consistent, with comparable systematic errors on the Alcock-Paczynski parameters and the growth of structure, and robust to different HOD prescriptions - thus validating the robustness of the models and the pipelines used for the baryon acoustic oscillation (BAO) and full shape clustering analysis. In particular, all the techniques are able to recover and alpha(11) to within 0.9 per cent, and f sigma(8) to within 1.5 per cent. As a by-product of our work, we are also able to gain interesting insights on the galaxy-halo connection. Our study is relevant for the final eBOSS DR16 'consensus cosmology', as the systematic error budget is informed by testing the results of analyses against these high-resolution mocks. In addition, it is also useful for future large-volume surveys, since similar mock-making techniques and systematic corrections can be readily extended to model for instance the Dark Energy Spectroscopic Instrument (DESI) galaxy sample.

cosmology: theory, large-scale structure of Univer↗

Pursuing Heteroleptic Ligand Design Principles for Photoactive Fe Complexes with Ultrafast X-ray Emission and Variable-Temperature Optical Spectroscopies

Understanding the key parameters that govern the photophysical and photochemical properties of transition metal complexes is essential for the development of efficient photosensitizers for photocatalytic applications. Achieving this objective necessitates clear and detailed investigations of their electronic excited states, for which time-resolved metal Kβ X-ray emission spectroscopy (XES) has proven highly effective. Here, we present a time-resolved Fe Kβ XES study of a heteroleptic Fe(II) polypyridyl carbene complex, [Fe(phen) 2 (C 4 H 10 N 4 )] 2+ (1; phen = 1,10-phenanthroline), utilizing both the valence-to-core and Kβ mainline spectral regions, complemented by variable-temperature transient optical absorption (VT-TA) spectroscopy. Detailed analysis of the time-resolved Kβ XES data, supported by density functional theory (DFT) calculations and an Eyring analysis of the VT-TA data, reveals parallel excited state relaxation dynamics that support an assignment of the long-lived excited state to a triplet metal-centered state. Placing these results in the context of prior studies of heteroleptic Fe(II) polypyridyl cyanide complexes motivated a series of DFT calculations to investigate the effects of ligand structural flexibility and arrangement. These calculations reinforce the experimentally derived conclusion that constraining structural flexibility with multidentate ligands significantly impacts the excited state relaxation dynamics. Furthermore, our study emphasizes that the arrangement of strong field ligands in heteroleptic complexes substantially affects the energy of Jahn–Teller active triplet metal-centered states in low-spin d 6 metal complexes. Together, these findings provide synthetic design principles for extending metal-to-ligand charge transfer excited state lifetimes of heteroleptic Fe complexes.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Real-space visualization of short-range antiferromagnetic correlations in a magnetically enhanced thermoelectric

Short-range magnetic correlations can significantly increase the thermopower of magnetic semiconductors, representing a noteworthy development in the decades-long effort to develop high-performance thermoelectric materials. Here, we reveal the nature of the thermopower-enhancing magnetic correlations in the antiferromagnetic semiconductor MnTe. Using magnetic pair distribution function analysis of neutron scattering data, we obtain a detailed, real-space view of robust, nanometer-scale, antiferromagnetic correlations that persist into the paramagnetic phase above the Neel temperature $T_N$ = 307 K. In this work, the magnetic correlation length in the paramagnetic state is significantly longer along the crystallographic c axis than within the ab plane, pointing to anisotropic magnetic interactions. Ab initio calculations of the spin-spin correlations using density functional theory in the disordered local moment approach reproduce this result with quantitative accuracy. These findings constitute the first real-space picture of short-range spin correlations in a magnetically enhanced thermoelectric and inform future efforts to optimize thermoelectric performance by magnetic means.

36 MATERIALS SCIENCE↗

Low-energy interband transition in the infrared response of the correlated metal SrVO 3 in the ultraclean limit

We studied the low-energy electronic response of the prototypical correlated metal SrVO 3 in the ultraclean and disordered limit using infrared spectroscopy and density functional theory plus dynamical mean field theory calculations (DFT+DMFT). A strong optical excitation at 70 meV is observed in the optical response of the ultraclean samples but is hidden by the low-energy Drude-like response from intraband excitations in the more disordered samples. DFT+DMFT calculations reveal that this optical excitation originates from interband transitions between the bands split by orbital off-diagonal hopping, which has often been ignored in cubic systems, such as SrVO 3 . A memory function analysis of the optical data shows that this interband transition can lead to deviations of optical self-energy from the expected Fermi-liquid behavior. Our findings demonstrate that analysis schemes employed to extract many-body effects from optical spectra may be oversimplified to study the true electronic ground state and that improvements in material quality can guide efforts to refine theoretical approaches.

36 MATERIALS SCIENCE↗

Enzyme Engineering Database (EnzEngDB): a platform for sharing and interpreting sequence–function relationships across protein engineering campaigns

The discovery and engineering of new enzymes is important across the bioeconomy, with diverse applications from foods to pharmaceuticals, sensors to agriculture. However, enzyme engineering, in particular machine learning-guided engineering, is hampered by a lack of data. Currently there exists no database designed to capture and interpret datasets created in this domain, nor are there easy analysis and visualisation tools. We developed the Enzyme Engineering Database to provide a centralized resource and an online analysis tool to consolidate sequence-function data from enzyme engineering campaigns, thereby making three contributions: (i) a database into which researchers can deposit public data, (ii) visualisation and analysis tools for protein engineers to analyse their own data or compare enzyme variants to other engineering campaigns, and (iii) a gold-standard dataset for benchmarking automated extraction along with the first large language model extraction pipeline specific for enzyme engineering campaigns. The Enzyme Engineering Database is accessible at http://enzengdb.org/.

Long, Yueming [California Institute of Technology ↗

Trimming and Decontamination of Metagenomic Data can Significantly Impact Assembly and Binning Metrics, Phylogenomic and Functional Analysis

Background: Investigators using metagenomic sequencing to study microbiomes often trim and decontaminate reads without knowing their effect on downstream analyses. Objective: This study was designed to evaluate the impacts JGI trimming and decontamination procedures have on assembly and binning metrics, placement of MAGs into species trees, and functional profiles of MAGs extracted from complex rhizosphere metagenomes, as well as how more aggressive trimming impacts these binning metrics. Methods: Twenty-three Miscanthus x giganteus rhizosphere metagenomes were subjected to different combinations and thresholds of force, kmer, and quality trimming and decontamination using BBDuk. Reads were assembled and binned in KBase. Phylogenomic and statistical analyses were applied to evaluate the effects of trimming and decontamination on downstream analyses. Results: We found that JGI trimmed and decontaminated reads had significant impacts on assembly and binning metrics compared to raw reads, including significantly higher total contig counts, more contigs greater than 10k bp in length, and larger total lengths of raw assemblies compared to QC assemblies, and 2.0% lower average contamination of QC MAGs compared to raw MAGs. We also found that differences in the placement of MAGs in species trees increased with decreasing completeness and contamination thresholds. Furthermore, aggressive trimming (Q20) was found to significantly reduce MAG counts. Conclusion: Trimming and decontamination of metagenomics reads prior to assembly can change an investigator’s answer to the questions, “Who is there and what are they doing?” However, mild trimming and decontamination of metagenomic reads with high-quality scores are recommended for removing sample processing and sequencing artifacts.

Whitham, Jason M.↗

A Robotics Enabled Eddy Current Testing System for Autonomous Inspection of Heat Exchanger Tubes

The objective of the project is to develop a robotics enabled eddy current testing system (REECTS) in automatic probe deployment, inspection, and data acquisition and analysis. The main functions of the REECTS are to: 1) identify geometry and locations of heat exchange tubes with assistance of an imaging recognition system; 2) precisely control the position and motion speed of ECT probes by an adaptive control system; 3) facilitate data analysis and real-time decision making for autonomous inspection assisted by machine learning algorithms.

20 FOSSIL-FUELED POWER PLANTS↗

Guiding the choice of informatics software and tools for lipidomics research applications

Progress in mass spectrometry lipidomics has led to a rapid proliferation of studies across biology and biomedicine. These generate extremely large raw datasets requiring sophisticated solutions to support automated data processing. To address this, numerous software tools have been developed and tailored for specific tasks. However, for researchers, deciding which approach best suits their application relies on ad hoc testing, which is inefficient and time consuming. Here we first review the data processing pipeline, summarizing the scope of available tools. Next, to support researchers, LIPID MAPS provides an interactive online portal listing open-access tools with a graphical user interface. This guides users towards appropriate solutions within major areas in data processing, including (1) lipid-oriented databases, (2) mass spectrometry data repositories, (3) analysis of targeted lipidomics datasets, (4) lipid identification and (5) quantification from untargeted lipidomics datasets, (6) statistical analysis and visualization, and (7) data integration solutions. Detailed descriptions of functions and requirements are provided to guide customized data analysis workflows.

59 BASIC BIOLOGICAL SCIENCES↗

Local structure of Mott insulating iron oxychalcogenides La 2 O 2 Fe 2 OM2(M=S,Se)

Here, we describe the local structural properties of the iron oxychalcogenides, La 2 O 2 Fe 2 OM 2 (M=S,Se), by using pair distribution function analysis applied to total scattering data. Our results from neutron powder diffraction show that M = S and Se possess similar nuclear structures at low and room temperatures. The local crystal structures were studied by investigating deviations in atomic positions and the extent of the formation of orthorhombicity. Analysis of the total scattering data suggests that buckling of the Fe 2 O plane occurs below 100 K. The buckling may occur concomitantly with a change in octahedral height. Furthermore, within a typical range of 1-2 nm, we observed a short-range orthorhombiclike structure suggestive of nematic fluctuations in both of these materials.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Identification and characterization of proteins of unknown function (PUFs) in Clostridium thermocellum DSM 1313 strains as potential genetic engineering targets

Abstract Background Mass spectrometry-based proteomics can identify and quantify thousands of proteins from individual microbial species, but a significant percentage of these proteins are unannotated and hence classified as proteins of unknown function (PUFs). Due to the difficulty in extracting meaningful metabolic information, PUFs are often overlooked or discarded during data analysis, even though they might be critically important in functional activities, in particular for metabolic engineering research. Results We optimized and employed a pipeline integrating various “guilt-by-association” (GBA) metrics, including differential expression and co-expression analyses of high-throughput mass spectrometry proteome data and phylogenetic coevolution analysis, and sequence homology-based approaches to determine putative functions for PUFs in Clostridium thermocellum . Our various analyses provided putative functional information for over 95% of the PUFs detected by mass spectrometry in a wild-type and/or an engineered strain of C. thermocellum . In particular, we validated a predicted acyltransferase PUF (WP_003519433.1) with functional activity towards 2-phenylethyl alcohol, consistent with our GBA and sequence homology-based predictions. Conclusions This work demonstrates the value of leveraging sequence homology-based annotations with empirical evidence based on the concept of GBA to broadly predict putative functions for PUFs, opening avenues to further interrogation via targeted experiments.

09 BIOMASS FUELS↗

Strangeness in the proton from $W+$ charm production and SIDIS data

We perform a global QCD analysis of unpolarized parton distribution functions (PDFs) in the proton, including new 𝑊+⁢ charm production data from 𝑝⁢𝑝 collisions at the LHC and semi-inclusive pion and kaon production data in lepton-nucleon deep-inelastic scattering, both of which have been suggested for constraining the strange quark PDF. Compared with a baseline global fit that does not include these datasets, the new analysis reduces the uncertainty on the strange quark distribution over the range 0.01 < 𝑥 < 0.3, and provides a consistent description of processes sensitive to strangeness in the proton. Including the new datasets, the ratio of strange to nonstrange sea quark distributions is $R_s = (s + \bar{s})/(\bar{u} +\bar{d})$ $=$ {$0.7⁢2^{+0.52}_{−0.34}, 0.4⁢6^{+0.30}_{−0.20}, 0.3⁢2^{+0.23}_{−0.15}$} for 𝑥 ={$0.01, 0.04, 0.1$} at 𝑄 2 $=$ 4 GeV 2 . The data place more stringent constraints on the strange asymmetry $(s - \bar{s})$, which is found to be consistent with zero in this range.

Anderson, Trey [College of William and Mary, Willi↗

Studying baryon acoustic oscillations using photometric redshifts from the DESI Legacy Imaging survey DR9

Context. The Dark Energy Spectroscopic Instrument (DESI) Legacy Imaging Survey DR9 (DR9 hereafter), with its extensive dataset of galaxy locations and photometric redshifts, presents an opportunity to study baryon acoustic oscillations (BAOs) in the region covered by the ongoing spectroscopic survey with DESI. Aims. We aim to investigate differences between different parts of the DR9 footprint. Furthermore, we want to measure the BAO scale for luminous red galaxies within them. Our selected redshift range of 0.6–0.8 corresponds to the bin in which a tension between DESI Y1 and eBOSS was found. Methods. We calculated the anisotropic two-point correlation function in a modified binning scheme to detect the BAOs in DR9 data. We then used template fits based on simulations to measure the BAO scale in the imaging data. Results. Our analysis reveals the expected correlation function shape in most of the footprint areas, showing a BAO scale consistent with Planck’s observations. Aside from identified mask-related data issues in the southern region of the South Galactic Cap, we find a notable variance between the different footprints. Conclusions. We find that this variance is consistent with the difference between the DESI Y1 and eBOSS data, and it supports the argument that that tension is caused by sample variance. Additionally, we also uncovered systematic biases not previously accounted for in photometric BAO studies. We emphasize the necessity of adjusting for the systematic shift in the BAO scale associated with typical photometric redshift uncertainties to ensure accurate measurements.

79 ASTRONOMY AND ASTROPHYSICS↗

Scalable Volume Visualization for Big Scientific Data Modeled by Functional Approximation

Considering the challenges posed by the space and time complexities in handling extensive scientific volumetric data, various data representations have been developed for the analysis of large-scale scientific data. Multivariate functional approximation (MFA) is an innovative data model designed to tackle substantial challenges in scientific data analysis. It computes values and derivatives with high-order accuracy throughout the spatial domain, mitigating artifacts associated with zero- or first-order interpolation. However, the slow query time through MFA makes it less suitable for interactively visualizing a large MFA model. In this work, we develop the first scalable interactive volume visualization pipeline, MFA-DVV, for the MFA model encoded from large-scale datasets. Our method achieves low input latency through distributed architecture, and its performance can be further enhanced by utilizing a compressed MFA model while still maintaining a high-quality rendering result for scientific datasets. We conduct comprehensive experiments to show that MFA-DVV can decrease the input latency and achieve superior visualization results for big scientific data compared with existing approaches.

big scientific dataset↗

Asc-Seurat: analytical single-cell Seurat-based web application

Abstract Background Single-cell RNA sequencing (scRNA-seq) has revolutionized the study of transcriptomes, arising as a powerful tool for discovering and characterizing cell types and their developmental trajectories. However, scRNA-seq analysis is complex, requiring a continuous, iterative process to refine the data and uncover relevant biological information. A diversity of tools has been developed to address the multiple aspects of scRNA-seq data analysis. However, an easy-to-use web application capable of conducting all critical steps of scRNA-seq data analysis is still lacking. Summary We present Asc-Seurat, a feature-rich workbench, providing an user-friendly and easy-to-install web application encapsulating tools for an all-encompassing and fluid scRNA-seq data analysis. Asc-Seurat implements functions from the Seurat package for quality control, clustering, and genes differential expression. In addition, Asc-Seurat provides a pseudotime module containing dozens of models for the trajectory inference and a functional annotation module that allows recovering gene annotation and detecting gene ontology enriched terms. We showcase Asc-Seurat’s capabilities by analyzing a peripheral blood mononuclear cell dataset. Conclusions Asc-Seurat is a comprehensive workbench providing an accessible graphical interface for scRNA-seq analysis by biologists. Asc-Seurat significantly reduces the time and effort required to analyze and interpret the information in scRNA-seq datasets.

60 APPLIED LIFE SCIENCES↗

IER 484: AFRRI Field Characterization Measurements [Slides]

This lecture includes discussion on the future work of finalizing the data analysis with dose as a function of radial distance and ion chamber integral. Further on agenda is to write manuscript of AFRRI dose characterization and write manuscripts of all NCSP-funded NAD work. Finally, it talks on the goal to publish in Radiation Measurements special issue.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗