Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data analysis methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Linking international technical specifications for acoustic characterization of marine energy converter sounds with environmental compliance criteria

As new ocean energy technologies emerge and are deployed for testing and operations, sound emissions are a potential concern for environmental effects to marine life. Consistent acoustic measurement and data analysis methods can help promote comparisons of technologies and transferability between project sites. In 2022, acoustic emissions from a prototype scale wave energy converter (WEC) were characterized for a range of environmental conditions and power generation states in the coastal waters off southern California using a set of international technical specifications. Results from the international technical specification analyses were applied to United States regulatory threshold criteria for acoustic impacts to marine mammals and examined in the context of European underwater noise monitoring guidelines. Weighted 24 hour cumulative sound exposure levels SEL24h calculated from the highest power generation state WEC sound pressure levels were often more than 20 dB below threshold criteria for temporary threshold shifts in five relevant marine mammal hearing groups. Following European Union recommendations for analyses and reporting, WEC sound characterization in third octave bands centered at 63 Hz and 125 Hz show clear spatial decay of WEC-generated noise, with more pronounced attenuation at 63 Hz, and a less marked but still detectable gradient at 125 Hz, collectively suggesting a relatively confined acoustic footprint under the observed conditions. The value of the international technical specification approach is highlighted by the isolation of WEC sounds from the surrounding soundscape. This allows for a robust characterization of acoustic emissions through a range of device power generation and sea states. Furthermore, in threshold-based regulatory contexts like the U.S., this facilitates direct evaluation of source contributions, while in broader monitoring frameworks used in the E.U. it provides a reproducible foundation for assessing the contribution of emerging ocean energy technologies to the underwater acoustic environment.

Haxel, Joseph H. (ORCID:0000000273864761)↗

Status and prospect of in situ and operando characterization of solid-state batteries

Electrification of the transportation sector relies on radical re-imagining of energy storage technologies to provide affordable, high energy density, durable and safe systems. Next generation energy storage systems will need to leverage high energy density anodes and high voltage cathodes to achieve the required performance metrics (longer vehicle range, long life, production costs, safety). Solid-state batteries (SSBs) are promising materials technology for achieving these metrics by enabling these electrode systems due to the underlying material properties of the solid electrolyte (viz. mechanical strength, electrochemical stability, ionic conductivity). Electro-chemo-mechanical degradation in SSBs detrimentally impact the Coulombic efficiencies, capacity retention, durability and safety in SSBs restricting their practical implementation. Solid|solid interfaces in SSBs are hot-spots of dynamics that contribute to the degradation of SSBs. Characterizing and understanding the processes at the solid|solid interfaces in SSBs is crucial towards designing of resilient, durable, high energy density SSBs. This work provides a comprehensive and critical summary of the SSB characterization with a focus on in situ and operando studies. Additionally, perspectives on experimental design, emerging characterization techniques and data analysis methods are provided. Furthermore, this work provides a thorough analysis of current status of SSB characterization as well as highlights important avenues for future work.

25 ENERGY STORAGE↗

Clustering of red sequence galaxies in the fourth data release of the Kilo-Degree Survey

We present a sample of luminous red sequence galaxies as the basis for a study of the large-scale structure in the fourth data release of the Kilo-Degree Survey. The selected galaxies are defined by a red sequence template, in the form of a data-driven model of the colour-magnitude relation conditioned on redshift. In this work, the red sequence template was built using the broad-band optical+near infrared photometry of KiDS-VIKING and the overlapping spectroscopic data sets. The selection process involved estimating the red sequence redshifts, assessing the purity of the sample and estimating the underlying redshift distributions of redshift bins. After performing the selection, we mitigated the impact of survey properties on the observed number density of galaxies by assigning photometric weights to the galaxies. We measured the angular two-point correlation function of the red galaxies in four redshift bins and constrain the large-scale bias of our red sequence sample assuming a fixed ΛCDM cosmology. We find consistent linear biases for two luminosity-threshold samples (‘dense’ and ‘luminous’). We find that our constraints are well characterised by the passive evolution model.

79 ASTRONOMY AND ASTROPHYSICS↗

Photometric redshift-aided classification using ensemble learning

We present SHEEP, a new machine learning approach to the classic problem of astronomical source classification, which combines the outputs from the XGBoost, LightGBM, and CatBoost learning algorithms to create stronger classifiers. A novel step in our pipeline is that prior to performing the classification, SHEEP first estimates photometric redshifts, which are then placed into the data set as an additional feature for classification model training; this results in significant improvements in the subsequent classification performance. SHEEP contains two distinct classification methodologies: (i) Multi-class and (ii) one versus all with correction by a meta-learner. We demonstrate the performance of SHEEP for the classification of stars, galaxies, and quasars using a data set composed of SDSS and WISE photometry of 3.5 million astronomical sources. The resulting F1 -scores are as follows: 0.992 for galaxies; 0.967 for quasars; and 0.985 for stars. In terms of the F1-scores for the three classes, SHEEP is found to outperform a recent RandomForest-based classification approach using an essentially identical data set. Our methodology also facilitates model and data set explainability via feature importances; it also allows the selection of sources whose uncertain classifications may make them interesting sources for follow-up observations.

79 ASTRONOMY AND ASTROPHYSICS↗

The edges of galaxies: Tracing the limits of star formation

The outskirts of galaxies have been studied from multiple perspectives for the past few decades. However, it is still unknown if all galaxies have clear-cut edges similar to everyday objects. We address this question by developing physically motivated criteria to define the edges of galaxies. Based on the gas density threshold required for star formation, we define the edge of a galaxy as the outermost radial location associated with a significant drop in either past or ongoing in situ star formation. We explore ~1000 low-inclination galaxies with a wide range in morphology (dwarfs to ellipticals) and stellar mass (10 7 M ⊙ < M * < 10 12 M ⊙ ). The location of the edges of these galaxies (R edge ) were visually identified as the outermost cutoff or truncation in their radial profiles using deep multi-band optical imaging from the IAC Stripe82 Legacy Project. We find this characteristic feature at the following mean stellar mass density, which varies with galaxy morphology: 2.9 ± 0.10 M ⊙ pc -2 for ellipticals, 1.1 ± 0.04 M ⊙ pc -2 for spirals, and 0.6 ± 0.03 M ⊙ pc -2 for present-day star-forming dwarfs. Additionally, we find that R edge depends on its age (colour) where bluer galaxies have larger R edge at a fixed stellar mass. The resulting stellar mass–size plane using R edge as a physically motivated galaxy size measure has a very narrow intrinsic scatter (≲0.06 dex). These results highlight the importance of new deep imaging surveys to explore the growth of galaxies and trace the limits of star formation in their outskirts.

79 ASTRONOMY AND ASTROPHYSICS↗

VarIabiLity seLection of AstrophysIcal sources iN PTF (VILLAIN): I. Structure function fits to 71 million objects

Light-curve variability is well-suited to characterizing objects in surveys with high cadence and a long baseline. This is especially relevant in view of the large datasets to be produced by the Vera C. Rubin Observatory Legacy Survey of Space and Time (LSST). We aim to determine variability parameters for objects in the Palomar Transient Factory (PTF) and explore differences between quasars (QSOs), stars, and galaxies. We relate variability and colour information in preparation for future surveys. We fit joint likelihoods to structure functions (SFs) of 71 million PTF light curves with a Markov chain Monte Carlo method. For each object, we assume a power-law SF and extract two parameters: the amplitude on timescales of one year, A, and a power-law index, γ. With these parameters and colours in the optical (Pan-STARRS1) and mid-infrared (WISE), we identify regions of parameter space dominated by different types of spectroscopically confirmed objects from SDSS. Candidate QSOs, stars, and galaxies are selected to show their parameter distributions. QSOs show high-amplitude variations in the R band, and the highest γ values. Galaxies have a broader range of amplitudes and their variability shows relatively little dependency on timescale. With variability and colours, we achieve a photometric selection purity of 99.3% for QSOs. Even though hard cuts in monochromatic variability alone are not as effective as seven-band magnitude cuts, variability is useful in characterizing object subclasses. Through variability, we also find QSOs that were erroneously classified as stars in the SDSS. We discuss perspectives and computational solutions in view of the upcoming LSST.

79 ASTRONOMY AND ASTROPHYSICS↗

Informed total-error-minimizing priors: Interpretable cosmological parameter constraints despite complex nuisance effects

While Bayesian inference techniques are standard in cosmological analyses, it is common to interpret resulting parameter constraints with a frequentist intuition. This intuition can fail, for example, when marginalizing high-dimensional parameter spaces onto subsets of parameters, because of what has come to be known as projection effects or prior volume effects. We present the method of informed total-error-minimizing (ITEM) priors to address this problem. An ITEM prior is a prior distribution on a set of nuisance parameters, such as those describing astrophysical or calibration systematics, intended to enforce the validity of a frequentist interpretation of the posterior constraints derived for a set of target parameters (e.g., cosmological parameters). Our method works as follows. For a set of plausible nuisance realizations, we generate target parameter posteriors using several different candidate priors for the nuisance parameters. We reject candidate priors that do not accomplish the minimum requirements of bias (of point estimates) and coverage (of confidence regions among a set of noisy realizations of the data) for the target parameters on one or more of the plausible nuisance realizations. Of the priors that survive this cut, we select the ITEM prior as the one that minimizes the total error of the marginalized posteriors of the target parameters. As a proof of concept, we applied our method to the density split statistics measured in Dark Energy Survey Year 1 data. We demonstrate that the ITEM priors substantially reduce prior volume effects that otherwise arise and that they allow for sharpened yet robust constraints on the parameters of interest.

79 ASTRONOMY AND ASTROPHYSICS↗

CIRCLEZ : Reliable photometric redshifts for active galactic nuclei computed solely using photometry from Legacy Survey Imaging for DESI

Photometric redshifts for galaxies hosting an accreting supermassive black hole in their center, known as active galactic nuclei (AGNs), are notoriously challenging. At present, they are most optimally computed via spectral energy distribution (SED) fittings, assuming that deep photometry for many wavelengths is available. However, for AGNs detected from all-sky surveys, the photometry is limited and provided by a range of instruments and studies. This makes the task of homogenizing the data challenging, presenting a dramatic drawback for the millions of AGNs that wide surveys such as SRG/eROSITA are poised to detect. This work aims to compute reliable photometric redshifts for X-ray-detected AGNs using only one dataset that covers a large area: the tenth data release of the Imaging Legacy Survey (LS10) for DESI. LS10 provides deep grizW1-W4 forced photometry within various apertures over the footprint of the eROSITA-DE survey, which avoids issues related to the cross-calibration of surveys. We present the results from CIRCLEZ, a machine-learning algorithm based on a fully connected neural network. CIRCLEZ is built on a training sample of 14 000 X-ray-detected AGNs and utilizes multi-aperture photometry, mapping the light distribution of the sources. The accuracy (σNMAD) and the fraction of outliers (η) reached in a test sample of 2913 AGNs are equal to 0.067 and 11.6%, respectively. The results are comparable to (or even better than) what was previously obtained for the same field, but with much less effort in this instance. We further tested the stability of the results by computing the photometric redshifts for the sources detected in CSC2 and Chandra-COSMOS Legacy, reaching a comparable accuracy as in eFEDS when limiting the magnitude of the counterparts to the depth of LS10. The method can be applied to fainter samples of AGNs using deeper optical data from future surveys (for example, LSST, Euclid), granting LS10-like information on the light distribution beyond the morphological type. Along with this paper, we have released an updated version of the photometric redshifts (including errors and probability distribution functions) for eROSITA/eFEDS.

79 ASTRONOMY AND ASTROPHYSICS↗

Counterpart identification and classification for eRASS1 and characterisation of the active galactic nuclei content

Context. Accurately accounting for the Active Galactic Nucleus (AGN) phase in galaxy evolution requires a large, clean AGN sample. This is now possible with SRG/eROSITA, which completed its first all-sky X-ray survey (eRASS1) on June 12, 2020. The public Data Release 1 (DR1, Jan 31, 2024) includes 930,203 sources from the western Galactic hemisphere. Aims. The data enable the selection of a large AGN sample and the discovery of rare sources. However, scientific return depends on accurate characterisation of the X-ray emitters, requiring high-quality multi-wavelength data. This paper presents the identification and classification of optical and infrared counterparts to eRASS1 sources. Methods. Counterparts to eRASS1 X-ray point sources were identified using Gaia DR3, CatWISE2020, and Legacy Survey DR10 (LS10) with the Bayesian NWAY algorithm and trained priors. Sources were classified as Galactic or extragalactic via a machine-learning model combining optical/IR and X-ray properties, trained on a reference sample. For extragalactic LS10 sources, photometric redshifts were computed using CIRCLEZ. Results. Within the LS10 footprint, all 656,614 eROSITA/DR1 sources have at least one possible optical counterpart; ∼570 000 are extragalactic and likely AGN. Half are new detections compared to AllWISE, Gaia, and Quaia AGN catalogues. Gaia and CatWISE2020 counterparts are less reliable, due to the survey’s shallowness and the limited amount of features available to assess the probability of being an X-ray emitter. In the Galactic plane, where the overdensity of stellar sources also increases the chance of associations, using conservative reliability cuts, we identified approximately 18 000 Gaia and 55 000 CatWISE2020 extragalactic sources. Conclusions. We have released three high-quality counterpart catalogues – plus the training and validation sets – as a benchmark for the field. These datasets have many applications, but in particular, they empower researchers to build AGN samples tailored for completeness and purity, accelerating the hunt for the Universe’s most energetic engines.

X-rays: general↗

An accurate measurement of the spectral resolution of the JWST Near Infrared Spectrograph

The spectral resolution (R ≡ λ/Δλ) of spectroscopic data is crucial information for accurate kinematic measurements. In this letter we present a robust measurement of the spectral resolution of the JWST Near Infrared Spectrograph (NIRSpec) in fixed slit (FS) and integral field spectroscopy (IFS) modes. Due to the similarity of the utilized slit dimension in the FS mode to that of the shutters in the multi-object spectroscopy (MOS) mode, our resolution measurements in the FS mode can also be used for the MOS mode in principle. We modeled H and He lines of the planetary nebula SMP LMC 58 using a Gaussian line spread function (LSF) to estimate the wavelength-dependent resolution for multiple disperser and filter combinations. We corrected for the intrinsic width of the planetary nebula’s H and He lines due to its expansion velocity by measuring it from a higher-resolution X-shooter spectrum. We find that NIRSpec’s in-flight spectral resolutions exceed the pre-launch estimates provided in the JWST User Documentation by 11–53% in the FS mode and by 1–24% in the IFS mode across the covered wavelengths. We recover the expected trend that the resolution increases with the wavelength within a configuration. The robust and accurate LSFs presented in this letter will enable high-accuracy kinematic measurements using NIRSpec for applications in cosmology and galaxy evolution.

methods: data analysis↗

Model independent approach for calculating galaxy rotation curves for low S/N MaNGA galaxies

Internal kinematics of galaxies, traced through the stellar rotation curve or two dimensional velocity map, carry important information on galactic structure and dark matter. With upcoming surveys, the velocity map may play a key role in the development of kinematic lensing as an astrophysical probe. Here, we improve techniques for extracting velocity information from integral field spectroscopy at low signal-to-noise (S/N), without a template, and demonstrate substantial advantages over the standard Penalized PiXel-Fitting method (pPXF) approach. Robust rotation curves can be derived down to S/N ≈ 2 using our method.

79 ASTRONOMY AND ASTROPHYSICS↗

DeepSZ: identification of Sunyaev–Zel’dovich galaxy clusters using deep learning

Galaxy clusters identified via the Sunyaev–Zel’dovich (SZ) effect are a key ingredient in multiwavelength cluster cosmology. In this paper, we present and compare three methods of cluster identification: the standard matched filter (MF) method in SZ cluster finding, a convolutional neural networks (CNN), and a ‘combined’ identifier. We apply the methods to simulated millimeter maps for several observing frequencies for a survey similar to SPT-3G, the third-generation camera for the South Pole Telescope. The MF requires image pre-processing to remove point sources and a model for the noise, while the CNN requires very little pre-processing of images. Additionally, the CNN requires tuning of hyperparameters in the model and takes cut-out images of the sky as input, identifying the cut-out as cluster-containing or not. We compare differences in purity and completeness. The MF signal-to-noise ratio depends on both mass and redshift. Our CNN, trained for a given mass threshold, captures a different set of clusters than the MF, some with signal-to-noise-ratio below the MF detection threshold. However, the CNN tends to mis-classify cut-out whose clusters are located near the edge of the cut-out, which can be mitigated with staggered cut-out. We leverage the complementarity of the two methods, combining the scores from each method for identification. The purity and completeness are both 0.61 for MF, and 0.59 and 0.61 for CNN. The combined method yields 0.60 and 0.77, a significant increase for completeness with a modest decrease in purity. We advocate for combined methods that increase the confidence of many low signal-to-noise clusters.

79 ASTRONOMY AND ASTROPHYSICS↗

Uncertainty quantification of the virial black hole mass with conformal prediction

Precise measurements of the black hole mass are essential to gain insight on the black hole and host galaxy co-evolution. A direct measure of the black hole mass is often restricted to nearest galaxies and instead, an indirect method using the single-epoch virial black hole mass estimation is used for objects at high redshifts. However, this method is subjected to biases and uncertainties as it is reliant on the scaling relation from a small sample of local active galactic nuclei. In this study, we propose the application of conformalized quantile regression (CQR) to quantify the uncertainties of the black hole predictions in a machine learning setting. We compare CQR with various prediction interval techniques and demonstrated that CQR can provide a more useful prediction interval indicator. In contrast to baseline approaches for prediction interval estimation, we show that the CQR method provides prediction intervals that adjust to the black hole mass and its related properties. That is it yields a tighter constraint on the prediction interval (hence more certain) for a larger black hole mass, and accordingly, bright and broad spectral line width source. Using a combination of neural network model and CQR framework, the recovered virial black hole mass predictions and uncertainties are comparable to those measured from the Sloan Digital Sky Survey. The code is publicly available.

79 ASTRONOMY AND ASTROPHYSICS↗

On the impact of the galaxy window function on cosmological parameter estimation

One important source of systematics in galaxy redshift surveys comes from the estimation of the galaxy window function. Up until now, the impact of the uncertainty in estimating the galaxy window function on parameter inference has not been properly studied. In this paper, we show that the uncertainty and the bias in estimating the galaxy window function will be salient for ongoing and next-generation galaxy surveys using a simulation-based approach. With a specific case study of cross-correlating emission-line galaxies from the DESI Legacy Imaging Surveys and the Planck cosmic microwave background lensing map, we show that neural network-based regression approaches to modelling the window function are superior in comparison to linear regression-based models. Here, we additionally show that the definition of the galaxy overdensity estimator can impact the overall signal-to-noise of observed power spectra. Finally, we show that the additive biases coming from the window functions can significantly bias the modes of the inferred parameters and also degrade their precision. Thus, a careful understanding of the window functions will be essential to conduct cosmological experiments.

79 ASTRONOMY AND ASTROPHYSICS↗

Decoding the age–chemical structure of the Milky Way disc: an application of copulas and elicitable maps

In the Milky Way, the distribution of stars in the [α/Fe] versus [Fe/H] and [Fe/H] versus age planes holds essential information about the history of star formation, accretion, and dynamical evolution of the Galactic disc. We investigate these planes by applying novel statistical methods called copulas and elicitable maps to the ages and abundances of red giants in the Apache Point Observatory Galactic Evolution Experiment survey. We find that the high- and low-α disc stars have a clean separation in copula space and use this to provide an automated separation of the α sequences using a purely statistical approach. This separation reveals that the high-α disc ends at the same [α/Fe] and age at high [Fe/H] as the low-[Fe/H] start of the low-α disc, thus supporting a sequential formation scenario for the high- and low-α discs. We then combine copulas with elicitable maps to precisely obtain the correlation between stellar age τ and metallicity [Fe/H] conditional on Galactocentric radius R and height z in the range 0 < R < 20 kpc and |z| < 2 kpc. The resulting trends in the age–metallicity correlation with radius, height, and [α/Fe] demonstrate a ≈0 correlation wherever kinematically cold orbits dominate, while the naively expected negative correlation is present where kinematically hot orbits dominate. This is consistent with the effects of spiral-driven radial migration, which must be strong enough to completely flatten the age–metallicity structure of the low-α disc.

79 ASTRONOMY AND ASTROPHYSICS↗

Emission line predictions for mock galaxy catalogues: a new differentiable and empirical mapping from DESI

ABSTRACT We present a simple, differentiable method for predicting emission line strengths from rest-frame optical continua using an empirically determined mapping. Extensive work has been done to develop mock galaxy catalogues that include robust predictions for galaxy photometry, but reliably predicting the strengths of emission lines has remained challenging. Our new mapping is a simple neural network implemented using the JAX Python automatic differentiation library. It is trained on Dark Energy Spectroscopic Instrument Early Release data to predict the equivalent widths (EWs) of the eight brightest optical emission lines (including H α, H β, [O ii], and [O iii]) from a galaxy’s rest-frame optical continuum. The predicted EW distributions are consistent with the observed ones when noise is accounted for, and we find Spearman’s rank correlation coefficient ρs > 0.87 between predictions and observations for most lines. Using a non-linear dimensionality reduction technique, we show that this is true for galaxies across the full range of observed spectral energy distributions. In addition, we find that adding measurement uncertainties to the predicted line strengths is essential for reproducing the distribution of observed line-ratios in the BPT diagram. Our trained network can easily be incorporated into a differentiable stellar population synthesis pipeline without hindering differentiability or scalability with GPUs. A synthetic catalogue generated with such a pipeline can be used to characterize and account for biases in the spectroscopic training sets used for training and calibration of photo-z’s, improving the modelling of systematic incompleteness for the Rubin Observatory LSST and other surveys.

79 ASTRONOMY AND ASTROPHYSICS↗

Validating sequential Monte Carlo for gravitational-wave inference

Nested sampling (NS) is the preferred stochastic sampling algorithm for gravitational-wave inference for compact binary coalescences. It can handle the complex nature of the gravitational-wave likelihood surface and provides an estimate of the Bayesian model evidence. However, there is another class of algorithms that meets the same requirements, but has not been used for gravitational-wave analyses: sequential Monte Carlo (SMC), an extension of importance sampling that maps samples from an initial density to a target density via a series of intermediate densities. In this work, we validate a type of SMC algorithm, called persistent sampling (PS), for gravitational-wave inference. We consider a range of different scenarios including binary black holes and binary neutron stars and real and simulated data and show that PS produces results that are consistent with NS whilst being, on average, 2 times more efficient and 2.74 times faster. This demonstrates that PS is a viable alternative to NS that should be considered for future gravitational-wave analyses.

black hole mergers↗

The AXEAP2 program for K β X-ray emission spectra analysis using artificial intelligence

The processing and analysis of synchrotron data can be a complex task, requiring specialized expertise and knowledge. Our previous work addressed the challenge of X-ray emission spectrum (XES) data processing by developing a standalone application using unsupervised machine learning. However, the task of analyzing the processed spectra remains another challenge. Although the non-resonant K β XES of 3 d transition metals are known to provide electronic structure information such as oxidation and spin state, finding appropriate parameters to match experimental data is a time-consuming and labor-intensive process. Here, a new XES data analysis method based on the genetic algorithm is demonstrated, applying it to Mn, Co and Ni oxides. This approach is also implemented as a standalone application, Argonne X-ray Emission Analysis 2 ( AXEAP2 ), which finds a set of parameters that result in a high-quality fit of the experimental spectrum with minimal intervention. AXEAP2 is able to find a set of parameters that reproduce the experimental spectrum, and provide insights into the 3 d electron spin state, 3 d –3 p electron exchange force and K β emission core-hole lifetime.

36 MATERIALS SCIENCE↗