Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Gaussian processes regression”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Aemulus ν: precision halo mass functions in wνCDM cosmologies

Precise and accurate predictions of the halo mass function for cluster mass scales in wνCDM cosmologies are crucial for extracting robust and unbiased cosmological information from upcoming galaxy cluster surveys. Here, we present a halo mass function emulator for cluster mass scales (≳ 1013 M ⊙/h) up to redshift z = 2 with comprehensive support for the parameter space of wνCDM cosmologies allowed by current data. Based on the Aemulus ν suite of simulations, the emulator marks a significant improvement in the precision of halo mass function predictions by incorporating both massive neutrinos and non-standard dark energy equation of state models. This allows for accurate modeling of the cosmology dependence in large-scale structure and galaxy cluster studies. We show that the emulator, designed using Gaussian Process Regression, has negligible theoretical uncertainties compared to dominant sources of error in future cluster abundance studies. Our emulator is publicly available (https://github.com/DelonShen/aemulusnu_hmf), providing the community with a crucial tool for upcoming cosmological surveys such as LSST and Euclid.

cluster counts↗

Nonlinear gyrokinetic predictions of SPARC burning plasma profiles enabled by surrogate modeling

Multi-channel, nonlinear predictions of core temperature and density profiles are performed for the SPARC tokamak accounting for both kinetic neoclassical and fully nonlinear gyro-kinetic turbulent fluxes. A series of flux-tube, nonlinear, electromagnetic simulations using the CGYRO code with six gyrokinetic species are coupled to a nonlinear optimizer using Gaussian process regression techniques. The simultaneous evolution of energy sources, including alpha heat, radiation, and energy exchange, coupled with these high fidelity models and techniques, leads to a converged solution in electron temperature, ion temperature and electron density channels with a minimal number of expensive gyrokinetic simulations without compromising accuracy.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Using AI to predict calibration constants for the central drift chamber in GlueX at Jefferson Lab

The AI for Experimental Controls project team at Jefferson Lab has developed an AI system to control and calibrate a large drift chamber system in near-real time. The AI system will monitor environmental and experimental variables to recommend voltage settings that maintain consistent dE/dx gain and optimal resolution throughout the experiment. At present, calibrations are performed after data have been recorded and require a considerable amount of time and attention from experts. The calibrations currently require multiple iterations and depend on accurate tracking information. Our approach uses environmental data, such as atmospheric pressure and gas temperature, and beam conditions, such as the flux of incident particles, as inputs to a Gaussian Process Regression (GPR) model. For the data taken during the GlueX 2020 run period, the GPR is able to predict the existing gain correction factors to within 3.5%. This talk will briefly describe the development, testing, and future plans for this system at Jefferson Lab.

Jeske, Torri↗

Machine learning for seismic low-frequency extrapolation

The cycle-skipping problem that plagues full waveform inversion (FWI) can be at least partially mitigated if low frequencies (which encode the kinematics of wave propagation in seismic data) are recorded. However, seismic sources and receivers are band-limited, so seismic data does not generally include signals down to 0 Hz. To improve our ability to solve the seismic inverse problem, one can synthesize this missing low-frequency (LF) content from the recorded high-frequency (HF) data using machine learning (ML) models. Deep learning models such as convolutional neural networks (CNNs) demonstrate impressive ability to perform low frequency extrapolation. However, such models require powerful hardware (GPU machines) and careful training. We assess the extrapolation capabilities of three different ML models that do not require GPU machines, namely, random forest, Gaussian process regression and gradient boosting, on both synthetic and real data. Experimental results on two synthetic data sets (generated from a low velocity lens embedded in a homogeneous medium, and the Marmousi model) demonstrate that FWI applied to the extrapolated data consistently improves inversion accuracy relative to FWI applied to the original data sets that do not contain low frequencies. Application of low-frequency extrapolation to real data from the Northwest Shelf of Australia demonstrates that tree-based ML models such as gradient boosting can outperform CNNs in terms of both accuracy and computational cost on non-GPU architectures.

58 GEOSCIENCES↗

Combined full shape analysis of BOSS galaxies and eBOSS quasars using an iterative emulator

ABSTRACT Standard full-shape clustering analyses in Fourier space rely on a fixed power spectrum template, defined at the fiducial cosmology used to convert redshifts into distances, and compress the cosmological information into the Alcock–Paczynski parameters and the linear growth rate of structure. In this paper, we propose an analysis method that operates directly in the cosmology parameter space and varies the power spectrum template accordingly at each tested point. Predictions for the power spectrum multipoles from the TNS model are computed at different cosmologies in the framework of $\Lambda \rm {CDM}$. Applied to the final eBOSS QSO and LRG samples together with the low-z DR12 BOSS galaxy sample, our analysis results in a set of constraints on the cosmological parameters Ωcdm, H0, σ8, Ωb, and ns. To reduce the number of computed models, we construct an iterative process to sample the likelihood surface, where each iteration consists of a Gaussian process regression. This method is validated with mocks from N-body simulations. From the combined analysis of the (e)BOSS data, we obtain the following constraints: σ8 = 0.877 ± 0.049 and $\Omega _{\rm m}=0.304^{+0.016}_{-0.010}$ without any external prior. The eBOSS quasar sample alone shows a 3.1σ discrepancy compared to the Planck prediction.

79 ASTRONOMY AND ASTROPHYSICS↗

Accelerating cosmological inference with Gaussian processes and neural networks – an application to LSST Y1 weak lensing and galaxy clustering

ABSTRACT Studying the impact of systematic effects, optimizing survey strategies, assessing tensions between different probes and exploring synergies of different data sets require a large number of simulated likelihood analyses, each of which cost thousands of CPU hours. In this paper, we present a method to accelerate cosmological inference using emulators based on Gaussian process regression and neural networks. We iteratively acquire training samples in regions of high posterior probability which enables accurate emulation of data vectors even in high dimensional parameter spaces. We showcase the performance of our emulator with a simulated 3×2 point analysis of LSST-Y1 with realistic theoretical and systematics modelling. We show that our emulator leads to high-fidelity posterior contours, with an order of magnitude speed-up. Most importantly, the trained emulator can be re-used for extremely fast impact and optimization studies. We demonstrate this feature by studying baryonic physics effects in LSST-Y1 3×2 point analyses where each one of our MCMC runs takes approximately 5 min. This technique enables future cosmological analyses to map out the science return as a function of analysis choices and survey strategy.

Astronomy & Astrophysics↗

Model independent comparison of supernova and strong lensing cosmography: Implications for the Hubble constant tension

We use supernovae measurements, calibrated by the local determination of the Hubble constant $H_0$ by SH0ES, to interpolate the distance-redshift relation using Gaussian process regression. We then predict, independent of the cosmological model, the distances that are measured with strong lensing time delays to test their mutual agreement. In this work we find excellent agreement between these predictions and the measurements. The agreement holds when we consider only the redshift dependence of the distance-redshift relation, independent of the value of $H_0$. Our results disfavor the possibility that lens mass modeling contributes a 10% bias or uncertainty in the strong lensing analysis, as suggested recently in the literature. In general our analysis strengthens the case that residual systematic errors in both measurements are below the level of the current discrepancy with the CMB determination of $H_0$, and supports the possibility of new physical phenomena on cosmological scales. With additional data our methodology can provide more stringent tests of unaccounted for systematics in the determinations of the distance-redshift relation in the late universe.

79 ASTRONOMY AND ASTROPHYSICS↗

Lattice QCD estimates of thermal photon production from the QGP

Thermal photons produced in heavy-ion collision experiments are an important observable for understanding quark-gluon plasma (QGP). The thermal photon rate from the QGP at a given temperature can be calculated from the spectral function of the vector current correlator. Extraction of the spectral function from the lattice correlator is known to be an ill-conditioned problem, as there is no unique solution for a spectral function for a given lattice correlator with statistical errors. The vector current correlator, on the other hand, receives a large ultraviolet contribution from the vacuum, which makes the extraction of the thermal photon rate difficult from this channel. We therefore consider the difference between the transverse and longitudinal part of the spectral function, only capturing the thermal contribution to the current correlator, simplifying the reconstruction significantly. The lattice correlator is calculated for light quarks in quenched QCD at T = 470 MeV ( ∼ 1.5 T c ), as well as in 2 + 1 flavor QCD at T = 220 MeV ( ∼ 1.2 T p c ) with m π = 320 MeV . In order to quantify the nonperturbative effects, the lattice correlator is compared with the corresponding NLO + LPM LO estimate of correlator. The reconstruction of the spectral function is performed in several different frameworks, ranging from physics-informed models of the spectral function to more general models in the Backus-Gilbert method and Gaussian process regression. We find that the resulting photon rates agree within errors. Published by the American Physical Society 2024

Astronomy & Astrophysics↗

Error mitigation in variational quantum eigensolvers using tailored probabilistic machine learning

Quantum computing technology has the potential to revolutionize the simulation of materials and molecules in the near future. A primary challenge in achieving near-term quantum advantage is effectively mitigating the noise effects inherent in current quantum processing units (QPUs). This challenge is also decisive in the context of quantum-classical hybrid schemes employing variational quantum eigensolvers (VQEs) that have attracted significant interest in recent years. In this paper, we present a method that employs parametric Gaussian process regression (GPR) within an active learning framework to mitigate noise in quantum computations, focusing on VQEs. Our approach, grounded in probabilistic machine learning, exploits a custom prior based on the VQE ansatz to capture the underlying correlations between VQE outputs for different variational parameters, thereby enhancing both accuracy and efficiency. We demonstrate the effectiveness of our method on a two-site Anderson impurity model and a eight-site Heisenberg model, using the IBM open-source quantum computing framework, Qiskit, showcasing substantial improvements in the accuracy of VQE outputs while reducing the number of direct QPU energy evaluations. This paper contributes to the ongoing efforts in quantum-error mitigation and optimization, bringing us a step closer to realizing the potential of quantum computing in quantum matter simulations. Published by the American Physical Society 2024

97 MATHEMATICS AND COMPUTING↗

A machine learning degradation model for electrochemical capacitors operated at high temperature

Electrochemical capacitors (ECs) have only recently been considered as an alternative power source for telemetry sensors of drilling equipment for geothermal or oil and gas exploration. The lifecycle analysis and modelling of ECs is underrepresented in literature in comparison to other storage devices e.g. Li-ion batteries. This paper investigates the degradation of ECs when cycled outside the manufacturer-specified operating temperature envelope and proposes a machine learning-based approach for modelling the degradation. Experimental results show that end of life, defined as a 30% decrease in capacitance, occurs at 1,000 cycles when the environmental temperature exceeds the maximum operating temperature by 30%. The life-cycle test data is then used as an input to a Gaussian process regression (GPR) algorithm to predict the capacitance fade trend. The GPR is validated on a total of nine commercial cells from two different manufacturers, achieving an average root mean squared percent error of less than 2% and a mean calibration score of 93% when referenced to a 95% confidence interval. The model can be utilized to determine the EC degradation rate at a range of operating temperature values.

42 ENGINEERING↗

Decentralized Voltage Control of Large-Scale Distribution System with PVs Based on MADRL

This paper proposes a model-free decentralized control framework for the voltage regulation of large-scale distribution systems through the coordinated control of PV inverters. This is achieved by developing a novel interaction mechanism between the surrogate model and the centralized training and decentralized execution multiagent deep reinforcement learning framework. Specifically, the sparse Gaussian processes regression method is first utilized to develop the surrogate model of the original distribution system for reward calculation during the training stage, where each agent represents a sub-region in the centralized fashion for coordination strategy learning. After that, the learned control rules are used to inform controllers within each sub-region for real-time decisions with only local measurements. Comparative tests among various methods on the EPRI Ckt5 test system demonstrate the effectiveness of the proposed method.

distribution system↗

S AP F LOWER : an automated tool for sap flow data preprocessing, gap-filling, and analysis using deep learning

Sap flow, a critical process in plant water use and ecosystem water cycles, is often measured using thermal dissipation probes (TDP) due to their ease of installation and continuous data collection. However, sap flow data frequently include noise, outliers, and gaps, creating challenges for analysis and requiring substantial manual processing. We developed S AP F LOWER , a tool that automates data preprocessing, model training, gap-filling, sapwood area scaling and modeling, and water use analysis. It integrates autocleaning, machine learning and deep learning models (e.g. random forest, Gaussian process regression, long short-term memory (LSTM), bidirectional LSTM (BiLSTM)), and efficient workflows to process sap flow data. S AP F LOWER can remove over 90% of noisy data while preserving legitimate variations and achieve high accuracy in gap-filling based on user-determined parameters. Random forest, LSTM, and BiLSTM models reduced root mean square error to 10% or less for long-term gaps. Model training and prediction can be performed efficiently within seconds. S AP F LOWER significantly enhances the efficiency and accessibility of TDP data analysis by automating complex tasks, enabling researchers without programming expertise to employ advanced techniques. Future improvements will focus on species-specific corrections for TDP and support for additional measurement methods. S AP F LOWER is openly available on GitHub (https://github.com/JiaxinWang123/SapFlower) and Zenodo (doi: 10.5281/zenodo.13665919).

ecosystem water balance↗

Tsuchinoko v1.0.0

Tsuchinoko is a Qt application for adaptive experiment execution and tuning. Live visualizations show details of measurements, and provide feedback on the adaptive engine's decision-making process. The parameters of the adaptive engine can also be tuned live to explore and optimize the search procedure. While Tsuchinoko is designed to allow custom adaptive engines to drive experiments, the gpCAM engine is a featured inclusion. This tool is based on a flexible and powerful Gaussian process regression at the core. A Tsuchinoko system includes 4 distinct components: the GUI client, an adaptive engine, and execution engine, and a core service. These components are separable to allow flexibility with a variety of distributed designs.

Pandolfi, Ronald↗

Physics Model For The Agn-201 Dt

The surrogate model is a Gaussian Process Regression model based on sci-kit learn. The model takes the coarse and fine control rod position, along with the temperature of the reactor, and produces a corresponding k-eff value. Given k-eff over time, deviations can be determine and flagged for review at a later date.

Stewart, RyanH. [Idaho National Laboratory (INL), ↗

GP-BayesOpInf

SAND2025-01851O GP-BayesOpInf is a software tool that uses algorithms to combine Gaussian process regression, principal component analysis, and linear Bayesian inference to produce a probabilistic reduced-order model for time-dependent systems. Numerical examples include the compressible Euler equations for an ideal gas, a heat diffusion process with a nonlinear reaction term, and a set of ordinary differential equations describing a compartmental model in epidemiology. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

SciDAC↗

Machine Learning Approach for Spatiotemporal Multivariate Optimization of Environmental Monitoring Sensor Locations

Abstract Long-term environmental monitoring is critical for managing the soil and groundwater at contaminated sites. Recent improvements in state-of-the-art sensor technology, communication networks, and artificial intelligence have created opportunities to modernize this monitoring activity for automated, fast, robust, and predictive monitoring. In such modernization, it is required that sensor locations be optimized to capture the spatiotemporal dynamics of all monitoring variables as well as to make it cost-effective. The legacy monitoring datasets of the target area are important to perform this optimization. In this study, we have developed a machine-learning approach to optimize sensor locations for soil and groundwater monitoring based on ensemble supervised learning and majority voting. For spatial optimization, Gaussian process regression (GPR) is used for spatial interpolation, while the majority voting is applied to accommodate the multivariate temporal dimension. Results show that the algorithms significantly outperform the random selection of the sensor locations for predictive spatiotemporal interpolation. While the method has been applied to a four-dimensional dataset (with two-dimensional space, time, and multiple contaminants), we anticipate that it can be generalizable to higher-dimensional datasets for environmental monitoring sensor location optimization.

Siddiquee, Masudur R.↗

Extratropical Cloud Feedback Constrained by Cloud Sources and Sinks in Cyclones

Constraining cloud feedback in global climate models (GCMs) using observations is important for establishing accurate predictions of future climate. Uncertainty in shortwave cloud feedback (SW FB ) dominates uncertainty in total cloud feedback. Recent studies show a shift toward more positive extratropical SW FB in the latest generations of GCMs leading to the emergence of very high equilibrium climate sensitivity (ECS). In this study, we use precipitation efficiency and albedo susceptibility to constrain liquid water path (LWP) response to warming and SW FB in the Southern Ocean (SO; 50°–80°S). We analyze precipitation in extratropical cyclones (ECs) to learn about extratropical condensed water sink processes, combined with observations of clouds and moisture convergence, and use the analysis to better understand and constrain SW FB . We utilize a perturbed parameter ensemble (PPE) hosted in the Community Atmosphere Model, version 6 (CAM6), to provide a constraint on SW FB based on observations from Clouds and the Earth’s Radiant Energy System (CERES) and Multisensor Advanced Climatology of LWP (MAC-LWP). We apply Gaussian process regression to emulate the model response to all parameters perturbed in the PPE. Confronting the emulator output with observations provides a new estimated response of Earth to global warming. Furthermore, our new estimates of SO LWP reduce the PPE range by 66%–72%, which results in a shortwave cloud radiative effect estimated range that is 27%–34% less than the PPE range. Observations suggest a more positive SO SW FB than the Community Earth System Model, version 2 (CESM2), and consequently do not reject the high climate sensitivity GCMs emerging from the Coupled Model Intercomparison Project phase 6 (CMIP6).

Atmosphere↗

Machine learning enhanced characterization and optimization of photonic cured MAPbI 3 for efficient perovskite solar cells

Photonic curing (PC) can facilitate high-speed perovskite solar cell (PSC) manufacturing because it uses high-intensity light pulses to crystallize perovskite films in milliseconds. However, optimizing PC conditions is challenging due to its many variables, and using power conversion efficiency (PCE) as the optimization metric is both time-consuming and labor-intensive. This work presents a machine learning (ML) approach to optimize PC conditions for fabricating methylammonium lead iodide (MAPbI 3 ) films by quantitatively comparing their ultraviolet-visible (UV-vis) absorbance spectra to thermal annealed (TA) films using four similarity metrics. We perform Bayesian optimization coupled with Gaussian process regression (BO-GP) to minimize the similarity metrics. Refining PC conditions using active learning based on BO-GP models, we achieve a PC MAPbI3 film with an absorbance spectrum closely matching a TA reference film, which is further verified by its crystalline and morphological properties. Thus, we demonstrate that the UV-vis absorption spectrum can accurately proxy film quality. Additionally, we use an AI-based segmentation model for a more efficient grain size analysis. However, when we use the optimized PC condition to fabricate PSCs, we find that interaction between MAPbI 3 and the hole transport layer (HTL) during PC critically degrades the PSC performance. By adding a buffer layer between the HTL and MAPbI 3 , the optimized PC PSCs produce a champion PCE of 11.8%, comparable to the TA reference of 11.7%. Using UV-vis similarity metrics instead of device PCE as the objective in our BO-GP method accelerates the optimization of PC processing conditions for MAPbI 3 films.

14 SOLAR ENERGY↗