Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Gaussian processes regression”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Machine learning with bond information for local structure optimizations in surface science

Local optimization of adsorption systems inherently involves different scales: within the substrate, within the molecule, and between the molecule and the substrate. In this work, we show how the explicit modeling of different characteristics of the bonds in these systems improves the performance of machine learning methods for optimization. Furthermore, we introduce an anisotropic kernel in the Gaussian process regression framework that guides the search for the local minimum, and we show its overall good performance across different types of atomic systems. The method shows a speed-up of up to a factor of two compared with the fastest standard optimization methods on adsorption systems. Additionally, we show that a limited memory approach is not only beneficial in terms of overall computational resources but can also result in a further reduction of energy and force calculations.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Optimization of artificial viscosity in production codes based on Gaussian Regression surrogate models

To accurately model flows with shock waves using staggered-grid Lagrangian hydrodynamics, artificial viscosity has to be introduced to convert kinetic energy into internal energy, thereby increasing the entropy across shocks. Determining the appropriate strength of the artificial viscosity is an art and strongly depends on the particular problem and experience of the researcher. The objective of this study is to pose the problem of finding the appropriate strength of artificial viscosity as an optimization problem and solve this problem using machine learning (ML) tools, specifically using surrogate models based on Gaussian Process regression and Bayesian analysis. We describe the optimization method and discuss various practical details of its implementation. The shock-containing problems for which we apply this method all have been implemented in the LANL code FLAG. First, we apply ML to find optimal values to isolated shock problems of different strengths. Second, we apply ML to optimize viscosity for a 1D propagating detonation problem based on Zel’dovich-von Neumann-Doring (ZND) detonation theory using a reactive burn model. We compare results for default (currently used values in FLAG) and optimized values of artificial viscosity for these problems demonstrating the potential for significant improvement in the accuracy of computations.

42 ENGINEERING↗

Gaussian Process for Flight Delay Prediction: Learning a Stochastic Process

This paper presents a machine-learning approach to predict flight delays. Whereas neural networks are extensively studied for predictive capabilities, they involve non-intuitive design and extensive analysis, particularly in training and optimization processes. Instead, the proposed framework employs Gaussian Processes as a supervised learning technique for flight delay prediction. This data-driven approach trains the model using prior information, specifically the mean and covariance tied to existing data. The proposed Gaussian Process Regression (GPR) model employs the day of flight as a pivotal feature for delay forecasting. We analyze flights from various routes and gauge the accuracy of the presented learning technique by comparing the predicted delays with the actual ones. Given the inherent challenges in precisely forecasting delays, we predict the delays with a 95 % confidence interval. Also, an error propagation analysis in the prediction horizon is carried out to determine the optimal time frame for prediction. The proposed method for flight delay prediction is important as airlines can strategize flight operations and issue timely advisories.

stochastic↗

Conditional Karhunen–Loève regression model with Basis Adaptation for high-dimensional problems: Uncertainty quantification and inverse modeling

Here, we propose a methodology for improving the accuracy of surrogate models of the observable response of physical systems as a function of the systems’ spatially heterogeneous parameter fields, with applications to uncertainty quantification and parameter estimation in high-dimensional problems. Practitioners often formulate finite-dimensional representations of spatially heterogeneous parameter fields using truncated unconditional Karhunen–Loève expansions (KLEs) for a certain choice of unconditional covariance kernel and construct surrogate models of the observable response with respect to the KLE coefficients. When direct measurements of the parameter fields are available, we propose improving the accuracy of these surrogate models by representing the parameter fields via conditional Karhunen-Loève expansions (CKLEs). CKLEs are constructed by conditioning the covariance kernel of the unconditional expansion on the direct measurements of the parameter field via Gaussian process regression, and then truncating the corresponding KLE. We apply the proposed methodology to constructing surrogate models via the Basis Adaptation (BA) method of the stationary hydraulic head response, measured at spatially discrete observation locations, of a groundwater flow model of the Hanford Site, as a function of the 1000-dimensional representation of the model’s log-transmissivity field. We find that BA surrogate models of the hydraulic head based on CKLEs are more accurate than BA surrogate models based on unconditional expansions for forward uncertainty quantification tasks. Furthermore, we find that inverse estimates of the hydraulic transmissivity field computed using CKLE-based BA surrogate models are more accurate than those computed using unconditional BA surrogate models.

97 MATHEMATICS AND COMPUTING↗

Uncertainty quantification in multivariable regression for material property prediction with Bayesian neural networks

With the increased use of data-driven approaches and machine learning-based methods in material science, the importance of reliable uncertainty quantification (UQ) of the predicted variables for informed decision-making cannot be overstated. UQ in material property prediction poses unique challenges, including multi-scale and multi-physics nature of materials, intricate interactions between numerous factors, limited availability of large curated datasets, etc. In this work, we introduce a physics-informed Bayesian Neural Networks (BNNs) approach for UQ, which integrates knowledge from governing laws in materials to guide the models toward physically consistent predictions. To evaluate the approach, we present case studies for predicting the creep rupture life of steel alloys. Experimental validation with three datasets of creep tests demonstrates that this method produces point predictions and uncertainty estimations that are competitive or exceed the performance of conventional UQ methods such as Gaussian Process Regression. Additionally, we evaluate the suitability of employing UQ in an active learning scenario and report competitive performance. The most promising framework for creep life prediction is BNNs based on Markov Chain Monte Carlo approximation of the posterior distribution of network parameters, as it provided more reliable results in comparison to BNNs based on variational inference approximation or related NNs with probabilistic outputs.

36 MATERIALS SCIENCE↗

Thinking Bayesian for plasma physicists

Bayesian statistics offers a powerful technique for plasma physicists to infer knowledge from the heterogeneous data types encountered. To explain this power, a simple example, Gaussian Process Regression, and the application of Bayesian statistics to inverse problems are explained. The likelihood is the key distribution because it contains the data model, or theoretic predictions, of the desired quantities. By using prior knowledge, the distribution of the inferred quantities of interest based on the data given can be inferred. Because it is a distribution of inferred quantities given the data and not a single prediction, uncertainty quantification is a natural consequence of Bayesian statistics. The benefits of machine learning in developing surrogate models for solving inverse problems are discussed, as well as progress in quantitatively understanding the errors that such a model introduces.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Time-series forecasting using manifold learning, radial basis function interpolation, and geometric harmonics

We address a three-tier numerical framework based on nonlinear manifold learning for the forecasting of high-dimensional time series, relaxing the “curse of dimensionality” related to the training phase of surrogate/machine learning models. At the first step, we embed the high-dimensional time series into a reduced low-dimensional space using nonlinear manifold learning (local linear embedding and parsimonious diffusion maps). Then, we construct reduced-order surrogate models on the manifold (here, for our illustrations, we used multivariate autoregressive and Gaussian process regression models) to forecast the embedded dynamics. Finally, we solve the pre-image problem, thus lifting the embedded time series back to the original high-dimensional space using radial basis function interpolation and geometric harmonics. The proposed numerical data-driven scheme can also be applied as a reduced-order model procedure for the numerical solution/propagation of the (transient) dynamics of partial differential equations (PDEs). In conclusion, we assess the performance of the proposed scheme via three different families of problems: (a) the forecasting of synthetic time series generated by three simplistic linear and weakly nonlinear stochastic models resembling electroencephalography signals, (b) the prediction/propagation of the solution profiles of a linear parabolic PDE and the Brusselator model (a set of two nonlinear parabolic PDEs), and (c) the forecasting of a real-world data set containing daily time series of ten key foreign exchange rates spanning the time period 3 September 2001–29 October 2020.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Taming nuclear mass models with Gaussian processes

We propose a new set of nuclear mass predictions based on multiple theoretical mass models. By employing Gaussian process regression with the Matérn kernel, we achieved root-mean-square (rms) deviations below 100 keV for the training dataset. The best-performing mass models achieved rms deviations below 150 keV for the new precise mass data from AME2020, whereas the ensemble average showed robust performance across the nuclear chart. Our approach uniquely combines: (1) systematic refinement of eight mass models through their residuals, (2) physics-informed features, including magic numbers, nucleon parity numbers, neutron excess, and nuclear collectivity, and (3) theory-to-theory validation demonstrating robust extrapolation capability. We find that the Matérn kernel provides superior uncertainty quantification compared to the RBF kernel, with a length-scale analysis revealing enhanced inter-nuclei correlations. We provide complete mass predictions for all unknown nuclides in AME2020, offering valuable constraints for nuclear structure studies and astrophysical modeling when used with proper uncertainty propagation.

Gaussian processes↗

Wavelength Dependence of Activity-induced Photometric Variations for Young Cool Stars in Hyades

We investigate photometric variations due to stellar activity that induce systematic radial-velocity errors (so-called “jitter”) for the four targets in the Hyades open cluster observed by the K2 mission (EPIC 210721261, EPIC 210923016, EPIC 247122957, and EPIC 247783757). Applying Gaussian process regressions to the K2 light curves and the near-infrared (NIR) light curves observed with the IRSF 1.4 m telescope, we derive the wavelength dependences of the photometric signals due to stellar activity. To estimate the temporal variations in the photometric variability amplitudes between the two observation periods of K2 and IRSF, separated by more than 2 yr, we analyze a number of K2 targets in Hyades that have also been observed in Campaigns 4 and 13 and find a representative variation rate over 2 yr of 38% ± 71%. Taking this temporal variation into account, we constrain projected sizes and temperature contrast properties of the starspots in the stellar photosphere to be approximately 10% and 0.95%, respectively. These starspot properties can induce relatively large differences in the variability amplitude over different observational passbands, and we find that radial-velocity jitter may be more suppressed in the NIR than previously expected. Our result supports profits of ongoing exoplanet search projects that are attempting to detect or confirm young planets in open clusters via radial-velocity measurements in the NIR.

47 OTHER INSTRUMENTATION↗

Optical Transmission Spectroscopy of the Terrestrial Exoplanet LHS 3844b from 13 Ground-based Transit Observations

Atmospheric studies of spectroscopically accessible terrestrial exoplanets lay the groundwork for comparative planetology between these worlds and the solar system terrestrial planets. LHS 3844b is a highly irradiated terrestrial exoplanet (R = 1.303 ± 0.022R {sub ⊕}) orbiting a mid-M dwarf 15 parsecs away. Work based on near-infrared Spitzer phase curves ruled out atmospheres with surface pressures ≥10 bars on this planet. We present 13 transit observations of LHS 3844b taken with the Magellan Clay telescope and the LDSS3C multi-object spectrograph covering 620–1020 nm. We analyze each of the 13 data sets individually using a Gaussian process regression, and present both white and spectroscopic light curves. In the combined white light curve we achieve an rms precision of 65 ppm when binning to 10 minutes. The mean white light-curve value of (R {sub p}/R {sub s}){sup 2} is 0.4170 ± 0.0046%. To construct the transmission spectrum, we split the white light curves into 20 spectrophotometric bands, each spanning 20 nm, and compute the mean values of (R {sub p}/R {sub s}){sup 2} in each band. We compare the transmission spectrum to two sets of atmospheric models. We disfavor a clear, solar composition atmosphere (μ = 2.34) with surface pressures ≥0.1 bar to 5.2σ confidence. We disfavor a clear, H{sub 2}O steam atmosphere (μ = 18) with surface pressures ≥0.1 bar to low confidence (2.9σ). Our observed transmission spectrum favors a flat line. For solar composition atmospheres with surface pressures ≥1 bar we rule out clouds with cloud-top pressures of 0.1 bar (5.3σ), but we cannot address high-altitude clouds at lower pressures. Our results add further evidence that LHS 3844b is devoid of an atmosphere.

79 ASTRONOMY AND ASTROPHYSICS↗

Predicting molecular dipole moments by combining atomic partial charges and atomic dipoles

The molecular dipole moment ( μ ) is a central quantity in chemistry. It is essential in predicting infrared and sum-frequency generation spectra as well as induction and long-range electrostatic interactions. Furthermore, it can be extracted directly—via the ground state electron density—from high-level quantum mechanical calculations, making it an ideal target for machine learning (ML). Here, we choose to represent this quantity with a physically inspired ML model that captures two distinct physical effects: local atomic polarization is captured within the symmetry-adapted Gaussian process regression framework which assigns a (vector) dipole moment to each atom, while the movement of charge across the entire molecule is captured by assigning a partial (scalar) charge to each atom. The resulting “MuML” models are fitted together to reproduce molecular μ computed using high-level coupled-cluster theory and density functional theory (DFT) on the QM7b dataset, achieving more accurate results due to the physics-based combination of these complementary terms. The combined model shows excellent transferability when applied to a showcase dataset of larger and more complex molecules, approaching the accuracy of DFT at a small fraction of the computational cost. We also demonstrate that the uncertainty in the predictions can be estimated reliably using a calibrated committee model. The ultimate performance of the models—and the optimal weighting of their combination—depends, however, on the details of the system at hand, with the scalar model being clearly superior when describing large molecules whose dipole is almost entirely generated by charge separation. These observations point to the importance of simultaneously accounting for the local and non-local effects that contribute to μ ; furthermore, they define a challenging task to benchmark future models, particularly those aimed at the description of condensed phases.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Hybrid Data‐Driven Discovery of High‐Performance Silver Selenide‐Based Thermoelectric Composites

Optimizing material compositions often enhances thermoelectric performances. However, the large selection of possible base elements and dopants results in a vast composition design space that is too large to systematically search using solely domain knowledge. To address this challenge, a hybrid data-driven strategy that integrates Bayesian optimization (BO) and Gaussian process regression (GPR) is proposed to optimize the composition of five elements (Ag, Se, S, Cu, and Te) in AgSe-based thermoelectric materials. Data is collected from the literature to provide prior knowledge for the initial GPR model, which is updated by actively collected experimental data during the iteration between BO and experiments. Within seven iterations, the optimized AgSe-based materials prepared using a simple high-throughput ink mixing and blade coating method deliver a high power factor of 2100 µW m −1 K −2 , which is a 75% improvement from the baseline composite (nominal composition of Ag 2 Se 1 ). In conclusion, the success of this study provides opportunities to generalize the demonstrated active machine learning technique to accelerate the development and optimization of a wide range of material systems with reduced experimental trials.

36 MATERIALS SCIENCE↗

A Comprehensive Comparative Study of Active Learning Schemes for Nanophotonics Design

We present a benchmarking study of active learning (AL) schemes for designing planar multilayer nanophotonic metamaterials, where the design tasks are formulated as binary optimization problems. Different surrogate models, including factorization machine (FM), Gaussian process regression (GPR), and convolutional neural network (CNN), combined with different optimization methods, including exhaustive enumeration, discrete particle swarm optimization (DPSO), quantum annealing (QA), hybrid QA, and simulated annealing are studied. The benchmark cases investigated range from small problems with short binary lengths (N = 25) to large problems with N up to 100, focusing on the design of two classes of photonic structures, including antireflective coatings for the long-wavelength infrared region and transparent radiative coolers. For small problems, CNN coupled with DPSO in AL achieves the best performance. As N increases, FM with QA outperforms GPR and CNN. For FM-based AL, hybrid QA yields the best optimization results, particularly in high-dimensional cases (N = 100). These results demonstrate that the optimization method can significantly affect in AL performance as N increases, and that QA-based optimization can provide practical routes for mitigating the optimization bottleneck in high-dimensional problems.

Jung, Serang [Kyung Hee University, Korea]↗

Data‐Driven Engineering of Thermostable Collagen‐Mimetic Peptoid Triple Helices

Collagen-mimetic peptides (CMPs) are engineered molecules designed to replicate the triple-helical structure of natural collagen. A repeating x–y-Gly sequence is the defining motif of CMPs and is critical to their triple-helical structure and stability. Substitutions to the residues occupying the x and y positions present a means to modulate the CMP structure and properties. Peptoid residues—N-substituted glycine derivatives—present an attractive potential substitution due to their thermal stability, proteolytic resistance, biocompatibility, and diverse palette of non-natural side chains, but also tend to introduce a high degree of backbone flexibility that can diminish the stability of the triple helix. In this work, we report a computational active learning cycle comprising molecular dynamics simulation, Gaussian process regression, and Bayesian optimization to computationally identify a number of promising peptoid substitutions predicted to stabilize the desired quaternary structure through side chain interactions and produce stable peptoid-based collagen-like triple helices. To experimentally test the computational predictions, a top candidate identified by the screen was synthesized and imaged using scanning electron microscopy to resolve fibril-like bundles consistent with collagen-like triple helices. This work predicts a number of CMP peptoid substitutions capable of forming stable triple-helical structures, presents a generalizable design strategy for engineering desired peptoid structures, and opens new avenues for the design of peptoid-based biomimetic materials.

active learning↗

Autonomous Synthesis of Thin Film Materials with Pulsed Laser Deposition Enabled by In Situ Spectroscopy and Automation

Autonomous systems that combine synthesis, characterization, and artificial intelligence can greatly accelerate the discovery and optimization of materials, however platforms for growth of macroscale thin films by physical vapor deposition techniques have lagged far behind others. Here this study demonstrates autonomous synthesis by pulsed laser deposition (PLD), a highly versatile synthesis technique, in the growth of ultrathin WSe 2 films. Further, by combing the automation of PLD synthesis and in situ diagnostic feedback with a high-throughput methodology, this study demonstrates a workflow and platform which uses Gaussian process regression and Bayesian optimization to autonomously identify growth regimes for WSe 2 films based on Raman spectral criteria by efficiently sampling 0.25% of the chosen 4D parameter space. With throughputs at least 10x faster than traditional PLD workflows, this platform and workflow enables the accelerated discovery and autonomous optimization of the vast number of materials that can be synthesized by PLD.

36 MATERIALS SCIENCE↗

Search for the associated production of charm quarks and a Higgs boson decaying into a photon pair with the ATLAS detector

A search for the production of a Higgs boson and one or more charm quarks, in which the Higgs boson decays into a photon pair, is presented. This search uses proton-proton collision data with a centre-of-mass energy of $\sqrt{s}$ = 13 TeV and an integrated luminosity of 140 fb −1 recorded by the ATLAS detector at the Large Hadron Collider. The analysis relies on the identification of charm-quark-containing jets, and adopts an approach based on Gaussian process regression to model the non-resonant di-photon background. The observed (expected, assuming the Standard Model signal) upper limit at the 95% confidence level on the cross-section for producing a Higgs boson and at least one charm-quark-containing jet that passes a fiducial selection is found to be 10.6 pb (8.8 pb). The observed (expected) measured cross-section for this process is 5.3 ± 3.2 pb (2.9 ± 3.1 pb).

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Search for dimuon resonance in the 35 to 75 GeV mass range using 140 fb−1 of 13 TeV pp collisions with the ATLAS detector

A model-independent search for low-mass resonances decaying into pairs of oppositely charged muons is presented. The analysis uses proton-proton collision data corresponding to an integrated luminosity of 140 fb−1, recorded by the ATLAS detector at the Large Hadron Collider between 2015 and 2018. The search targets hypothetical dimuon resonances in the invariant mass range from 35 GeV to 75 GeV. The modelling of this mass region is particularly challenging for conventional analytic background parameterisations. To address this, a Gaussian process regression technique is used to model the background. The dimuon mass spectrum is analysed for potential signals, and no statistically significant excess is observed. Upper limits at the 95% confidence level are set on the fiducial production cross-section of new resonances decaying promptly into muons, ranging from 20 fb to 110 fb, depending on the resonance mass. These results are further interpreted in the context of dark-photon and dark-matter-mediator models, leading to new constraints on their parameter spaces.

Aad, G↗

Data-Driven Surrogate Modeling with Microstructure-Sensitivity of Viscoplastic Creep in Grade 91 Steel

Abstract To support the development of advanced steel alloys tailored to withstand extreme conditions, it is imperative to account for the mechanical performance of components, while considering the influence of local microstructure on the macroscopic response. To this end, this study focuses on the development of microstructure-sensitive constitutive models for the mechanical response of Grade 91 steel exposed to extreme thermo-mechanical environments. Polynomial chaos expansion (PCE) surrogates are used to emulate high-fidelity polycrystal simulations of the viscoplastic response of Grade 91 steel as a function of the microstructure fingerprint (e.g., dislocations and precipitates). To cover a wide temperature–stress domain, two separate PCE surrogates—one that captures softening and the other that captures hardening behavior—are combined using another (sparse) Gaussian process regression model. The resulting constitutive creep surrogate model is integrated within the MOOSE finite element framework to simulate the intricate effects of microstructure, in particular MX-phase precipitates, on a component with a graded microstructure. Surrogate sensitivity analysis is applied to quantify the relevant impact of spatially varying microstructure on the creep response in a test-case involving a Grade 91 alloy with a prototypical weld.

36 MATERIALS SCIENCE↗