Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Bayesian parameter estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Development of an open-source regional data assimilation system in PEcAn v. 1.7.2: application to carbon cycle reanalysis across the contiguous US using SIPNET

Abstract. The ability to monitor, understand, and predict the dynamics of the terrestrial carbon cycle requires the capacity to robustly and coherently synthesize multiple streams of information that each provide partial information about different pools and fluxes. In this study, we introduce a new terrestrial carbon cycle data assimilation system, built on the PEcAn model–data eco-informatics system, and its application for the development of a proof-of-concept carbon “reanalysis” product that harmonizes carbon pools (leaf, wood, soil) and fluxes (GPP, Ra, Rh, NEE) across the contiguous United States from 1986–2019. We first calibrated this system against plant trait and flux tower net ecosystem exchange (NEE) using a novel emulated hierarchical Bayesian approach. Next, we extended the Tobit–Wishart ensemble filter (TWEnF) state data assimilation (SDA) framework, a generalization of the common ensemble Kalman filter which accounts for censored data and provides a fully Bayesian estimate of model process error, to a regional-scale system with a calibrated localization. Combined with additional workflows for propagating parameter, initial condition, and driver uncertainty, this represents the most complete and robust uncertainty accounting available for terrestrial carbon models. Our initial reanalysis was run on an irregular grid of ∼ 500 points selected using a stratified sampling method to efficiently capture environmental heterogeneity. Remotely sensed observations of aboveground biomass (Landsat LandTrendr) and leaf area index (LAI) (MODIS MOD15) were sequentially assimilated into the SIPNET model. Reanalysis soil carbon, which was indirectly constrained based on modeled covariances, showed general agreement with SoilGrids, an independent soil carbon data product. Reanalysis NEE, which was constrained based on posterior ensemble weights, also showed good agreement with eddy flux tower NEE and reduced root mean square error (RMSE) compared to the calibrated forecast. Ultimately, PEcAn's new open-source regional data assimilation framework provides a scalable workflow for harmonizing multiple data constraints and providing a uniform synthetic platform for carbon monitoring, reporting, and verification (MRV) as well as accelerating terrestrial carbon cycle research.

54 ENVIRONMENTAL SCIENCES↗

Detecting outbreaks using a spatial latent field

In this paper, we present a method for estimating the infection-rate of a disease as a spatial-temporal field. Our data comprises time-series case-counts of symptomatic patients in various areal units of a region. We extend an epidemiological model, originally designed for a single areal unit, to accommodate multiple units. The field estimation is framed within a Bayesian context, utilizing a parameterized Gaussian random field as a spatial prior. We apply an adaptive Markov chain Monte Carlo method to sample the posterior distribution of the model parameters condition on COVID-19 case-count data from three adjacent counties in New Mexico, USA. Our results suggest that the correlation between epidemiological dynamics in neighboring regions helps regularize estimations in areas with high variance (i.e., poor quality) data. Using the calibrated epidemic model, we forecast the infection-rate over each areal unit and develop a simple anomaly detector to signal new epidemic waves. Our findings show that anomaly detector based on estimated infection-rates outperforms a conventional algorithm that relies solely on case-counts.

Safta, Cosmin [Sandia National Laboratories (SNL-C↗

Cosmological parameter estimation with a joint-likelihood analysis of the cosmic microwave background and big bang nucleosynthesis

Here, we present a joint-likelihood analysis of big bang nucleosynthesis (BBN) and cosmic microwave background (CMB) data, consistently combining likelihoods and taking into account uncertainties in nuclear reaction rates for the first time. Bayesian inference is performed on the baryon abundance and the effective number of neutrino species, 𝑁 eff , using a CMB Boltzmann solver in combination with LINX , a new flexible and efficient BBN code. We marginalize over Planck nuisance parameters and nuclear rates to find 𝑁 eff =3.0⁢8$^{+0.15}_{−0.14}$, 2.9⁢4$^{+0.16}_{−0.15}$, or 2.96$^{+0.13}_{−0.14}$, for three separate reaction networks. This framework enables robust testing of the lambda cold dark matter paradigm and its variants with CMB and BBN data.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Joint Bayesian Inference for Near-Surface Explosion Yield and Height-of-Burst

Forensic capabilities to understand chemical and nuclear explosions are greatly aided by an accurate estimate of explosive yield with uncertainty. The relationship between explosive size and geophysical observations of seismic, acoustic, and optical waves can be exploited to provide an estimate of yield. Any near-surface yield estimate is complicated by the surface interaction, so an estimate for the explosion height-of-burst is necessarily included in the relationship. Additionally, the relationship dictates a trade-off between estimates of yield and height-of-burst. Fortunately, the surface interaction for each type of observation is different, which breaks the trade-off, and the inclusion of height-of-burst with multiple data types improves yield estimation. We define simple parametric forward models to relate seismoäcoustoöptic observations from a data set of known explosive yields and height-of-bursts. The parameters of the models and a prediction for the yield and height-of-burst of a new event can then be estimated given new observations via Bayesian inference. We report posterior distribution estimates of the parametric models using a Markov chain Monte Carlo sampling technique. These models are then used to predict the yield and height-of-burst of SUGAR, a historical near-surface nuclear explosion, using its reported historical observations. The reported yield of 1.2 ktonne Trinitrotoluene (TNT)-equivalent (Department of Energy, 2015) is within the estimated posterior. Yield uncertainty can be estimated from the spread of the posterior, which is between 0.9 and 2.1 ktonne TNT-equivalent. The posterior for height-of-burst has a wider range between 10 m below and 8 m above ground that includes the true height-of-burst of 1 m.

58 GEOSCIENCES↗

Mechanistic within-host mathematical model of inhalational anthrax

We present a mathematical model of the dynamics of Bacillus anthracis bacteria within the lymph nodes and blood of a host, following inhalation of an initial dose of spores. We also incorporate the dynamics of protective antigen, which is the binding component of the anthrax toxin produced by the bacteria. The model offers a mechanistic description of the early infection dynamics of inhalational anthrax, while its stochastic nature allows us to study the probabilities of different outcomes (for example, how likely it is that the infection will be cleared for a given inhaled dose of spores) in order to explain dose-response data for inhalational anthrax. The model is calibrated via a Bayesian approach, using in vivo data from New Zealand white rabbit and guinea pig infection studies, enabling within-host parameters to be estimated. We also leverage incubation-period data from the Sverdlovsk 1979 anthrax outbreak to show that the model can accurately describe human time-to-symptoms data under reasonable parameter regimes. Finally, we derive a simple approximate formula for the probability of symptom onset before time t, assuming that the number of inhaled spores has a Poisson distribution.

59 BASIC BIOLOGICAL SCIENCES↗

Estimation of stagnation performance metrics in magnetized liner inertial fusion experiments using Bayesian data assimilation

Here we present a new analysis methodology that allows for the self-consistent integration of multiple diagnostics including nuclear measurements, x-ray imaging, and x-ray power detectors to determine the primary stagnation parameters, such as temperature, pressure, stagnation volume, and mix fraction in magnetized liner inertial fusion (MagLIF) experiments. The analysis uses a simplified model of the stagnation plasma in conjunction with a Bayesian inference framework to determine the most probable configuration that describes the experimental observations while simultaneously revealing the principal uncertainties in the analysis. We validate the approach by using a range of tests including analytic and three-dimensional MHD models. An ensemble of MagLIF experiments is analyzed, and the generalized Lawson criterion χ is estimated for all experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Performing Bayesian Analyses With AZURE2 Using BRICK: An Application to the 7 Be System

Phenomenological R-matrix has been a standard framework for the evaluation of resolved resonance cross section data in nuclear physics for many years. It is a powerful method for comparing different types of experimental nuclear data and combining the results of many different experimental measurements in order to gain a better estimation of the true underlying cross sections. Yet a practical challenge has always been the estimation of the uncertainty on both the cross sections at the energies of interest and the fit parameters, which can take the form of standard level parameters. Frequentist (χ 2 -based) estimation has been the norm. In this work, a Markov Chain Monte Carlo sampler, emcee, has been implemented for the R-matrix code AZURE2, creating the Bayesian R-matrix Inference Code Kit (BRICK). Bayesian uncertainty estimation has then been carried out for a simultaneous R-matrix fit of the 3 He (α,γ) 7 Be and 3 He (α,α) 3 He reactions in order to gain further insight into the fitting of capture and scattering data. Both data sets constrain the values of the bound state α-particle asymptotic normalization coefficients in 7 Be. The analysis highlights the need for low-energy scattering data with well-documented uncertainty information and shows how misleading results can be obtained in its absence.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Robust Importance Sampling for Bayesian Model Calibration with Spatio-Temporal Data

This paper addresses two challenges in Bayesian calibration: 1) computational speed of existing sampling algorithms, and 2) calibration with spatio-temporal responses. The commonly used Markov Chain Monte Carlo (MCMC) approaches require many sequential model evaluations making the computational expense prohibitive. This paper proposes an efficient sampling algorithm: iterative importance sampling with genetic algorithm (IISGA). While iterative importance sampling enables computational efficiency, the genetic algorithm enables robustness by preventing sample degeneration and avoids getting stuck in multimodal search spaces. An inflated likelihood further enables robustness in high-dimensional parameter spaces by enlarging the target distribution. Spatio-temporal data complicate both surrogate modeling, which is necessary for expensive computational models, and the likelihood estimation. In this work, singular value decomposition is investigated for reducing the high-dimensional field data to a lower-dimensional space prior to Bayesian calibration. Then the likelihood is formulated and Bayesian inference is performed in the lower-dimension, latent space. An illustrative example is provided to demonstrate IISGA relative to existing sampling methods, and then IISGA is employed to calibrate a thermal battery model with 26 uncertain calibration parameters and spatio-temporal response data.

97 MATHEMATICS AND COMPUTING↗

Constraining Bedrock Groundwater Residence Times in a Mountain System With Environmental Tracer Observations and Bayesian Uncertainty Quantification

Groundwater residence time distributions provide fundamental insights on the hydrological processes within watersheds. Yet, observations that can constrain groundwater residence times over broad timescales remain scarce in mountain catchment studies. We use environmental tracers (CFC-12, SF 6 , 3 H, and 4 He) to investigate groundwater residence times along a hillslope in the East River Watershed, Colorado, USA. We develop a Bayesian inference framework that applies a Markov-chain Monte Carlo (MCMC) approach to estimate noble gas recharge temperature, elevation, and excess-air parameters and the resulting environmental tracer concentrations. MCMC is then used to propagate the environmental tracer uncertainties to estimates of groundwater mean residence times inferred with lumped parameter models. All samples contain 3 H, CFC-12, and SF 6 in addition to terrigenic 4 He, suggesting a mixture of water characterized by modern and premodern residence times. 4He exponential mean residence times range from hundreds of years at the upslope well to thousands of years at the toe-slope well assuming average crustal production rates. We find that binary mixing residence time distributions with separate young and old mixing fractions are needed to predict the 4 He, CFC-12, SF 6 , and 3 H observations, supporting the importance of flow path mixing in this bedrock system. Our findings that the fractured bedrock hosts groundwater with a mixture of residence times ranging from decades to millennia suggest variable recharge dynamics and flow path mixing along the hillslope and highlight the importance of characterizing groundwater systems with observations that are sensitive to transport over a broad range of residence times.

54 ENVIRONMENTAL SCIENCES↗

A Bayesian Approach for Quantifying Data Scarcity when Modeling Human Behavior via Inverse Reinforcement Learning

Computational models that formalize complex human behaviors enable study and understanding of such behaviors. However, collecting behavior data required to estimate the parameters of such models is often tedious and resource intensive. Thus, estimating dataset size as part of data collection planning (also known as Sample Size Determination) is important to reduce the time and effort of behavior data collection while maintaining an accurate estimate of model parameters. In this paper, we present a sample size determination method based on Uncertainty Quantification (UQ) for a specific Inverse Reinforcement Learning (IRL) model of human behavior, in two cases: 1) pre-hoc experiment design—conducted in the planning stage before any data is collected, to guide the estimation of how many samples to collect; and 2) post-hoc dataset analysis—performed after data is collected, to decide if the existing dataset has sufficient samples and whether more data is needed. Here, we validate our approach in experiments with a realistic model of behaviors of people with Multiple Sclerosis (MS) and illustrate how to pick a reasonable sample size target. Our work enables model designers to perform a deeper, principled investigation of effects of dataset size on IRL.

97 MATHEMATICS AND COMPUTING↗

Optimal Power Management for Large-Scale Battery Energy Storage Systems via Bayesian Inference

Large-scale battery energy storage systems (BESS) have found ever-increasing use across industry and society to accelerate clean energy transition and improve energy supply reliability and resilience. However, their optimal power management poses significant challenges: the underlying high-dimensional nonlinear nonconvex optimization lacks computational tractability in real-world implementation, and the uncertainty of the exogenous power demand makes exact optimization difficult. This paper presents a new solution framework to address these bottlenecks. The solution pivots on introducing power-sharing ratios to specify each cell’s power quota from the output power demand. To find the optimal power-sharing ratios, we formulate a nonlinear model predictive control (NMPC) problem to achieve power-loss-minimizing BESS operation while complying with safety, cell balancing, and power supply-demand constraints. We then propose a parameterized control policy for the power-sharing ratios, which utilizes only three parameters, to reduce the computational demand in solving the NMPC problem. This policy parameterization allows us to translate the NMPC problem into a Bayesian inference problem for the sake of 1) computational tractability, and 2) overcoming the nonconvexity of the optimization problem. We leverage the ensemble Kalman inversion technique to solve the parameter estimation problem. Concurrently, a low-level control loop is developed to seamlessly integrate our proposed approach with the BESS to ensure practical implementation. This low-level controller receives the optimal power-sharing ratios, generates output power references for the cells, and maintains a balance between power supply and demand despite uncertainty in output power. We conduct extensive simulations and experiments on a 20-cell prototype to validate the proposed approach.

Battery energy storage systems (BESSs)↗

Pervaporative Dehydration of 2,3-Butanediol by Dense Poly(vinylidene fluoride) Hollow Fiber Membranes: Parameter Estimation, Process Design, and Technoeconomic Evaluation under Uncertainty

Pervaporation, combined with other separation processes, can effectively remove water from fermentation product streams, making it highly suitable for purifying alcohols like 2,3-butanediol (BDO). In this study, a dense poly(vinylidene fluoride) (PVDF) hollow fiber membrane module prototype was fabricated for BDO dehydration, achieving >0.2 LMH total flux and >95% BDO rejection. With a Markov chain Monte Carlo (MCMC) approach, Bayesian inference was used to quantify the uncertainty of the permeance parameters. A membrane cascade model was developed to scale up a process that purifies a preconcentrated BDO feed (70 wt %) to high purity (90 wt %). Through propagation of the uncertainty of the parameters and sensitivity analyses of the process variables, a cascade design was recommended. Despite data and model limitations, the framework enabled a reliable system analysis and economic evaluation, validated through tight confidence intervals in key process metrics, establishing the foundation for future applications of Bayesian methods in membrane-based processes.

Animal feed↗

Uncertainty quantification of mass models using ensemble Bayesian model averaging

Developments in the description of the masses of atomic nuclei have led to various nuclear mass models that provide predictions for masses across the whole chart of nuclides. These mass models play an important role in understanding the synthesis of heavy elements in the rapid neutron capture ( r ) process. However, it is still a challenging task to estimate the size of uncertainty associated with the predictions of each mass model. In this work, a method called ensemble Bayesian model averaging (EBMA) is introduced to quantify the uncertainty of one-neutron separation energies (S 1 n ) which are directly relevant in the calculations of r -process observables. Here, this Bayesian method provides a natural way to perform model averaging, selection, and uncertainty quantification, by combining the mass models as a mixture of normal distributions whose parameters are optimized against the experimental data, employing the Markov chain Monte Carlo method using the no-u-turn sampler. The EBMA model optimized with all the experimental S 1 n from the AME2003 nuclides are shown to provide reliable uncertainty estimates when tested with the new data in the AME2020.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

BayesMT: A Probabilistic Bayesian Framework for the Seismic Moment Tensor

Moment tensors (MTs) have long been used in earthquake and explosion source analysis, and there has been a renewed interest in how they can inform us about the seismic source, particularly in the geophysical monitoring community due to its application in event identification and yield analysis. However, parameter uncertainties in seismic MT inversion are rarely available. The inverse procedure often does not quantify MT model errors such as event location, data noise and Earth model that are essential for estimating solution robustness. To address this need, we propose to adopt the Bayesian probabilistic framework to incorporate uncertainties in MT inversions. In this study, we present the theoretical background of a probabilistic Bayesian framework for MT inversion accounting for model and measurements errors and illustrate the implementation of the method using a synthetic example.

58 GEOSCIENCES↗

Strong Lensing Parameter Estimation on Ground-Based Imaging Data Using Simulation-Based Inference

Current ground-based cosmological surveys, such as the Dark Energy Survey (DES), are predicted to discover thousands of galaxy-scale strong lenses, while future surveys, such as the Vera Rubin Observatory Legacy Survey of Space and Time (LSST) will increase that number by 1-2 orders of magnitude. The large number of strong lenses discoverable in future surveys will make strong lensing a highly competitive and complementary cosmic probe. To leverage the increased statistical power of the lenses that will be discovered through upcoming surveys, automated lens analysis techniques are necessary. We present two Simulation-Based Inference (SBI) approaches for lens parameter estimation of galaxy-galaxy lenses. We demonstrate the successful application of Neural Posterior Estimation (NPE) to automate the inference of a 12-parameter lens mass model for DES-like ground-based imaging data. We compare our NPE constraints to a Bayesian Neural Network (BNN) and find that it outperforms the BNN, producing posterior distributions that are for the most part both more accurate and more precise; in particular, several source-light model parameters are systematically biased in the BNN implementation.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

MFNets: Multifidelity data-driven networks for Bayesian learning and prediction

This paper presents a multifidelity uncertainty quantification framework called MFNets. We seek to address three existing challenges that arise when experimental and simulation data from different sources are used to enhance statistical estimation and prediction with quantified uncertainty. Specifically, we demonstrate that MFNets can (1) fuse heterogeneous data sources arising from simulations with different parameterizations, e.g simulation models with different uncertain parameters or data sets collected under different environmental conditions; (2) encode known relationships among data sources to reduce data requirements; and (3) improve the robustness of existing multi-fidelity approaches to corrupted data. MFNets construct a network of latent variables (LVs) to facilitate the fusion of data from an ensemble of sources of varying credibility and cost. These LVs are posited as explanatory variables that provide the source of correlation in the observed data. Furthermore, MFNets provide a way to encode prior physical knowledge to enable efficient estimation of statistics and/or construction of surrogates via conditional independence relations on the LVs. We highlight the utility of our framework with a number of theoretical results which assess the quality of the posterior mean as a frequentist estimator and compare it to standard sampling approaches that use single fidelity, multilevel, and control variate Monte Carlo estimators. We also use the proposed framework to derive the Monte Carlo-based control variate estimator entirely from the use of Bayes rule and linear-Gaussian models -- to our knowledge the first such derivation. Finally, we demonstrate the ability to work with different uncertain parameters across different models.

97 MATHEMATICS AND COMPUTING↗

Functional Data Analysis for Extracting the Intrinsic Dimensionality of Spectra: Application to Chemical Homogeneity in the Open Cluster M67

High-resolution spectroscopic surveys of the Milky Way have entered the Big Data regime and have opened avenues for solving outstanding questions in Galactic archeology. However, exploiting their full potential is limited by complex systematics, whose characterization has not received much attention in modern spectroscopic analyses. In this work, we present a novel method to disentangle the component of spectral data space intrinsic to the stars from that due to systematics. Using functional principal component analysis on a sample of 18,933 giant spectra from APOGEE, we find that the intrinsic structure above the level of observational uncertainties requires ≈10 functional principal components (FPCs). Our FPCs can reduce the dimensionality of spectra, remove systematics, and impute masked wavelengths, thereby enabling accurate studies of stellar populations. To demonstrate the applicability of our FPCs, we use them to infer stellar parameters and abundances of 28 giants in the open cluster M67. We employ Sequential Neural Likelihood, a simulation-based Bayesian inference method that learns likelihood functions using neural density estimators, to incorporate non-Gaussian effects in spectral likelihoods. By hierarchically combining the inferred abundances, we limit the spread of the following elements in M67: Fe ≲ 0.02 dex; C ≲ 0.03 dex; O, Mg, Si, Ni ≲ 0.04 dex; Ca ≲ 0.05 dex; N, Al ≲ 0.07 dex (at 68% confidence). Our constraints suggest a lack of self-pollution by core-collapse supernovae in M67, which has promising implications for the future of chemical tagging to understand the star formation history and dynamical evolution of the Milky Way.

79 ASTRONOMY AND ASTROPHYSICS↗

Developing a Continuous Nuclear Data Evaluation Method for the Resolved and Unresolved Resonance Regions With FITAPI and SAMMY [Poster]

SAMMY is an R-matrix Bayesian fitting program used in the analysis of nuclear cross section data. SAMMY produces separate cross section models in the resolved and unresolved resonance regions. This project develops a new method of continuous evaluation joining both energy regions with FITAPI. Full energy spectrum model introduces cross covariance between physics models allowing more accurate description of uncertainty. Consistency developed between physics models by the use of a unique set of prior parameters. An unresolved region cross section developed with sensitivity to resolved cross section. Incorporation into ORNL FITAPI provides access to greater number of parameter estimation algorithms.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗