Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Bayesian parameter estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

An error covariance model for sea surface topography and velocity derived from TOPEX/POSEIDON altimetry

In order to facilitate the use of satellite-derived sea surface topography and velocity oceanographic models, methodology is presented for deriving the total error covariance and its geographic distribution from TOPEX/POSEIDON measurements. The model is formulated using a parametric model fit to the altimeter range observations. The topography and velocity modeled with spherical harmonic expansions whose coefficients are found through optimal adjustment to the altimeter range residuals using Bayesian statistics. All other parameters, including the orbit, geoid, surface models, and range corrections are provided as unadjusted parameters. The maximum likelihood estimates and errors are derived from the probability density function of the altimeter range residuals conditioned with a priori information. Estimates of model errors for the unadjusted parameters are obtained from the TOPEX/POSEIDON postlaunch verification results and the error covariances for the orbit and the geoid, except for the ocean tides. The error in the ocean tides is modeled, first, as the difference between two global tide models and, second, as the correction to the present tide model, the correction derived from the TOPEX/POSEIDON data. A formal error covariance propagation scheme is used to derive the total error. Our global total error estimate for the TOPEX/POSEIDON topography relative to the geoid for one 10-day period is found tio be 11 cm RMS. When the error in the geoid is removed, thereby providing an estimate of the time dependent error, the uncertainty in the topography is 3.5 cm root mean square (RMS). This level of accuracy is consistent with direct comparisons of TOPEX/POSEIDON altimeter heights with tide gauge measurements at 28 stations. In addition, the error correlation length scales are derived globally in both east-west and north-south directions, which should prove useful for data assimilation. The largest error correlation length scales are found in the tropics. Errors in the velocity field are smallest in midlatitude regions. For both variables the largest errors caused by uncertainty in the geoid. More accurate representations of the geoid await a dedicated geopotential satellite mission. Substantial improvements in the accuracy of ocean tide models are expected in the very near future from research with TOPEX/POSEIDON data.

Tsaoussi, Lucia S.↗

Semi-Analytical Hierarchical Bayesian Inference of Nonlinear Model Structure in Stochastic Dynamics: Applied to Compartmental Models of Infectious Diseases

A Bayesian computational framework for parsimonious inference in stochastic nonlinear dynamical systems is presented. This framework enables the concurrent estimation of system states, time-varying parameters, time-invariant parameters, and the optimal sparsity structure of the model parameters. Because differential equation-based models are often simplified mechanistic or phenomenological representations, robust inference from noisy measurement data requires explicit treatment of model error and uncertainty. Model error and time-varying parameters can be represented as random processes, enabling inference while making minimal assumptions about the underlying sources of discrepancy and variability. Adopting stochastic differential equation representations affords the model significant flexibility, but can also render it susceptible to overfitting during statistical inversion, where the inferred model may track noise rather than the underlying signal. To alleviate the effects of overfitting and to enable the discovery of the optimal sparse representation of the time-invariant parameters, a Bayesian sparse learning algorithm is embedded within the framework. This sparse learning framework adopts an approximate hierarchical Bayesian setting defined by a series of semi-analytical expressions. The model structure inference framework is validated using a stochastic compartmental model for tracking and forecasting active cases of an infectious disease. Compartmental models describe population-level infectious disease dynamics through interactions among population fractions grouped by disease state. Mathematically, such models consist of a system of coupled ordinary differential equations. This example adopts an expressive compartmental model that includes multiple possible interactions between disease states, motivated by early uncertainty surrounding COVID-19 reinfection dynamics and their implications for long-term epidemic forecasting. The sparse learning exercise permits the inference of a priori unknown epidemiological dynamics from simulated public health data, discovering the nested compartmental model that optimizes the trade-off between average data-fit and model complexity. It is shown that inducing sparsity among the model parameters eliminates redundant interactions between compartments, equivalently revealing the optimal coupling structure between differential equations.

97 MATHEMATICS AND COMPUTING↗

Multi-fidelity thermal modeling of laser powder bed additive manufacturing

Laser powder bed fusion (LPBF) Additive manufacturing (AM) has attracted interest as an agile method of building production metal parts to reduce design-build-test cycle times for systems. However, predicting part performance is difficult due to inherent process variabilities. This makes qualification challenging. Computational process models have attempted to address some of these challenges, including mesoscale, full physics models and reduced fidelity conduction models. The goal of this work is credible multi-fidelity modeling of the LPBF process by investigating methods for estimating the error between models of two different fidelities. Two methods of error estimation are investigated, adjoint-based error estimation and Bayesian calibration. Adjoint-based error estimation is found to effectively bounding the error between the two models, but with very conservative bounds, making predictions highly uncertain. Bayesian parameter calibration applied to conduction model heat source parameters is found to effectively bound the observed error between the models for melt pool morphology quantities of interest. However, the calibrations do not effectively bound the error in heat distribution.

36 MATERIALS SCIENCE↗

Uncertainty quantification for equations of state: copper as an example

Equations of state are essential for providing a fundamental description of materials properties in thermodynamic equilibrium and are used to provide closure relations for hydrodynamics simulations. Generally, equations of state rely on simple physics-based parameterized materials models to inform on the free energy of a material through out a given thermodynamic state space. Historically the parameters of these models have been tuned by hand to fit various experimental data. However, modern optimization and uncertainty quantification techniques allow us to quickly test thousands of parameter combinations and obtain meaningful uncertainty estimates on the parameters, opening opportunities for assessing systematic uncertainties in experiments, assessing model adequacy, and more. In this report, we use Bayesian inference to fit the solid (fcc) equation of state of copper. We focus on fitting five different experimental datasets, including the isobaric density, isobaric heat capacity, room temperature isotherm, principal isentrope, and principal Hugoniot. We fit all five data types simultaneously, and then explore the extent to which combinations of 2 subsets of the 5 datasets can constrain the EOS parameters, as compared to the fit to all 5. This information is useful for investigating the extent to which different datasets can con strain EOS models and thereby help guide experimental investigations in order to best constrain the EOS. We also discuss ways that the methodologies can be used to investigate systematic discrepancies between experiments, as well as how the methods can be used to assess model uncertainty. The framework we develop is general, in that it can be used with a variety of optimization or uncertainty quantification techniques and with a variety of data sources, including both experimental and ab-inito data.

97 MATHEMATICS AND COMPUTING↗

Development of an open-source regional data assimilation system in PEcAn v. 1.7.2: application to carbon cycle reanalysis across the contiguous US using SIPNET

Abstract. The ability to monitor, understand, and predict the dynamics of the terrestrial carbon cycle requires the capacity to robustly and coherently synthesize multiple streams of information that each provide partial information about different pools and fluxes. In this study, we introduce a new terrestrial carbon cycle data assimilation system, built on the PEcAn model–data eco-informatics system, and its application for the development of a proof-of-concept carbon “reanalysis” product that harmonizes carbon pools (leaf, wood, soil) and fluxes (GPP, Ra, Rh, NEE) across the contiguous United States from 1986–2019. We first calibrated this system against plant trait and flux tower net ecosystem exchange (NEE) using a novel emulated hierarchical Bayesian approach. Next, we extended the Tobit–Wishart ensemble filter (TWEnF) state data assimilation (SDA) framework, a generalization of the common ensemble Kalman filter which accounts for censored data and provides a fully Bayesian estimate of model process error, to a regional-scale system with a calibrated localization. Combined with additional workflows for propagating parameter, initial condition, and driver uncertainty, this represents the most complete and robust uncertainty accounting available for terrestrial carbon models. Our initial reanalysis was run on an irregular grid of ∼ 500 points selected using a stratified sampling method to efficiently capture environmental heterogeneity. Remotely sensed observations of aboveground biomass (Landsat LandTrendr) and leaf area index (LAI) (MODIS MOD15) were sequentially assimilated into the SIPNET model. Reanalysis soil carbon, which was indirectly constrained based on modeled covariances, showed general agreement with SoilGrids, an independent soil carbon data product. Reanalysis NEE, which was constrained based on posterior ensemble weights, also showed good agreement with eddy flux tower NEE and reduced root mean square error (RMSE) compared to the calibrated forecast. Ultimately, PEcAn's new open-source regional data assimilation framework provides a scalable workflow for harmonizing multiple data constraints and providing a uniform synthetic platform for carbon monitoring, reporting, and verification (MRV) as well as accelerating terrestrial carbon cycle research.

54 ENVIRONMENTAL SCIENCES↗

On determining the spectrum of primordial inhomogeneity from the COBE DMR sky maps: Results of two-year data analysis

A new technique of Fourier analysis on a cut sky has been applied to the two-year Cosmic Background Explorer (COBE) Differential Microwave Radiometer (DMR) 53 and 90 GHz sky maps. The Bayesian power spectrum estimation results are consistent with the Harrison-Zel'dovich n = 1 model. The maximum likelihood estimates of the usual parameters defining the power spectrum of primordial perturbations are n = 1.22 (1.02) and Q(sub rms-PS) = 17 (20) microK including (excluding) the quadrupole. A spectral-index-independent normalization is naturally expressed for the two-year maps in terms of the multipole amplitude a(sub 9) = 8.2 (8.3) microK (to approximately 12 sigma significance). The marginal likelihood function on n obtained by intergration with respect to a(sub 9) renders n = 1.17 +/- 0.31 (0.96 +/- 0.36).

Gorski, K. M.↗

Detecting outbreaks using a spatial latent field

In this paper, we present a method for estimating the infection-rate of a disease as a spatial-temporal field. Our data comprises time-series case-counts of symptomatic patients in various areal units of a region. We extend an epidemiological model, originally designed for a single areal unit, to accommodate multiple units. The field estimation is framed within a Bayesian context, utilizing a parameterized Gaussian random field as a spatial prior. We apply an adaptive Markov chain Monte Carlo method to sample the posterior distribution of the model parameters condition on COVID-19 case-count data from three adjacent counties in New Mexico, USA. Our results suggest that the correlation between epidemiological dynamics in neighboring regions helps regularize estimations in areas with high variance (i.e., poor quality) data. Using the calibrated epidemic model, we forecast the infection-rate over each areal unit and develop a simple anomaly detector to signal new epidemic waves. Our findings show that anomaly detector based on estimated infection-rates outperforms a conventional algorithm that relies solely on case-counts.

Safta, Cosmin [Sandia National Laboratories (SNL-C↗

Cosmological parameter estimation with a joint-likelihood analysis of the cosmic microwave background and big bang nucleosynthesis

Here, we present a joint-likelihood analysis of big bang nucleosynthesis (BBN) and cosmic microwave background (CMB) data, consistently combining likelihoods and taking into account uncertainties in nuclear reaction rates for the first time. Bayesian inference is performed on the baryon abundance and the effective number of neutrino species, 𝑁 eff , using a CMB Boltzmann solver in combination with LINX , a new flexible and efficient BBN code. We marginalize over Planck nuisance parameters and nuclear rates to find 𝑁 eff =3.0⁢8$^{+0.15}_{−0.14}$, 2.9⁢4$^{+0.16}_{−0.15}$, or 2.96$^{+0.13}_{−0.14}$, for three separate reaction networks. This framework enables robust testing of the lambda cold dark matter paradigm and its variants with CMB and BBN data.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Joint Bayesian Inference for Near-Surface Explosion Yield and Height-of-Burst

Forensic capabilities to understand chemical and nuclear explosions are greatly aided by an accurate estimate of explosive yield with uncertainty. The relationship between explosive size and geophysical observations of seismic, acoustic, and optical waves can be exploited to provide an estimate of yield. Any near-surface yield estimate is complicated by the surface interaction, so an estimate for the explosion height-of-burst is necessarily included in the relationship. Additionally, the relationship dictates a trade-off between estimates of yield and height-of-burst. Fortunately, the surface interaction for each type of observation is different, which breaks the trade-off, and the inclusion of height-of-burst with multiple data types improves yield estimation. We define simple parametric forward models to relate seismoäcoustoöptic observations from a data set of known explosive yields and height-of-bursts. The parameters of the models and a prediction for the yield and height-of-burst of a new event can then be estimated given new observations via Bayesian inference. We report posterior distribution estimates of the parametric models using a Markov chain Monte Carlo sampling technique. These models are then used to predict the yield and height-of-burst of SUGAR, a historical near-surface nuclear explosion, using its reported historical observations. The reported yield of 1.2 ktonne Trinitrotoluene (TNT)-equivalent (Department of Energy, 2015) is within the estimated posterior. Yield uncertainty can be estimated from the spread of the posterior, which is between 0.9 and 2.1 ktonne TNT-equivalent. The posterior for height-of-burst has a wider range between 10 m below and 8 m above ground that includes the true height-of-burst of 1 m.

58 GEOSCIENCES↗

Mechanistic within-host mathematical model of inhalational anthrax

We present a mathematical model of the dynamics of Bacillus anthracis bacteria within the lymph nodes and blood of a host, following inhalation of an initial dose of spores. We also incorporate the dynamics of protective antigen, which is the binding component of the anthrax toxin produced by the bacteria. The model offers a mechanistic description of the early infection dynamics of inhalational anthrax, while its stochastic nature allows us to study the probabilities of different outcomes (for example, how likely it is that the infection will be cleared for a given inhaled dose of spores) in order to explain dose-response data for inhalational anthrax. The model is calibrated via a Bayesian approach, using in vivo data from New Zealand white rabbit and guinea pig infection studies, enabling within-host parameters to be estimated. We also leverage incubation-period data from the Sverdlovsk 1979 anthrax outbreak to show that the model can accurately describe human time-to-symptoms data under reasonable parameter regimes. Finally, we derive a simple approximate formula for the probability of symptom onset before time t, assuming that the number of inhaled spores has a Poisson distribution.

59 BASIC BIOLOGICAL SCIENCES↗

Adaptive statistical pattern classifiers for remotely sensed data

A technique for the adaptive estimation of nonstationary statistics necessary for Bayesian classification is developed. The basic approach to the adaptive estimation procedure consists of two steps: (1) an optimal stochastic approximation of the parameters of interest and (2) a projection of the parameters in time or position. A divergence criterion is developed to monitor algorithm performance. Comparative results of adaptive and nonadaptive classifier tests are presented for simulated four dimensional spectral scan data.

Gonzalez, R. C.↗

The Error Distribution of BATSE GRB Location

We develop empirical probability models for BATSE GRB location errors by a Bayesian analysis of the separations between BATSE GRB locations and locations obtained with the InterPlanetary Network (IPN). Models are compared and their parameters estimated using 394 GRBs with single IPN annuli and 20 GRBs with intersecting IPN annuli. Most of the analysis is for the 4B (rev) BATSE catalog; earlier catalogs are also analyzed. The simplest model that provides a good representation of the error distribution has 78% of the locations in a 'core' term with a systematic error of 1.85 degrees and the remainder in an extended tail with a systematic error of 5.36 degrees, implying a 68% confidence region for bursts with negligible statistical errors of 2.3 degrees. There is some evidence for a more complicated model in which the error distribution depends on the BATSE datatype that was used to obtain the location. Bright bursts are typically located using the CONT datatype, and according to the more complicated model, the 68% confidence region for CONT-located bursts with negligible statistical errors is 2.0 degrees.

Briggs, Michael S.↗

The Error Distribution of BATSE Gamma-Ray Burst Locations

Empirical probability models for BATSE gamma-ray burst (GRB) location errors are developed via a Bayesian analysis of the separations between BATSE GRB locations and locations obtained with the Interplanetary Network (IPN). Models are compared and their parameters estimated using 392 GRBs with single IPN annuli and 19 GRBs with intersecting IPN annuli. Most of the analysis is for the 4Br BATSE catalog; earlier catalogs are also analyzed. The simplest model that provides a good representation of the error distribution has 78% of the probability in a "core" term with a systematic error of 1.85 deg and the remainder in an extended tail with a systematic error of 5.1 deg, which implies a 68% confidence radius for bursts with negligible statistical uncertainties of 2.2 deg. There is evidence for a more complicated model in which the error distribution depends on the BATSE data type that was used to obtain the location. Bright bursts are typically located using the CONT data type, and according to the more complicated model, the 68% confidence radius for CONT-located bursts with negligible statistical uncertainties is 2.0 deg.

Briggs, Michael S.↗

Estimation of stagnation performance metrics in magnetized liner inertial fusion experiments using Bayesian data assimilation

Here we present a new analysis methodology that allows for the self-consistent integration of multiple diagnostics including nuclear measurements, x-ray imaging, and x-ray power detectors to determine the primary stagnation parameters, such as temperature, pressure, stagnation volume, and mix fraction in magnetized liner inertial fusion (MagLIF) experiments. The analysis uses a simplified model of the stagnation plasma in conjunction with a Bayesian inference framework to determine the most probable configuration that describes the experimental observations while simultaneously revealing the principal uncertainties in the analysis. We validate the approach by using a range of tests including analytic and three-dimensional MHD models. An ensemble of MagLIF experiments is analyzed, and the generalized Lawson criterion χ is estimated for all experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Performing Bayesian Analyses With AZURE2 Using BRICK: An Application to the 7 Be System

Phenomenological R-matrix has been a standard framework for the evaluation of resolved resonance cross section data in nuclear physics for many years. It is a powerful method for comparing different types of experimental nuclear data and combining the results of many different experimental measurements in order to gain a better estimation of the true underlying cross sections. Yet a practical challenge has always been the estimation of the uncertainty on both the cross sections at the energies of interest and the fit parameters, which can take the form of standard level parameters. Frequentist (χ 2 -based) estimation has been the norm. In this work, a Markov Chain Monte Carlo sampler, emcee, has been implemented for the R-matrix code AZURE2, creating the Bayesian R-matrix Inference Code Kit (BRICK). Bayesian uncertainty estimation has then been carried out for a simultaneous R-matrix fit of the 3 He (α,γ) 7 Be and 3 He (α,α) 3 He reactions in order to gain further insight into the fitting of capture and scattering data. Both data sets constrain the values of the bound state α-particle asymptotic normalization coefficients in 7 Be. The analysis highlights the need for low-energy scattering data with well-documented uncertainty information and shows how misleading results can be obtained in its absence.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Application of a Bayesian Framework for Plasticity Model Selection

Interpretable Machine Learning (IML) has performed well when tasked with deriving constitutive material models. However, IML has been shown to prefer models that overfit noise in data, which tends to lead to bloat and a decrease in interpretability. Due to these issues, the ability of IML to reliably derive models that fit the data and are both interpretable and generalizable is limited. A method developed recently has shown promise to improve upon traditional IML by using a Bayesian fitness definition for the evolution of free-form models with non-deterministic parameters. This framework was developed for genetic-programming-based symbolic regression(GPSR) and involves model parameter estimation using Sequential Monte Carlo sampling (SMC).The method has demonstrated a reduction in bloat when dealing with noisy data in comparison to conventional GPSR. The results of this framework applied to stress-strain data for copper show models that more effectively predict the experimental data better than was previously shown with GPSR.

plasticity↗

Robust Importance Sampling for Bayesian Model Calibration with Spatio-Temporal Data

This paper addresses two challenges in Bayesian calibration: 1) computational speed of existing sampling algorithms, and 2) calibration with spatio-temporal responses. The commonly used Markov Chain Monte Carlo (MCMC) approaches require many sequential model evaluations making the computational expense prohibitive. This paper proposes an efficient sampling algorithm: iterative importance sampling with genetic algorithm (IISGA). While iterative importance sampling enables computational efficiency, the genetic algorithm enables robustness by preventing sample degeneration and avoids getting stuck in multimodal search spaces. An inflated likelihood further enables robustness in high-dimensional parameter spaces by enlarging the target distribution. Spatio-temporal data complicate both surrogate modeling, which is necessary for expensive computational models, and the likelihood estimation. In this work, singular value decomposition is investigated for reducing the high-dimensional field data to a lower-dimensional space prior to Bayesian calibration. Then the likelihood is formulated and Bayesian inference is performed in the lower-dimension, latent space. An illustrative example is provided to demonstrate IISGA relative to existing sampling methods, and then IISGA is employed to calibrate a thermal battery model with 26 uncertain calibration parameters and spatio-temporal response data.

97 MATHEMATICS AND COMPUTING↗

Constraining Bedrock Groundwater Residence Times in a Mountain System With Environmental Tracer Observations and Bayesian Uncertainty Quantification

Groundwater residence time distributions provide fundamental insights on the hydrological processes within watersheds. Yet, observations that can constrain groundwater residence times over broad timescales remain scarce in mountain catchment studies. We use environmental tracers (CFC-12, SF 6 , 3 H, and 4 He) to investigate groundwater residence times along a hillslope in the East River Watershed, Colorado, USA. We develop a Bayesian inference framework that applies a Markov-chain Monte Carlo (MCMC) approach to estimate noble gas recharge temperature, elevation, and excess-air parameters and the resulting environmental tracer concentrations. MCMC is then used to propagate the environmental tracer uncertainties to estimates of groundwater mean residence times inferred with lumped parameter models. All samples contain 3 H, CFC-12, and SF 6 in addition to terrigenic 4 He, suggesting a mixture of water characterized by modern and premodern residence times. 4He exponential mean residence times range from hundreds of years at the upslope well to thousands of years at the toe-slope well assuming average crustal production rates. We find that binary mixing residence time distributions with separate young and old mixing fractions are needed to predict the 4 He, CFC-12, SF 6 , and 3 H observations, supporting the importance of flow path mixing in this bedrock system. Our findings that the fractured bedrock hosts groundwater with a mixture of residence times ranging from decades to millennia suggest variable recharge dynamics and flow path mixing along the hillslope and highlight the importance of characterizing groundwater systems with observations that are sensitive to transport over a broad range of residence times.

54 ENVIRONMENTAL SCIENCES↗