Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Bayesian Statistics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Bayesian Estimation of Earth’s Undiscovered Mineralogical Diversity Using Noninformative Priors

Recently, statistical distributions have been explored to provide estimates of the mineralogical diversity of Earth, and Earth-like planets. In this paper, a Bayesian approach is introduced to estimate Earth’s undiscovered mineralogical diversity. Samples are generated from a posterior distribution of the model parameters using Markov chain Monte Carlo simulations such that estimates and inference are directly obtained. It was previously shown that the mineral species frequency distribution conforms to a generalized inverse Gauss–Poisson (GIGP) large number of rare events model. Even though the model fit was good, the population size estimate obtained by using this model was found to be unreasonably low by mineralogists. In this paper, several zero-truncated, mixed Poisson distributions are fitted and compared, where the Poisson-lognormal distribution is found to provide the best fit. Subsequently, the population size estimates obtained by Bayesian methods are compared to the empirical Bayes estimates. Species accumulation curves are constructed and employed to estimate the population size as a function of sampling size. Finally, the relative abundances, and hence the occurrence probabilities of species in a random sample, are calculated numerically for all mineral species in Earth’s crust using the Poisson-lognormal distribution. These calculations are connected and compared to the calculations obtained in a previous paper using the GIGP model for which mineralogical criteria of an Earth-like planet were given.

Bayesian statistics↗

Improved naive Bayesian probability classifier in predictions of nuclear mass

Recently, novel statistical methods such as neural networks and Bayesian learning methods are implemented to describe the nuclear masses. Based on previous studies, an improved naive Bayesian probability (iNBP) classifier is proposed to study the nuclear masses by refining the results of sophisticated nuclear models. In the iNBP method, the prediction for nuclear masses is treated as a classification problem. The residuals are classified into several groups to generate prior and conditional probabilities, and the posterior probabilities are further determined by the Bayesian formula. We choose the expectation with maximum probability as the final prediction. Reliability of the iNBP method is assessed by analyzing the global optimizations and the extrapolating capabilities. Here, the iNBP method exhibits impressive improvements on global descriptions for different mass models. Moreover, the method shows robust extrapolating capabilities. Results demonstrate the iNBP method can be applied to predict the nuclear masses of unknown regions. Considering the local mass relations, the iNBP method can offer considerable fine-tuning of the mass descriptions from nuclear models. The methodology proposed in this paper can also be applied to other model-based extrapolations of nuclear observables.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Bayesian prior construction for uncertainty quantification in first-principles statistical mechanics

First-principles statistical mechanics enables the prediction of thermodynamic and kinetic properties of materials, but is computationally expensive. Many approaches require surrogate models to calculate energies within Monte Carlo or molecular dynamics simulations. Inexpensive surrogates such as cluster expansions enable otherwise intractable calculations by interpolating data from higher accuracy methods, such as Density Functional Theory (DFT). Surrogate models introduce uncertainty into downstream calculations, in addition to any uncertainty inherent to DFT calculations. Bayesian frameworks address this by quantifying uncertainty and incorporating expert knowledge through priors. However, constructing effective priors remains challenging. This work introduces and describes practical strategies for building Bayesian cluster expansions, focusing on basis truncation, hyperparameter selection, and ground state replication. We analyze multiple basis truncation schemes, compare cross-validation to the evidence-approximation for hyperparameter optimization, and provide methods to find and enforce ground-state-preserving models through priors. Additionally, we compare the uncertainties between different approximations to DFT (LDA, PBE, SCAN) against the uncertainty introduced with the use of cluster expansion surrogate models. These approaches are demonstrated on the BCC Li x Mg 1-x and Li x Al 1-x alloys, which are both of interest for solid-state Li batteries. Our results provide guidelines for constructing and utilizing Bayesian cluster expansions, thereby improving the transparency of materials modeling. Furthermore, the approaches and insights developed in this work can be transferred to a wide range of cluster expansion surrogate models, including the atomic cluster expansion and related machine-learned interatomic potential architectures.

Alloy theory↗

Expanding neutrino oscillation parameter measurements in NOvA using a Bayesian approach

NOvA is a long-baseline neutrino oscillation experiment that measures oscillations in charged-current ν μ → ν μ (disappearance) and ν μ → ν e (appearance) channels, and their antineutrino counterparts, using neutrinos of energies around 2 GeV over a distance of 810 km. In this work we reanalyze the dataset first examined in our previous paper [] using an alternative statistical approach based on Bayesian Markov chain Monte Carlo. We measure oscillation parameters consistent with the previous results. We also extend our inferences to include the first NOvA measurements of the reactor mixing angle θ 13 , where we find 0.071 ≤ sin 2 2 θ 13 ≤ 0.107 , and the Jarlskog invariant, where we observe no significant preference for the C P -conserving value J = 0 over values favoring C P violation. We use these results to examine the effects of constraints from short-baseline measurements of θ 13 using antineutrinos from nuclear reactors when making NOvA measurements of θ 23 . Our long-baseline measurement of θ 13 is shown to be consistent with the reactor measurements, supporting the general applicability and robustness of the Pontecorvo-Maki-Nakagawa-Sakata framework for neutrino oscillations. Published by the American Physical Society 2024

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Bayesian Inference and the Effects of Varying Uncertainty Models in Charring Ablator Calibration and Uncertainty Quantification Problems

The Mars Science Laboratory (MSL) vehicle utilized a heat shield constructed from NASA’s Phenolic-Impregnated Carbon Ablator (PICA) material to protect the main structure from the high enthalpy environment encountered during hypersonic atmospheric entry. During the vehicle’s descent through Martian atmosphere, multiple thermocouples embedded within the heat shield captured in-depth material temperature data that allow for studies to be conducted on current material response reconstruction tools. In the present work, material temperature data obtained from thermocouples within the MISP-4 plug (MEDLI (Mars Science Laboratory Entry, Descent, and Landing Instrument) Integrated Sensor Plug) are utilized in the calibration of Theoretical Ablative Composite for Open Testing (TACOT) model parameters in conjunction with NASA’s Porous material Analysis Toolbox (PATO) through Bayesian inference where uncertainty due to parametric, modeling, and experimental sources is simultaneously quantified. Prior to the study, a sensitivity analysis is performed through computation of the robust Sobol indices in an effort to study the relationship between input space and model response and to reduce the dimensionality of the statistical inverse problem. The Bayesian inference methodology necessitates an a-priori choice to be made for the uncertainty model for which numerous possibilities are available. Across most works, however, only basic additive or multiplicative models are utilized with pre-defined magnitudes of uncertainty based on a-priori knowledge or to-be-calibrated multipliers of static covariance matrix structures. The present effort explores the effects of informed uncertainty models, ones with temporal dependence that are simultaneously calibrated through Bayesian inference, on calibrated results for parameters that make up the uncertain input space.

Sensitivity Analysis↗

Statistical Uncertainty in Paleoclimate Proxy Reconstructions

A quantitative analysis of any environment older than the instrumental record relies on proxies. Uncertainties associated with proxy reconstructions are often underestimated, which can lead to artificial conflict between different proxies, and between data and models. In this paper, using ordinary least squares linear regression as a common example, we describe a simple, robust and generalizable method for quantifying uncertainty in proxy reconstructions. We highlight the primary controls on the magnitude of uncertainty, and compare this simple estimate to equivalent estimates from Bayesian, nonparametric and fiducial statistical frameworks. We discuss when it may be possible to reduce uncertainties, and conclude that the unexplained variance in the calibration must always feature in the uncertainty in the reconstruction. This directs future research toward explaining as much of the variance in the calibration data as possible. We also advocate for a “data-forward” approach, that clearly decouples the presentation of proxy data from plausible environmental inferences.

58 GEOSCIENCES↗

Measurement of charge state distributions using a scintillation screen

Absolute cross sections measured using electromagnetic devices to separate and detect heavy recoiling ions need to be corrected for charge state fractions. Accurate prediction of charge state distributions using theoretical models is not always a possibility, especially in energy and mass regions where data is sparse. As such, it is often necessary to measure charge state fractions directly. In this paper we present a novel method of using a scintillation screen along with a CMOS camera to image the charge dispersed beam after a set of magnetic dipoles. A measurement of the charge state distribution for 88 Sr passing through a natural carbon foil is performed. Using a Bayesian model to extract statistically meaningful uncertainties from these images, we find agreement between the new method and a more traditional method using Faraday cups. Additional future work is need to better understand systematic uncertainties. Our technique offers a viable method to measure charge state distributions.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Bayesian projection pursuit regression

In projection pursuit regression (PPR), a univariate response variable is approximated by the sum of $M$ “ridge functions,” which are flexible functions of one-dimensional projections of a multivariate input variable. Traditionally, optimization routines are used to choose the projection directions and ridge functions via a sequential algorithm, and $M$ is typically chosen via cross-validation. Here, we introduce a novel Bayesian version of PPR, which has the benefit of accurate uncertainty quantification. To infer appropriate projection directions and ridge functions, we apply novel adaptations of methods used for the single ridge function case ($M$=1), called the Bayesian Single Index Model; and use a Reversible Jump Markov chain Monte Carlo algorithm to infer the number of ridge functions $M$. We evaluate the predictive ability of our model in 20 simulated scenarios and for 23 real datasets, in a bake-off against an array of state-of-the-art regression methods. Finally, we generalize this methodology and demonstrate the ability to accurately model multivariate response variables. Its effective performance indicates that Bayesian Projection Pursuit Regression is a valuable addition to the existing regression toolbox.

97 MATHEMATICS AND COMPUTING↗

Improved information criteria for Bayesian model averaging in lattice field theory

Bayesian model averaging is a practical method for dealing with uncertainty due to model specification. Use of this technique requires the estimation of model probability weights. Here, we revisit the derivation of estimators for these model weights. Use of the Kullback-Leibler divergence as a starting point leads naturally to a number of alternative information criteria suitable for Bayesian model weight estimation. We explore three such criteria, known to the statistics literature before, in detail: a Bayesian analog of the Akaike information criterion which we call the BAIC, the Bayesian predictive information criterion, and the posterior predictive information criterion (PPIC). We compare the use of these information criteria in numerical analysis problems common in lattice field theory calculations. We find that the PPIC has the most appealing theoretical properties and can give the best performance in terms of model-averaging uncertainty, particularly in the presence of noisy data, while the BAIC is a simple and reliable alternative.

97 MATHEMATICS AND COMPUTING↗

Differentiable Multiphysics Codes: A Breakthrough Technology for Simulation and Computing

This document summarizes the findings of a strategic planning exercise commissioned by the Weapons Simulation and Computing, Computational Physics (WSC/CP) program at the Lawrence Livermore National Laboratory (LLNL) in FY24. During the year, the committee met with multiple stakeholder communities to gather input, opinions, suggestions and concerns which have been incorporated throughout this document. The key findings from this exercise are summarized: • The development of multiphysics modelling and simulation (mod/sim) codes and software technologies, their deployment on exascale compute platforms, and their broad adoption across the NNSA is a major success of the Advanced Simulation and Computing (ASC) program and the Exascale Computing Project (ECP). Sustained investment in these core technologies is essential. • Today’s state of the art involves running ensembles of O(100K) simulations to perform uncertainty quantification (UQ) and design studies using multiple statistical methods such as Bayesian optimization to understand sensitivities of our models and explore parameterized design spaces. Even with exascale computing, we are practically limited to O(10) parameters in these studies since the number of simulations required to sample the space scales exponentially with the number of design parameters. • The data from these simulation ensembles is increasingly being used to train machine learned (ML) surrogates (or reduced order models, ROMs) which can then be used for optimization or real time design exploration. However, the trained surrogates are still limited in the number of parameters they can represent due to the sampling limitations previously noted. • Augmenting our suite of integrated multiphysics simulation codes, both current and emerging, with the ability to compute gradients (solution derivatives) of arbitrary simulation outputs with respect to (some or all) simulation inputs would be a breakthrough technology, opening the door to a new era of efficient and automated inverse design based on verified and validated mod/sim capabilities. • This capability, which we refer to as differentiable multiphysics codes (DMCs), would revolutionize both UQ and optimization studies by breaking the curse of dimensionality that presently limits our “gradient-free” ensemble based computing approach. A similar breakthrough occurred in the AI/ML community once the ability to compute gradients of arbitrary loss functions using back-propagation became commonplace. Gradient information from the multiphysics codes can also be used to dramatically improve the efficiency and scale of training of ML/ROM surrogates for rapid assessments. • Achieving this in our suite of codes will be a grand challenge, similar to the amount of effort that was required to transition from CPU to GPU computing. It will require buy-in from the entire WSC/CP program and beyond, including all integrated codes, physics and engineering models, third-party library dependencies and performance portability abstractions. It will also require investment in research and development of numerical methods for computing adjoints of coupled physics across multiple adaptively refined moving meshes and of stochastic (Monte Carlo) and mesh free (SPH) methods. • New software and numerical techniques, largely pioneered by the AI/ML community, make this feasible. Chief among these is automatic differentiation (AD), the ability to employ AD at point-wise locations in a physics calculation (instead of traditional black-box approaches) and the ability to perform “back-propagation in time” (or reverse mode AD) for non-linear partial differential equations (PDEs). Fundamentally, the conclusion of this strategic planning exercise is that the time is right to undertake a large scale effort in WSC, centered on the existing integrated codes, to continue the natural evolution of mod/sim in the age of AI/ML. Instead of attempting to replace mod/sim with purely data driven AI/ML models, we believe the key to success is to integrate AI/ML by building on top of the decades of hard-won knowledge and the verified/validated multiphysics modelling capability that is the hallmark of the ASC program.

97 MATHEMATICS AND COMPUTING↗

Nondeterministic data base for computerized visual perception

A description is given of the knowledge representation data base in the perception subsystem of the Mars robot vehicle prototype. Two types of information are stored. The first is generic information that represents general rules that are conformed to by structures in the expected environments. The second kind of information is a specific description of a structure, i.e., the properties and relations of objects in the specific case being analyzed. The generic knowledge is represented so that it can be applied to extract and infer the description of specific structures. The generic model of the rules is substantially a Bayesian representation of the statistics of the environment, which means it is geared to representation of nondeterministic rules relating properties of, and relations between, objects. The description of a specific structure is also nondeterministic in the sense that all properties and relations may take a range of values with an associated probability distribution.

Yakimovsky, Y.↗

Climate Change Adaptation Activities at the NASA John F. Kennedy Space Center, FL., USA

In 2010, the Office of Strategic Infrastructure and Earth Sciences established the Climate Adaptation Science Investigators (CASI) program to integrate climate change forecasts and knowledge into sustainable management of infrastructure and operations needed for the NASA mission. NASA operates 10 field centers valued at $32 billion dollars, occupies 191,000 acres and employs 58,000 people. CASI climate change and sea-level rise forecasts focus on the 2050 and 2080 time periods. At the 140,000 acre Kennedy Space Center (KSC) data are used to simulate impacts on infrastructure, operations, and unique natural resources. KSC launch and processing facilities represent a valued national asset located in an area with high biodiversity including 33 species of special management concern. Numerical and advanced Bayesian and Monte Carlo statistical modeling is being conducted using LiDAR digital elevation models coupled with relevant GIS layers to assess potential future conditions. Results are provided to the Environmental Management Branch, Master Planning, Construction of Facilities, Engineering Construction Innovation Committee and our regional partners to support Spaceport development, management, and adaptation planning and design. Potential impacts to natural resources include conversion of 50% of the Center to open water, elevation of the surficial aquifer, alterations of rainfall and evapotranspiration patterns, conversion of salt marsh to mangrove forest, reductions in distribution and extent of upland habitats, overwash of the barrier island dune system, increases in heat stress days, and releases of chemicals from legacy contamination sites. CASI has proven successful in bringing climate change planning to KSC including recognition of the need to increase resiliency and development of a green managed shoreline retreat approach to maintain coastal ecosystem services while maximizing life expectancy of Center launch and payload processing resources.

plant communities↗

Using Gaia Data 2 to Constrain Local Dark Matter Density and Thin Dark Disk

We use stellar kinematics from the latest Gaia data release (DR2) to measure the local dark matter (DM) density ρDM in a heliocentric cylinder of radius R = 150 pc and half-height z = 200 pc. We also explore the prospect of using our analysis to estimate the DM density in local substructure by setting constraints on the surface density and scale height of a thin dark disk aligned with the baryonic disk and formed due to dissipative dark matter self-interactions. Performing the statistical analysis within a Bayesian framework for three types of tracers, we obtain ρDM = 0.016 ± 0.010 Mꙩ/p c 3 for A stars; early G stars give a similar result, while F stars yield a significantly higher value. For a thin dark disk, A stars set the strongest constraint: excluding surface densities (5–12) Mꙩ/pc 2 for scale heights below 100 pc with 95% confidence. The upper bound of this constraint implies . 1% of the Milky Way DM mass is present in a dissipative dark sector. Comparing our results with those derived using Tycho-Gaia Astrometric Solution (TGAS) data, we find that the uncertainty in our measurements of the local DM content is dominated by systematic errors that arise from assumptions of our dynamical analysis in the low z region. Furthermore, there will only be a marginal reduction in these uncertainties with more data in the Gaia era. We comment on the robustness of our method and discuss potential improvements for future work.

dark matter theory↗

Applications of emulation and Bayesian methods in heavy-ion physics

Abstract Heavy-ion collisions provide a window into the properties of many-body systems of deconfined quarks and gluons. Understanding the collective properties of quarks and gluons is possible by comparing models of heavy-ion collisions to measurements of the distribution of particles produced at the end of the collisions. These model-to-data comparisons are extremely challenging, however, because of the complexity of the models, the large amount of experimental data, and their uncertainties. Bayesian inference provides a rigorous statistical framework to constrain the properties of nuclear matter by systematically comparing models and measurements. This review covers model emulation and Bayesian methods as applied to model-to-data comparisons in heavy-ion collisions. Replacing the model outputs (observables) with Gaussian process emulators is key to the Bayesian approach currently used in the field, and both current uses of emulators and related recent developments are reviewed. The general principles of Bayesian inference are then discussed along with other Bayesian methods, followed by a systematic comparison of seven recent Bayesian analyses that studied quark-gluon plasma properties, such as the shear and bulk viscosities. The latter comparison is used to illustrate sources of differences in analyses, and what it can teach us for future studies.

Paquet, Jean-François (ORCID:0000000187368171)↗

Cluster expansion by transfer learning for phase stability predictions

Recent progress towards universal machine-learned interatomic potentials holds considerable promise for materials discovery. Yet the accuracy of these potentials for predicting phase stability may still be limited. In contrast, cluster expansions provide accurate phase stability predictions but are computationally demanding to parameterize from first principles, especially for structures of low dimension or with a large number of components, such as interfaces or multimetal catalysts. We overcome this trade-off via transfer learning. Using Bayesian inference, we incorporate prior statistical knowledge from machine-learned and physics-based potentials, enabling us to sample the most informative configurations and to efficiently fit first-principles cluster expansions. Furthermore, this algorithm is tested on Pt:Ni, showing robust convergence of the mixing energies as a function of sample size with reduced statistical fluctuations.

36 MATERIALS SCIENCE↗

Exoplanet Biosignatures: Future Directions

We introduce a Bayesian method for guiding future directions for detection of life on exoplanets. We describe empirical and theoretical work necessary to place constraints on the relevant likelihoods, including those emerging from better understanding stellar environment, planetary climate and geophysics, geochemical cycling, the universalities of physics and chemistry, the contingencies of evolutionary history, the properties of life as an emergent complex system, and the mechanisms driving the emergence of life. We provide examples for how the Bayesian formalism could guide future search strategies, including determining observations to prioritize or deciding between targeted searches or larger lower resolution surveys to generate ensemble statistics and address how a Bayesian methodology could constrain the prior probability of life with or without a positive detection.

Bayesian analysis↗

Bayesian Classification Scheme

Scheme derived via statistical approach identifies classes in sets of data. Determines probable number of classes, probabilistic descriptions of classes, and probability that each object is member of each class. Scheme applicable to real-valued, discrete or continuous data presented in form of parameter vectors representing attributes of objects.

Stutz, John↗

Trust Your Gut: Comparing Human and Machine Inference from Noisy Visualizations

People commonly utilize visualizations not only to examine a given dataset, but also to draw generalizable conclusions about the underlying models or phenomena. Prior research has compared human visual inference to that of an optimal Bayesian agent, with deviations from rational analysis viewed as problematic. However, human reliance on non-normative heuristics may prove advantageous in certain circumstances. We investigate scenarios where human intuition might surpass idealized statistical rationality. In two experiments, we examine individuals’ accuracy in characterizing the parameters of known data-generating models from bivariate visualizations. Our findings indicate that, although participants generally exhibited lower accuracy compared to statistical models, they frequently outperformed Bayesian agents, particularly when faced with extreme samples. Participants appeared to rely on their internal models to filter out noisy visualizations, thus improving their resilience against spurious data. However, participants displayed overconfidence and struggled with uncertainty estimation. They also exhibited higher variance than statistical machines. Our findings suggest that analyst gut reactions to visualizations may provide an advantage, even when departing from rationality. These results carry implications for designing visual analytics tools, offering new perspectives on how to integrate statistical models and analyst intuition for improved inference and decision-making. The data and materials for this paper are available at https://osf.io/qmfv6

human-machine collaboration↗