Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Bayesian parameter estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

The decay of HIV under anti-retroviral therapy is biphasic even in humanized mice with just T cells

HIV-1 plasma viral load decays in a biphasic manner during antiretroviral therapy (ART). It was hypothesized that this is due to infection of different cell types, namely CD4+ T cells and macrophages. We studied this possibility directly by modeling the decay of HIV-1 in humanized mice. We utilized previously published data from humanized T-cell only mice (TOM) and myeloid-only mice (MOM) infected with HIV-1 and treated with a potent ART regimen. Viral load decay dynamics were modeled using either a single or a biexponential decay fitted using nonlinear mixed effects techniques. Fits were compared using the corrected Bayesian information criterion (BICc). In TOM, the biphasic model was significantly better than a single-phase decay model (ΔBICc ≈ 16) despite additional parameters. In MOM, the biphasic decay was statistically better, but there was substantial uncertainty because the virus goes below detection very fast. The first-phase half-life was consistent between groups (1.2 days in MOM and 1.3 days in TOM) and similar to the half-life estimated in human infection. The second-phase decay in these mice was minimal likely due to low initial viral loads. Additional analyses with mice containing both CD4+ T cells and macrophages or X4-tropic virus-infected MOM mice confirmed the biphasic pattern, demonstrating the robustness of this result. The biphasic decline in HIV-1 occurs, even with only CD4+ T cells, refuting the hypothesis that distinct cell populations (CD4+ T cells and macrophages) drive each decay phase. These findings support an alternative model in which the observed dynamics arise from intrinsic properties of the viral infection lifecycle rather than from cellular compartmentalization.

59 BASIC BIOLOGICAL SCIENCES↗

Hierarchical Bayesian Thermonuclear Rate for the 7 Be(n,p) 7 Li Big Bang Nucleosynthesis Reaction

Big Bang nucleosynthesis provides the earliest probe of standard model physics, at a time when the universe was less than 1000 seconds old. It determines the abundances of the lightest nuclides, which give rise to the subsequent history of the visible matter in the universe. This work derives new 7 Be(n,p) 7 Li thermonuclear reaction rates based on all available experimental information. This reaction sensitively impacts the primordial abundances of 7 Be and 7 Li during big bang nucleosynthesis. We critically evaluate all available data and disregard experimental results that are questionable. For the nuclear model, we adopt an incoherent sum of single-level, two-channel, R -matrix approximation expressions, which are implemented into a hierarchical Bayesian model, to analyze the remaining six data sets we deem most reliable. In the fitting of the data, we consistently model all known sources of uncertainty, including discrepant absolute normalizations of different data sets, and also take the variation of the neutron and proton channel radii into account, hence providing less biased estimates of the 7 Be(n,p) 7 Li thermonuclear rates. From the resulting posteriors, we extract R -matrix parameters ($E_r, γ^2_n, γ^2_p$) and derive excitation energies and partial and total widths. Our fit is sensitive to the contributions of the first three levels above the neutron threshold. Reaction rates were computed by integrating 10,000 samples of the reduced cross section. Our 7 Be(n,p) 7 Li thermonuclear rates have uncertainties between 1.5% and 2.0% at temperatures of ≤1 GK. Finally, we compare our rates to previous results and find that the 7 Be(n,p) 7 Li rates most commonly used in big bang simulations have uncertainties that are too optimistic.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Uncertainty-Aware Machine Learning for Small-Angle X-ray Scattering Analysis in Autonomous Experimentation

Small-angle X-ray scattering (SAXS) is a powerful high-throughput characterization tool for probing nanoscale structure in native sample environments, providing real-time morphological information such as nanoparticle size and shape during synthesis. However, automated SAXS data analysis for extracting meaningful structural parameters is non-trivial and remains a bottleneck in closed-loop experimentation towards autonomous materials discovery, which demands fast, reliable, and uncertainty-aware data analysis. Here, we develop a machine-learning approach for automated SAXS analysis tailored to closed-loop nanoparticle synthesis. A Random Forest (RF) regression model is trained on 100,000 synthetic SAXS curves generated from polydisperse spherical nanoparticles with realistic background contributions. Using normalized one-dimensional SAXS intensity profiles as input, the RF model directly predicts nanoparticle radius, size polydispersity, and background parameters, while the ensemble standard deviation across trees provides built-in uncertainty quantification (UQ). On synthetic data, we show that combining fit-quality metrics (R 2 , MAE) with thresholds on prediction uncertainty reliably identifies accurate parameter estimates without access to ground truth. We then apply the trained model to 365 experimental SAXS profiles of citrate-reduced gold nanoparticles synthesized using an automated droplet-flow microreactor with in situ SAXS at a synchrotron beamline, classifying the results into high- and low-confidence subsets based on UQ metrics. Finally, we integrate RF-based SAXS analysis into a simulated closed-loop optimization campaign using Gaussian process Bayesian optimization to minimize nanoparticle polydispersity, benchmarking against conventional automated Levenberg–Marquardt fitting. The RF-guided campaign exhibits substantially faster convergence and lower relative opportunity cost (∼0.07 vs ∼0.3), demonstrating that uncertainty-aware machine-learning SAXS analysis significantly enhances the efficiency and robustness of autonomous nanomaterials synthesis workflows.

Bayesian optimization↗

Host Star Metallicity of Directly Imaged Wide-orbit Planets: Implications for Planet Formation

Directly imaged planets (DIPs) are self-luminous companions of pre-main-sequence and young main-sequence stars. They reside in wider orbits (∼tens to thousands of astronomical units) and generally are more massive compared to the close-in (≲10 au) planets. Determining the host star properties of these outstretched planetary systems is important to understand and discern various planet formation and evolution scenarios. We present the stellar parameters and metallicity ([Fe/H]) for a subsample of 18 stars known to host planets discovered by the direct imaging technique. We retrieved the high-resolution spectra for these stars from public archives and used the synthetic spectral fitting technique and Bayesian analysis to determine the stellar properties in a uniform and consistent way. For eight sources, the metallicities are reported for the first time, while the results are consistent with the previous estimates for the other sources. Our analysis shows that metallicities of stars hosting DIPs are close to solar with a mean [Fe/H] = −0.04 ± 0.27 dex. The large scatter in metallicity suggests that a metal-rich environment may not be necessary to form massive planets at large orbital distances. We also find that the planet mass–host star metallicity relation for the directly imaged massive planets in wide orbits is very similar to that found for the well-studied population of short-period (≲1 yr) super-Jupiters and brown dwarfs around main-sequence stars.

36 MATERIALS SCIENCE↗

Measuring the Hubble Constant with Dark Neutron Star–Black Hole Mergers

Abstract Detection of gravitational waves (GWs) from neutron star-black hole (NSBH) standard sirens provides local measurements of the Hubble constant (H 0 ), regardless of the detection of an electromagnetic (EM) counterpart, given that matter effects can be exploited to break the redshift degeneracy of the GW waveforms. The distinctive merger morphology and the high-redshift detectability of tidally disrupted NSBH make them promising candidates for this method. Also, the detection prospects of an EM counterpart for these systems will be limited toz< 0.8 in the optical, in the era of future GW detectors. Using recent constraints on the equation of state of NSs from multi-messenger observations of NICER and LIGO/Virgo/KAGRA, we show the prospects of measuringH 0 solely from GW observation of NSBH systems, achievable by the Einstein telescope (ET) and Cosmic Explorer (CE) detectors. We first analyze individual events to quantify the effect of high-frequency (≥500 Hz) tidal distortions on the inference of NS tidal deformability parameter (Λ) and hence onH 0 . We find that disruptive mergers can constrain Λ up to  ( 60 % ) more precisely than nondisruptive ones. However, this precision is not sufficient to place stringent constraints on theH 0 from individual events. By performing Bayesian analysis on simulated NSBH data (up toN= 100 events, corresponding to a day of observation) in the ET+CE detectors, we find that NSBH systems enable unbiased 4%–13% precision on the estimate ofH 0 (68% credible interval). This is a similar measurement precision found in studies analyzing NSBH mergers with EM counterparts in the LVKC O5 era.

Astronomy & Astrophysics↗

PINN surrogate of Li-ion battery models for parameter inference, Part II: Regularization and application of the pseudo-2D model

Bayesian parameter inference is useful to improve Li-ion battery diagnostics and can help formulate battery aging models. However, it is computationally intensive and cannot be easily repeated for multiple cycles, multiple operating conditions, or multiple replicate cells. To reduce the computational cost of Bayesian calibration, numerical solvers for physics-based models can be replaced with faster surrogates. A physics-informed neural network (PINN) is developed as a surrogate for the pseudo-2D (P2D) battery model calibration. For the P2D surrogate, additional training regularization was needed as compared to the PINN single-particle model (SPM) developed in Part I. Both the PINN SPM and P2D surrogate models are exercised for parameter inference and compared to data obtained from a direct numerical solution of the governing equations. A parameter inference study highlights the ability to use these PINNs to calibrate scaling parameters for the cathode Li diffusion and the anode exchange current density. By realizing computational speed-ups of ~2250x for the P2D model, as compared to using standard integrating methods, the PINN surrogates enable rapid state-of-health diagnostics. Finally, in the low-data availability scenario, the testing error was estimated to ~2 mV for the SPM surrogate and ~10 mV for the P2D surrogate which could be mitigated with additional data.

25 ENERGY STORAGE↗

Bayesian optimization to design a novel x-ray shaping device

In radiation therapy, x-ray dose must be precisely sculpted to the tumor, while simultaneously avoiding surrounding organs at risk. This requires modulation of x-ray intensity in space and/or time. Typically, this is achieved using a multi leaf collimator (MLC) - a complex mechatronic device comprising over one hundred individually powered tungsten ‘leaves’ that move in or out of the radiation field as required. Here, an all-electronic x-ray collimation concept with no moving parts is presented, termed “SPHINX”: Scanning Pencil-beam High-speed Intensity-modulated X-ray source. SPHINX utilizes a spatially distributed bremsstrahlung target and collimator array in conjunction with magnetic scanning of a high energy electron beam to generate a plurality of small x-ray “beamlets.” A simulation framework was developed in Topas Monte Carlo incorporating a phase space electron source, transport through user defined magnetic fields, bremsstrahlung x-ray production, transport through a SPHINX collimator, and dose in water. This framework was completely parametric, meaning a simulation could be built and run for any supplied geometric parameters. This functionality was coupled with Bayesian optimization to find the best parameter set based on an objective function which included terms to maximize dose rate for a user defined beamlet width while constraining inter-channel cross talk and electron contamination. Designs for beamlet widths of 5, 7, and 10 mm 2 were generated. Each optimization was run for 300 iterations and took approximately 40 h on a 24-core computer. For the optimized 7-mm model, a simulation of all beamlets in water was carried out including a linear scanning magnet calibration simulation. Finally, a back-of-envelope dose rate formalism was developed and used to estimate dose rate under various conditions. The optimized 5–, 7–, and 10-mm models had beamlet widths of 5.1 , 7.2 , and 10.1 mm 2 and dose rates of 3574, 6351, and 10 015 Gy/C, respectively. The reduction in dose rate for smaller beamlet widths is a result of both increased collimation and source occlusion. For the simulation of all beamlets in water, the scanning magnet calibration reduced the offset between the collimator channels and beam centroids from 2.9 ±1.9 mm to 0.01 ±0.03 mm. A slight reduction in dose rate of approximately 2% per degree of scanning angle was observed. Based on a back-of-envelope dose rate formalism, SPHINX in conjunction with next-generation linear accelerators has the potential to achieve substantially higher dose rates than conventional MLC-based delivery, with delivery of an intensity modulated 100 x 100 mm 2 field achievable in 0.9 to 10.6 s depending on the beamlet widths used. Bayesian optimization was coupled with Monte Carlo modeling to generate SPHINX geometries for various beamlet widths. A complete Monte Carlo simulation for one of these designs was developed, including electron beam transport of all beamlets through scanning magnets, x-ray production and collimation, and dose in water. These results demonstrate that SPHINX is a promising candidate for sculpting radiation dose with no moving parts, and has the potential to vastly improve both the speed and robustness of radiotherapy delivery. A multi-beam SPHINX system may be a candidate for delivering magavoltage FLASH RT in humans.

62 RADIOLOGY AND NUCLEAR MEDICINE↗

Diagnostics: Chapter 8 of the special issue: on the path to tokamak burning plasma operation

This chapter presents the activity conducted by the ITPA topical group (TG) on Diagnostics over about the last 15 years. Following a general introduction of the ITER Diagnostics led by their measurement roles, the document is organized in several subchapters detailing the design support, research and development activity conducted by each of the specialist working groups (WGs) of the TG. Please note that the magnetic diagnostics were supported at the TG without a specific WG. Their status is included in the general introduction. In the following some highlights of the subchapter’s contents are provided. Recent advances in ITER first wall (FW) diagnostics for the measurements of plasma-metallic wall interaction in support of the ITER research plan are reported. An InfraRed imaging Video Bolometer for ITER has been developed and tested on several tokamaks to measure the radiated power loss. A laser-induced breakdown spectroscopy (LIBS) technique which utilizes a pulsed laser beam to ablate locally by forming a crater, will measure local tritium inventory in the FW material. Real-time Residual Gas Analyzers will measure the neutral gas composition in a divertor port and an equatorial port during plasma operation. Due to the full metallic FW environment, the plasma-wall interaction in ITER will face several challenges such as the compromised radiated power and divertor heat flux measurements by reflection. Ray tracing and analysis codes have been developed to eliminate and correct the effects of reflection in the measurements. The characteristics of the reflecting surfaces depending on the roughness and angle of the incidence have been measured by dedicated experiments, and the results were applied to the reflection elimination. For the measurement of the metallic impurity radiation induced by eroded metallic atoms, a vacuum ultraviolet spectrometer has been developed and tested. An extensive thermonuclear diagnostic suite will be required to support the operation of ITER and the planned experimental program for future burning plasma experiments. Due to the harsh environmental conditions, the implementation of diagnostic systems in ITER is a major challenge. These conditions include high levels of neutron and gamma fluxes, neutron heating, particle bombardment. Therefore, the selection and design of diagnostic systems must take into account a number of phenomena previously unseen in diagnostic design. For this reason, the measurement of neutrons and confined or lost fast ions, with particular emphasis on alpha particles, is critical to ITER. The diagnostics associated with these measurements will be important for future plasma-burning experiments at ITER. The high neutron emission and very large plasma size in ITER make neutron diagnostics the main diagnostic method used to measure plasma parameters such as fusion power, fusion power density, ion temperature, energy of fast ions and their spatial distributions in the plasma core. Active spectroscopy techniques are methods where a neutral particle beam is injected into the plasma and information on plasma parameters is extracted from the measurement of line emission resulting from the beam-plasma interaction, either by plasma ions or by beam atoms. Spatial localization is achieved by crossing the beamline and multiple observation lines. The ITER plasma will be a high temperature, moderately dense, fully ionized collisional plasma. The plasma facing surfaces are principally metallic being fashioned from beryllium or tungsten but many other elements, arising from either structural or from operational needs, may enter this plasma. The energy range of the emitted photons range from meV (infra-red) to multi keV (x-rays) and originate from all areas of the plasma volume. The primary role of passive emission diagnostics is to identify what is in the plasma from spectral signatures. Extracting quantitative information from these measurements such as impurity content, ion temperature, rotation, degree of detachment and radiated power depends on calibrated instruments, a physics model of the atomic and molecular processes and plasma transport and an analysis workflow that takes into account environmental effects such as reflections. The particular needs for ITER have prompted a multi-machine, many-year effort to address all these aspects and this chapter reviews the work on diagnostic design, experiments and new analysis techniques. An overview of the laser diagnostics to be implemented on ITER is also provided in this paper. This includes descriptions of the Thomson scattering in the core, edge and divertor regions, polarimetry and interferometry diagnostics used for measuring plasma density and also measurements of helium density in the divertor using Laser Induced Flourescence. Techniques which can allow improvements on current measurements are also addressed in particular expanding poloidal polarimetry measurements to measure field fluctuations and proposed use of dispersion interferometery which has a number of advantages over existing methods. This paper identifies particular areas where further research and testing on existing tokamaks is useful even at this advanced stage to inform the design of diagnostics for ITER. Outstanding areas of concern for the implementation of laser diagnostics, in particular with a view to reliable operation are identified. An overview of the latest developments of microwave diagnostic systems and techniques is given. The primary focus is the contributions for ITER—the next step burning plasma experiment—which is supplemented by describing recent progress of techniques applicable for fusion experiments beyond ITER. The contributions are intentionally kept concise, and are being supplemented by a rich list of references for further studies. Radiation induced effects are receiving continuous and well-deserved attention of the ITER diagnostic community and they are in many cases one of the primary design drivers of the ITER diagnostic systems. The paper summarizes recent progress in this area focusing primarily on the ITER diagnostics but in some cases provides also outlook for the possible solutions for even more demanding radiation environment of fusion reactors beyond ITER. Despite advancements in the area of modeling and simulation of various radiation induced effects, experimental testing in a nuclear environment as close as possible to the target one is still seen as unavoidable for proper qualification of particular diagnostic functional elements. Recent advancement within three diagnostic areas: optical diagnostics, magnetics and bolometers is covered. Encouraging results on qualification of silica glass vacuum window assemblies are presented. In the area of magnetic sensors, progress of irradiation tests performed on ITER in-vessel LTCC inductive sensors is presented with outlook for novel technological approaches to inductive sensors utilizing thick printing and photolithography technologies being highlighted. Summary of advancements in the area of steady state magnetic field sensors based on Hall effect is given. New results of neutron irradiation test of the ITER borosilicate glass inserts for vacuum electrical feedthroughs are summarized finding negligible swelling at target level of neutron fluence. Off-line irradiation tests of fiber optic current sensors for plasma current measurement demonstrated that both for gamma doses up to 5 MGy and a total neutron fluence up to 10 15 cm −2 , radiation induced changes are still compatible with required measurement accuracy on ITER. The ITER bolometers are given as an example how considering radiation effects may influence the diagnostic design. Finally, outlook for future main R&D directions is outlined. All optical and laser-based diagnostics in ITER will be using mirrors to guide plasma radiation toward detectors, cameras and sensors. In the hostile plasma, radiation and particle environment the optical characteristics of diagnostic mirrors will degrade directly affecting the entire performance of involved diagnostic systems. An assessment of factors affecting mirror performance is provided. Among the prime adverse factors are deposition of plasma impurities, sputtering of mirror surface and steam ingress in the vicinity of mirrors. Within the International Tokamak Physics Activity with active support by ITER central team and domestic agencies, the structured research and development (R&D) program on mitigation of risks for diagnostic mirrors is underway. Within this program the mirror material development, the passive mitigation of mirror degradation by using diagnostic ducts and shutters along with an active mirror recovery program comprising the in-situ mirror cleaning and calibration is underway. Recent developments in diagnostic mirror R&D are described in this Chapter along with an example of their implementation of R&D solutions in ITER Infrared Thermography diagnostic. An assessment of still open engineering and physics questions, considerations on mirror risks during an early phase of ITER operation are given along with an overview of diagnostic mirror evolution in the late ITER operation stage toward the demonstration fusion power plant. Several crucial areas of diagnostic R&D outlined in ITER Research Plan are addressed. The basic control groups in a fusion reactor can be broken-down in five categories: (1) plasma position, magnetic configuration, and plasma current control, (2) profile control and confinement optimization, (3) MHD control and suppression, (4) edge dissipation control, radiation and plasma exhaust control and (5) break-down optimization. These categories are coupled via the physics (a control action in one domain will affect the other domains) and via shared actuators (e.g. ECRH for impurity accumulation avoidance, current density distribution control and MHD suppression). Consequently, a supervisory control system should determine the priority of the various control tasks, their couplings, and the interfaces with the safety and interlock system. For the systematic development of the various controllers taking the complexity of the plasma and the control system into account, a model-based approach is required. A short historical overview is given of the developments in systems and control theory and control engineering with special emphasis on those developments that are most relevant for Nuclear Fusion research and operation. An overview is given of the state of the field of fusion plasma control for the control categories. It will be shown how synthetic diagnostics are being developed in ITER and how they are used in diagnostic design and design validation and how they can be in model-based controller synthesis using relatively simple models. In modern control methods, multiple diagnostics are used to constrain relatively simple models. The constrained models provide an estimate for the state. This opens the route to state controllers, such as model predictive control. A major challenge in nuclear fusion research is the coherent combination of data from heterogeneous diagnostics and modeling codes for machine control and safety as well as physics studies. Measured data from different diagnostics often provide information about the same subset of physical parameters. Additionally, information provided by some diagnostics might be needed for the analysis of other diagnostics. A joint analysis of complementary and redundant data allows, e.g. to improve the reliability of parameter estimation, to increase the spatial and temporal resolution of profiles, to obtain synergistic effects, to consider diagnostics interdependencies and to find and resolve data inconsistencies. Physics-based modeling and parameter relationships provide additional information improving the treatment of ill-posed inversion problems. A coherent combination of all kind of available information within a probabilistic framework allows for improved data analysis results. The concept of integrated data analysis (IDA) in the framework of Bayesian probability theory is outlined and contrasted with conventional data analysis. Components of the probabilistic approach are summarized and specific ingredients beneficial for data analysis at fusion devices are discussed.

ITER↗

Using PyBioNetFit to leverage qualitative and quantitative data in biological model parameterization and uncertainty quantification

Data generated in studies of cellular regulatory systems are often qualitative. For example, measurements of signaling readouts in the presence and absence of mutations may reveal a rank ordering of responses across conditions but not the precise extents of mutation-induced differences. Qualitative data are often ignored by mathematical modelers or are considered in an ad hoc manner, as in the study of Kocieniewski and Lipniacki (2013) [Phys Biol 10: 035006], which was focused on the roles of MEK isoforms in ERK activation. In this earlier study, model parameter values were tuned manually to obtain consistency with a combination of qualitative and quantitative data. This approach is not reproducible, nor does it provide insights into parametric or prediction uncertainties. Here, starting from the same data and the same ordinary differential equation (ODE) model structure, we generate formalized statements of qualitative observations, making these observations more reusable, and we improve the model parameterization procedure by applying a systematic and automated approach enabled by the software package PyBioNetFit. We also demonstrate uncertainty quantification (UQ), which was absent in the original study. Our results show that PyBioNetFit enables qualitative data to be leveraged, together with quantitative data, in parameterization of systems biology models and facilitates UQ. These capabilities are important for reliable estimation of model parameters and model analyses in studies of cellular regulatory systems and reproducibility.

59 BASIC BIOLOGICAL SCIENCES↗

Encoding nonlinear and unsteady aerodynamics of limit cycle oscillations using nonlinear sparse Bayesian learning

This article investigates the applicability of a recently proposed, nonlinear sparse Bayesian learning (NSBL) algorithm to identify and estimate the complex aerodynamics of limit cycle oscillations. NSBL provides a semi-analytical framework for determining the data-optimal sparse model nested within a (potentially) over-parameterized model. This is particularly relevant to nonlinear dynamical systems where modelling approaches involve the use of physics-based and data-driven components. In such cases, the data-driven components, where analytical descriptions of the physical processes are not readily available, are often prone to overfitting, meaning that the empirical aspects of these models will often involve the calibration of an unnecessarily large number of parameters. While an overparameterized model may fit the observed data well, such models may be inadequate for making predictions in regimes that are different from those wherein the data were recorded. In view of this, it is desirable to not only calibrate the model parameters, but also identify the optimal compromise between data fit and model complexity. In this article, we exhibit the optimal model discovery for an aeroelastic system wherein the structural dynamics are well-known and described by a differential equation model, coupled with a semi-empirical aerodynamic model for laminar separation flutter, resulting in low-amplitude limit cycle oscillations (LCO). To illustrate the performance of the algorithm, in this article, we use synthetic data and demonstrate the ability of the algorithm to correctly rediscover the optimal model and model parameters, given a known data-generating model. The synthetic data are generated from a forward simulation of a known differential equation model with parameters selected so as to mimic the dynamics observed in wind-tunnel experiments. Subsequently, we demonstrate the performance of the algorithm for model selection using noisy LCO data from wind tunnel experiments. As there is no ground truth available for the experimental data case, we provide a comparison between NSBL and Bayesian model selection to validate the results, and demonstrate the use of NSBL as an efficient alternative to traditional methods.

97 MATHEMATICS AND COMPUTING↗

Quantifying modeling uncertainty in simplified beam models for building response prediction

The use of simple models for response prediction of building structures is preferred in earthquake engineering for risk evaluations at regional scales, as they make computational studies more feasible. The primary impediment in their gainful use presently is the lack of viable methods for quantifying (and reducing upon) the modeling errors/uncertainties they bear. This study presents a Bayesian calibration method wherein the modeling error is embedded into the parameters of the model. Here, the method is specifically described for coupled shear-flexural beam models here, but it can be applied to any parametric surrogate model. The major benefit the method offers is the ability to consider the modeling uncertainty in the forward prediction of any degree-of-freedom or composite response regardless of the data used in calibration. The method is extensively verified using two synthetic examples. In the first example, the beam model is calibrated to represent a similar beam model but with enforced modeling errors. In the second example, the beam model is used to represent the detailed finite element model of a 52-story building. Both examples show the capability of the proposed solution to provide realistic uncertainty estimation around the mean prediction.

47 OTHER INSTRUMENTATION↗

Probabilistic Context Neighborhood model for lattices

Here we present the Probabilistic Context Neighborhood model designed for two-dimensional lattices as a variation of a Markov random field assuming discrete values. In this model, the neighborhood structure has a fixed geometry but a variable order, depending on the neighbors’ values. Our model extends the Probabilistic Context Tree model, originally applicable to one-dimensional space. It retains advantageous properties, such as representing the dependence neighborhood structure as a graph in a tree format, facilitating an understanding of model complexity. Furthermore, we adapt the algorithm used to estimate the Probabilistic Context Tree to estimate the parameters of the proposed model. We illustrate the accuracy of our estimation methodology through simulation studies. Additionally, we apply the Probabilistic Context Neighborhood model to spatial real-world data, showcasing its practical utility.

97 MATHEMATICS AND COMPUTING↗

Enhancing Gaussian Process Surrogates for Optimization and Posterior Approximation via Random Exploration

This paper proposes novel noise-free Bayesian optimization strategies that rely on a random exploration step to enhance the accuracy of Gaussian process surrogate models. The new algorithms retain the ease of implementation of the classical GP-UCB algorithm, but the additional random exploration step accelerates their convergence, nearly achieving the optimal convergence rate. Furthermore, to facilitate Bayesian inference with intractable likelihoods, we propose to utilize optimization iterates for maximum a posteriori estimation to build a Gaussian process surrogate model for the unnormalized log-posterior density. We provide bounds for the Hellinger distance between the true and the approximate posterior distributions in terms of the number of design points. We demonstrate the effectiveness of our Bayesian optimization algorithms in nonconvex benchmark objective functions, in a machine learning hyperparameter tuning problem, and in a black-box engineering design problem. The effectiveness of our posterior approximation approach is demonstrated in two Bayesian inference problems for parameters of dynamical systems.

Bayesian inference↗

Thoughts on Analyzing Neutron Multiplicity Data

Bayesian methods offer many advantages for analyzing neutron multiplicity data. Ideally one would make nonparametric estimates of the the posterior distributions, the probabilities - given the data - for unknowns to take on particular values. Once one has the posterior, uncertainty quantification is very natural. We describe how different ways of formulating the statistical theory of fission chains are mathematically equivalent. When confronted with actual data and the burden of finite statistics, different methods of expressing the theory offer advantages for understanding the underlying source parameters from an information theory standpoint.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Inferring Type II-P Supernova Progenitor Masses from Plateau Luminosities

Abstract Connecting observations of core-collapse supernova explosions to the properties of their massive star progenitors is a long-sought, and challenging, goal of supernova science. Recently, Barker et al. presented bolometric light curves for a landscape of progenitors from spherically symmetric neutrino-driven core-collapse supernova (CCSN) simulations using an effective model. They find a tight relationship between the plateau luminosity of the Type II-P CCSN light curve and the terminal iron-core mass of the progenitor. Remarkably, this allows us to constrain progenitor properties with photometry alone. We analyze a large observational sample of Type II-P CCSN light curves and estimate a distribution of iron-core masses using the relationship of Barker et al. The inferred distribution matches extremely well with the distribution of iron-core masses from stellar evolutionary models and namely, contains high-mass iron cores that suggest contributions from very massive progenitors in the observational data. We use this distribution of iron-core masses to infer minimum and maximum masses of progenitors in the observational data. Using Bayesian inference methods to locate optimal initial mass function parameters, we find M min = 9.8 − 0.27 + 0.37 and M max = 24.0 − 1.9 + 3.9 solar masses for the observational data.

79 ASTRONOMY AND ASTROPHYSICS↗

Taylor approximation variance reduction for approximation errors in PDE-constrained Bayesian inverse problems

In numerous applications, surrogate models are used as a replacement for accurate parameter-to-observable mappings when solving large-scale inverse problems governed by partial differential equations (PDEs). The surrogate model may be a computationally cheaper alternative to the accurate parameter-to-observable mappings and/or may ignore additional unknowns or sources of uncertainty. The Bayesian approximation error (BAE) approach provides a means to account for the induced uncertainties and approximation errors, i.e. the errors between the accurate parameter-to-observable mapping and the surrogate. The statistics of these errors are, however, in general unknown a priori, and are thus calculated using Monte Carlo sampling. Although the sampling is typically carried out offline, i.e. before considering the data, the process can still represent a computational bottleneck. In this work, we develop a scalable computational approach for reducing the costs associated with the sampling stage of the BAE approach. Specifically, we consider the Taylor expansion of the accurate and surrogate forward models with respect to the uncertain parameter fields either as a control variate for variance reduction or as a means to directly and efficiently approximate the mean and covariance of the approximation errors. We propose efficient methods for evaluating the expressions for the mean and covariance of the Taylor approximations based on linear(-ized) PDE solves. Furthermore, the proposed approach is independent of the dimension of the uncertain parameter, depending instead on the intrinsic dimension of the data, ensuring scalability to high-dimensional problems. The potential benefits of the proposed approach are demonstrated for two high-dimensional inverse problems governed by PDE examples, namely for the estimation of a distributed Robin boundary coefficient in a linear diffusion problem, and for a coefficient estimation problem governed by a nonlinear diffusion problem.

Bayesian approximation error↗

Unbinned extraction of $γ$ from $B\to DK$ with normalizing flows

We introduce an unbinned method for extracting the CKM angle $γ$ from the decay chain $B^\pm \to (D \to K_S π^+ π^-) K^\pm$ using normalizing flows (NFs). The NFs, trained on $D$ decay data, learn a faithful continuous representation of the amplitude and strong phase variation over the $D\to K_Sπ^+π^-$ Dalitz plot whose fidelity improves with increased data sample sizes. With this input, the $B$ decay data can be used to extract the parameters $r_B$, $δ_B$, and $γ$. We test the method on Monte Carlo generated data, where it successfully recovers the injected value of $γ$ within uncertainties. The present implementation propagates statistical uncertainties from finite training data via an ensemble of independently trained flows, and does not attempt to capture the effects of systematic experimental errors. We explore two versions of the method that differ in how the trigonometric constraint on phase variation is encoded, and comment on the possible extension to Bayesian NFs, which would provide direct uncertainty estimates on the learned densities without requiring ensemble training.

Grossman, Yuval [Cornell U., LEPP]↗

Hierarchical Bayesian Modeling for Cosmology: Can NPE reliably replace MCMC?

Hierarchical neural posterior estimation has its place Hierarchical Bayesian Modeling (HBM) combined with MCMC algorithms has been shown to provide more robust and accurate inference for real-world phenomena in which nature takes a nested form. However, MCMC-based inference can be computationally expensive, and its performance often suffers for complex posterior geometries. These costs are especially pertinent for HBM. Studies have recently demonstrated the potential for a flexible, expressive, and amortized hierarchical neural posterior estimator (HNPE) built on Normalizing Flows. These studies have mostly been performed on simple datasets, or they focus on a single parameter from each level of the hierarchy. A systematic study analyzing how both hierarchical methods compare for more complex and realistic datasets is necessary before applying HNPE for scientific measurements. Here, we re-explore the theory behind HNPE and conduct comparative numerical experiments of HNPE and MCMC-based HBM methods on real and synthetic data, including strong gravitational lensing simulations. In particular, we use a suite of diagnostics to show trade-offs in terms of accuracy, precision, time to train or sample, reproducibility, and the need for expert domain knowledge. Especially for higher dimensional and complex posteriors, HNPE is expected to drastically improve on time for inference, accuracy, and precision with an upfront training time cost.

Hur, Rachel [Chicago U.] (ORCID:000900089890445X)↗