Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Maximum likelihood estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

34 records · Page 2

Discovery of Probabilistic Dirichlet-to-Neumann Maps on Graphs

Dirichlet-to-Neumann maps enable the coupling of multiphysics simulations across computational subdomains by ensuring continuity of state variables and fluxes at artificial interfaces. We present a novel method for learning Dirichlet-to-Neumann maps on graphs using Gaussian processes, specifically for problems where the data obey a conservation law arising from an underlying partial differential equation. Our approach combines discrete exterior calculus and nonlinear optimal recovery to infer relationships between vertex and edge values. This framework yields data-driven predictions with uncertainty quantification across the entire graph, even when observations are limited to a subset of vertices and edges. By minimizing the reproducing kernel Hilbert space norm while penalizing kernel complexity through maximum likelihood estimation, our method ensures that the resulting surrogate strictly enforces conservation laws without overfitting. We demonstrate our method on two representative applications: subsurface flow in fracture networks and arterial blood flow. Finally, the results demonstrate that the method maintains high accuracy and well-calibrated uncertainty estimates even under severe data scarcity, highlighting its potential for scientific applications where limited data and reliable uncertainty quantification are critical.

Dirichlet-to-Neumann map

Elliptically-Contoured Tensor-variate Distributions with Application to Image Learning

Statistical analysis of tensor-valued data has largely used the tensor-variate normal (TVN) distribution that may be inadequate for data arising from distributions with heavier or lighter tails. We study a general family of elliptically contoured (EC) TV distributions and derive its characterizations, moments, marginal, and conditional distributions. We describe procedures for maximum likelihood estimation from data that are (1) uncorrelated draws from an EC distribution, (2) from a scale mixture of the TVN distribution, and (3) from an underlying but unknown EC distribution, for which we extend Tyler’s robust estimator. A detailed simulation study highlights the benefits of choosing an EC distribution over the TVN for heavier-tailed data. We develop TV classification rules using discriminant analysis and EC errors and show that they better predict cats and dogs from images in the Animal Faces-HQ dataset than the TVN-based rules. A novel tensor-on-tensor regression and TV analysis of variance (TANOVA) framework under EC errors is also demonstrated to better characterize gender, age, and ethnic origin than the usual TVN-based TANOVA in the celebrated labeled faces of the wild dataset.

97 MATHEMATICS AND COMPUTING

Poisson-response Tensor-on-Tensor Regression and Applications

We introduce Poisson-response tensor-on-tensor regression (PToTR), a novel regression framework designed to handle tensor responses composed element-wise of random Poisson-distributed counts. Tensors, or multi-dimensional arrays, composed of counts are common data in fields such as inter national relations, social networks, epidemiology, and medical imaging, where events occur across multiple dimensions like time, location, and dyads. PToTR accommodates such tensor responses alongside tensor covariates, providing a versatile tool for multi dimensional data analysis. We propose algorithms for maximum likelihood estimation under a canonical polyadic (CP) structure on the regression coefficient tensor that satisfy the positivity of Poisson parameters and then provide an initial theoretical error analysis for PToTR estimators. We also demonstrate the utility of PToTR through three concrete applications: longitudinal data analysis of the Integrated Crisis Early Warning System database, positron emission tomography (PET) image reconstruction, and change-point detection of communication patterns in longitudinal dyadic data. These applications highlight the versatility of PToTR in addressing complex, structured count data across various domains.

97 MATHEMATICS AND COMPUTING

ASCR Workshop Position Paper: Challenges and Opportunities in High Energy Physics

High energy particle physics and cosmology concern themselves with estimating fundamental parameters of nature, such as the masses and interactions of fundamental particles like the Higgs boson and the rate of expansion of the universe. In doing so, they analyze exabyte-scale datasets, some of the largest in all of science, and face many challenges in subsequent data analysis. These challenges are shared between the two disciplines, but we focus on particle physics to highlight one specific domain. In particle physics, the standard method for estimating parameters involves performing Monte Carlo (MC) integration as a function of both parameters of interest and nuisance parameters using an expensive simulator, counting the number of observed collision events (i.i.d. samples) from an experiment in the corresponding integration domains, and forming a Poisson likelihood function. This likelihood function is then used in a Frequentist manner to construct a maximum likelihood point estimate (MLE) and confidence set for the parameters. To sufficiently populate the high-dimensional integration domains, simulators consume billions of CPU-hours annually and produce hundreds of petabytes of intermediate output data. Several techniques have been developed to: optimize definitions of the integration domains so as to be maximally sensitive to a particular subset of parameters, efficiently estimate the integrals, and build robust surrogate models by interpolating between integral evaluations at different parameter points. One can view this whole endeavor as classical Simulation-Based Inference (SBI).

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Using PyBioNetFit to leverage qualitative and quantitative data in biological model parameterization and uncertainty quantification

Data generated in studies of cellular regulatory systems are often qualitative. For example, measurements of signaling readouts in the presence and absence of mutations may reveal a rank ordering of responses across conditions but not the precise extents of mutation-induced differences. Qualitative data are often ignored by mathematical modelers or are considered in an ad hoc manner, as in the study of Kocieniewski and Lipniacki (2013) [Phys Biol 10: 035006], which was focused on the roles of MEK isoforms in ERK activation. In this earlier study, model parameter values were tuned manually to obtain consistency with a combination of qualitative and quantitative data. This approach is not reproducible, nor does it provide insights into parametric or prediction uncertainties. Here, starting from the same data and the same ordinary differential equation (ODE) model structure, we generate formalized statements of qualitative observations, making these observations more reusable, and we improve the model parameterization procedure by applying a systematic and automated approach enabled by the software package PyBioNetFit. We also demonstrate uncertainty quantification (UQ), which was absent in the original study. Our results show that PyBioNetFit enables qualitative data to be leveraged, together with quantitative data, in parameterization of systems biology models and facilitates UQ. These capabilities are important for reliable estimation of model parameters and model analyses in studies of cellular regulatory systems and reproducibility.

59 BASIC BIOLOGICAL SCIENCES

Open World Dempster-Shafer Theory/The Transferable Belief Model with Intervals: A Practitioner's Guide to DST and TBM

Dempster-Shafer theory (DST) is a mathematical framework that allows for uncertainty or ignorance to be quantified and included when making predictions from evidence. This is in contrast to Bayesian theory, which does not allow for any quantification of ignorance. The framework is described in great detail in [7]. DST is particularly useful for problems where the inclusion of additional evidence (for example, data from another sensor) could lead to a different conclusion. Thus, it is a useful data fusion method, especially in applications not suited to maximum likelihood or maximum a posteriori estimations due to limited samples or incomplete prior knowledge.

97 MATHEMATICS AND COMPUTING

BICEP/Keck. XX. Component-separated Maps of the Polarized Cosmic Microwave Background and Thermal Dust Emission Using Planck and BICEP/Keck Observations through the 2018 Observing Season

We present component-separated polarization maps of the cosmic microwave background (CMB) and Galactic thermal dust emission, derived using data from the BICEP/Keck experiments through the 2018 observing season and Planck. By employing a maximum-likelihood method that utilizes observing matrices, we produce unbiased maps of the CMB and dust signals. We outline the computational challenges and demonstrate an efficient implementation of the component map estimator. We show methods to compute and characterize power spectra of these maps, opening up an alternative way to infer the tensor-to-scalar ratio from our data. We compare the results of this map-based separation method with the baseline BICEP/Keck analysis. Our analysis demonstrates consistency between the two methods, finding an 84% correlation between the pipelines.

cosmic inflation

A Simulator for Neyer Tests of Explosives

Explosives and explosive devices such as detonators are typically tested by applying a range of stimuli such as voltage or mechanical shock, and recording binary “detonated/did not detonate” responses. These are analyzed using maximum likelihood or generalized linear models to provide estimates of quantities such as the all-fire and no-fire points. Given that the true threshold for detonation is unknown a priori , sequential design methods are typically used to optimize the set of test points. One popular method, implemented in commercial software, is Neyer’s algorithm. To support simulation and experimental design, we have developed code in the R programming language to duplicate the functions of the Neyer software. We provide code for the simulator along with a description and examples of usage.

42 ENGINEERING

Enhancing Gaussian Process Surrogates for Optimization and Posterior Approximation via Random Exploration

This paper proposes novel noise-free Bayesian optimization strategies that rely on a random exploration step to enhance the accuracy of Gaussian process surrogate models. The new algorithms retain the ease of implementation of the classical GP-UCB algorithm, but the additional random exploration step accelerates their convergence, nearly achieving the optimal convergence rate. Furthermore, to facilitate Bayesian inference with intractable likelihoods, we propose to utilize optimization iterates for maximum a posteriori estimation to build a Gaussian process surrogate model for the unnormalized log-posterior density. We provide bounds for the Hellinger distance between the true and the approximate posterior distributions in terms of the number of design points. We demonstrate the effectiveness of our Bayesian optimization algorithms in nonconvex benchmark objective functions, in a machine learning hyperparameter tuning problem, and in a black-box engineering design problem. The effectiveness of our posterior approximation approach is demonstrated in two Bayesian inference problems for parameters of dynamical systems.

Bayesian inference

An iterative CMB lensing estimator minimizing instrumental noise bias

Noise maps from cosmic microwave background (CMB) experiments are generally statistically anisotropic, due to scanning strategies, atmospheric conditions, or instrumental effects. Any mismodeling of this complex noise can bias the reconstruction of the lensing potential and the measurement of the lensing power spectrum from the observed CMB maps. We introduce a new CMB lensing estimator based on the maximum (MAP) reconstruction that is minimally sensitive to these instrumental noise biases. By modifying the likelihood to rely exclusively on correlations between CMB map splits with independent noise realizations, we minimize autocorrelations that contribute to biases. In the regime of many independent splits, this maximum closely approximates the optimal MAP reconstruction of the lensing potential. In simulations, we demonstrate that this method is able to determine lensing observables that are immune to any noise mismodeling with a negligible cost in signal-to-noise ratio. Our estimator enables unbiased and nearly optimal lensing reconstruction for next-generation CMB surveys.

Legrand, Louis [University of Cambridge (United Ki

A Measurement of the Largest-scale CMB E -mode Polarization with CLASS

We present measurements of large-scale cosmic microwave background E-mode polarization from the Cosmology Large Angular Scale Surveyor 90 GHz data. Using 115 det-yr of observations collected through 2024 with a variable-delay polarization modulator, we achieved a polarization sensitivity of 82 μK arcimin, comparable to Planck at similar frequencies (100 and 143 GHz ). The analysis demonstrates effective mitigation of systematic errors and addresses challenges to large-angular-scale power recovery posed by time-domain filtering in maximum-likelihood map-making. A novel implementation of the pixel-space transfer matrix is introduced, which enables efficient filtering simulations and bias correction in the power spectrum using the quadratic cross-spectrum estimator. Overall, we achieved an unbiased time-domain filtering correction to recover the largest angular scale polarization, with the only power deficit, arising from map-making nonlinearity, being characterized as <3%. Through cross-correlation with Planck, we detected the cosmic reionization at 99.4% significance and measured the reionization optical depth τ = $0.053^{+0.018}_{-0.019}$, marking the first ground-based attempt at such a measurement. At intermediate angular scales (ℓ > 30), our results, both independently and in cross-correlation with Planck, remain fully consistent with Planck’s measurements.

79 ASTRONOMY AND ASTROPHYSICS

Measurements of W + W − production cross-sections in pp collisions at $\sqrt{s}=13$ TeV with the ATLAS detector

Measurements of W + W − → e ± νμ ∓ ν production cross-sections are presented, providing a test of the predictions of perturbative quantum chromodynamics and the electroweak theory. The measurements are based on data from pp collisions at $\sqrt{s}$ = 13 TeV recorded by the ATLAS detector at the Large Hadron Collider in 2015–2018, corresponding to an integrated luminosity of 140 fb −1 . The number of events due to top-quark pair production, the largest background, is reduced by rejecting events containing jets with b-hadron decays. An improved methodology for estimating the remaining top-quark background enables a precise measurement of W + W − cross-sections with no additional requirements on jets. The fiducial W + W − cross-section is determined in a maximum-likelihood fit with an uncertainty of 3.1%. The measurement is extrapolated to the full phase space, resulting in a total W + W − cross-section of 127 ± 4 pb. Differential cross-sections are measured as a function of twelve observables that comprehensively describe the kinematics of W + W − events. The measurements are compared with state-of-the-art theory calculations and excellent agreement with predictions is observed. A charge asymmetry in the lepton rapidity is observed as a function of the dilepton invariant mass, in agreement with the Standard Model expectation. A CP-odd observable is measured to be consistent with no CP violation. Limits on Standard Model effective field theory Wilson coefficients in the Warsaw basis are obtained from the differential cross-sections.

Accelerator Physics

A dendritic strontium river isoscape for fisheries applications in the Sacramento River basin, California, USA

Objective Understanding the origins and movements of fish is fundamental to effective conservation and fisheries management. Strontium isotope ratios ( 87 Sr/ 86 Sr) in otoliths provide a powerful tracer of natal origin and migratory pathways. However, existing 87 Sr/ 86 Sr isoscapes for the Sacramento River basin, an ecosystem that supports ecologically and economically important salmon populations, rely on discrete classification approaches that overlook unsampled habitats and do not incorporate spatial uncertainty. Our objective was to develop a continuous, network-explicit 87 Sr/ 86 Sr isoscape with quantified uncertainty to fill in data gaps and enable probabilistic assignments of fish origin and movement. Methods We used river water 87 Sr/ 86 Sr data from 106 sites (1997–2021) to develop spatial stream network models that use dendritic connectivity and watershed characteristics (lithology, bedrock age, and land cover) to predict river water 87 Sr/ 86 Sr throughout the basin. Models were fitted using maximum and restricted likelihood and were evaluated via Akaike’s information criterion and leave-one-out cross validation. We produced both historical (pre-dam) and present-day (below-dam) isoscapes, delineated uncertainty-informed isotopic ranges using k -means clustering, and applied a proof-of-concept Bayesian assignment to estimate natal origins and early rearing habitats for two endangered winter-run Chinook Salmon Oncorhynchus tshawytscha. Results Cross validation indicated strong performance of the 87 Sr/ 86 Sr model (leave-one-out cross validation: R 2 = 0.91; root mean square error = 0.0005). Uncertainty-informed clustering identified 19 isotopic “suites” (reaches with indistinguishable 87 Sr/ 86 Sr values) in present-day anadromous habitats and 25 suites in the historical network. Example natal and early rearing assignments included predictions that challenged expectations for juvenile salmon migration based on predicted river 87 Sr/ 86 Sr compositions. Conclusions This study developed a continuous, network-explicit 87 Sr/ 86 Sr isoscape that integrates existing river data to predict 87 Sr/ 86 Sr in unsampled reaches and the likely achievable range and resolution of otolith-based origin and life history inference. The resulting river isoscape provides a valuable tool to predict salmon movements and identify habitats supporting their survival and growth that otherwise might remain undetected. Coupling these predictions with complementary approaches that ground-truth juvenile presence (e.g., targeted fish surveys) represents an important step toward science-informed restoration and management of critical habitats throughout the Sacramento River basin.

Environmental sciences

CMB-S4: Foreground-cleaning Pipeline Comparison for Measuring Primordial Gravitational Waves

We compare multiple foreground-cleaning pipelines for estimating the tensor-to-scalar ratio, r, using simulated maps of the planned CMB-S4 experiment within the context of the South Pole Deep Patch. To evaluate robustness, we analyze bias and uncertainty on r across various foreground suites using map-based simulations. The foreground-cleaning methods include: a parametric maximum likelihood approach applied to auto- and cross-power spectra between frequency maps; a map-based parametric maximum-likelihood method; and a harmonic-space internal linear combination using frequency maps. We summarize the conceptual basis of each method to highlight their similarities and differences. To better probe the impact of foreground residuals, we implement an iterative internal delensing step, leveraging a map-based pipeline to generate a lensing B-mode template from the large aperture telescope frequency maps. Our results show that the performance of the three approaches is comparable for simple and intermediate-complexity foregrounds, with σ(r) ranging from 3–5 ×10 −4 . However, biases at the 1σ–2σ level appear when analyzing more complex forms of foreground emission. By extending the baseline pipelines to marginalize over foreground residuals, we demonstrate that contamination can be reduced to within statistical uncertainties, albeit with a pipeline-dependent impact on σ(r), which translates to a detection significance between 2σ and 4σ for an input value of r = 0.003. These findings suggest varying levels of maturity among the tested pipelines, with the auto- and cross-spectra-based approach demonstrating the best stability and overall performance. Moreover, given the extremely low noise levels, mutual validation of independent foreground-cleaning pipelines is essential to ensure the robustness of any potential detection.

astronomy data analysis

Alleviating prior dependencies for DESI DR1 clustering fits through reparameterization

Bayesian analyses of the full-shape clustering of Dark Energy Spectroscopic Instrument (DESI) Data Release 1 (DR1) exhibit prior-volume projection effects, whereby weakly constrained nuisance parameters of the Effective Field Theory of Large Scale Structure (EFTofLSS) shift marginalized cosmological posteriors away from the posterior maximum. We reanalyze DESI DR1 power spectrum multipoles using two complementary mitigation strategies: (i) nonlinear orthogonalization to decorrelate nuisance and cosmological parameter priors, and (ii) a fully reparameterization-invariant Jeffreys prior over all EFTofLSS coefficients, evaluated on-the-fly via closed-form Jacobians. Including data from DESI, Big-Bang Nuclesynthesis and a constraint on $n_{\mathrm{s}}$, baseline priors lead to multi-$σ$ projection in the Hubble parameter $H_{0}$ and dark energy equation of state parameters $w_{0}$ and $w_{a}$; the Jeffreys prior successfully recenters these posteriors to enclose the maximum a posteriori estimate within the 68% credible regions, demonstrating clear mitigation of projection effects for these late-time expansion parameters. A hybrid Jeffreys+baseline-Gaussian configuration controls residual over-broad tails in the physical cold dark matter density $ω_{\mathrm{c}}$ while preserving the volume correction, and is our favoured approach. We compare the credible intervals derived using our methodology to those obtained using Halo Occupation Distribution (HOD)-informed priors and to confidence intervals derived using frequentist profile likelihood analyses, finding agreement in both central values and degeneracy directions in the $w_{0}$--$w_{a}$ plane. This demonstrates that, once projection effects are properly controlled, we can make robust inferences about the late-time cosmological expansion independent of the statistical framework adopted.

Bonici, M. [Waterloo U.; Perimeter Inst. Theor. Ph

A Deep Look at the Ultra-Faint Milky Way Satellite Virgo III with Rubin Observatory Data Preview 2

We analyze the ultra-faint Milky Way satellite Virgo III using data from the Vera C. Rubin Observatory Data Preview 2 (DP2). Virgo III was observed in the Rubin "Cosmic Treasure Chest" (M49) First Look field, which contains 924 visits in the u,g,r,i bands comprising ~10.5hrs of exposure time with LSSTCam. These data are considerably deeper than the majority of DP2, with a $5σ$ limiting magnitude that approaches the expected 10-year depth of LSST (~25.2-26.5mag, depending on band). We report the morphological and stellar population parameters of Virgo III measured with the maximum-likelihood-based package ugali. The depth of the Rubin imaging yields more than a factor of four increase in the number of candidate member stars ($N_* = 114^{+11}_{-11}$) relative to the Virgo III discovery results ($N_* = 25^{+5}_{-4}$), enabling significantly more precise morphological constraints. Our best-fit parameters broadly agree with previous measurements, further confirming that Virgo III has properties that are consistent with an ultra-faint dwarf galaxy ($M_V = -2.72^{+0.49}_{-0.70}$; $r_{1/2} = 53^{+10}_{-8}$) located at a heliocentric distance of $D_\odot = 151^{+8}_{-8}$. We also demonstrate that the depth and photometric quality of the DP2 data are sufficient to separate metal-poor and metal-rich stars in color-color space. We further present period estimates for the three known RR Lyrae member stars derived from the DP2 forced photometry; we use theoretical Period-Luminosity-Metallicity (PLZ) and Period-Wesenheit-Metallicity (PWZ) relations to obtain independent distance estimates. We find that our period and distance estimates are broadly consistent with previous measurements for these RR Lyrae. These results demonstrate the power of LSST data for the discovery and characterization of ultra-faint dwarf galaxies and motivate future searches for new satellites across the southern sky.

Pai, Aashay [Chicago U.; Chicago U., KICP; SkAI, C