Engineering PapersSearch

DOE OSTI · 3009256

Optimizers for stabilizing likelihood-free inference

Abstract

A growing number of applications in particle physics and beyond use neural networks as unbinned likelihood ratio estimators applied to real or simulated data. Precision requirements on the inference tasks demand a high-level of stability from these networks, which are affected by the stochastic nature of training. We show how physics concepts can be used to stabilize network training through a physics-inspired optimizer. In particular, the energy conserving descent (ECD) optimization framework uses classical Hamiltonian dynamics on the space of network parameters to reduce the dependence on the initial conditions while also stabilizing the result near the minimum of the loss function. We develop a version of this optimizer known as , which has few free hyperparameters with limited ranges guided by physical reasoning. We apply to representative likelihood-ratio estimation tasks in particle physics and find on average that it out-performs the widely used Adam optimizer. We expect that ECD will be a useful tool for wide array of data-limited problems, where it is computationally expensive to exhaustively optimize hyperparameters and mitigate fluctuations with ensembling.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

De Luca, G Bruno, Nachman, Benjamin, Silverstein, Eva, Zheng, Henry. 2025-11-01. Optimizers for stabilizing likelihood-free inference. https://doi.org/10.1103/x88j-mv39

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related reports

First Wall Design of a Tokamak Pilot Plant Using a Monte Carlo Model for 3-D Heat Flux Deposition

We present a method for calculating the heat fluxes deposited on nonaxisymmetric tokamak first wall components, allowing for a first-of-its-kind model for power handling in the tokamak far scrape-off layer (SOL). The DIV3D Monte Carlo model features strict global power conservation and can calculate the finite cross-field plasma transport into magnetically-shadowed regions, which is significant when dealing with meter-scale shadows introduced by components such as poloidal limiters or antennas. As a case study, we apply the DIV3D model to inform the distribution of first wall poloidal limiters in an ARC-class reactor device. We demonstrate that discrete protection limiters can efficiently reduce peak heat fluxes on recessed breeder wall components in the presence of significant far-SOL plasma fluxes. By varying the toroidal periodicity and radial standoff depth of the limiters, we demonstrate one of the tradeoffs that must be considered in first wall design: more limiters provide greater protection, but at the cost of reduced breeding performance. We also present the impact that radial misalignments between limiters would have on first wall power loading.

Monte Carlo methods

Fokker-Planck Equation Governing the Distribution of Walkers in Auxiliary-Field Quantum Monte Carlo

Auxiliary-field quantum Monte Carlo (AFQMC) is typically formulated as an open-ended random walk in an overcomplete space of Slater determinants, implemented through a Langevin equation. However, the explicit form of the underlying Fokker-Planck equation governing the walker population distribution has remained unknown. Here, in this Letter, we derive the Fokker-Planck equation for AFQMC and propose a novel numerical scheme to solve it. The solution of the Fokker-Planck equation reveals the wave function actually sampled by the AFQMC algorithm. Interestingly, we find that even when the exact ground state is used as a guiding wave function in constrained path AFQMC, contrary to the common assumption, the wave function sampled by AFQMC is not exact. Beyond clarifying several fundamental aspects of AFQMC, the availability of a Fokker-Planck equation formulation opens new avenues for systematically improving its accuracy, which we outline in this Letter.

Monte Carlo methods

Codimension-two spiral spin liquid in the effective honeycomb-lattice compound Cs 3 ⁢Fe 2 ⁢Cl 9

A codimension-two spiral spin liquid is a correlated paramagnetic state with one-dimensional ground state degeneracy hosted within a three-dimensional lattice. Here, in this work, via neutron scattering experiments and numerical simulations, we establish the existence of a codimension-two spiral spin liquid in the effective honeycomb-lattice compound Cs 3 ⁢Fe 2⁢ Cl 9 , which demonstrates an alternate path to spiral spin liquids by overcoming the long-standing impediment of weak further-neighbor interactions. In the long-range ordered regime, competing spiral and spin density wave orders emerge as a function of applied magnetic field, among which a possible order-by-disorder transition is identified.

Monte Carlo methods