Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “hard constraint”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Hydroxylpyromorphite, a mineral important to lead remediation: Modern description and characterization

Abstract Hydroxylpyromorphite, Pb5(PO4)3(OH), has been documented in the literature as a synthetic and naturally occurring phase for some time but has not previously been formally described as a mineral. It is fully described here for the first time using crystals collected underground in the Copps mine, Gogebic County, Michigan. Hydroxylpyromorphite occurs as aggregates of randomly oriented hexagonal prisms, primarily between about 20–35 μm in length and 6–10 μm in diameter. The mineral is colorless and translucent with vitreous luster and white streak. The Mohs hardness is ~3½–4; the tenacity is brittle, the fracture is irregular, and indistinct cleavage was observed on {001}. Electron microprobe analyses provided the empirical formula Pb4.97(PO4)3(OH0.69F0.33Cl0.06)Σ1.08. The calculated density using the measured composition is 7.32 g/cm3. Powder X-ray diffraction data for the type material is compared to data previously reported for hydroxylpyromorphite from the talc mine at Rabenwald, Austria, and from Whytes Cleuch, Wanlockhead, Scotland. Hydroxylpyromorphite is hexagonal, P63/m, at 100 K, a = 9.7872(14), c = 7.3070(10) Å, V = 606.16(19) Å3, and Z = 2. The structure [R1 = 0.0181 for 494 F>4σ(F) reflections] reveals that hydroxylpyromorphite adopts a column anion arrangement distinct from other members of the apatite supergroup due to the presence of fluorine and steric constraints imposed by stereoactive lone-pair electrons of Pb2+ cations. The F– anion sites are displaced slightly from hydroxyl oxygen anions, which allows for stronger hydrogen-bonding interactions that may in turn stabilize the observed column-anion arrangement and overall structure. Our modern characterization of hydroxylpyromorphite provides deeper understanding to a mineral useful for remediation of lead-contaminated water.

Geochemistry & Geophysics↗

A cost comparison of various hourly-reliable and net-zero hydrogen production pathways in the United States

Hydrogen (H 2 ) as an energy carrier may play a role in various hard-to-abate subsectors, but to maximize emission reductions, supplied hydrogen must be reliable, low-emission, and low-cost. Here, we build a model that enables direct comparison of the cost of producing net-zero, hourly-reliable hydrogen from various pathways. To reach net-zero targets, we assume upstream and residual facility emissions are mitigated using negative emission technologies. For the United States (California, Texas, and New York), model results indicate next-decade hybrid electricity-based solutions are lower cost ($2.02-$2.88/kg) than fossil-based pathways with natural gas leakage greater than 4% ($2.73-$5.94/kg). These results also apply to regions outside of the U.S. with a similar climate and electric grid. However, when omitting the net-zero emission constraint and considering the U.S. regulatory environment, electricity-based production only achieves cost-competitiveness with fossil-based pathways if embodied emissions of electricity inputs are not counted under U.S. Tax Code Section 45V guidance.

08 HYDROGEN↗

High-energy Radiation and Ion Acceleration in Three-dimensional Relativistic Magnetic Reconnection with Strong Synchrotron Cooling

Abstract We present the results of 3D particle-in-cell simulations that explore relativistic magnetic reconnection in pair plasma with strong synchrotron cooling and a small mass fraction of nonradiating ions. Our results demonstrate that the structure of the current sheet is highly sensitive to the dynamic efficiency of radiative cooling. Specifically, stronger cooling leads to more significant compression of the plasma and magnetic field within the plasmoids. We demonstrate that ions can be efficiently accelerated to energies exceeding the plasma magnetization parameter, ≫ σ , and form a hard power-law energy distribution, f i ∝ γ −1 . This conclusion implies a highly efficient proton acceleration in the magnetospheres of young pulsars. Conversely, the energies of pairs are limited to either σ in the strong cooling regime or the radiation burnoff limit, γ syn , when cooling is weak. We find that the high-energy radiation from pairs above the synchrotron burnoff limit, ε c ≈ 16 MeV, is only efficiently produced in the strong cooling regime, γ syn < σ . In this regime, we find that the spectral cutoff scales as ε cut ≈ ε c ( σ / γ syn ) and the highest energy photons are beamed along the direction of the upstream magnetic field, consistent with the phenomenological models of gamma-ray emission from young pulsars. Furthermore, our results place constraints on the reconnection-driven models of gamma-ray flares in the Crab Nebula.

79 ASTRONOMY AND ASTROPHYSICS↗

Extraction of Beam-Spin Asymmetries from the Hard Exclusive π + Channel off Protons in a Wide Range of Kinematics

We have measured beam-spin asymmetries to extract the $\sin\phi$ moment $A_{LU}^{\sin\phi}$ from the hard exclusive $\vec{e} p \to e^\prime n \pi^+$ reaction above the resonance region, for the first time with nearly full coverage from forward to backward angles in the center-of-mass. The $A_{LU}^{\sin\phi}$ moment has been measured up to 6.6 GeV$^{2}$ in $-t$, covering the kinematic regimes of Generalized Parton Distributions (GPD) and baryon-to-meson Transition Distribution Amplitudes (TDA) at the same time. The experimental results in very forward kinematics demonstrate the sensitivity to chiral-odd and chiral-even GPDs. In very backward kinematics where the TDA framework is applicable, we found $A_{LU}^{\sin\phi}$ to be negative, while a sign change was observed near 90$^\circ$ in the center-of-mass. The unique results presented in this paper will provide critical constraints to establish reaction mechanisms that can help to further develop the GPD and TDA frameworks.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

On the Diffusivity of Moist Static Energy and Implications for the Polar Amplification Response to Climate Warming

Energy balance models (EBMs) have been widely used in a range of climate problems, but the assumption of constant diffusivity in the parameterization of the moist static energy (MSE) flux can be hardly justified. We demonstrate in this study that the diffusive MSE flux can be derived from the basic energy balance equation with a few tolerable assumptions. The estimated diffusivity is both spatially and seasonally dependent, and its midlatitude average is then tested against several scaling theories for the midlatitude eddy diffusivity. The result supports the diffusivity theory of Held and Larichev (1996) modified for the moist atmosphere, affording a dynamics-based parameterization of MSE diffusivity. The implementation of the parameterization in an EBM leads to an interactive MSE diffusivity that accounts for the midlatitude eddy response to climate forcing perturbations. Under a uniform radiative forcing, the EBM with a diffusivity so parameterized produces a weakening of the midlatitude diffusivity and a modestly polar-amplified surface temperature response as an inevitable outcome under the dual constraints of the nonlinear Clausius-Clapeyron relation and the temperature gradient-dependent diffusivity, even in the absence of any poleward amplifying radiative feedbacks. As the consequence of more isothermal temperature and reduced diffusivity, the variance of the midlatitude surface temperature also decreases with warming.

54 ENVIRONMENTAL SCIENCES↗

Accelerating Hanford Site Cleanup through Operations Research Modeling - 20238

The Hanford Site cleanup effort will require the integration of dozens of unique facilities and processes, many of which will be first-of-a-kind in implementation and design. Each facility will be governed by its own set of operating logic, configured with a unique array of unit operations, and subject to a set of constraints that will affect its behavior. The collection of facilities have multiple points of interface, making the operations of any one facility potentially significant to the operations of other up- or downstream processes. It is therefore highly desirable to accurately predict these operations, as it allows for Site officials to identify and preempt bottlenecks and vulnerabilities before they unexpectedly inhibit the cleanup mission. With the quantity and complexity of the processes that will be on Site, building a pen-and-paper or even a spreadsheet-assisted model of the cleanup mission quickly becomes overwhelming in scope and inaccurate in execution. The Engineering organization for the Site's Tank Operations Contract (TOC) has therefore implemented the use of operations research (OR) modeling to simulate and predict future operations of Site facilities. These models are created using a discrete event simulation tool that allows for the development of detailed, versatile, and robust models. Not only can these models account for complex logical behaviors, but they can also simulate process details down to the level of vessel sizing, labor utilization, equipment reliability, and resource availability. To date, the TOC has developed OR models for several facilities on Site, including for single-shell tank (SST) farms, double-shell tank (DST) farms, the Effluent Treatment Facility (ETF), and the waste transfer system. These models have focused on identifying bottlenecks and operational constraints, and have been used to quantify the effects of implementing process changes. This latter point is particularly valuable, as it allows for several alternatives to be studied in a virtual setting before committing resources to making a change in the field. The decision to develop OR models has gained tremendous support from the Site's stakeholders and the U.S. Department of Energy (DOE) management, and has prompted the use of the tool to support additional internal and external initiatives. Recently, an initiative was proposed to use the models to help identify and provide quantitative backing for risks and opportunities for the TOC. This application of OR could not only help inform how the TOC manages its risks (e.g. quantities and types of spare parts), but could also help drive process improvements whose benefits might otherwise be hard to quantify. The models have also been used to drive the TOC's cloud computing, artificial intelligence (AI), and machine learning (ML) initiatives. These initiatives will not only improve the ability of the TOC to more rapidly respond to the needs of its customers, but it will also aid in the ability of the TOC to analyze and improve the processes it studies. Partnership with two external software development and consulting companies (Lanner and Ynformed) has furthered not only the application of AI and ML within the TOC, but has also spurred the development of new/improved software tools and platforms used by the companies. These partnerships have proven to be mutually beneficial and productive, and have set a precedent for the types of gains that can be made by exploring such options. (authors)

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Faster-than-real-time Simulation with Demonstration for Resilient DER Integration

The US electric grid is facing operational, stability, and security challenges. Transmission system operators need some measure of visibility into distribution system renewable generation. Distribution system generation needs to support transmission system voltage. The grid is experiencing an expansion in measurement systems. How to take full advantage of this expansion and defend against attacks, both cyber and physical, poses additional challenges. The Faster-than-real-time Simulation with demonstration for Resilient DER Integration project set out to do the following: a. Flatten the voltage profile through the feeders and system for cost saving and voltage stabilization needs. b. Increase the amount of intermittent distributed energy resources (IDERs) that could be deployed on a utility feeder and provide 100% or more energy needed for the demands on that feeder, and c. based on an accurate model (Digital Twin) of the utilities system, be able to detect any abnormalities on the utilities distribution system. To manage the voltage and increase IDER penetration (a,b), Graph Trace Analysis is employed in a time-series, optimal power flow to coordinate the time-varying feedback control setpoints of a distribution feeder’s utility control devices. Under the coordinated control are a Load Tap Changing Transformer, a voltage regulator, and five switched capacitor banks. The feeder serves over 2000 customers, the feeder secondaries are modeled, and the feeder has 2.3 MW of PV generation, corresponding to a 17.4% penetration of PV generation. The feeder model has over 12,000 components, where every customer load bus and PV generator are modeled. The accuracy of the power flow solution is compared against historical meter voltage measurements, the improvement in conservation voltage reduction energy savings as a function of the coordinated control desired voltage profile is investigated, and the increase in PV penetration of the coordinated control over the existing control is presented. To achieve improved control performance while observing system operation constraints, bellwether Advanced Metering Infrastructure (AMI) voltage measurements are used to adjust the desired voltage profile used by the optimal power flow analysis. To detect and alleviate or negate attacks or failures on the distribution and transmission utility grids (c) the grid needs to be resilient and self-healing. In this project software was designed to do just that. At the center of the software is an Integrated System Model (ISM) that spans from transmission to secondary distribution. The ISM is employed in real-time abnormality detection, voltage stability forecasting, and multi-mode control. Testing results are presented for: 1—attacks on utility infrastructure; 2—energy savings from optimal control; 3—distribution system control response during a low voltage transmission system event; 4—cyber-attacks on PV inverters, where physical inverters are used in hard-ware-in-the-simulation-loop studies. Contributions of this work include real-time analysis that spans from three-phase transmission through secondary distribution; an approach for detecting abnormalities that employs measurements from three independent measurement systems; and a multi-mode distribution system control that responds to cyber-attacks, physical attacks, equipment failures, and transmission system needs.

Integrated System Model, Graph Trace Analysis, Adv↗

An optimized search for dark matter in the galactic halo with HAWC

The Galactic Halo is a key target for indirect dark matter detection. The High Altitude Water Cherenkov (HAWC) observatory is a high-energy (~300 GeV to >100 TeV) gamma-ray detector located in central Mexico. HAWC operates via the water Cherenkov technique and has both a wide field of view of ~ 2 sr and a >95% duty cycle, making it ideal for analyses of highly extended sources. We made use of these properties of HAWC and a new background-estimation technique optimized for extended sources to probe a large region of the Galactic Halo for dark matter signals. With this approach, we set improved constraints on dark matter annihilation and decay between masses of 10 and 100 TeV. Due to the large spatial extent of the HAWC field of view, these constraints are robust against uncertainties in the Galactic dark matter spatial profile.

79 ASTRONOMY AND ASTROPHYSICS↗

Enrichment by the first stars in a relic dwarf galaxy

Stars that contain trace amounts of elements heavier than helium (that is, ‘metallicity’) preserve the chemical fingerprints of the first stars. In the Milky Way, nearly all of the lowest-metallicity stars show an extreme over-abundance of carbon. The origin of this signature has remained a mystery owing to the lack of observational constraints on the environments in which it originates. Here, in this work, we present observations of a star in the >10-billion-year-old ultrafaint dwarf galaxy Pictor II, showing the lowest iron and calcium abundances outside the Milky Way (<1/43,000th solar and ~1/160,000th solar), with >3,000× relative carbon enhancement. The star’s exceptional paucity in iron and calcium make it clearly preserve enrichment from the first stars in a relic dwarf galaxy; Pictor II is one of the smallest, most chemically primordial systems known. This star supports the hypothesis that extreme carbon enhancement results from low-energy supernovae from the first stars, as the yields of energetic supernovae are harder to retain in small-scale environments. This signature of enrichment by the first stars may trace a regime inaccessible to current high-redshift observations, which can hardly detect the initial enrichment of the smallest galaxies.

chemical evolution↗

QoS-aware edge AI placement and scheduling with multiple implementations in FaaS-based edge computing

Resource constraints on the computing continuum require that we make smart decisions for serving AI-based services at the network edge. AI-based services typically have multiple implementations (e.g., image classification implementations include SqueezeNet, DenseNet, and others) with varying trade-offs (e.g., latency and accuracy). The question then is how should AI-based services be placed across Function-as-a-Service (FaaS) based edge computing systems in order to maximize total Quality-of-Service (QoS). To address this question, we propose a problem that jointly aims to solve (i) edge AI service placement and (ii) request scheduling. These are done across two time-scales (one for placement and one for scheduling). Here we first cast the problem as an integer linear program. We then decompose the problem into separate placement and scheduling subproblems and prove that both are NP-hard. We then propose a novel placement algorithm that places services while considering device-to-device communication across edge clouds to offload requests to one another. Our results show that the proposed placement algorithm is able to outperform a state-of-the-art placement algorithm for AI-based services, and other baseline heuristics, with regard to maximizing total QoS. Additionally, we present a federated learning-based framework, FLIES, to predict the future incoming service requests and their QoS requirements. Our results also show that our FLIES algorithm is able to outperform a standard decentralized learning baseline for predicting incoming requests and show comparable predictive performance when compared to centralized training.

97 MATHEMATICS AND COMPUTING↗

Progress in Constraining Nuclear Symmetry Energy Using Neutron Star Observables Since GW170817

The density dependence of nuclear symmetry energy is among the most uncertain parts of the Equation of State (EOS) of dense neutron-rich nuclear matter. It is currently poorly known especially at suprasaturation densities partially because of our poor knowledge about isovector nuclear interactions at short distances. Because of its broad impacts on many interesting issues, pinning down the density dependence of nuclear symmetry energy has been a longstanding and shared goal of both astrophysics and nuclear physics. New observational data of neutron stars including their masses, radii, and tidal deformations since GW170817 have helped improve our knowledge about nuclear symmetry energy, especially at high densities. Based on various model analyses of these new data by many people in the nuclear astrophysics community, while our brief review might be incomplete and biased unintentionally, we learned in particular the following: (1) The slope parameter L of nuclear symmetry energy at saturation density ρ0 of nuclear matter from 24 new analyses of neutron star observables was about L≈57.7±19 MeV at a 68% confidence level, consistent with its fiducial value from surveys of over 50 earlier analyses of both terrestrial and astrophysical data within error bars. (2) The curvature Ksym of nuclear symmetry energy at ρ0 from 16 new analyses of neutron star observables was about Ksym≈−107±88 MeV at a 68% confidence level, in very good agreement with the systematics of earlier analyses. (3) The magnitude of nuclear symmetry energy at 2ρ0, i.e., Esym(2ρ0)≈51±13 MeV at a 68% confidence level, was extracted from nine new analyses of neutron star observables, consistent with the results from earlier analyses of heavy-ion reactions and the latest predictions of the state-of-the-art nuclear many-body theories. (4) While the available data from canonical neutron stars did not provide tight constraints on nuclear symmetry energy at densities above about 2ρ0, the lower radius boundary R2.01=12.2 km from NICER’s very recent observation of PSR J0740+6620 of mass 2.08±0.07M⊙ and radius R=12.2–16.3 km at a 68% confidence level set a tight lower limit for nuclear symmetry energy at densities above 2ρ0. (5) Bayesian inferences of nuclear symmetry energy using models encapsulating a first-order hadron–quark phase transition from observables of canonical neutron stars indicated that the phase transition shifted appreciably both L and Ksym to higher values, but with larger uncertainties compared to analyses assuming no such phase transition. (6) The high-density behavior of nuclear symmetry energy significantly affected the minimum frequency necessary to rotationally support GW190814’s secondary component of mass (2.50–2.67) M⊙ as the fastest and most massive pulsar discovered so far. Overall, thanks to the hard work of many people in the astrophysics and nuclear physics community, new data of neutron star observations since the discovery of GW170817 have significantly enriched our knowledge about the symmetry energy of dense neutron-rich nuclear matter.

79 ASTRONOMY AND ASTROPHYSICS↗

Fermi-LAT Detection of the Supernova Remnant G312.4-0.4 in the Vicinity of 4FGL J1409.1-6121e

Abstract Gamma-ray emission provides constraints on the nonthermal radiation processes at play in astrophysical particle accelerators. This allows both the nature of accelerated particles and the maximum energy that they can reach to be determined. Notably, it remains an open question to what extent supernova remnants contribute to the sea of Galactic cosmic rays. In the Galactic plane, at around 312° of Galactic longitude, Fermi-LAT observations show an extended source (4FGL J1409.1−6121e) around five powerful pulsars. This source is described by one large disk of 0.°7 radius with a high significance of 45σin the 4FGL-DR3 catalog. Using 14 yr of Fermi-LAT observations, we revisited this region with a detailed spectro-morphological analysis in order to disentangle its underlying structure. Three sources have been distinguished, including the supernova remnant G312.4−0.4 whose gamma-ray emission correlates well with the shell observed at radio energies. The hard spectrum detected by the LAT, extending up to 100 GeV without any sign of a cutoff, is well reproduced by a purely hadronic model.

Astronomy & Astrophysics↗

Reconfigurable Framework for Resilient Semantic Segmentation for Space Applications

Deep learning (DL) presents new opportunities for enabling spacecraft autonomy, onboard analysis, and intelligent applications for space missions. However, DL applications are computationally intensive and often infeasible to deploy on radiation-hardened (rad-hard) processors, which traditionally harness a fraction of the computational capability of their commercial-off-the-shelf counterparts. Commercial FPGAs and system-on-chips present numerous architectural advantages and provide the computation capabilities to enable onboard DL applications; however, these devices are highly susceptible to radiation-induced single-event effects (SEEs) that can degrade the dependability of DL applications. In this article, we propose Reconfigurable ConvNet (RECON), a reconfigurable acceleration framework for dependable, high-performance semantic segmentation for space applications. In RECON, we propose both selective and adaptive approaches to enable efficient SEE mitigation. In our selective approach, control-flow parts are selectively protected by triple-modular redundancy to minimize SEE-induced hangs, and in our adaptive approach, partial reconfiguration is used to adapt the mitigation of dataflow parts in response to a dynamic radiation environment. Combined, both approaches enable RECON to maximize system performability subject to mission availability constraints. We perform fault injection and neutron irradiation to observe the susceptibility of RECON and use dependability modeling to evaluate RECON in various orbital case studies to demonstrate a 1.5–3.0× performability improvement in both performance and energy efficiency compared to static approaches.

97 MATHEMATICS AND COMPUTING↗

High-yield implosion modeling using the Frustraum: Assessing and controlling the formation of polar jets and enhancing implosion performance with applied magnetization

Frustraums have a higher laser-to-capsule x-ray radiation coupling efficiency and can accommodate a large capsule, thus potentially generating a higher yield with less laser energy than cylindrical Hohlraums for a given Hohlraum volume [Amendt et al., Phys. Plasmas 26, 082707 (2019]. Frustraums are expected to have less m = 4 azimuthal asymmetries arising from the intrinsic inner-laser-beam geometry on the National Ignition Facility. An experimental campaign at Lawrence Livermore National Laboratory to demonstrate the high-coupling efficiency and radiation symmetry tuning of the Frustraum has been under way since 2021. Simulations benchmarked against experimental data show that implosions using Frustraums can achieve more yield with higher ignition margins than cylindrical Hohlraums using the same laser energy. Hydrodynamic jets in capsules along the Hohlraum axis, driven by radiation-flux asymmetries in a Hohlraum with a gold liner on a depleted uranium (DU) wall, are present around stagnation, and these “polar” jets can cause severe yield degradation. The early-time Legendre mode P4<0 radiation-flux asymmetry is a leading cause of these jets, which can be reduced by using an unlined DU Hohlraum because the shape of the shell is predicted to be more prolate. Magnetization can increase the implosion robustness and reduce the required hotspot ρR for ignition; therefore, magnetizing the Frustraum can maintain the same yield while reducing the required laser energy or increase the yield using the same laser energy—all under the constraint that the ignition margin is preserved. Reducing polar jets is particularly important for magnetized implosions because of the intrinsic toroidal hotspot ion temperature topology.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Integral equation theory of thermodynamics, pair structure, and growing static length scale in metastable hard sphere and Weeks-Chandler-Andersen fluids

Here, we employ the Ornstein-Zernike integral equation theory with the Percus-Yevick (PY) and modified-Verlet (MV) closures to study the equilibrium structural and thermodynamic properties of metastable monodisperse hard sphere and continuous repulsion Weeks-Chandler-Andersen (WCA) fluids under density and temperature conditions where the system is strongly overcompressed or supercooled, respectively. The theoretical results are compared to crystal-avoiding simulations of these dense monodisperse model one-component fluids. The equation of state (EOS) and dimensionless compressibility are computed using both the virial and compressibility routes. For hard spheres, the MV-based virial route EOS and dimensionless compressibility are in very good agreement with simulation for all packing fractions, much better than the PY analogs. The corresponding MV-based predictions for the static structure factor are also very good. The amplitude of density fluctuations on the local cage scale and in the long wavelength limit, and three technically different measures of the density correlation length, are studied with both closures. All five properties grow in a roughly exponential manner with density in the metastable regime up to packing fractions of 58% with no sign of saturation. The MV-based results are in good agreement with our crystal-avoiding simulations. Interestingly, the density dependences of long and short wavelength quantities are closely related. The MV-based theory is also quite accurate for the thermodynamics and structure of supercooled monodisperse WCA fluids. Overall our findings are also relevant as critical input to microscopic theories that relate the equilibrium pair correlation function or static structure factor to dynamical constraints, barriers, and activated relaxation in glass-forming liquids.

36 MATERIALS SCIENCE↗

Global constraint on the jet transport coefficient from single-hadron, dihadron, and γ -hadron spectra in high-energy heavy-ion collisions

Modifications of large transverse momentum single-hadron, dihadron, and γ -hadron spectra in relativistic heavy-ion collisions are direct consequences of parton-medium interactions in the quark-gluon plasma (QGP). The interaction strength and underlying dynamics can be quantified by the jet transport coefficient q ̂ . We carry out the first global constraint on q ̂ using a next-to-leading order pQCD parton model with higher-twist parton energy loss and combining world experimental data on single-hadron, dihadron, and γ -hadron suppression at both RHIC and LHC energies with a wide range of centralities. The global Bayesian analysis using the information field (IF) priors provides the most stringent constraint on q ̂ ( T ) . We demonstrate in particular the progressive constraining power of the IF Bayesian analysis on the strong temperature dependence of q ̂ using data from different centralities and colliding energies. We also discuss the advantage of using both inclusive and correlation observables with different geometric biases. As a verification, the obtained q ̂ ( T ) is shown to describe data on single-hadron anisotropy at high transverse momentum well. Predictions for future jet quenching measurements in oxygen-oxygen collisions are also provided. Published by the American Physical Society 2024

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

AI Model Benchmarking for Nonproliferation Applications: Steel Thread Benchmarking Task Force Technical Report (Rev. 2)

Steel Thread is a NA-22 venture that seeks to build trustworthy, reliable AI models that can be used in a wide variety of nonproliferation tasks. A key aspect of building these models is developing appropriate benchmarks and evaluation methods, which will enable the venture to identify and adapt models to provide the most value in the nonproliferation domain. Benchmarks must be relevant to key tasks in this domain, such as question answering, information retrieval, document summarization and classification, consensus analysis, and image and data analysis. This report 1) provides an overview of benchmark design, evaluation, and challenges; 2) reviews a variety of open benchmarks, with a focus on language models and tasks; and 3) identifies benchmarks that are most relevant to Steel Thread. This report is intended to serve as a basis for further efforts to classify and evaluate benchmarks and their correlation with success on nonproliferation-specific tasks. The Steel Thread venture has defined benchmarks to be a particular combination of a dataset (or datasets) and a metric (or metrics) conceptualized as representing one or more specific tasks or sets of abilities for a specific modality. It is adopted by a research community as a shared framework for comparing methods.1 It includes 1) Data: Labeled (a designated subset not used for training, which could be all the data), 2) Metric: A way to quantify performance, 3) Task/Ability: What the benchmark is testing, 4) Protocol: A structured and repeatable evaluation process, 5) Baseline/Reference Model: For comparison; could be statistical, rule-based, SME-derived, or another model, and 6) Maintenance Plan: to update with new information over time; important for long-term utility. For further clarity, the definition includes what a benchmark, in this context, is not. It is not a corpus of training data, specific to a model (it is intended to apply to a range of models), a universal evaluation of performance, a guarantee that the ‘top’ model on the leaderboard will be the best fit for every specific use case, an all-encompassing proof of a model’s universal quality, nor is it a one-size-fits-all measure of success. It does not cover every real-world constraint (like operational, ethical, or cost considerations), a systems integration test, or a unit test. This definition was inspired by and resulted from discussions within the Steel Thread Benchmarking Task Force. This group was formed to define what we would mean as a benchmark within Steel Thread but persisted as the need to develop a thorough understanding of the large and expanding existing benchmarking space. This technical report is a result of the group’s divide and conquer approach to exploring this space. The release of benchmarks might not be progressing as quickly as model development, but it is moving very fast, as many benchmarks quickly become saturated, when state-of-the-art models score so close to the benchmark’s ceiling that their results are virtually indistinguishable. At that point, the test no longer differentiates between new systems, so researchers usually stop reporting scores as the benchmark no longer informs about improvements from the next generation of models. In the OpenAI announcement of GPT-5, they reported results on six flagship public benchmarks (AIME 2025, SWE-bench Verified, Aider Polyglot, MMMU, HealthBench Hard, GPQA) but the full system-card covers roughly thirty-five separate evaluations, comprising hundreds of test task items in total. There have been some efforts to summarize benchmarks in specific fields, like for text-to-image generation, but these surveys have had a narrow methodology scope. Therefore, a comprehensive survey of all benchmarks or even all benchmarks that could be relevant to Steel Thread is outside of the scope of this report. We chose some specific benchmarks to investigate in detail.

97 MATHEMATICS AND COMPUTING↗

Modified gravity constraints from the full shape modeling of clustering measurements from DESI 2024

We present cosmological constraints on deviations from general relativity (GR) from the first-year of clustering observations from the Dark Energy Spectroscopic Instrument (DESI) in combination with other available datasets including the CMB data from Planck with CMB-lensing from Planck and ACT, BBN constraints on the physical baryon density, the galaxy weak lensing and clustering from DESY3 and supernova data from DESY5. We first consider the μ(a,k)–Σ(a,k) modified gravity (MG) parameterization (as well as η(a,k)) in a ΛCDM and a w 0 w a CDM cosmological backgrounds. Using a functional form for time-only evolution gives μ 0 = 0.11 +0.44 -0.54 from DESI(FS+BAO)+BBN and a wide prior on n s . Using DESI(FS+BAO)+CMB+DESY3+DESY5-SN, we obtain μ 0 = 0.05 ± 0.22 and Σ 0 = 0.008 ± 0.045 and similarly μ 0 = 0.02 +0.19 -0.24 and η 0 = 0.09 +0.36 -0.60 , in an ΛCDM background. In w 0 w a CDM we obtain μ 0 = -0.24 +0.32 -0.28 and Σ 0 = 0.006 ± 0.043, consistent with GR, and we still find a preference of the data for a dynamical dark energy with w 0 > -1 and w a < 0. Using functional dependencies in both time and scale gives μ 0 and Σ 0 with a same level of precision as above but other scale MG parameters remain hard to constrain. We then move to binned parameterizations in a ΛCDM background starting with two bins in redshift and obtain, μ 1 = 1.02 ± 0.13, μ 2 = 1.04 ± 0.11, Σ 1 = 1.021 ± 0.029 and Σ 2 = 1.022 +0.027 -0.023 , all consistent with the unity value of GR in the binning formalism. We then extend the analysis to combine two bins in redshift and two in scale giving 8 MG parameters that we find all consistent with GR. We note that we find here that the tension reported in previous studies about Σ 0 being inconsistent with GR when using Planck PR3 data goes away when we use the recent LoLLiPoP+HiLLiPoP likelihoods. As noted in previous studies, this seems to indicate that the tension is indeed related to the CMB lensing anomaly in PR3 which is also resolved when using the recent likelihoods. We then constrain the class of Horndeski theory in the effective field theory of dark energy approach. We consider both EFT-basis and α-basis in the analysis. Assuming a power law parameterization for the EFT function Ω, which controls non-minimal coupling, we obtain Ω 0 = 0.012 +0.001 -0.012 and s 0 = 0.996 +0.54 -0.20 from the combination of DESI(FS+BAO)+DESY5SN+CMB in a ΛCDM background, which are consistent with GR. Similar results are obtained when using the α-basis and assuming no-braiding (α B = 0) giving c M < 1.14 at 95% CL in a ΛCDM background, also in agreement with GR. However, we see a mild yet consistent indication for c B > 0 when α B is allowed to vary which will require further study to determine whether this is due to systematics or new physics.

79 ASTRONOMY AND ASTROPHYSICS↗