Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Error Budget”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Harmonizing direct and indirect anthropogenic land carbon fluxes indicates a substantial missing sink in the global carbon budget since the early 20th century

Inconsistencies in the calculation of the two anthropogenic land flux terms of the global carbon cycle are investigated. The two terms—the direct anthropogenic flux (caused by direct human disturbance in anthromes, currently a carbon source to the atmosphere) and the indirect anthropogenic flux (caused indirectly by human activities that lead to global change and affecting all biomes, currently an atmospheric carbon sink)—are typically calculated independently, resulting in inconsistent underlying assumptions. We harmonize the estimation of the two anthropogenic land flux terms by incorporating previous estimates of these inconsistencies. We recalculate the global carbon budget (GCB) and apply change-point analysis to the cumulative budget imbalance. Cumulative over 1850–2018 (1959–2018), harmonization results in a 13% lesser (4% greater) land use source from anthromes and a 20% (23%) lesser land sink. This recalculation yields a greater non-closure of the GCB, indicating a missing carbon sink averaging 0.65 Pg C year -1 since the early 20th century. The imbalance likely results from a combination of method discontinuity and structural errors in the assessment of the direct anthropogenic land use flux, greater ocean carbon uptake, structural errors in land models, and in how these land terms are quantified for the budget. We caution against overconfidence in considering the GCB a solved problem and recommend further study of methodological discontinuities in budget terms. We strongly recommend studies that quantify the direct and indirect anthropogenic land fluxes simultaneously to ensure consistency, with a deeper understanding of human disturbance and legacy effects in anthromes.

54 ENVIRONMENTAL SCIENCES

GeoLoRA: Geometric integration for parameter efficient fine-tuning

Low-Rank Adaptation (LoRA) has become a widely used method for parameter-efficient fine-tuning of large-scale, pre-trained neural networks. However, LoRA and its extensions face several challenges, including the need for rank adaptivity, robustness, and computational efficiency during the fine-tuning process. We introduce GeoLoRA, a novel approach that addresses these limitations by leveraging dynamical low-rank approximation theory. GeoLoRA requires only a single backpropagation pass over the small-rank adapters, significantly reducing computational cost as compared to similar dynamical low-rank training methods and making it faster than popular baselines such as AdaLoRA. This allows GeoLoRA to efficiently adapt the allocated parameter budget across the model, achieving smaller low-rank adapters compared to heuristic methods like AdaLoRA and LoRA, while maintaining critical convergence, descent, and error-bound theoretical guarantees. The resulting method is not only more efficient but also more robust to varying hyperparameter settings. We demonstrate the effectiveness of GeoLoRA on several state-of-the-art benchmarks, showing that it outperforms existing methods in both accuracy and computational efficiency.

Schotthoefer, Steffen [ORNL] (ORCID:00000002156965

Water Mass Transformation Budgets in Finite‐Volume Generalized Vertical Coordinate Ocean Models

Water Mass Transformation (WMT) theory provides conceptual tools that in principle enable innovative analyses of numerical ocean models; in practice, however, these methods can be challenging to implement and interpret, and therefore remain under-utilized. Our aim is to demonstrate the feasibility of diagnosing all terms in the water mass budget and to exemplify their usefulness for scientific inquiry and model development by quantitatively relating water mass changes, overturning circulations, boundary fluxes, and interior mixing. We begin with a pedagogical derivation of key results of classical WMT theory. We then describe best practices for diagnosing each of the water mass budget terms from the output of Finite-Volume Generalized Vertical Coordinate (FV-GVC) ocean models, including the identification of a non-negligible remainder term as the spurious numerical mixing due to advection scheme discretization errors. We illustrate key aspects of the methodology through the analysis of a polygonal region of the Greater Baltic Sea in a regional demonstration simulation using the Modular Ocean Model v6 (MOM6). We verify the convergence of our WMT diagnostics by brute-force, comparing time-averaged (“offline”) diagnostics on various vertical grids to timestep-averaged (“online”) diagnostics on the native model grid. Finally, we briefly describe a stack of xarray-enabled Python packages for evaluating WMT budgets in FV-GVC models (culminating in the new xwmb package), which is intended to be model-agnostic and available for community use and development.

54 ENVIRONMENTAL SCIENCES

Analyzing and Exploring Training Recipes for Large-Scale Transformer-Based Weather Prediction

Abstract The rapid rise of deep learning (DL) in numerical weather prediction (NWP) has led to a proliferation of models which forecast atmospheric variables with comparable or superior skill than traditional physics-based NWP. However, among these leading DL models, there is a wide variance in both the training settings and architecture used. Further, the lack of thorough ablation studies makes it hard to discern which components are most critical to success. In this work, we show that it is possible to attain high forecast skill even with relatively off-the-shelf architectures, simple training procedures, and moderate compute budgets. Specifically, we train a minimally modified Swin Transformer V2 (SwinV2) on ERA5 data and find that it attains superior skill in terms of mean-square errors of deterministic forecasts when compared against the European Centre for Medium-Range Weather Forecasts’ Integrated Forecasting System (IFS). Almost all DL–NWP systems share a core set of hyperparameters and design decisions. To aid and expedite future DL–NWP research, we present an in-depth, systematic exploration of different loss functions, model sizes and depths, patch sizes, and multistep training objectives. We also examine the model performance with metrics beyond the typical accuracy (ACC) and RMSE and investigate how the performance scales with model size. Through our open-source code, scoring pipelines, and models, we share our findings on key aspects of the training pipeline. These ablations reduce the necessity for expensive hyperparameter tuning and lower the barrier to entry for future DL–NWP research. Significance Statement This study investigates the potential of using large-scale transformer-based models for weather prediction, showing that it is possible to achieve high forecast accuracy with simpler, off-the-shelf architectures. By training a minimally modified SwinV2 transformer on ERA5 data, we show that the model achieves competitive forecast skill in terms of mean-square error for key variables, outperforming the European Centre for Medium-Range Weather Forecasts’ Integrated Forecasting System (IFS) at all lead times. Our findings suggest that effective training strategies, such as multistep fine-tuning and channel-weighted losses, significantly enhance the model’s performance. However, we also highlight that these improvements come with trade-offs in other areas, such as ensemble spread and high-frequency spatial detail. This work highlights the promise of deep learning in improving weather forecasts, which could lead to better preparedness and response to weather events, ultimately benefiting society by providing more reliable weather predictions.

Willard, Jared D. [Lawrence Berkeley National Labo

Measurements of the 239 Pu(n,f)/ 235 U(n,f) and 238 U(n,f)/ 235 U(n,f) Cross-Section Ratios Using Quasi-Monoenergetic Neutron Beams

Neutron-induced fission cross-section ratio measurements were carried out at Triangle Universities Nuclear Laboratory (TUNL) over multiple experiment campaigns from 2021-2023. The total beam time for these measurements across all the experimental campaigns was approximately three weeks. This work was intended to serve as an independent validation of the fissionTPC cross-section ratio measurements. In contrast to the white spectrum neutron source and time-projection chamber utilized in the fissionTPC, these measurements utilized pulsed, quasi-monoenergetic neutron beams and fission ionization chambers. Therefore, this work has different sources of systematic error and can be used as a complementary measurement. In this report we describe our fission cross-section ratio measurements, analysis, results and provide a detailed uncertainty budget.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Simulation budgeting for hybrid effective field theories

In this work, we forecast the number of, and requirements on, N-body simulations needed to train hybrid effective field theory (HEFT) emulators for a range of use cases, using a hybrid of HMcode and perturbation theory as a surrogate model. Our accuracy goals, determined with careful consideration of statistical and systematic uncertainties, are 1% accurate in the high-likelihood range of cosmological parameters, and 2% accurate over a broader parameter space volume for k < 1 h Mpc -1 and z < 3. Focusing in part on the 8-parameter w 0 w a CDM+ m ν cosmological model, we find that < 225 simulations are required to meet our error goals over our wide parameter space, including models with rapidly evolving dark energy, given our simulation and emulator recommendations. For a more restricted parameter space volume, as few as 80 simulations are sufficient. We additionally present simulation forecasts for example use cases, and make the code used in our analyses publicly available. These results offer practical guidance for efficient emulator design and simulation budgeting in future cosmological analyses.

cosmological parameters from LSS

ZFP: A compressed array representation for numerical computations

HPC trends favor algorithms and implementations that reduce data motion relative to FLOPS. We investigate the use of lossy compressed data arrays in place of traditional IEEE floating point arrays to store the primary data of calculations. Simulation is fundamentally an exercise in controlled approximation, and error introduced by finite-precision arithmetic (or lossy compression) is just one of several sources of error that need to be managed to ensure sufficient accuracy in a computed result. We describe ZFP, a compressed numerical format designed for in-memory storage of multidimensional arrays, and summarize theoretical results that demonstrate that the error of repeated lossy compression can be bounded and controlled. Furthermore, we establish a relationship between grid resolution and compression-induced errors and show that, contrary to conventional floating point, ZFP reduces finite-difference errors with finer grids. We present example calculations that demonstrate data reduction by 4x or more with negligible impact on solution accuracy. Our results further demonstrate several orders-of-magnitude increase in accuracy using ZFP over IEEE floating point and Posits for the same storage budget.

Lindstrom, Peter

Autonomous Hydrogen Fueling Station

This project “Autonomous Hydrogen Fueling Station” covered the autonomous refueling with both gaseous hydrogen and liquid hydrogen. The part on gaseous hydrogen focused on the development of an autonomous robotic fueling arm that would couple to a fuel cell engine for hydrogen refueling without guidance from the forklift operator and budget period. Research was also covered for the robotic fueling with a commercial vehicle. The second phase of the project created the baseline for an autonomous liquid hydrogen transfer system that would minimize boil off losses by operating at thermodynamically efficient state points. For the development of the robotic fueling arm, testing was conducted to establish a baseline measurement of the accuracy and repeatability of a human operator positioning a lift truck in front of a dispenser. The goal was to establish the range of motion required for an autonomous fueling mechanism to mate a hydrogen nozzle with a receptacle on a fuel cell system installed in a forklift. The final design comprised a selective compliance articulated robot arm (SCARA)-type mechanism with two arms for horizontal motion and a ball screw for vertical movement and color and LIDAR cameras were used for marker identification and proximity awareness to guide the robotic arm to its target receptacle. Initial tests resulted in 199 out of 200 successful attempts at autonomous coupling of the dispensing coupler and a fuel cell engine, without hydrogen. The dispenser prototype was modified to include tubing for both hydrogen fuel and air purge lines, but subsequent tests were confounded by the shoulder motor over current errors which limited the robot from getting to the fully inserted position to achieve a positive latch. Robotic hydrogen refueling was successfully demonstrated over 1.5 hours of testing, Plug completed 29 successful latches with an average number of 4 sequential latches before failure. However, a robot capable of placing the nozzle with more force is required for higher reliability. For budget period two, a small scale (10 kg / transfer) automated control system was designed that would operate valves to control pressure and flow of liquid nitrogen between a source and receiving tank with an aim to minimize boil off losses by operating at the most thermodynamically efficient state points. Control system logic flow and a P&ID were developed prior to system safety characterization via HAZOP. A control narrative and system state points were defined. Delays in approval for a change of project objective and procurement issues precluded the construction and test of the final prototype system.

08 HYDROGEN

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE

Understanding the Biases in Global Monsoon Simulations from the Perspective of Atmospheric Energy Transport

Understanding global monsoon (GM) variability and projecting its future changes rely heavily on climate models. However, climate models generally show pronounced biases in GM simulations, and the reasons for this remain unclear. Here, in this study, we evaluate the performance of 20 pairs of climate models that participated in both phase 5 of the Coupled Model Intercomparison Project (CMIP5) and phase 6 of CMIP (CMIP6) and identify the sources of their GM simulation biases from an energy transport perspective. The multimodel mean improvement in CMIP6 compared to CMIP5 is demonstrated by the increasing skill scores for various GM metrics from 0.20–0.79 to 0.48–0.83. More specifically, the dry biases in the Northern Hemisphere Summer Monsoon (NHSM) precipitation in CMIP5 [root-mean-square error (RMSE): 1.85 mm day −1 ] are reduced in CMIP6 (RMSE: 1.66 mm day −1 ). This higher simulation skill is associated with higher skill in simulating the precipitation-solstitial mode, monsoon intensity, and monsoon domains. The improvement in the NHSM precipitation simulation results from that in the meridional transport of atmospheric energy. Atmospheric energy budget analysis shows that the negative biases in downward surface longwave radiation and northward energy transport are smaller in CMIP6 than in CMIP5 in the boreal summer, resulting in a more realistic interhemispheric thermal contrast and meridional gradient of moist static energy. However, a major weakness of the CMIP6 models is found in the Southern Hemisphere Summer Monsoon precipitation simulation due to the positive bias in the top-of-the-atmosphere downward longwave radiation. This study shows that reasonably reproducing the meridional global atmospheric energy transportation is necessary for skillful GM simulation.

54 ENVIRONMENTAL SCIENCES

Assessing the design of integrated methane sensing networks

Abstract While methane is the second largest contributor to global warming after carbon dioxide, it has a larger warming effect over a much shorter lifetime. Despite accelerated technological efforts to radically reduce global carbon dioxide emissions, rapid reductions in methane emissions are needed to limit near-term warming. Being primarily emitted as a byproduct from agricultural activities and energy extraction, methane is currently monitored via bottom–up (i.e. activity level) or top–down (via airborne or satellite retrievals) approaches. However, significant methane leaks remain undetected and emission rates are challenging to characterize with current monitoring frameworks. In this paper, we study the design of a layered monitoring approach that combines bottom–up and top–down approaches as an integrated sensing network. By recognizing that varying meteorological conditions and emission rates impact the efficacy of bottom–up monitoring, we develop a probabilistic approach to optimal sensor placement in its bottom–up network. Subsequently, we derive an inverse Bayesian framework to quantify the improvement that a design-optimized integrated framework has on emission-rate quantifications and their uncertainties. We find that under realistic meteorological conditions, the overall error in estimating the true emission rates is approximately 1.3 times higher, with their uncertainties being approximately 2.4 times higher, when using a randomized network over an optimized network, highlighting the importance of optimizing the design of integrated methane sensing networks. Further, we find that optimized networks can improve scenario coverage fractions by more than a factor of 2 over experimentally-studied networks, and identify a budget threshold beyond which the rate of optimized-network coverage improvement exhibits diminishing returns, suggesting that strategic sensor placement is also crucial for maximizing network efficiency.

54 ENVIRONMENTAL SCIENCES

Optimal Client Sampling in Federated Learning with Client-level Heterogeneous Differential Privacy

Federated Learning with client-level differential privacy (DP) provides a promising framework for collaboratively training models while rigorously protecting clients’ privacy. However, classic approaches like DP-FedAvg struggle when clients have heterogeneous privacy requirements, as they must uniformly enforce the strictest privacy level across all clients, leading to excessive DP noise and significant degradation in model utility. Existing methods to improve the model utility in such heterogeneous privacy settings often assume a trusted server and are largely heuristic, resulting in suboptimal performance and lacking strong theoretical foundations. Here, in this work, we address these challenges under a practical attack model where both clients and the server are honest-but-curious. We propose GDPFed, which partitions clients into groups based on their privacy budgets and achieves client-level DP within each group to reduce the privacy budget waste and hence improve the model utility. Based on the privacy and convergence analysis of GDPFed, we find that the magnitude of DP noise depends on both model dimensionality and the per-group client sampling ratios. To further improve the performance of GDPFed, we introduce GDPFed+, which integrates model sparsification to eliminate unnecessary noise and optimizes per-group client sampling ratios to minimize convergence error. Extensive empirical evaluations on multiple benchmark datasets demonstrate the effectiveness of GDPFed+, showing substantial performance gains compared with state-of-the-art methods.

Xu, Jiahao [Univ. of Nevada, Reno, NV (United Stat

Improving the efficiency of learning-based error mitigation

Error mitigation will play an important role in practical applications of near-term noisy quantum computers. Current error mitigation methods typically concentrate on correction quality at the expense of frugality (as measured by the number of additional calls to quantum hardware). To fill the need for highly accurate, yet inexpensive techniques, we introduce an error mitigation scheme that builds on Clifford data regression (CDR). The scheme improves the frugality by carefully choosing the training data and exploiting the symmetries of the problem. We test our approach by correcting long range correlators of the ground state of XY Hamiltonian on IBM Toronto quantum computer. We find that our method is an order of magnitude cheaper while maintaining the same accuracy as the original CDR approach. The efficiency gain enables us to obtain a factor of 10 improvement on the unmitigated results with the total budget as small as 2 ⋅ 10 5 shots. Furthermore, we demonstrate orders of magnitude improvements in frugality for mitigation of energy of the LiH ground state simulated with IBM's Ourense-derived noise model.

97 MATHEMATICS AND COMPUTING

Validating a Dynamic PWR Safety and Security Model?

Nuclear power plants (NPPs) are assessed for safety and security using separate models that cannot capture how an attacker's decisions and a plant's response unfold together in real time, leaving regulators and operators without a complete picture of true plant vulnerability. Traditional probabilistic risk assessment (PRA) methods treat adversarial events as fixed initiators with predetermined outcomes, and are structurally incapable of representing the time-dependent interplay between physical security events, safety system response, and operator mitigative actions. At Idaho National Laboratory (INL), I contributed to the development and validation of Modeling and Analysis for Safety and Security using the Dynamic EMRALD Framework (MASS-DEF). Where static PRA relies on event-tree logic that cannot evolve mid-scenario, MASS-DEF couples a time-dependent dynamic PRA tool EMRALD (Event Modeling Risk Assessment using Linked Diagrams) with attack simulation software, allowing attacker behavior, plant system states, and operator actions to interact across time. My work focused on validating a general Pressurized Water Reactor (PWR) model. I traced model logic against PWR plant to identified errors in logic and confirm accuracy. I then built and tested attack scenarios against a general PWR model to verify that the model produced expected outcomes across all logical pathways. I also contributed a section to a related technical paper applying the same EMRALD platform to radiation dose modeling. Results show that MASS-DEF can quantitatively demonstrate that many plants exceed their regulatory security thresholds. This demonstrated margin provides a technically defensible basis for reducing the number of guards without compromising regulatory compliance. Physical security costs represent roughly 10% of annual operating budgets, making such reductions directly meaningful to INL's mission of sustaining existing commercial NPPs. This internship strengthened my understanding of nuclear systems, probabilistic modeling, and technical writing, and has solidified my pursuit of a career at a national laboratory.

98 - NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL

Better practices for inferring ecosystem water use strategy from eddy covariance data

Eddy covariance data are critical for inferring ecosystem water use strategies. Yet, such inferences are sensitive to a range of assumptions applied across studies, hindering our understanding of water use strategies within and across eddy covariance sites. A recent analysis across 151 FLUXNET2015 and AmeriFlux-FLUXNET datasets found that poor model performance was the key driver of non-robust inferences of ecosystem water use strategies. Here, we leverage this previous analysis to (i) identify the specific assumptions that improve inference model performance across most sites, (ii) explain the mechanisms behind the performance improvements, and (iii) check whether better performance improves water use inference. We find that the common practice of fitting a model to canopy conductance (G c ) derived from the evapotranspiration (ET) observations, rather than to observed ET itself, artificially amplifies data errors and degrades the model performance. Next, accounting for vegetation dynamics by applying a growing season filter or incorporating satellite LAI data improves performance, but the former practice may remove soil water stress periods. Lastly, using the leaf-to-air vapor pressure deficit (VPD l ) derived from ET observations as a model input may artificially inflate performance. Based on these results, we recommend selecting observed ET (rather than derived G c ) as the response variable, carefully accounting for vegetation dynamics, and avoiding derived VPD l as a model input; these best practices improve model performance by c. 20% and robustness by c. 80% across all eddy covariance sites. Nevertheless, the performance improvements do not always correspond to more robust inference of water use strategies, as model parameter selection and surface energy budget closure corrections still strongly influence the ecosystem water use parameter estimation in a site-specific manner.

AmeriFlux

Constrained variational optimization of counting-time allocation in sequential scattering measurements: Application to Bonse–Hart USANS

Sequential scattering measurements are often performed under a fixed experimental-time budget, even though the expected count rate varies strongly across the measured coordinate. When the dwell time at each measurement position can be controlled independently, this variation creates a general resource-allocation problem: how should the available time be distributed to minimize the uncertainty of the reconstructed profile? We formulate this problem as a constrained variational optimization for measurements governed by Poisson counting statistics. When each measurement is treated independently, minimizing the averaged squared relative uncertainty yields an inverse-square-root intensity allocation. The formulation is then generalized to include correlations between neighboring measurements and an instrumental resolution operator, leading to an allocation criterion that equalizes the marginal reduction in posterior uncertainty per unit measurement time. Bonse–Hart ultra-small-angle neutron scattering (USANS), in which reciprocal space is sampled sequentially through analyzer-angle stepping, provides an experimentally grounded application. Computational benchmarking shows that the optimized allocation outperforms uniform-time and constant-relative-error strategies, while application to an experimentally measured graphite USANS profile from the Spallation Neutron Source, using Poisson resampling under alternative schedules, demonstrates how counting time should be redistributed toward weak-intensity regions under an identical total duration. The resulting framework applies to sequential scattering and related scanning measurements whenever local dwell times are adjustable and directly determine the measurement uncertainties, and when the relevant correlation and instrumental-response models are available.

Tung, Chi-Huan [ORNL] (ORCID:0000000221972074)

Mesoscale Convective Systems Represented in High Resolution E3SMv2 and Impact of New Cloud and Convection Parameterizations

Mesoscale convective systems (MCSs) play an important role in modulating the global hydrological cycle, general circulation, and radiative energy budget. In this study, we evaluate MCS simulations in the second version of U.S. Department of Energy (DOE) Energy Exascale Earth System Model (E3SMv2). E3SMv2 atmosphere model (EAMv2) is run at the uniform 0.25? horizontal resolution. We track MCSs consistently in the model and observations using the PyFLEXTRKR algorithm, which defines MCS based on both cloud-top brightness temperature (Tb) and surface precipitation. Results from using Tb only to define MCS, commonly used in previous studies, are also discussed. Furthermore, sensitivity experiments are performed to examine the impact of new cloud and convection parameterizations developed for EAMv3 on simulated MCSs. Our results show that EAMv2 simulated MCS precipitation is largely underestimated in the tropics and contiguous United States. This is mainly attributed to the underestimated precipitation intensity in EAMv2. In contrast, the simulated MCS frequency becomes more comparable to observations if MCSs are defined only based on cloud-top Tb. The Tb-based MCS tracking method, however, includes many cloud systems with very weak precipitation which conflicts with the MCS definition. This result illustrates the importance of accounting for precipitation in evaluating simulated MCSs. We also find that the new physics parameterizations help increase the relative contribution of convective precipitation to total precipitation in the tropics, but the simulated MCS properties are generally not improved. This suggests that simulating MCSs will remain a challenge for the next version of E3SM.

Zhang, Meng

A New Coupled Biogeochemical Modeling Approach Provides Accurate Predictions of Methane and Carbon Dioxide Fluxes Across Diverse Tidal Wetlands

Abstract Tidal wetlands provide valuable ecosystem services, including storing large amounts of carbon. However, the net exchanges of carbon dioxide (CO 2 ) and methane (CH 4 ) in tidal wetlands are highly uncertain. While several biogeochemical models can operate in tidal wetlands, they have yet to be parameterized and validated against high‐frequency, ecosystem‐scale CO 2 and CH 4 flux measurements across diverse sites. We paired the Cohort Marsh Equilibrium Model (CMEM) with a version of the PEPRMT model called PEPRMT‐Tidal, which considers the effects of water table height, sulfate, and nitrate availability on CO 2 and CH 4 emissions. Using a model‐data fusion approach, we parameterized the model with three sites and validated it with two independent sites, with representation from the three marine coasts of North America. Gross primary productivity (GPP) and ecosystem respiration (R eco ) modules explained, on average, 73% of the variation in CO 2 exchange with low model error (normalized root mean square error (nRMSE) <1). The CH 4 module also explained the majority of variance in CH 4 emissions in validation sites ( R 2 = 0.54; nRMSE = 1.15). The PEPRMT‐Tidal‐CMEM model coupling is a key advance toward constraining estimates of greenhouse gas emissions across diverse North American tidal wetlands. Further analyses of model error and case studies during changing salinity conditions guide future modeling efforts regarding four main processes: (a) the influence of salinity and nitrate on GPP, (b) the influence of laterally transported dissolved inorganic C on R eco , (c) heterogeneous sulfate availability and methylotrophic methanogenesis impacts on surface CH 4 emissions, and (d) CH 4 responses to non‐periodic changes in salinity.

54 ENVIRONMENTAL SCIENCES