Engineering PapersSearch

SEARCH · Engineering Papers

Results for “large-scale systems”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Electrode Erosion and Prefire Studies Towards Fusion Scale Pulsed Power

This study presents a comprehensive investigation of electrode erosion and discharge behavior in spark gap switches over long switching cycle lifetimes. Brass, copper–tungsten (CuW), and stainless steel electrodes are tested under controlled conditions to quantify material degradation, debris accumulation, and changes in breakdown voltage. High-resolution imaging and statistical analysis of spark channel locations and gap breakdown voltages reveal how surface evolution influences long-term performance and reliability. These results provide essential data for lifetime modeling and inform design strategies for pulsed power systems in emerging applications such as private sector fusion energy and large-scale facilities like Sandia’s Z Machine and proposed ZX upgrades, where high repetition reliability and predictable behavior are critical.

electrical breakdown

Toward an event-level analysis of hadron structure using differential programming

Reconstructing the internal properties of hadrons in terms of fundamental quark and gluon de- grees of freedom is a central goal in nuclear and particle physics. This effort lies at the core of major experimental programs, such as the Jefferson Lab 12 GeV program and the upcoming Electron-Ion Collider. A primary challenge is the inherent inverse problem: converting large-scale observational data from collision events into the fundamental QCD-defined densities that characterize the micro- scopic structure of hadronic systems. Recent advances in AI and machine learning have opened new avenues for addressing this challenge using deep learning techniques. A particularly promising direction is the integration of complex theoretical calculations and experimental simulations into a unified framework capable of reconstructing these densities directly from event-level information. In this document, we introduce a key algorithm called LOITS, which enables differentiable program- ming within such a framework, facilitating the use of AI/ML techniques to solve the inverse problem of QCF reconstruction at the event level.

Braga, Kevin [College of William and Mary, William

SimH 2 : an integrated techno-economic modeling framework for hydrogen pipeline infrastructure and network optimization

Large-scale hydrogen (H 2 ) pipeline transport design and network optimization have seldom been reported due to the lack of a cost model accounting for the relationship between transport cost and hydrogen mass flow rate. Here, this work introduced a system-level cost model for hydrogen pipeline transport at supercritical state and integrated it with an existing CO 2 pipeline network tool, SimCCS, for hydrogen-specific pipeline design and optimization. The Intermountain West (I-West) region of the U.S., historically dependent on fossil fuel-based economies, is chosen to demonstrate the capabilities of our H 2 pipeline cost model and transport network optimization platform called SimH 2 . Two scenarios are examined: one where the pipeline is not allowed to pass through disadvantaged communities and the other where it is permitted. The results highlight that incorporating disadvantaged-community constraints lead to longer pipeline routes and increased transport costs, reflecting the trade-offs involved in equitable infrastructure development. It is demonstrated that the newly developed SimH 2 tool not only enables the efficient design of H 2 transportation pipelines but also optimizes the network by accounting for local terrain and the presence of disadvantaged areas.

08 HYDROGEN

The S-PLUS Fifth Data Release: Over 4500 Square Degrees of the Southern Sky and a Multicolor View of the Hydra and Antlia Galaxy Clusters

Abstract We present the fifth data release (DR5) of the Southern Photometric Local Universe Survey (S-PLUS), covering 4592 deg 2 across 2491 fields. Observations were conducted with the T80-South, a Brazilian robotic telescope equipped with the Javalambre 12-filter system, containing five broadband and seven narrowband filters. Data products feature FITS images and extensive catalogs containing fluxes, magnitudes, and shape parameters for over 113 million detections. In addition, several value-added catalogs are provided, offering photometric redshifts (photo- zs ), object classifications, masks, and extinction coefficients. For the first time, this release includes coverage of 110 deg 2 along the Galactic disk, facilitating new research into Galactic structure and stellar populations. The release also provides full coverage of the Hydra Supercluster and numerous other nearby clusters with improved data reduction and calibration, enhancing photo- z accuracy, which is vital for large-scale structure studies. A preliminary analysis of the Hydra and Antlia galaxy clusters up to 5 × R 200 yields an updated catalog of 1706 cluster members based on both spectroscopic and high-quality photo- zs . Our photometric data shows that both clusters have a similar proportion of galaxies with an H α excess relative to their clustercentric distance, though Hydra has a higher fraction near its center. Additionally, the spatial distribution of all objects in our sample highlights a bridge connecting both clusters. We verify that S-PLUS DR5 provides a solid foundation for future scientific investigations, ranging from solar system studies to cosmology.

Vinicius Rodrigues de Lima, Erik [Universidade de

Ensemble Kalman filter for data assimilation coupled with low-resolution computations techniques applied in fluid dynamics

This paper presents an innovative Reduced-order model (ROM) for merging experimental and simulation data using data assimilation (DA) to estimate the "True" state of a fluid dynamics system, leading to more accurate predictions. Our methodology introduces a novel approach by implementing the ensemble Kalman filter (EnKF) within a reduced-dimensional framework, grounded in a robust theoretical foundation and applied to fluid dynamics. To address the substantial computational demands of DA, the proposed ROM employs low-resolution (LR) techniques to drastically reduce computational costs. This innovative approach involves downsampling datasets for DA computations, followed by an advanced reconstruction technique based on low-cost singular value decomposition (lcSVD). The lcSVD method, a key innovation in this paper, has never been applied to DA before and offers a highly efficient way to enhance resolution with minimal computational resources. Our results demonstrate significant reductions in both computation time and RAM usage through these LR techniques without compromising the accuracy of the estimations. For instance, in a turbulent test case, for a data compression rate of 15.9, the LR approach can achieve a speed-up of 13.7 and a RAM compression of 90.9% while maintaining a low relative root mean square error (RRMSE) of 2.6%, compared to 0.8% in the high-resolution (HR) reference. Furthermore, we highlight the effectiveness of the EnKF in estimating and predicting the state of fluid flow systems based on limited observations and given low-fidelity numerical data. This paper highlights the potential of the proposed DA method in fluid dynamics applications, particularly for improving computational efficiency in CFD and related fields. Its ability to balance accuracy with low computational and memory costs makes it especially suitable for large-scale and real-time applications, such as environmental monitoring or engineering design. This method will be incorporated into ModelFLOWs-app.

Data Assimilation

Photosynthetic capacity is reduced by warming but unaffected by elevated CO2 in seedlings of five boreal tree species

Abstract Increasing atmospheric CO2 concentrations fuel global warming, with boreal regions warming at a faster rate than many other areas. Boreal forests are an important component of the global carbon cycle, yet we have little data on photosynthetic responses of boreal trees to elevated CO2 (EC) and warming. We grew seedlings of 5 widespread North American boreal tree species (from Betula, Larix, Picea, and Pinus) under current (410 ppm) or elevated (750 ppm) CO2 and either ambient (+0 °C) or increased (+4 °C or +8 °C) temperature, then measured photosynthetic traits over a range of leaf temperatures. Our results were generally consistent across species: photosynthetic capacity (maximum rates of Rubisco carboxylation, Vcmax, and electron transport, Jmax) was unaffected by EC but decreased under +8 °C warming. Accordingly, net photosynthesis measured at the growth CO2 concentration (Agrowth) was reduced under warming and increased under EC. The thermal optimum for Agrowth (ToptA) increased by ∼1.8 °C with EC but increased with warming in only two species. In contrast, the activation energies and thermal optima for Vcmax and Jmax, which are used to estimate photosynthesis in Earth System Models, were unaffected by growth environment. There were a few interactions between growth, CO2, and warming. These results suggest increased photosynthesis of widespread boreal tree species under EC may be offset by future reductions in photosynthetic capacity related to warming. We also show that the temperature sensitivities of parameters used to estimate global photosynthesis in large-scale models are generally unaffected by simulated climate change in these species.

Plant Sciences

Integrated Transmission-Distribution Multi-Period Switching for Wildfire Risk Mitigation: Improving Speed and Scalability with Distributed Optimization: Preprint

With increasingly severe wildfire conditions driven by climate change, utilities must manage the risk of wildfire ignitions from electric power lines. During "public safety power shutoff'" events, utilities de-energize power lines to reduce wildfire ignition risk, which may result in load shedding. Distributed energy resources provide flexibility that can help support the system to reduce load shedding when lines are de-energized. We investigate a coordinated transmission-distribution optimization problem that balances wildfire risk mitigation and load shedding. We model distribution systems that include battery energy storage systems which may support loads when transmission lines are de-energized. This multi-period integrated transmission-distribution optimal switching problem jointly optimizes line switching decisions, the generators' setpoints, load shedding, and the batteries' states of charge, resulting in significant computational challenges. To improve scalability, we decompose the problem over both space and time and apply a distributed optimization algorithm. Using a large-scale synthetic California test case with realistic distribution models and real wildfire risk data, we show that distributed optimization can solve large-scale multi-period switching problems that are otherwise intractable for centralized solvers. We also discuss challenges and future directions for improving the distributed algorithm's convergence performance as the number of time periods increases.

24 POWER TRANSMISSION AND DISTRIBUTION

Mode Multiplexing for Scalable Cavity-Enhanced Operations in Neutral-Atom Arrays

Neutral-atom arrays provide a versatile platform for quantum information processing. However, in large-scale arrays, efficient photon collection remains a bottleneck for key tasks such as fast, nondestructive qubit readout and remote entanglement distribution. We propose a cavity-based approach that enables fast, parallel operations over many atoms using multiple modes of a single optical cavity. By selectively shifting the relevant atomic transitions, each atom can be coupled to a distinct cavity mode, allowing independent simultaneous processing. We present practical system designs that support cavity-mode multiplexing with up to 50 modes, enabling rapid mid-circuit syndrome extraction and significantly enhancing entanglement distribution rates between remote atom arrays. This approach offers a scalable solution to core challenges in neutral-atom arrays, advancing the development of practical quantum technologies.

Aqua, Ziv [Massachusetts Institute of Technology (

Interactions Between Climate Policy and Technology-influenced Travel Behavior: Mitigating Induced Demand from CACC

Advances in vehicle technology have influenced the development of automated vehicle systems, where vehicles that do not require human intervention are already deployed in the roadway networks. While these advances are proved to increase roadway safety and highway capacity, more research is needed to understand the long-term and regional-level impacts on mobility, land use, energy consumption, and emissions. This study proposes a multi-model approach to analyze the effect of vehicle automation and deep decarbonization policies over a period from 2020 to 2040 in Austin, Texas. We use the Global Change Analysis Model (GCAM) to develop internally the scenarios that are then passed to the SMART Mobility modeling workflow, a large-scale simulation framework combining the POLARIS activity-based travel demand model and mesoscopic traffic simulator with the Autonomie vehicle energy consumption model and the UrbanSim land use simulator. Results suggest that the introduction of vehicles with advanced automation could increase fuel consumption when no decarbonization policies are implemented. Also, advances in vehicle technology research and development could lead to a decline in energy use in the long-term. Energy pricing and vehicle electrification incentives could help reduce the impact of vehicle automation. Finally, our analysis indicates the relevance of introducing land use processes in longterm vehicle automation studies.

land use

Explainable machine learning reveals that local structural motifs encode the thermodynamic state across the CuZr metallic glass-forming range

Metallic glasses derive their properties from the statistics of local atomic motifs rather than from long-range order, yet a quantitative, chemistry-specific link between motif populations and the underlying glassy state has remained elusive. In this work we combine large-scale molecular dynamics, Voronoi tessellation, deep neural networks, and SHapley Additive exPlanations (SHAP) to identify which local structural motifs define the glassy state of Cu—Zr metallic glasses. A dataset of 17,180 atomistic configurations spanning ten compositions (Cu 20 Zr 80 –Cu 80 Zr 20 ) and four quench rates (10 9 –10 12 K/s) is used to train a feed-forward neural network that regresses temperature across the 50–2000 K liquid–supercooled–glass range, achieving a mean absolute error of 19.89 K and R 2 = 0.9974, confirming that the local structural state is faithfully encoded in motif-level structure. SHAP analysis then reveals that a tightly coupled near-icosahedral family of motifs (coordination numbers (CN) 11–13, including the full icosahedron 001200 and its single-atom-perturbation sibling 10930) collectively encodes the thermodynamic state of the system across the full glass-forming range. The CN = 11–13 ordered members carry negative SHAP values at high populations, tracking the most deeply-quenched configurations, while 10930 shows the reversed signature consistent with its role as a soft-spot host whose population shrinks as the icosahedral network deepens. The analysis demonstrates that explainable machine learning can isolate the minimal motif vocabulary defining the glassy state and recovers the near-icosahedral building blocks previously identified by data-driven analyses of Cu—Zr. The approach provides a general, chemistry-specific route for characterizing the structural state of disordered materials.

36 MATERIALS SCIENCE

Hydrogeological assessment of CO2 containment assurance and wellbore integrity at a Gulf Coast storage site

Abstract A large-scale carbon capture and storage (CCS) initiative on the Texas Gulf Coast serves as a premier demonstration of the U.S. Department of Energy’s CarbonSAFE program. Targeting deep saline formations, specifically Oligo-Miocene deltaic sequences, the project aims to establish technical and commercial viability for geologic CO2 storage within a major industrial corridor. This study provides a rigorous hydrogeological assessment to support Class VI permitting by quantifying the high degree of containment security. Utilizing a compositional reservoir simulator, we developed a suite of 27 distinct simulation cases to evaluate vertical plume dynamics near both planned injection wells and proximal legacy infrastructure. To ensure numerical accuracy near wellbores, we implemented a refined mesh strategy, determining that a 5.6 ft × 5.6 ft grid refinement offered the optimal balance between computational efficiency and descriptive precision. The modeling framework utilized a systematic sensitivity-based approach to evaluate the mechanical redundancy of the subsurface system by performing a bounding analysis of wellbore interfaces against hypothetical high-permeability microannuli. By systematically isolating competing physical drivers, including permeability, porosity, gas hysteresis, thermal gradients, salinity, and solubility trapping (quantified via Henry’s law with dynamically adjusted coefficients), this work moves beyond binary assessments to establish a nuanced hierarchy of containment factors. The results confirm that primary trapping mechanisms (e.g., gas hysteresis and solubility), combined with the site's unique geomechanical stratigraphy, significantly restrict vertical mobility and reinforce the robust containment security of the reservoir. Baseline results demonstrate substantial vertical separation between the CO2 plume and the upper confining system, ensuring robust containment. Sensitivity analysis reveals that even under highly conservative bounding scenarios—assuming theoretical 10-Darcy pathways at specific wellbore locations—the 2,900-ft thick multi-layered confining zone remains a reliable barrier. In these hypothetical upper-bound cases, peak upward fluxes of CO2 and saltwater after 15 years of injection remain localized and dissipate rapidly within the lower sections of the confining interval, leaving the integrity of the seal uncompromised. Furthermore, the study identifies that while localized wellbore pathways define theoretical upper bounds of vertical migration, the Area of Review (AoR) is primarily sensitive to regional thermal gradients and hysteresis, which can influence the AoR by over 3,000 acres in pessimistic configurations. Also, primary trapping mechanisms, specifically gas hysteresis and solubility, work in tandem with the Gulf Coast’s unique geomechanical stratigraphy to significantly restrict vertical mobility. Ductile, smectite-rich mudstones facilitate natural borehole convergence and the self-healing of potential conduits, creating a natural geomechanical bridge that effectively mitigates migration potential at both current injection points and legacy-well locations. This comprehensive modeling effort demonstrates that the integration of high-resolution wellbore simulations and regional geomechanical observations confirms the long-term storage security of the studied site, providing a physics-based foundation for industrial-scale CCS deployments. This modeling framework establishes a baseline for future research into coupled geomechanical effects, such as time-dependent borehole convergence, to further refine long-term containment projections. Acknowledgements We thank the Gulf Coast Carbon Center (GCCC) at the Bureau of Economic Geology for foundational research support. We appreciate Alex Bump for technical guidance and David Hoffman for model mesh generation. This work used TACC’s Frontera cluster for simulations and CMG Ltd. software licenses provided to UT-Austin. This material is based upon work supported by the Department of Energy under Award Number DE-FE0032338. Disclaimer This material is based upon work supported by the U.S. Department of Energy’s Fossil Energy and Carbon Management Office under the CarbonSAFE program, award Number DE-FE0032338. The views expressed herein do not necessarily represent the views of the U.S. Department of Energy or the United States Government.

58 GEOSCIENCES

Hydrogeological assessment of CO2 containment assurance and wellbore integrity at a Gulf Coast storage site

Abstract A large-scale carbon capture and storage (CCS) initiative on the Texas Gulf Coast serves as a premier demonstration of the U.S. Department of Energy’s CarbonSAFE program. Targeting deep saline formations, specifically Oligo-Miocene deltaic sequences, the project aims to establish technical and commercial viability for geologic CO2 storage within a major industrial corridor. This study provides a rigorous hydrogeological assessment to support Class VI permitting by quantifying the high degree of containment security. Utilizing a compositional reservoir simulator, we developed a suite of 27 distinct simulation cases to evaluate vertical plume dynamics near both planned injection wells and proximal legacy infrastructure. To ensure numerical accuracy near wellbores, we implemented a refined mesh strategy, determining that a 5.6 ft × 5.6 ft grid refinement offered the optimal balance between computational efficiency and descriptive precision. The modeling framework utilized a systematic sensitivity-based approach to evaluate the mechanical redundancy of the subsurface system by performing a bounding analysis of wellbore interfaces against hypothetical high-permeability microannuli. By systematically isolating competing physical drivers, including permeability, porosity, gas hysteresis, thermal gradients, salinity, and solubility trapping (quantified via Henry’s law with dynamically adjusted coefficients), this work moves beyond binary assessments to establish a nuanced hierarchy of containment factors. The results confirm that primary trapping mechanisms (e.g., gas hysteresis and solubility), combined with the site's unique geomechanical stratigraphy, significantly restrict vertical mobility and reinforce the robust containment security of the reservoir. Baseline results demonstrate substantial vertical separation between the CO2 plume and the upper confining system, ensuring robust containment. Sensitivity analysis reveals that even under highly conservative bounding scenarios—assuming theoretical 10-Darcy pathways at specific wellbore locations—the 2,900-ft thick multi-layered confining zone remains a reliable barrier. In these hypothetical upper-bound cases, peak upward fluxes of CO2 and saltwater after 15 years of injection remain localized and dissipate rapidly within the lower sections of the confining interval, leaving the integrity of the seal uncompromised. Furthermore, the study identifies that while localized wellbore pathways define theoretical upper bounds of vertical migration, the Area of Review (AoR) is primarily sensitive to regional thermal gradients and hysteresis, which can influence the AoR by over 3,000 acres in pessimistic configurations. Also, primary trapping mechanisms, specifically gas hysteresis and solubility, work in tandem with the Gulf Coast’s unique geomechanical stratigraphy to significantly restrict vertical mobility. Ductile, smectite-rich mudstones facilitate natural borehole convergence and the self-healing of potential conduits, creating a natural geomechanical bridge that effectively mitigates migration potential at both current injection points and legacy-well locations. This comprehensive modeling effort demonstrates that the integration of high-resolution wellbore simulations and regional geomechanical observations confirms the long-term storage security of the studied site, providing a physics-based foundation for industrial-scale CCS deployments. This modeling framework establishes a baseline for future research into coupled geomechanical effects, such as time-dependent borehole convergence, to further refine long-term containment projections. Acknowledgements We thank the Gulf Coast Carbon Center (GCCC) at the Bureau of Economic Geology for foundational research support. We appreciate Alex Bump for technical guidance and David Hoffman for model mesh generation. This work used TACC’s Frontera cluster for simulations and CMG Ltd. software licenses provided to UT-Austin. This material is based upon work supported by the Department of Energy under Award Number DE-FE0032338. Disclaimer This material is based upon work supported by the U.S. Department of Energy’s Fossil Energy and Carbon Management Office under the CarbonSAFE program, award Number DE-FE0032338. The views expressed herein do not necessarily represent the views of the U.S. Department of Energy or the United States Government.

58 GEOSCIENCES

Analytical Model for Atomic Relaxation in Twisted Moiré Materials

By virtue of being atomically thin, the electronic properties of heterostructures built from two-dimensional materials are strongly influenced by atomic relaxation. The atomic layers behave as flexible membranes rather than rigid crystals. Here we develop an analytical theory of lattice relaxation in twisted moiré materials. We obtain analytical results for the lattice displacements and corresponding pseudo gauge fields, as a function of twist angle. We benchmark our results for twisted bilayer graphene and twisted WSe 2 bilayers using large-scale molecular dynamics simulations. Our single-parameter theory is valid in graphene bilayers for twist angles 𝜃 ≳ 0.7°, and in twisted WSe 2 for 𝜃 ≳ 1.6°. Furthermore, we also investigate how relaxation alters the electronic structure in twisted bilayer graphene, providing a simple extension to the continuum model to account for lattice relaxation.

36 MATERIALS SCIENCE

Refractory-based thermal energy storage for industrial process heat: one-dimensional modeling, control, and optimization

The variable and weather-dependent output of wind and solar power plants present a substantial challenge for planning and operating electricity-systems, particularly in the absence of cost-effective and dispatchable energy storage technologies. This study investigates a high-temperature, electrically heated, refractory-based thermal energy storage (RTES) system that stores electrical energy as sensible heat in dense ceramic bricks over the 950–1800 °C range. The stored heat can be discharged as a controlled hot-gas stream for industrial heating, fuel substitution in high-temperature processes, or electricity generation. The main novelty is a comprehensive modelling, control, mapping, and optimization framework that integrates one-dimensional transient gas–solid heat transfer, fan-assisted discharge, bypass-flow regulation, reheating logic, fan-power evaluation, insulation-loss assessment, and genetic-algorithm-based design optimization. The model uses feedback from outlet temperature and delivered power to regulate discharge, while a two-stage genetic algorithm optimizes brick-channel geometry, gas-flow operation, and multilayer insulation thicknesses. Storage capacities below 50 MWh and discharge powers of 5–30 MW are analyzed to evaluate hold time, thermal delivery, fan-power penalty, heat loss, state-of-charge evolution, and indicative capital cost. Results demonstrate that optimized and well-insulated refractory-based thermal energy storage units can provide stable, efficient, and repeatable heat delivery over multiple discharge cycles. The generated performance and cost maps support modular refractory thermal energy storage as a practical option for large-scale integration of wind and solar generation and for high-temperature industrial process heat.

25 ENERGY STORAGE

HydraGNN_Predictive_GFM_2026 - Ensemble of predictive graph foundation models for atomistic materials modeling

This release contains data and parameters of HydraGNN-based graph foundation models trained as a result of the work published in the pre-print "Exascale Multi-Task Graph Foundation Models for Imbalanced, Multi-Fidelity Atomistic Data" by M. Lupo Pasini et al. (https://arxiv.org/abs/2604.15380). We jointly train on 16 open first-principles datasets (544+ million structures covering 85+ elements) using a multi-task architecture with per-dataset heads and a scalable ADIOS2/DDStore data pipeline. On Frontier, we execute six large-scale DeepHyper hyperparameter optimization campaigns in FP64 and promote the top-performing message-passing models to sustained 2,048-node training, yielding a PaiNN-based lead model. The version of HydraGNN used to generate the outputs provided in this release is HydraGNN v5.0 (https://github.com/ORNL/HydraGNN/releases/tag/v5.0) The list of datasets used for the training of the graph foundation model is the following: 1) Alexandria [1] 2) ANI1x [2] 3) MPTrj [3] 4) Open Catalyst 2020 (OC20) [4] 5) Open Catalyst 2022 (OC22) [5] 6) Open Catalyst 2025 (OC25) [6] 7) Open Direct ir Capture 2023 (ODAC23) [7] 8) Open Materials 2024 (OMat24) [8] 9) Open Molecules 2025 (OMol25) [9] 10) OMol25-neutral (subset of OMol25 that contains only molecules with zero total charge) 11) OMol25-non-neutral (subset of OMol25 that contains only molecules with non-zero total charge) 12) Open Polymers 2026 (OPoly2026) [10] 13) Nabla2DFT [11] 14) QCML [12] 15) QM7X [reference 13] 16) transition1x [14] Dataset references: [1] J. Schmidt et al., “A dataset of 175k stable and metastable materials calculated with the PBEsol and SCAN functionals,” Scientific Data, vol. 9, p. 64, 2022. [2] J. S. Smith et al., “The ANI-1ccx and ANI-1x data sets, coupled-cluster and density functional theory properties for molecules,” Scientific Data, vol. 7, p. 134, 2020. [Online]. Available: https: //www.nature.com/articles/s41597-020-0473-z [3] A. Jain et al., “Commentary: The Materials Project: A materials genome approach to accelerating materials innovation,” APL Materials, vol. 1, no. 1, p. 011002, 07 2013. [Online]. Available: https://doi.org/10.1063/1.4812323 [4] L. Chanussot et al., “Open catalyst 2020 (oc20) dataset and community challenges,” ACS Catalysis, vol. 11, no. 10, pp. 6059–6072, 2021. [Online]. Available: https://doi.org/10.1021/acscatal.0c04525 [5] K. Tran et al., “Open catalyst 2022 (oc22) dataset and challenges for oxidation electrocatalysts,” ACS Catalysis, vol. 13, no. 5, pp. 3066–3084, 2023. [Online]. Available: https://doi.org/10.1021/acscatal.2c05426 [6] S. J. Sahoo et al., “The open catalyst 2025 (oc25) dataset and models for solid-liquid interfaces,” arXiv preprint arXiv:2509.17862, 2025. [Online]. Available: https://arxiv.org/abs/2509.17862 [7] A. Sriram et al., “The open DAC 2023 dataset and challenges for sorbent discovery in direct air capture,” ACS Central Science, vol. 10, no. 5, pp. 923–941, 2024. [8] L. Barroso-Luque et al., “Open materials 2024 (omat24) inorganic materials dataset and models,” 2024. [Online]. Available: https://arxiv.org/abs/2410.12771 [9] D. S. Levine et al., “The open molecules 2025 (OMol25) dataset, evaluations, and models,” 2025. [Online]. Available: https://arxiv.org/abs/2505.08762 [10] D. S. Levine et al., The open polymers 2026 (OPoly26) dataset and evaluations,” arXiv preprint arXiv:2512.23117, 2025. [Online]. Available: https://arxiv.org/abs/2512.23117 [11] K. Khrabrov et al., “Nabla2dft: A universal quantum chemistry dataset of drug-like molecules and a benchmark for neural network potentials,” in NeurIPS 2024 Datasets and Benchmarks Track, 2024. [Online]. Available: https://openreview.net/forum?id=ElUrNM9U8c [12] S. Ganscha et al., “The QCML dataset, quantum chemistry reference data from 33.5M DFT and 14.7B semi-empirical calculations,” Scientific Data, vol. 12, p. 406, 2025. [13] J. Hoja et al., “QM7-X, a comprehensive dataset of quantum-mechanical properties spanning the chemical space of small organic molecules,” Scientific Data, vol. 8, p. 43, 2021. [Online]. Available: https://www.nature.com/articles/s41597-021-00812-2 [14] M. Schreiner et al., “Transition1x - a dataset for building generalizable reactive machine learning potentials,” Scientific Data, vol. 9, p. 779, 2022. The folder "datasets_ADIOS2_format" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "datasets_ADIOS2_format" directory contains 2 sub-directories, one for the version "v1" of the datasets and one for the version "v2" of the datasets. The version "v1" of the datasets provides values of the total energy as they are extracted from the original data as it was released by the respective institutions. The version "v2" of the datasets provides values of the energy that have been realigned. The realignment was performed by training a linear regression model that predicts the total energy as a function of the chemical composition of the atomistic structure, and then subtract such prediction from the original value of the total energy. Both folders "v1" and "v2" contain 16 sub-directories, each corresponding to an ADIOS2-formatted dataset The folder "DeepHyper-results" contains the configurational files and model's parameters for all the 186 HPO trials that were successfully completed by the scalable hyperparameter optimization (HPO) runs on Frontier. The content of the folder "DeepHyper-results" I structured as follows: 1) task-list.txt: list of mpnn name, jobid, and deephyper task id 2) gfm_${MPNN}_${JOBID}_0.${TASKID}: run directory with checkpoint files 3) gfm_${MPNN}: deephyper summary directory (*.csv) for each specific MPNN type 4) deephyper-experiment-${JOBID}: output and error logs for each job The file "deephyper-sorted.csv" contains the details of each HydraGNN model built and tested by HPO, obtained by merging the (*.csv) filed from each HPO run executed. Out of all the HPO trials, we selected 10 to continue the training of the respective HydraGNN models. Due to limited computational budget available in the LRN070 allocation we could not complete the training till convergence for all these 10 selected models. The folder "models" contains multiple sub-folders, one per each HydraGNN model trained. Each model sub-folder contains the parameters of each HydraGNN model, with multiple checkpoint-restarts. The list of sub-folders are as follows: 1) multidataset_hpo-BEST1-fp64 2) multidataset_hpo-BEST2-fp64 3) multidataset_hpo-BEST3-fp64 4) multidataset_hpo-BEST4-fp64 5) multidataset_hpo-BEST5-fp64 6) multidataset_hpo-BEST6-fp64 7) multidataset_hpo-BEST7-fp64 8) multidataset_hpo-BEST8-fp64 9) multidataset_hpo-BEST9-fp64 10) multidataset_hpo-BEST10-fp64 Within each one of these folders, additional auxiliary log files are provided with descriptions about how the training proceeded. The lead PaiNN-model is contained inside "multidataset_hpo-BEST6-fp64". The file "mlp_branch_weights" contains the parameters of the multi-layer perceptron (MLP) used to reconcile the predictions of the 16 output decoding heads of the HydragNN architectures. The MLP takes in input the chemical composition of the atomistic structure and predicts averaging weights to linearly mix the predictions of each output decoding head toward consolidating them into a single one. The folder "1.1billion-structure-inference" contains 1.1 billion atomistic structures randomly generated. Each structures is associated with energy and forces predicted with the lead-PaiNN model combined with the MLP model for reconciliation of the multi-branch predictions generated by the 16 output decoding heads. The folder "1.1billion-structure-inference" contains 9,300 (*.tar.gz) subdirectories, one per Frontier compute node used to execute the inference at exascale. Once uncompressed, each (*.tar.gz) subdirectory contains an ADIOS2 (*.bp) file container, where each atomistic structure is stored as a PyTorch-Geometric Data object. The file "export_dataset_environment_variables.sh" contains the environment variables that need to be set before running the HydraGNN code to reproduce the results provided in this dataset release. The code that can be used to load the ADIOS2 files, load HydraGNN models, and run inference is available at: https://github.com/ORNL/HydraGNN/releases/tag/v5.0

36 MATERIALS SCIENCE

From zonal to nodal capacity expansion planning: Spatial aggregation impacts on a realistic test-case

Solving power system capacity expansion planning (CEP) problems at realistic spatial resolutions is computationally challenging. Thus, a common practice is to solve CEP over zonal models with low spatial resolution rather than over full-scale nodal power networks. Due to improvements in solving large-scale stochastic mixed integer programs, these computational limitations are becoming less relevant, and the assumption that zonal models are realistic and useful approximations of nodal CEP is worth revisiting. Here, this work is the first to conduct a systematic computational study on the assumption that spatial aggregation can reasonably be used for ISO-scale CEP. By considering a realistic, large-scale test network based on the state of California with over 8000 buses, we find that well-designed small spatial aggregations can yield good approximations but that coarser zonal models may result in large distortions of investment decisions, e.g., capacity under-investment of up to 41% for the lowest resolution model considered.

24 POWER TRANSMISSION AND DISTRIBUTION

Distributed quantum approximate optimization algorithm on a quantum-centric supercomputing architecture

Quantum approximate optimization algorithm (QAOA) has shown promise in solving combinatorial optimization problems by providing quantum speedup on near-term gate-based quantum computing systems. However, QAOA faces challenges for high-dimensional problems due to the large number of qubits required and the complexity of deep circuits, limiting its scalability for real-world applications. In this study, we present a distributed QAOA (DQAOA), which leverages distributed computing strategies to decompose a large computational workload into smaller tasks that require fewer qubits and shallower circuits than are necessary to solve the original problem. These sub-problems are processed using a combination of high-performance and quantum computing resources. The global solution is iteratively updated by aggregating sub-solutions, allowing convergence toward the optimal solution. We demonstrate that DQAOA can handle considerably large-scale optimization problems (e.g., 1000-bit problem), achieving a high solution quality and short time-to-solution, outperforming existing strategies. Furthermore, we realize DQAOA on a quantum-centric supercomputing architecture, paving the way for practical applications of gate-based quantum computers in real-world optimization tasks. To extend DQAOA’s applicability to materials science, we further develop an active learning algorithm integrated with our DQAOA (AL-DQAOA), which involves machine learning, DQAOA, and active data production in an iterative loop. We successfully optimize photonic structures using AL-DQAOA, indicating that solving real-world optimization problems using gate-based quantum computing is feasible. We expect the proposed DQAOA to be applicable to a wide range of optimization problems and AL-DQAOA to find broader applications in material design.

Kim, Seongmin [ORNL] (ORCID:0000000159063004)

Pre- and post-processing of cluster galaxies out to 5 × R 200: the extreme case of A2670

ABSTRACT We study galaxy interactions in the large-scale environment around A2670, a massive (M200 = $8.5 \pm 1.2~\times 10^{14} \, \mathrm{{M}_{\odot }}$) and interacting galaxy cluster at z = 0.0763. We first characterize the environment of the cluster out to 5× R200 and find a wealth of substructures, including the main cluster core, a large infalling group, and several other substructures. To study the impact of these substructures (pre-processing) and their accretion into the main cluster (post-processing) on the member galaxies, we visually examined optical images to look for signatures indicative of gravitational or hydrodynamical interactions. We find that ∼21 per cent of the cluster galaxies have clear signs of disturbances, with most of those (∼60 per cent) likely being disturbed by ram pressure. The number of ram-pressure stripping candidates found (101) in A2670 is the largest to date for a single system, and while they are more common in the cluster core, they can be found even at >4 × R200, confirming cluster influence out to large radii. In support of a pre-processing scenario, most of the disturbed galaxies follow the substructures found, with the richest structures having more disturbed galaxies. Post-processing also seems plausible, as many galaxy–galaxy mergers are seen near the cluster core, which is not expected in relaxed clusters. In addition, there is a comparable fraction of disturbed galaxies in and outside substructures. Overall, our results highlight the complex interplay of gas stripping and gravitational interactions in actively assembling clusters up to 5 × R200, motivating wide-area studies in larger cluster samples.

Astronomy & Astrophysics