Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “tensor factorization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

196 records · Page 11

Calibrating the BHB star distance scale and the halo kinematic distance to the Galactic Centre

ABSTRACT We report the first determination of the distance to the Galactic Centre based on the kinematics of halo objects. We apply the statistical-parallax technique to the sample of ∼2500 blue horizontal branch (BHB) stars compiled by Xue et al. to simultaneously constrain the correction factor to the photometric distances of BHB stars as reported by those authors and the distance to the Galactic Centre to find R = 8.2 ± 0.6 kpc. We also find that the average velocity of our BHB star sample in the direction of Galactic rotation, V0 = −240 ± 4 km s−1, is greater by about 20 km s−1 in absolute value than the corresponding velocity for halo RR Lyrae type stars (V0 = −222 ± 4 km s−1) in the Galactocentric distance interval from 6 to 18 kpc, whereas the total (σV) and radial (σr) velocity dispersion of the BHB sample are smaller by about 40–45 km s−1 than the corresponding parameters of the velocity dispersion ellipsoid of halo RR Lyrae type variables. The velocity dispersion tensor of halo BHB stars proved to be markedly less anisotropic than the corresponding tensor for RR Lyrae type variables: the corresponding anisotropy parameter values are equal to βBHB = 0.51 ± 0.02 and βRR = 0.71 ± 0.03, respectively.

Utkin, Nikita D.↗

The Cosmological Bootstrap: Spinning correlators from symmetries and factorization

We extend the cosmological bootstrap to correlators involving massless spinning particles, focusing on spin-1 and spin-2. In de Sitter space, these correlators are constrained both by symmetries and by locality. In particular, the de Sitter isometries become conformal symmetries on the future boundary of the spacetime, which are reflected in a set of Ward identities that the boundary correlators must satisfy. We solve these Ward identities by acting with weight-shifting operators on scalar seed solutions. Using this weight-shifting approach, we derive three- and four-point correlators of massless spin-1 and spin-2 fields with conformally coupled scalars. Four-point functions arising from tree-level exchange are singular in particular kinematic configurations, and the coefficients of these singularities satisfy certain factorization properties. We show that in many cases these factorization limits fix the structure of the correlators uniquely, without having to solve the conformal Ward identities. The additional constraint of locality for massless spinning particles manifests itself as current conservation on the boundary. We find that the four-point functions only satisfy current conservation if the s, t, and u-channels are related to each other, leading to nontrivial constraints on the couplings between the conserved currents and other operators in the theory. For spin-1 currents this implies charge conservation, while for spin-2 currents we recover the equivalence principle from a purely boundary perspective. For multiple spin-1 fields, we recover the structure of Yang--Mills theory. Finally, we apply our methods to slow-roll inflation and derive a few phenomenologically relevant scalar-tensor three-point functions.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

The magnetic gradient scale length explains why certain plasmas require close external magnetic coils

Abstract The separation between the last closed flux surface of a plasma and the external coils that magnetically confine it is a limiting factor in the construction of fusion-capable plasma devices. This plasma-coil separation must be large enough so that components such as a breeding blanket and neutron shielding can fit between the plasma and the coils. Plasma-coil separation affects reactor size, engineering complexity, and particle loss due to field ripple. For some plasmas it can be difficult to produce the desired flux surface shaping with distant coils, and for other plasmas it is infeasible altogether. Here, we seek to understand the underlying physics that limits plasma-coil separation and explain why some configurations require close external coils. In this paper, we explore the hypothesis that the limiting plasma-coil separation is set by the shortest scale length of the magnetic field as expressed by the ∇ B tensor. We tested this hypothesis on a database of > 40 stellarator and tokamak configurations. Within this database, the coil-to-plasma distance compared to the minor radius varies by over an order of magnitude. The magnetic scale length is well correlated to the coil-to-plasma distance of actual coil designs generated using the REGCOIL method (Landreman 2017 Nucl. Fusion 57 046003). Additionally, this correlation reveals a general trend that larger plasma-coil separation is possible with a small number of field periods.

Physics↗

Drell-Yan q T resummation of fiducial power corrections at N 3 LL

We consider Drell-Yan production pp → V*X → LX at small q T << Q, where q T and Q are the total transverse momentum and invariant mass of the leptonic final state L. Experimental measurements require fiducial cuts on L, which in general introduce enhanced, linear power corrections in q T /Q. We show that they can be unambiguously predicted from factorization, and resummed to the same order as the leading-power contribution. For the fiducial q T spectrum, they constitute the complete linear power corrections. We thus obtain predictions for the fiducial q T spectrum to N 3 LL and next-to-leading-power in q T /Q. Matching to full NNLO ($α$$^{2}_{s}$), we find that the linear power corrections are indeed the dominant ones, and once included by factorization, the remaining fixed-order corrections become almost negligible below q T ≲ 40 GeV. We also discuss the implications for more complicated observables, and provide predictions for the fiducial Φ* spectrum at N 3 LL+NNLO. We find excellent agreement with ATLAS and CMS measurements of q T and Φ*. We also consider the $p$$^{ℓ}_{T}$ spectrum. We show that it develops leptonic power corrections in q T /(Q – 2$p$$^{ℓ}_{T}$), which diverge near the Jacobian peak $p$$^{ℓ}_{T}$ ~ Q/2 and must be kept to all powers to obtain a meaningful result there. Doing so, we obtain for the first time an analytically resummed result for the $p$$^{ℓ}_{T}$ spectrum around the Jacobian peak at N 3 LL+NNLO. Our method is based on performing a complete tensor decomposition for hadronic and leptonic tensors. We show that in practice this is equivalent to often-used recoil prescriptions, for which our results now provide rigorous, formal justification. Our tensor decomposition yields nine Lorentz-scalar hadronic structure functions, which for Z/γ* → ℓℓ or W → ℓν directly map onto the commonly used angular coefficients, but also holds for arbitrary leptonic final states. In particular, for suitably defined Born-projected leptons it still yields a LO-like angular decomposition even when including QED final-state radiation. Finally, we also discuss the application to q T subtractions. Including the unambiguously predicted fiducial power corrections significantly improves their performance, and in particular makes them applicable near kinematic edges where they otherwise break down due to large leptonic power corrections.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Electron Heating in 2D Particle-in-cell Simulations of Quasi-perpendicular Low-beta Shocks

Abstract We measure the thermal electron energization in 1D and 2D particle-in-cell simulations of quasi-perpendicular, low-beta ( β p = 0.25) collisionless ion–electron shocks with mass ratio m i / m e = 200, fast Mach number M ms = 1 –4, and upstream magnetic field angle θ Bn = 55°–85° from the shock normal n ˆ . It is known that shock electron heating is described by an ambipolar, B -parallel electric potential jump, Δ ϕ ∥ , that scales roughly linearly with the electron temperature jump. Our simulations have Δ ϕ ∥ / ( 0.5 m i u sh 2 ) ∼ 0.1 –0.2 in units of the pre-shock ions’ bulk kinetic energy, in agreement with prior measurements and simulations. Different ways to measure ϕ ∥ , including the use of de Hoffmann–Teller frame fields, agree to tens-of-percent accuracy. Neglecting off-diagonal electron pressure tensor terms can lead to a systematic underestimate of ϕ ∥ in our low- β p shocks. We further focus on two θ Bn = 65° shocks: a M s = 4 ( M A = 1.8 ) case with a long, 30 d i precursor of whistler waves along n ˆ , and a M s = 7 ( M A = 3.2 ) case with a shorter, 5 d i precursor of whistlers oblique to both n ˆ and B ; d i is the ion skin depth. Within the precursors, ϕ ∥ has a secular rise toward the shock along multiple whistler wavelengths and also has localized spikes within magnetic troughs. In a 1D simulation of the M s = 4 , θ Bn = 65° case, ϕ ∥ shows a weak dependence on the electron plasma-to-cyclotron frequency ratio ω pe /Ω ce , and ϕ ∥ decreases by a factor of 2 as m i / m e is raised to the true proton–electron value of 1836.

Tran, Aaron (ORCID:0000000334834890)↗

Phase-field modeling of orientation-dependent crack growth in ductile single crystals with anisotropic elasticity

Crack growth in ductile single crystals (DuSCs) is orientation dependent due to the anisotropies of crystal plasticity and elastic tensor. This study develops a phase-field model incorporating both crystal plasticity and crack growth and proposes a general method to decompose the elastic energy into compressive and tensile parts to prevent crack growth under compression in the phase-field description. The phase-field model, in combination with three Euler angles, is employed to simulate orientation-dependent crack growth in DuSCs. The contributions from crystal plasticity and anisotropic elasticity are compared, and the former is found to dominate in the anisotropy of crack growth in copper single crystals. Furthermore, the simulation results demonstrate that crystal orientation strongly affects the heterogeneous distribution of plastic strain and the interaction between plastic strain and crack growth. High-throughput phase-field simulations are performed with exhaustive crystal orientations, and the results are explained based on the anisotropy of the Taylor factor.

Computational Solid Mechanics↗

Simulating spin dynamics of supersolid states in a quantum Ising magnet

Motivated by a recent experimental study on the quantum Ising magnet K 2 Co(SeO 3 ) 2 that presented spectroscopic evidence of zero-field supersolidity (Chen et al., arXiv:2402.15869), we simulate the excitation spectrum of the corresponding microscopic XXZ model for the compound, using the recently developed excitation ansatz for infinite projected entangled-pair states. Here, we map out the ground state phase diagram and compute the dynamical spin structure factors across a range of magnetic field strengths, focusing especially on the two supersolid phases found near zero and saturation fields. Our simulated excitation spectra for the zero-field supersolid “Y” phase are in excellent agreement with the experimental data.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Synchronization effects on rest frame energy and momentum densities in the proton

We obtain two-dimensional relativistic densities and currents of energy and momentum in a proton at rest. These densities are obtained at surfaces of fixed light front time, which physically corresponds to using an alternative synchronization convention. Mathematically, this is done using tilted light front coordinates, which consist of light front time and ordinary spatial coordinates. In this coordinate system, all 16 components of the energy-momentum tensor (EMT) obtain clear physical interpretations, and the nine Galilean components reproduce results from standard light front coordinates. We find angular modulations in several densities that are absent in the corresponding instant form results, which are explained as optical effects arising from using fixed light front time when motion is present within the target. Additionally, transversely polarized spin-half targets exhibit an energy dipole moment—which evaluates to −1/4 for all targets if the Belinfante EMT is used, but which is target dependent and vanishes for pointlike fermions if the asymmetric EMT is instead used.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Requirements on the gain calibration for LiteBIRD polarisation data with blind component separation

The detection of primordial B modes of the cosmic microwave background (CMB) could provide information about the early stages of the Universe's evolution. The faintness of this signal requires exquisite calibration accuracy and control of instrumental systematic effects which otherwise could bias the measurements. In this work, we study the impact of an imperfect relative polarisation gain calibration on the recovered value of the tensor-to-scalar ratio r for the LiteBIRD experiment, through the application of the blind Needlet Internal Linear Combination (NILC) foreground-cleaning method. We derive requirements on the relative calibration accuracy of the overall polarisation gain (Δg ν ) for each LiteBIRD frequency channel. Our results show that minimum variance techniques, as NILC, are less sensitive to systematic gain calibration uncertainties compared to a parametric approach, if the latter is not equipped with a proper modelling of these instrumental effects. In this study, the most stringent requirements are found in the channels where the CMB signal is relatively brighter, with the tightest constraints at 166 GHz (Δg ν ≈ 0.16%). This differs from the outcome of an analogous analysis performed with a parametric method, where the tightest requirements are obtained for the foreground-dominated channels. Gain calibration uncertainties, corresponding to the derived requirements, are then simultaneously propagated into all frequency channels. By doing so, we find that the overall impact on estimated r is lower than the total gain systematic budget for LiteBIRD approximately by a factor 5, due to the correlations of the impacts of gain calibration uncertainties in different frequency channels. In order to decouple the systematic effect from the specific choice of the model, we derive the requirements assuming constant spectral parameters for the foreground emission. To assess the robustness of the obtained results against more realistic scenarios, we repeat the analysis assuming sky models of intermediate and high complexity. In these further cases, we adopt an optimised NILC pipeline, called the Multi-Clustering NILC (MC-NILC). We find that the impact of gain calibration uncertainties on r is lower than the LiteBIRD gain systematics budget for the intermediate-complexity sky model. For the high-complexity case, instead, it would be necessary to tighten the requirements by a factor 1.8.

79 ASTRONOMY AND ASTROPHYSICS↗

Polarization angle requirements for CMB B-mode experiments. Application to the LiteBIRD satellite

A methodology to provide the polarization angle requirements for different sets of detectors, at a given frequency of a CMB polarization experiment, is presented. The uncertainties in the polarization angle of each detector set are related to a given bias on the tensor-to-scalar ratio r parameter. The approach is grounded in using a linear combination of the detector sets to obtain the CMB polarization signal. In addition, assuming that the uncertainties on the polarization angle are in the small angle limit (lower than a few degrees), it is possible to derive analytic expressions to establish the requirements. The methodology also accounts for possible correlations among detectors, that may originate from the optics, wafers, etc. The approach is applied to the LiteBIRD space mission. We show that, for the most restrictive case (i.e., full correlation of the polarization angle systematics among detector sets), the requirements on the polarization angle uncertainties are of around 1 arcmin at the most sensitive frequency bands (i.e., ≈ 150 GHz) and of few tens of arcmin at the lowest (i.e., ≈ 40 GHz) and highest (i.e., ≈ 400 GHz) observational bands. Conversely, for the least restrictive case (i.e., no correlation of the polarization angle systematics among detector sets), the requirements are ≈ 5 times less restrictive than for the previous scenario. At the global and the telescope levels, polarization angle knowledge of a few arcmins is sufficient for correlated global systematic errors and can be relaxed by a factor of two for fully uncorrelated errors in detector polarization angle. The reported uncertainty levels are needed in order to have the bias on r due to systematics below the limit established by the LiteBIRD collaboration.

79 ASTRONOMY AND ASTROPHYSICS↗

Dimerization and spin decoupling in a two-leg Heisenberg ladder with frustrated trimer rungs

We study the antiferromagnetic spin-half Heisenberg ladder in the presence of an additional frustrating rung spin which is motivated and relevant also for the description of real two-dimensional materials such as the two-dimensional trimer magnet Ba 4 Ir 3 O 10 . We study the zero-temperature phase diagram, where we combine numerical and analytical methods into an overall consistent description. All numerical simulations are also accompanied by studies of the dynamical spin structure factor obtained via the density matrix renormalization group. Overall, we find in the regime of strong rung coupling a gapped dimerized phase related to competing symmetry sectors in Hilbert space that ultimately results in frustration-driven spin-Peierls transition. In the weak rung-coupling regime, the system is uniform, yet shows a gapped spinon continuum together with a sharp coherent low-energy branch which renders the system critical overall. In either case, the additional rung spin quickly get sidelined and nearly decouple in the regime when their bare coupling to the ladder is weakened

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

BICEP Array: 150 GHz Detector Module Development

The Background Imaging of Cosmic Extragalactic Polarization (BICEP)/Keck (BK) collaboration is currently leading the quest for the highest-sensitivity measurements of the polarized cosmic microwave background (CMB) anisotropies on a degree scale with a series of cryogenic telescopes, of which BICEP Array (BA) is the latest Stage-3 upgrade with a total of ~ 32,000 detectors. The instrument comprises 4 receivers spanning 30–270 GHz, with the low-frequency 30/40 GHz deployed to the South Pole Station in late 2019. The full complement of receivers is forecast to set the most stringent constraints on the tensor-to-scalar ratio r. Building on these advances, the overarching small-aperture telescope concept is already being used as the reference for further Stage-4 experiment design. This paper describes the development of the BICEP Array 150 GHz detector module and its fabrication requirements, with highlights on the high-density time division multiplexing (TDM) design of the cryogenic circuit boards. The low-impedance wiring required between the detectors and the frst stage of superconducting quantum interference device amplifers is crucial to maintaining a stable bias current on the detectors. Here, a novel multilayer FR4 Printed Circuit Board with superconducting traces, capable of reading out up to 648 detectors, is detailed along with its validation tests. An ultra-high-density TDM detector module concept we developed for a CMB-S4-like experiment that allows up to 1920 detectors to be read out is also presented. TDM has been chosen as the detector readout technology for the Cosmic Microwave Background Stage-4 (CMB-S4) experiment based on its proven low-noise performance, predictable costs, and overall maturity of the architecture. The heritage for TDM is rooted in mm- and sub-mm-wave experiments dating back 20 years and has since evolved to support a multiplexing factor of 64x in Stage-3 experiments.

79 ASTRONOMY AND ASTROPHYSICS↗

Partons as unique ground states of quantum Hall parent Hamiltonians: The case of Fibonacci anyons

We present microscopic, multiple Landau level, (frustration-free and positive semi-definite) parent Hamiltonians whose ground states, realizing different quantum Hall fluids, are parton-like and whose excitations display either Abelian or non-Abelian braiding statistics. We prove ground state energy monotonicity theorems for systems with different particle numbers in multiple Landau levels, demonstrate S-duality in the case of toroidal geometry, and establish complete sets of zero modes of special Hamiltonians stabilizing parton-like states, specifically at filling factor \nu=2/3 ν = 2 / 3 . The emergent Entangled Pauli Principle (EPP), introduced in [Phys. Rev. B 98, 161118(R) (2018)] and which defines the “DNA” of the quantum Hall fluid, is behind the exact determination of the topological characteristics of the fluid, including charge and braiding statistics of excitations, and effective edge theory descriptions. When the closed-shell condition is satisfied, the densest (i.e., the highest density and lowest total angular momentum) zero-energy mode is a unique parton state. We conjecture that parton-like states generally span the subspace of many-body wave functions with the two-body M M -clustering property within any given number of Landau levels, that is, wave functions with M M th-order coincidence plane zeroes and both holomorphic and anti-holomorphic dependence on variables. General arguments are supplemented by rigorous considerations for the M=3 M = 3 case of fermions in four Landau levels. For this case, we establish that the zero mode counting can be done by enumerating certain patterns consistent with an underlying EPP. We apply the coherent state approach of [Phys. Rev. X 1, 021015 (2011)] to show that the elementary (localized) bulk excitations are Fibonacci anyons. This demonstrates that the DNA associated with fractional quantum Hall states encodes all universal properties. Specifically, for parton-like states, we establish a link with tensor network structures of finite bond dimension that emerge via root level entanglement.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Correlating Protein Dynamics and Catalytic Activity of a Model Hydrogenase Using Paramagnetic and Biological Nuclear Magnetic Resonance Spectroscopy

Rational catalyst design remains a significant challenge, with electronic structure, steric, and electrostatic effects known to contribute to activity. Recently, dynamics has been recognized as another factor that impacts catalysis, though identifying and predicting these effects has remained out of reach. Nickel-substituted rubredoxin (NiRd), a protein-based mimic of a hydrogenase enzyme, serves as a model catalytic system in which dynamics can be systematically investigated with respect to activity. While over 30 secondary-sphere mutants of NiRd have been shown to be catalytically active, no significant correlation was observed between the rates and catalytic overpotential or electronic structure, prompting questions about the protein-derived factors that modulate activity. Here, in this work, NMR spectroscopy was used to investigate the roles of substrate accessibility, protein dynamics, and protein stability in controlling catalysis. Significant paramagnetic effects from the nickel center (S = 1) isolate the methylene proton resonances of the metal-coordinating cysteine residues. The sensitivity of resonance positions and linewidths to local environment offers an opportunity to study dynamical molecular changes around the metal center with high resolution. Machine learning algorithms were employed to identify correlations between the catalytic activity and the paramagnetic NMR spectra. These analyses revealed spectroscopic features of specific cysteine protons that report on catalytic overpotential and increased turnover rates, which are further supported by the results obtained using high-field NMR techniques. Collectively, these studies indicate the potential for multifrequency NMR techniques to resolve key contributors to catalytic activity and highlight the importance of local and outer-sphere dynamics.

Protein Engineering↗

Geothermal Fault Zone and Fluid Imaging through Joint Airborne ZTEM and Ground MT Data Inversion Analysis

This project has aimed to achieve detailed electrical resistivity resolution at geothermal reservoir scales by combining airborne natural electromagnetic (EM) field surveying (ZTEM) with ground magnetotelluric (MT) measurements to approximate an airborne MT geophysical method. MT alone is relatively expensive and may have permitting challenges in sensitive areas. Airborne ZTEM field data contains only the magnetic field, requires a background assumption, and has been limited to relatively high frequencies, thus suffering uniqueness problems. Based on proto-type 2D simulations, ZTEM ambiguities may be reduced through formal incorporation with possibly sparse ground MT soundings, which we pursued in full 3D for this project. The methodology was tested at the high-temperature Roosevelt Hot Springs geothermal system, Utah, which was considered advantageous given the near total exposure of crystalline reservoir rocks across the project area. ZTEM and ground MT survey data were acquired in 2017, subcontracted to outside parties with which we have worked in the past. These included 80 remote-referenced tensor MT soundings over the Mineral Mountains and adjacent Roosevelt Hot Spring producing geothermal system. These MT stations abut later coverage of a similar number of MT stations taken for the Utah FORGE project providing excellent total data aperture to re-solve structure beneath both project areas better than either set alone. The airborne ZTEM survey covered 704 line kilometers in E-W flight lines with a 250 m line spacing. Although this survey was timed during a maintenance-related shutdown of power production at the Roosevelt Hot Springs, other noise sources difficult to identify but including two high-voltage state-scale transmission lines compromised the ZTEM survey badly leading to unusable responses. Thus, with DOE management concurrence, the project proceeded to emphasize inversion and interpretation of the joint SubTER-FORGE MT data sets with regard to the Roosevelt Hot Springs reservoir recharge and to deep heat sources for both it and the Utah FORGE EGS project area. We also investigated the joint ZTEM-MT sampling concept with data sets from the Eleven Mile Canyon prospect area donated by the U.S. Navy (A. Sabin, PoC). Inversion of the SubTER-FORGE MT data using the HexMT 3D finite element algorithm reveals a large, low-resistivity anomaly extending sub-vertically through the depth range of the crust beneath the western Mineral Mountains. The steep conductive zone connects in the lower crust to a more tabular conductor characteristic of much of the Great Basin that generally is ascribed to current mafic magmatic underplating, hybridization and fluid release. The location of the resolved anomaly relative to the recent (0.5-0.8 Ma) eruptive centers of the Mineral Mountains implicates it as remnants of the magma body which fed these centers. This structure appears to be currently feeding heat and fluids upward into the Roosevelt Hot Springs hydrothermal system, as well as heat laterally to the FORGE project area. Separate and joint inversion models were carried out for the donated Eleven Mile Canyon MT-ZTEM data set to demonstrate concept. ZTEM only inversion showed two main alteration zones in the western portion of the project area known from geological mapping. Joint inversion including an E-W profile of MT soundings sharpened these features considerably. It also resolved in much greater detail the graben related normal faulting structure of the central project area which lies at depths exceeding the sensitivity of ZTEM alone. The sparse number of MT da-ta relative to the ZTEM required upweighting the former by a factor of several, but an exact procedure awaits future research. Our final impression is that sparse MT data can improve resolution of the subsurface over that of ZTEM alone. However, well sampled MT data are to be preferred and offer the simplicity of interpreting just one data type, and possess the superior resolution capability coming with the electric field everywhere, and from their high bandwidth.

15 GEOTHERMAL ENERGY↗

HydraGNN_Predictive_GFM_2024 - Ensemble of predictive graph foundation models for ground state atomistic materials modeling

We provide the ensemble of fifteen pre-trained graph foundation models (GFMs) for atomistic materials modeling applications. Each one of the fifteen GFMs has been trained on five open-source datasets that (once aggregated) amount to over 154 million atomistic structures, which cover over two-thirds of the natural elements of the periodic table and that comprises a broad set of organic and inorganic compounds. This vast set of atomistic structures comprises ground state configurations that are dynamically stable (i.e., equilibrated structures with atomic forces approximately close to zero values) as well as dynamically unstable structures (i.e., non-equilibrium structures with non-negligible non-zero values of atomic forces). The ensemble of datasets aggregated does NOT include excited states. The datasets have been curated to remove atomistic structures with spectral norm of the force tensor above 100 eV/angstrom. Moreover, a linear term of the energy was computed for each dataset using a linear regression model that uses the chemical concentration of each natural element as regressor. The linear term predicted by the linear regression model has been subtracted from each original energy value to perform a re-alignment of the energy values across different electronic structures approximation theories performed to generate the diverse multi-source, multi-fidelity datasets. The folder "ADIOS_files" contains the set of pre-processed datasets in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used for the development and training of GFMs in this work. The "ADIOS_files" directory contains 6 sub-directories named as follows: - ANI1x-v3.bp - MPTrj-v3.bp - OC2020-20M-v3.bp - OC2020-v3.bp - OC2022-v3.bp - qm7x-v3.bp Each sub-directory contains the pre-processed datasets converted in Adaptable I/O System (ADIOS) format (https://www.exascaleproject.org/research-project/adios/) that have been used to the development, training, and performance testing of the ensemble go predictive graph foundation models. Each GFM was developed using HydraGNN (https://github.com/ORNL/HydraGNN) as underlying graph neural network (GNN) architecture. The multi-task learning (MTL) capability of HydraGNN was used to simultaneously train the GFMs on labeled values for direct predictions of energy (a total system property of an atomistic structure that measures the chemical stability) and atomic forces (an atomic level property of an atomistic structure that measures the dynamical stability). The hyper parameters of the GFM have been tuned using scalable hyperparameter optimization (HPO) algorithms implemented in the software DeepHyper (https://github.com/deephyper/deephyper). The pre-training of each HPO trial was performed using distributed data parallelism (DDP) to scale the training across 128 compute nodes of the exascale OLCF supercomputer Frontier. Each HPO trial was trained only for 10 epochs and an early stopping was performed to avoid wasting significant computational resources on GNN architectures that were clearly underperforming. For each HPO trial, the 'omnistat' tool developed by (AMD Research - Advanced Micro Device) was used to measure the total energy consumption in kWh. The ensemble of GFMs was obtained by selecting the fifteen best performing HPO trials. Four models have been selected for their clear advantage in accuracy, and these are the GFMs with IDs 229, 156, 147, 260. Additional eleven models have been selected based on judicious balance between accuracy and energy consumption needed for training, and these are the GFMs with IDs 165, 78, 137, 1, 175, 171, 181, 67, 179, 167, 351. Each selected GFM of the ensemble was continued to cumulate a total of at most 30 epochs. In some cases, the total number of epochs actually performed was les than 30 due to two combined factors: (1) the size of the GFM (i.e., the number of model parameters to train) and (2) the total wall-clock time for which the computational resources could be allocated on OLCF-Frontier. The "Ensemble_of_models" directory contains 15 sub-directories named as follows: - gfm_0.229 - gfm_0.156 - gfm_0.147 - gfm_0.260 - gfm_0.165 - gfm_0.78 - gfm_0.137 - gfm_0.1 - gfm_0.175 - gfm_0.171 - gfm_0.181 - gfm_0.67 - gfm_0.179 - gfm_0.167 - gfm_0.351 Each one of these sub-directories refers to one of the fifteen HPO trials that have been selected to continue the pre-training with at most 30 epochs. With each sub-directory associated with a specific HPO trial, the following files can be found: - config.json: file for argument parsing to develop and train an HydraGNN architecture - gfm_0.ID_epoch_N.pk: file with model parameters for HPO ID trial after N epochs of training The ensemble of fifteen GFM architectures was used for (1) ensemble averaging to stabilize the predictions of energy and atomic forces after pre-training for post-processing analysis and (2) ensemble uncertainty quantification (UQ). The code used to develop, pre-train, and load the pre-trained models for post-processing analysis is available on the ORNL-GitHub at the following link: https://github.com/ORNL/HydraGNN/tree/Predictive_GFM_2024

36 MATERIALS SCIENCE↗