Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Statistical modeling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Predicting battery capacity from impedance at varying temperature and state of charge using machine learning

Prediction of battery health from electrochemical impedance spectroscopy (EIS) data can enable rapid measurement of battery state in real-world applications without using additional sensors or time-consuming performance measurements. However, deconvoluting the effect of capacity, state of charge, and temperature on EIS response is complicated analytically. Here, various machine-learning models, such as linear, Gaussian process, random forest, and artificial neural network regression, are utilized to predict capacity from EIS using hundreds of capacity, direct current (DC) resistance, and EIS measurements recorded under varying conditions of health, temperature, and state of charge (SOC). Several feature extraction and selection methods from traditional electrochemical analysis and statistical modeling are explored using machine-learning pipelines. EIS data from just two frequencies can accurately predict capacity, and interrogation shows that the optimal set of frequencies is not usually intuitive. Best results are achieved with an ensemble model, which predicts battery capacity with a mean absolute error of 1.9% on data from unobserved cells.

25 ENERGY STORAGE↗

Considering uncertainties expands the lower tail of maize yield projections

Crop yields are sensitive to extreme weather events. Improving the understanding of the mechanisms and the drivers of the projection uncertainties can help to improve decisions. Previous studies have provided important insights, but often sample only a small subset of potentially important uncertainties. Here we expand on a previous statistical modeling approach by refining the analyses of two uncertainty sources. Specifically, we assess the effects of uncertainties surrounding crop-yield model parameters and climate forcings on projected crop yield. We focus on maize yield projections in the eastern U.S.in this century. We quantify how considering more uncertainties expands the lower tail of yield projections. We characterized the relative importance of each uncertainty source and show that the uncertainty surrounding yield model parameters is the main driver of yield projection uncertainty.

59 BASIC BIOLOGICAL SCIENCES↗

New 26 P( p, γ ) 27 S Thermonuclear Reaction Rate and Its Astrophysical Implications in the rp -process

Accurate nuclear reaction rates for 26 P(p, γ) 27 S are pivotal for a comprehensive understanding of the rp-process nucleosynthesis path in the region of proton-rich sulfur and phosphorus isotopes. However, large uncertainties still exist in the current rate of 26 P(p, γ) 27 S because of the lack of nuclear mass and energy level structure information for 27 S. We reevaluate this reaction rate using the experimentally constrained 27 S mass, together with the shell model predicted level structure. It is found that the 26 P(p, γ) 27 S reaction rate is dominated by a direct capture reaction mechanism despite the presence of three resonances at E = 1.104, 1.597, and 1.777 MeV above the proton threshold in 27 S. The new rate is overall smaller than the other previous rates from the Hauser–Feshbach statistical model by at least 1 order of magnitude in the temperature range of X-ray burst interest. In addition, we consistently update the photodisintegration rate using the new 27 S mass. The influence of new rates of forward and reverse reaction in the abundances of isotopes produced in the rp-process is explored by postprocessing nucleosynthesis calculations. The final abundance ratio of 27 S/ 26 P obtained using the new rates is only 10% of that from the old rate. The abundance flow calculations show that the reaction path 26 P(p, γ) 27 S(β + ,ν) 27 P is not as important as previously thought for producing 27 P. The adoption of the new reaction rates for 26 P(p, γ) 27 S only reduces the final production of aluminum by 7.1% and has no discernible impact on the yield of other elements.

79 ASTRONOMY AND ASTROPHYSICS↗

Power Electronics Materials and Bonded Interfaces - Reliability and Lifetime

High temperature operation of wide bandgap devices continue to be a challenge for the power electronics packages. Thermal performance and reliability are important factors that determine the viability of a bonded interface for operation at high temperatures. In this presentation, we present the technical approach and key results from the research on sintered silver, transient liquid phase alloy, and polymeric materials. A lifetime prediction model that incorporates the thermomechanical behavior of sintered silver at 200C was developed. The copper-aluminum transient alloy completed 350 thermal cycles from -40C to 200C and little increase in the defect level was observed. In addition to material research, we initiated a time-series analysis on the scanning acoustic microscope images of eutectic solder to explore statistical forecasting methods and machine learning techniques. Initial results that report the accuracy of a few different statistical models are presented.

ADVANCED PROPULSION SYSTEMS↗

Research Needs for Trusted Analytics in National Security Settings

As artificial intelligence, machine learning, and statistical modeling methods become commonplace in national security applications, the drive to create trusted analytics becomes increasingly important. The goal of this report is to identify areas of research that can provide the foundational understanding and technical prerequisites for the development and deployment of trusted analytics in national security settings. Our review of the literature covered several disjoint research communities, including computer science, statistics, human factors, and several branches of psychology and cognitive science, which tend not to interact with one another or cite each other's literatures. As a result, there exists no agreed-upon theoretical framework for understanding how various factors influence trust and no well-established empirical paradigm for studying these effects. This report therefore takes three steps. First, we define several key terms in an effort to provide a unifying language for trusted analytics and to manage the scope of the problem. Second, we outline an empirical perspective that identifies key independent, moderating, and dependent variables in assessing trusted analytics. Though not a substitute for a theoretical framework, the empirical perspective does support research and development of trusted analytics in the national security domain. Finally, we discuss several research gaps relevant to developing trusted analytics for the national security mission space.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Efficient method for estimation of fission fragment yields of $\textit{r}$-process nuclei

Background: More than half of all the elements heavier than iron are made by the rapid neutron capture process (or $\textit{r}$ process). For very-neutron-rich astrophysical conditions, such at those found in the tidal ejecta of neutron stars, nuclear fission determines the $\textit{r}$-process endpoint, and the fission-fragment yields shape the final abundances of 110 ≤ $\textit{A}$ ≤ 170 nuclei. The knowledge of fission-fragment yields of hundreds of nuclei inhabiting very-neutron-rich regions of the nuclear landscape is thus crucial for the modeling of heavy-element nucleosynthesis. Purpose: In this study, we propose a model for the fast calculation of fission-fragment yields based on the concept of shell-stabilized prefragments defined with help of the nucleonic localization functions. Methods: To generate realistic potential-energy surfaces and nucleonic localizations, we apply Skyrme density-functional theory. In this work, the distribution of the neck nucleons among the two prefragments is obtained by means of a statistical model. Results: We benchmark the method by studying the fission yields of 178 Pt, 240 Pu, 254 Cf, and 254,256,258 Fm and show that it satisfactorily explains the experimental data. We then make predictions for 254 Pu and 290 Fm as two representative cases of fissioning nuclei that are expected to significantly contribute during the $\textit{r}$-process nucleosynthesis occurring in neutron-star mergers. Conclusions: The proposed framework provides an efficient alternative to microscopic approaches based on the evolution of the system in a space of collective coordinates all the way to scission. It can be used to carry out global calculations of fission-fragment distributions across the $\textit{r}$-process region.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Leveraging Triphenylphosphine-Containing Polymers to Explore Design Principles for Protein-Mimetic Catalysts

Complex interactions between noncoordinating residues are significant yet commonly overlooked components of macromolecular catalyst function. While these interactions have been demonstrated to impact binding affinities and catalytic rates in metalloenzymes, the roles of similar structural elements in synthetic polymeric catalysts remain underexplored. Using a model Suzuki–Miyuara cross-coupling reaction, we performed a series of systematic studies to probe the interconnected effects of metal–ligand cross-links, electrostatic interactions, and local rigidity in polymer catalysts. To achieve this, a novel bifunctional triphenylphosphine acrylamide (BisTPPAm) monomer was synthesized and evaluated alongside an analogous monofunctional triphenylphosphine acrylamide (TPPAm). In model copolymer catalysts, increased initial reaction rates were observed for copolymers untethered by Pd complexation (BisTPPAm-containing) compared to Pd-cross-linked catalysts (TPPAm-containing). Further, incorporating local rigidity through secondary structure-like and electrostatic interactions revealed nonmonotonic relationships between composition and the reaction rate, demonstrating the potential for tunable behavior through secondary-sphere interactions. Lastly, through rigorous cheminformatics featurization strategies and statistical modeling, we quantitated relationships between chemical descriptors of the substrate and reaction conditions on catalytic performance. Collectively, these results provide insights into relationships among the composition, structure, and function of protein-mimetic catalytic copolymers.

Catalysts↗

Partial Least Squares, Experimental Design, and Near-Infrared Spectrophotometry for the Remote Quantification of Nitric Acid Concentration and Temperature

Near-infrared spectrophotometry and partial least squares regression (PLSR) were evaluated to create a pleasantly simple yet effective approach for measuring HNO3 concentration with varying temperature levels. A training set, which covered HNO3 concentrations (0.1–8 M) and temperature (10–40 °C), was selected using a D-optimal design to minimize the number of samples required in the calibration set for PLSR analysis. The top D-optimal-selected PLSR models had root mean squared error of prediction values of 1.4% for HNO 3 and 4.0% for temperature. The PLSR models built from spectra collected on static samples were validated against flow tests including HNO 3 concentration and temperature gradients to test abnormal conditions (e.g., bubbles) and the model performance between sample points in the factor space. Based on cross-validation and prediction modeling statistics, the designed near-infrared absorption approach can provide remote, quantitative analysis of HNO 3 concentration and temperature for production-oriented applications in facilities where laser safety challenges would inhibit the implementation of other optical techniques (e.g., Raman spectroscopy) and in which space, time, and/or resources are constrained. The experimental design approach effectively minimized the number of samples in the training set and maintained or improved PLSR model performance, which makes the described chemometric approach more amenable to nuclear field applications.

07 ISOTOPE AND RADIATION SOURCES↗

Comparison of All Solid Cancer Mortality and Incidence Dose-Response in the Life Span Study of Atomic Bomb Survivors, 1958–2009

Recent analysis of all solid cancer incidence (1958–2009) in the Life Span Study (LSS) revealed evidence of upward curvature in the radiation dose response among males but not females. Upward curvature in sex-averaged excess relative risk (ERR) for all solid cancer mortality (1950–2003) was also observed in the 0–2 Gy dose range. As reasons for non-linearity in the LSS are not completely understood, we conducted dose response analyses for all solid cancer mortality and incidence applying similar methods (1958–2009 follow-up, DS02R1 doses, including subjects notin-city (NIC) at the time of the bombing) and statistical models. Incident cancers were ascertained from Hiroshima and Nagasaki cancer registries, while cause of death was ascertained from death certificates over entire Japan. The study included 105,444 LSS subjects who were alive and not known to have cancer before Jan 1, 1958 (80,205 with dose estimates and 25,239 NIC subjects). Between 1958 and 2009, there were 3.1 million person-years (PY) and 22,538 solid cancers for incidence analysis and 3.8 million PY and 15,419 solid cancer deaths for mortality analysis. We fitted sex-specific ERR models adjusted for smoking to both types of data. Over the entire range of doses, solid cancer mortality dose response exhibited a borderline significant upward curvature among males (P=0.062) and significant upward curvature among females (P=0.010); for solid cancer incidence, as before, we found a significant upward curvature among males (P=0.001) but not among females (P=0.624). The sex difference in magnitude of dose response curvature was statistically significant for cancer incidence (P=0.017) but not for cancer mortality (P=0.781). The results of analyses in the 0–2 Gy range and restricted lower dose ranges generally supported inferences made about the sex-specific dose response shape over the entire range of doses for each outcome. Patterns of sex-specific curvature by calendar period (1958–1987 vs 1988–2009) and age at exposure (0–19 vs 20–83) varied between mortality and incidence data, particularly among females, although for each outcome there was an indication of curvature among 0–19 year old male survivors in both calendar periods and among 0–19 year old female survivors in the recent period. Collectively, our findings indicate that the upward curvature in all solid cancer dose response in the LSS is neither specific to males nor to incidence data; it appears to depend on composition of case series and age at exposure or time. Further follow-up and site-specific analyses of cancer mortality and incidence will be important to confirm the emerging trend in dose response curvature among young survivors and unveil the contributing factors and sites.

63 RADIATION, THERMAL, AND OTHER ENVIRON. POLLUTAN↗

Detection Limits of Low-mass, Long-period Exoplanets Using Gaussian Processes Applied to HARPS-N Solar Radial Velocities

Radial velocity (RV) searches for Earth-mass exoplanets in the habitable zone around Sun-like stars are limited by the effects of stellar variability on the host star. In particular, suppression of convective blueshift and brightness inhomogeneities due to photospheric faculae/plage and starspots are the dominant contribution to the variability of such stellar RVs. Gaussian process (GP) regression is a powerful tool for statistically modeling these quasi-periodic variations. We investigate the limits of this technique using 800 days of RVs from the solar telescope on the High Accuracy Radial velocity Planet Searcher for the Northern hemisphere (HARPS-N) spectrograph. These data provide a well-sampled time series of stellar RV variations. Into this data set, we inject Keplerian signals with periods between 100 and 500 days and amplitudes between 0.6 and 2.4 m s{sup −1}. We use GP regression to fit the resulting RVs and determine the statistical significance of recovered periods and amplitudes. We then generate synthetic RVs with the same covariance properties as the solar data to determine a lower bound on the observational baseline necessary to detect low-mass planets in Venus-like orbits around a Sun-like star. Our simulations show that discovering planets with a larger mass (∼0.5 m s{sup −1}) using current-generation spectrographs and GP regression will require more than 12 yr of densely sampled RV observations. Furthermore, even with a perfect model of stellar variability, discovering a true exo-Venus (∼0.1 m s{sup −1}) with current instruments would take over 15 yr. Therefore, next-generation spectrographs and better models of stellar variability are required for detection of such planets.

47 OTHER INSTRUMENTATION↗

Assessing the Influence of Climate on the Spatial Pattern of West Nile Virus Incidence in the United States

West Nile virus (WNV) is the leading cause of mosquito-borne disease in humans in the United States. Since the introduction of the disease in 1999, incidence levels have stabilized in many regions, allowing for analysis of climate conditions that shape the spatial structure of disease incidence. Our goal was to identify the seasonal climate variables that influence the spatial extent and magnitude of WNV incidence in humans. We developed a predictive model of contemporary mean annual WNV incidence using U.S. county-level case reports from 2005 to 2019 and seasonally averaged climate variables. We used a random forest model that had an out-of-sample model performance of R 2 =0.61. Our model accurately captured the V-shaped area of higher WNV incidence that extends from states on the Canadian border south through the middle of the Great Plains. It also captured a region of moderate WNV incidence in the southern Mississippi Valley. The highest levels of WNV incidence were in regions with dry and cold winters and wet and mild summers. The random forest model classified counties with average winter precipitation levels <23.3 mm/month as having incidence levels over 11 times greater than those of counties that are wetter. Among the climate predictors, winter precipitation, fall precipitation, and winter temperature were the three most important predictive variables. We consider which aspects of the WNV transmission cycle climate conditions may benefit the most and argued that dry and cold winters are climate conditions optimal for the mosquito species key to amplifying WNV transmission. Our statistical model may be useful in projecting shifts in WNV risk in response to climate change.

60 APPLIED LIFE SCIENCES↗

Calibration of the Diffusivity Predictions of Centipede Using Approximate Bayesian Computation and Applications in Nyx (Engineering Scale) and Xolotl-MARMOT (Meso-Scale) Simulations

Fission gas evolution and release in UO 2 nuclear fuel are important fuel performance metrics and occur in several distinct stages: 1) nucleation, growth and resolution of intra-granular bubbles, 2) diffusion to grain boundaries and 3) nucleation and growth of bubbles at grain boundaries, which eventually form a connected network (percolation) enabling release of gas from grain boundaries through connections to triple junctions, grain edges or free surfaces. The NE-SciDAC project is developing several computational tools to model this problem, which are connected in a hierarchical multi-scale framework. The information transfer in the multi-scale framework is a critical step that, in addition to best-estimates, should include uncertainty quantification. Despite taking a first-principles multi-scale approach, there is a need to perform parameter calibration to ensure consistency with available experimental data. In the present study, uncertainty quantification (UQ) and parameter calibration is demonstrated for one of the lower length scale codes in the multi-scale framework (Centipede) and then the results, including instances of the propagated uncertainties, are used in other codes within the framework, specifically Nyx and Xolotl-MARMOT. We calibrated the model parameters in Centipede, a computer code used to predict diffusivities of uranium (U) and xenon (Xe) in the context of the simulation of fission gas in uranium oxide (UO 2 ) nuclear fuel. The Centipede code depends on 183 parameters, all of which are subject to uncertainty. The three data sets used in our calibration effort are taken from the literature. This data is available as a set of measurements, including measurement errors. Our goal is to calibrate a statistical model that predicts both the value of the measurement and the uncertainty associated with the measurement. We perform a Bayesian calibration of the model parameters using a dedicated approximate Bayesian computation (ABC) likelihood function. To avoid excessive computational costs, we replace the expensive Centipede simulation code by a higher-order surrogate model, constructed using only the 9 most important parameters. These important parameters are identified by a preliminary global sensitivity analysis (GSA) study. Among the important parameters are T0 (the temperature at which UO 2 is perfectly stoichiometric) and Hf_pO2 (the temperature dependence of the oxygen (O) partial pressure) that should be considered as operating conditions to be estimated along with the other parameters. We consider two different cases: one where we define one set of these operating conditions for all data sets, and one where we define distinct operating condition parameters for each data set. The Xe diffusivities predicted by the latter case show distinct features that could not be observed in the former. Next, we use the diffusivity predictions by Centipede as input to Nyx, a reduced order fuel performance code focused on gas behavior alone, in order to estimate quantities associated with inter-granular bubble formation at conditions specified by the experiments. Finally, the diffusivities obtained from the calibrated Centipede runs were used in coupled Xolotl-MARMOT simulations of intra- and inter-granular gas evolution. The results are compared to simulations using the baseline diffusivities from Turnbull et al.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Lethality is Local, but Survival is Systemic: Temporal and Multi-Organ Responses to Chlorine Gas Exposure in a Murine Model

Chlorine gas (Cl2) is a highly toxic chemical associated with both localized lung injury and systemic health effects. While pulmonary damage has been well characterized, the systemic inflammatory and metabolic responses remain poorly understood. We aimed to define the temporal and multi-organ responses to Cl2 exposure in a murine model, with a focus on identifying spatiotemporal inflammation and its impact on survival and lethality. SKH1 mice were exposed for 10 min to varying concentrations of Cl2 (94.4–810 ppm, representative of non-lethal, LD10, and LD50 doses) and monitored for respiratory function, perfusion, and acidosis using organ-specific imaging. At multiple time points (40 min, 6 h, 24 h, and 7 d), we measured phosphoproteins, cytokines, chemokines, growth factors, and metabolic hormones in the lungs, heart, cortex, and plasma. Statistical modeling and logistic regression were used to identify biomarkers associated with lethality and survival. We found that lung injury was the primary cause of potential lethality, particularly via early phosphoprotein signaling disruptions. However, survival correlated with early systemic coordination of inflammatory and metabolic signals across organs. Perfusion and acidosis imaging were strongly associated with chemokine and hormone responses. Key survival-associated plasma biomarkers included decreased insulin, increased ghrelin, and decreased eotaxin. While potential lethality from Cl2 exposure is locally driven by pulmonary injury, survival depends on systemic, multi-organ responses that occur rapidly post-exposure. Within this model, our findings identify a potential therapeutic window to enhance survival and suggest candidate biomarkers that may be explored translationally for both triage and treatment of chlorine-related incidents.

chlorine gas↗

Probabilistic Forward Modeling of Galaxy Catalogs with Normalizing Flows

Abstract Evaluating the accuracy and calibration of the redshift posteriors produced by photometric redshift (photo- z ) estimators is vital for enabling precision cosmology and extragalactic astrophysics with modern wide-field photometric surveys. Evaluating photo- z posteriors on a per-galaxy basis is difficult, however, as real galaxies have a true redshift but not a true redshift posterior. We introduce PZFlow, a Python package for the probabilistic forward modeling of galaxy catalogs with normalizing flows. For catalogs simulated with PZFlow, there is a natural notion of “true” redshift posteriors that can be used for photo- z validation. We use PZFlow to simulate a photometric galaxy catalog where each galaxy has a redshift, noisy photometry, shape information, and a true redshift posterior. We also demonstrate the use of an ensemble of normalizing flows for photo- z estimation. We discuss how PZFlow will be used to validate the photo- z estimation pipeline of the Dark Energy Science Collaboration, and the wider applicability of PZFlow for statistical modeling of any tabular data.

Astronomy & Astrophysics↗

Learning-based approach to plasticity in athermal sheared amorphous packings: Improving softness

The plasticity of amorphous solids undergoing shear is characterized by quasi-localized rearrangements of particles. While many models of plasticity exist, the precise relationship between the plastic dynamics and the structure of a particle’s local environment remains an open question. Previously, machine learning was used to identify a structural predictor of rearrangements called “softness.” Although softness has been shown to predict which particles will rearrange with high accuracy, the method can be difficult to implement in experiments where data are limited and the combinations of descriptors it identifies are often difficult to interpret physically. Here, we address both of these weaknesses, presenting two major improvements to the standard softness method. First, we present a natural representation of each particle’s observed mobility, allowing for the use of statistical models that are both simpler and provide greater accuracy in limited datasets. Second, we employ persistent homology as a systematic means of identifying simple, topologically informed, structural quantities that are easy to interpret and measure experimentally. We test our methods on two-dimensional athermal packings of soft spheres under quasi-static shear. We find that the same structural information that predicts small variations in the response is also predictive of where plastic events will localize. We also find that an excellent accuracy is achieved in athermal sheared packings using simply a particle’s species and the number of nearest neighbor contacts.

36 MATERIALS SCIENCE↗

Mountain Basin Controls on the Snow-to-Streamflow Signal: An AIC-Weighted Multiple Linear Regression Framework

A regression-based analysis quantifies how basin characteristics modulate the snow-to-streamflow signal. First, we use the ERA5-Land reanalysis gridded product (European Centre for Medium Range Weather Forecasts reanalysis 5 -Land component) for 4,655 hydrologic unit code - 10 (HUC10) mountain basins across the western United States (US) for water years 1987–2024. Linear regressions are performed for peak snow water equivalent (SWE) and annual streamflow for each mountain basin. Models use ordinary least squares in Python’s statsmodels package. After which, an Akaike Information Criterion (AIC)–weighted ensemble multiple linear regression (MLR) framework with 47 watershed traits is used to predict the linear regression coefficient of determination (r-squared) defining the ability of peak SWE to predict annual streamflow across all mountain basin. Predictor sets are constrained to avoid multicollinearity by excluding models with variance inflation factors (VIF) greater than 5. Mountain basin traits included in the MLR include seasonal climate, topography, vegetation type and structure, and bedrock geology. Accepted models are considered if their AIC is within 2.0 of the model with the minimum AIC, or best model. To compare predictor influence across acceptable models, we computed standardized regression coefficients. To evaluate structural redundancy among models, we constructed binary inclusion vectors for each acceptable model, denoting whether a predictor was present (1) or absent (0). Core predictor variables are defined as occurring in at least 67% of the acceptable models. For this regional analysis, only one model was found acceptable, with higher snow-to-streamflow translation (higher r-squared) occurring in colder mountain basins with higher relative winter precipitation, more snow accumulation and a lower fraction of annual precipitation that falls in the spring and summer. The second component of the data package uses previously published, high-resolution output from an integrated hydrological model of the East River watershed using the U.S. Geological Survey Groundwater and Surface water Flow model (GSFLOW, doi:10.15485/1998576). East River MLR expands upon the approach described above to explore the response of five streamflow metrics—annual streamflow, runoff efficiency, 7-day minimum flow, low-flow duration, and non-perennial stream fraction to snow system indicators including peak SWE, snow-covered area, snow disappearance date, and the fraction of basin area characterized by low-to-no snow, as well as seasonal precipitation and temperature, and annual hydrologic variables representing soil moisture, evapotranspiration (ET), the partitioning of incoming precipitation to evapotranspiration (ET/P), groundwater storage, and groundwater inflow to streams. MLR was done on all water years (P0: 1987-2024) and for each period as determined in the split analysis using pooled regression techniques (P1: 1987-2011 and P2: 2012-2024) to evaluate shifting predictor variable emphasis on streamflow generation. Results indicate that since 2012, peak SWE has lost statistical strength in its prediction of annual streamflow and runoff efficiency, and the indirect influence of spring temperature has emerged as critically important. Low-flow metrics remain largely influenced by soil moisture, vegetation water use and groundwater inflows with summer precipitation becoming a direct influence on minimum summer flow. Together, these data and Python-based analysis tools provide a framework for identifying the key watershed characteristics that control how streamflow responds to snow from year to year. The package also helps quantify uncertainty in statistical models and assess how snow–streamflow relationships vary across regions and over time. This dataset contains comma-separated values files (.csv), text files (.txt), python code files (.py), figure files (.png), and shapefiles (.cpg, .dbf, .prj, .sbn, .sbx, .shp, .xml). Further details on file contents and MLR execution can be found in the readme file and the FLMD files. Work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

54 ENVIRONMENTAL SCIENCES↗

On decoupling the integrals of cosmological perturbation theory

ABSTRACT Perturbation theory (PT) is often used to model statistical observables capturing the translation and rotation-invariant information in cosmological density fields. PT produces higher order corrections by integration over linear statistics of the density fields weighted by kernels resulting from recursive solution of the fluid equations. These integrals quickly become high dimensional and naively require increasing computational resources the higher the order of the corrections. Here, we show how to decouple the integrands that often produce this issue, enabling PT corrections to be computed as a sum of products of independent 1D integrals. Our approach is related to a commonly used method for calculating multiloop Feynman integrals in quantum field theory, the Gegenbauer Polynomial x-Space Technique. We explicitly reduce the three terms entering the 2-loop power spectrum, formally requiring 9D integrations, to sums over successive 1D radial integrals. These 1D integrals can further be performed as convolutions, rendering the scaling of this method Nglog Ng with Ng the number of grid points used for each fast Fourier transform. This method should be highly enabling for upcoming large-scale structure redshift surveys where model predictions at an enormous number of cosmological parameter combinations will be required by Monte Carlo Markov Chain searches for the best-fitting values.

Slepian, Zachary↗

A Machine Learning Framework for Modeling Ensemble Properties of Atomically Disordered Materials

Atomic disorder can strongly influence material properties such as charge transport, optical response, and catalytic activity. However, efficiently modeling these disorder effects remains challenging for first-principles methods due to the cost of sampling large configurational spaces and computing complex physical quantities. Recent advances of machine learning techniques, particularly graph neural networks (GNNs), has enabled the efficient and accurate predictions of complex material properties, offering promising tools for studying disordered systems. In this work, we present a general machine-learning-assisted computational framework that integrates equivariant GNNs with Monte Carlo simulations to compute the thermodynamic and ensemble-averaged functional properties of disordered materials. Using the surface-termination-disordered MXene monolayer Ti 3 C 2 T 2–x as a representative system, we find that electrical conductivity exhibits an emergent peak near the order–disorder phase transition temperature due to the interplay between electron scattering and doping. In contrast, optical conductivity remains largely insensitive to local atomic disorder and reflects the global surface chemical composition. These results highlight the role of atomic disorder in affecting material properties and demonstrate the potential of our approach for statistically modeling disorder effects in a wide range of materials such as high-entropy alloys and spin liquids.

MXene↗