Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Traditional Machine Learning Models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 541 records · Page 30

Provably efficient machine learning for quantum many-body problems

Classical machine learning (ML) provides a potentially powerful approach to solving challenging quantum many-body problems in physics and chemistry. However, the advantages of ML over traditional methods have not been firmly established. In this work, we prove that classical ML algorithms can efficiently predict ground-state properties of gapped Hamiltonians after learning from other Hamiltonians in the same quantum phase of matter. By contrast, under a widely accepted conjecture, classical algorithms that do not learn from data cannot achieve the same guarantee. We also prove that classical ML algorithms can efficiently classify a wide range of quantum phases. Extensive numerical experiments corroborate our theoretical results in a variety of scenarios, including Rydberg atom systems, two-dimensional random Heisenberg models, symmetry-protected topological phases, and topologically ordered phases.

Science & Technology - Other Topics↗

Interpreting Transformers for Jet Tagging

Machine learning (ML) algorithms, particularly attention-based transformer models, have become indispensable for analyzing the vast data generated by particle physics experiments like ATLAS and CMS at the CERN LHC. Particle Transformer (ParT), a state-of-the-art model, leverages particle-level attention to improve jet-tagging tasks, which are critical for identifying particles resulting from proton collisions. This study focuses on interpreting ParT by analyzing attention heat maps and particle-pair correlations on the $\eta$-$\phi$ plane, revealing a binary attention pattern where each particle attends to at most one other particle. At the same time, we observe that ParT shows varying focus on important particles and subjets depending on decay, indicating that the model learns traditional jet substructure observables. These insights enhance our understanding of the model's internal workings and learning process, offering potential avenues for improving the efficiency of transformer architectures in future high-energy physics applications.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Computing for the DUNE Long-Baseline Neutrino Oscillation Experiment

This is a talk given at Computers in High Energy Physics in Adelaide, South Australia, Australia in November 2019. It is partially intended to explain the context of DUNE Computing for computing specialists. The DUNE collaboration consists of over 180 institutions from 33 countries. The experiment is in preparation now with commissioning of the first 10kT fiducial volume Liquid Argon TPC expected over the period 2025-2028 and a long data taking run with 4 modules expected from 2029 and beyond. An active prototyping program is already in place with a short test beam run with a 700T, 15,360 channel prototype of single-phase readout at the neutrino platform at CERN in late 2018 and tests of a similar sized dual-phase detector scheduled for mid-2019. The 2018 test beam run was a valuable live test of our computing model. The detector produced raw data at rates of up to ~2GB/s. These data were stored at full rate on tape at CERN and Fermilab and replicated at sites in the UK and Czech Republic. In total 1.2 PB of raw data from beam and cosmic triggers were produced and reconstructed during the six week test beam run. Baseline predictions for the full DUNE detector data, starting in the late 2020's are 30-60 PB of raw data per year. In contrast to traditional HEP computational problems, DUNE's Liquid Argon TPC data consist of simple but very large (many GB) 2D data objects which share many characteristics with astrophysical images. This presents opportunities to use advances in machine learning and pattern recognition as a frontier user of High Performance Computing facilities capable of massively parallel processing.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Exploring the frontiers of condensed-phase chemistry with a general reactive machine learning potential

Abstract Atomistic simulation has a broad range of applications from drug design to materials discovery. Machine learning interatomic potentials (MLIPs) have become an efficient alternative to computationally expensive ab initio simulations. For this reason, chemistry and materials science would greatly benefit from a general reactive MLIP, that is, an MLIP that is applicable to a broad range of reactive chemistry without the need for refitting. Here we develop a general reactive MLIP (ANI-1xnr) through automated sampling of condensed-phase reactions. ANI-1xnr is then applied to study five distinct systems: carbon solid-phase nucleation, graphene ring formation from acetylene, biofuel additives, combustion of methane and the spontaneous formation of glycine from early earth small molecules. In all studies, ANI-1xnr closely matches experiment (when available) and/or previous studies using traditional model chemistry methods. As such, ANI-1xnr proves to be a highly general reactive MLIP for C, H, N and O elements in the condensed phase, enabling high-throughput in silico reactive chemistry experimentation.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Co-Firing Switchgrass and Waste Coal in A Power Plant: A Techno-Economic and Life Cycle Evaluation for The Ohio River Valley (SWITCH) (Final Technical Report for Ohio State/FE0032204)

Abandoned coal mine lands (AMLs) represent one of the most persistent environmental challenges in the United States. Prior to the enactment of the Surface Mining Control and Reclamation Act (SMCRA) in 1977, coal mining operations were not legally required to reclaim disturbed lands, leaving behind approximately 500,000 AML sites nationwide. These sites pose severe environmental and health risks, including acid mine drainage, soil and water contamination, and spontaneous combustion of waste coal piles. Millions of Americans live within one mile of these AMLs, underscoring the urgency of remediation. Traditional reclamation practices, such as planting cool-season grasses, often fail to fully restore ecological function or leverage the economic potential of these lands. This project addressed these challenges by developing integrated strategies for resource recovery, land reclamation, and sustainable energy production. This project evaluated an integrated strategy to convert this liability into an opportunity by recovering waste coal and co-firing it with switchgrass (Panicum virgatum L.) cultivated on reclaimed or marginal AML areas in existing coal-fired power plants. Switchgrass not only provides a renewable feedstock but also aids in land reclamation and carbon sequestration. 1) Remote Sensing and Machine Learning for Waste Coal Identification Using Sentinel-2 satellite imagery and supervised classification, we applied four machine learning models to detect historical waste coal piles. Random Forest achieved the highest accuracy (precision: 86%, recall: 77%). Time-series analysis revealed gradual vegetation recovery since 1986, indicating natural reclamation processes in historical sites, while active mining areas showed ongoing disturbance. This workflow enables scalable monitoring and prioritization of reclamation efforts. 2) UAS-Based Stockpile Volume Estimation To quantify recoverable waste coal, we evaluated Unmanned Aerial Systems (UAS) equipped with Light Detection and Ranging (LiDAR) and multispectral sensors. Structure-from-Motion (SfM) photogrammetry combined with interpolated Digital Terrain Models (DTMs) achieved strong agreement with LiDAR reference volumes (Root Mean Square Error (RMSE) ≈147 m 3 , Mean Absolute Percentage Error (MAPE) ≈2%). Sensitivity analysis confirmed that spatial resolution significantly influences accuracy, emphasizing the need for high-resolution data for precise volume estimation. This approach offers a scalable, cost-effective, and accurate alternative to conventional ground-based surveys. 3) Switchgrass Cultivation for Bioenergy and Water Quality Improvement We assessed the hydrological and water quality impacts of converting AMLs to switchgrass production areas using the Soil and Water Assessment Tool (SWAT). Results showed that converting 10% of the watershed area into the switchgrass production zone reduced streamflow by 3.1%, total suspended solids by 18.1%, total nitrogen by 7.6%, and total phosphorus by 6.2%, while achieving biomass yields of 8.6–9.2 metric tons per hectare. These findings highlight switchgrass as a dual-benefit strategy for land reclamation and bioenergy feedstock production. 4) Integrated Co-Firing and CCS for Carbon-Negative Power Generation We modeled co-firing scenarios using the Power Plant Flexible Model (PPFM) to evaluate plant efficiency, greenhouse gas (GHG) emissions, and levelized cost of electricity (LCOE). Without carbon capture and storage (CCS), increasing switchgrass co-firing ratios reduced LCOE from $\$$150/MWh at 0% biomass to $\$$110/MWh at full substitution. Under CCS, costs remained higher (~$\$$250/MWh at 0% biomass) but decreased to $\$$200/MWh at 100% biomass, while enabling net-zero or carbon-negative electricity due to switchgrass sequestration benefits. Although CCS introduced efficiency penalties, pairing it with biomass co-firing offset these impacts and maximized climate benefits. Overall, optimizing co-firing ratios between 60-100%, supported by reliable logistics and storage strategies, emerged as a practical pathway to balance affordability, sustainability, and net-zero or negative GHG emissions while promoting productive reuse of AMLs.

01 COAL, LIGNITE, AND PEAT↗

ProtoDUNE-VD for Beyond the Standard Model Searches: Initial Studies and Future Prospects

The Deep Underground Neutrino Experiment (DUNE) is a next-generation long-baseline neutrino program designed to address fundamental questions in neutrino and astroparticle physics. ProtoDUNE, operating at the CERN Neutrino Platform, serves as a full-scale prototype for the DUNE Far Detector. In particular, the ProtoDUNE Vertical Drift (ProtoDUNE-VD) detector provides a powerful testbed for validating reconstruction and event selection techniques for future DUNE operations. In addition to detector R&D, ProtoDUNE enables a novel parasitic beam-dump search for beyond-the-Standard-Model (BSM) particles. However, it faces several challenges. Most notably, the ProtoDUNE-VD modules operate on the surface and are consequently exposed to an intense flux of cosmic rays, which requires a dedicated trigger. In addition, standard neutrinos are also produced in the T2 target area from the decay of unstable mesons, constituting a relevant background, which needs to be well understood and characterized a priori. We present the first studies based on 2025 data taken with a trigger designed to identify neutrino candidates at ProtoDUNE-VD. ProtoDUNE-VD’s high-resolution LArTPC imaging allows detailed reconstruction of decay and scattering signatures. This work demonstrates the complementarity of traditional tools such as Pandora and modern machine-learning approaches, providing key input for atmospheric neutrino and rare-event searches in the DUNE Vertical Drift program.

Bagdu, Halit [U. Iowa, Iowa City]↗

Using Machine Learning Tools to Predict Compressor Stall

Clean energy has become an increasingly important consideration in today’s power systems. As the push for clean energy continues, many coal-fired power plants are being decommissioned in favor of renewable power sources such as wind and solar. However, the intermittent nature of renewables means that dynamic load following traditional power systems is crucial to grid stability. With high flexibility and fast response at a wide range of operating conditions, gas turbine systems are poised to become the main load following component in the power grid. Yet, rapid changes in load can lead to fluid flow instabilities in gas turbine power systems. These instabilities often lead to compressor surge and stall, which are some of the most critical problems facing the safe and efficient operation of compressors in turbomachinery today. Although the topic of compressor surge and stall has been extensively researched, no methods for early prediction have been proven effective. This study explores the utilization of machine learning tools to predict compressor stall. The long short-term memory (LSTM) model, a form of recurrent neural network (RNN), was trained using real compressor stall datasets from a 100 kW recuperated gas turbine power system designed for hybrid configuration. Two variations of the LSTM model, classification and regression, were tested to determine optimal performance. The regression scheme was determined to be the most accurate approach, and a tool for predicting compressor stall was developed using this configuration. Overall, results show that the tool is capable of predicting stalls 5–20 ms before they occur. With a high-speed controller capable of 5 ms time-steps, mitigating action could be taken to prevent compressor stall before it occurs.

42 ENGINEERING↗

HPC Digital Twins for Evaluating Scheduling Policies, Incentive Structures and their Impact on Power and Cooling

Schedulers are critical for optimal resource utilization in high-performance computing. Traditional methods to evaluate sched- ulers are limited to post-deployment analysis, or simulators, which do not model associated infrastructure. In this work, we present the first-of-its-kind integration of scheduling and digital twins in HPC. This enables what-if studies to understand the impact of parameter configurations and scheduling decisions on the physical assets, even before deployment, or regarching changes not easily realizable in production. We (1) provide the first digital twin framework extended with scheduling capabilities, (2) integrate various top-tier HPC systems given their publicly available datasets, (3) implement extensions to integrate external scheduling simulators. Finally, we show how to (4) implement and evaluate incentive structures, as- well-as (5) evaluate machine learning based scheduling, in such novel digital-twin based meta-framework to prototype scheduling. Our work enables what-if scenarios of HPC systems to evaluate sustainability, and the impact on the simulated system.

Maiterth, Matthias [ORNL] (ORCID:000000018698460X)↗

A machine learning-based fast frequency response control for a VSC-HVDC system

An HVDC system can realize a very fast frequency response to the disturbed system under a contingency because its active power control is decoupled from the frequency deviation. However, most of existing HVDC frequency control strategies are coupled with system primary frequency control and secondary frequency control. Since the traditional system frequency control is dominated by the thermal generators, the advantage of the fast response of the HVDC system is not made fully used. The development of a frequency response estimation based on a machine learning algorithm provides another approach to improve the frequency response capability of the HVDC system. Different from other frequency deviation tracking strategies, a machine learning based HVDC frequency response control can directly increase the power flow of a HVDC system by estimation of the system generator or load lost. In this paper, a fast frequency response control using a HVDC system for a large power system disturbance based on the multivariate random forest regression (MRFR) algorithm is proposed. The simulation is carried out with an integrated power system model based on the North American interconnections. The simulation results indicate that the proposed MRFR based frequency response control can significantly improve the frequency low point during an event, while stabilizing the frequency in advance.

42 ENGINEERING↗

Model Driven Deception for Defense of Operational Technology Environments

Due to the strong integration of real-world physics, OT deception platforms must operate differently than traditional IT deceptions. For instance, turning off a valve will be detected downstream by other sensors because the flow will reduce and stop. Additionally, controllers and applications leverage data from sensors to send control commands to each other. A believable deception must be integrated with the system to project the effects of events. An attack will likely attempt to control the physical process in a negative manner. To make the attacker believe they are achieving their objective, it must predict the effects of these actions, to a reasonable degree. Our approach to simulating a model to generate realistic decoy behavior is explored including description of two approaches: a physics model-based approach and a data driven approach. The performance of two machine learning techniques are investigated in their ability to learn a good enough model of the physics of the system.

97 MATHEMATICS AND COMPUTING↗

Exploring PV Circularity by Modeling Socio-Technical Dynamics of Modules' End-of-Life Management: Preprint

The circular economy (CE) tackles environmental and resource scarcity issues by maximizing value retention in the economy. The concept implies design strategies such as reducing the use of materials or improving products’ durability and end-of-life (EOL) strategies such as reusing products and components and recycling materials. With an estimated 80 million tons of global cumulative EOL photovoltaic (PV) modules, applying CE principles to the PV industry could alleviate resource scarcity issues while also providing economic benefits. However, transitioning to a CE may imply changes in organizations and consumer behaviors. In this context, assessment of CE strategies may require accounting for behavioral change, a requirement that methods from complex system science such as agent-based modeling meet. Thus, this paper uses an agent-based modeling (ABM) approach to study circularity in the photovoltaics supply chain. Four types of agents are represented in the ABM: PV owners, installers, recyclers, and manufacturers. Moreover, five possible EOL options – including three CE strategies – are modeled. Departing from traditional techno-economic analysis, the model includes techno-economic factors as well as social factors to model EOL management decisions. Results show that each dollar decrease in the recycling fees improves the recycling rate by roughly 1.1%. However, excluding social factors underestimates the effect that lower recycling prices have on material circularity.

28 EE - Advanced Manufacturing Office (EE-5A)↗

Modeling laser-driven ion acceleration with deep learning

Developments in machine learning promise to ameliorate some of the challenges of modeling complex physical systems through neural-network-based surrogate models. High-intensity, short-pulse lasers can be used to accelerate ions to mega-electronvolt energies, but to model such interactions requires computationally expensive techniques such as particle-in-cell simulations. Multilayer neural networks allow one to take a relatively sparse ensemble of simulations and generate a surrogate model that can be used to rapidly search the parameter space of interest. In this study, we created an ensemble of over 1,000 simulations modeling laser-driven ion acceleration and developed a surrogate to study the resulting parameter space. A neural-network-based approach allows for rapid feature discovery not possible for traditional parameter scans given the computational cost. A notable observation made during this study was the dependence of ion energy on the pre-plasma gradient length scale. While this methodology harbors great promise for ion acceleration, it has ready application to all topics in which large-scale parameter scans are restricted by significant computational cost or relatively large, but sparse, domains.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Ultra-fast interpretable machine-learning potentials

Abstract All-atom dynamics simulations are an indispensable quantitative tool in physics, chemistry, and materials science, but large systems and long simulation times remain challenging due to the trade-off between computational efficiency and predictive accuracy. To address this challenge, we combine effective two- and three-body potentials in a cubic B-spline basis with regularized linear regression to obtain machine-learning potentials that are physically interpretable, sufficiently accurate for applications, as fast as the fastest traditional empirical potentials, and two to four orders of magnitude faster than state-of-the-art machine-learning potentials. For data from empirical potentials, we demonstrate the exact retrieval of the potential. For data from density functional theory, the predicted energies, forces, and derived properties, including phonon spectra, elastic constants, and melting points, closely match those of the reference method. The introduced potentials might contribute towards accurate all-atom dynamics simulations of large atomistic systems over long-time scales.

36 MATERIALS SCIENCE↗

Multiscale Flow for robust and optimal cosmological analysis

We propose Multiscale Flow, a generative Normalizing Flow that creates samples and models the field-level likelihood of two-dimensional cosmological data such as weak lensing. Multiscale Flow uses hierarchical decomposition of cosmological fields via a wavelet basis and then models different wavelet components separately as Normalizing Flows. The log-likelihood of the original cosmological field can be recovered by summing over the log-likelihood of each wavelet term. This decomposition allows us to separate the information from different scales and identify distribution shifts in the data such as unknown scale-dependent systematics. The resulting likelihood analysis can not only identify these types of systematics, but can also be made optimal, in the sense that the Multiscale Flow can learn the full likelihood at the field without any dimensionality reduction. We apply Multiscale Flow to weak lensing mock datasets for cosmological inference and show that it significantly outperforms traditional summary statistics such as power spectrum and peak counts, as well as machine learning–based summary statistics such as scattering transform and convolutional neural networks. We further show that Multiscale Flow is able to identify distribution shifts not in the training data such as baryonic effects. Finally, we demonstrate that Multiscale Flow can be used to generate realistic samples of weak lensing data.

79 ASTRONOMY AND ASTROPHYSICS↗

High-bandwidth image-based predictive laser stabilization via optimized Fourier filters

Controlling the delivery of kHz-class pulsed lasers is of interest in a variety of industrial and scientific applications, from next-generation laser-plasma acceleration to laser-based x-ray emission and high-precision manufacturing. The transverse position of the laser pulse train on the application target is often subject to fluctuations by external drivers (e.g., room cooling and heating systems, motorized optics stages and mounts, vacuum systems, chillers, and/or ground vibrations). For typical situations where the disturbance spectrum exhibits discrete peaks on top of a broad-bandwidth lower-frequency background, traditional PID (proportional-integral-derivative) controllers may struggle, since as a general rule PID controllers can be used to suppress vibrations up to only about 5%–10% of the sampling frequency. Here, a predictive feed-forward algorithm is presented that significantly enhances the stabilization bandwidth in such laser systems (up to the Nyquist limit at half the sampling frequency) by online identification and filtering of one or a few discrete frequencies using optimized Fourier filters. Furthermore, the system architecture demonstrated here uses off-the-shelf CMOS cameras and piezo-electric actuated mirrors connected to a standard PC to process the alignment images and implement the algorithm. To avoid high-end, high-cost components, a machine-learning-based model of the piezo mirror’s dynamics was integrated into the system, which enables high-precision positioning by compensating for hysteresis and other hardware-induced effects. A successful demonstration of the method was performed on a 1 kHz laser pulse train, where externally-induced vibrations of up to 400 Hz were attenuated by a factor of five, far exceeding what could be done with a standard PID scheme.

Natal, Joseph↗

Optimizing chemistry for designing oxidation resistant FeCrAl alloys

Abstract Traditionally, FeCrAl alloys played an important role in high-temperature applications due to their ability to form a passive Al oxide film at temperatures above ~ 800 °C. Recently, FeCrAl alloys became of interest for the application of accident tolerant nuclear fuel cladding. This study covers work done at GE Research for better understanding the role of Al, Cr, and Mo in oxidation kinetics and thermodynamics. Several models and commercial prototype alloys have been tested in hydrothermal corrosion autoclave loops, at low temperature steam exposure (~ 400 °C), high temperature steam exposure (~ 1000 °C or higher), and high temperature air exposures. The results provide insights on how chromium and aluminum play a significant role in both high temperature and low temperature oxidation of FeCrAl. Additionally, machine learning tools are used to gain further insights on both predicting future optimized chemistries for balancing the properties of hydrothermal corrosion, low and high temperature steam oxidation, and thermal aging (which is exacerbated due to radiation in a nuclear reactor environment). GE plans to use this framework to further optimize the FeCrAl alloy system for use in nuclear reactor environments. Graphical abstract

Roy, Indranil (ORCID:0000000336124323)↗

End-To-End Decentralized Transmission Line Protection in IBR-Dominated Weak Grids Using Interpretable Data-Driven Methods

Traditional transmission line protection relies on predictable synchronous-based fault signatures, which frequently fail under the non-standard, current-limited fault characteristics of Inverter-Based Resources (IBRs). This study investigates how to achieve secure, communication-free fault isolation in IBR-dominated weak grids without relying on opaque, computationally heavy "black-box" machine learning algorithms. To address this, we propose a novel, standalone, and inherently interpretable data-driven protection framework. Unlike centralized methods requiring multi-terminal communication, this decentralized approach relies solely on local measurements using a hierarchical linear-kernel Support Vector Machine (SVM). The methodology decomposes the protection task into four sequential stages that mimic traditional protection elements: fault detection and fault direction identification, fault type classification, zone classification, and location estimation. This multi-stage architecture allows for specialized feature engineering at each stage, combining high computational efficiency with logic traceability. The framework's end-to-end performance was validated via C-code and PSCAD/EMTDC co-simulation, utilizing a real-world utility network and an OEM black-box IBR model. The proposed relay achieves 97.2% overall accuracy and provides a reliable trip decision within a 2.5-cycle window. The results confirm 100% accuracy in fundamental fault detection, reliable zone selectivity across low to moderate fault resistances, and robust security against non-fault transients, proving its immediate viability for integration into commercial numerical relays.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Accelerating catalytic advancements through the precision of high-throughput experiments & calculations

The growing demand for energy-efficient processes to support a sustainable future drives the need for research to rapidly explore chemical and material space through accelerated catalyst discovery initiatives. Recent breakthroughs in high-throughput experimental and computational methods are transforming the catalysis field, surpassing traditional approaches to manipulating variables in catalytic processes. Key advancements in innovation include the integration of machine learning for efficient catalyst screening, high-throughput experimentation, data-driven methodologies employing comprehensive databases, and in situ and in operando techniques for realistic observations. This progress has undoubtedly been intertwined with a collaborative framework across disciplines, reshaping catalyst discovery methods in both industry and academia. This Opinion article presents a multifaceted perspective from coauthors with expertise spanning various stages of the Technology Readiness Level spectrum, highlighting both opportunities and persistent challenges in integrating computational and experimental approaches in catalysis. These challenges span from obtaining high-quality experimental data, scaling simulations to industrially relevant materials and process conditions to navigating the complexity and predictive accuracy of computational models.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗