Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “learning rate”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Framework and Tool for Artificial Intelligence & Machine Learning (AI/ML) Enabled Automated Non-Destructive Inspection of Composites Aerostructures Manufacturing

Vehicles and systems in the field of aerospace have two major requirements: a high demand for a large quantity and an expectation to perform for their lifetime with little to no failures. Thus, there is a need for a fast production rate of aerospace products with high quality. Improvements to production rate have many benefits, including a reduction in energy consumption per unit manufactured. This would be from factory energy usage, which is required to build and verify a product. Manufacturing process specifications require inspection of parts to determine if any flaws are present. Depending on factory planning and product quality, especially at higher rates, the evaluation process can pose a production rate bottleneck. This project was comprised of using artificial intelligence and machine learning (AI/ML) methods on inspection evaluations with the objective of reducing the required time to produce an aerospace structure or product and without reducing the final quality.

42 ENGINEERING↗

Ephemeral Learning - Augmenting Triggers with Online-Trained Normalizing Flows

The large data rates at the LHC require an online trigger system to select relevant collisions. Rather than compressing individual events, we propose to compress an entire data set at once. We use a normalizing flow as a deep generative model to learn the probability density of the data online. The events are then represented by the generative neural network and can be inspected offline for anomalies or used for other analysis purposes. We demonstrate our new approach for a toy model and a correlation-enhanced bump hunt.

97 MATHEMATICS AND COMPUTING↗

Race-Specific Risk Factors for Homeownership Disparity in the Continental United States

The United States has a racial homeownership gap due to a legacy of historic inequality and discriminatory policies, but factors that contribute to the racial disparity in homeownership rates between White Americans and people of color have not been fully characterized. In order to alleviate this issue, policymakers need a better understanding of how risk factors affect the homeownership rates of racial and ethnic groups differently. In this study, data from several publicly available surveys, including the American Community Survey and United States Census, were leveraged in combination with statistical learning models to investigate potential factors related to homeownership rates across racial and ethnic categories, with a focus on how risk factors vary by race or ethnicity. Our models indicated that job availability for specific demographics, and specific regions of the United States were factors that affect homeownership rates in Black, Hispanic, and Asian populations in different ways. Based on the results of this study, it is recommended policymakers promote strategies to increase access to jobs for people of color (POC), such as vocational training and programs to reduce implicit bias in hiring practices. These interventions could ultimately increase homeownership rates for POC and be a step toward reducing the racial wealth gap.

99 GENERAL AND MISCELLANEOUS↗

ELM‐MOSART‐DOC: A Large‐Scale Riverine Dissolved Organic Carbon Model and Its Application Over the United States

Riverine dissolved organic carbon (DOC), primarily sourced from soil organic carbon (SOC), plays a crucial role in regional and global carbon cycles. However, the complexities of the underlying mechanisms and limited observations present significant challenges for predictive understanding of DOC at regional or larger scales. Recently, we developed a machine learning‐based (ML) map of DOC transformation rates, bridging the gap between SOC and DOC leaching flux and simplifying terrestrial DOC representation. Building on this advancement, we introduce ELM‐MOSART‐DOC, a DOC module integrated into the riverine component of the Energy Exascale Earth System Model (E3SM)—the Model for Scale Adaptive River Transport (MOSART). ELM‐MOSART‐DOC simulates DOC transport and transformation across both headwater streams and river networks, including those managed. Model validation demonstrates the ability of ELM‐MOSART‐DOC to accurately capture long‐term average DOC concentrations, with Kling‐Gupta Efficiency (KGE) scores of 0.58 and 0.76 at large and local stations, respectively. We further assess the impact of reservoirs through different simulation schemes, revealing that reservoirs significantly alter DOC fluxes by regulating streamflow patterns and promoting DOC mineralization. Model simulations indicate that reservoirs reduce total DOC flux from the Mississippi River into the ocean by 7.5%, with the long‐term average annual export decreasing from 3.34 to 3.14 teragrams (Tg) per year. ELM‐MOSART‐DOC integrates process‐based modeling with ML parameterization to enhance the predictive understanding of riverine biogeochemical processes. This approach reduces uncertainties in modeling regional and global carbon cycle ESMs and provides new insights into carbon cycling and its implications for global environmental change.

Li, Lingbo [Univ. of Houston, TX (United States); ↗

Utilizing machine learning to improve the precision of fluorescence imaging of cavity-generated spin squeezed states

We present a supervised learning model to calibrate the photon collection rate during the fluorescence imaging of cold atoms. The linear regression model finds the collection rate at each location on the sensor such that the atomic population difference equals that of a highly precise optical cavity measurement. This 192 variable regression results in a measurement variance 27% smaller than our previous single variable regression calibration. The measurement variance is now in agreement with the theoretical limit due to other known noise sources. This model efficiently trains in less than a minute on a standard personal computer's CPU and requires less than 10 min of data collection. Furthermore, the model is applicable across a large change in population difference and across data collected on different days.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

FutureTense

Protective vaccines and reliable diagnostics are essential tools for controlling viral diseases. However, the efficacy of these tools can be diminished by mutations in viral genomes. The delay between the emergence of new viral strains and the redesign of vaccines and diagnostics allows for continued viral transmission. Is it possible to address this challenge by computationally predicting viral genome sequence evolution? Can we “future-proof” vaccines and diagnostics by targeting both current and anticipated future sequence variants? While predicting viral evolution is still an unsolved, “grand challenge” problem in biology, the large, and rapidly growing, number of SARS-CoV-2 genome sequences provide an opportunity to quantify the ability of machine learning to predict viral genome sequence evolution. Towards this end, we have developed a simple computational model for predicting viral evolution at the level of individual nucleotides. The key metric for quantifying the per-base, prediction accuracy for viral evolution is the Mann-Whitney U statistic (or, equivalently, the area under the receiver operator curve). Since the Mann-Whitney U statistic is not a differentiable function, existing deep leaning packages (like Pytorch and Keras/TensorFlow) are not useful, as they require that the accuracy metric/objective function be analytically differentiable with respect to the model parameters. To overcome this challenge, we have implemented custom software, “FutureTense”, that can train a machine learning model by maximizing the non-differentiable Mann-Whitney U statistic. This software trains a machine learning model by exploring along the direction of the discrete gradient of the Mann-Whitney U statistic in the model parameter space. Parallel computing and genome sequence-specific optimizations are used to accelerate model training. The resulting machine learning model learns the observed high C->U mutation rates in the SARS-CoV-2 genome (which are potentially induced by host defenses) and provides prediction accuracies that are significantly better than one would expect from random chance. While predicting viral evolution is still quite far from a solved problem, the surprising performance of this simple model gives hope that the accuracy of predicting viral genome evolution can be further increased by more sophisticated approaches.

Gans, Jason↗

Conduct of Operations and Nuclear Criticality Safety Standards

Since the beginning of the nuclear age in 1943, there have been 22 criticality accidents in process facilities worldwide in which fissionable materials were processed by hand for various purposes. Of these 22 process criticality accidents, 16 involved faulty or flawed conduct of operations. Most of these process accidents (17 of 22) occurred before 1970, mostly due to the implementation of formal conduct of operations—such as the use of operating procedures and training—and the use of consensus standards for nuclear criticality safety (NCS). Obviously, the nuclear facilities applied lessons learned from their accident history, and the rate of criticality accidents was significantly reduced as a result. Over the years, it has been emphasized in NCS training courses and other venues in the United States that the consensus standards, such as ASA N6.1-1964, were the key contributors to this reduction in accidents. The formality of operations required to implement and apply the standards at a facility is also a crucial prerequisite that must be sufficiently robust to ensure NCS. This prerequisite is crucial to success fully implement the NCS consensus and other safety standards at a site.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Cyber-Attack Detection and Accommodation for the Energy Delivery System

The goals of this project were to create a software system with a suite of key algorithms for cyber-attack detection and accommodation providing domain layer protection for critical power generation assets. Example assets included gas and steam turbines, heat recovery steam generators, and electrical generators. The aggressive algorithm goals were aimed at reducing the false positive rates in threat detection to <1% using learnings from many evolving disciplines (power turbine and generator physics, power system modeling, modern control theory, system identification, machine learning, deep learning, mathematics and data science). Additional goals for the algorithms involved localizing threats on-the-fly to know in which monitoring node the effects of attacks are present, and then providing accommodation to keep the system running uninterrupted much of the time in the presence of the attack. Accommodation had a performance goal of providing resiliency when up to 50% of monitoring nodes are in an attack state.

cybersecurity, cyber-physical↗

Research Introduction [Slides]

In the topic of Reverse Time Imaging, we proposed a new IC to reduce computation cost but reserve image resolution for distributed sensor networks. For Induced Seismicity in Oklahoma, we analyzed fault stress state analysis at state scale, and we applied machine learning techniques to polarity picking and seismicity rate forecasting. The results provide better understanding of fault properties, stress field, and the relationships among fault, stress state, injections, and potential seismic hazards. Lastly, for Microseismic Monitoring, we detected and located 770 low-frequency events (5-50 Hz): (1) Shallow events are highly clustered, consistent with the pathway from injection well 13-10A to monitoring well; moment tensor inversion shows dominant tensile cracking; (2) Deep events are scattered and show migration pattern to the basement; moment tensors show that most events are shear cracks.

58 GEOSCIENCES↗

Learning to Trigger: Reinforcement Learning at the Large Hadron Collider

High-throughput scientific facilities such as the Large Hadron Collider depend on real-time event filtering (\textit{triggering}) under tight constraints on bandwidth, latency, and storage. In practice, trigger menus are largely static and hand-tuned and can become suboptimal as detector conditions, pileup, and background composition drift over time. We cast online threshold tuning as a sequential decision-making problem: a reinforcement learning agent ingests streaming summaries of recent rates and signal-sensitive features and updates trigger thresholds to maximize signal efficiency while tracking a target background rate within a tolerance band. We adapt Group-Filtered Policy Optimization (GFPO) to streaming control and introduce two variants (GFPO-F, GFPO-FR) that enforce background rate feasibility during training. On a benchmark that emulates realistic collider operation, we study two representative triggers: a total transverse energy ($H_{T}$) trigger sensitive to pileup variation, and an anomaly-detection (AD) trigger based on reconstruction loss for rare or non-standard signatures. On Monte Carlo streams, our agent increases the fraction of in-tolerance time intervals by 48% ($H_T$) and 28% (AD), with a cumulative gain of up to 2% in signal efficiency on those in-tolerance intervals. Transferring from simulation to \emph{real} collision data (CMS Run 283408), the same agent, without fine-tuning, achieves a 56% ($H_T$) and 28% (AD) in-tolerance improvement over baselines, with further signal-efficiency gain on both triggers. To our knowledge, this is the \emph{first} demonstration of RL-based trigger control on real Large Hadron Collider collision data. Code is available at https://github.com/Zixind/GFPO_LHC (see repo for details).

Ding, Zixin [Chicago U.]↗

Multitask Machine Learning of Collective Variables for Enhanced Sampling of Rare Events

Computing accurate reaction rates is a central challenge in computational chemistry and biology because of the high cost of free energy estimation with unbiased molecular dynamics. In this work, a data-driven machine learning algorithm is devised to learn collective variables with a multitask neural network, where a common upstream part reduces the high dimensionality of atomic configurations to a low dimensional latent space and separate downstream parts map the latent space to predictions of basin class labels and potential energies. Here, the resulting latent space is shown to be an effective low-dimensional representation, capturing the reaction progress and guiding effective umbrella sampling to obtain accurate free energy landscapes. This approach is successfully applied to model systems including a 5D Müller Brown model, a 5D three-well model, the alanine dipeptide in vacuum, and an Au(110) surface reconstruction unit reaction. It enables automated dimensionality reduction for energy controlled reactions in complex systems, offers a unified and data-efficient framework that can be trained with limited data, and outperforms single-task learning approaches, including autoencoders.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

GeoNEX-ML: A Machine Learning System for Earth Observations

Improved capabilities of earth monitoring satellites are enabling a wide range of studies on the environmental effects of climate change, often leveraging the recent advancements in machine learning. At the same time, the new capabilities, including higher spatial resolution and temporal frequency, are expanding the amount of data generated at exponential rates. At the NASA Earth eXchange (NEX), we build deep learning methods to learn from cross sensor satellite-based Earth observations for generating new datasets with efficient processing techniques. Using current generation geostationary satellites on NEX, we present an interchangeable set of machine models to perform spectral adjustment, physical model emulation, LEO-GEO emulation, and optical flow. These tools are used to generate consistent virtual observations across sensors, perform atmospheric correction and cloud detection, and estimate land surface temperature and atmospheric winds. This approach aims to improve the robustness of remotely sensed data processing by learning from diverse sets of observations while enabling near real-time and on-demand capabilities.

Geostationary↗

Atmospheric winds with deep optical flow

Improved capabilities of earth monitoring satellites are enabling a wide range of studies on the environmental effects of climate change, often leveraging the recent advancements in machine learning. At the same time, the new capabilities, including higher spatial resolution and temporal frequency, are expanding the amount of data generated at exponential rates. At the NASA Earth eXchange (NEX), we build deep learning methods to learn from cross sensor satellite-based Earth observations for generating new datasets with efficient processing techniques. Using current generation geostationary satellites on NEX, we present an interchangeable set of machine models to perform spectral adjustment, physical model emulation, LEO-GEO emulation, and optical flow. These tools are used to generate consistent virtual observations across sensors, perform atmospheric correction and cloud detection, and estimate land surface temperature and atmospheric winds. This approach aims to improve the robustness of remotely sensed data processing by learning from diverse sets of observations while enabling near real-time and on-demand capabilities.

Atmospheric winds↗

Adaptive Activation Functions Accelerate Convergence in Deep and Physics-informed Neural Networks

We employ adaptive activation functions for regression in deep and physics-informed neural networks (PINNs) to approximate smooth and discontinuous functions as well as solutions of linear and nonlinear partial differential equations. In particular, we solve the nonlinear Klein-Gordon equation, which has smooth solutions, the nonlinear Burgers equation, which can admit high gradient solutions, and the Helmholtz equation. We introduce a scalable hyper-parameter in the activation function, which can be optimized to achieve best performance of the network as it changes dynamically the topology of the loss function involved in the optimization process. The adaptive activation function has better learning capabilities than the traditional one (fixed activation) as it improves greatly the convergence rate, especially at early training, as well as the solution accuracy. To better understand the learning process, we plot the neural network solution in the frequency domain to examine how the network captures successively different frequency bands present in the solution. We consider both forward problems, where the approximate solutions are obtained, as well as inverse problems, where parameters involved in the governing equation are identified. Our simulation results show that the proposed method is a very simple and effective approach to increase the efficiency, robustness and accuracy of the neural network approximation of nonlinear functions as well as solutions of partial differential equations, especially for forward problems. We theoretically prove that in the proposed method, gradient descent algorithms are not attracted to suboptimal critical points or local minima.

machine leaning, Bad minima, Inverse problems, Phy↗

Machine learning analysis reveals relationship between pomacentrid calls and environmental cues

Sound production rates of fishes can be used as an indicator for coral reef health, providing an opportunity to utilize long-term acoustic recordings to assess environmental change. As acoustic datasets become more common, computational techniques need to be developed to facilitate analysis of the massive data files produced by long-term monitoring. Machine learning techniques demonstrate an advantage in the identification of fish sounds over manual sampling approaches. Here we evaluated the ability of convolutional neural networks to identify and monitor call patterns for pomacentrids (damselfishes) in a tropical reef region of the western Pacific. A stationary hydrophone was deployed for 39 mo (2014-2018) in the National Park of American Samoa to continuously record the local marine acoustic environment. A neural network was trained—achieving 94% identification accuracy of pomacentrids—to demonstrate the applicability of machine learning in fish acoustics and ecology. The distribution of sound production was found to vary on diel and interannual timescales. Additionally, the distribution of sound production was correlated with wind speed, water temperature, tidal amplitude, and sound pressure level. This research has broad implications for state-of-the-art acoustic analysis and promises to be an efficient, scalable asset for ecological research, environmental monitoring, and conservation planning.

59 BASIC BIOLOGICAL SCIENCES↗

Explainable and Differentiable Reinforcement Learning for Multi-objective Optimization in Particle Accelerators

Operating particle accelerators involves optimizing multiple goals simultaneously, which can be challenging due to trade-offs among objectives. While evolutionary algorithms like the genetic algorithm (GA) have been used for various Multi-Objective Optimization (MOO) tasks, they are not inherently suited for complex control problems. This talk highlights two variations of Reinforcement Learning (RL) for concurrently optimizing heat load and trip rates at the Continuous Electron Beam Accelerator Facility (CEBAF). The problem involves strict constraints on individual states, actions, and overall energy requirements of the beam. First, this talk highlights how differentiability can be harnessed through a Deep Differentiable Reinforcement Learning (DDRL) approach to address MOO issues within particle accelerators. We examine the DDRL method alongside Model Free Reinforcement Learning (MFRL), GA, and Bayesian Optimization (BO). The performance of these methods is assessed by generating a Pareto-front for two objectives. Our findings indicate that DDRL excels in handling high-dimensional problems more effectively than MFRL, BO, and GA. Next, we will show integration of explainable physics-based constraints into RL algorithms to enhance trans- parency and trust in decision-making processes by enabling users to verify that agents adhere to established physical principles. This surrogate function can be modeled using neural networks or sparse dictionary mod- els. By examining the mathematical form of the learned constraint function, we are able to confirm the agent has learned to use the established physics of each environment provided but the surrogate model. In addi- tion, we find that the introduction of a mathematical functional dictionary based surrogate model enables our reinforcement learning algorithms to reliably converge for difficult high-dimensional accelerator controls environments.

Rajput, Kishansingh [Thomas Jefferson National Acc↗

A Reinforcement Learning Approach to Augment Conventional PID Control in Nuclear Power Plant Transient Operation

The ability of nuclear reactors to operate their power conversion cycles more flexibly will enhance their value to energy grids with variable pricing. Current nuclear control systems are typically classical controllers that are often based on proportional-integral-derivative (PID) control. This paper presents a method of augmenting the existing PID control for difficult transient operations in nuclear power plants using a reinforcement learning–derived feedforward signal applied in real time. The agents, which are trained on a test thermal load-following problem, are designed to improve steam generator outlet temperature control for a range of fast load-following scenarios covering ramp rates from 9%/min to 15%/min. Several reinforcement learning algorithms were initially investigated for the training of the feedforward agents with deep Q-learning (DQN) and proximal policy optimization (PPO) networks, which were found to be the most promising. The DQN controllers utilize discrete actions, giving them a better disturbance rejection at steady state but inconsistent response to initial temperature deviations. In contrast, PPO-trained agents, which take continuous actions except for a dead zone around zero, were shown to have the best combination of high disturbance rejection at steady state and good tracking of the desired temperature value. The ability of the PPO agent was also examined, with the average time of decision making found to be on the order of 1 ms. The fault properties of the controller under the loss of the reinforcement learning agent feedforward signal were also examined. The controller showed strong performance in situations of “no-signal” faults. but was less good at handling “stuck-at” faults, where the feedforward signal remains at a set value. In both cases, however, the PID was able to successfully maintain stability, eventually returning the system to a steady state. It is hoped that this work will allow for the proposed control architecture to be examined for more difficult control problems such that it may eventually be used to adapt existing nuclear plants for more aggressive load-following on grids of the future.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Pointing stabilization of a 1 Hz high-power laser via machine learning

Abstract High-power lasers are vital for particle acceleration, imaging, fusion and materials processing, requiring precise control and high-energy delivery. Laser plasma accelerators (LPAs) demand laser positional stability at focus to ensure consistent electron beams in applications such as X-ray free-electron lasers and high-energy colliders. Achieving this stability is especially challenging for the low-repetition-rate lasers in current LPAs. We present a machine learning method that predicts and corrects laser pointing instabilities in real-time using a high-frequency pilot beam. By preemptively adjusting a correction mirror, this approach overcomes traditional feedback limits. Demonstrated on the BELLA petawatt laser operating at the terawatt level (30 mJ amplification), our method achieved root mean square pointing stabilization of 0.34 and 0.59 $\unicode{x3bc} \mathrm{rad}$ in the x and y directions, reducing jitter by 65% and 47%, respectively. This is the first successful application of predictive control for shot-to-shot stabilization in low-repetition-rate laser systems, paving the way for full-energy petawatt lasers and transformative advances across science, industry and security.

Amodio, Alessio↗