Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Statistical learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

LCSL-II Cryomodule Testing at Fermilab

Cold powered testing of all LCLS-II production cryomodules at Fermilab is complete as of February 2021. A total of twenty-five tests on both 1.3 GHz and 3.9 GHz cryomodules were conducted over a nearly five year time span beginning in the summer of 2016. During the course of this campaign cutting-edge results for cavity Q₀ and gradient in continuous wave operation were achieved. A summary of all test results will be presented, with a comparison to established acceptance criteria, as well as overall test stand statistics and lessons learned.

43 PARTICLE ACCELERATORS↗

DeepBench: A simulation package for physical benchmarking data

We introduce **DeepBench**, a python library that generates simple simulated image data from first principles, such as basic geometric shapes and astronomical objects. These data are highly valuable for developing (calibration, testing, and benchmarking) statistical and machine learning models because they make it possible to connect the final data product to physically interpretable inputs. This software includes tools to curate and store the datasets to maximize reproducibility.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Hydrodynamic and Radiographic Toolbox (HART)

With a multi-lab and university team, we propose to develop new methods for a Hydrodynamic and Radiographic Toolbox (HART) that will enable a fuller and more extensive use of experimental radiographic data towards better characterizing and reducing uncertainties in predictive modeling of weapons performance. We will achieve this by leveraging recent developments in areas of computational imaging, statistical and machine learning, and reduced order modeling of hydrodynamics. In terms of software practices, by partnering with XCP (RISTRA Project) we will conform to recent XCP standards that are inline with modern software practices and standards and ensure compatibility and ease of inter-operability with existing codes. Three new activities under this proposal include (a) the development and use of deep learning-based surrogates to accelerate reconstruction and variational inference of density fields from radiographs of hydrotests, (b) a model-data fusion strategy that couples deep learning-based density reconstructions with fast hydrodynamics simulators to better constrain the reconstruction, and (c) a method for treating asymmetries using techniques adopted from limited view tomography. All three new activities will be based on improved treatment of scatter, noise, beam spot movement, detector blur, and flat fielding in the forward model, and a use of sophisticated priors to aid in the re construction. The improvements to the forward model and improved algorithmic design of the reconstruction when complete will be contained in the iterative reconstruction code SHIVA—a code project that we have recently initiated. The many ways in which machine learning can be used in the reconstruction work will be contained in a code HERMES that has been initiated with DTRA support. For example, the significant levels of acceleration that will likely be achieved by the use of machine learning techniques will permit us (and are required) to quantify uncertainties in density retrievals. Next, the two-way coupling between density reconstruction and model-based simulation of the hydrodynamics will be contained in code EREBUS, and will permit a fuller realization of the potential of the data to constrain the hydrodynamic model and better address issues related to asymmetries in the problem. Finally, we anticipate that the better consistency with physics achieved in our reconstructions will allow them to be used by X-Division more so than today.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF↗

Development of Prognostic Models Using Plant Asset Data

The recent growth of machine learning and artificial intelligence technologies provides opportunities for leveraging data-driven algorithms to address the problems of diagnostics and prognostics in the nuclear power industry. The use of machine learning and other statistical methods as prognostic models is of particular interest in the nuclear industry to accurately predict future equipment or plant state given a set of measurements. Such predictive capability will enable predictive assessment of component condition and remaining life and allow for condition-based predictive maintenance. The resulting optimization of maintenance scheduling and reduction in unnecessary maintenance activities will lower overall maintenance costs and improve the economics of nuclear power. This report discusses the various aspects of data processing and model development that are likely to influence the performance of prognostic models. Data from a boiling-water reactor was used to evaluate several prognostic models to identify key considerations for developing such models to predict data-driven plant state and equipment degradation condition. Preliminary results indicate the need for data sets that are relevant to the problem at hand and contain signatures that may be correlated to the prediction problem. Assuming such data exist, development of prognostic models using data-driven methods requires an understanding of the various sources of influence on the prediction accuracy (such as the model architecture, data preprocessing approaches, and potentially external factors influencing the equipment or plant system under assessment). Ongoing research is evaluating these factors in greater detail and examining techniques for calculating prediction uncertainty bounds.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Coordinated Ramping Product and Regulation Reserve Procurements in CAISO and MISO using Multi-Scale Probabilistic Solar Power Forecasts (Pro2R)

How can probabilistic solar forecasts lower costs and improve reliability for independent system operator (ISO) markets? We tackle this question in three steps. First, we enhance an existing solar forecasting system to provide well-calibrated hours-ahead probabilistic forecasts. We then relate the degree of uncertainty in those forecasts to error distributions for net load ramps for the California ISO (CAISO) using statistical and machine learning methods. Projected net load errors conditioned on solar uncertainty are translated into flexible ramp requirements that therefore reflect real-time meteorological and solar conditions, improving on typical ISO procedures. Finally, a multi-period look-ahead production cost model quantifies how conditional ramp requirements can a) decrease operating costs by lowering requirements compared to often conservative unconditional methods, and b) reduce generation scarcity events and consequently improve reliability by increasing flexibility requirements at times when unconditional forecast-based requirements understate actual ramp uncertainty. In addition to the products just described (quantification of solar uncertainty, its translation into requirements for ramp capability product, and quantification of the benefits of more accurate ramp requirements), this project also developed a visualization system that alerts system operators of ramp and uncertainty conditions within the network based on solar forecasts. The system is called Resource Forecast and Ramp Visualization for Situational Awareness (RaVIS). These four products represent significant advances in the state-of-the-art of probabilistic solar forecasting, development of weather-informed reserve requirements, production costing methods for estimating the benefits of more accurate reserve requirements, and visualization of system status, respectively. Yet the products are also practical and can be immediately implemented, potentially enabling system operators to save millions of dollars in ramp product procurement costs per year.

14 SOLAR ENERGY↗

Technical Report on the Belle II Summer Workshop and Explorer Workshop 2023

The 2023 Belle II Summer workshop took place July 24-28, 2023 at Duke University in Durham. The meeting webpage can be found at https://indico.belle2.org/event/8841/. The meeting had 58 registered participants with the overwhelming majority attending in person. The event at giving beginning graduate students and postdocs an overview over the Belle II physics program and detector as well as an in-depth exploration of the Belle II software. For the latter, several hands-on sessions were organized to introduce participants to the Belle II software as well as more specialized topics in statistics and Machine Learning. The workshop also featured an ML/AI competition. In addition to the DOE support, the workshop was also supported by the Duke Physics department. We were able to host almost all students that so wished in the Duke Dorms and provide a meal plan. The support enabled us to waive the registration fee for all participants and cover also part of the dorm costs. Furthermore, we covered travel costs for external speakers on ML/AI topics. Figures 1 and 2 show the group picture and a scene from the hands-on sessions, respectively.

99 GENERAL AND MISCELLANEOUS↗

The Bias-Variance-Correlation Tradeoff and Its Implications for ML Applications in HEP

The bias-variance tradeoff is a well-recognized phenomenon in statistics and machine learning. In this talk, I will discuss an extension, dubbed the bias-variance-correlation tradeoff. Roughly speaking, as the flexibility of a model decreases, the correlations in the outputs of a trained model for different inputs increases. Such correlations have implications for several applications of machine learning in high energy physics, e.g., the use generative models for event generation. In particular, I will argue that claims in the literature of data amplification by generative models stem from ignoring important correlations between the model's outputs for different inputs.

Shyamsundar, Prasanth [Fermilab] (ORCID:0000000227↗

Hazard Detection Detector Cards

This report presents a comprehensive summary of five advanced anomaly detection tools developed and deployed by Oak Ridge National Laboratory in support of the VA’s Health Information Technology modernization. These detectors—Order Path Tracker, Trend Watcher, Pain Pointer, Performance Monitor, and Patient Record Flag Detector—leverage statistical and machine learning methods to monitor workflow disruptions, detect anomalies in care sequences and volumes, identify bottlenecks, and track system-level performance metrics across VistA and Millennium systems. All detectors have been integrated into the Health Data Analytics Platform (HDAP), with most having completed deployment and testing using live data from targeted stations in cardiology and oncology domains. This work enhances VA’s capacity for proactive system surveillance, promotes patient safety, and informs data-driven operational improvements across the EHR ecosystem.

97 MATHEMATICS AND COMPUTING↗

An Overview of Electric Vehicle Load Modeling Strategies for Grid Integration Studies

The adoption of electric vehicles (EVs) has emerged as a solution to reduce greenhouse gas emissions in the transportation sector, which has motivated the implementation of public policies to promote their use in several countries. However, the high adoption of EVs poses challenges for the electricity sector, as it would imply an increase in energy demand and possible impacts on the power quality (PQ) of the power grid. Therefore, it is important to conduct EV integration studies in the power grid to determine the amount that can be incorporated without causing problems and identify the areas of the power sector that will require reinforcements. Accurate EV load patterns are required for this type of study that, through mathematical modeling, reflect both the dynamic behavior and the factors that influence the decision to recharge EVs. This article aims to present an overview of EVs, examine the different factors considered in the literature for modeling EV load patterns, and review modeling methods. EV load modeling methods are classified into deterministic, statistical, and machine learning. The article shows that each modeling method has its advantages, disadvantages, and data requirements, ranging from simple load modeling to more accurate models requiring large datasets.

Computer Science↗

Data from: Understanding the biogeochemical and spatial drivers of methane and carbon dioxide fluxes in a large temperate reservoir

This dataset contains spatially resolved measurements of CO₂ and CH₄ fluxes and associated environmental variables collected across 200 sites in Douglas Reservoir (Tennessee, USA) between July 29-August 2, 2024. Measurements include diffusive fluxes of CO₂ and CH₄, CH₄ ebullition, and biogeochemical and spatial variables such as dissolved oxygen, temperature, conductivity, chlorophyll-a, pH, water depth, and distance from the dam. Sampling was conducted using a spatially balanced design to capture longitudinal and depth-related gradients throughout the reservoir. The dataset is structured to support analyses of spatial variability, flux pathway comparisons, and modeling approaches (e.g., spatial statistics and machine learning) aimed at understanding controls on reservoir CO₂ and CH₄ fluxes and improving upscaling to whole-reservoir and regional estimates.

Neeper, Jamie [ORNL] (ORCID:0009000842342101)↗

Data exploration systems for databases

Data exploration systems apply machine learning techniques, multivariate statistical methods, information theory, and database theory to databases to identify significant relationships among the data and summarize information. The result of applying data exploration systems should be a better understanding of the structure of the data and a perspective of the data enabling an analyst to form hypotheses for interpreting the data. This paper argues that data exploration systems need a minimum amount of domain knowledge to guide both the statistical strategy and the interpretation of the resulting patterns discovered by these systems.

Greene, Richard J.↗

Ceramic processing: Experimental design and optimization

The objectives of this paper are to: (1) gain insight into the processing of ceramics and how green processing can affect the properties of ceramics; (2) investigate the technique of slip casting; (3) learn how heat treatment and temperature contribute to density, strength, and effects of under and over firing to ceramic properties; (4) experience some of the problems inherent in testing brittle materials and learn about the statistical nature of the strength of ceramics; (5) investigate orthogonal arrays as tools to examine the effect of many experimental parameters using a minimum number of experiments; (6) recognize appropriate uses for clay based ceramics; and (7) measure several different properties important to ceramic use and optimize them for a given application.

Weiser, Martin W.↗

Using Historical Data to Automatically Identify Air-Traffic Control Behavior

This project seeks to develop statistical-based machine learning models to characterize the types of errors present when using current systems to predict future aircraft states. These models will be data-driven - based on large quantities of historical data. Once these models are developed, they will be used to infer situations in the historical data where an air-traffic controller intervened on an aircraft's route, even when there is no direct recording of this action.

trajectory generation↗

Multispectral Imagery Research and Applications

The NASA Short-term Prediction Research and Transition (SPoRT) Center developed techniques to improve the quality and interpretation of multispectral imagery derived from NASA/NOAA geostationary satellites using statistical and machine learning approaches. A physically-based machine learning approach, DustTracker-AI, was developed to overcome the problem of night-time dust detection and to augment dust analysis with satellite products such as the Dust RGB. Additionally, an updated limb-correction and intercalibration methodology for short-wave, near-infrared, thermal infrared, and water vapor bands was developed for the purpose of developing a suite of high-quality RGB imagery that can be used at high viewing angles and across the constellation of geostationary sensors. This presentation will briefly highlight the techniques developed to detect dust in difficult night-time scenes and improve the quality and interpretation of multispectral imagery.

Emily Berndt↗

Assessing Current and Future Infrastructure Hazards

Project Objective Execute intelligent analytics via an advanced analytical framework, to assess the current state of offshore infrastructure, evaluate infrastructure life, and identify technologies to reduce infrastructure hazards, costs, and extend infrastructure life. Approach &amp; Results thus Far• Build comprehensive dataset• Perform data-driven analytics to evaluate infrastructure integrity<p> 1. Remaining lifespan</p><p> 2. Likelihood of future risk</p><p>• Apply data-driven advanced spatial, statistical, and Machine Learning (ML) models to quantify existing infrastructure integrity</p><p>• Release data and models through a smart, online platform hosted by Energy Data eXchange (EDX)</p>

Romeo, Lucy F.↗

Utility-scale Building Type Assignment Using Smart Meter Data

United States building energy use accounted for 40% of total energy use, 74% of peak demand, and $412 billion in 2019. Building energy modeling allows researchers to simulate building physics, gain insights into possible energy/demand saving opportunities, and assess cost-effective resilience amidst climate change. Many building features needed to create building energy models are readily available such as 2D footprints and LiDAR (height). A critical feature that is not generally obtainable is the building type. In partnership with a utility, a years worth of real-world, 15-minute electrical use data has been examined. The smart meter data is compared to 97 different prototype building energy models to assign building type. Real-world considerations including data preparation, quality assurance, and handling of missing values for advanced metering infrastructure data are addressed. Euclidean distance for pattern-matching of energy use, dynamic time warping, and time-window statistics with machine learning are compared for determining building type from measured electricity use.

Bass, Brett↗

Power Electronics Materials and Bonded Interfaces - Reliability and Lifetime

Advanced packaging technologies are currently being designed and developed by the power electronics industry however, the maximum operating temperature is still limited to 175 degrees Celsius for the silicon carbide devices. Bonded materials such as sintered copper and polymeric materials are potential candidates for high temperature operation, but it is critical to characterize and evaluate its reliability under harsh operating conditions. In this project, we discuss the results of the accelerated experiments conducted on sintered copper and polymeric materials. Additionally, a novel framework to develop the lifetime prediction model of bonded interfaces through employing statistical and machine learning models are described. In this task, scanning acoustic microscope images of bonded interfaces obtained under thermal cycling experiments are used as the data.

ENGINEERING↗

A Causal Approach to Model Validation and Calibration

This poster presents a novel method for validation and verification that focuses on identifying causal relationships between data elements, moving beyond traditional statistical and machine learning approaches. These methods employ causal discovery techniques to reveal the underlying mechanisms of data generation. The research utilizes structural causal models and directed acyclic graphs to depict causal relationships. This approach assists in achieving alignment between simulation models and reality.

97 MATHEMATICS AND COMPUTING↗