Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “CNN”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Dynamical Sketching for Enhanced Communication Efficiency in Federated Learning

Federated learning (FL) has revolutionized distributed machine learning by enabling collaborative model training without sharing local data. However, communication efficiency and privacy guarantees remain significant challenges. This paper introduces a dynamic sketching mechanism in FL, optimizing the trade-off between communication efficiency and model accuracy. By dynamically selecting the sketch matrix size, our approach adapts to the evolving characteristics of the data and the model, ensuring optimal performance across diverse scenarios. We leverage Bayesian optimization to systematically tune the sketch parameters, achieving an effective balance between resource efficiency and model performance. Experimental results on the MNIST dataset using a convolutional neural network (CNN) architecture validate the proposed method's efficiency and scalability. Our dynamic sketching approach significantly outperforms fixed-size sketching techniques, achieving higher compression ratios (up to 62x) and providing better privacy guarantees while maintaining high model accuracy. These findings highlight the robustness and versatility of our approach and make it a valuable solution for privacy-preserving, communication-efficient federated learning.

Afrose, Sharmin [ORNL]↗

Crack Identification and Characterization in Deformed Nb3Sn Rutherford Cable Stacks Using Machine Learning

An investigation of instance segmentation of cracks in Nb3Sn 4-stack 40-strand Rutherford cables using machine learning is presented. Three samples were uniaxially and biaxially loaded before metallographic inspections were performed. The Mask R-CNN model was used in the Detectron2 framework with pre-trained weights but fine-tuned to detect and segment cracks. The model detected cracks with bounding box and mask average precisions (AP) of 42.8 and 27.9, respectively, and was used for instance segmentation of all cracks in the three samples. More cracks were found in the sample pre-loaded along the z-axis (i.e., along the cable length). Pre-loading along the x-axis (i.e., on the cables edges) reduced the number of cracks and changed the crack orientation distribution, away from being highly aligned with the y-axis (i.e., normal to the cables broad faces), i.e., the direction with the highest applied load. Fine-tuning of the Segment Anything Model (SAM) was also studied but performed poorly without human-provided prompts. However, the zero-shot capability of SAM showed high promises to accelerate the image annotation process for applications beyond this study.

Croteau, Jean-Francois↗

A Weakly Supervised Machine Learning Procedure for Magnet Quench Diagnostics

Voltage taps remain the standard and reliable diagnostic tool for detecting quenches in superconducting magnets. However, they identify a quench only at the time of voltage rise and do not provide information on earlier physical precursors. In this work, we investigate whether acoustic emission data can reveal precursor activity that occurs before conventional voltage detection using machine learning techniques. We introduce an event selection method and a weakly supervised machine learning procedure to learn data-driven criteria for identifying potential acoustic precursors to quenches. Two Convolutional Neural Network (CNN) architectures are trained: one on acoustic sensor events from our selection procedure and one on the Fast Fourier Transforms (FFTs) of these events. Both networks are trained iteratively using confidence-weighted loss functions to associate certain subsets of training data with a precursor label. We evaluate the performance of these models by examining the time distribution of events classified as potential precursors relative to the quench onset. Results indicate that the proposed approach can possibly distinguish acoustic emission events occurring closer to the quench from earlier acoustic activity during ramping, suggesting the potential for flagging quench precursors in acoustic data.

Khan, Maira [Fermilab] (ORCID:0009000891602387)↗

A Two-Stage Quantum Reinforcement Learning Method for Multi-Objective Transmission Switching

Multi-objective transmission switching (MO-TS) problems involve the strategic reconfiguration of network topology to simultaneously optimize multiple objectives. As the system scale increases, finding feasible solutions becomes increasingly challenging due to the problem's nonlinearity and high computational complexity. To address these challenges, this paper proposes a two-stage quantum reinforcement learning method that leverages potential quantum advantages for MO-TS. In the first stage, candidate switching lines are identified using a graph-theoretical approach to reduce the problem's dimensionality. The second stage introduces a quantum-classical reinforcement learning framework, where a learnable measurement-based CNN-ResVQC architecture is developed to effectively reduce the input dimension for quantum processing, mitigate vanishing gradients, and enhance trainability while improving the quantum circuit's flexibility in modeling complex decision policies for MO-TS. Numerical studies on IEEE 14-bus, 57-bus, and 118-bus systems demonstrate that the proposed algorithm achieves superior training stability and faster convergence with approximately 1% of the network parameters required by classical algorithms, highlighting its effectiveness, efficiency, and scalability. Furthermore, the practicality is validated through its stable convergence under three common quantum noise channels.

99 GENERAL AND MISCELLANEOUS↗

NeurIPSCosmicNoonMergerID

This work uses TNG50, HST CANDELS imaging, and Zoobot to create a CNN to identify galaxy mergers near cosmic noon. This was accepted to NeurIPS ML4PS 2025.

Schechter, AimeeL. [Univ. of Colorado, Boulder, CO↗

Surrogate modeling of Cellular-Potts agent-based models as a segmentation task using the U-Net neural network architecture

The Cellular-Potts model is a powerful and ubiquitous framework for developing computational models for simulating complex multicellular biological systems. Cellular-Potts models (CPMs) are often computationally expensive due to the explicit modeling of interactions among large numbers of individual model agents and diffusive fields described by partial differential equations (PDEs). In this work, we develop a convolutional neural network (CNN) surrogate model using a U-Net architecture that accounts for periodic boundary conditions. We use this model to accelerate the evaluation of a mechanistic CPM previously used to investigate in vitro vasculogenesis. The surrogate model was trained to predict 100 computational steps ahead (Monte-Carlo steps, MCS), accelerating simulation evaluations by a factor of 562 times compared to single-core CPM code execution on CPU. Over short timescales of up to 3 recursive evaluations, or 300 MCS, our model captures the emergent behaviors demonstrated by the original Cellular-Potts model such as vessel sprouting, extension and anastomosis, and contraction of vascular lacunae. This approach demonstrates the potential for deep learning to serve as a step toward efficient surrogate models for CPM simulations, enabling faster evaluation of computationally expensive CPM simulations of biological processes.

97 MATHEMATICS AND COMPUTING↗

Status of the Muon Neutrino Charged-Current Mesonless Cross Section Measurement in the NOvA Near Detector

NOvA is a long-baseline accelerator neutrino experiment at Fermilab whose physics goals include precision neutrino oscillation as well as cross-section measurements. We present the status of the measurement of a muon neutrino charged-current cross section with zero mesons in the final state at the NOvA near detector. This measurement is being made with respect to the kinematics of the final state muon. The chosen interaction channel is especially sensitive to quasielastic and meson exchange current interactions and aims to provide experimental constraints for the development of models of neutrino interactions. It will also provide a handle for constraining cross section systematic uncertainties in oscillation analyses in current and future experiments. For particle identification, we use a convolutional neural network (CNN) trained on individual particles simulated in the NOvA near detector that allows us to select the desired signal while reducing the potential bias from neutrino interaction modeling. Charged pion background constraining is further improved via Michel electron tagging.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Deep learning with mixup augmentation for improved pore detection during additive manufacturing

In additive manufacturing (AM), process defects such as keyhole pores are difficult to anticipate, affecting the quality and integrity of the AM-produced materials. Hence, considerable efforts have aimed to predict these process defects by training machine learning (ML) models using passive measurements such as acoustic emissions. This work considered a dataset in which keyhole pores of a laser powder bed fusion (LPBF) experiment were identified using X-ray radiography and then registered both in space and time to acoustic measurements recorded during the LPBF experiment. Due to AM’s intrinsic process controls, where a pore-forming event is relatively rare, the acoustic datasets collected during monitoring include more non-pores than pores. In other words, the dataset for ML model development is imbalanced. Moreover, this imbalanced and sparse data phenomenon remains ubiquitous across many AM monitoring schemes since training data is nontrivial to collect. Hence, we propose a machine learning approach to improve this dataset imbalance and enhance the prediction accuracy of pore-labeled data. Specifically, we investigate how data augmentation helps predict pores and non-pores better. This imbalance is improved using recent advances in data augmentation called Mixup, a weak-supervised learning method. Convolutional neural networks (CNNs) are trained on original and augmented datasets, and an appreciable increase in performance is reported when testing on five different experimental trials. When ML models are trained on original and augmented datasets, they achieve an accuracy of 95% and 99% on test datasets, respectively. We also provide information on how dataset size affects model performance. Lastly, we investigate the optimal Mixup parameters for augmentation in the context of CNN performance.

42 ENGINEERING↗

Status of the Muon Neutrino Charged-Current Zero Mesons Cross Section at the NOvA Near Detector

NOvA is a long-baseline accelerator neutrino experiment at Fermilab whose physics goals include precision neutrino oscillation as well as cross-section measurements. We present the status of the measurement of a muon neutrino charged-current cross section with zero mesons in the final state at the NOvA near detector. This measurement is being made with respect to the kinematics of the final state muon. The chosen interaction channel is especially sensitive to quasielastic and meson exchange current interactions and aims to provide experimental constraints for the development of models of neutrino interactions. It will also provide a handle for constraining cross section systematic uncertainties in oscillation analyses in current and future experiments. For particle identification, we use a convolutional neural network (CNN) trained on individual particles simulated in the NOvA near detector that allows us to select the desired signal while reducing the potential bias from neutrino interaction modeling. Charged pion background constraining is further improved via Michel electron tagging.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Computer vision models and advanced TEM imaging for microstructures of irradiated AM316 stainless steels

Advancements were made in automating microscopy-based material characterization, particularly in studying irradiation effects on additively manufactured (AM) materials using machine learning (ML) and computer vision (CV). These automation efforts address the challenges of analyzing complex microstructures, accelerating the detection of irradiation-induced defects. Two CV models were developed at Argonne National Laboratory (ANL) to enhance transmission electron microscopy (TEM) analysis of irradiated AM 316 stainless steel. The first model focused on the detection of irradiation-induced dislocation loops, which contribute to material hardening and embrittlement. These loops, categorized as faulted or perfect, were automatically detected and classified using a Mask R-CNN model trained on TEM images from both in-situ and ex-situ ion irradiation experiments. The model achieved high accuracy, with precision, recall, and F1 scores of 0.839, 0.734, and 0.776, respectively, demonstrating its effectiveness in analyzing dislocation loops in irradiated AM materials. The second CV model was developed to analyze the size and wall thickness of dislocation cells in laser powder bed fusion (LPBF) 316 stainless steel. Using a U-Net++ architecture with EfficientNet as the encoder, the model was trained on TEM images to segment and measure cell size and wall thickness.

36 MATERIALS SCIENCE↗

Towards AI Based Data Classification for Decision Making During Testing

During the development of high-consequence items, test systems should be capable of differentiating between test failures resulting from narrowly missing requirements versus those indicating potentially catastrophic faults. In many instances, classifying the data corresponds to simply identifying whether measured waveforms have approximately the anticipated shape. Cast in this light, the problem reduces to converting raw data into a form optimal for use with neural network classifiers. This manuscript investigates different means of representing raw data for image classification. Raw data plots and Short Time Fourier Transform (STFT) spectrograms are classified by both custom built, small-scale, Convolution Neural Networks (CNN) and open-source, multi-million parameter, pre-trained deep CNNs. In the case of time varying frequency content, the STFTs provide images with greater detail and can be accurately classified with simpler networks. This requires less memory and runs faster than classifying the raw data using the more sophisticated options—making STFTs optimal for applications with memory constraints. STFTs are not a panacea. In some cases the time-domain signal contains useful information that should not be discarded. Rather than using raw data or STFTs, the images can be constructed from both by using red and green channels of an RGB image to visualize the real and imaginary components of the transform, with the raw data occupying the blue channel.

97 MATHEMATICS AND COMPUTING↗

Investigating spatial variability of aerosol, cloud condensation nuclei, and ice nucleating particles in mountainous terrain

The ASR-supported Surface Atmosphere Integrated field Laboratory (SAIL) in the East River Watershed (ERW) of the Upper Colorado River Basin in southwestern Colorado ran from fall 2021 to spring 2023. Two monitoring sites were deployed in the East River Watershed as part of SAIL. The two sites were the Aerosol Observation System (AOS) located on Crested Butte Ski Mountain, and the ARM Mobile Facility (AMF-2), located at the Rocky Mountain Biological Laboratory in Gothic, Colorado. To gain a more comprehensive understanding of aerosols in complex, mountainous terrain, Handix Scientific deployed SAIL-Net, a distributed network of six measurement nodes spanning the domain of the SAIL research area from October 2021 to July 2023. Each node measured aerosol particles between 140 nm and 3.4 μm in diameter using a small particle counter (POPS, (Gao et al., 2016)), CNN using a miniature CCN counter (CloudPuck), and INP using the Time-Resolved Aerosol Filter Sampler (TRAPS, Creamean et al. (2018)). Our approach was similar to other studies that aimed to better characterize and understand aerosols and gas-phase pollutants using networks of lower-cost sensors (Caubel et al., 2019; Kelly et al., 2021; Asher et al., 2022). Such studies have identified neighborhood-level variations in pollutant concentrations (Schneider et al., 2017; Popoola et al., 2018; Caubel et al., 2019). Small-scale variations such as this are poorly represented in models and poorly measured by a single monitoring system (Caubel et al., 2019). Previous work has shown the representation error (the ability of measurements to represent a larger area) increases with complex orography, leading to decreases in model accuracy (Schutgens et al., 2017). The overall goal of SAIL-Net was to improve our understanding of the variability of aerosol in ERW, thus increasing our knowledge of aerosol-cloud interactions in this region and informing the usefulness of distributed networks of measurements for future studies.

54 ENVIRONMENTAL SCIENCES↗

Fusion of Experiments and Simulations for Real-Time Identification of Pipeline Defects

In this study, we explored fusion of experiments and simulations for real time identification of pipeline defects across physical and non-physical domains. The challenges associated to data processing were addressed and a combined classification models was presented via CNN models. In addition, regression model based on XGBOOST is built to determine the defect location and defect dimension from data-driven features of guided wave signals captured by SMS fiber optic sensor.

deep learning↗

Web-Based Tools for Data-Informed Remedy Optimization: Software Theory and User Guide

This report documents the development and application of two web-based decision-support tools for pump-and-treat (P&T) groundwater remediation systems: PTOLEMY (Pump-and-Treat Optimized Location Evaluation to Maximize Yields) and OPTIMA (Optimization for Pump-and-Treat Implementation, Management, & Assessment). These tools enhance remedy design and management by leveraging advanced computational methods – specifically deep learning and multi-objective optimization – within a user-friendly platform. By integrating data-driven models with established hydrogeological knowledge, PTOLEMY and OPTIMA enable more efficient evaluation of well placement and operational strategies, helping site managers balance multiple remediation objectives under complex conditions. Both tools are implemented as modules within the SOCRATES (Suite Of Comprehensive Rapid Analysis Tools for Environmental Sites) web platform, which provides data access, visualization, and analytics to support remedy optimization across sites in the U.S. Department of Energy Office of Environmental Management complex. PTOLEMY is a rapid screening module designed to identify promising locations for new extraction wells. It employs a multi-channel three-dimensional convolutional neural network (MC3D-CNN) trained on high-fidelity simulation data to predict the relative performance (in terms of contaminant mass recovery) of potential well sites. Through an interactive web interface, PTOLEMY visualizes the probability of high performance across a site, highlighting areas where an extraction well is likely to yield above-threshold contaminant removal over a multi-year period. PTOLEMY’s map-based displays and exportable results support transparent communication of screening analyses. By focusing attention on the most favorable candidate locations, the tool augments traditional engineering judgment and physics-based modeling, providing a data informed basis for subsequent detailed evaluations. OPTIMA is a multi objective optimization module designed to find wellfield layouts and operating schedules that meet various cleanup goals. It quickly evaluates thousands of candidate setups – combinations of well locations, timing, and rates – and returns a small set of best trade-off options for comparison. At its core, OPTIMA uses a U-Net-based surrogate model – a deep-learning emulator of a groundwater flow and transport simulator – to dramatically accelerate scenario evaluations. Coupling this fast surrogate with the NSGA-II (Non-dominated Sorting Genetic Algorithm II) evolutionary algorithm, OPTIMA explores a wide decision space of well locations and schedules to identify Pareto-optimal solutions that trade off key objectives (e.g., minimizing cleanup time, maximizing contaminant mass removal, and minimizing plume extent). The tool outputs a family of optimal configurations and visualizes their trade-offs (Pareto frontiers of cleanup metrics and maps of optimized well placements). Site managers can use these results to understand the range of viable strategies and to select candidate designs for more detailed verification. OPTIMA is currently under active development and not yet fully released; this guide provides early documentation to support planning and gather user feedback.

54 ENVIRONMENTAL SCIENCES↗

Investigating Spatial Variability of Aerosol, Cloud Condensation Nuclei, and Ice Nucleating Particles in Mountainous Terrain Field Campaign Report

The U.S. Department of Energy Atmospheric System Research (ASR)-supported Surface Atmosphere Integrated Field Laboratory (SAIL) campaign in the East River Watershed (ERW) of the Upper Colorado River Basin in southwestern Colorado ran from fall 2021 to spring 2023. Two monitoring sites were deployed in the ERW as part of SAIL. The two sites were the Aerosol Observation System (AOS) located on Crested Butte Ski Mountain, and the second ARM Mobile Facility (AMF2), located at the Rocky Mountain Biological Laboratory in Gothic, Colorado. To gain a more comprehensive understanding of aerosols in complex, mountainous terrain, Handix Scientific deployed SAIL-Net, a distributed network of six measurement nodes spanning the domain of the SAIL research area from October 2021 to July 2023. Each node measured aerosol particles between 140 nm and 3.4 μm in diameter using a small portable optical particle spectrometer (POPS; Gao et al. 2016), cloud condensation nuclei (CNN) using a miniature CCN counter (CloudPuck), and ice nucleating particles (INP) using the time-resolved aerosol filter sampler (TRAPS; Creamean et al. 2018). Our approach was similar to other studies that aimed to better characterize and understand aerosols and gas-phase pollutants using networks of lower-cost sensors (Caubel et al. 2019, Kelly et al. 2021, Asher et al. 2022). Such studies have identified neighborhood-level variations in pollutant concentrations (Schneider et al. 2017, Popoola et al. 2018, Caubel et al. 2019). Small-scale variations such as this are poorly represented in models and poorly measured by a single monitoring system (Caubel et al. 2019). Previous work has shown the representation error (the ability of measurements to represent a larger area) increases with complex orography, leading to decreases in model accuracy (Schutgens et al. 2017). The overall goal of SAIL-Net was to improve our understanding of the variability of aerosol in the ERW, thus increasing our knowledge of aerosol-cloud interactions in this region and informing the usefulness of distributed networks of measurements for future studies. We met this goal by answering the following science questions: 1. What is the aerosol temporal variability, and how does aerosol inhomogeneity vary seasonally? Is there significant seasonal variability in sources, or are short-term meteorological conditions the most important determining factor in sources for cloud nuclei? 2. What is the aerosol spatial variability? What are the aerosol characteristics at cloud base, presumably the particles most representative of those acting as cloud nuclei? 3. How should measurement networks be designed to capture aerosol-cloud interactions, and what do they need to measure? Can a single measurement site accurately represent aerosol properties in regions of complex terrain? SAIL-Net consisted of six measurement nodes spread across the ERW near Crested Butte, Colorado. The primary objective in site placement was to select locations that captured the vertical variation in aerosol properties while also spanning the domain of the SAIL campaign. The elevation of the sites ranged from roughly 2750 m along the valley floor of the ERW to approximately 3500 m near the top of Crested Butte Mountain, which is one of the taller peaks in the ERW. The farthest distance between sites was 14 km, while the closest two sites were approximately 1 km apart. Two of the sites were collocated with the ARM SAIL sites; our instruments sat on top of one of the trailers at AOS and another one of our sites was located in a meadow just above AMF2.

54 ENVIRONMENTAL SCIENCES↗

Prong Segmentation using Point Set Transformers in Multiple View Neutrino Detectors

NOvA is a long-baseline neutrino experiment studying neutrino oscillations by detecting neutrinos from the NuMI beam at Fermilab. Its physics analysis relies on accurate prong segmentation, which involves matching each hit to its source particle and identifying the particle type. This task has commonly been addressed using a combination of traditional clustering algorithms and convolutional neural networks (CNNs). However, NOvA’s detector design presents data as two sparse and decoupled 2D images (XZ and YZ views) rather than a native 3D representation, posing a significant challenge for traditional CNN-based models. In this talk, we propose a novel neural network based on the Point Set Transformer. By treating detector hits as sparse point clouds and implementing a cross-view attention mechanism, our model enables efficient information mixing between both views. Evaluated on NOvA simulated data, our model achieves superior accuracy while requiring significantly fewer computational resources compared to other models. Furthermore, the model demonstrates great performance when applied to Liquid Argon Time Projection Chamber (LArTPC) data, which shows its potential as a universal prong segmentation algorithm for multiple view neutrino detectors.

Liu, Jiaxi [UC, Irvine]↗

SBND Shower Reconstruction with SPINE

The Short-Baseline Near Detector (SBND) is a liquid argon time projection chamber (LArTPC) neutrino detector in the Short-Baseline Neutrino (SBN) program at Fermilab. SBND is designed to investigate the Low-Energy Excess (LEE), an unexplained excess of electron-like events observed by previous short-baseline neutrino experiments that may point to physics beyond the Standard Model. In LArTPC detectors, precise shower reconstruction is essential for distinguishing electrons from photons, a key requirement for testing possible explanations of the LEE and improving $\nu_e$ event selection. In this poster, the reconstruction studies using the Scalable Particle Imaging with Neural Embeddings (SPINE), a machine learning based reconstruction framework for particle imaging detectors will be presented. SPINE combines sparse convolutional neural networks (CNN) and graph neural networks (GNN) to enable detailed reconstruction and characterization of neutrino interactions in LArTPC detectors. Shower calorimetry and kinematic reconstruction are performed in dedicated post-processing stages. Strong agreement between data and Monte Carlo simulation will be demonstrated, indicating high-precision detector calibration and reconstruction performance. The agreement between reconstructed and true electron shower energy will also be discussed, emphasizing the robustness of the shower reconstruction performance. These results demonstrate the unprecedented precision achievable with SPINE in SBND, highlighting their potential for future high-resolution neutrino measurements.

Fan, Castaly [Florida U.; Fermilab] (ORCID:0000000↗