Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “machine vision”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Identifying Neutrino Final States and Energies in MicroBooNE with New Deep-Learning Based LArTPC Reconstruction Frameworks

MicroBooNE, a Liquid Argon Time Projection Chamber (LArTPC) located in the $\nu_{\mu}$-dominated Booster Neutrino Beam at Fermilab, has been studying $\nu_{e}$ charged-current (CC) interaction rates to shed light on the MiniBooNE low energy excess. The LArTPC technology employed by MicroBooNE provides the capability to image neutrino interactions with mm-scale precision. Computer vision and other machine learning techniques are promising tools for image processing that could boost efficiencies for selecting $\nu_{e}$-CC and other rare signals, reduce cosmic and beam-induced backgrounds, and improve the reconstruction of neutrino energies. The MicroBooNE experiment has been at the forefront of developing and testing such techniques for use in physics analyses. In this poster we overview deep-learning based reconstruction methods. We will showcase the use of a recurrent neural network to estimate neutrino energies and present a new reconstruction framework that uses convolutional neural networks to locate neutrino interaction vertices, tag pixels with track and shower labels, and perform particle identification on reconstructed clusters. We will present studies characterizing the performance of these new tools and demonstrate their effectiveness through their use in an inclusive $\nu_{e}$-CC event selection.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Fourier-based three-dimensional multistage transformer for aberration correction in multicellular specimens

High-resolution tissue imaging is often compromised by sample-induced optical aberrations that degrade resolution and contrast. Although wavefront sensor-based adaptive optics (AO) can measure these aberrations, such hardware solutions are typically complex, expensive to implement and slow when serially mapping spatially varying aberrations across large fields of view. Here we introduce AOViFT (adaptive optical vision Fourier transformer)—a machine learning-based aberration sensing framework built around a three-dimensional multistage vision transformer that operates on Fourier domain embeddings. AOViFT infers aberrations and restores diffraction-limited performance in puncta-labeled specimens with substantially reduced computational cost, training time and memory footprint compared to conventional architectures or real-space networks. We validated AOViFT on live gene-edited zebrafish embryos, demonstrating its ability to correct spatially varying aberrations using either a deformable mirror or postacquisition deconvolution. By eliminating the need for the guide star and wavefront sensing hardware and simplifying the experimental workflow, AOViFT lowers technical barriers for high-resolution volumetric microscopy across diverse biological samples.

Alshaabi, Thayer [Howard Hughes Medical Institute,↗

Challenges and Vision for Standardization of Biopolymer Data Sets for Machine Learning

Machine learning (ML) is transforming materials research, yet potential for biopolymer discovery remains constrained by fragmented data and nonstandardized reporting. Biopolymers differ significantly from synthetic polymers, requiring specialized approaches to represent their biosynthetic origins, hierarchical structures, and application-specific metrics. In this Perspective, we identify three core challenges limiting biopolymer representation: information encoding, data quality, and data sharing. We describe the most pressing issues and propose commensurate approaches to address each key challenge. Recommendations include the design and adoption of biopolymer-specific fingerprinting and representation frameworks, development of hybrid human-large language model (LLM) data extraction strategies, and expanding Findable, Accessible, Interoperable, Reusable (FAIR)-compliant repositories. We propose a robust foundation to define interoperable, high-quality data sets that capture the full context of biopolymer materials. Standardized metadata, shared ontologies, and community-driven infrastructure would enable scalable, reproducible workflows and accelerate the ML-driven development of biopolymers.

36 MATERIALS SCIENCE↗

A Smart Vision-Aided RICH (Robotic Interface Control and Handling) System for VULCAN

High-flux neutron beams and high-efficiency detectors enable rapid neutron diffraction measurements at the Engineering Materials Diffractometer (VULCAN) at the Spallation Neutron Source (SNS), Oak Ridge National Laboratory (ORNL). To optimize beam time utilization, efficient sample exchange, alignment, and automated measurements are essential. Recent advances in artificial intelligence (AI) have expanded the capabilities of robotic systems. Here, we report the development of a Robotic Interactive Control and Handling (RICH) system for sample handling at VULCAN, designed to support high-throughput experiments and reduce overhead time. The RICH system employs a six-axis desktop robot integrated with AI-based computer vision models capable of recognizing and localizing samples in real time from instrument and depth-resolving cameras. Vision algorithms combine these detections to align samples with designated measurement positions or place them within complex sample environments such as furnaces. This integration of machine learning-assisted vision with robotic handling demonstrates the feasibility of autonomous sample detection and preparation, offering a pathway toward fully unmanned neutron scattering experiments.

automation↗

Toward Urban Water Security: Broadening the Use of Machine Learning Methods for Mitigating Urban Water Hazards

Due to the complex interactions of human activity and the hydrological cycle, achieving urban water security requires comprehensive planning processes that address urban water hazards using a holistic approach. However, the effective implementation of such an approach requires the collection and curation of large amounts of disparate data, and reliable methods for modeling processes that may be co-evolutionary yet traditionally represented in non-integrable ways. In recent decades, many hydrological studies have utilized advanced machine learning and information technologies to approximate and predict physical processes, yet none have synthesized these methods into a comprehensive urban water security plan. In this paper, we review ways in which advanced machine learning techniques have been applied to specific aspects of the hydrological cycle and discuss their potential applications for addressing challenges in mitigating multiple water hazards over urban areas. We also describe a vision that integrates these machine learning applications into a comprehensive watershed-to-community planning workflow for smart-cities management of urban water resources.

54 ENVIRONMENTAL SCIENCES↗

Machine learning-based real-time monitoring system for smart connected worker to improve energy efficiency

Recent advances in machine learning and computer vision brought to light technologies and algorithms that serve as new opportunities for creating intelligent and efficient manufacturing systems. In this study, the real-time monitoring system of manufacturing workflow for the Smart Connected Worker (SCW) is developed for the small and medium-sized manufacturers (SMMs), which integrates state-of-the-art machine learning techniques with the workplace scenarios of advanced manufacturing systems. Specifically, object detection and text recognition models are investigated and adopted to ameliorate the labor-intensive machine state monitoring process, while artificial neural networks are introduced to enable real-time energy disaggregation for further optimization. The developed system achieved efficient supervision and accurate information analysis in real-time for prolonged working conditions, which could effectively reduce the cost related to human labor, as well as provide an affordable solution for SMMs. The competent experiment results also demonstrated the feasibility and effectiveness of integrating machine learning technologies into the realm of advanced manufacturing systems.

42 ENGINEERING↗

A baseline structure inventory with critical attribution for the US and its territories

Leveraging high performance computing, remote sensing, geographic data science, machine learning, and computer vision, Oak Ridge National Laboratory has partnered with Federal Emergency Management Agency (FEMA) to build a baseline structure inventory covering the US and its territories to support disaster preparedness, response, and recovery. The dataset contains more than 125 million structures with critical attribution, and is ready to be used by federal agencies, local government and first responders to accelerate on-the-ground response to disasters, further identify vulnerable areas, and develop strategies to enhance the resilience of critical structures and communities. Data can be freely and openly accessed through Figshare data repository, ESRI’s Living Atlas or FEMA’s Geodata platform.

96 KNOWLEDGE MANAGEMENT AND PRESERVATION↗

WEBAT (Wind Energy with Bat AI-based Tracker) [SWR-24-121]

The WEBAT (Wind Energy with Bat AI-based Tracker) is a Python-based bat tracking software, integrating machine learning and computer vision with infrared thermal sensors to enhance the monitoring and protection of bats in proximity to wind turbines.

Ryu, Sora [National Renewable Energy Laboratory (N↗

Providing Geospatial Intelligence through a Scalable Imagery Pipeline

This chapter describes ORNL’s (Oak Ridge National Laboratory’s) contributions to imagery preprocessing for geospatial intelligence research and development (R&D) in four sections. First, we discuss challenges involved in building an effective imagery preprocessing workflow and the world-class high-performance computing (HPC) resources at ORNL available to process petabytes of imagery data. Second, we highlight how we developed imagery preprocessing tools over three decades while paving the way for our current cutting-edge machine learning and computer vision algorithms that are impacting humanitarian and disaster response efforts. Third, we discuss how PIPE modules work together to turn raw images into analysis-ready datasets. Fourth, we look toward the future and discuss planned advancements to PIPE and computing trends that will affect geospatial intelligence R&D.

Reith, Andrew↗

Open Data and Deep Semantic Segmentation for Automated Extraction of Building Footprints

Advances in machine learning and computer vision, combined with increased access to unstructured data (e.g., images and text), have created an opportunity for automated extraction of building characteristics, cost-effectively, and at scale. These characteristics are relevant to a variety of urban and energy applications, yet are time consuming and costly to acquire with today’s manual methods. Several recent research studies have shown that in comparison to more traditional methods that are based on features engineering approach, an end-to-end learning approach based on deep learning algorithms significantly improved the accuracy of automatic building footprint extraction from remote sensing images. However, these studies used limited benchmark datasets that have been carefully curated and labeled. How the accuracy of these deep learning-based approach holds when using less curated training data has not received enough attention. The aim of this work is to leverage the openly available data to automatically generate a larger training dataset with more variability in term of regions and type of cities, which can be used to build more accurate deep learning models. In contrast to most benchmark datasets, the gathered data have not been manually curated. Thus, the training dataset is not perfectly clean in terms of remote sensing images exactly matching the ground truth building’s foot-print. A workflow that includes data pre-processing, deep learning semantic segmentation modeling, and results post-processing is introduced and applied to a dataset that include remote sensing images from 15 cities and five counties from various region of the USA, which include 8,607,677 buildings. The accuracy of the proposed approach was measured on an out of sample testing dataset corresponding to 364,000 buildings from three USA cities. The results favorably compared to those obtained from Microsoft’s recently released US building footprint dataset.

97 MATHEMATICS AND COMPUTING↗

Computer vision models and advanced TEM imaging for microstructures of irradiated AM316 stainless steels

Advancements were made in automating microscopy-based material characterization, particularly in studying irradiation effects on additively manufactured (AM) materials using machine learning (ML) and computer vision (CV). These automation efforts address the challenges of analyzing complex microstructures, accelerating the detection of irradiation-induced defects. Two CV models were developed at Argonne National Laboratory (ANL) to enhance transmission electron microscopy (TEM) analysis of irradiated AM 316 stainless steel. The first model focused on the detection of irradiation-induced dislocation loops, which contribute to material hardening and embrittlement. These loops, categorized as faulted or perfect, were automatically detected and classified using a Mask R-CNN model trained on TEM images from both in-situ and ex-situ ion irradiation experiments. The model achieved high accuracy, with precision, recall, and F1 scores of 0.839, 0.734, and 0.776, respectively, demonstrating its effectiveness in analyzing dislocation loops in irradiated AM materials. The second CV model was developed to analyze the size and wall thickness of dislocation cells in laser powder bed fusion (LPBF) 316 stainless steel. Using a U-Net++ architecture with EfficientNet as the encoder, the model was trained on TEM images to segment and measure cell size and wall thickness.

36 MATERIALS SCIENCE↗

Discovering Strong Gravitational Lenses in the Dark Energy Survey with Interactive Machine Learning and Crowd-sourced Inspection with Space Warps

We conduct a search for strong gravitational lenses in the Dark Energy Survey (DES) Year 6 imaging data. We implement a pre-trained Vision Transformer (ViT) for our machine learning (ML) architecture and adopt interactive machine learning to construct a training sample with multiple classes to address common types of false positives. Our ML model reduces ∼236 million DES cutout images to 22,564 targets of interest, including ∼85% of previously reported galaxy–galaxy lens candidates discovered in DES. These targets were visually inspected by citizen scientists, who ruled out ∼90% as false positives. Of the remaining 2618 candidates, 149 were expert-classified as “definite” lenses and 516 as “probable” lenses, for a total of 665 systems, with 147 of these candidates being newly identified. Additionally, we trained a second ViT to find double-source plane lens systems, finding at least one double-source system. Our main ViT excels at identifying galaxy–galaxy lenses, consistently assigning high scores to candidates with high expert assessments. The top 800 ViT-scored images include ∼100 of our “definite” lens candidates. This selection is an order of magnitude higher in purity than previous convolutional neural-network-based lens searches and demonstrates the feasibility of applying our methodology for discovering large samples of lenses in future surveys.

79 ASTRONOMY AND ASTROPHYSICS↗

Vision-based localization for cooperative robot-CNC hybrid manufacturing

Wire and arc additive manufacturing (WAAM) has shown promise in recent years for producing large-scale parts with higher deposition rates than other additive processes. WAAM is often combined with subtractive machining to form a hybrid manufacturing process. This hybrid process can be realized by retrofitting computer numerical control (CNC) machines with deposition heads, adding spindles and deposition heads to robots, or developing part localization methods to transfer parts from an additive cell to a CNC machine. Here, a novel, robot-CNC hybrid configuration is introduced where a maneuverable robot is placed in front of a CNC machine to deposit material within the machine envelop. Furthermore, this method removes the need for part localization and the extensive machine modifications required for retrofitting; however, the problem of robot localization is also added. In this work, the effects of error in vision-based, contactless robot localization on machining parameters in a robot-machine hybrid process were studied. Performance was characterized on an implementation of this system using classical computer vision techniques. In addition, machining simulations were conducted to evaluate the effects of image-induced error on chip thickness, material removal rate, and machining allowance. Initial tests show that computer vision could adequately locate a robot for the hybrid WAAM process within .5 mm.

Hybrid manufacturing↗

3D reconstruction and neural rendering for adversarial machine learning

While evasion attacks on computer vision systems have been widely studied, creating attacks that remain effective under significant changes in viewpoint continues to be challenging. Traditional approaches often rely on affine transformations of images, but these approaches degrade at larger perspective shifts and often produce unrealistic or ineffective perturbations. Recent methods use differentiable renderers to improve viewpoint robustness, but they typically depend on manually constructed 3D models. We introduce a semi-automated pipeline that generates physically printable and perspective-invariant adversarial patches using only a small set of 2D images. Our method integrates 3D reconstruction, neural rendering, adversarial patch optimization, and an object detection victim model into a unified workflow. We use 2D Gaussian Splatting for high fidelity mesh reconstruction and FlexPara for surface parameterization that produces texture maps suitable for patch editing. Together, these components form a fully differentiable pipeline in PyTorch3D that links texture modification to model outputs, enabling efficient optimization of patches that remain effective across many viewpoints. The complete process, from image capture to patch printing and physical evaluation, can be completed within a few hours. We demonstrate the effectiveness of the resulting patches through attacks on the YOLOv8 object detection model and discuss remaining challenges and opportunities for improving robustness and scalability.

Singhvi, Vivaan [ORNL] (ORCID:0009000586288221)↗

Out-of-distribution detection with non-parametric density estimation for models predicting processing history of uranium ore concentrates

The rapid advancement in machine learning (ML) and computer vision (CV) coincides with the growth of interest in deploying these ML/CV models in numerous fields from medicine to social science. Similar to those areas, we have witnessed a great number of works in materials science employing ML/CV models – neural networks in particular – in their studies in recent years. These models have proven to obtain accurate performance in various tasks. However, these models struggle to attain a similar performance when encountering test samples coming from a distribution that is different from the training set. More importantly, they fail without providing any warning to the users. Therefore, we propose a framework for detecting out-of-distribution (OOD) samples to alert users when a human intervention might be necessary in this work. Specifically, we explore the use of a non-parametric density estimation method to detect OOD samples. Here, we assess OOD detection capability of the proposed framework on ML models developed for categorizing precipitation routes of U 3 O 8 when encountering OOD datasets that contain samples (1) undergone different imaging acquisition process, (2) undergone different material synthesis process, and (3) different materials than ID set. Through those experiments, we achieve an average area under the receiver operating characteristic (AUROC) of at least 91% on average in detecting OOD samples. With minimal overhead cost and superior performance, the proposed framework enables a reliable and safe system when deploying in real-world scenarios.

Convolutional neural networks↗

Identifying coins from scrap

A system classifies materials utilizing a vision system that implements a machine learning system, such as a neural network, in order to identify or classify each of the materials as either a monetary coin or not a monetary coin, which may then be sorted into separate groups based on such an identification or classification. Such a system can sort monetary coins from other forms of scrap, which may have been produced from a shredding of end of life vehicles.

Kumar, Nalin↗

Multiple stage sorting

A material sorting system sorts materials utilizing multiple stages of classification and sorting, including a vision system that implements a machine learning system in order to identify or classify each of the materials, and Laser Induced Breakdown Spectroscopy to perform a subsequent classification and sorting of the remaining materials.

97 MATHEMATICS AND COMPUTING↗

Metal sorter

A material sorting system sorts materials utilizing a vision system that implements a machine learning system in order to identify or classify each of the materials, which are then sorted into separate groups based on such an identification or classification determining that the materials are composed of either wrought aluminum, extruded aluminum, or cast aluminum.

Kumar, Nalin↗