Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “image recognition”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

De novo synthesis and near atomic resolution imaging of host immune receptors critical for pathogen recognition (Abbreviated Final Report)

The innate immune system serves as the body’s first line of defense against invading pathogens, responding rapidly through the deployment of immune cells at common sites of infection, such as the skin and airways. These immune-cell sentinels express a range of highly conserved receptors, including Toll-like receptors (TLRs), which recognize and bind to pathogen-derived components. This recognition event triggers intracellular signaling cascades that initiate and coordinate immune responses. Despite their significance, the complete structural characterization of full-length TLRs, including their extracellular domain (ECD), transmembrane domain (TMD), and Toll/interleukin-1 receptor (TIR) domain, remains incomplete. In this study, we examined the expression and isolation of human TLR4 incorporated into nanodiscs using two approaches: a cell-free synthesis system and a cell-based transfection strategy employing Expi293F cells derived from the human embryonic kidney lineage. Our findings demonstrate that NLP-bound human TLR4 produced via the cell-based method yielded functional protein suitable for time-resolved single-particle cryo-electron microscopy (cryo-EM). This advancement enables structural characterization of full-length TLR4, maintaining the integrity of its extracellular domain, transmembrane domain, and Toll/Interleukin-1 receptor (TIR) domain.

59 BASIC BIOLOGICAL SCIENCES↗

Object Detection and Recognition with PointPillars in LiDAR Point Clouds – Comparisions

In the field of autonomous systems, neural networks have been leveraged for object detection and recognition in 2-dimensional images captured by cameras. Other types of sensors are available for sensing surroundings, including LiDAR sensors, and corresponding networks have been developed to perform detection and recognition in the point clouds generated by these sensors. The approaches are similar, both perform convolutions, but have distinct characteristics and challenges. In designing and configuring autonomous systems, a variety of LiDAR sensors are available, along with configurable deep neural networks to leverage their data. This work presents a review of the PointPillars network, an evolution of the seminal PointNet, comparing accuracy and training time relative to different LiDAR sensors, network and training parameters, CPU and GPU hardware, and the criticality of the use of reflective intensity as a feature. The value of using reflectivity as a predictive feature is explored and quantified to determine if it makes a significant difference in accuracy of the PointPillars network. Two separate LiDAR sensors are utilized, a 16-plane and a 32-plane, and corresponding accuracies and training times with the PointPillars network are evaluated.

LiDAR, machine learning, neural network, object re↗

Integrative Quantitative-Phase and Airy Light-Sheet Imaging

Light-sheet microscopy enables considerable speed and phototoxicity gains, while quantitative-phase imaging confers label-free organelle recognition and metabolic information that are inaccessible by conventional methods. We report the fusion of these two modalities onto a standard inverted microscope that retains compatibility with microfluidics. We describe the utilization of an accelerating Airy-beam light-sheet yielding identical imaging areas with interferometry, and an application in unmasking the effects of cellular noise on metabolic compartmentalization.

Biological sciences, Biological techniques, Micros↗

Machine Vision-based Robot Manipulators for Nuclear Applications - 20370

Decommissioning and dismantling of nuclear facilities are major challenges facing the nuclear industry. Robot manipulators with capabilities of restoring nuclear structures and significantly prolonging nuclear power generation in addition to their applications in decommissioning and dismantling (D and D) would immensely ease the above-mentioned challenge. Ability to remotely sense/monitor and analyze dangerous and hazardous environments such as nuclear reactors as well as performing reparatory tasks in such environments using robot manipulators require additional information from vision sensors. Structural sample collection and inspection as well as restoration require object and position information. Computer vision is used to obtain precise 3D position information. The current work uses the Sawyer robot with the integrated Cognex camera. The research presents an approach to facilitate and improve operations in a nuclear reactor using robotic vision control. The vision control system is modeled with Augmented Image Space-based visual servoing approach. Results showing accurate robot control. Object recognition is used to recognize the image and calculate it poses. To transform coordinates of the object's pose from the camera frame to the robotic frame transformation matrices are employed. Future work will employ a 3D camera (with depth information) using Denso robot. (authors)

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Photon-Sparse, Poisson Light-Sheet Microscopy

Light-sheet microscopy has revolutionized bioimaging by enabling approximately an order of magnitude reduction in specimen irradiance compared to confocal imaging. Here, we introduce a light-sheet imaging system that enables an additional order of magnitude reduction in specimen irradiance by operating at the Poisson limit. To operate at this limit, we integrated classical illumination with single-photon detection and wavelet-based image reconstruction. This integration enabled brightness quantification and object recognition from fewer than one detected photon per image pixel, corresponding to more than 10-fold lower irradiance levels than modern systems. We demonstrate how such photon-sparse imaging can eradicate photobleaching and enable both dim and bright object imaging, thus, further enhancing the related gains of light-sheet microscopy.

60 APPLIED LIFE SCIENCES↗

System and method for identifying objects of interest in images based on likelihood map decluttering

An automatic threat recognition system and method is disclosed for scanning the x-ray CT image of an article to identify the objects of interest (OOIs) contained within the article, which are otherwise not always quickly apparent or discernable to an individual. The system uses a computer to receive information from two-dimensional (2D) image slices from a reconstructed computed tomography (CT) scan image and to produce a plurality of voxels for each slice of the 2D image. The computer analyzes the voxels to create a likelihood map (LM) representing likelihoods that voxels making up the CT image are associated with a material of interest (MOI). The computer further analyzes the LM to construct neighborhoods of voxels within the LM, and classifies each voxel neighborhood based on its features, thereby decluttering the LM to facilitate the process of connecting voxels of a like MOI together to form segments. The computer classifies each candidate segment based on its features, thereby identifying those segments that correspond to objects of interest.

Paglieroni, David W.↗

LIvermore SEM image Tools

Scanning Electron Microscopy (SEM) images provide a variety of structural and morphological information for the characterization of the nanomaterials. This code offers automatic recognition and quantitative analysis of SEM images in a high-throughput manner using computer vision and machine learning techniques. The main function of this application is to extract particle size and morphology information of overlapping nanoparticles and core-shell nanostructures in a user friendly interface. The code is written in C++ with QT environment, and has been tested on MacOSX.

KIM, HYOJIN↗

Completion design improvement using a deep convolutional network

Maximizing stimulated natural and hydraulic fracture network is one of the primary hydraulic fracturing concerns for economic production from a horizontal shale gas well. Geomechanical facies and preexisting fractures in each stage are identified based on similarities in formation characteristics to optimize the locations of perforation clusters. This often requires analyzing large volumes of drilling, Logging While Drilling (LWD) and Measurement While Drilling (MWD) data. In this paper, we develop a methodology that calculates the mechanical specific energy (MSE) using real-time drill string acceleration signals directly from its definition. High resolution vibration signals have been collected using a tri-axial accerlometer, which was an auxiliary tool included in acoustic borehole imager. This technique provides a cost-efficient solution for engineered completion design. Furthermore, we adopt deep Convolutional Neural Network (CNN) with signal processing to build a data pipeline that effectively extracts patterns from dynamic acceleration signals for rock lateral MSE classification. First, we apply discrete wavelet transform and Short-Time Fourier Transform (STFT) for signal denoising and pattern recognition. Then we construct an image dataset using multi-scale image fusion at pixel level from 3 sensor channels, including axial, lateral acceleration spectrograms and zero-padded revolutions per minute (RPM). The resulted RGB image dataset includes 4,000 images of 5 MSE ranges with various rock strength conditions. Our results demonstrate that the proposed deep learning model can achieve more than 90% classification accuracy. The deep learning results, as a reference source, were applied in selected Marcellus Shale Energy and Environmental Lab (MSEEL) wells engineered completion located in the Marcellus shale gas site.

03 NATURAL GAS↗

ChemPix: automated recognition of hand-drawn hydrocarbon structures using deep learning

Inputting molecules into chemistry software, such as quantum chemistry packages, currently requires domain expertise, expensive software and/or cumbersome procedures. Leveraging recent breakthroughs in machine learning, we develop ChemPix: an offline, hand-drawn hydrocarbon structure recognition tool designed to remove these barriers. A neural image captioning approach consisting of a convolutional neural network (CNN) encoder and a long short-term memory (LSTM) decoder learned a mapping from photographs of hand-drawn hydrocarbon structures to machine-readable SMILES representations. We generated a large auxiliary training dataset, based on RDKit molecular images, by combining image augmentation, image degradation and background addition. Additionally, a small dataset of ~600 hand-drawn hydrocarbon chemical structures was crowd-sourced using a phone web application. These datasets were used to train the image-to-SMILES neural network with the goal of maximizing the hand-drawn hydrocarbon recognition accuracy. By forming a committee of the trained neural networks where each network casts one vote for the predicted molecule, we achieved a nearly 10 percentage point improvement of the molecule recognition accuracy and were able to assign a confidence value for the prediction based on the number of agreeing votes. The ensemble model achieved an accuracy of 76% on hand-drawn hydrocarbons, increasing to 86% if the top 3 predictions were considered.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Scaling the training of particle classification on simulated MicroBooNE events to multiple GPUs

Measurements in Liquid Argon Time Projection Chamber (LArTPC) neutrino detectors, such as the MicroBooNE detector at Fermilab, feature large, high fidelity event images. Deep learning techniques have been extremely successful in classification tasks of photographs, but their application to LArTPC event images is challenging, due to the large size of the events. Events in these detectors are typically two orders of magnitude larger than images found in classical challenges, like recognition of handwritten digits contained in the MNIST database or object recognition in the ImageNet database. Ideally, training would occur on many instances of the entire event data, instead of many instances of cropped regions of interest from the event data. However, such efforts lead to extremely long training cycles, which slow down the exploration of new network architectures and hyperparameter scans to improve the classification performance. We present studies of scaling a LArTPC classification problem on multiple architectures, spanning multiple nodes. The studies are carried out on simulated events in the MicroBooNE detector. We emphasize that it is beyond the scope of this study to optimize networks or extract the physics from any results here. Institutional computing at Pacific Northwest National Laboratory and the SummitDev machine at Oak Ridge National Laboratory’s Leadership Computing Facility have been used. To our knowledge, this is the first use of state-of-the-art Convolutional Neural Networks for particle physics and their attendant compute techniques onto the DOE Leadership Class Facilities. We expect benefits to accrue particularly to the Deep Underground Neutrino Experiment (DUNE) LArTPC program, the flagship US High Energy Physics (HEP) program for the coming decades.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Artifact identification in X-ray diffraction data using machine learning methods

In situ synchrotron high-energy X-ray powder diffraction (XRD) is highly utilized by researchers to analyze the crystallographic structures of materials in functional devices ( e.g. battery materials) or in complex sample environments ( e.g. diamond anvil cells or syntheses reactors). An atomic structure of a material can be identified by its diffraction pattern along with a detailed analysis of the Rietveld refinement which yields rich information on the structure and the material, such as crystallite size, microstrain and defects. For in situ experiments, a series of XRD images is usually collected on the same sample under different conditions ( e.g. adiabatic conditions) yielding different states of matter, or is simply collected continuously as a function of time to track the change of a sample during a chemical or physical process. In situ experiments are usually performed with area detectors and collect images composed of diffraction patterns. For an ideal powder, the diffraction pattern should be a series of concentric Debye–Scherrer rings with evenly distributed intensities in each ring. For a realistic sample, one may observe different characteristics other than the typical ring pattern, such as textures or preferred orientations and single-crystal diffraction spots. Textures or preferred orientations usually have several parts of a ring that are more intense than the rest, whereas single-crystal diffraction spots are localized intense spots owing to diffraction of large crystals, typically >10 µm. In this work, an investigation of machine learning methods is presented for fast and reliable identification and separation of the single-crystal diffraction spots in XRD images. The exclusion of artifacts during an XRD image integration process allows a precise analysis of the powder diffraction rings of interest. When it is trained with small subsets of highly diverse datasets, the gradient boosting method can consistently produce high-accuracy results. The method dramatically decreases the amount of time spent identifying and separating single-crystal diffraction spots in comparison with the conventional method.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Long-Range Biometric Identification in Real World Scenarios: A Comprehensive Evaluation Framework Based on Missions

The considerable body of data available for evaluating biometric recognition systems in Research and Development (R&D) environments has contributed to the increasingly common problem of target performance mismatch. Biometric algorithms are frequently tested against data that may not reflect the real world applications they target. From a Testing and Evaluation (T&E) standpoint, this domain mismatch causes difficulty assessing when improvements in State-of-the-Art (SOTA) research actually translate to improved applied outcomes. This problem can be addressed with thoughtful preparation of data and experimental methods to reflect specific use-cases and scenarios.To that end, this paper evaluates research solutions for identifying individuals at ranges and altitudes, which could support various application areas such as counterterrorism, protection of critical infrastructure facilities, military force protection, and border security. We address challenges including image quality issues and reliance on face recognition as the sole biometric modality. By fusing face and body features, we propose developing robust biometric systems for effective long-range identification from both the ground and steep pitch angles. Preliminary results show promising progress in whole-body recognition. This paper presents these early findings and discusses potential future directions for advancing long-range biometric identification systems based on mission-driven metrics.

Aykac, Deniz↗

Evaluating Automated Face Identity-Masking Methods with Human Perception and a Deep Convolutional Neural Network

Face de-identification (or “masking”) algorithms have been developed in response to the prevalent use of video recordings in public places. Here, we evaluated the success of face identity masking for human perceivers and a deep convolutional neural network (DCNN). Eight de-identification algorithms were applied to videos of drivers’ faces, while they actively operated a motor vehicle. These masks were pre-selected to be applicable to low-quality video and to maintain coarse information about facial actions. Humans studied high-resolution images to learn driver identities and were tested on their recognition of active drivers in low-resolution videos. Faces in the videos were either unmasked or were masked by one of the eight algorithms. When participants were tested immediately after learning (Experiment 1), all masks reduced identification, with six of eight masks reducing identification to extremely poor performance. In a second experiment, two of the most effective masks were tested after a delay of 7 or 28 days. The delay did not further reduce identification of the masked faces. In all masked conditions, participants maintained stringent decision criteria, with low confidence in recognition, further indicating the effectiveness of the masks. Next, the DCNN performed an identity-matching task between high-resolution images and masked videos—a task analogous to that done by humans. The pattern of accuracy for the DCNN mirrored some, but not all, aspects of human performance, highlighting the need to test the effectiveness of identity masking for both humans and machines. The DCNN was also tested on its ability to match identity between masked and unmasked versions of the same video, based only on the face. DCNN performance for the eight masks offers insight into the nature of the information in faces that is coded in these networks.

97 MATHEMATICS AND COMPUTING↗

Neutrino event selection in the MicroBooNE liquid argon time projection chamber using Wire-Cell 3D imaging, clustering, and charge-light matching

An accurate and efficient event reconstruction is required to realize the full scientific capability of liquid argon time projection chambers (LArTPCs). The current and future neutrino experiments that rely on massive LArTPCs create a need for new ideas and reconstruction approaches. Wire-Cell, proposed in recent years, is a novel tomographic event reconstruction method for LArTPCs. The Wire-Cell 3D imaging approach capitalizes on charge, sparsity, time, and geometry information to reconstruct a topology-agnostic 3D image of the ionization electrons prior to pattern recognition. A second novel method, the many-to-many charge-light matching, then pairs the TPC charge activity to the detected scintillation light signal, thus enabling a powerful rejection of cosmic-ray muons in the MicroBooNE detector. A robust processing of the scintillation light signal and an appropriate clustering of the reconstructed 3D image are fundamental to this technique. In this paper, we describe the principles and algorithms of these techniques and their successful application in the MicroBooNE experiment. A quantitative evaluation of the performance of these techniques is presented. Using these techniques, a 95% efficient pre-selection of neutrino charged-current events is achieved with a 30-fold reduction of non-beam-coincident cosmic-ray muons, and about 80% of the selected neutrino charged-current events are reconstructed with at least 70% completeness and 80% purity.

3D imaging↗

Quantum optical classifier with superexponential speedup

Abstract Classification is a central task in deep learning algorithms. Usually, images are first captured and then processed by a sequence of operations, of which the artificial neuron represents one of the fundamental units. This paradigm requires significant resources that scale (at least) linearly in the image resolution, both in terms of photons and computational operations. Here, we present a quantum optical pattern recognition method for binary classification tasks. It classifies objects without reconstructing their images, using the rate of two-photon coincidences at the output of a Hong-Ou-Mandel interferometer, where both the input and the classifier parameters are encoded into single-photon states. Our method exhibits the behaviour of a classical neuron of unit depth. Once trained, it shows a constant $${{\mathcal{O}}}(1)$$ O ( 1 ) complexity in the number of computational operations and photons required by a single classification. This is a superexponential advantage over a classical artificial neuron.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Semi-automatic image annotation using 3D LiDAR projections and depth camera data

Efficient image annotation is necessary to utilize deep learning object recognition neural networks in nuclear safeguards, such as for the detection and localization of target objects like nuclear material containers (NMCs). This capability can help automate the inventory accounting of different types of NMCs within nuclear storage facilities. The conventional manual annotation process is labor-intensive and time-consuming, hindering the rapid deployment of deep learning models for NMC identifications. This paper introduces a novel semi-automatic method for annotating 2D images of nuclear material containers (NMCs) by combining 3D light detection and ranging (LiDAR) data with color and depth camera images collected from a handheld scan system. The annotation pipeline involves an operator manually marking new target objects on a LiDAR-generated map, and projecting these 3D locations to images, thereby automatically creating annotations from the projections. The semi-automatic approach significantly reduces manual efforts and the expertise in image annotation that is required to perform the task, allowing deep learning models to be trained on-site within a few hours. The paper compares the performance of models trained on datasets annotated through various methods, including semi-automatic, manual, and commercial annotation services. The evaluation demonstrates that the semi-automatic annotation method achieves comparable or superior results, with a mean average precision (mAP) above 0.9, showcasing its efficiency in training object recognition models. Additionally, the paper explores the application of the proposed method to instance segmentation, achieving promising results in detecting multiple types of NMCs in various formations.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

Nanomaterial Synthesis Insights from Machine Learning of Scientific Articles by Extracting, Structuring, and Visualizing Knowledge

Nanomaterials of varying compositions and morphologies are of interest for many applications from catalysis to optics, but the synthesis of nanomaterials and their scale-up are most often time-consuming and Edisonian processes. Information gleaned from the scientific literature can help inform and accelerate nanomaterials development, but again, searching the literature and digesting the information are time-consuming manual processes for researchers. To help address these challenges, here we developed scientific article-processing tools that extract and structure information from the text and figures of nanomaterials articles, thereby enabling the creation of a personalized knowledgebase for nanomaterials synthesis that can be mined to help inform further nanomaterials development. Starting with a corpus of ~35k nanomaterials-related articles, we developed models to classify articles according to the nanomaterial composition and morphology, extract synthesis protocols from within the articles’ text, and extract, normalize, and categorize chemical terms within synthesis protocols. We demonstrate the efficiency of the proposed pipeline on an expert-labeled set of nanomaterials synthesis articles, achieving 100% accuracy on composition prediction, 95% accuracy on morphology prediction, 0.99 AUC on protocol identification, and up to a 0.87 F1-score on chemical entity recognition. In addition to processing articles’ text, microscopy images of nanomaterials within the articles are also automatically identified and analyzed to determine the nanomaterials’ morphologies and size distributions. To enable users to easily explore the database, we developed a complementary browser-based visualization tool that provides flexibility in comparing across subsets of articles of interest. We use these tools and information to identify trends in nanomaterials synthesis, such as the correlation of certain reagents with various nanomaterial morphologies, which is useful in guiding hypotheses and reducing the potential parameter space during experimental design.

36 MATERIALS SCIENCE↗

Anti-distortion bioinspired camera with an inhomogeneous photo-pixel array

The bioinspired camera, comprising a single lens and a curved image sensor—a photodiode array on a curved surface—, was born of flexible electronics. Its economical build lends itself well to space-constrained machine vision applications. The curved sensor, much akin to the retina, helps image focusing, but the curvature also creates a problem of image distortion, which can undermine machine vision tasks such as object recognition. Here we report an anti-distortion single-lens camera, where 4096 silicon photodiodes arrayed on a curved surface in a nonuniform pattern assimilated to the distorting optics are the key to anti-distortion engineering. That is, the photo-pixel distribution pattern itself is warped in the same manner as images are warped, which correctively reverses distortion. Acquired images feature no appreciable distortion across a 120° horizontal view, as confirmed by their neural-network recognition accuracies. This distortion correction via photo-pixel array reconfiguration is a form of in-sensor computing.

77 NANOSCIENCE AND NANOTECHNOLOGY↗