Engineering PapersSearch

SEARCH · Engineering Papers

Results for “VIDEO DATA”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

YOLO11 to SAM2 pipeline for feature extraction from nuclear test films

The response to the effects of nuclear detonations is supported by models that describe the evolution of the nuclear fireball and cloud and the associated transport of active debris. Validation of those descriptions relies on data from the nuclear test operations. Video records of those events offer a rich source of information that was exploited to a limited extent in historic analyses. Computer vision and machine learning techniques are powerful tools that can be used to increase the number of measurements that can be obtained from those films. In this work, we apply computer vision techniques to automatically track the temporal evolution of the nuclear fireball. In particular, we apply You Only Look Once 11 (YOLO11) and Segment Anything Model 2 (SAM2) in combination with minimal human intervention to digitized versions of the original nuclear test films. As part of the proposed workflow, the YOLO11 model is applied to films to determine bounding boxes for the fireball within each frame. These are then used as inputs to SAM2, which uses image segmentation to determine the fireball boundaries and their temporal evolution. We assess the accuracy of our approach by using it to determine the energy released during the Trinity nuclear test and comparing the results with previous analyses based on manual measurements.

Van Exel, Kimberly [ORNL] (ORCID:0009000877463894)

Automated Vehicle Feasibility Study

This study collected automated vehicle (AV) performance data on public roadways in Athens, Ohio. The route for the study contained a combination of roads with different functional classifications, conditions, annual average daily traffic, and ownership responsibilities for maintenance and repair. Preparation for the public road deployment was done in a controlled environment at Transportation Research Center’s SMARTCenter, a dedicated AV test facility in East Liberty, Ohio. Researchers analyzed data and extracted insights relevant for both AV developers and infrastructure owners and operators. The study found that rural environments offer a unique set of roadway features such as hills and curves, which can challenge the driving behavior of an AV. Rural regions can also contain a large number of low-traffic gravel roads that lack pavement markings, which appear to be a crucial infrastructure element for operation of current generation AVs. Similarly, the presence of well-maintained lane lines along curves can influence the AV’s roadway departure tendencies. The study found that curvature-related behavior of an AV is also influenced by driving speed on the roadway segment. Such findings were consistent regardless of the time of day along the route or season of data collection. However, commentary about AV performance in active adverse weather cannot be made, as this is still an area of active research.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

MultiTaskDeltaNet: change detection-based image segmentation for operando ETEM with application to carbon gasification kinetics

Transforming in situ transmission electron microscopy (TEM) imaging into a tool for spatially-resolved operando characterization of solid-state reactions requires automated, high-precision semantic segmentation of dynamically evolving features. However, traditional deep learning methods for semantic segmentation often face limitations due to the scarcity of labeled data, visually ambiguous features of interest, and scenarios involving small objects. To tackle these challenges, we introduce MultiTaskDeltaNet (MTDN), a novel deep learning architecture that creatively reconceptualizes the segmentation task as a change detection problem. By implementing a unique Siamese network with a U-Net backbone and using paired images to capture feature changes, MTDN effectively leverages minimal data to produce high-quality segmentations. Furthermore, MTDN utilizes a multi-task learning strategy to exploit correlations between physical features of interest. In an evaluation using data from in situ environmental TEM (ETEM) videos of filamentous carbon gasification, MTDN demonstrated a significant advantage over conventional segmentation models, particularly in accurately delineating fine structural features. Notably, MTDN achieved a 10.22% performance improvement over conventional segmentation models in predicting small and visually ambiguous physical features. This work bridges key gaps between deep learning and practical TEM image analysis, advancing automated characterization of nanomaterials in complex experimental settings.

08 HYDROGEN

Frontal Slice Approaches for Tensor Linear Systems

Inspired by the row and column action methods for solving large-scale linear systems, in this work, we explore the use of frontal slices for solving tensor linear systems. In particular, this paper presents a novel approach for using frontal slices of a tensor $\mathcal{A}$ to solve tensor linear systems $\mathcal{A} ∗\mathcal{X} = \mathcal{B}$ where ∗ denotes the $t$-product. In addition, we consider variations of this method, including cyclic, block, and randomized approaches, each designed to optimize performance in different operational contexts. Our primary contribution lies in the development and convergence analysis of these methods. Experimental results on synthetically generated and real-world data, including applications such as image and video deblurring, demonstrate the efficacy of our proposed approaches and validate our theoretical findings.

Luo, Hengrui

Spatiotemporally Registered In-Situ and Ex-Situ Datasets for Laser-based Blown Powder Directed Energy Deposition

This dataset is comprised of in situ sensing data collected during laser-based, blown powder directed energy deposition (DED) of Inconel 718 representing eight different printing conditions: (1) nominal, (2) +15% scan speed, (3) +12% laser power, (4) +42% powder feed rate, (5) +100% jerk limit, (6) +10% layer height, (7) +20% hatch spacing, (8) +20% carrier gas flow. All eight DED builds constructed an identical test coupon geometry consisting of geometric features representative of industrial print requirements (e.g., bulk deposition, thin walls, overhangs). In situ data consists of xyz-coordinates (100 Hz) and on-axis melt pool camera video (60 Hz), both of which have been temporally synchronized to spatially map the melt pool camera data. In addition, post-build X-ray computed tomography (XCT) data for each of the eight test geometries have been spatially registered to the recorded xyz-coordinates, allowing for comparisons between melt pool camera data and flaws identified in the XCT data.

additive manufacturing

WHONDRS River Corridor Sediment and Water Geochemistry and In Situ Sensor Data from 7 Perennial and 7 Intermittent Streams across San Antonio, Texas (v3)

This dataset supports a broader study examining the effects of intermittency on sediment respiration. The dataset provides sediment and surface water geochemistry and in situ sensor data from 7 perennial and 7 intermittent streams in San Antonio, Texas. Each stream/site was visited both in summer during base flow (July-September 2023) and winter during peak flow (January-February 2024). Related data were collected and will be published separately in collaboration with A. Veach. The data package was originally published in April 2025. It was updated in June 2025 (v2; modified and new files) and September 2025 (v3; modified files). See the change history section in the readme for more details. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. This dataset is comprised of two folders of field photos and videos, one folder of raw Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS) data and one main data folder containing (1) file-level metadata; (2) data dictionary; (3) field metadata; (4) readme; (5) international generic sample number (IGSN) mapping file; (6) field protocol; (7) a subfolder with sample data; and (8) a subfolder with sensor data. The sample data subfolder contains (1) surface water and sediment dissolved organic carbon (DOC, measured as non-purgeable organic carbon, NPOC) data and averages; (2) surface water and sediment total nitrogen data and averages; (3) sediment grain size data; (4) sediment iron (II) data and averages; (5) wet sediment mass, dry sediment mass, water mass, and wet sediment volume in incubation and sediment ICR vials; (7) sediment incubation respiration rate data and averages; (8) normalized respiration rate data and averages; (9) methods codes; (10) sediment percent carbon and nitrogen; (11) sediment X-ray diffraction (XRD) data; (12) gravimetric moisture and averages; (13) a subfolder with sediment incubation respiration data, scripts, and plots; (14) surface water and sediment FTICR methods; and (15) a subfolder of 9.4 Tesla (9.4T) FTICR-MS data. This folder contains five subfolders, one containing the sediment .xml data files, one containing the water .xml files, one containing the sediment CoreMS output files, one containing the water CoreMS output files, and the other containing instructions and scripts for processing the files in CoreMS (https://github.com/EMSL-Computing/CoreMS). The sensor data subfolder contains (1) a subfolder with miniDOT dissolved oxygen and temperature data and plots; (2) miniDOT dissolved oxygen and temperature summary data; and (3) miniDOT installation methods. All files are .csv, .pdf, .R, .xml, .d, .html, .Rmd, .py, .cal, .json, .jpg, .jpeg, .png, .mov, or .mp4. CORRECTION: The data processing methods for FTICR described in “v3_WHONDRS_AV1_Methods_Codes.csv” mistakenly indicate that users should process the data in Formultitude. The corrected description should read: “Both unprocessed and processed data are provided to allow users flexibility in data processing. Instructions and scripts for processing the data using CoreMS are included.” CORRECTION: Carbon and nitrogen content are reported as percentages. The current column headers "01395_C_percent_per_mg" and "01397_N_percent_per_mg" are incorrect. These should read "01395_C_percent" and "01397_N_percent" and will be corrected in the next version of this data package.

54 ENVIRONMENTAL SCIENCES

Dataset for "A Microfluidic Spore Chamber for Long-Term Imaging of Single-Spore Hyphal Development"

Understanding the life cycle of fungal spores is essential for elucidating their roles in pathogenesis, dispersal, and survival. However, studying spore development under controlled, spatially defined conditions remains challenging. Here, we present the Spore Chamber, a custom-built microfluidic platform engineered for parallel trapping and long-term imaging of individual spores under defined media conditions, enabling real-time visualization of hyphal development. Using Aspergillus fumigatus as a model organism, we demonstrate that sparse trapping of individual spores within size-matched trap geometries enables long-term time-lapse imaging of key developmental stages, including germination, polarized hyphal elongation, branching, and conidiophore formation. To assess the device’s capacity to resolve morphogenetic responses to exogenous signals, we introduced lipochitooligosaccharides (LCOs) and short-chain chitooligosaccharides (COs). Rhizobium-derived, non-sulfated LCO (nsLCO) mixtures induced enhanced secondary branching (hyperbranching), a response not previously reported in A. fumigatus under these signal conditions, to our knowledge, whereas sulfated LCOs and CO4 did not significantly alter branching patterns. In addition, long-term confinement and imaging revealed rare developmental morphologies previously described primarily in mutant strains, including split conidiophore formation, elongated phialides, microcyclic conidiation, and chlamydospore development. Together, these results establish the Spore Chamber as a targeted microfluidic platform for single-spore phenotyping and long-term developmental analysis, with applications in fungal biology, chemical signaling studies, and host–microbe interaction research. Videos of the observed phenomena are included in this data set.

59 BASIC BIOLOGICAL SCIENCES

Advanced Interactive 3D Visualization Tool for Customizable Analyses of Tomography Datasets in Material Science

Current methods for visualizing and analyzing 3D tomography datasets in materials science often lack the interactivity and depth required for detailed structural insights. This limitation restricts a researchers' ability to accurately interpret complex data, which is critical for advancing material innovations and understanding structural properties. To address this issue, we have developed a novel, web-based interactive 3D visualization and analysis tool from the Trame framework that offers customizable features to enhance data interpretability. The tool allows users to adjust parameters such as visible range, slice planes, data rotation, and layering, providing a more detailed and dynamic view of complex structures. Its user-friendly web interface increases the accessibility and ease of use for both novice and experienced researchers, to visualize large volumetric datasets. The tool supports a diverse range of data formats, making it versatile for various research applications. Unique capabilities include real-time data manipulation, automated feature detection, context-sensitive feedback, and real-time volume calculations and distributions per sliced region or layer, alongside the ability to quickly generate high-quality screenshots and videos for presentations and reports. These advancements offer a comprehensive solution for enhanced 3D data exploration, significantly improving the analysis process and communication of results in materials science.

36 - MATERIALS SCIENCE

An AI-Based 3D Bat Movement Tracking System at Wind Energy Facilities Using Multi-Thermal Video Cameras

The poster at the 15th Wind Wildlife Research Meeting discusses how to leverage the potential of real-time thermal-imaging methodologies in quantifying nocturnal bat activities at wind turbines, using 3D computer vision techniques within a deep learning framework. This innovation enables the automatic detection and classification of bats, birds, and insects in thermal-imaging videos captured at wind turbine sites, facilitating efficient and accurate data analysis for enhanced understanding and mitigation of bat-wind turbine interactions.

AI

An AI-Based 3D Bat Movement Tracking System at Wind Energy Facilities Using Multi-Thermal Video Cameras

The talk at the NAWEA Wind Tech 2024 conference discusses how to leverage the potential of real-time thermal-imaging methodologies in quantifying nocturnal bat activities at wind turbines, using 3D computer vision techniques within a deep learning framework. This innovation enables the automatic detection and classification of bats, birds, and insects in thermal-imaging videos captured at wind turbine sites, facilitating efficient and accurate data analysis for enhanced understanding and mitigation of bat-wind turbine interactions.

AI

Artificial Intelligence-Assisted Daytime Video Monitoring for Bird, Insect, and Other Wildlife Interactions with Photovoltaic Solar Energy Facilities

Studying bird, insect, and other wildlife interactions with photovoltaic (PV) solar energy facilities is difficult due to limited multi-season, multi-site data. Researchers can address such data gaps by combining passive monitoring and artificial intelligence (AI). As a part of the development of AI-enabled avian–solar monitoring software, we collected over 19,000 h of daytime videos at five PV sites across three U.S. regions between 2019 and 2024. We applied a moving object detection and tracking (MODT Version 1) AI model we developed earlier to 4373 h of the footage to extract moving objects in video frames, and human reviewers interpreted the model output and identified 68,646 bird, 25,968 insect, and 169 other wildlife instances to generate the training/validation dataset. We analyzed the data by site, region, and season, considering ground cover and landscapes. Songbirds were most common, with raptors as the next most frequent group. Most notably, no bird collisions were confirmed in our observations collected from the videos. Birds most often flew over or near panels, with the highest observations in the Midwest and Northeast (approximately 30 observations per hour on average) and fewer in the desert Southwest. Other behaviors included perching, foraging, and nesting. Bird abundance peaked during breeding and migration seasons. AI-assisted video monitoring proved effective for non-invasively studying flying wildlife at solar facilities to inform ecologically mindful energy development.

avian mortality

Kivalina Biomass Reactor

This report summarizes work performed under DOE Award DE-EE00010149 to support the reliable operation of a community-scale biochar reactor system in Kivalina, Alaska. The project focused on improving sanitation and waste management in a remote community by assessing the installed system, identifying spare parts, defining key performance indicators (KPIs), preparing operator and maintenance manuals, and developing mobile reporting tools for operational data and KPI tracking. The team also produced training materials and recorded videos to support operator onboarding and continuity. The project demonstrated progress in system readiness, documentation, and digital reporting, while also identifying challenges common to remote deployments, including travel constraints, upstream system failures, and local resource limitations. This work provides a practical framework for improving the operation, monitoring, and future replication of biomass reactor systems in remote communities.

09 BIOMASS FUELS

Deep Learning for Fish Identification from Sonar Data (CRADA 481 Final Report)

In eastern regions of the United States, the American eel is a species of management and regulatory concern because of significant population declines, despite the species’ previous abundance in all tributaries of rivers flowing into the Atlantic Ocean. The American eel is also a candidate for listing under the U.S. Endangered Species Act. While hydropower construction and operation are only one of several factors contributing to this population decline, such a listing could impose additional regulatory challenges for a large number of hydropower projects. In this CRADA project, we improved technologies for identifying migrating eels with the goal of reducing the cost and time required for future American eel hydropower impact assessment and mitigation studies, while maintaining accuracy. We built on results from a previous FOA project (FOA# DE-FOA-0001662), led by the Electric Power Research Institute (EPRI), which developed a highly accurate, deep-learning method for identifying migrating eels from imaging sonar data. The current study aimed to further optimize this deep-learning model, originally designed for image classification, and to develop an object detection software capable of identifying fish from sonar videos in real time, enabling the detection of events like fish migrations and specific species, such as the American eel, at hydropower dams. The data conversion algorithms were packaged as software with a graphical user interface, and the software is evaluated by external collaborators. We focused on the American eel in this project and explored the transferability of the developed deep learning models to the sea lamprey, given the similar body shape and swimming behavior between the two species.

13 HYDRO ENERGY

The Sensor Dilemma in Intelligent Transportation Systems

Intelligent Transportation Systems (ITS) are at the forefront in advancing the way we interact and perceive with the transportation network. This revolution is fueled by the significant advancement in sensor perception technologies such as Radar, LiDAR and Video Imaging which are the most popular modalities for ITS. Real-time perception data from these sensors allows intelligent infrastructure side decision making to improve the energy, efficiency and safety at traffic intersections. As traffic departments across the United States are transitioning from traditional loop detectors / emulators and embracing newer technologies, they are often left with a dilemma in choosing a sensor technology for infrastructure-based perception which is reliable, inexpensive, easy to setup and has robust performance in varying weather conditions. However, choosing a sensor which checks all boxes is not straightforward as every sensor type has unique benefits and drawbacks. Radar is excellent at detecting long range vehicles and weather resistance but lacks high resolution. LiDAR is expensive and weather-sensitive, while cameras provide rich visual data at a low cost but are constrained by lighting and visibility. This study examines Radar, LiDAR and camera sensors capabilities to ascertain whether any of these qualifies as the "best" sensor for ITS perception. Through this evaluation, we hope to draw attention to the necessity of National Renewable Energy Laboratory's (NREL) Infrastructure Perception and Control (IPC) framework which presents a multi-sensor track data fusion engine to assimilate multiple data streams in order to provide robust and reliable perception. While no single sensor can meet all the demands of ITS, a hybrid approach combining multiple sensor modalities like Radar, LiDAR and cameras, offers the most robust solution for enhancing the safety and efficiency in intelligent transportation systems.

33 ADVANCED PROPULSION SYSTEMS

The Sensor Dilemma in Intelligent Transportation Systems: Evaluating Radar, Lidar and Camera: Preprint

Intelligent transportation systems (ITS) are at the forefront in advancing the way we interact with and perceive the transportation network. This revolution is fueled by the significant advancement in sensor perception technologies such as radar, lidar, and video imaging, which are the most popular modalities for ITS. Real-time perception data from these sensors allow intelligent infrastructure-side decision-making to improve the energy, efficiency, and safety at traffic intersections. As traffic departments across the United States transition from traditional loop detectors and emulators and embrace newer technologies, they are often left with a dilemma in choosing a sensor technology for infrastructure-based perception that is reliable, inexpensive, and easy to set up and that has robust performance in varying weather conditions. However, choosing a sensor that checks all these boxes is not straightforward, as every sensor type has unique benefits and drawbacks. Radar is excellent at detecting long-range vehicles and weather resistance but lacks high resolution. Lidar is expensive and weather-sensitive, while cameras provide rich visual data at a low cost but are constrained by lighting and visibility. This study examines radar, lidar, and camera sensor capabilities to ascertain whether any of these qualifies as the "best" sensor for ITS perception. While no single sensor can meet all the demands of ITS, a hybrid approach combining multiple sensor modalities like radar, lidar, and cameras offers the most robust solution for enhancing the safety and efficiency of ITS. Through this evaluation, we hope to draw attention to the necessity of the National Renewable Energy Laboratory's infrastructure perception and control framework, which presents a multisensor track data fusion engine to assimilate multiple data streams in order to provide robust and reliable perception.

33 ADVANCED PROPULSION SYSTEMS

Real-Time Characterization of Salt Aerosols Generated from Static and Sparged Molten Salt

The formation of radionuclide-bearing aerosols in the respirable size range has the potential to significantly influence offsite dose consequences and is, therefore, an important consideration in nuclear facility safety assessments. Molten salt reactor (MSR) developers will likely need to demonstrate an understanding of the conditions under which radionuclide-bearing aerosols may be generated from their reactor under normal operating and accident conditions, as well as the characteristics and transport behavior of these aerosols, to demonstrate to the U.S. Nuclear Regulatory Commission (NRC) that the facility can be operated safely. Recent reviews of the literature identified a lack of experimental data describing the mechanisms of formation and properties (size, concentration, and composition) of salt aerosol particles that are produced from molten salts. Experiments that identify the conditions that lead to radionuclide-bearing salt aerosol releases and quantify the characteristics of salt aerosols formed by different mechanisms are high-priority needs to support MSR licensing. This report describes tests that were conducted within the Argonne Salt Aerosol Test Stand (a sealed vessel and measurement system) to generate salt aerosols from static and sparged molten salts and measure their size and concentration in real-time. The results provide insight into salt aerosol formation by the vapor condensation and bubble bursting mechanisms and inform the potential radiological consequences of aerosol formation from molten fuel salt. Videos of the salt surface were taken during salt sparge tests to observe surface bubble behavior. The data in this report can be used to develop mechanistic source term and accident progression models for MSRs. The real-time salt aerosol characterization technique used in this study will be employed in future integral effects tests that are conducted at an engineering scale to simulate realistic MSR accidents and in future separate effects tests to address additional variables that may impact salt aerosol characteristics (e.g., presence of fission products in salt and humidity in atmosphere).

22 GENERAL STUDIES OF NUCLEAR REACTORS

Automated Classification of Vehicle Movements at Signalized Intersections Using Vehicle Trajectories

Accurate vehicle movement classification through signalized intersections is of paramount importance to the analysis of intersection performance and the optimization of traffic control strategies. Conventional techniques for tracking vehicle turning movements depend on infrastructure-based strategies like human counts, loop detectors, and video analytics, all of which are costly, prone to errors, and spatially constrained. High-frequency trajectory data can be utilized to determine vehicle movement patterns in a scalable and infrastructure-independent method due to the adoption of connected vehicles (CVs). In recent years, several studies have utilized connected vehicle data to generate performance measures. Most of the trajectory-based performance measures approaches, however, require map matching-i.e., extracting geospatial references from maps to identify the movements that individual vehicles make at a signalized intersection. These approaches are often time-consuming and hinder scalability since geographic features need to be provided for an analysis to be conducted. Map matching methods are prone to errors as different map versions change these geographic features. This research presents a novel automatic classification pipeline that uses CV trajectory data to classify vehicle movements at signalized crossings, specifically pass-through left-turn and right-turn maneuvers. The process starts by filtering trips that cross a spatial bounding box that has been defined at the target intersection. Approach and departure headings for each trajectory crossing the boundary are computed and are clustered together to identify dominant movements. The proposed algorithm is used to classify the movement of vehicles at 10 intersections in the state of California, and the results indicate that the algorithm can classify movements at these intersections with varying traffic volumes and road network configurations, all in a map-less framework with no need for conflation of vehicle trajectories to a digital base map.

24 POWER TRANSMISSION AND DISTRIBUTION

Generalist multimodal AI: A review of architectures, challenges and opportunities

Multimodal models are expected to be a critical component to future advances in artificial intelligence. Here, this field is starting to grow rapidly with a surge of new design elements motivated by the success of foundation models in natural language processing (NLP) and vision. It is widely hoped that further extending the foundation models to multiple modalities (e.g., text, image, video, sensor, time series, graph, etc.) will ultimately lead to generalist multimodal models, i.e. one model across different data modalities and tasks. However, there is little research that systematically analyzes recent multimodal models (particularly the ones that work beyond text and vision) with respect to the underling architecture proposed. Therefore, this work provides a fresh perspective on generalist multimodal models (GMMs) via a novel architecture and training configuration specific taxonomy. This includes factors such as Unifiability, Modularity, and Adaptability that are pertinent and essential to the wide adoption and application of GMMs. The review further highlights key challenges and prospects for the field and guide the researchers into the new advancements.

Artificial intelligence (AI)