Engineering PapersSearch

SEARCH · Engineering Papers

Results for “VIDEO DATA”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Utah FORGE 3-2417: Fiber-Optic Geophysical Monitoring of Reservoir Evolution - 2024 Annual Workshop Presentation

This is a presentation on the Fiber-Optic Geophysical Monitoring of Reservoir Evolution by Rice University, presented by Jonathan Ajo-Franklin. This video slide presentation discusses the development of an end-to-end fiber-optic sensing approach for EGS to track the (1) initial zone of fracture creation, (2) zones of connected mechanically compliant fractures, (3) zones of flowing fractures, and (4) the integration of the data into an improved THM model. This presentation was featured in the Utah FORGE R&D Annual Workshop on August 13, 2024.

15 GEOTHERMAL ENERGY

Sea spray aerosol production flux retrieval based on Doppler lidar measurements

Supermicron sea spray aerosols (SSA) play a crucial role in stratocumulus cloud microphysics and drizzle formation. However, our understanding of SSA's spatiotemporal distribution, production mechanisms, and fluxes in the marine boundary layer remains limited. This study introduces a novel approach to determine supermicron SSA production flux and number concentration at various heights above the ocean surface. Our method leverages Doppler lidar data, specifically attenuated backscatter and vertical velocity, collected at a 105 m range gate during the Eastern Pacific Clouds and Aerosols for Precipitation Experiment (EPCAPE) at Scripps Pier, La Jolla, California. To focus on nascent SSA, we analyzed data from periods with wind directions between 225° and 315° and wind speeds exceeding 4 m s −1 . Cloud and precipitation events were excluded using relative humidity, backscatter thresholds, and 2D Video Disdrometer observations. Results show that calculated supermicron SSA number concentrations at 105 m ranged from 0.26 to 2.93 cm −3 for surface wind speeds between 4 and 6 ms −1 . When exceeding the method's Limit of Detection, estimated production fluxes ranged from 0.20 to 1.53 cm −2 s −1 . This methodology has potential for application in various oceanic locations, likely enhancing our understanding of SSA's role in atmospheric processes and climate interactions.

54 ENVIRONMENTAL SCIENCES

Edge AI-Enhanced Traffic Monitoring and Anomaly Detection Using Multimodal Large Language Models

This paper addresses the challenge of traffic monitoring and incident detection in remote areas, utilizing multimodal large language models (LLMs) deployed on edge AI devices. The key novelty of the LLM is to convert real-time video streams into descriptive texts, enabling low-bandwidth transmissions and reliable detection of anomalies and incidents in environments of intermittent connectivity. The model is developed based on fine-tuning open-source LLMs and extending it with multi-modal capabilities to analyze video frames. Our work also involves deploying this model on edge devices such as Nvidia IGX Orin and is planned to be tested in realistic environments in future work. The methodology includes data set curation, iterative model fine-tuning and compression, and hardware-based optimization. This approach aims to enhance traffic safety and response speed in remote areas, marking a significant advancement in the application of AI for traffic monitoring and safety management.

Peruski, Ryan [University of Tennessee, Knoxville

Flow pattern and void fraction characterization in nitrogen/water flows through a diamond-type triply periodic minimal surface lattice

Here, this study presents, to the authors knowledge, the first experimental investigation on the void fraction and flow patterns in two-phase flows through a diamond-type Triply Periodic Minimal Surface (TPMS) lattice. An additively manufactured TPMS structure was tested under upward co-current flow of a water/nitrogen mixture. Superficial velocities varied for a total of 42 test conditions (gas: 0.01–2.4 m/s; liquid: 0.01–2.4 m/s; mass flux: 20–2370 kg/m 2 ·s). High-speed video and X-ray imaging enabled time-averaged void fraction measurements and identified six distinct flow regimes which were used to develop a flow pattern map. Comparison of the void fraction data with correlations from literature demonstrated the Rouhani and Axelsson (1970) [49] correlation modified by Steiner (1993) [52] provided the best agreement, which was improved with empirically fit coefficients. This approach predicted the void fraction with errors < ±20% for 64% of the data, with a mean absolute percent deviation (MAPD) of 23%. The measured frictional pressure drop was compared to correlations from literature which captured the observed trends but did not provide good accuracy. The best agreement was found after optimizing the empirical coefficients of the Muller-Steinhagen & Heck (1986) [56] correlation; this approach captured 33% of the data within ±20% with a MAPD of 46% over the full range and captured 71% of the data within ±20% a MAPD of 13.3% at mass fluxes >1100 kg/m 2 -s. These results provide foundational insight into TPMS two-phase flow behavior and inform modeling and design of advanced heat exchange components incorporating TPMS geometries.

42 - ENGINEERING

Utah FORGE 6-3629: Application of Machine Learning, Geomechanics, and Seismology for Real-Time Decision Making Tools During Stimulation - 2024 Annual Workshop Presentation

This is a presentation on the Cutting Edge Application of Machine Learning, Geomechanics, and Seismology for Real-Time Decision Making Tools During Stimulation by the University of Utah, presented by No'am Zach Dvory. This video slide presentation, by the University of Utah, discussed the technical objectives of developing a real-time decision-making platform to enhance seismic monitoring and risk management during stimulation activities. This presentation was featured in the Utah FORGE R&D Annual Workshop on August 15, 2024.

15 GEOTHERMAL ENERGY

From Data to Insights: A Covariate Analysis of the IARPA BRIAR Dataset for Multimodal Biometric Recognition Algorithms at Altitude and Range

This paper examines covariate effects on fused whole body biometrics performance in the IARPA BRIAR dataset, specifically focusing on UAV platforms, elevated positions, and distances up to 1000 meters. The dataset includes outdoor videos compared with indoor images and controlled gait recordings. Normalized raw fusion scores relate directly to predicted false accept rates (FAR), offering an intuitive means for interpreting model results. A linear model is developed to predict biometric algorithm scores, analyzing their performance to identify the most influential covariates on accuracy at altitude and range. Weather factors like temperature, wind speed, solar loading, and turbulence are also investigated in this analysis. The study found that resolution and camera distance best predicted accuracy and findings can guide future research and development efforts in long-range/elevated/UAV biometrics and support the creation of more reliable and robust systems for national security and other critical domains.

Bolme, David

Uncertainty quantification of fireball features extracted from nuclear test films using computer vision

Films from the US’s historic nuclear testing era comprise the only extensive collection of imagery depicting high-yield detonations. These films offer unique insights into the characteristics of flows occurring on scales that are difficult to replicate experimentally, and they are a valuable source of data for the validation of models used to describe nuclear detonations. In recent work, we implemented modern computer vision and machine learning techniques to extract features of the fireball following nuclear detonation. With a training dataset of fireball films, we fine-tuned a You Only Look Once 11 (YOLO11) model to detect and track the fireball. Applied to a video, the outer bounding box produced in each frame by YOLO11 is used as an input prompt to Meta’s Segment Anything Model 2 (SAM2), which is shown to accurately predict the boundary of the fireball over time with high resolution. These state-of-the-art computer vision foundation models exhibit impressive visual accuracy in their results but lack an output of values that robustly quantify uncertainty in scientific applications. In this paper, we develop procedures for uncertainty quantification of extracted fireball features. We outline the application of a parallel attention mechanism to calculate uncertainty ranges that complement and better pose model validation data. This higher quality fireball validation data may serve to improve prognostic models describing nuclear detonations in support of nuclear forensic and emergency response activities.

Khristy, Joel [ORNL] (ORCID:0000000209963060)

TEAMER: Twin Ocean Power Wave Energy Converter Comprehensive Overview

These files collectively provide a comprehensive overview of the testing process, data analysis, and validation for the Twin Ocean Power device tested at the O.H. Hinsdale Wave Research Laboratory, supported by TEAMER funding. This resource includes an overview of power results for a series of 7 trials. The files included in this comprehensive overview include a comprehensive log sheet for each trial, a summary of all trials, and processing scripts for the raw data. It includes all raw data in .tsv and MATLAB compatible formats, an average power chart, angular velocity charts for each trial, trial metrics, and power output files. This resource includes images of the Twin Ocean Power Wave Energy Converter device components and movement during testing and video recordings of each trial.

16 TIDAL AND WAVE POWER

Developing a Deep Learning-Computer Vision Framework to Monitor Avian Interactions with Solar Energy Facility Infrastructure (Final Technical Report)

The project addressed an inability to monitor avian interactions with photovoltaic (PV) solar energy facilities necessary for understanding PV solar impacts on birds. In the project, machine-vision technology that continuously monitors avian activities at PV solar facilities was developed. The technology includes four machine-learning (ML) models, each of which accomplishes a specific task in detecting birds and classifying their activities in live or recorded videos—detecting and tracking moving objects, differentiating birds from other objects, detecting bird collisions with solar panels, and classifying non-collision bird activities around PV facilities. Major project outcomes include adoption by two of DOE SETO’s SolWEB projects, providing novel observational data on birds to promote co-location of PV solar development and habitat conservation, known as ecovoltaics.

14 SOLAR ENERGY

Virtual Reality for Shoot/No-Shoot Decision Training in Law Enforcement: A Literature Review and Research Agenda

Virtual reality (VR) can materially improve “shoot / no-shoot” (SNS) training by giving officers realistic, repeatable practice making high-stakes decisions under pressure. Traditional tools—live-fire ranges and video simulators—build basics, but they cannot adapt to each officer in real time or fully mirror the complexity of the field. VR closes that gap by creating immersive scenarios that are safer, more flexible, easier to scale across units, and able to capture objective performance data. SNS decisions are not just about marksmanship; they rely on perception, judgment, memory, and the ability to hold fire when a threat is uncertain. Effective training therefore needs realism, decision complexity, and branching outcomes that reflect the true consequences of choices. These elements strengthen recognition of hostile intent while reducing false positives and building the self-control required in ambiguous situations. VR brings specific advantages: dynamic environments, full-body interaction, and the ability to measure performance with precision—enabling targeted feedback and better transfer of learning to the street. At the same time, responsible deployment must address scenario quality (credible environments and behaviors), lawful decision models, and user wellbeing (appropriate stress levels, comfort, and safety). Sandia’s VIPER Lab is positioned to lead this work. The team combines human-performance science, AI/ML, and VR/AR development with a deep equipment bench (e.g., omnidirectional treadmill, eye-tracking, haptics, multiple HMDs). This ecosystem supports building and validating next-generation SNS training that is immersive, measurable, and trustworthy. Bottom line: Investment in VR-enabled SNS training that blends evidence-based design with careful validation and legal safeguards is expected to pay off in safer, more consistent decision-making and improved community trust, delivered through training that is practical to deploy at scale.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF

Unrolled Video Super-Resolution Network with Autoregressive Prior for the Case of Known Motion

Real-time detection and classification of distant objects is necessary for many national security applications. However, when objects are far from the sensor, they occupy only a small number of pixels in the captured video, limiting the amount of visual detail available for recognition. State-of-the-art classification methods typically rely on high-resolution (HR) video streams to capture characteristic object features, but obtaining such detail is challenging for distant objects that occupy only a few pixels. This motivates the development of video super-resolution (VSR) methods that enhance object classification by recovering fine details from low-pixel representations. Current VSR methods rely either on model-based optimization, which is interpretable but computationally expensive, or on learning-based approaches, which are efficient and high-performing but often lack flexibility and interpretability. In this report, we propose an end-to-end trainable unrolled VSR network, UVSRNet, which super-resolves each frame in a video by exploiting sub-pixel motion between neighboring low-resolution (LR) frames as well as incorporating high-frequency detail from previously super-resolved frames. In particular, by unrolling a plug-and-play (PnP) half-quadratic splitting (HQS) algorithm, we leverage a model-based data-fitting module alongside a learning-based autoregressive prior module. This combination yields a method that maintains the flexibility and interpretability of model-based methods while achieving the performance advantages of learning-based methods.

45 MILITARY TECHNOLOGY, WEAPONRY, AND NATIONAL DEF

A multi-scale cognitive interaction model of instrument operations at the Linac Coherent Light Source

The Linac Coherent Light Source (LCLS) is the world’s first x-ray free electron laser. It is a scientific user facility operated by the SLAC National Accelerator Laboratory, at Stanford, for the U.S. Department of Energy. As beam time at LCLS is extremely valuable and limited, experimental efficiency—getting the most high quality data in the least time—is critical. Our overall project employs cognitive engineering methodologies with the goal of improving experimental efficiency and increasing scientific productivity at LCLS by refining experimental interfaces and workflows, simplifying tasks, reducing errors, and improving operator safety and stress. Here, in this study, we describe a multi-agent, multi-scale computational cognitive interaction model of instrument operations at LCLS. Our model simulates the aspects of human cognition at multiple cognitive and temporal scales, ranging from seconds to hours, and among agents playing multiple roles, including instrument operator, real time data analyst, and experiment manager. The model can roughly predict impacts stemming from proposed changes to operational interfaces and workflows. Example results demonstrate the model’s potential in guiding modifications to improve operational efficiency. We discuss the implications of our effort for cognitive engineering in complex experimental settings and outline future directions for research. The model is open source, and the videos of the supplementary material provide extensive detail.

47 OTHER INSTRUMENTATION

Deep Generative Models in Energy System Applications: Review, Challenges, and Future Directions

In recent years, with the advent of mature machine learning products like ChatGPT, Stable Diffusion, and Sora, the world has witnessed tremendous changes driven by the rapid development of generative artificial intelligence (GAI). Beyond applications in text, speech, image, and video creation, deep generative models (DGMs) underpinning these cutting-edge technologies have also been employed by domain researchers to address scientific and engineering challenges. This paper aims to fill a gap in the research community by providing a systematic review of how DGMs have been utilized in energy system applications. After introducing four most popular DGMs, we review and categorize 196 research articles into five focus areas: data generation, forecasting, situational awareness, modeling, and optimal decision-making. Through this classification, we uncover trends in how DGMs are employed for each type of problem, highlighting GAI techniques that contribute to breakthroughs over traditional methods. We discuss limitations in existing literature, engineering challenges, and propose future directions, all tailored to the unique nature of problems in energy system engineering. Our goal is to offer insights for energy system domain researchers, providing a comprehensive view of existing studies and potential future opportunities.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Utah FORGE 3-2514: Strain Sensing Array for Deformation Characterization - 2024 Annual Workshop Presentation

This is a presentation on the Strain Sensing Array to characterize deformation at the FORGE site project by Clemson University, presented by Lawrence Murdoch. The project's objective was to evaluate the feasibility of measuring and interpreting tensor strain data to improve the performance of EGS. This presentation was featured in the Utah FORGE R&D Annual Workshop on August 23, 2024.

15 GEOTHERMAL ENERGY

The infrared imaging video bolometer at Wendelstein 7-X

The radiated power distribution is a crucial aspect of heat transport, heat load mitigation, and plasma exhaust performance. An imaging video bolometer camera has been installed to measure the plasma radiated power in the divertor region of the Wendelstein 7-X (W7-X) stellarator. This diagnostic offers a wide-angle (40° × 68°) sampling of the plasma volume in both the poloidal and toroidal directions. The field-of-view is covered with a large number (> 500) of bolometer channels, providing imaging capability. The diagnostic design is introduced here together with its data analysis procedure. A set of laboratory experiments is performed to assess the thermal properties of the gold absorber foil and their spatial uniformity. Following installation, a heat source originating from the inertially cooled front of the diagnostic is identified and filtered out. The discharge data indicate a satisfactory signal-to-noise ratio as well as spatiotemporal resolution. These represent the first toroidally resolved images of the line-integrated radiated power in the W7-X island divertor. The diagnostic was then upgraded with a thinner platinum absorber and adjusted mirrors. Early data from the most recent experimental campaign employing the upgraded design show a considerable improvement in the diagnostic performance with more bolometer channels (> 1400), extended coverage, and increased spatial resolution.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Segmentation Model Distillation [Poster]

The process of training object detection (OD) or image segmentation model requires both a substantial amount of data and technical knowledge, which often creates challenges in applying these types of models to their full potential. In order to streamline the process of developing these models, we propose a new pipeline where a foundation model assists in the dataset generation. Then this resulting dataset is used to fine-tune a fast light-weight model to perform the custom segmentation or OD. This resulting model is also fit for real-time image segmentation, such as in a video stream.

97 MATHEMATICS AND COMPUTING

Quantifying Operational Drivers of Multimodal Biometric Verification in Aerial Surveillance

Multimodal biometric verification is increasingly applied across operational contexts ranging from close-range security cameras and building-mounted surveillance to long-range ground sensors and unmanned aerial system (UAS) imagery. Variations in acquisition conditions—such as image resolution, viewing geometry, and motion artifacts—pose significant challenges for cross-domain algorithmic generalization. This study evaluates two independent multimodal biometric verification systems developed under the Intelligence Advanced Research Projects Activity (IARPA) Biometric Recognition and Identification at Altitude and Range (BRIAR) program, comparing performance on close-range and aerial datasets. Close-range video served as a baseline to quantify the decline in verification performance on aerial footage. The dataset included six UAS platforms, spanning small quadcopters at 10m altitude to medium-sized fixed-wing aircraft at 360m. Mixed-effects logistic regression identified image resolution (head and body pixel counts), head height, sensor characteristics, and algorithm selection as primary determinants of verification success, whereas demographic attributes and mission gait were not significant predictors. Activity type and collection site influenced performance in close-range data but had negligible impact on UAS imagery. These results clarify modality-specific strengths and limitations and highlight opportunities to enhance cross-domain biometric verification.

Peluso, Alina [ORNL] (ORCID:0000000328950406)

AMVOS: Additive Manufacturing Video Object Segmentation Dataset

This dataset provides labeled video frames from four additive manufacturing (AM) processes for video object segmentation (VOS) tasks. It contains 90 video segments comprising 900 individually annotated frames across five AM datasets: laser hot-wire directed energy deposition (LHW-DED), tungsten inert gas wire arc additive manufacturing (TIG-WAAM), plasma arc welding (PAW), visible-light polymer extrusion (visPolymer), and near-infrared polymer extrusion (irPolymer). Each video segment consists of 10 contiguous frames with corresponding pixel-level object instance annotations. Depending on the process, two of four object classes are labeled per frame: Melt Pool, Feed Wire, Nozzle, or Material. Raw frames are provided as .jpg files and annotations as palettized .png files. The dataset follows the directory structure of established VOS benchmarks (DAVIS, YouTube-VOS, MOSE), enabling direct integration into VOS model training and evaluation pipelines for foundation model fine-tuning, domain adaptation, or zero-shot performance benchmarking. Data was collected at Oak Ridge National Laboratory's Manufacturing Demonstration Facility.

Wetzel, Jon [ORNL]