Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Data segmentation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

Supervisory Control and Data Acquisition for Electrochemical Separation Experimentation

The Python-based program is a laboratory automation tool designed to control and monitor electrochemical systems. The tool was developed for capacitive deionization (CDI) experiments, but it can be used for any system that requires controlled voltage or current segments and multi-parameter monitoring. The program integrates hardware components to run user-defined experimental parameters, providing operational control of a programmable power supply, peristaltic pump, and data acquisition devices. Currently, the program is structured with a workflow that includes an initialization (or pre-run) phase, a main loop, and a post-experiment stabilization (or post-run) phase. The initialization phase prepares and stabilizes the cell, ensuring that the electrodes and solution reach a baseline state before the experiment begins. The main loop consists of multiple voltage segments that repeat, controlling the experiment while recording key parameters such as time, voltage, current, pH, and conductivity. Finally, the post-experiment stabilization phase allows the system to stabilize after the experiment, returning the cell and solution to equilibrium conditions before ending the sequence. The program is designed with four variations, each tailored to different experimental needs. All variations include both the initialization and post-experiment stabilization stages, which run for a set amount of time, voltage, current, and flow rate before and after the main experiment block. The main loop runs for a set number of cycles, as defined by the user input, and each cycle is composed of 2 or 4 segments. The 4 program variations are described as follows: Program 1: The main program includes 2 segments. Each segment is defined to have a set duration, flow rate, voltage, and current. This program measures conductivity, flow rate, voltage, and current. Program 2: The main program expands Program 1 to include 4 segments. Each segment has a specified duration, flow rate, voltage, and current. Like Program 1, it measures conductivity, flow rate, voltage, and current. Program 3: The main program consists of 2 segments, each defined by time, flow rate, voltage, and current. In addition to conductivity, flow rate, voltage, and current, Program 3 collects pH and temperature data through a 4-channel data acquisition device. Program 4: This program independently controls two channels of a multi-channel power supply simultaneously. While conductivity can only be measured for one cell at a time, the dual-channel control makes it possible to operate two cells simultaneously under different voltage/current conditions. The main program includes 2 segments.For each program, all measurements are automatically logged and integrated into a single Excel output file. Data are displayed in numerical format and plotted, both in real time, to track system performance. A key feature of the program is its ability to synchronize all outputs so that every measurement shares a single timestamp, ensuring accurate alignment of voltage, current, pH, conductivity, and pH data.By combining hardware control, real-time monitoring, and unified data collection, this program significantly reduces manual workload and minimizes errors, making it a reliable platform for researchers, engineers, and laboratory technicians conducting CDI experiments, among other electrochemical tests.

Valentino, Lauren [Argonne National Laboratory (AN↗

Super-Resolution Ptychography with Small Segmented Detectors

To overcome the spatial resolution limit set by aperture-limited diffraction in traditional scanning transmission electron microscopy, microscopists have developed ptychography enabled by iterative phase retrieval algorithms and high-dynamic-range pixel array detectors. Current detector designs are limited by the data rate off chip, so a high-pixel-count detector has a proportionally lower frame rate than the few-segment detectors used for differential phase contrast (DPC) imaging. This slower acquisition speed leads to heightened vulnerability to scan noise, drift, and potential sample damage. This creates opportunities for repurposing fast segmented detectors for ptychography by trading a reduction in reciprocal space pixels for an increase in real space pixels. Here, we explore a strategy of oversampling in real space and instead apply detector pixel upsampling during the reconstruction process. Further, we demonstrate the viability of achieving super-resolution ptychography on thin objects using only 2 × 2 detector pixels, surpassing the resolution of integrated DPC (iDPC) imaging. With optimization using simulated datasets and experiments on MoTe 2 /WSe 2 bilayer moiré superlattices, we achieved super-resolution ptychography reconstructions under rapid acquisition conditions (37.5 pA, 1 μs dwell time), yielding over 50% improvements in contrast and information limit compared to annular dark field and iDPC imaging on the same detectors.

2D materials↗

Latent Mechanisms of Polarization Switching from In Situ Electron Microscopy Observations

In situ scanning transmission electron microscopy enables observation of the domain dynamics in ferroelectric materials as a function of externally applied bias and temperature. The resultant data sets contain a wealth of information on polarization switching and phase transition mechanisms. However, identification of these mechanisms from observational data sets has remained a problem due to a large variety of possible configurations, many of which are degenerate. Here, an approach based on a combination of deep learning-based semantic segmentation, rotationally invariant variational autoencoder (VAE), and non-negative matrix factorization to enable learning of a latent space representation of the data with multiple real-space rotationally equivalent variants mapped to the same latent space descriptors is introduced. By varying the size of training sub-images in the VAE, the degree of complexity in the structural descriptors is tuned from simple domain wall detection to the identification of switching pathways. Importantly, this yields a powerful tool for the exploration of the dynamic data in mesoscopic electron, scanning probe, optical, and chemical imaging. Moreover, this work adds to the growing body of knowledge of incorporating physical constraints into the machine and deep-learning methods to improve learned descriptors of physical phenomena.

36 MATERIALS SCIENCE↗

Clustering earthquake signals and background noises in continuous seismic data with unsupervised deep learning

The continuously growing amount of seismic data collected worldwide is outpacing our abilities for analysis, since to date, such datasets have been analyzed in a human-expert intensive, supervised fashion. Moreover, analyses that are conducted can be strongly biased by the standard models employed by seismologists. In response to both of these challenges, we develop a new unsupervised machine learning framework for detecting and clustering seismic signals in continuous seismic records. Our approach combines a deep scattering network and a Gaussian mixture model to cluster seismic signal segments and detect novel structures. To illustrate the power of the framework, we analyze seismic data acquired during the June 2017 Nuugaatsiaq, Greenland landslide. We demonstrate the blind detection and recovery of the repeating precursory seismicity that was recorded before the main landslide rupture, which suggests that our approach could lead to more informative forecasting of the seismic activity in seismogenic areas.

59 BASIC BIOLOGICAL SCIENCES↗

(U) Segmented Scintillator Pitch, Thickness, and Septa Material Effects on the Swank Factor, Quantum Efficiency, and DQE(0) for High-Energy X-Ray Radiography

Scintillator design is an important aspect in constructing a radiographic facility in order to collect the most robust data possible. For the proposed Enhanced Capabilities for Sub-critical Experiment (ECSE) facility in Nevada, we explored a permutation of 13 segmented scintillator designs in order to understand the best performance and trade offs for each design. We calculated the quantum efficiency, Swank factor, and detective quantum efficiency at zero-frequency (DQE(0)) for each scintillator design in order to compare performance.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Categorizing distributed wind energy installations in the United States to inform research and stakeholder priorities

Abstract Background Distributed wind energy adoption in the United States can contribute to the diverse portfolio of energy technologies needed to achieve ambitious decarbonization goals. However, with limited deployment to date, the current distributed wind market must be better understood; these efforts will support the range of stakeholders who will drive successful deployment. This article first distinguishes three categories of distributed wind from existing literature: (1) behind the meter, (2) intended for explicit local load, and (3) physically distributed. A novel methodology to classify individual wind installations into each of these categories is then presented and applied to two data sets of wind installations in the United States to categorize and illuminate distinct segments in the distributed wind market. Results Physically distributed installations, constituted by small to moderately sized projects serving local loads on distribution systems solely because of their proximity to them, account for the highest amount of capacity but the lowest number of installations out of the three categories. The inverse is true for behind-the-meter installations, which are used to serve on-site loads. Installations intended for explicit local load, which are interconnected on the utility side of the distribution system and intentionally built to provide energy to loads on the same distribution system, rank in the middle for both installed capacity and number of installations. Conclusions Distributed wind energy deployment in the United States is geographically widespread, but the extent to which a single category is developed in each state varies. Policies, wind resources, and broad energy technology trends contribute to these deployment patterns. By identifying the extent to which each category of installations exists, decision-makers are empowered with data necessary to tailor research and development programs and address stakeholder priorities through policy and other means, ultimately supporting future deployment.

17 WIND ENERGY↗

MusMorph, a database of standardized mouse morphology data for morphometric meta-analyses

Complex morphological traits are the product of many genes with transient or lasting developmental effects that interact in anatomical context. Mouse models are a key resource for disentangling such effects, because they offer myriad tools for manipulating the genome in a controlled environment. Unfortunately, phenotypic data are often obtained using laboratory-specific protocols, resulting in self-contained datasets that are difficult to relate to one another for larger scale analyses. To enable meta-analyses of morphological variation, particularly in the craniofacial complex and brain, we created MusMorph, a database of standardized mouse morphology data spanning numerous genotypes and developmental stages, including E10.5, E11.5, E14.5, E15.5, E18.5, and adulthood. To standardize data collection, we implemented an atlas-based phenotyping pipeline that combines techniques from image registration, deep learning, and morphometrics. Alongside stage-specific atlases, we provide aligned micro-computed tomography images, dense anatomical landmarks, and segmentations (if available) for each specimen (N = 10,056). Our workflow is open-source to encourage transparency and reproducible data collection.

59 BASIC BIOLOGICAL SCIENCES↗

In-situ TEM investigation of void swelling in nickel under irradiation with analysis aided by computer vision

Understanding the stability of irradiation-induced voids in materials is important for engineering material's swelling behavior under irradiation. In-situ TEM offers a spatial and temporal resolution that is suitable for investigating the evolution of voids under irradiation. However, the in-situ videos have often been too large to be analyzed manually, leaving the valuable data underutilized. For this work, we developed a deep learning-based semantic segmentation model to consistently study the growth and shrinkage of voids in nickel under 1 MeV krypton ion irradiation at various temperatures from 525 °C to 650 °C. With a foil thickness near 100 nm and ion flux of 6.3 × 10 11 ions∙ cm –2 ∙s –1 , the pre-existing voids, which were created beforehand by irradiation at 600 °C to 0.5 dpa, shrank at low temperatures and grew at high temperatures under irradiation, where the transition occurred at 575 °C (~0.5 T M ). The observed stability transition provided new insight for the shrinkage mechanism of voids under irradiation. In addition, an annealing experiment on nickel, previously irradiated at 600 °C to 3 dpa, was performed sequentially at 650 °C to 720 °C to reveal the shrinkage rate of void as a function of temperature and void size. The advantage of combining computer vision and in-situ TEM to obtain comprehensive void evolution was demonstrated.

36 MATERIALS SCIENCE↗

Global normal spectral irradiance in Albuquerque - Data and Resources

The Photovoltaic Systems Evaluation Laboratory (PSEL) at Sandia National Laboratories (SNL) in Albuquerque, NM has measured global normal spectral irradiance nearly continuously from August 2013 to April 2018. During this time other broadband irradiance measurements (global horizontal, direct normal, diffuse horizontal and global normal) and weather variables were also recorded. For this dataset PV Performance Labs (PVPL) has pulled together data from both sources to assemble a full calendar year spectral data set for use in photovoltaic research. It is composed of eight continuous segments of different durations taken from the two-year period September 2013 to August 2015.

photovoltaics↗

Spread Spectrum Time Domain Reflectometry (SSTDR) Digital Twin Simulation of Photovoltaic Systems for Fault Detection and Location

Utilizing spread spectrum time domain reflectometry (SSTDR) to detect, locate, and characterize faults in photovoltaic (PV) systems is examined in this paper. We present a method to obtain the model parameters that are needed to produce digital twin SSTDR responses for PV systems. The digital twin SSTDR responses could be used to predict faults within the PV systems. Here, the model parameters are the reflection and transmission coefficients at each impedance discontinuity in the PV system along with the propagation coefficients across each PV cable segment. We obtain model parameter by applying inverse modeling techniques to experimental SSTDR data associated with PV systems. Our model parameters can be used in any digital twin simulation method for modeling reflectometry in frequency-dependent and complex loads. For validation, we used the model parameters in a graph network simulation engine and adapted it to be used for SSTDR digital twin simulations in PV systems. We produced simulations for 0 to 10 PV modules connected in series. We also simulated SSTDR responses for open circuit disconnections in a PV setup containing 10 PV modules in series. Results show that all but one simulated disconnect locations match experimental disconnection locations of the same setup with an error of less than 5%.

14 SOLAR ENERGY↗

Acoustic Rocket Signatures Collected by Smartphones

Rockets generate complex acoustic signatures that can be detected over a thousand kilometers from their source. While many far-field acoustic rocket signatures have been collected and released to the public, very few signatures collected at distances less than 100 km are available. This work presents a curated and annotated dataset of acoustic signatures of 243 rocket launches collected by a network of smartphones stationed at distances between 10 and 70 km from the launch sites, resulting in 1089 individual recordings. Due to the frequency dependence of atmospheric attenuation and the relatively short propagation distances, higher-frequency features not preserved in most publicly available data are observed. The signals are time-aligned to allow for different segments of the signal (ignition, launch, trajectory, chronology) to be more easily examined and compared. Initial analysis of the features of these rocket launch stages is performed, observed features are compared to those found in the existing literature, and comparisons between signals from launches of different rocket types are made. The dataset is annotated and made available to the public to aid future analysis of the characteristics and source mechanisms of rocket acoustics as well as applications such as rocket detection and classification models.

33 ADVANCED PROPULSION SYSTEMS↗

Data supporting the manuscript "Nanometer Scale Imaging to Develop Quantitative Descriptors of Bipolar Membrane Junction Structure"

This dataset contains atomic force microscopy images associated with the manuscript "Nanometer Scale Imaging to Develop Quantitative Descriptors of Bipolar Membrane Junction Structure" by Maria Kelly, Emily R. Dunn, Ellis A. Spickermann, Josephine N. Gruber, César A. Lasalde-Ramírez, P. N. Romero Zavala, Éowyn Lucas, Ankur Gupta, Harry A. Atwater, and Wilson A. Smith. The dataset contains both raw images as well as segmented images produced by the image processing workflow described in the manuscript. A readme file and meta data file are included to provide additional details regarding the sample identity, image acquisition parameters, and file naming scheme.

36 MATERIALS SCIENCE↗

Efficient Implementation of Artificial Neural Networks for Sensor Data Analysis Based on a Genetic Algorithm

The reliability of many industrial processes depends on the sensor system. However, these sensors can be affected by noise, perturbations and failures. Hence, sensor monitoring and diagnosis are fundamental to guarantee the quality of an industrial process. Nowadays, artificial neural networks (ANN) are widely used in sensor signal processing and diagnosis. However, those ANNs usually require many artificial neurons, being difficult to implement in software and hardware due to their high computational costs. This paper presents an optimized implementation of artificial neurons in ANNs for sensor data analysis using a Genetic Algorithm (GA). The objective of GA is to find an adequate segmentation to reduce the activation function approximation error. One of the advantages of the proposed approach is that the cost function used in GA considers the effect of factors such as the ANN architecture or the number of bits used in arithmetic operations. The proposed ANN implementation technique aims to get the best possible approximation for a specific ANN architecture, making easier its implementation in software and hardware. Simulation and experimental results using FPGA (Field Programmable Gate Array) prove the advantages of the proposed approach for implementing sensor data analysis systems based on ANNs.

D estefani, André↗

RCSB Protein Data Bank: improved annotation, search and visualization of membrane protein structures archived in the PDB

Abstract Motivation Membrane proteins are encoded by approximately one fifth of human genes but account for more than half of all US FDA approved drug targets. Thanks to new technological advances, the number of membrane proteins archived in the PDB is growing rapidly. However, automatic identification of membrane proteins or inference of membrane location is not a trivial task. Results We present recent improvements to the RCSB Protein Data Bank web portal (RCSB PDB, rcsb.org) that provide a wealth of new membrane protein annotations integrated from four external resources: OPM, PDBTM, MemProtMD and mpstruc. We have substantially enhanced the presentation of data on membrane proteins. The number of membrane proteins with annotations available on rcsb.org was increased by ∼80%. Users can search for these annotations, explore corresponding tree hierarchies, display membrane segments at the 1D amino acid sequence level, and visualize the predicted location of the membrane layer in 3D. Availability and implementation Annotations, search, tree data and visualization are available at our rcsb.org web portal. Membrane visualization is supported by the open-source Mol* viewer (molstar.org and github.com/molstar/molstar). Supplementary information Supplementary data are available at Bioinformatics online.

59 BASIC BIOLOGICAL SCIENCES↗

Graph-learning approach to combine multiresolution seismic velocity models

SUMMARY The resolution of velocity models obtained by tomography varies due to multiple factors and variables, such as the inversion approach, ray coverage, data quality, etc. Combining velocity models with different resolutions can enable more accurate ground motion simulations. Toward this goal, we present a novel methodology to fuse multiresolution seismic velocity maps with probabilistic graphical models (PGMs). The PGMs provide segmentation results, corresponding to various velocity intervals, in seismic velocity models with different resolutions. Further, by considering physical information (such as ray path density), we introduce physics-informed probabilistic graphical models (PIPGMs). These models provide data-driven relations between subdomains with low (LR) and high (HR) resolutions. Transferring (segmented) distribution information from the HR regions enhances the details in the LR regions by solving a maximum likelihood problem with prior knowledge from HR models. When updating areas bordering HR and LR regions, a patch-scanning policy is adopted to consider local patterns and avoid sharp boundaries. To evaluate the efficacy of the proposed PGM fusion method, we tested the fusion approach on both a synthetic checkerboard model and a fault zone structure imaged from the 2019 Ridgecrest, CA, earthquake sequence. The Ridgecrest fault zone image consists of a shallow (top 1 km) high-resolution shear-wave velocity model obtained from ambient noise tomography, which is embedded into the coarser Statewide California Earthquake Center Community Velocity Model version S4.26-M01. The model efficacy is underscored by the deviation between observed and calculated traveltimes along the boundaries between HR and LR regions, 38 per cent less than obtained by conventional Gaussian interpolation. The proposed PGM fusion method can merge any gridded multiresolution velocity model, a valuable tool for computational seismology and ground motion estimation.

Geochemistry & Geophysics↗

Network Slicing for Federated Learning in Operational Technology Environment

Industrial Control Systems (ICS) and Supervisory Control and Data Acquisition (SCADA) environments are essential to modern infrastructure, facing challenges in ensuring low-latency, high-throughput communication while mitigating cyber threats. This paper presents a framework integrating Federated Learning (FL) and network slicing with Quality of Service (QoS) to enable real-time monitoring without disrupting OT operations. Leveraging digital twin technology and Network Function Virtualization (NFV), the architecture supports predictive analytics and Industry 4.0 requirements. FL facilitates decentralized model training, preserving data privacy and scalability, though it introduces potential throughput constraints. Network slicing addresses this by creating dedicated virtualized segments optimized for performance and security. Advanced fault tolerance at the container and instance levels enhances system reliability. The proposed architecture ensures high throughput, low latency, and secure orchestration for real-time anomaly detection in OT networks. Performance evaluations validate its efficiency in throughput, deployment, and learning accuracy, providing a robust foundation for future ICS automation and data-driven decision-making.

Delgado, Brian G. Rodiles [University of Texas at ↗

A data driven approach for cross-slip modelling in continuum dislocation dynamics

Cross-slip is a thermally activated process by which screw dislocation changes its glide plane to another slip plane sharing the same Burgers vector. The rate at which this process happens is determined by a Boltzmann type expression that is a function of the screw segment length and the stress acting on the dislocation. In continuum dislocation dynamics (CDD), the information regarding the length of the screw dislocation segment and local stress state on dislocations are lost due to the coarse-grained representation of the density. Here, in this work, a data driven approach to characterize the lost information by analyzing the discrete dislocation configurations is proposed to enable cross-slip modeling in the CDD framework in terms of the coarse-grained dislocation density and stress fields. The analysis showed that the screw segment length follows an exponential distribution, and the stress fluctuations, defined as the difference between the stress on the dislocations and the mean field stress in CDD, follows a Lorentzian distribution. A novel approach for cross slip implementation in CDD employing the screw segment length and stress fluctuation statistics was proposed and rigorously tested by comparing the CDD cross-slip rates with discrete dislocation dynamics (DDD) rates. This approach has been applied in conjunction with three cross-slip models used in DDD simulations differing mainly in the functional form of cross slip activation energy. It was found that different cross-slip activation energy formulations yielded different cross-slip rates, yet the effect on mechanical stress-strain response and dislocation density evolution was minimal for the [001] type loading.

42 ENGINEERING↗

Block segmentation in feature space for realtime object detection in high granularity images

Computer vision has applications in object detection, image recognition and classification, and object tracking. One of the challenges of computer vision is the presence of useful information at multiple distance scales. Filtering techniques may sacrifice details at small scales in order to prioritize the analysis of large-scale features of the image. We present a strategy for coarse-graining multidimensional data while maintaining fine-grained detail for subsequent analysis. The algorithm is based on fixed-size block segmentation in the feature space. We apply this strategy to solve the long-standing challenge of detecting particle trajectories at the Large Hadron Collider in real time.

Computer vision↗