Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “label propagation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Direction-optimizing Label Propagation Framework for Structure Detection in Graphs: Design, Implementation, and Experimental Analysis

Label Propagation is not only a well-known machine learning algorithm for classification but also an effective method for discovering communities and connected components in networks. We propose a new Direction-optimizing Label Propagation Algorithm (DOLPA) framework that enhances the performance of the standard Label Propagation Algorithm (LPA), increases its scalability, and extends its versatility and application scope. As a central feature, the DOLPA framework relies on the use of frontiers and alternates between label push and label pull operations to attain high performance. It is formulated in such a way that the same basic algorithm can be used for finding communities or connected components in graphs by only changing the objective function used. Additionally, DOLPA has parameters for tuning the processing order of vertices in a graph to reduce the number of edges visited and improve the quality of solution obtained. We present the design and implementation of the enhanced algorithm as well as our shared-memory parallelization of it using OpenMP. We also present an extensive experimental evaluation of our implementations using the LFR benchmark and real-world networks drawn from various domains. Compared with an implementation of LPA for community detection available in a widely used network analysis software, we achieve at most five times the F-Score while maintaining similar runtime for graphs with overlapping communities. We also compare DOLPA against an implementation of the Louvain method for community detection using the same LFR-graphs and show that DOLPA achieves about three times the F-Score at just 10% of the runtime. For connected component decomposition, our algorithm achieves orders of magnitude speedups over the basic LP-based algorithm on large-diameter graphs, up to 13.2× speedup over the Shiloach-Vishkin algorithm, and up to 1.6× speedup over Afforest on an Intel Xeon processor using 40 threads.

97 MATHEMATICS AND COMPUTING↗

Advanced Semi-Supervised Learning with Uncertainty Estimation for Phase Identification in Distribution Systems

The integration of advanced metering infrastructure (AMI) into power distribution networks generates valuable data for tasks such as phase identification; however, the limited and unreliable availability of labeled data in the form of customer phase connectivity presents challenges. To address this issue, we propose a semi-supervised learning (SSL) framework that effectively leverages labeled and unlabeled data. Our approach incorporates self-training, label spreading, and Bayesian neural networks (BNNs) to enhance phase identification with AMI data. Our method uses an ensemble of multilayer perceptron classifiers in a self-training setup, iteratively adding high-confidence pseudo-labels to improve robustness. We also apply label spread to propagate labels based on data similarity, which enhances generalization across diverse distributions. In addition, we employ a BNNs with uncertainty estimation, boosting confidence in predictions and reducing phase identification errors. In our case study, we achieved approximately 98% +/- 0.08 accuracy with uncertainty using minimal and unreliable labeled data from a real U.S. utility, Duquesne Light Company. Our SSL approach, combined with uncertainty estimation, provides an efficient solution for phase identification in AMI data, ultimately improving the reliability of smart grid applications.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Advanced Semi-Supervised Learning With Uncertainty Estimation for Phase Identification in Distribution Systems

The integration of advanced metering infrastructure (AMI) into power distribution networks generates valuable data for tasks such as phase identification; however, the limited and unreliable availability of labeled data in the form of customer phase connectivity presents challenges. To address this issue, we propose a semi-supervised learning (SSL) framework that effectively leverages labeled and unlabeled data. Our approach incorporates self-training, label spreading, and Bayesian neural networks (BNNs) to enhance phase identification with AMI data. Our method uses an ensemble of multilayer perceptron classifiers in a self-training setup, iteratively adding high-confidence pseudo-labels to improve robustness. We also apply label spread to propagate labels based on data similarity, which enhances generalization across diverse distributions. In addition, we employ a BNNs with uncertainty estimation, boosting confidence in predictions and reducing phase identification errors. In our case study, we achieved approximately 98% +/- 0.08 accuracy with uncertainty using minimal and unreliable labeled data from a real U.S. utility, Duquesne Light Company. Our SSL approach, combined with uncertainty estimation, provides an efficient solution for phase identification in AMI data, ultimately improving the reliability of smart grid applications.

24 POWER TRANSMISSION AND DISTRIBUTION↗

SNM Radiation Signature Classification Using Different Semi-Supervised Machine Learning Models

The timely detection of special nuclear material (SNM) transfers between nuclear facilities is an important monitoring objective in nuclear nonproliferation. Persistent monitoring enabled by successful detection and characterization of radiological material movements could greatly enhance the nuclear nonproliferation mission in a range of applications. Supervised machine learning can be used to signal detections when material is present if a model is trained on sufficient volumes of labeled measurements. However, the nuclear monitoring data needed to train robust machine learning models can be costly to label since radiation spectra may require strict scrutiny for characterization. Therefore, this work investigates the application of semi-supervised learning to utilize both labeled and unlabeled data. As a demonstration experiment, radiation measurements from sodium iodide (NaI) detectors are provided by the Multi-Informatics for Nuclear Operating Scenarios (MINOS) venture at Oak Ridge National Laboratory (ORNL) as sample data. Anomalous measurements are identified using a method of statistical hypothesis testing. After background estimation, an energy-dependent spectroscopic analysis is used to characterize an anomaly based on its radiation signatures. In the absence of ground-truth information, a labeling heuristic provides data necessary for training and testing machine learning models. Supervised logistic regression serves as a baseline to compare three semi-supervised machine learning models: co-training, label propagation, and a convolutional neural network (CNN). In each case, the semi-supervised models outperform logistic regression, suggesting that unlabeled data can be valuable when training and demonstrating value in semi-supervised nonproliferation implementations.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

Anomaly Detection in Flight Operational Data Using Deep Learning

In this session, we demonstrate two recently developed deep learning models for anomaly detection in flight operational data by the Data Sciences Group at NASA Ames Research Center. The first model is Convolutional Variational Auto-Encoder (CVAE) [1], which is an unsupervised deep encoder-decoder model, designed specifically for finding anomalies in heterogeneous multivariate time series data. We will demonstrate its application to finding anomalies in streaming data from NASA’s Digital Information Platform’s Fuser source. CVAE identifies data instances that are not representative of expected nominal behavior as anomalous. Since it is an unsupervised approach, the flagged anomalies will need to be reviewed by the subject matter experts (SMEs) for validation and labeling and is designed to assist with vulnerability discovery within Safety Monitoring System programs. The second model is Robust and Explainable Semi-supervised Anomaly Detection (RESAD) model [2], which builds on CVAE to allow learning from both minimally labeled data (previously reviewed by the SMEs) as well as majority unlabeled data. RESAD takes advantage of graph theoretic techniques to propagate the labels from the labeled data to the unlabeled data based on a pre-defined similarity metric and structures the learned feature space from flight time-series so that data of the same class would cluster tightly together. This model characteristic is enabled by training with an augmented loss function and allows learning of a more informative feature space for down-stream tasks such as search and active learning. We demonstrate RESAD using data from the NASA DASHlink project [3].

anomaly detection↗

Analyzing Cyber Security Threats on Cyber-Physical Systems Using Model-Based Systems Engineering

The spectre of cyber attacks on aerospace systems can no longer be ignored given that many of the components and vulnerabilities that have been successfully exploited by the adversary on other infrastructures are the same as those deployed and used within the aerospace environment. An important consideration with respect to the mission/safety critical infrastructure supporting space operations is that an appropriate defensive response to an attack invariably involves the need for high precision and accuracy, because an incorrect response can trigger unacceptable losses involving lives and/or significant financial damage. A highly precise defensive response, considering the typical complexity of aerospace environments, requires a detailed and well-founded understanding of the underlying system where the goal of the defensive response is to preserve critical mission objectives in the presence of adversarial activity. In this paper, a structured approach for modeling aerospace systems is described. The approach includes physical elements, network topology, software applications, system functions, and usage scenarios. We leverage Model-Based Systems Engineering methodology by utilizing the Object Management Group's Systems Modeling Language to represent the system being analyzed and also utilize model transformations to change relevant aspects of the model into specialized analyses. A novel visualization approach is utilized to visualize the entire model as a three-dimensional graph, allowing easier interaction with subject matter experts. The model provides a unifying structure for analyzing the impact of a particular attack or a particular type of attack. Two different example analysis types are demonstrated in this paper: a graph-based propagation analysis based on edge labels, and a graph-based propagation analysis based on node labels.

MBSE↗

Return of the Lemnaceae: duckweed as a model plant system in the genomics and postgenomics era

Abstract The aquatic Lemnaceae family, commonly called duckweed, comprises some of the smallest and fastest growing angiosperms known on Earth. Their tiny size, rapid growth by clonal propagation, and facile uptake of labeled compounds from the media were attractive features that made them a well-known model for plant biology from 1950 to 1990. Interest in duckweed has steadily regained momentum over the past decade, driven in part by the growing need to identify alternative plants from traditional agricultural crops that can help tackle urgent societal challenges, such as climate change and rapid population expansion. Propelled by rapid advances in genomic technologies, recent studies with duckweed again highlight the potential of these small plants to enable discoveries in diverse fields from ecology to chronobiology. Building on established community resources, duckweed is reemerging as a platform to study plant processes at the systems level and to translate knowledge gained for field deployment to address some of society’s pressing needs. This review details the anatomy, development, physiology, and molecular characteristics of the Lemnaceae to introduce them to the broader plant research community. We highlight recent research enabled by Lemnaceae to demonstrate how these plants can be used for quantitative studies of complex processes and for revealing potentially novel strategies in plant defense and genome maintenance.

Biochemistry & Molecular Biology↗

The excitation of plasma waves by a current source moving in a magnetized plasma - Two-dimensional propagation

The steady motion of a large conducting body in a magnetized plasma has been shown to create a spatial structure of charges and currents in the plasma which propagate without attenuation along the magnetic field direction (the label 'Alfven wings' has been given to this structure because of its resemblance to the wings of a biplane). It has been found that the Alfven wing structure breaks up into individual Fourier components which propagate away from the source at an angle slightly greater with respect to the magnetic field than the alignment of the original structure. This process depends on wave frequency. A given frequency is lost from the Alfven wing structure at a distance which varies as 1/omega-cubed. The amplitude of these traveling waves decreases at a rate which is inversely proportional to the square root of the distance from the source.

Rasmussen, Craig E.↗

Studies on propagation of microbes in the airborne state

An investigation was conducted to demonstrate whether airborne microbes could propagate. The procedure consisted of: (1) looking for dilution of a labelled base in DNA; (2) looking for labelling of DNA by mixing aerosols of the label and the cells; (3) examining changes in cell size; (4) testing the possibility of spore germination; and (5) seeking evidence of an increase in cell number. Results indicate that growth and propagation can occur under special conditions, principally at temperatures of approximately 30 C (87 F) and water activity equivalents of 0.95 to 0.98.

Dimmick, R. L.↗

Self-Supervised T-GCN for Detection of Disturbance and Propagation in Power Grid

Urban power systems increasingly rely on dense sensing to monitor grid reliability, yet disturbance labels are scarce and events are rare. We present a self-supervised spatio-temporal method that detects, localizes, and characterizes grid frequency disturbances across urban areas using only unlabeled data. Our approach trains a tiny Temporal Graph Convolutional Network (T-GCN) to forecast per-site frequency residuals (deviation from 60 Hz). The sensor graph is constructed directly from signals using pre-event Pearson correlation with a cross-correlation lag penalty without geocoding. At inference, node-level anomalies are the model's forecast errors; region-level alarms arise from connected components of high-score nodes. We estimate disturbance propagation by computing per-node arrival times (first persistent exceedance), then fit a planar or time-of-arrival model to obtain direction, speed, and an epicenter proxy. With only three real events collected at decisecond resolution across U.S. cities, we evaluate the T-GCN and report time-to-detect, footprint size, and propagation consistency. We further show that short-window embeddings from the T-GCN's hidden states enable few-shot event-vs-background recognition via a simple prototypical classifier. Despite minimal data and no labels, our system yields fast, spatially coherent detection and interpretable propagation maps, offering a practical, lightweight pathway to city-scale grid resilience analytics.

Niu, Haoran [ORNL] (ORCID:0000000155228297)↗

Count Every Trip: Finding the Uncertainty in Energy Estimates Made from Inferred Travel Modes

To properly inform transport policy and infrastructure changes, transportation related metrics need both measured values and uncertainties of those values. Travel monitoring smartphone apps can record people's travel behavior, but trip data quality is limited by sensor errors, user labeling rates and the accuracy of inference algorithms used for travel diary creation. We discuss the use of phone app recorded travel diary data to estimate energy consumption, and propose the use of propagation of variance to find error bars for such estimates. We define energy consumption for one trip as trip length times the energy intensity per distance unit of the travel mode used. We characterize trip length errors with relative error and inferred trip mode errors with confusion matrix columns. The resulting variances of each measurement are then propagated to the final calculated energy consumption. We tested our uncertainty methods on a dataset that used phone app data combined with prompted recall, consisting of 92,234 labeled trips for over 500,000 miles. Accounting for uncertainty using expected energy intensities and variance propagation gives a dataset-wide aggregate energy consumption percent error of about 9%, within one standard deviation from the truth. Future work could involve applying similar methods to other travel diary based metrics.

ADVANCED PROPULSION SYSTEMS,MATHEMATICS AND COMPUT↗

Count Every Trip: Finding the Uncertainty in Energy Estimates Made from Inferred Travel Modes

To properly inform transport policy and infrastructure changes, transportation related metrics need both measured values and uncertainties of those values. Travel monitoring smartphone apps can record people's travel behavior, but trip data quality is limited by sensor errors, user labeling rates and the accuracy of inference algorithms used for travel diary creation. We discuss the use of phone app recorded travel diary data to estimate energy consumption, and propose the use of propagation of variance to find error bars for such estimates. We define energy consumption for one trip as trip length times the energy intensity per distance unit of the travel mode used. We characterize trip length errors with relative error and inferred trip mode errors with confusion matrix columns. The resulting variances of each measurement are then propagated to the final calculated energy consumption. We tested our uncertainty methods on a dataset that used phone app data combined with prompted recall, consisting of 92,234 labeled trips for over 500,000 miles. Accounting for uncertainty using expected energy intensities and variance propagation gives a dataset-wide aggregate energy consumption percent error of about 9%, within one standard deviation from the truth. Future work could involve applying similar methods to other travel diary based metrics.

ADVANCED PROPULSION SYSTEMS↗

Estimating Travel Energy Consumption Uncertainty Based on Inferred Travel Mode and Sensed Travel Length

To properly inform transport policy and infrastructure changes, transportation related metrics need both measured values and uncertainties of those values. Travel monitoring smartphone apps can record people's travel behavior, but trip data quality is limited by sensor errors, user labeling rates and the accuracy of inference algorithms used for travel diary creation. We discuss the use of phone app recorded travel diary data to estimate energy consumption, and propose the use of propagation of variance to find error bars for such estimates. We define energy consumption for one trip as trip length times the energy intensity per distance unit of the travel mode used. We characterize trip length errors with relative error and inferred trip mode errors with confusion matrix columns. The resulting variances of each measurement are then propagated to the final calculated energy consumption. We tested our uncertainty methods on a dataset that used phone app data combined with prompted recall, consisting of 92,234 labeled trips for over 500,000 miles. Accounting for uncertainty using expected energy intensities and variance propagation gives a dataset-wide aggregate energy consumption percent error of about 8%, within one standard deviation from the truth. Future work could involve applying similar methods to other travel diary based metrics.

ADVANCED PROPULSION SYSTEMS,ENERGY PLANNING, POLIC↗

Mass Detection for Heavy-Duty Vehicles using Gaussian Belief Propagation

Predicting vehicle mass is critical to accurately estimate energy use and emissions of commercial trucks. However, data from vehicle telematics is often not at sufficient temporal resolution or accuracy for use in model-based detection methods. In this work, a new statistical mass prediction technique is described for heavy-duty vehicles that incorporates the use Gaussian Belief Propagation (GBP) for probabilistic inference. Similar to Bayesian inference models, the GBP model typically requires less labeled training data than other contemporary machine learning techniques. First, a factor graph is constructed, and a set of Gaussian belief nodes with associated means and variances are fitted to the training data. To better handle noisy input data, the GBP mass prediction model utilizes a k-nearest factors (kNF) algorithm for probabilistic inference on unseen testing data. The proposed method is compared with a classical weighted k-nearest neighbors (kNN) regressor. This statistical kNF-GBP model works even with low-quantity, low-quality initial training data, while being capable of realtime mass estimation. Unlike the kNN regressor, the GBP model produces a measure of uncertainty with its predictions. The proposed method is validated using curve-sampled driving data collected from multiple cloud-connected Class 8 regional haul diesel trucks. Both the kNN regressor and the kNF-GBP mass prediction model were able to predict payload mass with coefficients of determination above 0.97 with minimal data preprocessing.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Computerized aerodynamic design of a transonically 'quiet' blade

The high noise levels produced by helicopters are major sources of concern. There are many sources of the noise, but during high-speed forward flight, impulsive noise dominates the noise spectrum. The cause of the high-speed impulsive noise is the propagation into the far field of shock waves that form on the advancing blade. This mechanism has been labeled 'delocalization'. It has been shown, however, that by judicious design of the blade-tip planform, delocalization can be prevented. The objective of the present study is to illustrate how blade-tip configurations (both planform and airfoil shape) can be systematically varied to identify shapes that avoid delocalization and simultaneously improve aerodynamic performance. This has been done using the latest version of the ROT22 transonic, full-potential, quasi-steady, rotor flow-field code. A hypothetical modern rotor blade was postulated, and tip modifications consisting of taper, sweep, and airfoil section alterations were investigated. Planform modifications were found to be most effective in eliminating delocalization.

Tauber, M. E.↗

Iterative self-organizing SCEne-LEvel sampling (ISOSCELES) for large-scale building extraction

Convolutional neural networks (CNN) provide state-of-the-art performance in many computer vision tasks, including those related to remote-sensing image analysis. Successfully training a CNN to generalize well to unseen data, however, requires training on samples that represent the full distribution of variation of both the target classes and their surrounding contexts. With remote sensing data, acquiring a sufficiently representative training set is a challenge due to both the inherent multi-modal variability of satellite or aerial imagery and the general high cost of labeling data. To address this challenge, we have developed ISOSCELES, an Iterative Self-Organizing SCEne LEvel Sampling method for hierarchical sampling of large image sets. Using affinity propagation, ISOSCELES automates the selection of highly representative training images. Compared to random sampling or using available reference data, the distribution of the training is principally data driven, reducing the chance of oversampling uninformative areas or undersampling informative ones. In comparison to manual sample selection by an analyst, ISOSCELES exploits descriptive features, spectral and/or textural, and eliminates human bias in sample selection. Using a hierarchical sampling approach, ISOSCELES can obtain a training set that reflects both between-scene variability, such as in viewing angle and time of day, and within-scene variability at the level of individual training samples. We verify the method by demonstrating its superiority to stratified random sampling in the challenging task of adapting a pre-trained model to a new image and spatial domain for country-scale building extraction. Using a pair of hand-labeled training sets comprising 1,987 sample image chips, a total of 496,000,000 individually labeled pixels, we show, across three distinct model architectures, an increase in accuracy, as measured by F1-score, of 2.2–4.2%.

42 ENGINEERING↗

Error Modeling of Multi-baseline Optical Truss: Application to SIM Metrology Truss Field Dependent Error - Part II

The current design of the Space Interferometry Mission (SIM) employs a 19 laser-metrology-beam system (also called L19 external metrology truss) to monitor changes of distances between the fiducials of the flight system's multiple baselines. The function of the external metrology truss is to aid in the determination of the time-variations of the interferometer baseline. The largest contributor to truss error occurs in SIM wide-angle observations when the articulation of the siderostat mirrors (in order to gather starlight from different sky coordinates) brings to light systematic errors due to offsets at levels of instrument components (which include comer cube retro-reflectors, etc.). This error is labeled external metrology wide-angle field-dependent error. Physics-based model of field-dependent error at single metrology gauge level is developed and linearly propagated to errors in interferometer delay. In this manner delay error sensitivity to various error parameters or their combination can be studied using eigenvalue/eigenvector analysis. Also validation of physics-based field-dependent model on SIM testbed lends support to the present approach. As a first example, dihedral error model is developed for the comer cubes (CC) attached to the siderostat mirrors. Then the delay errors due to this effect can be characterized using the eigenvectors of composite CC dihedral error. The essence of the linear error model is contained in an error-mapping matrix. A corresponding Zernike component matrix approach is developed in parallel, first for convenience of describing the RMS of errors across the field-of-regard (FOR), and second for convenience of combining with additional models. Average and worst case residual errors are computed when various orders of field-dependent terms are removed from the delay error. Results of the residual errors are important in arriving at external metrology system component requirements. Double CCs with ideally co-incident vertices reside with the siderostat. The non-common vertex error (NCVE) is treated as a second example. Finally combination of models, and various other errors are discussed.

comer cube retro-reflector↗