Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “network data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Regressing Nuclear Reactor Power Level Using Low-Cost Sensor Network Data

Multisensor networks deployed at nuclear facilities can be leveraged to collect data used as inputs to machine learning models predicting nuclear safeguard relevant information. This work demonstrates an application of this idea by regressing nuclear reactor power levels, a key indicator for nuclear safeguard verification, at the McClellan Nuclear Research Center using data collected by five Merlyn multisensor platforms with LASSO and LSTM models. This work also demonstrates the use of Leave One Node Out to measure the importance of each multisensor for this regression problem providing insight into model explainability and allowing inferential hypotheses about the nuclear facility to be made. This work can be used as a starting point for future development of methods for regression on reactor power levels at nuclear facilities using multisensor network data.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Differentially Private Synthesis and Sharing of Network Data Via Bayesian Exponential Random Graph Models

Abstract Network data often contain sensitive relational information. One approach to protecting sensitive information while offering flexibility for network analysis is to share synthesized networks based on the information in originally observed networks. We employ differential privacy (DP) and exponential random graph models (ERGMs) and propose the DP-ERGM method to synthesize network data. We apply DP-ERGM to two real-world networks. We then compare the utility of synthesized networks generated by DP-ERGM, the DyadWise Randomized Response (DWRR) approach, and the Synthesis through Conditional distribution of Edge given nodal Attribute (SCEA) approach. In general, the results suggest that DP-ERGM preserves the original information significantly better than two other approaches in network structural statistics and inference for ERGMs and latent space models. Furthermore, DP-ERGM satisfies node DP through modeling the global network structure with ERGM, a stronger notion of privacy than the edge DP under which DWRR and SCEA operate.

graph synthesis↗

CAN-D: A Modular Four-Step Pipeline for Comprehensively Decoding Controller Area Network Data

Controller area networks (CANs) are a broadcast protocol for real-time communication of critical vehicle subsystems. Original equipment manufacturers of passenger vehicles hold secret their mappings of CAN data to vehicle signals, and these definitions vary according to make, model, and year. Without these mappings, the wealth of real-time vehicle information hidden in the CAN packets is uninterpretable, severely impeding vehicle-related research, including CAN cybersecurity and privacy studies, aftermarket tuning, efficiency and performance monitoring, and fault diagnosis to name a few. Guided by the four-part CAN signal definition, we present CAN-D (CAN-Decoder), a modular, four-step pipeline for identifying each signal's boundaries (start bit and length), endianness (byte ordering), signedness (bit-to-integer encoding), and by leveraging diagnostic standards, augmenting a subset of the extracted signals with meaningful, physical interpretation. En route to CAN-D, we provide a comprehensive review of the CAN signal reverse engineering research. All previous methods ignore endianness and signedness, rendering them incapable of decoding many standard CAN signal definitions. Incorporating endianness grows the search space from 128 to 4.72E21 signal tokenizations and introduces a web of changing dependencies. In response, we formulate, formally analyze, and provide an efficient solution to an optimization problem, allowing identification of the optimal set of signal boundaries and byte orderings. In addition, we provide two novel, state-of-the-art signal boundary classifiers—both of which are superior to previous approaches in precision and recall in three different test scenarios—and the first signedness classification algorithm, which exhibits a $>$ 97% F-score. Altogether, CAN-D is the only solution with the potential to extract any CAN signal that is also the state of the art. In evaluation on 10 vehicles of different makes, CAN-D's average $\ell ^1$ error is five times better (81% less) than all previous methods and exhibits lower average error, even when considering only signals that meet prior methods’ assumptions. Finally, CAN-D is implemented in lightweight hardware, allowing for an on-board diagnostic (OBD-II) plugin for real-time in-vehicle CAN decoding.

42 ENGINEERING↗

Mitigate: An Adaptive Network Data Anonymization Tool Using Condensation-Based Differential Privacy

Modern network devices collect a large amount of data that can be analyzed to identify bottlenecks, anomalies, cyber-attacks, etc. Therefore, there is often a need to analyze such collections of network data quite often by an external expert or by the research community. However, these collections of data contain sensitive, proprietary information. In order for the network data to be shared, it must first be anonymized. The overall objective of this project is to develop an innovative privacy management tool to anonymize network data and achieve sufficient privacy, acceptable data utility, and efficient data analysis at the same time. No existing anonymization methods can achieve all of these at the same time. The core of this technology is a differential private clustering algorithm that provides strong privacy protection, preserves data properties important for subsequent analysis, and allows the party receiving the anonymized data to conduct analysis directly on anonymized data without the need of decryption or any extra processing. The research carried out was to design, implement and verify a solution to this problem by completing the following tasks: 1) developing the core technology; 2) developing a context based method that automatically recommends fields that must be anonymized; 3) conducted experiments showing superior results using our approach compared to existing tools, and 4) developed an intuitive but basic user interface. The research that was conducted generated novel algorithmic techniques that utilize state-of-the-art methods such as condensation, differential privacy preservation, clustering, automated tuning based on contextual awareness, and recommendation techniques to specify columns to users for anonymization leading to optimal privacy that allows research analysis on the dataset. Experiments were conducted to evaluate the efficacy of these novel algorithmic techniques by performing analysis on original non-anonymized datasets, then conducting analysis on the same yet anonymized datasets and comparing the results of the analyses. Overall, the anonymized analysis results were within 1% of the original results, verifying that the generated technology not only guarantees a high level of privacy but also enables research analysis as if it were conducted on the original dataset. Potential applications of this technology include anonymization of any type of structured network datasets that contain sensitive identifiers, such as IP addresses, that can be used in multiple applications. For example, to create an AI or machine learning model for cyber security, e.g., to detect attacks, or for performance analysis, e.g., identify bottlenecks or predict performance. In addition, a market analysis that was conducted for potential applications of this technology identified a broader range of applications of our anonymization technology beyond the network sector that includes healthcare, banking, insurance, securities, finance (FISB), data brokering, cloud services, ad sales, and government.

97 MATHEMATICS AND COMPUTING↗

NDNSEG (Named Data Network Scalable Environment Generator) [SWR-21-91]

NDNSEG (Named Data Network Scalable Environment Generator) automates the deployment of NDN environments in NREL's Cyber-Energy Emulation Platform (CEEP). Specifying the number of clients, servers, gateways, and pv_inverters and running NDNSEG will automate the generation of required files to deploy the number of specified virtual machines, the underlying networking infrastructure, and configuration for GitLab CI/CD.

Peterson, Jordan↗

Enriching OpenStreetMap network data for transportation applications: Insights into the impact of urban congestion on accessibility

OpenStreetMap (OSM) data is a valuable open-source resource for various transportation, traffic, and planning applications. However, OSM network data lack operating traffic speed information, which is critical for transport planning and operations. Addressing this shortcoming, this study leverages commercial vendor data (to serve as ground truth) with exogenous, open-source variables characterizing local transport infrastructure, land use, and demographic information to predict average congested traffic speeds on OSM networks. Three machine-learning models were tested and estimated for OSM links with and without speed limit information in the Denver metropolitan region. Among these, XGBoost performed best, with mean absolute errors of 3.27 and 3.62 mph for links with and without speed limits, respectively. The developed models accurately predicted traffic speeds for different hours and days of the week compared to ground truth data. Using these predicted speeds, drive accessibility scores were computed for the Denver region for different time periods using the Mobility Energy Productivity (MEP) metric to understand the impact of congestion on energy-efficient accessibility. Results show that congestion-adjusted drive accessibility can be significantly lower compared to accessibility calculated using free flow speeds. Specifically, weekday evening hours saw a 42 % drop in accessibility due to reduced speeds, particularly around downtown Denver. Across the Denver metro region, approximately half as many opportunities and jobs are accessible in under 20 min by car during the evening peak period relative to free flow conditions. These findings underscore the importance of using congestion-adjusted operating speeds rather than speed limits in accessibility calculations, as reliance on speed limits can substantially overestimate energy-efficient drive accessibility in large, car-centric cities susceptible to significant congestion. In conclusion, the methodology presented here could further enrich OSM network data, making them useful for an even broader range of transportation applications.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

NCAR-RAL Surface Hydrometeorological Observation Network Data for LASSO-CACTI Overview Paper

This data set contains the 15 minute resolution surface meteorology and soils data from the 15 NCAR/RAL weather stations that were operated around central Argentina during the RELAMPAGO (Remote sensing of Electrification, Lightning, And Meso-scale/micro-scale Processes with Adaptive Ground Observations) Extended Observing Period (EOP). Data providence, citation, and acknowledgement This ARM data set is a copy of v1.0 of the NCAR data set obtained in June 2024 from https://doi.org/10.26023/KW8Z-F2WX-H0Y. The citation for the original data source is: Gochis, D., et al. 2019. NCAR-RAL Surface Hydrometeorological Observation Network Data. Version 1.0. UCAR/NCAR - Earth Observing Laboratory. https://doi.org/10.26023/KW8Z-F2WX-H0Y Accessed June 2024. In addition to the citation reference and any other acknowledgements, please acknowledge NCAR/EOL in your publications with text such as: “Data provided by NCAR/EOL under the sponsorship of the National Science Foundation. https://data.eol.ucar.edu/”

air temperature↗

Improved lithospheric attenuation structure of the Arabian Peninsula through the use of national network data

We characterize the attenuation structure of the Arabian Peninsula through the measurement of regional phase amplitudes. High-resolution is achieved by combining stations from global networks with national network data through the cooperative effort of several countries in the region, including Saudi Arabia, Oman, Iraq, and Kuwait. The result is an improved attenuation model of the crust and upper mantle for a broad frequency band that extends from 0.5 to 10 Hz. The observed attenuation is in accordance with various elements of earth structure, including plate boundary type, style of tectonism, thermo-tectonic age, and temperature. Emerging features from the model include details in the structure along the Red Sea, and improved imaging of the southern Arabian Peninsula extending north from the Gulf of Aden. Finally, the resulting attenuation model can be employed for better magnitude estimates, in isolating tectonic and structural features, and in characterizing strong ground motion in the Arabian Peninsula.

58 GEOSCIENCES↗

A product data network to enable faster, easier, and better planning of building envelopes

The building envelopes contributes significantly to the energy-efficiency of the building. Building performance simulation has made it possible to compare façade technologies regarding energy demand, daylighting, thermal and visual comfort in detail. Planners, such as architects and engineers, need experience to find product data with the right quality and level of detail, and to process the data to fit the calculation and the application. In the available time, planners can compare only a limited number of products, which means that better solutions could go unnoticed. This paper presents a new concept for making product data easily accessible for building façade planning. The concept consists of a network of databases for the efficient exchange and use of optical and calorimetric data of glazing units, shading devices, and combinations of both. The paper presents the research questions, an analysis of the current challenges, six design goals for the product data network and its implementation together with a discussion. Many product data sources can be connected to many planning software applications via the specified application programming interface. When planning software connects to the product data network, the planning of building envelopes can be much faster because planners do not need to spend so much time to search and process product data manually. The planning of building envelopes can also become much easier, especially for planners with limited experience. They do not need to understand all the details about which data fits which calculation if the software company implements this. The planning of building envelopes can become much more reliable when software companies validate their use of the product data network, because the current manual process is prone to errors. The planning of building envelopes can also improve because more products can be compared in the available time, allowing better solutions to be found.

Maurer, Christoph↗

Data for The Stem Cell-Type Transcriptome of Bioenergy Sorghum Reveals the Spatial Regulation of Secondary Cell Wall Networks

Bioenergy sorghum is a low-input, drought-resilient, deep-rooting annual crop that has high biomass yield potential enabling the sustainable production of biofuels, biopower, and bioproducts. Bioenergy sorghum’s 4-5 m stems account for ~80% of the harvested biomass. Stems accumulate high levels of sucrose that could be used to synthesize bioethanol and useful biopolymers if information about stem cell-type gene expression and regulation was available to enable engineering. To obtain this information, Laser Capture Microdissection (LCM) was used to isolate and collect transcriptome profiles from five major cell types that are present in stems of the sweet sorghum Wray. Transcriptome analysis identified genes with cell-type specific and cell-preferred expression patterns that reflect the distinct metabolic, transport, and regulatory functions of each cell type. Analysis of cell-type specific gene regulatory networks (GRNs) revealed that unique TF families contribute to distinct regulatory landscapes, where regulation is organized through various modes and identifiable network motifs. Cell-specific transcriptome data was combined with a stem developmental transcriptome dataset to identify the GRN that differentially activates the secondary cell wall (SCW) formation in stem xylem sclerenchyma and epidermal cells. The cell-type transcriptomic dataset provides a valuable source of information about the function of sorghum stem cell types and GRNs that will enable the engineering of bioenergy sorghum stems.

Software↗

Utilizing the Dynamic Networks Data Processing and Analysis Experiment (DNE18) to Establish Methodologies for the Comparison of Automatic Infrasonic Signal Detectors

The Dynamic Networks Experiment 2018 (DNE18) was a collaborative effort between Los Alamos National Laboratory (LANL), Sandia National Laboratories (SNL), Lawrence Livermore National Laboratory (LLNL) and Pacific Northwest National Laboratory (PNNL) designed to evaluate methodologies for multi-modal data ingestion and processing. One component of this virtual experiment was a quantitative assessment of current capabilities for infrasound data processing, beginning with the establishment of a baseline for infrasound signal detection. To produce such baselines, SNL and LANL exploited a common dataset of infrasound data recorded across a regional network in Utah from December 2010 through February 2011. We utilize two automated signal detectors, the Adaptive F-Detector (AFD) and the Multivariate Adaptive Learning Detector (MALD) to produce automated signal detection catalogs and an analyst-produced catalog. Comparisons indicate that automatic detectors may be able to identify small amplitude, low SNR events that cannot be identified by analyst review. We document detector performance in terms of precision and recall, demonstrating that the AFD is more precise, but the MALD has higher recall. We use a synthetic dataset of signals embedded in pink noise in order to highlight shortcomings in assessing detection algorithms for low signal to noise ratio signals which are commonly of interest to the nuclear monitoring community. For comparisons utilizing the synthetic dataset, the AFD has higher recall while precision is equal for both detectors. These results indicate that both detectors perform well across a variety of background noise environments; however, both detectors fail to identify repetitive, short duration signals arriving from similar backazimuths. These failures represent specific scenarios that could be targeted for further detector development.

97 MATHEMATICS AND COMPUTING↗

A guide to the BRAIN Initiative Cell Census Network data ecosystem

Characterizing cellular diversity at different levels of biological organization and across data modalities is a prerequisite to understanding the function of cell types in the brain. Classification of neurons is also essential to manipulate cell types in controlled ways and to understand their variation and vulnerability in brain disorders. The BRAIN Initiative Cell Census Network (BICCN) is an integrated network of data-generating centers, data archives, and data standards developers, with the goal of systematic multimodal brain cell type profiling and characterization. Emphasis of the BICCN is on the whole mouse brain with demonstration of prototype feasibility for human and nonhuman primate (NHP) brains. Here, we provide a guide to the cellular and spatial approaches employed by the BICCN, and to accessing and using these data and extensive resources, including the BRAIN Cell Data Center (BCDC), which serves to manage and integrate data across the ecosystem. We illustrate the power of the BICCN data ecosystem through vignettes highlighting several BICCN analysis and visualization tools. Finally, we present emerging standards that have been developed or adopted toward Findable, Accessible, Interoperable, and Reusable (FAIR) neuroscience. The combined BICCN ecosystem provides a comprehensive resource for the exploration and analysis of cell types in the brain.

59 BASIC BIOLOGICAL SCIENCES↗

Blue Keanu: A Scientific Visualization Tool For Network Data

This software allows the user to visualize complex PCAP-ng files captured from network capture software such as Wireshark. The visualization runs in a GUI window that can be zoomed or moved to areas of interest in a waterfall type display. The user then can see an area of interest that looks different than the typical traffic visually, such as a human interaction or non-repetitive area of data. The program will tell the user the packet number and byte offset of interest for fast analysis of discrete atomic or non-random events. This is particularly useful for visualization of unknown binary format data, such as in PLC or SCADA protocols that may have human or other non-repetitive activity for further analysis, reverse engineering, or fast forensic analysis.

Durller, MichaelGeorge↗

Utah FORGE Well 16A(78)-32 Simplified Discrete Fracture Network Data

The FORGE team is making these fracture models available to researchers wanting a set of natural fractures in the FORGE reservoir for use in their own modeling work. They have been used to predict stimulation distances during hydraulic stimulation at the open toe section of well 16A(78)-32. This is a simplified DFN (discrete fracture network) dataset, that was generated using FracMan, for Utah FORGE well 16A(78)-32. A short, well-illustrated, report describing the data is also included in the provided archive file.

15 GEOTHERMAL ENERGY↗

Probing the D-region ionosphere globally with Earth Networks Total Lightning Network data

An existing technique to use broadband lightning waveforms to probe the D-region ionosphere (60–90 km altitude) is shown to be extendable to a global scale using the Earth Networks Total Lightning Network (ENTLN). This paper demonstrates the technique in detail on a region of the Southeastern United States. This demonstration shows that diurnal D-region height variation and smaller time-scale variations on the order of tens of minutes to hours are evident in the measurement. The technique is then extended to three additional global regions on this same day: Northeastern U.S., India, and Japan. The diurnal behavior between these different regions is compared to a D-region model from the International Reference Ionosphere.

D-region ionosphere↗