Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Complex Network Data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Fast Gaussian Process Estimation for Large-Scale In Situ Inference using Convolutional Neural Networks

Exascale computing will bring with it significant I/O limitations. One foreseeable consequence of such restrictions is that the user can save only a small fraction of complex simulation data to disk for subsequent analysis. An alternative is to fit statistical models to data in situ, that is, inside the simulation as it runs. This option requires extremely fast statistical estimation to avoid slowing down the simulation. Gaussian processes (GPs) have state-of-the-art predictive performance for modeling spatial data. However, standard estimation methods for GPs scale quite poorly to large data sets as parameter estimation requires inverting a covariance matrix to the size of the data set. In the presented work, we use a convolutional neural network (CNN) to predict the GP parameters for a spatial data set, from a simulation or otherwise, rather than optimize the parameters directly. Here, our presented case study models spatial data from E3SM, the Department of Energy’s Exascale climate model. The CNN is trained on synthetic data simulated from GP models with known parameters and then applied to data from the climate simulation. In the presented examples, the neural network scheme produces parameter estimates that compare well with standard methods such as maximum likelihood estimation in predictive performance but is obtained four orders of magnitude faster.

big data↗

Phase Identification in Synchrotron X-ray Diffraction Patterns of Ti–6Al–4V Using Computer Vision and Deep Learning

X-ray diffraction patterns contain information about the atomistic structure and microstructure (defect population) of materials, extracting detailed information from diffraction patterns is complex, demanding and relies on prior knowledge. Here, we hypothesize that deep-learning techniques can help to perform an effective and accurate analysis with high throughput rates. To demonstrate this concept, we applied a novel deep learning framework to determine the evolution of the β-phase volume fraction in a Ti–6Al–4V alloy during heat-treatment from video sequences of 2D diffraction patterns recorded in transmission and with highly monochromatic radiation in a synchrotron beamline. In particular, we studied the impact of network design on prediction reliability and computational performance. Networks of different architectures were trained using 3008 experimental 2D patterns. A well-tuned model was found to reproduce the phase fractions of another experimental data set, consisting of 1100 diffraction patterns, with a mean-square error as small as 2.6 x 10 -4 . The average prediction error of β-phase volume fraction was within 1.6 x 10 -2 (in each diffraction pattern) of the values obtained by conventional methods. Our work demonstrates that convolutional neural networks can evaluate high energy X-ray diffraction patterns with a remarkable level of reliability. Furthermore, it demonstrates the significance of network design on the reliability of predictions and computational performance. The most complex models do not necessarily result in highest accuracy and may even fail to learn from the data.

36 MATERIALS SCIENCE↗

Million-scale data integrated deep neural network for phonon properties of heuslers spanning the periodic table

Existing machine learning potentials for predicting phonon properties of crystals are typically limited on a material-to-material basis, primarily due to the exponential scaling of model complexity with the number of atomic species. We address this bottleneck with the developed Elemental Spatial Density Neural Network Force Field, namely Elemental-SDNNFF. The effectiveness and precision of our Elemental-SDNNFF approach are demonstrated on 11,866 full, half, and quaternary Heusler structures spanning 55 elements in the periodic table by prediction of complete phonon properties. Self-improvement schemes including active learning and data augmentation techniques provide an abundant 9.4 million atomic data for training. Deep insight into predicted ultralow lattice thermal conductivity (<1 Wm –1 K –1 ) of 774 Heusler structures is gained by p–d orbital hybridization analysis. Additionally, a class of two-band charge-2 Weyl points, referred to as “double Weyl points”, are found in 68% and 87% of 1662 half and 1550 quaternary Heuslers, respectively.

36 MATERIALS SCIENCE↗

Multi-Level Structural Damage Characterization Using Sparse Acoustic Sensor Networks and Knowledge Transferred Deep Learning

Standard structural health monitoring techniques face well-known difficulties for comprehensive defect diagnosis in real-world structures that have structural, material, or geometric complexity. This motivates the exploration of machine-learning-based structural health monitoring methods in complex structures. However, creating sufficient training data sets with various defects is an ongoing challenge for data-driven machine (deep) learning algorithms. The ability to transfer the knowledge of a trained neural network from one component to another or to other sections of the same component would drastically reduce the required training data set. Also, it would facilitate computationally inexpensive machine learning based inspection systems. In this work, a machine-learning-based multi-level damage characterization is demonstrated with the ability to transfer trained knowledge within the sparse sensor network. A novel network spatial assistance and an adaptive convolution technique are proposed for efficient knowledge transfer within the deep learning algorithm. Proposed structural health monitoring method is experimentally evaluated on an aluminum plate with artificially induced defects. It was observed that the method improves the performance of knowledge transferred damage characterization by 50% during localization and 24% during severity assessment. Further, experiments using time windows with and without multiple edge reflections are studied. Results reveal that multiply scattered waves contain rich and deterministic defect signatures that can be mined using deep learning neural networks, improving the accuracy of both identification and quantification. In the case of a fixed sensor network, using multiply scattered waves shows 100% prediction accuracy at all levels of damage characterization.

36 MATERIALS SCIENCE↗

L-VISP: LSTM Visualization for Interpretable Symptom Prediction in Patient Cohorts

Symptom modelling in head and neck cancer is challenged by the complexity of heterogeneous patient data, leading to an interest in deep learning approaches. Although Long Short-Term Memory Networks (LSTMs) have shown great results in patient risk prediction, their low interpretability requires data modellers to collaborate with clinical experts to validate the results. We present L-VISP, a human–machine solution that uses visual analytics for LSTM modelling in clinical research. L-VISP uses custom visual encodings to make multiple LSTM variants interpretable, supporting a full range of analysis, from understanding model operations and evaluating performance to interpreting results in a clinical context. We evaluate L-VISP with data modellers and a clinical oncologist and present the takeaways from this multidisciplinary collaboration.

LSTM modeling↗

Day-Ahead Forecasting with Federated LSTM to Plan Energy Sharing in a Community Microgrid

Energy balancing in microgrids is a key enabler of resilience. Community microgrids located close to each other have the added benefit of networking and sharing surplus energy, if available. Such complex decision-making runs on optimization that requires reliable short-term (up to very-short-term) forecasts of energy generation and consumption for scheduling or trading. Each microgrid may also opt to not expose their sensitive data such as consumption patterns of individual businesses or residences. This paper investigates a federated approach to dayahead forecasting that trains naive long short-term memory (LSTM) at each business in a microgrid and aggregates weights at the microgrid controller using proximal regularization. This approach ensures that the controller has access only to energy surplus/deficit and not the actual generation or consumption values, avoiding unwanted exposure of sensitive data. A community microgrid in Adjuntas, Puerto Rico with 3 businesses is selected as a case study with a laboratory-scale computing setup. A central LSTM forecaster, where sensitive data from businesses are aggregated at the controller, is implemented as a baseline for qualifying the results. This work serves as a proof-of-concept for scaling the approach to networked and nested microgrids with more complex control options.

Sundararajan, Aditya [ORNL] (ORCID:000000033577854↗

Data-scarce surrogate modeling of shock-induced pore collapse process

Understanding the mechanisms of shock-induced pore collapse is of great interest in various disciplines in sciences and engineering, including materials science, biological sciences, and geophysics. However, numerical modeling of the complex pore collapse processes can be costly. To this end, a strong need exists to develop surrogate models for generating economic predictions of pore collapse processes. Here, in this work, we study the use of a data-driven reduced-order model, namely dynamic mode decomposition, and a deep generative model, namely conditional generative adversarial networks, to resemble the numerical simulations of the pore collapse process at representative training shock pressures. Since the simulations are expensive, the training data are scarce, which makes training an accurate surrogate model challenging. To overcome the difficulties posed by the complex physics phenomena, we make several crucial treatments to the plain original form of the methods to increase the capability of approximating and predicting the dynamics. In particular, physics information is used as indicators or conditional inputs to guide the prediction. In realizing these methods, the training of each dynamic mode composition model takes only around 30 s on CPU. In contrast, training a generative adversarial network model takes 8 h on GPU. Moreover, using dynamic mode decomposition, the final-time relative error is around 0.3% in the reproductive cases. We also demonstrate the predictive power of the methods at unseen testing shock pressures, where the error ranges from 1.3 to 5% in the interpolatory cases and 8 to 9% in extrapolatory cases.

97 MATHEMATICS AND COMPUTING↗

Identifying Critical Infrastructure in Imagery Data Using Explainable Convolutional Neural Networks

To date, no method utilizing satellite imagery exists for detailing the locations and functions of critical infrastructure across the United States, making response to natural disasters and other events challenging due to complex infrastructural interdependencies. This paper presents a repeatable, transferable, and explainable method for critical infrastructure analysis and implementation of a robust model for critical infrastructure detection in satellite imagery. This model consists of a DenseNet-161 convolutional neural network, pretrained with the ImageNet database. The model was provided additional training with a custom dataset, containing nine infrastructure classes. The resultant analysis achieved an overall accuracy of 90%, with the highest accuracy for airports (97%), hydroelectric dams (96%), solar farms (94%), substations (91%), potable water tanks (93%), and hospitals (93%). Critical infrastructure types with relatively low accuracy are likely influenced by data commonality between similar infrastructure components for petroleum terminals (86%), water treatment plants (78%), and natural gas generation (78%). Local interpretable model-agnostic explanations (LIME) was integrated into the overall modeling pipeline to establish trust for users in critical infrastructure applications. The results demonstrate the effectiveness of a convolutional neural network approach for critical infrastructure identification, with higher than 90% accuracy in identifying six of the critical infrastructure facility types.

97 MATHEMATICS AND COMPUTING↗

A simulation framework for evaluating electronic order workflows in integrated health records

Electronic health record (EHR) systems are critical to modern healthcare delivery, yet the dynamic workflows that govern electronic order processing remain underexplored. Inefficiencies in these digital pathways can cause delays in care, repetitive workloads, and even patient harm. This study presents a discrete-event simulation framework used to reconstruct and evaluate EHR-based order workflows in a large integrated healthcare system. Using real-world data extracted from the Veterans Health Administration’s Corporate Data Warehouse, the authors mapped order events to standardized state transitions and modeled their progression across different facilities of varying complexity levels. After being calibrated with empirical distributions of transition times and validated against observed time-in-system metrics, the simulation demonstrates close alignment with historical performance. Scenario analyses reveal that resource capacity constraints significantly amplify the impact of electronic order surges, which are reflected in the disproportionate growth in backlogs and processing delays. Adjustments in transition probabilities further increased recirculation and extended workflow paths. Network-based analysis identified Reserved, InProgress, and Completed as structurally critical states that function as hubs within the process network but the transitions in-between also act as major bottlenecks. These results showcased the effectiveness of simulation-based approaches in monitoring EHR order processing performance and evaluating consequences of workflow changes on healthcare network resources planning. The proposed simulation framework provides a scalable data-driven tool to support operational decision-making and improve the efficiency of electronic order management in complex healthcare environments.

Engineering↗

Trustworthiness modeling and evaluation for a nearly autonomous management and control system

The Nearly Autonomous Management and Control (NAMAC) system supports the advanced reactor operation by recommending control actions to operators based on real-time measurements and digital twins (DTs) learning from the knowledge base. To enable the safe and reliable use of autonomous technologies, NAMAC and its recommendations should be trustworthy to operators and regulators at both the design and operation stages. This study proposes a NAMAC trustworthiness modeling and evaluation framework supported by trustworthiness ontologies and evidence-based approaches. The development-time and run-time ontologies are separately constructed and then converted to Bayesian networks to quantitatively evaluate the NAMAC trustworthiness. This evaluation is demonstrated by collecting and characterizing evidence from NAMAC practices, such as the development and assessment of the NAMAC system, data coverage assessment, and the training and optimizations of neural-network-based DTs. Our proposed approach can aggregate various trustworthiness attributes of complex artificial-intelligence-supported systems for safety-critical applications. It also considers the interaction between different DTs and extends beyond the trustworthiness evaluation of a single DT. In conclusion, the evidence-based method enhances the transparency of the trustworthiness modeling and evaluation processes and helps identify uncertainties and subjectivity involved in the processes.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Methods for R&D Portfolio Analysis and Evaluation (Workshop Report)

The Workshop on Methods for R&D Portfolio Analysis and Evaluation convened on 17–18 July 2019 at the National Renewable Energy Laboratory in Golden, Colorado, and examined strengths and weaknesses of the various methodologies applicable to R&D portfolio modeling, analysis, and decision support, given pragmatic constraints such as data availability, uncertainties in estimating the impact of R&D spending, and practical operational overheads. Participants employed their deep expertise in approaches such as stochastic optimization, real options, Monte-Carlo analysis, Bayesian networks, decision theory, complex systems analysis, deep uncertainty, and technology-evolution modeling to critique the initial example models developed by the project’s core team and to conduct thought experiments grounded in real-life technology models, progress data, expert elicitation, and portfolio information. This engagement of participants’ methodological expertise with the practical requirements of real-life portfolio decision support yielded ideas for improved approaches, alternative methodological hypotheses, and hybridization of methodologies that are well-grounded theoretically, computationally sound, and realistically executable given data availability and other practical constraints.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Methods for R&D Portfolio Analysis and Evaluation (Workshop Report)

The Workshop on Methods for R&D Portfolio Analysis and Evaluation convened on 17-18 July 2019 at the National Renewable Energy Laboratory in Golden, Colorado, and examined strengths and weaknesses of the various methodologies applicable to R&D portfolio modeling, analysis, and decision support, given pragmatic constraints such as data availability, uncertainties in estimating the impact of R&D spending, and practical operational overheads. Participants employed their deep expertise in approaches such as stochastic optimization, real options, Monte-Carlo analysis, Bayesian networks, decision theory, complex systems analysis, deep uncertainty, and technology-evolution modeling to critique the initial example models developed by the project’s core team and to conduct thought experiments grounded in real-life technology models, progress data, expert elicitation, and portfolio information. This engagement of participants’ methodological expertise with the practical requirements of real-life portfolio decision support yielded ideas for improved approaches, alternative methodological hypotheses, and hybridization of methodologies that are well-grounded theoretically, computationally sound, and realistically executable given data availability and other practical constraints.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Efficient Scalable Contact Network Generation from Population Data

Modeling the contacts among a population is critical to understanding the dynamics of a disease outbreak. Contact networks, where nodes are individuals and edges are contacts among them, are used to represent these complex individual-level interactions. In this work, we are given the daily activity schedules of an urban population that represent the activity location and time of individuals in a population during a single twenty four hour period over multiple days. Using collocation to determine contact between individuals, our goal is to extract hourly contact networks from large-scale activity data. We improve upon the existing adjacency matrix-based method by implementing our custom sparse matrix multiplication algorithm. Starting with a Python implementation, we achieve a 1600x speed up in the computation with a fast custom designed sparse matrix multiplier algorithm implemented in the C++ language. This work is central to future parallel designs of the problem.

97 MATHEMATICS AND COMPUTING↗

Data for The Role of Social Support on Midwestern Farmers’ Willingness to Grow Perennial Bioenergy Crops

The lack of farmers’ willingness to grow perennial bioenergy crops (PBCs) presents a critical barrier to the emergence of cellulosic biofuel production. The willingness relies on a complex network of economic, environmental, and social drivers, among which the influence of social factors (e.g., the influence of neighborhood, community, and communication) is less understood. This study addresses this knowledge gap via a survey analysis of midwestern farmers. The survey data are analyzed through ordinary least square regression and structural equation model, which together investigate the individual and interactive impacts of multiple factors on farmers’ decisions to adopt PBCs. Based on a farm-scale analysis, six statistically significant predictors of farmer willingness to grow PBCs are identified: perception of PBCs’ environment benefits, education level, willingness to take risks, familiarity with PBCs, portion of peers already growing PBCs, and support of biorefineries locating in the local community. Among these, the latter three predictors are social support variables. It is found that familiarity with the crops is the most significant predictor of willingness; familiarity is also an important intermediate variable that mediates the influence of many other predictors. In addition, peer adoption can both directly and indirectly affect willingness via its influence on familiarity. These findings suggest that it is a pressing need to improve farmers’ knowledge of PBCs to promote the adoption of such crops.

Economics↗

DyG-DPCD: A Distributed Parallel Community Detection Algorithm for Large-Scale Dynamic Graphs

Dynamic (Temporal) graphs capture the valuable evolution of real-world systems, from the continuously evolving patterns of social interactions and genetic pathways to the dynamic fluctuations of economic forces. Detecting communities for such evolving networks poses unique challenges. Detecting and analyzing the evolution of communities within dynamic graphs unlocks valuable insights into the underlying structural and temporal patterns of real-world systems. However, the sheer volume of modern graph data and the inherent complexity of the temporal dimension pose significant challenges to scalable community detection algorithms. Addressing this gap, our work explores the limited landscape of scalable distributed-memory parallel methods specifically designed for dynamic network community detection. We propose a novel parallel algorithm, DyG-DPCD (Dynamic Graph Distributed Parallel Community Detection), to detect communities in dynamic networks using the Message Passing Interface (MPI) framework. We present a vertex-centric approach, allowing us to detect communities through local optimization. Furthermore, we enhance our baseline algorithm by incorporating three heuristics, which improve the algorithm’s performance significantly while maintaining the quality of the solutions. We demonstrate the efficiency of our algorithm by experimenting on several real-world large-scale networks with hundreds of millions of edges spanning diverse domains. Notably, DyG-DPCD achieves speedups between 25× and 30× for large networks that we experimented on using NERSC compute nodes. In conclusion, our algorithm outperforms the STINGER parallel re-agglomeration algorithm by 30×.

97 MATHEMATICS AND COMPUTING↗

Enhancing Unknown Waveform Detection by Learning Intra and Inter-domain Dependencies with Advanced Attention Fusion Mechanisms

Detection of unknown waveforms in mission-critical communications is a crucial area of interest for the Department of Energy (DoE). Traditional methods and recent deep learning-based approaches often assume that the training set includes all possible classes, which is impractical for detecting new waveforms. This limitation gives rise to the problem of open-set recognition (OSR), which involves correctly identifying known classes while detecting and rejecting unknown or unseen classes. To address this limitation, we propose a novel dual-domain complex-valued neural architecture that jointly processes time-domain and frequency-domain signal representations using transformer mechanisms. A transformer model is a deep learning architecture that uses self-attention mechanisms to process and learn relationships in sequential data. Our model employs a cosine similarity loss to extract domain-specific features and incorporates a transformer architecture in the latent space to weigh the importance of different features from the time and frequency domains. The transformer layer includes stacked self-attention and cross-attention modules to learn intra-domain and inter-domain dependencies, creating a more holistic signal representation. An attention-based fusion module intelligently combines the time and frequency-domain features using multi-head attention, enabling the network to learn the optimal feature for each domain in each input signal. Quantitative results demonstrate the impact of these architectural choices on overall performance, showing significant improvement after incorporating self and cross-attention modules and using complex attention fusion over simple weighted fusion. Our ongoing work will focus on addressing the limitations of threshold-based OSR methods by developing a novel generative framework that integrates a conditional diffusion probabilistic model (DPM). DPM is a generative framework that learns to synthesize complex data by reversing a gradual noising process using a neural network trained to denoise step-by-step. Our goal is to leverage the inherent strengths of DPMs for identifying unknown signals more robustly. One primary advantage of using a DPM is its ability to provide a more reliable anomaly score based on the model's reconstruction error, rather than relying solely on classifier confidence. Additionally, the iterative denoising process of DPMs makes this approach naturally resilient to low Signal-to-Noise Ratio (SNR) conditions, where traditional methods often fail. By implementing this generative framework, we aim to enhance the model's capability to accurately detect unknown waveforms and maintain performance in challenging environments.

99 - GENERAL AND MISCELLANEOUS↗

Genomics-enabled analysis of specialized metabolism in bioenergy crops: Current progress and challenges

Plants produce a staggering diversity of specialized small molecule metabolites that play vital roles in mediating environmental interactions and stress adaptation. This chemical diversity derives from dynamic biosynthetic pathway networks that are often species-specific and operate under tight spatiotemporal and environmental control. A growing divide between demand and environmental challenges in food and bioenergy crop production have intensified research on these complex metabolite networks and their contribution to crop fitness. High-throughput omics technologies provide access to ever-increasing data resources for investigating plant metabolism. However, the efficiency of using such system-wide data to decode the gene and enzyme functions controlling specialized metabolism has remained limited; due largely to the recalcitrance of many plants to genetic approaches and the lack of ‘user-friendly’ biochemical tools for studying the diverse enzyme classes involved in specialized metabolism. With emphasis on terpenoid metabolism in the bioenergy crop switchgrass as an example, this review aims to illustrate current advances and challenges in the application of DNA synthesis and synthetic biology tools for accelerating the functional discovery of genes, enzymes and pathways in plant specialized metabolism. These technologies have accelerated knowledge development on the biosynthesis and physiological roles of diverse metabolite networks across many ecologically and economically important plant species and can provide resources for application to precision breeding and natural product metabolic engineering.

59 BASIC BIOLOGICAL SCIENCES↗

Novel and Emerging Capabilities that Can Provide a Holistic Understanding of the Plant Root Microbiome

In recent years, the root microbiome (i.e., microorganisms growing inside, on, or in close proximity to plant roots) has been shown to play an important role in plant health and productivity. Despite its importance, the root microbiome is challenging to study because of its complexity, heterogeneity, and subterranean location. Fortunately, root microbiome research has seen a tremendous influx of novel technologies (e.g., imaging tools, robotics, and molecular analyses), experimental platforms (e.g., micro- and mesocosms), and data integration, modeling, and prediction tools in the past decade that have greatly increased our ability to dissect the complex network of interactions between above- and belowground environmental parameters, plants, bacteria, and fungi that dictate soil and broader ecosystem health. Herein, we discuss methods that are currently used in root microbiome research and that can be expanded to phytobiome research in general ranging from laboratory studies to mesocosm-scale studies and, finally, to field studies; evaluate their relevance to ecosystem studies; and discuss future root microbiome research directions.

59 BASIC BIOLOGICAL SCIENCES↗