Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “heterogeneous networks”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Heterogeneous Wireless Mesh Network Technology Evaluation for Space Proximity and Surface Applications

NASA has identified standardized wireless mesh networking as a key technology for future human and robotic space exploration. Wireless mesh networks enable rapid deployment, provide coverage in undeveloped regions. Mesh networks are also self-healing, resilient, and extensible, qualities not found in traditional infrastructure-based networks. Mesh networks can offer lower size, weight, and power (SWaP) than overlapped infrastructure-perapplication. To better understand the maturity, characteristics and capability of the technology, we developed an 802.11 mesh network consisting of a combination of heterogeneous commercial off-the-shelf devices and opensource firmware and software packages. Various streaming applications were operated over the mesh network, including voice and video, and performance measurements were made under different operating scenarios. During the testing several issues with the currently implemented mesh network technology were identified and outlined for future work.

DeCristofaro, Michael A.↗

Structural evolution during gelation of pea and whey proteins envisaged by time-resolved ultra-small-angle x-ray scattering (USAXS)

Hydrogels from plant proteins commonly exhibit inferior gel strength compared to those from dairy proteins partially due to their distinct gel networks. How protein aggregates to form such networks in response to heat remains largely unknown. In this work, pea (PPI) and whey (WPI) protein isolate gels were produced at the same protein content and similar heating/cooling rate. The process was monitored using rheology, microscopy and in situ ultra-small-angle x-ray scattering (USAXS). Rheology showed an initial decrease in G' and G'' in PPI followed by a steady increase when the temperature surpassed ~60 °C whereas a much higher temperature (~80 °C) was required for WPI, both using 2 °C/min heating rate. Microscopy showed a coarse and heterogenous network in PPI, whereas for WPI, the network was finer and more continuous. In both gels, nano-sized spherical or ellipsoidal particles were present as the basic constituents. USAXS found individual protein was dominant in PPI or WPI solution at temperature below 57 °C. Their proportions decreased together with appearance of aggregates with an average R g of 9–10 nm in PPI and 6–7 nm in WPI at higher temperature. The size of the aggregates changed slightly during further heating and cooling, but their proportions increased. Power law exponents revealed the aggregates were mass fractals for WPI and PPI gels, and they became more compact during heating. Our findings suggested formation of primary aggregates in protein gel networks is a more organized process and provided theoretical guidance for production of high protein food gels with desirable texture.

59 BASIC BIOLOGICAL SCIENCES↗

Energy–Performance Trade-offs in Privacy-Preserving Federated Learning on SmartNIC-Enabled HPC Systems

Federated learning (FL) is increasingly deployed on accelerator-rich high-performance computing (HPC) systems, yet the system-level energy cost of privacy-aware FL remains poorly understood, particularly across heterogeneous networking and server-placement options. We present a measurement-driven study of energy–performance trade-offs for FL on GH200-class nodes across three deployment configurations: CPU-Ethernet, CPU-InfiniBand (RDMA-capable), and a DPU-hosted FL server over InfiniBand using a BlueField-3 SmartNIC/DPU. Using NVIDIA FLARE (NVFLARE), we align node-level power telemetry with per-round timing extracted from NVFLARE logs to quantify time-to-solution (TTS), energy-to-solution (ETS), energy-delay product (EDP), and synchronization behavior for three transformer models (ALBERT, DistilBERT, BERT), trained with and without differential privacy (DP). We find that interconnect choice is the dominant driver of runtime and energy: host-managed InfiniBand consistently reduces communication overhead versus Ethernet, yielding lower TTS/ETS/EDP. In contrast, in our NVFLARE deployment, placing the FL server on the DPU does not consistently match CPU-InfiniBand performance and can be slower—especially for larger models—highlighting that server placement alone is not sufficient to guarantee end-to-end gains. Finally, under our fixed-round protocol, DP increases per-round cost and runtime variance; ETS increases largely in proportion to TTS because average node power remains relatively stable across configurations.

Kotevska, Olivera [ORNL] (ORCID:0000000316772243)↗

Non-resonant Bragg scattering four-wave mixing at near-visible wavelengths in low-confinement silicon nitride waveguides

Quantum state coherent frequency conversion processes—such as Bragg-scattering four-wave mixing (BSFWM)—hold promise as a flexible technique for networking heterogeneous and distant quantum systems. In this Letter, we demonstrate BSFWM within an extended (1.2-m) low-confinement silicon nitride waveguide and show that this system has the potential for near-unity frequency conversion in visible and near-visible wavelength ranges. Using sensitive classical heterodyne laser spectroscopy at low optical powers, we characterize the Kerr coefficient (∼1.55 W −1 m −1 ) and linear propagation loss (∼0.0175 dB/cm) of this non-resonant waveguide system, revealing a record-high nonlinear figure of merit (NFM = γ / α ≈ 3.85 W −1 ) for BSFWM of near-visible light in non-resonant silicon nitride waveguides. We predict how, at high yet achievable on-chip optical powers, this NFM would yield a comparatively large frequency conversion efficiency, opening the door to near-unity flexible frequency conversion without cavity enhancement and resulting bandwidth constraints.

Jaber, Nicholas (ORCID:000900070106809X)↗

A distributed data base management facility for the CAD/CAM environment

Current/PAD research in the area of distributed data base management considers facilities for supporting CAD/CAM data management in a heterogeneous network of computers encompassing multiple data base managers supporting a variety of data models. These facilities include coordinated execution of multiple DBMSs to provide for administration of and access to data distributed across them.

Balza, R. M.↗

Extensions to the Parallel Real-Time Artificial Intelligence System (PRAIS) for fault-tolerant heterogeneous cycle-stealing reasoning

Extensions to an architecture for real-time, distributed (parallel) knowledge-based systems called the Parallel Real-time Artificial Intelligence System (PRAIS) are discussed. PRAIS strives for transparently parallelizing production (rule-based) systems, even under real-time constraints. PRAIS accomplished these goals (presented at the first annual C Language Integrated Production System (CLIPS) conference) by incorporating a dynamic task scheduler, operating system extensions for fact handling, and message-passing among multiple copies of CLIPS executing on a virtual blackboard. This distributed knowledge-based system tool uses the portability of CLIPS and common message-passing protocols to operate over a heterogeneous network of processors. Results using the original PRAIS architecture over a network of Sun 3's, Sun 4's and VAX's are presented. Mechanisms using the producer-consumer model to extend the architecture for fault-tolerance and distributed truth maintenance initiation are also discussed.

Goldstein, David↗

Security aspects of space operations data

This paper deals with data security. It identifies security threats to European Space Agency's (ESA) In Orbit Infrastructure Ground Segment (IOI GS) and proposes a method of dealing with its complex data structures from the security point of view. It is part of the 'Analysis of Failure Modes, Effects Hazards and Risks of the IOI GS for Operations, including Backup Facilities and Functions' carried out on behalf of the European Space Operations Center (ESOC). The security part of this analysis has been prepared with the following aspects in mind: ESA's large decentralized ground facilities for operations, the multiple organizations/users involved in the operations and the developments of ground data systems, and the large heterogeneous network structure enabling access to (sensitive) data which does involve crossing organizational boundaries. An IOI GS data objects classification is introduced to determine the extent of the necessary protection mechanisms. The proposal of security countermeasures is oriented towards the European 'Information Technology Security Evaluation Criteria (ITSEC)' whose hierarchically organized requirements can be directly mapped to the security sensitivity classification.

Schmitz, Stefan↗

Advanced information processing system: Hosting of advanced guidance, navigation and control algorithms on AIPS using ASTER

This program demonstrated the integration of a number of technologies that can increase the availability and reliability of launch vehicles while lowering costs. Availability is increased with an advanced guidance algorithm that adapts trajectories in real-time. Reliability is increased with fault-tolerant computers and communication protocols. Costs are reduced by automatically generating code and documentation. This program was realized through the cooperative efforts of academia, industry, and government. The NASA-LaRC coordinated the effort, while Draper performed the integration. Georgia Institute of Technology supplied a weak Hamiltonian finite element method for optimal control problems. Martin Marietta used MATLAB to apply this method to a launch vehicle (FENOC). Draper supplied the fault-tolerant computing and software automation technology. The fault-tolerant technology includes sequential and parallel fault-tolerant processors (FTP & FTPP) and authentication protocols (AP) for communication. Fault-tolerant technology was incrementally incorporated. Development culminated with a heterogeneous network of workstations and fault-tolerant computers using AP. Draper's software automation system, ASTER, was used to specify a static guidance system based on FENOC, navigation, flight control (GN&C), models, and the interface to a user interface for mission control. ASTER generated Ada code for GN&C and C code for models. An algebraic transform engine (ATE) was developed to automatically translate MATLAB scripts into ASTER.

Brenner, Richard↗

PRAIS: Distributed, real-time knowledge-based systems made easy

This paper discusses an architecture for real-time, distributed (parallel) knowledge-based systems called the Parallel Real-time Artificial Intelligence System (PRAIS). PRAIS strives for transparently parallelizing production (rule-based) systems, even when under real-time constraints. PRAIS accomplishes these goals by incorporating a dynamic task scheduler, operating system extensions for fact handling, and message-passing among multiple copies of CLIPS executing on a virtual blackboard. This distributed knowledge-based system tool uses the portability of CLIPS and common message-passing protocols to operate over a heterogeneous network of processors.

Goldstein, David G.↗

Methodologies and systems for heterogeneous concurrent computing

Heterogeneous concurrent computing is gaining increasing acceptance as an alternative or complementary paradigm to multiprocessor-based parallel processing as well as to conventional supercomputing. While algorithmic and programming aspects of heterogeneous concurrent computing are similar to their parallel processing counterparts, system issues, partitioning and scheduling, and performance aspects are significantly different. In this paper, we discuss critical design and implementation issues in heterogeneous concurrent computing, and describe techniques for enhancing its effectiveness. In particular, we highlight the system level infrastructures that are required, aspects of parallel algorithm development that most affect performance, system capabilities and limitations, and tools and methodologies for effective computing in heterogeneous networked environments. We also present recent developments and experiences in the context of the PVM system and comment on ongoing and future work.

Sunderam, V. S.↗

An Adaptive Flow Solver for Air-Borne Vehicles Undergoing Time-Dependent Motions/Deformations

This report describes a concurrent Euler flow solver for flows around complex 3-D bodies. The solver is based on a cell-centered finite volume methodology on 3-D unstructured tetrahedral grids. In this algorithm, spatial discretization for the inviscid convective term is accomplished using an upwind scheme. A localized reconstruction is done for flow variables which is second order accurate. Evolution in time is accomplished using an explicit three-stage Runge-Kutta method which has second order temporal accuracy. This is adapted for concurrent execution using another proven methodology based on concurrent graph abstraction. This solver operates on heterogeneous network architectures. These architectures may include a broad variety of UNIX workstations and PCs running Windows NT, symmetric multiprocessors and distributed-memory multi-computers. The unstructured grid is generated using commercial grid generation tools. The grid is automatically partitioned using a concurrent algorithm based on heat diffusion. This results in memory requirements that are inversely proportional to the number of processors. The solver uses automatic granularity control and resource management techniques both to balance load and communication requirements, and deal with differing memory constraints. These ideas are again based on heat diffusion. Results are subsequently combined for visualization and analysis using commercial CFD tools. Flow simulation results are demonstrated for a constant section wing at subsonic, transonic, and a supersonic case. These results are compared with experimental data and numerical results of other researchers. Performance results are under way for a variety of network topologies.

Singh, Jatinder↗

Heterogeneous Distributed Computing for Computational Aerosciences

The research supported under this award focuses on heterogeneous distributed computing for high-performance applications, with particular emphasis on computational aerosciences. The overall goal of this project was to and investigate issues in, and develop solutions to, efficient execution of computational aeroscience codes in heterogeneous concurrent computing environments. In particular, we worked in the context of the PVM[1] system and, subsequent to detailed conversion efforts and performance benchmarking, devising novel techniques to increase the efficacy of heterogeneous networked environments for computational aerosciences. Our work has been based upon the NAS Parallel Benchmark suite, but has also recently expanded in scope to include the NAS I/O benchmarks as specified in the NHT-1 document. In this report we summarize our research accomplishments under the auspices of the grant.

Sunderam, Vaidy S.↗

Automated Concurrent Blackboard System Generation in C++

In his 1992 Ph.D. thesis, "Design and Analysis Techniques for Concurrent Blackboard Systems", John McManus defined several performance metrics for concurrent blackboard systems and developed a suite of tools for creating and analyzing such systems. These tools allow a user to analyze a concurrent blackboard system design and predict the performance of the system before any code is written. The design can be modified until simulated performance is satisfactory. Then, the code generator can be invoked to generate automatically all of the code required for the concurrent blackboard system except for the code implementing the functionality of each knowledge source. We have completed the port of the source code generator and a simulator for a concurrent blackboard system. The source code generator generates the necessary C++ source code to implement the concurrent blackboard system using Parallel Virtual Machine (PVM) running on a heterogeneous network of UNIX(trademark) workstations. The concurrent blackboard simulator uses the blackboard specification file to predict the performance of the concurrent blackboard design. The only part of the source code for the concurrent blackboard system that the user must supply is the code implementing the functionality of the knowledge sources.

Kaplan, J. A.↗

A Component-based Programming Model for Composite, Distributed Applications

The nature of scientific programming is evolving to larger, composite applications that are composed of smaller element applications. These composite applications are more frequently being targeted for distributed, heterogeneous networks of computers. They are most likely programmed by a group of developers. Software component technology and computational frameworks are being proposed and developed to meet the programming requirements of these new applications. Historically, programming systems have had a hard time being accepted by the scientific programming community. In this paper, a programming model is outlined that attempts to organize the software component concepts and fundamental programming entities into programming abstractions that will be better understood by the application developers. The programming model is designed to support computational frameworks that manage many of the tedious programming details, but also that allow sufficient programmer control to design an accurate, high-performance application.

Eidson, Thomas M.↗

Collectives for Multiple Resource Job Scheduling Across Heterogeneous Servers

Efficient management of large-scale, distributed data storage and processing systems is a major challenge for many computational applications. Many of these systems are characterized by multi-resource tasks processed across a heterogeneous network. Conventional approaches, such as load balancing, work well for centralized, single resource problems, but breakdown in the more general case. In addition, most approaches are often based on heuristics which do not directly attempt to optimize the world utility. In this paper, we propose an agent based control system using the theory of collectives. We configure the servers of our network with agents who make local job scheduling decisions. These decisions are based on local goals which are constructed to be aligned with the objective of optimizing the overall efficiency of the system. We demonstrate that multi-agent systems in which all the agents attempt to optimize the same global utility function (team game) only marginally outperform conventional load balancing. On the other hand, agents configured using collectives outperform both team games and load balancing (by up to four times for the latter), despite their distributed nature and their limited access to information.

Tumer, K.↗

Understanding Biomass and Polymer Feedstock Variability Through Analytical Pyrolysis and Two Dimensional Gas Chromatography and Mass Spectrometry

Moving toward a circular carbon economy requires enabling reuse of carbon-based macromolecules like those in biomass and polymers. Meeting this goal requires increased upcycling of waste products into higher value products. For biomass waste such as corn stover, pine needles, or bark, one approach is using pyrolysis to convert these feedstocks into bio-oil that can be used for liquid fuel or transformed into other hydrocarbon based products, like bio-polymers. Waste plastics can be upcycled or recycled into new products, or alternatively converted into bio-oil by pyrolysis. Macromolecules from biomass and polymers are challenging to convert into liquid fuel or chemical feedstocks, due to large and sometimes heterogeneous networks of polymeric bonds. We work with the Idaho National Laboratory’s Biomass Feedstock National User Facility and Department of Energy’s Feedstock Conversion Interface Consortium to develop a molecular understanding of biomass variability that occurs naturally and as a result of storage or preprocessing techniques applied to feedstocks, to understand chemical reactions occurring during pyrolysis and other biomass or polymer conversion techniques. Using two dimensional gas chromatography mass spectrometry coupled to analytical pyrolysis provides insight into the molecular changes that occur during the pyrolysis event, and these insights can extend to bench or pilot scale pyrolysis conversions. Understanding pyrolysis of biomass and polymers at a molecular level contributes to development of new tools to predict and manipulate the optimum conditions and identify the best storage, preprocessing, and conversion techniques to maximize output of specific high-value chemical species in the conversion to fuels, or to characterize feedstocks for blending before pyrolysis. Identification and understanding of the molecular level changes resulting from novel preprocessing techniques that can be used to control and manipulate the chemical output of bench or pilot scale pyrolysis is another focus of investigation, including the use of gamma irradiation, acid, or enzyme pre-treatments. We also focus on the impacts of storage conditions and techniques, which can affect moisture levels, degradation, and impacts from naturally occurring microbes in the environment.

09 BIOMASS FUELS↗

Physics-informed heterogeneous graph neural networks for DC blocker placement

The threat of geomagnetic disturbances (GMDs) to the reliable operation of the bulk energy system has spurred the development of effective strategies for mitigating their impacts. One such approach involves placing transformer neutral blocking devices, which interrupt the path of geomagnetically induced currents (GICs) to limit their impact. The high cost of these devices and the sparsity of transformers that experience high GICs during GMD events, however, calls for a sparse placement strategy that involves high computational cost. To address this challenge, we developed a physics-informed heterogeneous graph neural network (PIHGNN) for solving the graph-based dc-blocker placement problem. Our approach combines a heterogeneous graph neural network (HGNN) with a physics-informed neural network (PINN) to capture the diverse types of nodes and edges in ac/dc networks and incorporates the physical laws of the power grid. We train the PIHGNN model using a surrogate power flow model and validate it using case studies. Results demonstrate that PIHGNN can effectively and efficiently support the deployment of GIC dc-current blockers, ensuring the continued supply of electricity to meet societal demands. Furthermore, our approach has the potential to contribute to the development of more reliable and resilient power grids capable of withstanding the growing threat that GMDs pose.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Synchronization in electric power networks with inherent heterogeneity up to 100% inverter-based renewable generation

The synchronized operation of power generators is the foundation of electric power network stability and a key to the prevention of undesired power outages and blackouts. Here, we derive the conditions that guarantee synchronization in power networks with inherent generator heterogeneity when subjected to small perturbations, and perform a parametric sensitivity analysis to understand synchronization with varied types of generators. As inverter-based resources, which are the primary interfacing technology for many renewable sources of energy, have supplanted synchronous generators in ever growing numbers, the center of attention on associated integration challenges have resided primarily on the role of declining system inertia. Our results instead highlight the critical role of generator damping in achieving a stable state of synchronization. Additionally, we report the feasibility of operating interconnected electric grids with up to 100% power contribution from inverter-based renewable generation technologies. Our study has important implications as it sets the basis for the development of advanced control architectures and grid optimization methods that ensure synchronization and further pave the path towards the decarbonization of the electric power sector.

24 POWER TRANSMISSION AND DISTRIBUTION↗