Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “network partition”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Efficient partitioning and assignment on programs for multiprocessor execution

The general problem studied is that of segmenting or partitioning programs for distribution across a multiprocessor system. Efficient partitioning and the assignment of program elements are of great importance since the time consumed in this overhead activity may easily dominate the computation, effectively eliminating any gains made by the use of the parallelism. In this study, the partitioning of sequentially structured programs (written in FORTRAN) is evaluated. Heuristics, developed for similar applications are examined. Finally, a model for queueing networks with finite queues is developed which may be used to analyze multiprocessor system architectures with a shared memory approach to the problem of partitioning. The properties of sequentially written programs form obstacles to large scale (at the procedure or subroutine level) parallelization. Data dependencies of even the minutest nature, reflecting the sequential development of the program, severely limit parallelism. The design of heuristic algorithms is tied to the experience gained in the parallel splitting. Parallelism obtained through the physical separation of data has seen some success, especially at the data element level. Data parallelism on a grander scale requires models that accurately reflect the effects of blocking caused by finite queues. A model for the approximation of the performance of finite queueing networks is developed. This model makes use of the decomposition approach combined with the efficiency of product form solutions.

Standley, Hilda M.↗

Wiring Viterbi decoders (splitting deBruijn graphs)

A new Viterbi decoder, capable of decoding convolutional codes with constraint lengths up to 15, is under development for the Deep Space Network (DSN). A key feature of this decoder is a two-level partitioning of the Viterbi state diagram into identical subgraphs. The larger subgraphs correspond to circuit boards, while the smaller subgraphs correspond to Very Large Scale Integration (VLSI) chips. The full decoder is built from identical boards, which in turn are built from identical chips. The resulting system is modular and hierarchical. The decoder is easy to implement, test, and repair because it uses a single VLSI chip design and a single board design. The partitioning is completely general in the sense that an appropriate number of boards or chips may be wired together to implement a Viterbi decoder of any size greater than or equal to the size of the module.

Collins, O.↗

FD/DAMA Scheme For Mobile/Satellite Communications

Integrated-Adaptive Mobile Access Protocol (I-AMAP) proposed to allocate communication channels to subscribers in first-generation MSAT-X mobile/satellite communication network. Based on concept of frequency-division/demand-assigned multiple access (FD/DAMA) where partition of available spectrum adapted to subscribers' demands for service. Requests processed, and competing requests resolved according to channel-access protocol, or free-access tree algorithm described in "Connection Protocol for Mobile/Satellite Communications" (NPO-17735). Assigned spectrum utilized efficiently.

Yan, Tsun-Yee↗

Network design consideration of a satellite-based mobile communications system

Technical considerations for the Mobile Satellite Experiment (MSAT-X), the ground segment testbed for the low-cost spectral efficient satellite-based mobile communications technologies being developed for the 1990's, are discussed. The Network Management Center contains a flexible resource sharing algorithm, the Demand Assigned Multiple Access scheme, which partitions the satellite transponder bandwidth among voice, data, and request channels. Satellite use of multiple UHF beams permits frequency reuse. The backhaul communications and the Telemetry, Tracking and Control traffic are provided through a single full-coverage SHF beam. Mobile Terminals communicate with the satellite using UHF. All communications including SHF-SHF between Base Stations and/or Gateways, are routed through the satellite. Because MSAT-X is an experimental network, higher level network protocols (which are service-specific) will be developed only to test the operation of the lowest three levels, the physical, data link, and network layers.

Yan, T.-Y.↗

Thermodynamic analysis of hydrogel swelling in aqueous sodium chloride solutions

Here, a comprehensive thermodynamic model is presented for swelling and salt partitioning of poly (n-isopropyl acrylamide) hydrogels in aqueous sodium chloride (NaCl) solutions at 298 K. Together with a Helmholtz energy expression for network elastic energy, a modified electrolyte Nonradom Two-Liquid activity coefficient model is used to represent both the long-range electrostatic interactions and the short-range van der Waals interactions present in the gel phase. The model parameters include one network parameter per hydrogel system plus two binary parameters per interaction pair in the polymer–solvent-salt systems. With semiquantitave agreement with experimental data for both hydrogel swelling and salt partitioning up to salt saturation, the analysis suggests the ion hydration between NaCl and water is the root cause for hydrogel deswelling in aqueous NaCl solution at high NaCl concentrations, i.e., NaCl weight fraction >0.03. On the other hand, the strong repulsive interaction between NaCl and hydrogel polymer is the dominant factor for salt partitioning.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Automated Instrumentation, Monitoring and Visualization of PVM Programs Using AIMS

We present views and analysis of the execution of several PVM (Parallel Virtual Machine) codes for Computational Fluid Dynamics on a networks of Sparcstations, including: (1) NAS Parallel Benchmarks CG and MG; (2) a multi-partitioning algorithm for NAS Parallel Benchmark SP; and (3) an overset grid flowsolver. These views and analysis were obtained using our Automated Instrumentation and Monitoring System (AIMS) version 3.0, a toolkit for debugging the performance of PVM programs. We will describe the architecture, operation and application of AIMS. The AIMS toolkit contains: (1) Xinstrument, which can automatically instrument various computational and communication constructs in message-passing parallel programs; (2) Monitor, a library of runtime trace-collection routines; (3) VK (Visual Kernel), an execution-animation tool with source-code clickback; and (4) Tally, a tool for statistical analysis of execution profiles. Currently, Xinstrument can handle C and Fortran 77 programs using PVM 3.2.x; Monitor has been implemented and tested on Sun 4 systems running SunOS 4.1.2; and VK uses XIIR5 and Motif 1.2. Data and views obtained using AIMS clearly illustrate several characteristic features of executing parallel programs on networked workstations: (1) the impact of long message latencies; (2) the impact of multiprogramming overheads and associated load imbalance; (3) cache and virtual-memory effects; and (4) significant skews between workstation clocks. Interestingly, AIMS can compensate for constant skew (zero drift) by calibrating the skew between a parent and its spawned children. In addition, AIMS' skew-compensation algorithm can adjust timestamps in a way that eliminates physically impossible communications (e.g., messages going backwards in time). Our current efforts are directed toward creating new views to explain the observed performance of PVM programs. Some of the features planned for the near future include: (1) ConfigView, showing the physical topology of the virtual machine, inferred using specially formatted IP (Internet Protocol) packets: and (2) LoadView, synchronous animation of PVM-program execution and resource-utilization patterns.

Mehra, Pankaj↗

DRIPS: Dynamic Rebalancing of Pipelined Streaming Applications on CGRAs

Coarse-grained reconfigurable arrays (CGRAs) provide higher flexibility than application-specific integrated circuits (ASICs) and higher efficiency than fine-grained reconfigurable devices such as Field Programmable Gate Arrays (FPGAs). However, CGRAs are generally designed to support offloading of a single kernel. While their design, based on communicating functional units, appears to naturally suit data streaming applications composed of multiple cooperating kernels, current approaches only statically partition the resources across application kernels. However, emerging streaming applications at the edge (scientific instruments, sensor networks, network processing) perform much more than digital signal processing and often are data and input dependent. This leads to extremely variable kernel execution times, severely impacting the throughput of the entire pipeline if resources are only statically allocated. Therefore, in this paper, we propose DRIPS — a coarse-grained, dynamically, and partially reconfigurable array for data-dependent streaming applications. We present a unified compiler framework to facilitate the mapping of a given streaming application onto the DRIPS CGRA architecture. The experimental results show that DRIPS achieves an average throughput improvement of 1.46$\times$ across a set of representative applications over a statically partitioned solution. The additional area overhead to enable dynamic rebalancing consumes 16.34% of the entire area for a 5x5 CGRA prototype.

Tan, Cheng↗

Vegetation type is an important predictor of the arctic summer land surface energy budget

Despite the importance of high-latitude surface energy budgets (SEBs) for land-climate interactions in the rapidly changing Arctic, uncertainties in their prediction persist. Here, we harmonize SEB observations across a network of vegetated and glaciated sites at circumpolar scale (1994–2021). Our variance-partitioning analysis identifies vegetation type as an important predictor for SEB-components during Arctic summer (June-August), compared to other SEB-drivers including climate, latitude and permafrost characteristics. Differences among vegetation types can be of similar magnitude as between vegetation and glacier surfaces and are especially high for summer sensible and latent heat fluxes. The timing of SEB-flux summer-regimes (when daily mean values exceed 0 Wm –2 ) relative to snow-free and -onset dates varies substantially depending on vegetation type, implying vegetation controls on snow-cover and SEB-flux seasonality. Our results indicate complex shifts in surface energy fluxes with land-cover transitions and a lengthening summer season, and highlight the potential for improving future Earth system models via a refined representation of Arctic vegetation types.

54 ENVIRONMENTAL SCIENCES↗

Improved Multi-Partition Method for Line-Based Iteration Schemes

Regular 3-dimensional multi-partitioning has been shown to be an efficient domain decomposition method for the parallelization of ADI-type algorithms on MIMD architectures. This paper discusses further improvements that can be made to the scheme that increase the granularity and reduce the communication density. These improvements, which are illustrated by simulation and parallel benchmark results, make multi-partitioning the method of choice on systems with relatively poor communication capabilities, such as networks of workstations, or on massively parallel machines with very fast processors, such as the IBM SP2.

Smith, Merritt H.↗

Distributed simulation using a real-time shared memory network

The Advanced Control Technology Branch of the NASA Lewis Research Center performs research in the area of advanced digital controls for aeronautic and space propulsion systems. This work requires the real-time implementation of both control software and complex dynamical models of the propulsion system. We are implementing these systems in a distributed, multi-vendor computer environment. Therefore, a need exists for real-time communication and synchronization between the distributed multi-vendor computers. A shared memory network is a potential solution which offers several advantages over other real-time communication approaches. A candidate shared memory network was tested for basic performance. The shared memory network was then used to implement a distributed simulation of a ramjet engine. The accuracy and execution time of the distributed simulation was measured and compared to the performance of the non-partitioned simulation. The ease of partitioning the simulation, the minimal time required to develop for communication between the processors and the resulting execution time all indicate that the shared memory network is a real-time communication technique worthy of serious consideration.

Simon, Donald L.↗

D2NO: Efficient handling of heterogeneous input function spaces with distributed deep neural operators

Neural operators have been applied in various scientific fields, such as solving parametric partial differential equations, dynamical systems with control, and inverse problems. However, challenges arise when dealing with input functions that exhibit heterogeneous properties, requiring multiple sensors to handle functions with minimal regularity. To address this issue, discretization-invariant neural operators have been used, allowing the sampling of diverse input functions with different sensor locations. However, existing frameworks still require an equal number of sensors for all functions. We propose a novel distributed approach to further relax the discretization requirements and solve the heterogeneous dataset challenges. Our method involves partitioning the input function space and processing individual input functions using independent and separate neural networks. A centralized neural network is used to handle shared information across all output functions. This distributed methodology reduces the number of gradient descent back-propagation steps, improving efficiency while maintaining accuracy. Here, we demonstrate that the corresponding neural network is a universal approximator of continuous nonlinear operators and present three numerical examples to validate its performance.

97 MATHEMATICS AND COMPUTING↗

Neurocomputing strategies in decomposition based structural design

The present paper explores the applicability of neurocomputing strategies in decomposition based structural optimization problems. It is shown that the modeling capability of a backpropagation neural network can be used to detect weak couplings in a system, and to effectively decompose it into smaller, more tractable, subsystems. When such partitioning of a design space is possible, parallel optimization can be performed in each subsystem, with a penalty term added to its objective function to account for constraint violations in all other subsystems. Dependencies among subsystems are represented in terms of global design variables, and a neural network is used to map the relations between these variables and all subsystem constraints. A vector quantization technique, referred to as a z-Network, can effectively be used for this purpose. The approach is illustrated with applications to minimum weight sizing of truss structures with multiple design constraints.

Szewczyk, Z.↗

Partitioning Evapotranspiration in Semiarid Grassland and Shrubland Ecosystems Using Diurnal Surface Temperature Variation

The encroachment of woody plants in grasslands across the Western U.S. will affect soil water availability by altering the contributions of evaporation (E) and transpiration (T) to total evapotranspiration (ET). To study this phenomenon, a network of flux stations is in place to measure ET in grass- and shrub-dominated ecosystems throughout the Western U.S. A method is described and tested here to partition the daily measurements of ET into E and T based on diurnal surface temperature variations of the soil and standard energy balance theory. The difference between the mid-afternoon and pre-dawn soil surface temperature, termed Apparent Thermal Inertia (I(sub A)), was used to identify days when E was negligible, and thus, ET=T. For other days, a three-step procedure based on energy balance equations was used to estimate Qe contributions of daily E and T to total daily ET. The method was tested at Walnut Gulch Experimental Watershed in southeast Arizona based on Bowen ratio estimates of ET and continuous measurements of surface temperature with an infrared thermometer (IRT) from 2004- 2005, and a second dataset of Bowen ratio, IRT and stem-flow gage measurements in 2003. Results showed that reasonable estimates of daily T were obtained for a multi-year period with ease of operation and minimal cost. With known season-long daily T, E and ET, it is possible to determine the soil water availability associated with grass- and shrub-dominated sites and better understand the hydrologic impact of regional woody plant encroachment.

Moran, M. Susan↗

U-net architected deep material network training with microstructure local field information

The Deep Material Network (DMN) has recently emerged as a powerful reduced-order modeling framework for simulating the mechanical response of heterogeneous materials such as composites. Unlike most data-driven approaches that directly learn a material’s response under prescribed loading, the DMN acts as a homogenization operator, learning the kinematic constraints and mechanical interactions of the underlying microstructure. However, traditional DMN training relies exclusively on homogenized effective properties derived from Direct Numerical Simulations (DNS), discarding the rich local field data that govern microstructural interactions. In this work, we extend the DMN framework to incorporate such local field information into the offline training process. Utilizing a U-Net architecture, we augment the DMN training objective to include the first and second statistical moments of the local stress fields obtained from linear DNS. This ensures that the learned network topology not only fits the effective stiffness but also accurately reflects the internal local stress and strain partitioning of the microstructure. The results confirm that supervising the localization process during training yields a superior surrogate model, reducing local prediction errors by an order of magnitude and significantly improving generalization to unseen nonlinear constitutive behaviors compared to traditional DMNs.

36 MATERIALS SCIENCE↗

Capacities of Entanglement Distribution From a Central Source

Distribution of entanglement is an essential task in quantum information processing and the realization of quantum networks. In our work, we theoretically investigate the scenario where a central source prepares an N -partite entangled state and transmits each entangled subsystem to one of N receivers through noisy quantum channels. The receivers are then able to perform local operations assisted by unlimited classical communication to distill target entangled states from the noisy channel output. In this operational context, we define the EPR distribution capacity and the GHZ distribution capacity of a quantum channel as the largest rates at which Einstein-Podolsky-Rosen (EPR) states and Greenberger-Horne-Zeilinger (GHZ) states can be faithfully distributed through the channel, respectively. We establish lower and upper bounds on the EPR distribution capacity by connecting it with the task of assisted entanglement distillation. We also construct an explicit protocol consisting of a combination of a quantum communication code and a classical-post-processing-assisted entanglement generation code, which yields a simple achievable lower bound for generic channels. As applications of these results, we give an exact expression for the EPR distribution capacity over two erasure channels and bounds on the EPR distribution capacity over two generalized amplitude damping channels. We also bound the GHZ distribution capacity, which results in an exact characterization of the GHZ distribution capacity when the most noisy channel is a dephasing channel.

42 ENGINEERING↗

Trace elements and their isotopes in streams and rivers

The occurrence and speciation of trace elements in streams and rivers are controlled by interconnected hydrogeological, biogeochemical, and anthropogenic factors that include rock weathering, climate, vegetation, and land use within the watershed. This chapter provides a broad overview of trace element abundance, speciation, and mobility in streams and rivers, including an assessment of how trace elements are partitioned into dissolved, colloidal, and particulate phases. We also discuss how trace elements are mobilized into stream networks and transformed during transport. Select isotopic systems are reviewed to provide examples of how isotopic analyses can be used to understand geochemical and anthropogenic processes.

Herndon, Elizabeth [ORNL] (ORCID:0000000291945493)↗

A Principled Framework to Assess the Information-Theoretic Fitness of Brain Functional Sub-Circuits

In systems and network neuroscience, many common practices in brain connectomic analysis are often not properly scrutinized. One such practice is mapping a predetermined set of sub-circuits, like functional networks (FNs), onto subjects’ functional connectomes (FCs) without adequately assessing the information-theoretic appropriateness of the partition. Another practice that goes unchallenged is thresholding weighted FCs to remove spurious connections without justifying the chosen threshold. This paper leverages recent theoretical advances in Stochastic Block Models (SBMs) to formally define and quantify the information-theoretic fitness (e.g., prominence) of a predetermined set of FNs when mapped to individual FCs under different fMRI task conditions. Our framework allows for evaluating any combination of FC granularity, FN partition, and thresholding strategy, thereby optimizing these choices to preserve the important topological features of the human brain connectomes. By applying to the Human Connectome Project with Schaefer parcellations at multiple levels of granularity, the framework showed that the common thresholding value of 0.25 was indeed information-theoretically valid for group-average FCs, despite its previous lack of justification. Our results pave the way for the proper use of FNs and thresholding methods, and provide insights for future research in individualized parcellations.

Duong-Tran, Duy (ORCID:0009000944967575)↗

Using Superconducting Thin Films in Microwave Lines

High temperature superconductors(HTS) and microwaves devices form the ideal partnership. The application of superconductors in microwave devices, components and systems allows the reduction in size, power consumption and insertion loss. The surface resistance of high-Tc superconductors has been found to be two orders of magnitude lower than normal conducting copper materials. The reduction in size and power requirements, which together both lead to a reduction in system mass, coupled with reasonably accessible operating temperatures, suggest that HTS microwave components should find ready application in satellite communications systems. At present, multi- channeling communication networks demand filters with narrow bandwidth in order to allow the available RF frequency spectrum to be partitioned into small frequency bands, -and possible variation of dielectric constant from substrate to substrate is undesirable. Microwave multiplexers demand the fabrication of two identical filters in each channel. Thus, the filter with tuning function is preferable. Tunable filters are the critical component for phased array antennas in order to electronically steer the radiated beam. To fabricate a tunable filter that uses an electric field for operation, one would like a material that provides a large change on dielectric constant for a given electric field, yet has a relatively low tangent in order to minimize the insertion loss of the device. Ferroelectrics have been the materials of choice. Their large dielectric constant sufficiently increases the coupling between microwave resonators and its dependence on electric field provides timability. Development of technology promises to diminish tangent loss. The use of thin ferroelectric films sufficiently decreases insertion losses keeping considerable potential for applications. NASA Lewis Research Center is the one of the leading centers in investigation of superconductors/ferroelectric tunable components for microwave devices. A large number of possible microwave devices were fabricated and tested on the basis of thin film multilayer superconductor-ferroelectric structures. In major cases the systems with edge-coupling scheme were investigated. Dr. Genkin has recently focused on the new potentialities which implements the using of thin ferroelectric films in filters fabricated with end-coupled microstrip lines. Numerical modeling shows that these systems have large potential for application in tunable narrow- and wide-bandpass filters in the frequency range 10-20 GHz. The phase shifter with end-coupled resonant sections was fabricated and tested. Experimental results show large tunability, particular in low voltages. The possible optimization of this structure promises to improve the obtained result and to reach the low level of insertion losses.

Genkin, Varery↗