Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel communication”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

C-Band Airport Surface Communications System Engineering-Initial High-Level Safety Risk Assessment and Mitigation

This document is being provided as part of ITT's NASA Glenn Research Center Aerospace Communication Systems Technical Support (ACSTS) contract: "New ATM Requirements--Future Communications, C-Band and L-Band Communications Standard Development." ITT has completed a safety hazard analysis providing a preliminary safety assessment for the proposed C-band (5091- to 5150-MHz) airport surface communication system. The assessment was performed following the guidelines outlined in the Federal Aviation Administration Safety Risk Management Guidance for System Acquisitions document. The safety analysis did not identify any hazards with an unacceptable risk, though a number of hazards with a medium risk were documented. This effort represents an initial high-level safety hazard analysis and notes the triggers for risk reassessment. A detailed safety hazards analysis is recommended as a follow-on activity to assess particular components of the C-band communication system after the profile is finalized and system rollout timing is determined. A security risk assessment has been performed by NASA as a parallel activity. While safety analysis is concerned with a prevention of accidental errors and failures, the security threat analysis focuses on deliberate attacks. Both processes identify the events that affect operation of the system; and from a safety perspective the security threats may present safety risks.

Zelkin, Natalie↗

Improved algorithms for mapping pipelined and parallel computations

Recent work on the problem of mapping pipelined or parallel computations onto linear array, shared memory, and host-satellite systems is extended. It is shown how these problems can be solved even more efficiently when computation module execution times are bounded from below, intermodule communication times are bounded from above, and the processors satisfy certain homogeneity constraints. The improved algorithms have significantly lower time and space complexities than the more general algorithms: in one case, an O(nm3) time algorithm for mapping m modules onto n processors is replaced with an O(nm log m) time algorithm, and the space requirements are reduced from O(nm2) to O(m). Run-time complexity is reduced further with parallel mapping algorithms based on these improvements, which run on the architectures for which they create mappings.

Nicol, David M.↗

Systems and Methods for Transmitting Data Using Parallel Decode Capacity Optimized Symbol Constellations

Transmitters are described that use non-uniformly spaced symbol constellations that have increased capacity compared to conventional constellations. In many embodiments, a transmitter includes a coder configured to receive bits and output encoded bits, a mapper configured to map said encoded bits to symbols in a non-uniform symbol constellation selected from a plurality of symbol constellations, and a modulator configured to generate a signal for transmission via the communication channel using symbols generated by the mapper. In a variety of embodiments, wherein the non-uniform symbol constellation comprises a set of non-uniformly spaced constellation points, and the location points in the non-uniform symbol constellation are chosen to optimize parallel decode capacity of the non-uniform symbol constellation subject to at least one optimization constraint. In many embodiments, the non-uniform symbol constellation is a quadrature amplitude modulation constellation or a phase shift keyed constellation.

Barsoum, Maged F.↗

Distributed communications and control network for robotic mining

The application of robotics to coal mining machines is one approach pursued to increase productivity while providing enhanced safety for the coal miner. Toward that end, a network composed of microcontrollers, computers, expert systems, real time operating systems, and a variety of program languages are being integrated that will act as the backbone for intelligent machine operation. Actual mining machines, including a few customized ones, have been given telerobotic semiautonomous capabilities by applying the described network. Control devices, intelligent sensors and computers onboard these machines are showing promise of achieving improved mining productivity and safety benefits. Current research using these machines involves navigation, multiple machine interaction, machine diagnostics, mineral detection, and graphical machine representation. Guidance sensors and systems employed include: sonar, laser rangers, gyroscopes, magnetometers, clinometers, and accelerometers. Information on the network of hardware/software and its implementation on mining machines are presented. Anticipated coal production operations using the network are discussed. A parallelism is also drawn between the direction of present day underground coal mining research to how the lunar soil (regolith) may be mined. A conceptual lunar mining operation that employs a distributed communication and control network is detailed.

Schiffbauer, William H.↗

Antenna-array, phase quadrature tracking system

Phase relationship between input signals appearing on widely-spaced parallel connected antenna elements in array is automatically adjusted in phase quadrature tracking system. Compact and lightweight design permit use in wide variety of airborne communications networks.

Cubley, H. D.↗

A message passing kernel for the hypercluster parallel processing test bed

A Message-Passing Kernel (MPK) for the Hypercluster parallel-processing test bed is described. The Hypercluster is being developed at the NASA Lewis Research Center to support investigations of parallel algorithms and architectures for computational fluid and structural mechanics applications. The Hypercluster resembles the hypercube architecture except that each node consists of multiple processors communicating through shared memory. The MPK efficiently routes information through the Hypercluster, using a message-passing protocol when necessary and faster shared-memory communication whenever possible. The MPK also interfaces all of the processors with the Hypercluster operating system (HYCLOPS), which runs on a Front-End Processor (FEP). This approach distributes many of the I/O tasks to the Hypercluster processors and eliminates the need for a separate I/O support program on the FEP.

Blech, Richard A.↗

An information theory of image gathering

Shannon's mathematical theory of communication is extended to image gathering. Expressions are obtained for the total information that is received with a single image-gathering channel and with parallel channels. It is concluded that the aliased signal components carry information even though these components interfere with the within-passband components in conventional image gathering and restoration, thereby degrading the fidelity and visual quality of the restored image. An examination of the expression for minimum mean-square-error, or Wiener-matrix, restoration from parallel image-gathering channels reveals a method for unscrambling the within-passband and aliased signal components to restore spatial frequencies beyond the sampling passband out to the spatial frequency response cutoff of the optical aperture.

Fales, Carl L.↗

Parallel Processing of Adaptive Meshes with Load Balancing

Many scientific applications involve grids that lack a uniform underlying structure. These applications are often also dynamic in nature in that the grid structure significantly changes between successive phases of execution. In parallel computing environments, mesh adaptation of unstructured grids through selective refinement/coarsening has proven to be an effective approach. However, achieving load balance while minimizing interprocessor communication and redistribution costs is a difficult problem. Traditional dynamic load balancers are mostly inadequate because they lack a global view of system loads across processors. In this paper, we propose a novel and general-purpose load balancer that utilizes symmetric broadcast networks (SBN) as the underlying communication topology, and compare its performance with a successful global load balancing environment, called PLUM, specifically created to handle adaptive unstructured applications. Our experimental results on an IBM SP2 demonstrate that the SBN-based load balancer achieves lower redistribution costs than that under PLUM by overlapping processing and data migration.

Das, Sajal K.↗

Laser ablation spectrometry system

This disclosure provides systems, methods, and apparatus related to laser ablation spectrometry systems. In one aspect, a system comprises a microscope, a laser, a continuous flow probe, and a gas confinement device. The laser is positioned to emit light through an objective lens of the microscope. The continuous flow probe is coupled to a spectrometer. An end of the continuous flow probe is positioned proximate a sample and between the sample and the objective lens. The gas confinement device defines a gas inlet, a chamber, a platform, a wall surrounding the platform, a plurality of vents, and a plurality of channels. Each of the plurality of vents is positioned to direct a gas substantially parallel to the platform, and each of the plurality of vents is defined in the wall. The plurality of channels is operable to provide fluid communication between the chamber and the plurality of vents.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Situational Awareness of Grid Anomalies (SAGA)

The modern power industry becomes more vulnerable to cyber events due to the growing interconnectivity, interdependence, and complexity of the electric power grid. High-fidelity modeling and simulation tools that support the preventative risk analysis on potential cyber-relevant events is essential for ensuring the situational awareness of the system operator as it provides an inexpensive and risk-free environment to test the system responses under various cyber-relevant events and hereby can support research on cyber anomaly detection, optimal protective resource allocation, and mitigation measures. In this webinar, we will share NREL's cybersecurity research capabilities by highlighting the development of a scalable cyber-physical event test bed and demonstration with real hardware in the loop. The developed cyber-physical event test bed is backboned by an integrated transmission, distribution, and communication dynamic co-simulation framework and a plug-and-play cyber event generation module. It is designed to be modular and compatible with parallel computing, and thereby supports large-scale system simulations at an affordable computation cost. The test bed can capture millisecond-to-minutes dynamic frequency and voltage responses under cyber events from the bulk transmission system to the active distribution systems and distributed energy resources at the grid edge.

co-simulation↗

Massively parallel, multiple-organ perfusion control system

A fluidic cartridge comprises a fluidic disk having a plurality of alignment openings; a fluidic chip comprising a body, one or more channels formed in the body in fluidic communications with input ports and output ports for transferring one or more fluids between the input ports and the output ports, and a plurality of protrusions formed on the body and received in the alignment openings of the fluidic disk for aligning the fluidic chip to the fluidic disk; an actuator operably engaging with the one or more channels for selectively and individually transferring the one or more fluids through the one or more channels from at least one of the input ports to at least one of the output ports at desired flow rates; and a tube member defining a cylindrical housing for accommodating the fluidic disk, the fluidic chip and the actuator therein.

Reiserer, Ronald S.↗

DSMC analysis in a heterogeneous parallel computing environment

A methodology for implementing parallel DSMC codes in a heterogeneous computing environment is described. The methodology involves the use of a common message-passing software library together with recently developed software that handles the actual interprocessor communications in a standard manner across a variety of computing platforms. Benchmark tests using a simple DSMC model problem were performed on an Intel iPSC/860, a Cray-YMP and a group of Sun workstations. The approach was found to give speedups that scaled linearly with problem size on all the computing platforms tested. This methodology was then incorporated into a production-type DSMC code to allow the simulation of problems that would not otherwise have been practical. The application of this production code to simulations of hypersonic shear flows and shock-lip interactions under near-continuum conditions is described. Synchronous and asynchronous models for implementing parallelism into DSMC simulations are also described and both models are shown to produce the same steady-state result.

Wilmoth, R. G.↗

Parthenon—a performance portable block-structured adaptive mesh refinement framework

On the path to exascale the landscape of computer device architectures and corresponding programming models has become much more diverse. While various low-level performance portable programming models are available, support at the application level lacks behind. To address this issue, we present the performance portable block-structured adaptive mesh refinement (AMR) framework Parthenon, derived from the well-tested and widely used Athena++ astrophysical magnetohydrodynamics code, but generalized to serve as the foundation for a variety of downstream multi-physics codes. Parthenon adopts the Kokkos programming model, and provides various levels of abstractions from multidimensional variables, to packages defining and separating components, to launching of parallel compute kernels. Parthenon allocates all data in device memory to reduce data movement, supports the logical packing of variables and mesh blocks to reduce kernel launch overhead, and employs one-sided, asynchronous MPI calls to reduce communication overhead in multi-node simulations. Using a hydrodynamics miniapp, we demonstrate weak and strong scaling on various architectures including AMD and NVIDIA GPUs, Intel and AMD x86 CPUs, IBM Power9 CPUs, as well as Fujitsu A64FX CPUs. At the largest scale on Frontier (the first TOP500 exascale machine), the miniapp reaches a total of 1.7 × 10 13 zone-cycles/s on 9216 nodes (73,728 logical GPUs) at [Formula: see text] weak scaling parallel efficiency (starting from a single node). In combination with being an open, collaborative project, this makes Parthenon an ideal framework to target exascale simulations in which the downstream developers can focus on their specific application rather than on the complexity of handling massively-parallel, device-accelerated AMR.

97 MATHEMATICS AND COMPUTING↗

Smart Droplets Stabilized by Designer Surfactants: From Biomimicry to Active Motion to Materials Healing

The science and technologies of emulsion droplets have been a long‐term focus of extensive research endeavors for their practical utility across a breadth of industries, including pharmaceutical products, oil recovery processes, and the food sciences. However, with advances in materials chemistry and characterization tools, new emerging areas are arising with a focus on “smart droplets”. The versatility of emulsion droplets across is based on their ability to partition and create isolated systems with properties defined by the liquid–liquid interface, while preparative routes allow manipulation of droplet size, stability, and encapsulated contents. As described in this article, significant efforts are being devoted to creating new types of droplets by “activating” this interface through the incorporation of reactive structures that trigger droplet response to applied or environmental stimuli (e.g., pH, temperature, salt, or external fields). Moreover, parallels between droplets and live cells inspire efforts to conceive systems that resemble biological motifs or that can produce cellular behaviors that imitate biology (e.g., swarming, communication, or motion). Here, the authors highlight recent advances in smart droplets, with emphasis on organic, polymer, and/or particle surfactants that give rise to inter‐droplet communication (via aggregation, fusion, division, or mass transfer), droplet vehicles for controlled delivery, autonomous droplet motion, and tunable emulsion inversion. Especially emphasized is the macromolecular design to produce reactive and functional surfactants, which are crucial to responsive droplet behavior and their underlying mechanisms. More generally, the exquisite interplay between materials science and biology inspires the review of this research area that provides unique opportunities for insight and inspiration into the capabilities of new droplet designs.

36 MATERIALS SCIENCE↗

Refining HPCToolkit for application performance analysis at exascale

As part of the US Department of Energy’s Exascale Computing Project (ECP), Rice University has been refining its HPCToolkit performance tools to better support measurement and analysis of applications executing on exascale supercomputers. To efficiently collect performance measurements of GPU-accelerated applications, HPCToolkit employs novel non-blocking data structures to communicate performance measurements between tool threads and application threads. To attribute performance information in detail to source lines, loop nests, and inlined call chains, HPCToolkit performs parallel analysis of large CPU and GPU binaries involved in the execution of an exascale application to rapidly recover mappings between machine instructions and source code. To analyze terabytes of performance measurements gathered during executions at exascale, HPCToolkit employs distributed-memory parallelism, multithreading, sparse data structures, and out-of-core streaming analysis algorithms. To support interactive exploration of profiles up to terabytes in size, HPCToolkit’s hpcviewer graphical user interface uses out-of-core methods to visualize performance data. The result of these efforts is that HPCToolkit now supports collection, analysis, and presentation of profiles and traces of GPU-accelerated applications at exascale. These improvements have enabled HPCToolkit to efficiently measure, analyze and explore terabytes of performance data for executions using as many as 64K MPI ranks and 64K GPU tiles on ORNL’s Frontier supercomputer. HPCToolkit’s support for measurement and analysis of GPU-accelerated applications has been employed to study a collection of open-science applications developed as part of ECP. This paper reports on these experiences, which provided insight into opportunities for tuning applications, strengths and weaknesses of HPCToolkit itself, as well as unexpected behaviors in executions at exascale.

Adhianto, Laksono↗

Performance of battery reconditioning on the communications technology satellite

Some real life data on the flight performance and reconditioning on the Communications Technology Satellite are presented. There are two 24-cell nominal 5 amp-hour nickel cadmium batteries on board. The two batteries are operated in parallel with isolating diodes in between and permanently connected to the housekeeping buses. The recharge can be at C over 10 or C over 20 to a 1.4 charge/discharge ratio; and recharge is terminated either when a computed 1.4 charge/discharge ratio is reached, or a charge voltage peak is reached.

Lackner, J.↗

Sonic levitation apparatus

A sonic levitation apparatus is disclosed which includes a sonic transducer which generates acoustical energy responsive to the level of an electrical amplifier. A duct communicates with an acoustical chamber to deliver an oscillatory motion of air to a plenum section which contains a collimated hole structure having a plurality of parallel orifices. The collimated hole structure converts the motion of the air to a pulsed. Unidirectional stream providing enough force to levitate a material specimen. Particular application to the production of microballoons in low gravity environment is discussed.

Dunn, S. A.↗