Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallelization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,729 records · Page 96

Isolated and interacting round parallel heated jets

An experimental study of the flowfield of heated and unheated single and dual jet configurations was performed. This study of two parallel jets is unique since most previous aerodynamic structure experiments were limited to single round and two-dimensional jets. The present closely spaced dual jet geometry was motivated by the potential jet noise reduction available from this configuration. This geometry has shown promise as a method for redirecting jet noise away from ground based observers in side by side or over/under turbofan engine mountings on aircraft (Simonich et al., 1984). Since the effectiveness of this noise reduction technique is based on the existence of two independent jets, an understanding of the aerodynamics of the merging process is essential to establishing the acoustic benefits. The experimental program was structured so that Mach number, jet exit temperature, and spacing to diameter ratios could be independently varied to isolate each effect.

Simonich, J. C.↗

An experimental investigation of the parallel blade-vortex interaction

A scheme for investigating the parallel blade vortex interaction (BVI) has been designed and tested. The scheme involves setting a vortex generator upstream of a nonlifting rotor so that the vortex interacts with the blade at the forward azimuth. The method has revealed two propagation mechanisms: a type C shock propagation from the leading edge induced by the vortex at high tip speeds, and a rapid but continuous pressure pulse associated with the proximity of the vortex to the leading edge. The latter is thought to be the more important source. The effects of Mach number and vortex proximity are discussed.

Caradonna, F. X.↗

Solution of partial differential equations on vector and parallel computers

The present status of numerical methods for partial differential equations on vector and parallel computers was reviewed. The relevant aspects of these computers are discussed and a brief review of their development is included, with particular attention paid to those characteristics that influence algorithm selection. Both direct and iterative methods are given for elliptic equations as well as explicit and implicit methods for initial boundary value problems. The intent is to point out attractive methods as well as areas where this class of computer architecture cannot be fully utilized because of either hardware restrictions or the lack of adequate algorithms. Application areas utilizing these computers are briefly discussed.

Ortega, J. M.↗

Radiative transfer in a sphere illuminated by a parallel beam - An integral equation approach

The problem of multiple scattering of nonpolarized light in a planetary body of arbitrary shape illuminated by a parallel beam is formulated using the integral equation approach. There exists a simple functional whose stationarity condition is equivalent to solving the equation of radiative transfer and whose value at the stationary point is proportional to the differential cross section. The analysis reveals a direct relation between the microscopic symmetry of the phase function for each scattering event and the macroscopic symmetry of the differential cross section for the entire planetary body, and the interconnection of these symmetry relations and the variational principle. The case of a homogeneous sphere containing isotropic scatterers is investigated in detail. It is shown that the solution can be expanded in a multipole series such that the general spherical problem is reduced to solving a set of decoupled integral equations in one dimension. Computations have been performed for a range of parameters of interest, and illustrative examples of applications to planetary problems as provided.

Shia, R.-L.↗

Extended testing of a general contextual classifier using the massively parallel processor - Preliminary results and test plans

Earlier encouraging test results of a contextual classifier that combines spatial and spectral information employing a general statistical approach are expanded. The earlier results were of limited meaning because they were produced from small (50-by-50 pixel) data sets. An implementation of the contextual classifier on NASA Goddard's Massively Parallel Processor (MPP) is presented; for the first time the MPP makes feasible the testing of the classifier on large data sets (a 12-hour test on a VAX-11/780 minicomputer now takes 5 minutes on the MPP). The MPP is a Single-Instruction, Multiple Data Stream computer, consisting of 16,384 bit serial microprocessors connected in a 128-by-128 mesh array with each element having data transfer connections with its four nearest neighbors so that the MPP is capable of billions of operations per second. Preliminary results are given (with more expected for the conference) and plans are mentioned for extended testing of the contextual classifier on Thematic Mapper data sets.

Tilton, J. C.↗

The motion of ions specularly reflected off a quasi-parallel shock in the presence of large-amplitude, monochromatic MHD waves

A model is used to examine the motion of specularly reflected ions in the presence of large-amplitude, monochromatic, transverse MHD waves. The calculations of ion trajectories are described. The heating downstream from the quasi-parallel bow shock is analyzed. The relationship between the specularly reflected ions and their gyrospeeds and guiding center speeds is studied. The data reveal that the characteristics of the motion depend on the frequency, wavelength, phase, and amplitude of the wave that is converted into the shock.

Fuselier, S. A.↗

State-plane analysis of parallel resonant converter

A method for analyzing the complex operation of a parallel resonant converter is developed, utilizing graphical state-plane techniques. The comprehensive mode analysis uncovers, for the first time, the presence of other complex modes besides the continuous conduction mode and the discontinuous conduction mode and determines their theoretical boundaries. Based on the insight gained from the analysis, a novel, high-frequency resonant buck converter is proposed. The voltage conversion ratio of the new converter is almost independent of load.

Oruganti, R.↗

Transition to unstable ion flow in parallel electric fields

The stability of ionospheric O(+)-H(+) outflows accelerated by a nonambipolar parallel electric field is considered under conditions where the ion motion initially develops adiabatically and the ambient plasma is vertically stratified with an effective temperature that increases with altitude. Such conditions are expected near the bottom of the auroral acceleration region where ion and electron streaming instabilities first develop. It is shown for a particular equilibrium profile that the differentially accelerated ion flows become unstable within about 100 km from their entry point in the acceleration region. At O(+)/H(+) density ratios less than about 9, the instability is dominated by a violent H(+)-O(+) two-stream interaction which couples the O(+) and H(+) acoustic modes, and which mediates a transition to nonadiabatic acceleration. At higher altitudes and/or larger O(+)/H(+) density ratios, a much weaker resonant instability exists, which is driven by the relative drift between electrons and O(+) or H(+) ions. The results suggest that the H(+)-O(+) two-stream instability may be a viable mechanism for heating upflowing auroral ions.

Bergmann, R.↗

Optical Interferometric Parallel Data Processor

Image data processed faster than in present electronic systems. Optical parallel-processing system effectively calculates two-dimensional Fourier transforms in time required by light to travel from plane 1 to plane 8. Coherence interferometer at plane 4 splits light into parts that form double image at plane 6 if projection screen placed there.

Breckinridge, J. B.↗

Parallel Analog-to-Digital Image Processor

Proposed integrated-circuit network of many identical units convert analog outputs of imaging arrays of x-ray or infrared detectors to digital outputs. Converter located near imaging detectors, within cryogenic detector package. Because converter output digital, lends itself well to multiplexing and to postprocessing for correction of gain and offset errors peculiar to each picture element and its sampling and conversion circuits. Analog-to-digital image processor is massively parallel system for processing data from array of photodetectors. System built as compact integrated circuit located near local plane. Buffer amplifier for each picture element has different offset.

Lokerson, D. C.↗

Dynamic remapping decisions in multi-phase parallel computations

The effectiveness of any given mapping of workload to processors in a parallel system is dependent on the stochastic behavior of the workload. Program behavior is often characterized by a sequence of phases, with phase changes occurring unpredictably. During a phase, the behavior is fairly stable, but may become quite different during the next phase. Thus a workload assignment generated for one phase may hinder performance during the next phase. We consider the problem of deciding whether to remap a paralled computation in the face of uncertainty in remapping's utility. Fundamentally, it is necessary to balance the expected remapping performance gain against the delay cost of remapping. This paper treats this problem formally by constructing a probabilistic model of a computation with at most two phases. We use stochastic dynamic programming to show that the remapping decision policy which minimizes the expected running time of the computation has an extremely simple structure: the optimal decision at any step is followed by comparing the probability of remapping gain against a threshold. This theoretical result stresses the importance of detecting a phase change, and assessing the possibility of gain from remapping. We also empirically study the sensitivity of optimal performance to imprecise decision threshold. Under a wide range of model parameter values, we find nearly optimal performance if remapping is chosen simply when the gain probability is high. These results strongly suggest that except in extreme cases, the remapping decision problem is essentially that of dynamically determining whether gain can be achieved by remapping after a phase change; precise quantification of the decision model parameters is not necessary.

Nicol, D. M.↗

Low phase noise oscillator using two parallel connected amplifiers

A high frequency oscillator is provided by connecting two amplifier circuits in parallel where each amplifier circuit provides the other amplifier circuit with the conditions necessary for oscillation. The inherent noise present in both amplifier circuits causes the quiescent current, and in turn, the generated frequency, to change. The changes in quiescent current cause the transconductance and the load impedance of each amplifier circuit to vary, and this in turn results in opposing changes in the input susceptance of each amplifier circuit. Because the changes in input susceptance oppose each other, the changes in quiescent current also oppose each other. The net result is that frequency stability is enhanced.

Kleinberg, Leonard L.↗

Exploiting loop level parallelism in nonprocedural dataflow programs

Discussed are how loop level parallelism is detected in a nonprocedural dataflow program, and how a procedural program with concurrent loops is scheduled. Also discussed is a program restructuring technique which may be applied to recursive equations so that concurrent loops may be generated for a seemingly iterative computation. A compiler which generates C code for the language described below has been implemented. The scheduling component of the compiler and the restructuring transformation are described.

Gokhale, Maya B.↗

Problem size, parallel architecture and optimal speedup

The communication and synchronization overhead inherent in parallel processing can lead to situations where adding processors to the solution method actually increases execution time. Problem type, problem size, and architecture type all affect the optimal number of processors to employ. The numerical solution of an elliptic partial differential equation is examined in order to study the relationship between problem size and architecture. The equation's domain is discretized into n sup 2 grid points which are divided into partitions and mapped onto the individual processor memories. The relationships between grid size, stencil type, partitioning strategy, processor execution time, and communication network type are analytically quantified. In so doing, the optimal number of processors was determined to assign to the solution, and identified (1) the smallest grid size which fully benefits from using all available processors, (2) the leverage on performance given by increasing processor speed or communication network speed, and (3) the suitability of various architectures for large numerical problems.

Nicol, David M.↗

Solving the Cauchy-Riemann equations on parallel computers

Discussed is the implementation of a single algorithm on three parallel-vector computers. The algorithm is a relaxation scheme for the solution of the Cauchy-Riemann equations; a set of coupled first order partial differential equations. The computers were chosen so as to encompass a variety of architectures. They are: the MPP, and SIMD machine with 16K bit serial processors; FLEX/32, an MIMD machine with 20 processors; and CRAY/2, an MIMD machine with four vector processors. The machine architectures are briefly described. The implementation of the algorithm is discussed in relation to these architectures and measures of the performance on each machine are given. Simple performance models are used to describe the performance. These models highlight the bottlenecks and limiting factors for this algorithm on these architectures. Conclusions are presented.

Fatoohi, Raad A.↗

Phase space simulation of collisionless stellar systems on the massively parallel processor

A numerical technique for solving the collisionless Boltzmann equation describing the time evolution of a self gravitating fluid in phase space was implemented on the Massively Parallel Processor (MPP). The code performs calculations for a two dimensional phase space grid (with one space and one velocity dimension). Some results from calculations are presented. The execution speed of the code is comparable to the speed of a single processor of a Cray-XMP. Advantages and disadvantages of the MPP architecture for this type of problem are discussed. The nearest neighbor connectivity of the MPP array does not pose a significant obstacle. Future MPP-like machines should have much more local memory and easier access to staging memory and disks in order to be effective for this type of problem.

White, Richard L.↗

Block iterative restoration of astronomical images with the massively parallel processor

A method is described for algebraic image restoration capable of treating astronomical images. For a typical 500 x 500 image, direct algebraic restoration would require the solution of a 250,000 x 250,000 linear system. The block iterative approach is used to reduce the problem to solving 4900 121 x 121 linear systems. The algorithm was implemented on the Goddard Massively Parallel Processor, which can solve a 121 x 121 system in approximately 0.06 seconds. Examples are shown of the results for various astronomical images.

Heap, Sara R.↗

Chemical network problems solved on NASA/Goddard's massively parallel processor computer

The single instruction stream, multiple data stream Massively Parallel Processor (MPP) unit consists of 16,384 bit serial arithmetic processors configured as a 128 x 128 array whose speed can exceed that of current supercomputers (Cyber 205). The applicability of the MPP for solving reaction network problems is presented and discussed, including the mapping of the calculation to the architecture, and CPU timing comparisons.

Cho, Seog Y.↗