Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 739 records · Page 41

Instrumentation, performance visualization, and debugging tools for multiprocessors

The need for computing power has forced a migration from serial computation on a single processor to parallel processing on multiprocessor architectures. However, without effective means to monitor (and visualize) program execution, debugging, and tuning parallel programs becomes intractably difficult as program complexity increases with the number of processors. Research on performance evaluation tools for multiprocessors is being carried out at ARC. Besides investigating new techniques for instrumenting, monitoring, and presenting the state of parallel program execution in a coherent and user-friendly manner, prototypes of software tools are being incorporated into the run-time environments of various hardware testbeds to evaluate their impact on user productivity. Our current tool set, the Ames Instrumentation Systems (AIMS), incorporates features from various software systems developed in academia and industry. The execution of FORTRAN programs on the Intel iPSC/860 can be automatically instrumented and monitored. Performance data collected in this manner can be displayed graphically on workstations supporting X-Windows. We have successfully compared various parallel algorithms for computational fluid dynamics (CFD) applications in collaboration with scientists from the Numerical Aerodynamic Simulation Systems Division. By performing these comparisons, we show that performance monitors and debuggers such as AIMS are practical and can illuminate the complex dynamics that occur within parallel programs.

Yan, Jerry C.↗

Interaction of Aircraft Wakes From Laterally Spaced Aircraft

Large Eddy Simulations are used to examine wake interactions from aircraft on closely spaced parallel paths. Two sets of experiments are conducted, with the first set examining wake interactions out of ground effect (OGE) and the second set for in ground effect (IGE). The initial wake field for each aircraft represents a rolled-up wake vortex pair generated by a B-747. Parametric sets include wake interactions from aircraft pairs with lateral separations of 400, 500, 600, and 750 ft. The simulation of a wake from a single aircraft is used as baseline. The study shows that wake vortices from either a pair or a formation of B-747 s that fly with very close lateral spacing, last longer than those from an isolated B-747. For OGE, the inner vortices between the pair of aircraft, ascend, link and quickly dissipate, leaving the outer vortices to decay and descend slowly. For the IGE scenario, the inner vortices ascend and last longer, while the outer vortices decay from ground interaction at a rate similar to that expected from an isolated aircraft. Both OGE and IGE scenarios produce longer-lasting wakes for aircraft with separations less than 600 ft. The results are significant because concepts to increase airport capacity have been proposed that assume either aircraft formations and/or aircraft pairs landing on very closely spaced runways.

Proctor, Fred H.↗

Coupling Carbon Oxidation and Surface Recession in Direct-Simulation Monte Carlo Code, SPARTA

Ablative thermal protection system (TPS) materials for spacecraft are composites that are often made out of carbon-based reinforcement and a polymeric matrix. They endure high-temperature oxidation and surface recession when re-entering Earth’s atmosphere. Ablation is the result of many coupled and competing thermal, mechanical, and chemical phenomena, and it is difficult to isolate the role of each on the overall degradation of the TPS. Here we develop an ablation model for material recession coupled explicitly to finite rate carbon oxidation in complex microstructures. In this work, Stochastic PArallel Rarified-gas Time-accurate Analyzer (SPARTA), a direct-simulation Monte Carlo (DSMC) code, is modified to allow oxidation-driven ablation of implicitly defined carbon surfaces. In SPARTA, implicit surfaces are generated from the grid corner point values via a marching cubes algorithm, therefore creating a new set of surface elements every time ablation is performed. The finite-rate oxidation model developed by Gopalan et. al, was adapted to tally surface reactions and other surface data on a per-grid cell basis. The ablation functionality was also adjusted so once the reactions have occurred, the number of reactions leading to CO formation can be converted to corner point reduction values; therefore, carbon removal is directly proportional to surface recession. We also develop robust algorithms which handle the evolution of the flow cells and solid material regions, including split cells (flow cell divided in two by a solid surface). Finally, we demonstrate our implicit chemistry model for 2D and 3D geometries by producing reaction statistics and detailed visualization of oxidation-induced material recession at the microscale.

V Arias↗

A Simulation Testbed for Airborne Merging and Spacing

The key innovation in this effort is the development of a simulation testbed for airborne merging and spacing (AM&S). We focus on concepts related to airports with Super Dense Operations where new airport runway configurations (e.g. parallel runways), sequencing, merging, and spacing are some of the concepts considered. We focus on modeling and simulating a complementary airborne and ground system for AM&S to increase efficiency and capacity of these high density terminal areas. From a ground systems perspective, a scheduling decision support tool generates arrival sequences and spacing requirements that are fed to the AM&S system operating on the flight deck. We enhanced NASA's Airspace Concept Evaluation Systems (ACES) software to model and simulate AM&S concepts and algorithms.

Santos, Michel↗

Observations of Closed Structures at the Magnetopause: A Case for Multiple Reconnections

We further analyze a case of Interball LLBL crossing on the dusk flank of geomagnetosphere under southward magnetosheath magnetic field, previously categorized as an interval of highly structured LLBL. These conditions of highly structured LLBL include reconnection signatures. Observed ion velocity distributions with LLBL are quite variable. D-shaped distributions that are associated with the open reconnected flux tube are observed at the boundaries of LLBL transients and sometimes within the LLBL transients. In most cases the ion velocity distributions consist of two magnetosheath-type components with different velocities parallel to the magnetic field, or of three components one of which has nearly zero Vpar. The shapes of ion velocity distributions and their evolution with decreasing number density in LLBL indicate that most of LLBL is located on closed magnetic field lines. These observations strongly favor multiple reconnections between magnetosheath and magnetosphereric flux tubes, creating long spiral flux tube islands at the magnetopause. We report evidence for the simultaneous occurrence of magnetic reconnection at multiple points across the magnetopause, as has been proposed and found to occur in magnetopause simulations. The evidence is in the form of highly structured distributions of ions in velocity parallel to the local magnetic field direction, within the magnetopause and low latitude boundary layer region, from the Interball-Tall spacecraft. We interpret these distributions as a natural consequence of the formation of spiral magnetic flux tubes consisting of a mixture of alternating segments originating from the magnetosheath or interplanetary plasma and from the low latitude boundary layer or magnetospheric plasma. We further analyze a case of Interball LLBL crossing on the dusk flank of geomagnetosphere under southward magnetosheath magnetic field, previously categorized as an interval of highly structured LLBL. These conditions of highly structured LLBL include reconnection signatures. Observed ion velocity distributions with LLBL are quite variable. D-shaped distributions that are associated with the open reconnected flux tube are observed at the boundaries of LLBL transients and sometimes within the LLBL transients. In most cases the ion velocity distributions consist of two magnetosheath-type components with different velocities parallel to the magnetic field, or of three components one of which has nearly zero Vpar. The shapes of ion velocity distributions and their evolution with decreasing number density in LLBL indicate that most of LLBL is located on closed magnetic field lines. These observations strongly favor multiple reconnection between magnetosheath and magnetospheric flux tubes, creating long spiral flux tube islands at the magnetopause. We report evidence for the simultaneous occurrence of magnetic reconnection at multiple points across the magnetopause, as has been proposed and found to occur in magnetopause simulations. The evidence is in the form of highly structured distributions of ions in velocity parallel to the local magnetic field direction, within the magnetopause and low latitude boundary layer region, from the Interball-Tail spacecraft. We interpret these distributions as a natural consequence of the formation of spiral. We further analyze a case of Interball LLBL crossing on the dusk flank of geomagnetosphere under southward magnetosheath magnetic field, previously categorized as an interval of highly structured LLBL. These conditions of highly structured LLBL include reconnection signatures. Observed ion velocity distributions with LLBL are quite variable. D-shaped distributions that are associated with the open reconnected flux tube are observed at the boundaries of LLBL transients and sometimes within the LLBL transients. In most cases the ion velocity distributions consist of two magnetosheath-type components with different velocities parallel to the magnetic field, or of three components one of which has nearly zero Vpar.

Vaisberg, O. L.↗

Terminal Area Procedures for Paired Runways

Parallel runway operations have been found to increase capacity within the National Airspace but poor visibility conditions reduce the use of these operations. The NextGen and SESAR Programs have identified the capacity benefits from increased use of closely-space parallel runway. Previous research examined the concepts and procedures related to parallel runways however, there has been no investigation of the procedures associated with the strategic and tactical pairing of aircraft for these operations. This simulation study developed and examined the pilot and controller procedures and information requirements for creating aircraft pairs for parallel runway operations. The goal was to achieve aircraft pairing with a temporal separation of 15s (+/- 10s error) at a coupling point that was about 12 nmi from the runway threshold. Two variables were explored for the pilot participants: two levels of flight deck automation (current-day flight deck automation and auto speed control future automation) as well as two flight deck displays that assisted in pilot conformance monitoring. The controllers were also provided with automation to help create and maintain aircraft pairs. Results show the operations in this study were acceptable and safe. Subjective workload, when using the pairing procedures and tools, was generally low for both controllers and pilots, and situation awareness was typically moderate to high. Pilot workload was influenced by display type and automation condition. Further research on pairing and off-nominal conditions is required however, this investigation identified promising findings about the feasibility of closely-spaced parallel runway operations.

Lozito, Sandra↗

Onset of Reconnection in the near Magnetotail: PIC Simulations

Using 2.5-dimensional particle-in-cell (PIC) simulations of magnetotail dynamics, we investigate the onset of reconnection in two-dimensional tail configurations with finite Bz. Reconnection onset is preceded by a driven phase, during which magnetic flux is added to the tail at the high-latitude boundaries, followed by a relaxation phase, during which the configuration continues to respond to the driving. We found a clear distinction between stable and unstable cases, dependent on deformation amplitude and ion/electron mass ratio. The threshold appears consistent with electron tearing. The evolution prior to onset, as well as the evolution of stable cases, are largely independent of the mass ratio, governed by integral flux tube entropy conservation as imposed in MHD (magnetohydrodynamics). This suggests that ballooning instability in the tail should not be expected prior to the onset of tearing and reconnection. The onset time and other onset properties depend on the mass ratio, consistent with expectations for electron tearing. At onset,we found electron anisotropies T⊥⁄ T∥ (bottom tail divided by parallel tail) equals 1.1-1.3, raising growth rates and wavenumbers. Our simulations have provided a quantitative onset criterion that is easily evaluated in MHD simulations, provided the spatial resolution is sufficient. The evolution prior to onset and after the formation of a neutral line does not depend on the electron physics, which should permit an approximation by MHD simulations with appropriate dissipation terms.

PIC↗

Parallelization of Lower-Upper Symmetric Gauss-Seidel Method for Chemically Reacting Flow

Development of technologies for exploration of the solar system has revived an interest in computational simulation of chemically reacting flows since planetary probe vehicles exhibit non-equilibrium phenomena during the atmospheric entry of a planet or a moon as well as the reentry to the Earth. Stability in combustion is essential for new propulsion systems. Numerical solution of real-gas flows often increases computational work by an order-of-magnitude compared to perfect gas flow partly because of the increased complexity of equations to solve. Recently, as part of Project Columbia, NASA has integrated a cluster of interconnected SGI Altix systems to provide a ten-fold increase in current supercomputing capacity that includes an SGI Origin system. Both the new and existing machines are based on cache coherent non-uniform memory access architecture. Lower-Upper Symmetric Gauss-Seidel (LU-SGS) relaxation method has been implemented into both perfect and real gas flow codes including Real-Gas Aerodynamic Simulator (RGAS). However, the vectorized RGAS code runs inefficiently on cache-based shared-memory machines such as SGI system. Parallelization of a Gauss-Seidel method is nontrivial due to its sequential nature. The LU-SGS method has been vectorized on an oblique plane in INS3D-LU code that has been one of the base codes for NAS Parallel benchmarks. The oblique plane has been called a hyperplane by computer scientists. It is straightforward to parallelize a Gauss-Seidel method by partitioning the hyperplanes once they are formed. Another way of parallelization is to schedule processors like a pipeline using software. Both hyperplane and pipeline methods have been implemented using openMP directives. The present paper reports the performance of the parallelized RGAS code on SGI Origin and Altix systems.

Yoon, Seokkwan↗

Efficiently modeling neural networks on massively parallel computers

Neural networks are a very useful tool for analyzing and modeling complex real world systems. Applying neural network simulations to real world problems generally involves large amounts of data and massive amounts of computation. To efficiently handle the computational requirements of large problems, we have implemented at Los Alamos a highly efficient neural network compiler for serial computers, vector computers, vector parallel computers, and fine grain SIMD computers such as the CM-2 connection machine. This paper describes the mapping used by the compiler to implement feed-forward backpropagation neural networks for a SIMD (Single Instruction Multiple Data) architecture parallel computer. Thinking Machines Corporation has benchmarked our code at 1.3 billion interconnects per second (approximately 3 gigaflops) on a 64,000 processor CM-2 connection machine (Singer 1990). This mapping is applicable to other SIMD computers and can be implemented on MIMD computers such as the CM-5 connection machine. Our mapping has virtually no communications overhead with the exception of the communications required for a global summation across the processors (which has a sub-linear runtime growth on the order of O(log(number of processors)). We can efficiently model very large neural networks which have many neurons and interconnects and our mapping can extend to arbitrarily large networks (within memory limitations) by merging the memory space of separate processors with fast adjacent processor interprocessor communications. This paper will consider the simulation of only feed forward neural network although this method is extendable to recurrent networks.

Farber, Robert M.↗

Improving Fidelity of Launch Vehicle Liftoff Acoustic Simulations

Launch vehicles experience high acoustic loads during ignition and liftoff affected by the interaction of rocket plume generated acoustic waves with launch pad structures. Application of highly parallelized Computational Fluid Dynamics (CFD) analysis tools optimized for application on the NAS computer systems such as the Loci/CHEM program now enable simulation of time-accurate, turbulent, multi-species plume formation and interaction with launch pad geometry and capture the generation of acoustic noise at the source regions in the plume shear layers and impingement regions. These CFD solvers are robust in capturing the acoustic fluctuations, but they are too dissipative to accurately resolve the propagation of the acoustic waves throughout the launch environment domain along the vehicle. A hybrid Computational Fluid Dynamics and Computational Aero-Acoustics (CFD/CAA) modeling framework has been developed to improve such liftoff acoustic environment predictions. The framework combines the existing highly-scalable NASA production CFD code, Loci/CHEM, with a high-order accurate discontinuous Galerkin (DG) solver, Loci/THRUST, developed in the same computational framework. Loci/THRUST employs a low dissipation, high-order, unstructured DG method to accurately propagate acoustic waves away from the source regions across large distances. The DG solver is currently capable of solving up to 4th order solutions for non-linear, conservative acoustic field propagation. Higher order boundary conditions are implemented to accurately model the reflection and refraction of acoustic waves on launch pad components. The DG solver accepts generalized unstructured meshes, enabling efficient application of common mesh generation tools for CHEM and THRUST simulations. The DG solution is coupled with the CFD solution at interface boundaries placed near the CFD acoustic source regions. Both simulations are executed simultaneously with coordinated boundary condition data exchange.

Liever, Peter↗

Power combining in an array of microwave power rectifiers

This work analyzes the resultant efficiency degradation when identical rectifiers operate at different RF power levels as caused by the power beam taper. Both a closed-form analytical circuit model and a detailed computer-simulation model are used to obtain the output dc load line of the rectifier. The efficiency degradation is nearly identical with series and parallel combining, and the closed-form analytical model provides results which are similar to the detailed computer-simulation model.

Gutmann, R. J.↗

Weak, quasiparallel profiles of earth's bow shock - A comparison between numerical simulations and ISEE 3 observations on the far flank

Over 200 crossings of the distant downwind flanks of earth's magnetosonic bow shock by ISEE 3 included many cases of weak, or low Mach number, quasi-parallel shocks. A consistent feature of the magnetic field profiles was the presence of large amplitude, near periodic to irregular transverse oscillations downstream from even the weakest Q-parallel shocks. Large downstream perturbations with whistler-like features similar to those of the observations appear in 1D simulations when the Alfven Mach number M(A) is greater than 2.5 but not when M(A) = 2.1. The observed cases with downstream waves also occurred when M(A) is greater than about 2.5, suggesting the importance of the Alfven as opposed to magnetosonic Mach number in determining the signature of weak, Q-parallel shocks.

Greenstadt, E. W.↗

Performance evaluation of a simulated data-flow computer with low-resolution actors

Basic problems related to the exploitation of parallelism in a program include sequencing of the instructions and communication of the data. It is pointed out that the data-flow approach offers an elegant solution to the sequencing problem, since all data dependencies are automatically handled and only instructions with ready input sets are activated. It is shown that a change in the level of subcomputations (actors) affects communications costs. The concept of variable resolution is discussed, and the testbed environment is examined. Attention is given to the architecture of the processing elements, the communication network, and the simulators. A description of the analytical model is also provided. Simulation and results are discussed, taking into account test programs and allocation, the variation of the number of processing elements, the variation of the resolution in directed acyclic graphs, performance in processing loops, and array handling.

Gaudiot, J. L.↗

The quasiperpendicular environment of large magnetic pulses in Earth's quasiparallel foreshock - ISEE 1 and 2 observations

ULF waves in Earth's foreshock cause the instantaneous angle theta-B(n) between the upstream magnetic field and the shock normal to deviate from its average value. Close to the quasi-parallel (Q-parallel) shock, the transverse components of the waves become so large that the orientation of the field to the normal becomes quasi-perpendicular (Q-perpendicular) during applicable phases of each wave cycle. Large upstream pulses of B were observed completely enclosed in excursions of Theta-B(n) into the Q-perpendicular range. A recent numerical simulation included Theta-B(n) among the parameters examined in Q-parallel runs, and described a similar coincidence as intrinsic to a stage in development of the reformation process of such shocks. Thus, the natural environment of the Q-perpendicular section of Earth's bow shock seems to include an identifiable class of enlarged magnetic pulses for which local Q-perpendicular geometry is a necessary association.

Greenstadt, E. W.↗

Turbomachinery Flows Modeled

Last year, researchers at the NASA Lewis Research Center used the average passage code APNASA to complete the largest three-dimensional simulation of a multistage axial flow compressor to date. Consisting of 29 blade rows, the configuration is typical of those found in aeroengines today. The simulation, which was executed on the High Performance Computing and Communications (HPCC) Program IBM SP2 parallel computer located at the NASA Ames Research Center, took nearly 90 hr to complete. Since the completion of this activity, a fine-grain, parallel version of APNASA has been written by a team of researchers from General Electric, NASA Lewis, and NYMA. Timing studies performed on the SP2 have shown that, with eight processors assigned to each blade row, the simulation time is reduced by a factor of six. For this configuration, the simulation time would be 15 hr. The reduction in computing time indicates that an overnight turnaround of a multistage configuration simulation is feasible. In addition, average passage forms of two-equation turbulence models were formulated. These models are currently being incorporated into APNASA.

Adamczyk, John J.↗

Extensible Adaptable Simulation Systems: Supporting Multiple Fidelity Simulations in a Common Environment

Common practice in the development of simulation systems is meeting all user requirements within a single instantiation. The Joint Polar Satellite System (JPSS) presents a unique challenge to establish a simulation environment that meets the needs of a diverse user community while also spanning a multi-mission environment over decades of operation. In response, the JPSS Flight Vehicle Test Suite (FVTS) is architected with an extensible infrastructure that supports the operation of multiple observatory simulations for a single mission and multiple mission within a common system perimeter. For the JPSS-1 satellite, multiple fidelity flight observatory simulations are necessary to support the distinct user communities consisting of the Common Ground System development team, the Common Ground System Integration & Test team, and the Mission Rehearsal Team/Mission Operations Team. These key requirements present several challenges to FVTS development. First, the FVTS must ensure all critical user requirements are satisfied by at least one fidelity instance of the observatory simulation. Second, the FVTS must allow for tailoring of the system instances to function in diverse operational environments from the High-security operations environment at NOAA Satellite Operations Facility (NSOF) to the ground system factory floor. Finally, the FVTS must provide the ability to execute sustaining engineering activities on a subset of the system without impacting system availability to parallel users. The FVTS approach of allowing for multiple fidelity copies of observatory simulations represents a unique concept in simulator capability development and corresponds to the JPSS Ground System goals of establishing a capability that is flexible, extensible, and adaptable.

McLaughlin, Brian J.↗

A collision-selection rule for a particle simulation method suited to vector computers

A theory is developed for a selection rule governing collisions in a particle simulation of rarefied gas-dynamic flows. The selection rule leads to an algorithmic form highly compatible with fine grain parallel decomposition, allowing for efficient utilization of supercomputers having vector or massively parallel single instruction multiple data architectures. A comparison of shock-wave profiles obtained using both the selection rule and Bird's direct simulation Monte Carlo (DSMC) method show excellent agreement. The equation on which the selection rule is based is shown to be directly related to the time-counter procedure in the DSMC method. The results of several example simulations of representative rarefied flows are presented, for which the number of particles used ranged from 10 to the 6th to 10 to the 7th demonstrating the greatly improved computational efficiency of the method.

Baganoff, D.↗

Auroral plasma transport processes in the presence of kV potential structures

We have simulated plasma transport processes in the presence of a quasi-two-dimensional current filament, that generated kV potential structure in the auroral region. The simulation consists of a set of one-dimensional flux tube simulations with different imposed time-dependent, field-aligned currents. The model uses the 16 moment system of equations and simultaneously solves coupled continuity and momentum equations and equations describing the transport along the magnetic field lines of parallel and perpendicular thermal energy and heat flows for each species. The lower end of the simulation is at an altitude of 800 km, in the collisional topside ionosphere, while the upper end is at 10 R(sub E) in the magnetosphere. The plasma consists of hot electrons and protons of magnetospheric origin and low-energy electrons, protons, and oxygen ions of ionospheric origin. The dynamical interaction of the individual current filaments with ionospheric and magnetospheric plasma generates a potential structure in the horizontal direction and kilovolt field-aligned potential drops along the field lines. The side-by-side display exhibits the evolution of the implied potential structure in the horizontial direction. In the presence of this potential structure and parallel electric field ionospheric plasma density is depleted and velocity is reduced, while density enhancement and increased velocity is observed in magnetospheric plasma. The ionospheric and magnetospheric electron temperatures increase below 2 R(sub E) due to magnetic mirror force on converging geomagnetic field lines. The primary cross-field motion produced by the horizontal E field (E x B drift) is perpendicular to both of the significant spatial directions and is thus ignorable in this geometry. The effects of other cross-field drift processes are discussed. The simulation thus provides insight into the dynamical evolution of two-dimensional potential structures driven by an imposed finite width, field-aligned current profile.

Ganguli, Supriya B.↗