Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 775 records · Page 43

Programming approaches for scalability, performance, and portability of combustion physics codes

Here, this paper presents the process, strategy, and results associated with porting a typical combustion physics flow solver to current state-of-the-art and future massively-parallel computer architectures. Major focus is placed on the distinct algorithmic structure of these types of codes and how it can be integrated with modern programming paradigms for heterogeneous platforms (i.e., distributed many-core systems with accelerators). An end-to-end case study is presented that exemplifies the process in a generic manner, which then serves as a clear guide with respect to the strategy and best practices leading to a robust and adaptable framework that performs well, is durable over time, is portable, and requires minimal human-effort. This end is accomplished beginning with the use of a mature, validated, structured, multiblock code framework optimized for application of both Large Eddy Simulation (LES) and Direct Numerical Simulation (DNS). This code has been ported to a variety of platforms over the past decade, including most recently the Oak Ridge Leadership Computing Facility’s “Summit” Platform. The experience gained on these multiple platforms provides general insights and thus the results presented are not specific to any one code or platform other than the overarching trend toward distributed many-core systems with accelerators in order to move toward exascale performance. The resultant performance and scalability of the ported code is demonstrated on a real-world application; a state-of-the-art rotating detonation rocket engine simulation that matches the complex geometry and boundary conditions imposed as part of a companion experimental campaign.

97 MATHEMATICS AND COMPUTING↗

Self-consistent formation of parallel electric fields in the auroral zone

This paper presents results from a fully self-consistent kinetic particle simulation of the time-dependent formation of large scale parallel electric fields in the auroral zone. The results show that magnetic mirroring of the hot plasma that streams earthward from the magnetotail leads to a charge separation potential drop of many kilovolts, over an altitude range of a few thousand kilometers. Once the potential drop is formed, it remains relatively static and is maintained in time by the constant input of hot plasma from the tail; the parallel electric field accelerates ions away from Earth and ionospheric electrons towards the Earth. At altitudes above where the ions are mirror reflected and accelerated by the parallel electric field, low frequency waves are generated, possibly due to an ion/ion two-stream interaction.

Schriver, David↗

Optics Program Modified for Multithreaded Parallel Computing

A powerful high-performance computer program for simulating and analyzing adaptive and controlled optical systems has been developed by modifying the serial version of the Modeling and Analysis for Controlled Optical Systems (MACOS) program to impart capabilities for multithreaded parallel processing on computing systems ranging from supercomputers down to Symmetric Multiprocessing (SMP) personal computers. The modifications included the incorporation of OpenMP, a portable and widely supported application interface software, that can be used to explicitly add multithreaded parallelism to an application program under a shared-memory programming model. OpenMP was applied to parallelize ray-tracing calculations, one of the major computing components in MACOS. Multithreading is also used in the diffraction propagation of light in MACOS based on pthreads [POSIX Thread, (where "POSIX" signifies a portable operating system for UNIX)]. In tests of the parallelized version of MACOS, the speedup in ray-tracing calculations was found to be linear, or proportional to the number of processors, while the speedup in diffraction calculations ranged from 50 to 60 percent, depending on the type and number of processors. The parallelized version of MACOS is portable, and, to the user, its interface is basically the same as that of the original serial version of MACOS.

Lou, John↗

An Object Oriented Extensible Architecture for Affordable Aerospace Propulsion Systems

Driven by a need to explore and develop propulsion systems that exceeded current computing capabilities, NASA Glenn embarked on a novel strategy leading to the development of an architecture that enables propulsion simulations never thought possible before. Full engine 3 Dimensional Computational Fluid Dynamic propulsion system simulations were deemed impossible due to the impracticality of the hardware and software computing systems required. However, with a software paradigm shift and an embracing of parallel and distributed processing, an architecture was designed to meet the needs of future propulsion system modeling. The author suggests that the architecture designed at the NASA Glenn Research Center for propulsion system modeling has potential for impacting the direction of development of affordable weapons systems currently under consideration by the Applied Vehicle Technology Panel (AVT). This paper discusses the salient features of the NPSS Architecture including its interface layer, object layer, implementation for accessing legacy codes, numerical zooming infrastructure and its computing layer. The computing layer focuses on the use and deployment of these propulsion simulations on parallel and distributed computing platforms which has been the focus of NASA Ames. Additional features of the object oriented architecture that support MultiDisciplinary (MD) Coupling, computer aided design (CAD) access and MD coupling objects will be discussed. Included will be a discussion of the successes, challenges and benefits of implementing this architecture.

Follen, Gregory J.↗

Granular Simulation of NEO Anchoring

NASA is interested in designing a spacecraft capable of visiting a Near Earth Object (NEO), performing experiments, and then returning safely. Certain periods of this mission will require the spacecraft to remain stationary relative to the NEO. Such situations require an anchoring mechanism that is compact, easy to deploy and upon mission completion, easily removed. The design philosophy used in the project relies on the simulation capability of a multibody dynamics physics engine. On Earth it is difficult to create low gravity conditions and testing in low gravity environments, whether artificial or in space is costly and therefore not feasible. Through simulation, gravity can be controlled with great accuracy, making it ideally suited to analyze the problem at hand. Using Chrono::Engine [1], a simulation package capable of utilizing massively parallel GPU hardware, several validation experiments will be performed. Once there is sufficient confidence, modeling of the NEO regolith interaction will begin after which the anchor tests will be performed and analyzed. The outcome of this task is a study with an analysis of several different anchor designs, along with a recommendation on which anchor is better suited to the task of anchoring. With the anchors tested against a range of parameters relating to soil, environment and anchor penetration angles/velocities on a NEO.

multibody dynamics↗

Scaling neural simulations in STACS

Abstract As modern neuroscience tools acquire more details about the brain, the need to move towards biological-scale neural simulations continues to grow. However, effective simulations at scale remain a challenge. Beyond just the tooling required to enable parallel execution, there is also the unique structure of the synaptic interconnectivity, which is globally sparse but has relatively high connection density and non-local interactions per neuron. There are also various practicalities to consider in high performance computing applications, such as the need for serializing neural networks to support potentially long-running simulations that require checkpoint-restart. Although acceleration on neuromorphic hardware is also a possibility, development in this space can be difficult as hardware support tends to vary between platforms and software support for larger scale models also tends to be limited. In this paper, we focus our attention on Simulation Tool for Asynchronous Cortical Streams (STACS), a spiking neural network simulator that leverages the Charm++ parallel programming framework, with the goal of supporting biological-scale simulations as well as interoperability between platforms. Central to these goals is the implementation of scalable data structures suitable for efficiently distributing a network across parallel partitions. Here, we discuss a straightforward extension of a parallel data format with a history of use in graph partitioners, which also serves as a portable intermediate representation for different neuromorphic backends. We perform scaling studies on the Summit supercomputer, examining the capabilities of STACS in terms of network build and storage, partitioning, and execution. We highlight how a suitably partitioned, spatially dependent synaptic structure introduces a communication workload well-suited to the multicast communication supported by Charm++. We evaluate the strong and weak scaling behavior for networks on the order of millions of neurons and billions of synapses, and show that STACS achieves competitive levels of parallel efficiency.

59 BASIC BIOLOGICAL SCIENCES↗

Simulation study of Type 2 counterstreaming electrons along auroral field lines

The production of counterstreaming electrons associated with parallel fields along auroral field lines is examined through the use of computer simulation. A 2 1/2-dimensional (two spatial and three velocity dimensions) electrostatic particle algorithm and auroral boundary conditions are used to set up a self-consistent V potential structure. The simulation produces signatures of counterstreaming electrons resembling those observed by the Dynamics Explorer 1 satellite. The main signatures are as follows: (1) the phase space contours of the electron distribution function are elongated along the V-parallel axis, and (2) the energy of electrons streaming in the upward direction is comparable to the energy of the accelerated electron beam. The simulation indicates that a portion of the accelerated electron beam is trapped by large amplitude electrostatic waves produced through the two-stream instability. Strong wave-particle interactions then thermalize the trapped electrons to produce suprathermal electrons streaming in the direction opposite to that of the accelerated electron beam. These results suggest a possible mechanism of producing counterstreaming electron fluxes through nonlinear processes of the two-stream instability.

Wagner, J. S.↗

NASA Tech Briefs, February 2014

Topics include: JWST Integrated Simulation and Test (JIST) Core; Software for Non-Contact Measurement of an Individual's Heart Rate Using a Common Camera; Rapid Infrared Pixel Grating Response Testbed; Temperature Measurement and Stabilization in a Birefringent Whispering Gallery Resonator; JWST IV and V Simulation and Test (JIST) Solid State Recorder (SSR) Simulator; Development of a Precision Thermal Doubler for Deep Space; Improving Friction Stir Welds Using Laser Peening; Methodology of Evaluating Margins of Safety in Critical Brazed Joints; Interactive Inventory Monitoring; Sensor for Spatial Detection of Single-Event Effects in Semiconductor-Based Electronics; Reworked CCGA-624 Interconnect Package Reliability for Extreme Thermal Environments; Current-Controlled Output Driver for Directly Coupled Loads; Bulk Metallic Glasses and Matrix Composites as Spacecraft Shielding; Touch Temperature Coating for Electrical Equipment on Spacecraft; Li-Ion Electrolytes Containing Flame-Retardant Additives; Autonomous Robotic Manipulation (ARM); CARVE Log; Platform Perspective Toolkit; Convex Hull-Based Plume and Anomaly Detection; Pre-Filtration of GOSAT Data Using Only Level 1 Data and an Intelligent Filter to Remove Low Clouds; Affordability Comparison Tool - ACT; "Ascent - Commemorating Shuttle" for iPad; Cassini Mission App; Light-Weight Workflow Engine: A Server for Executing Generic Workflows; Model for System Engineering of the CheMin Instrument; Timeline Central Concepts; Parallel Particle Filter Toolkit; Particle Filter Simulation and Analysis Enabling Non-Traditional Navigation; Quasi-Terminator Orbits for Mapping Small Primitive Bodies; The Subgrid-Scale Scalar Variance Under Supercritical Pressure Conditions; Sliding Gait for ATHLETE Mobility; and Automated Generation of Adaptive Filter Using a Genetic Algorithm and Cyclic Rule Reduction.

Source record↗

Time-dependent simulation of cosmic-ray shocks, including Alfven transport

Time evolution of plane, cosmic-ray modified shocks has been simulated numerically for the case with parallel magnetic fields. Computations were done in a 'three-fluid' dynamical model incorporating cosmic-ray and Alfven-wave energy transport equations. Nonlinear feedback from the cosmic rays and Alfven waves is included in the equation of motion for the underlying plasma, as is the finite propagation speed and energy dissipation of the Alfven waves. Exploratory results confirm earlier, steady state analyses that found these Alfven transport effects to be potentially important when the upstream Alfven speed and gas sound speeds are comparable. As noted earlier, Alfven transport effects tend to reduce the transfer of energy through a shock from gas to energetic particles. These studies show as well that the timescale for modification of the shock is altered in nonlinear ways. It is clear, however, that the consequences of Alfven transport are strongly model dependent and that both advection of cosmic rays by the waves and dissipation of wave energy in the plasma will be important to model correctly when quantitative results are needed. Comparison is made between simulations based on a constant diffusion coefficient and more realistic diffusion models allowing the diffusion coefficient to vary in response to changes in Alfven wave intensity. No really substantive differences were found between them.

Jones, T. W.↗

Time dependent simulation of cosmic-ray shocks including Alfven transport

Time evolution of plane, cosmic-ray modified shocks was simulated numerically for the case with parallel magnetic fields. Computations were done in a 'three-fluid' dynamical model incorporating cosmic-ray and Alfven wave energy transport equations. Nonlinear feedback from the cosmic-rays and Alfven waves is included in the equation of motion for the underlying plasma, as is the finite propagation speed and energy dissipation of the Alfven waves. Exploratory results confirm earlier, steady state analyses that found these Alfven transport effects to be potentially important when the upstream Alfven speed and gas sound speeds are comparable. As noted earlier Alfven transport effects tend to reduce the transfer of energy through a shock from gas to energetic particles. These studies show as well that the time scale for modification of the shock is altered in nonlinear ways. It is clear, however, that the consequences of Alfven transport are strongly model dependent and that both advection of cosmic-rays by the waves and dissipation of wave energy in the plasma will be important to model correctly when quantitative results are needed. Comparison is made between simulations based on a constant diffusion coefficient and more realistic diffusion models allowing the diffusion coefficient to vary in response to changes in Alfven wave intensity. No really substantive differences were found between them.

Jones, T. W.↗

Numerical Simulation of Protoplanetary Vortices

The fluid dynamics within a protoplanetary disk has been attracting the attention of many researchers for a few decades. Previous works include, to list only a few among many others, the well-known prescription of Shakura & Sunyaev, the convective and instability study of Stone & Balbus and Hawley et al., the Rossby wave approach of Lovelace et al., as well as a recent work by Klahr & Bodenheimer, which attempted to identify turbulent flow within the disk. The disk is commonly understood to be a thin gas disk rotating around a central star with differential rotation (the Keplerian velocity), and the central quest remains as how the flow behavior deviates (albeit by a small amount) from a strong balance established between gravitational and centrifugal forces, transfers mass and momentum inward, and eventually forms planetesimals and planets. In earlier works we have briefly described the possible physical processes involved in the disk; we have proposed the existence of long-lasting, coherent vortices as an efficient agent for mass and momentum transport. In particular, Barranco et al. provided a general mathematical framework that is suitable for the asymptotic regime of the disk; Barranco & Marcus (2000) addressed a proposed vortex-dust interaction mechanism which might lead to planetesimal formation; and Lin et al. (2002), as inspired by general geophysical vortex dynamics, proposed basic mechanisms by which vortices can transport mass and angular momentum. The current work follows up on our previous effort. We shall focus on the detailed numerical implementation of our problem. We have developed a parallel, pseudo-spectral code to simulate the full three-dimensional vortex dynamics in a stably-stratified, differentially rotating frame, which represents the environment of the disk. Our simulation is validated with full diagnostics and comparisons, and we present our results on a family of three-dimensional, coherent equilibrium vortices.

Lin, H.↗

Combining machine-learned and empirical force fields with the parareal algorithm: application to the diffusion of atomistic defects

We numerically investigate an adaptive version of the parareal algorithm in the context of molecular dynamics. This adaptive variant has been originally introduced in [1]. We focus here on test cases of physical interest where the dynamics of the system is modelled by the Langevin equation and is simulated using the molecular dynamics software LAMMPS. In this work, the parareal algorithm uses a family of machine-learning spectral neighbor analysis potentials (SNAP) as fine, reference, potentials and embedded-atom method potentials (EAM) as coarse potentials. We consider a self-interstitial atom in a tungsten lattice and compute the average residence time of the system in metastable states. Our numerical results demonstrate significant computational gains using the adaptive parareal algorithm in comparison to a sequential integration of the Langevin dynamics. We also identify a large regime of numerical parameters for which statistical accuracy is reached without being a consequence of trajectorial accuracy.

36 MATERIALS SCIENCE↗

Rendezvous algorithms for large-scale modeling and simulation

Rendezvous algorithms encode a communication pattern that is useful when processors sending data do not know who the receiving processors should be, or vice versa. The idea is to define an intermediate decomposition where datums from different sending processors can ”rendezvous” to perform a computation, in a manner that both the senders and eventual receivers of the results can identify the appropriate rendezvous processor. Though they were originally designed for interpolating between overlaid grids with independent parallel decompositions (Plimpton et al., 2004), we have recently found rendezvous algorithms useful for a variety of operations in particle- or grid-based simulation codes when running large problems on large numbers of processors. In particular, we show they can perform well when a load-balanced intermediate decomposition is randomized and not spatial, requiring all-to-all communication to move data between processors. In this case rendezvous algorithms leverage the large bisection communication bandwidths which parallel machines provide. We describe how rendezvous algorithms work in a scientific computing context and give specific examples for molecular dynamics and Direct Simulation Monte Carlo codes which result in dramatic performance improvements versus simpler algorithms which do not scale as well. We explain how a generic rendezvous algorithm can be implemented, and also point out similarities with the MapReduce paradigm popularized by Google and Hadoop.

97 MATHEMATICS AND COMPUTING↗

Parallel sorting algorithm classification: is manual instrumentation necessary?

Understanding parallel algorithms is crucial for accelerating scientific simulations on complex, distributed memory, high-performance computers. Modern algorithm classification approaches learn semantics directly from source code to differentiate between algorithms, however, accessing source code is not always possible. We can learn about parallel algorithms from observing their performance, as programs running the same algorithms and using the same hardware should exhibit similar performance characteristics. We present an approach to learn algorithm classes from parallel performance data directly in order to classify algorithms without access to the source code. We extend previous work to enable classifying parallel sorting algorithms using automatic instrumentation instead of requiring manual region annotations in the source code. In this work, we design and demonstrate a study for classification of parallel sorting algorithms using parallel performance data collected from automatic instrumentation, and evaluate the performance of our new methodology on classification. We leverage Caliper to collect the performance data, Thicket for our exploratory data analysis (EDA), and PyTorch and Scikit-learn to evaluate the effectiveness of random forests, support vector machines (SVMs), decision trees, neural networks, and logistic regressions on parallel performance data. Additionally, we study noise in parallel performance data, whether the removal of noise and pre-processing of the data is necessary to accurately classify parallel sorting algorithms, and determine the effectiveness of features created from performance data. In conclusion, we demonstrate classification accuracy for these five different models of up to 97.7% across four different parallel algorithm classes.

Algorithm Classification↗

Advances in computational design and analysis of airbreathing propulsion systems

The development of commercial and military aircraft depends, to a large extent, on engine manufacturers being able to achieve significant increases in propulsion capability through improved component aerodynamics, materials, and structures. The recent history of propulsion has been marked by efforts to develop computational techniques that can speed up the propulsion design process and produce superior designs. The availability of powerful supercomputers, such as the NASA Numerical Aerodynamic Simulator, and the potential for even higher performance offered by parallel computer architectures, have opened the door to the use of multi-dimensional simulations to study complex physical phenomena in propulsion systems that have previously defied analysis or experimental observation. An overview of several NASA Lewis research efforts is provided that are contributing toward the long-range goal of a numerical test-cell for the integrated, multidisciplinary design, analysis, and optimization of propulsion systems. Specific examples in Internal Computational Fluid Mechanics, Computational Structural Mechanics, Computational Materials Science, and High Performance Computing are cited and described in terms of current capabilities, technical challenges, and future research directions.

Klineberg, John M.↗

Advances in computational design and analysis of airbreathing propulsion systems

The development of commercial and military aircraft depends, to a large extent, on engine manufacturers being able to achieve significant increases in propulsion capability through improved component aerodynamics, materials, and structures. The recent history of propulsion has been marked by efforts to develop computational techniques that can speed up the propulsion design process and produce superior designs. The availability of powerful supercomputers, such as the NASA Numerical Aerodynamic Simulator, and the potential for even higher performance offered by parallel computer architectures, have opened the door to the use of multi-dimensional simulations to study complex physical phenomena in propulsion systems that have previously defied analysis or experimental observation. An overview of several NASA Lewis research efforts is provided that are contributing toward the long-range goal of a numerical test-cell for the integrated, multidisciplinary design, analysis, and optimization of propulsion systems. Specific examples in Internal Computational Fluid Mechanics, Computational Structural Mechanics, Computational Materials Science, and High Performance Computing are cited and described in terms of current capabilities, technical challenges, and future research directions.

Klineberg, John M.↗

Nonlocal, diamagnetic electromagnetic effects in magnetically insulated transmission lines

We identify the time-dependent physics responsible for the critical reduction of current losses in magnetically insulated transmission lines (MITLs) due to uninsulated space charge-limited currents of electrons emitted by field stress. A drive current of sufficiently short pulse length introduces a strong enough time dependence that steady-state results alone become inadequate for the complete understanding of current losses. The time-dependent physics can be described as a nonlocal, diamagnetic electromagnetic response of space charge limited currents. As the pulse length is increased or equivalently, the MITL length reduced, these time-dependent effects diminish and current losses converge to those predicted by the well-known Child–Langmuir law in the external (vacuum) fields. We present a simple one-dimensional (1D) model that encapsulates the essence of this physics. We find excellent agreement with 2D particle-in-cell simulations for two MITL geometries, Cartesian parallel plate and azimuthally symmetric straight coaxial. Based on the 1D model, we explore various scaling dependencies of MITL losses with relevant parameters, e.g., peak current, pulse length, geometrical dimensions, etc. We propose an improved physics model of magnetic insulation in the form of a Hull curve, which could also help improve predictions of current losses by common circuit element codes, such as BERTHA. Finally, we describe how to calculate the temperature rise due to electron impact within the 1D model.

Computer simulation↗

Experimental determination of transient strain in a thermally-cycled simulated turbine blade utilizing a non-contact technique

A type of noncontacting electro-optical extensometer was used to measure the displacement between parallel targets mounted on the leading edge of a simulated turbine blade throughout a complete heating and cooling cycle. The blade was cyclically heated and cooled by moving it into and out of a Mach 1 hot gas stream. The principle of operation and measurement procedure of the electro-optics extensometer are described.

Calfo, F. D.↗