Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 469 records · Page 26

Simulation of auroral double layers

Some basic properties of plasma double layers are deduced from a particle-in-cell computer simulation and related to parallel electric-field structures above the auroral regions. The simulation results on the processes leading to double-layer formation are examined, particularly in relation to the transient stage and double-layer structure and stability. It is concluded that: (1) a large potential difference applied to a finite-length plasma will be concentrated in a shocklike localized region instead of occurring over the entire length of the system; (2) the initial stage in double-layer formation is dominated by a large-potential pulse propagating in the direction of the induced electrostatic drift; (3) the entire potential is dropped over a specific scale length once the double layer has formed; and (4) this scale length is expected to be of the order of 1 km for a double layer above a discrete auroral arc with a potential of 10 kV and the electric-field vector parallel to the magnetic-field vector.

Hubbard, R. F.↗

User's guide to the Reliability Estimation System Testbed (REST)

The Reliability Estimation System Testbed is an X-window based reliability modeling tool that was created to explore the use of the Reliability Modeling Language (RML). RML was defined to support several reliability analysis techniques including modularization, graphical representation, Failure Mode Effects Simulation (FMES), and parallel processing. These techniques are most useful in modeling large systems. Using modularization, an analyst can create reliability models for individual system components. The modules can be tested separately and then combined to compute the total system reliability. Because a one-to-one relationship can be established between system components and the reliability modules, a graphical user interface may be used to describe the system model. RML was designed to permit message passing between modules. This feature enables reliability modeling based on a run time simulation of the system wide effects of a component's failure modes. The use of failure modes effects simulation enhances the analyst's ability to correctly express system behavior when using the modularization approach to reliability modeling. To alleviate the computation bottleneck often found in large reliability models, REST was designed to take advantage of parallel processing on hypercube processors.

Nicol, David M.↗

A parallel and performance portable implementation of a full-field crystal plasticity model

We have developed a parallel implementation of an Elasto-Viscoplastic Fast Fourier Transform-based (EVPFFT) micromechanical solver to enable computationally efficient crystal plasticity modeling for polycrystalline materials. Our primary focus lies in achieving performance portability, allowing a single EVPFFT implementation to run optimally on various homogeneous architectures, including multi-core Central Processing Units (CPUs), as well as on heterogeneous computer architectures comprising multi-core CPUs and Graphics Processing Units (GPUs) from different vendors. To accomplish this goal, we have leveraged MATAR, a C++ software library that simplifies the creation and utilization of multidimensional dense or sparse matrix and array data structures. These data structures are designed to be portable across diverse architectures through the use of Kokkos, a performance-portable library. Additionally, we have employed the Message Passing Interface (MPI) to efficiently distribute the computational workload among processors. The heFFTe (Highly Efficient FFT for Exascale) library is used to facilitate the performance portability of the fast Fourier transforms (FFTs) computation. The computational performance of EVPFFT is evaluated and presented in terms of parallel scalability and simulation runtime on different high-performance computing (HPC) architectures. As a result, the utility of the developed framework to efficiently simulate the micro-mechanical fields in polycrystalline microstructures in engineering applications is discussed.

36 MATERIALS SCIENCE↗

Tool for Viewing Faults Under Terrain

Multi Surface Light Table (MSLT) is an interactive software tool that was developed in support of the QuakeSim project, which has created an earthquake- fault database and a set of earthquake- simulation software tools. MSLT visualizes the three-dimensional geometries of faults embedded below the terrain and animates time-varying simulations of stress and slip. The fault segments, represented as rectangular surfaces at dip angles, are organized into collections, that is, faults. An interface built into MSLT queries and retrieves fault definitions from the QuakeSim fault database. MSLT also reads time-varying output from one of the QuakeSim simulation tools, called "Virtual California." Stress intensity is represented by variations in color. Slips are represented by directional indicators on the fault segments. The magnitudes of the slips are represented by the duration of the directional indicators in time. The interactive controls in MSLT provide a virtual track-ball, pan and zoom, translucency adjustment, simulation playback, and simulation movie capture. In addition, geographical information on the fault segments and faults is displayed on text windows. Because of the extensive viewing controls, faults can be seen in relation to one another, and to the terrain. These relations can be realized in simulations. Correlated slips in parallel faults are visible in the playback of Virtual California simulations.

Siegel, Herbert, L.↗

Scheduling message processing for reducing rollback propagation

Traditional checkpointing and rollback recovery techniques for parallel systems have typically assumed the communication pattern is specified by program behavior. In this paper we exploit the property that the communication pattern can often be changed at run-time without affecting program correctness. A scheduling algorithm for message processing and its implementation for reducing rollback propagation are described. The algorithm incorporates a user-transparent prioritized scheme based upon the run-time communication and checkpointing history. Communication trace-driven simulation for several parallel programs written in the Chare Kernel language demonstrates that the probability of rollback propagation can be reduced at the cost of slight additional performance degradation.

Wang, Yi-Min↗

Optimal message log reclamation for independent checkpointing

Independent (uncoordinated) check pointing for parallel and distributed systems allows maximum process autonomy but suffers from possible domino effects and the associated storage space overhead for maintaining multiple checkpoints and message logs. In most research on check pointing and recovery, it was assumed that only the checkpoints and message logs older than the global recovery line can be discarded. It is shown how recovery line transformation and decomposition can be applied to the problem of efficiently identifying all discardable message logs, thereby achieving optimal garbage collection. Communication trace-driven simulation for several parallel programs is used to show the benefits of the proposed algorithm for message log reclamation.

Wang, Yi-Min↗

Optimal message log reclamation for independent checkpointing

Independent (uncoordinated) check pointing for parallel and distributed systems allows maximum process autonomy but suffers from possible domino effects and the associated storage space overhead for maintaining multiple checkpoints and message logs. In most research on check pointing and recovery, it was assumed that only the checkpoints and message logs older than the global recovery line can be discarded. It is shown how recovery line transformation and decomposition can be applied to the problem of efficiently identifying all discardable message logs, thereby achieving optimal garbage collection. Communication trace-driven simulation for several parallel programs is used to show the benefits of the proposed algorithm for message log reclamation.

Wang, Yi-Min↗

Parallelism and pipelining in high-speed digital simulators

The attainment of high computing speed as measured by the computational throughput is seen as one of the most challenging requirements. It is noted that high speed is cardinal in several distinct classes of applications. These classes are then discussed; they comprise (1) the real-time simulation of dynamic systems , (2) distributed parameter systems, and (3) mixed lumped and distributed systems. From the 1950s on, the quest for high speed in digital simulators concentrated on overcoming the limitations imposed by the so-called von Neumann bottleneck. Two major architectural approaches have made ig possible to circumvent this bottleneck and attain high speeds. These are pipelining and parallelism. Supercomputers, peripheral array processors, and microcomputer networks are then discussed.

Karplus, W. J.↗

Additive Manufactured Composite Phase-Change Material for Thermal Energy Storage Applications

Phase-change materials play a critical role in industrial energy storage applications to drive efficiency improvements, thermal energy management, and carbon emissions reductions. Recently, it has been shown that rapid solidification of alloys with metastable immiscibility in the liquid phase has the potential to form unique microstructures in which a low-melting phase is uniformly distributed in a high-melting matrix. This feature can be exploited using additive manufacturing to produce components with complex geometries containing such unique phase-change microstructures. Phase-field simulations utilizing high-performance computing were used to provide a detailed description of the evolution of the active phase during service in terms of their morphology and composition in different polycrystalline matrix grain morphologies that are typically produced during additive manufacturing. Phase field simulations were performed using, MEUMAPPS-SL (Microstructure Evolution Using Massively Parallel Phase-field Simulations – Solid Liquid) code that was developed in-house by the Oak Ridge National Laboratory. The simulations utilized the capabilities of the Kestrel supercomputer at the National Renewable Energy Laboratory. The simulation results were compared with experimental results generated at Siemens Energy, Inc. The results indicate that the kinetics of liquid spreading along grain boundaries is largely determined by the mobility of the triple line along the intersection of the grain boundary liquid and the grain boundary plane.

25 ENERGY STORAGE↗

Reduction of blob-filament radial propagation by parallel variation of flows: Analysis of a gyrokinetic simulation

Data from the XGC1 gyrokinetic simulation is analyzed to understand the three-dimensional spatial structure and the radial propagation of blob-filaments generated by quasisteady turbulence in the tokamak edge pedestal and scrape-off layer plasma. Spontaneous toroidal flows vary in the poloidal direction and shear the filaments within a flux surface resulting in a structure that varies in the parallel direction. Here, this parallel structure allows the curvature and grad-B induced polarization charge density to be shorted out via parallel electron motion. As a result, it is found that the blob-filament radial velocity is significantly reduced from estimates which neglect parallel electron kinetics, broadly consistent with experimental observations. Conditions for when this charge shorting effect tends to dominate blob dynamics are derived and compared with the simulation.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

MARTINI-Compatible Coarse-Grained Model for the Mesoscale Simulation of Peptoids

Peptoids (poly-N-substituted glycines) are a class of synthetic polymers that are regioisomers of peptides (poly-C-substituted glycines), in which the point of side-chain connectivity is shifted from the backbone C to the N atom. Peptoids have found diverse applications as peptidomimetic drugs, protein mimetic polymers, surfactants, and catalysts. Computational modeling is valuable in the understanding and design of peptoid-based nanomaterials. In this work, we report the bottom-up parameterization of coarse-grained peptoid force fields based on the MARTINI peptide force field against all-atom peptoid simulation data. Our parameterization pipeline iteratively refits coarse-grained bonded interactions using iterative Boltzmann inversion and nonbonded interactions by matching the potential of mean force for chain extension. We assure good sampling of the amide bond cis/trans isomerizations in the all-atom simulation data using parallel bias metadynamics. We develop coarse-grained models for two representative peptoids—polysarcosine (poly(N-methyl glycine)) and poly(N-((4-bromophenyl)ethyl)glycine)—and show their structural and thermodynamic properties to be in excellent accord with all-atom calculations but up to 25-fold more efficient and compatible with MARTINI force fields. Here, this work establishes a new rigorously parameterized coarse-grained peptoid force field for the understanding and design of peptoid nanomaterials at length and time scales inaccessible to all-atom calculations.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

NPSS on NASA's IPG: Using CORBA and Globus to Coordinate Multidisciplinary Aeroscience Applications

Within NASA's High Performance Computing and Communication (HPCC) program, the NASA Glenn Research Center is developing an environment for the analysis/design of aircraft engines called the Numerical Propulsion System Simulation (NPSS). The vision for NPSS is to create a "numerical test cell" enabling full engine simulations overnight on cost-effective computing platforms. To this end, NPSS integrates multiple disciplines such as aerodynamics, structures, and heat transfer and supports "numerical zooming" between O-dimensional to 1-, 2-, and 3-dimensional component engine codes. In order to facilitate the timely and cost-effective capture of complex physical processes, NPSS uses object-oriented technologies such as C++ objects to encapsulate individual engine components and CORBA ORBs for object communication and deployment across heterogeneous computing platforms. Recently, the HPCC program has initiated a concept called the Information Power Grid (IPG), a virtual computing environment that integrates computers and other resources at different sites. IPG implements a range of Grid services such as resource discovery, scheduling, security, instrumentation, and data access, many of which are provided by the Globus toolkit. IPG facilities have the potential to benefit NPSS considerably. For example, NPSS should in principle be able to use Grid services to discover dynamically and then co-schedule the resources required for a particular engine simulation, rather than relying on manual placement of ORBs as at present. Grid services can also be used to initiate simulation components on parallel computers (MPPs) and to address inter-site security issues that currently hinder the coupling of components across multiple sites. These considerations led NASA Glenn and Globus project personnel to formulate a collaborative project designed to evaluate whether and how benefits such as those just listed can be achieved in practice. This project involves firstly development of the basic techniques required to achieve co-existence of commodity object technologies and Grid technologies; and secondly the evaluation of these techniques in the context of NPSS-oriented challenge problems. The work on basic techniques seeks to understand how "commodity" technologies (CORBA, DCOM, Excel, etc.) can be used in concert with specialized "Grid" technologies (for security, MPP scheduling, etc.). In principle, this coordinated use should be straightforward because of the Globus and IPG philosophy of providing low-level Grid mechanisms that can be used to implement a wide variety of application-level programming models. (Globus technologies have previously been used to implement Grid-enabled message-passing libraries, collaborative environments, and parameter study tools, among others.) Results obtained to date are encouraging: we have successfully demonstrated a CORBA to Globus resource manager gateway that allows the use of CORBA RPCs to control submission and execution of programs on workstations and MPPs; a gateway from the CORBA Trader service to the Grid information service; and a preliminary integration of CORBA and Grid security mechanisms. The two challenge problems that we consider are the following: 1) Desktop-controlled parameter study. Here, an Excel spreadsheet is used to define and control a CFD parameter study, via a CORBA interface to a high throughput broker that runs individual cases on different IPG resources. 2) Aviation safety. Here, about 100 near real time jobs running NPSS need to be submitted, run and data returned in near real time. Evaluation will address such issues as time to port, execution time, potential scalability of simulation, and reliability of resources. The full paper will present the following information: 1. A detailed analysis of the requirements that NPSS applications place on IPG. 2. A description of the techniques used to meet these requirements via the coordinated use of CORBA and Globus. 3. A description of results obtained to date in the first two challenge problems.

Lopez, Isaac↗

Identifying the Growth Phase of Magnetic Reconnection Using Pressure‐Strain Interaction

Abstract Magnetic reconnection often initiates abruptly and then rapidly progresses to a nonlinear quasi‐steady state. While satellites frequently detect reconnection events, ascertaining whether the system has achieved steady‐state or is still evolving in time remains challenging. Here, we propose that the relatively rapid opening of the reconnection separatrices within the electron diffusion region serves as an indicator of the growth phase of reconnection. The opening of the separatrices is produced by electron flows diverging away from the neutral line downstream of the X‐line and flowing around a dipolarization front. This flow pattern leads to characteristic spatial structures in the electron pressure‐strain interaction that could be a useful indicator for the growth phase of a reconnection event. We employ two‐dimensional particle‐in‐cell numerical simulations of anti‐parallel magnetic reconnection to validate this prediction. We find that the signature discussed here, alongside traditional reconnection indicators, can serve as a marker of the growth phase. This signature is potentially accessible using multi‐spacecraft single‐point measurements, such as with NASA's Magnetospheric Multiscale satellites in Earth's magnetotail. Applications to other settings where reconnection occurs are also discussed.

Barbhuiya, M. Hasan [Department of Physics and Ast↗

Modeling of small tungsten dust grains in EAST tokamak with NDS-BOUT ++

In order to investigate the transport of small dusts as well as their evolution property along their trajectories, the NDS module is developed under the BOUT++ framework, a highly desirable C++ code package to perform parallel plasma fluid simulations with an arbitrary number of equations in three-dimensional curvilinear coordinates. Due to the severe dust ablation in fusion plasmas, the dust size would decrease from micrometer to nanometer, resulting in impurities. Small dusts in the simulations here are specified as tungsten spheres with the radii on or below the order of submicrometer. The Rayleigh limit is included in the charging process when the dust is ablated to the droplet phase. The simulation results from the NDS module show that a 200 nm radius spherical tungsten dust originated from upper divertor region of EAST Tokamak is ablated completely due to the intense heating from the incoming plasma inside the core region, well consistent with the CCD footage of EAST shot # 81459. Furthermore it is found that the magnetic field dominates the dust transport when the dust radius is below 100 nm during the ablation along the trajectory. Our simulations predict that a 10 nm radius spherical tungsten dust injected from the inner midplane is well constrained by the magnetic field, and it reaches the inner divertor target with a velocity on the order of km/s.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Magnetic island formation and rotation braking induced by low-Z impurity penetration in an EAST plasma

Abstract Recent observations of the successive formations of the 4 / 1 , 3 / 1 , and 2 / 1 magnetic islands as well as the subsequent braking of the 2 / 1 mode during a low- Z impurity penetration process in EAST experiments are well reproduced in our 3 D resistive MHD simulations. The enhanced parallel current perturbation induced by impurity radiation predominately contributes to the tearing mode growth, and the 2 / 1 island rotation is mainly damped by the impurity accumulation as results of the influence from high n modes.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Online and Scalable Data Compression Pipeline with Guarantees on Quantities of Interest

Data compression is becoming critical for data-intensive scientific applications. Scientists require compression techniques that accurately preserve derived quantities of interest (QoIs). Prior work has shown that a pipeline can be built to guarantee error on the primary data (PD) within user-defined bounds and achieve near-floating point QoI errors. In this paper, we present novel computational approaches for accelerating the pipeline and demonstrate results that enable concurrent execution of compression in parallel with the simulation nodes. This allows compression, including the writing of the required compression data, for the previous time step to be completed while the simulation proceeds with the current time step. Overall, the approach presented in this paper results in a 6–8 times improvement in computational overhead compared to previous work. These results were obtained using data generated by a large-scale fusion code called XGC, which produces hundreds of terabytes of data in a single day.

Banerjee, Tania↗

Evaluating Trade-offs in Potential Exascale Interconnect Technologies

This report details work to study trade-offs in topology and network bandwidth for potential interconnects in the exascale (2021-2022) timeframe. The work was done using multiple interconnect models across two parallel discrete event simulators. Results from each independent simulator are shown and discussed and the areas of agreement and disagreement are explored.

97 MATHEMATICS AND COMPUTING↗

Biologically Inspired Interception on an Unmanned System

Borrowing from nature, neural-inspired interception algorithms were implemented onboard a vehicle. To maximize success, work was conducted in parallel within a simulated environment and on physical hardware. The intercept vehicle used only optical imaging to detect and track the target. A successful outcome is the proof-of-concept demonstration of a neural-inspired algorithm autonomously guiding a vehicle to intercept a moving target. This work tried to establish the key parameters for the intercept algorithm (sensors and vehicle) and expand the knowledge and capabilities of implementing neural-inspired algorithms in simulation and on hardware.

42 ENGINEERING↗