Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,261 records · Page 70

A Faster-Than-Real-Time Framework for Reliability-Oriented Simulation of PV Inverters

Physics-of-Failure (PoF) based reliability assessment for photovoltaic (PV) inverters requires long-duration electrical and electrothermal stress histories, yet generating such stress histories with high-fidelity switching models over year long mission profiles is computationally prohibitive. Conventional methods either sacrifice modeling fidelity for speed or require runtimes that are impractical for design iteration and uncertainty studies. To address this bottleneck, this paper presents a High-Performance Computing (HPC) based simulation frame work for faster-than-real-time reliability-oriented simulation. The proposed framework integrates the Average-to-Switching (A2S) method with parallel computing techniques to accelerate switching-level waveform reconstruction. We further introduce optimization strategies, including cluster merging and sensitivity based mission profile screening, to reduce the computational burden. Evaluated using real-world mission profile inputs and a MATLAB/Simulink switching-model reference, the framework reduces the simulation time for a one-year mission from an intractable multi-year duration to approximately 7.3 minutes while maintaining low waveform error. This acceleration provides a practical reliability-oriented simulation engine that can be coupled with component-specific aging models for subsequent PV inverter PoF assessment.

High-performance Computing↗

A Simulation Testbed for Airborne Merging and Spacing

The key innovation in this effort is the development of a simulation testbed for airborne merging and spacing (AM&S). We focus on concepts related to airports with Super Dense Operations where new airport runway configurations (e.g. parallel runways), sequencing, merging, and spacing are some of the concepts considered. We focus on modeling and simulating a complementary airborne and ground system for AM&S to increase efficiency and capacity of these high density terminal areas. From a ground systems perspective, a scheduling decision support tool generates arrival sequences and spacing requirements that are fed to the AM&S system operating on the flight deck. We enhanced NASA's Airspace Concept Evaluation Systems (ACES) software to model and simulate AM&S concepts and algorithms.

Santos, Michel↗

A hybrid 3D/2D field response calculation for liquid argon detectors with PCB based anode plane

Liquid Argon Time Projection Chamber (LArTPC) technology is commonly utilized in neutrino detector designs. It enables detailed reconstruction of neutrino events with high spatial precision and low energy threshold. Its field response (FR) model describes the time-dependent electric currents induced in the anode-plane electrodes when ionization electrons drift nearby. An accurate and precise FR is a crucial input to LArTPC detector simulations and charge reconstruction. Established LArTPC designs have been based on parallel wire planes. It allows accurate and computationally economic two-dimensional (2D) FR models utilizing the translational symmetry along the direction of the wires. Recently, novel LArTPC designs utilize electrodes formed on printed circuit board (PCB) in the shape of strips with through holes. The translational symmetry is no longer a good approximation near the electrodes and a new FR calculation that employs regions with three dimensions (3D) has been developed. Extending the 2D models to 3D would be computationally expensive. Fortuitously, the nature of strips with through holes allows for a computationally economic approach based on the finite-difference method (FDM). In this paper, we present a new software package pochoir that calculates LArTPC field response for these new strip-based anode designs. This package combines 3D calculations in the volume near the electrodes with 2D far-field solutions to achieve fast and precise field response computation. We apply the resulting FR to simulate and reconstruct samples of cosmic-ray muons and 39 Ar decays from a Vertical Drift (VD) detector prototype operated at CERN. We find the difference between real and simulated data within 5%. Current state-of-the-art LArTPC software requires a 2D FR which we provide by averaging over one dimension and estimate that variations lost in this average are smaller than 7%.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Observations of Closed Structures at the Magnetopause: A Case for Multiple Reconnections

We further analyze a case of Interball LLBL crossing on the dusk flank of geomagnetosphere under southward magnetosheath magnetic field, previously categorized as an interval of highly structured LLBL. These conditions of highly structured LLBL include reconnection signatures. Observed ion velocity distributions with LLBL are quite variable. D-shaped distributions that are associated with the open reconnected flux tube are observed at the boundaries of LLBL transients and sometimes within the LLBL transients. In most cases the ion velocity distributions consist of two magnetosheath-type components with different velocities parallel to the magnetic field, or of three components one of which has nearly zero Vpar. The shapes of ion velocity distributions and their evolution with decreasing number density in LLBL indicate that most of LLBL is located on closed magnetic field lines. These observations strongly favor multiple reconnections between magnetosheath and magnetosphereric flux tubes, creating long spiral flux tube islands at the magnetopause. We report evidence for the simultaneous occurrence of magnetic reconnection at multiple points across the magnetopause, as has been proposed and found to occur in magnetopause simulations. The evidence is in the form of highly structured distributions of ions in velocity parallel to the local magnetic field direction, within the magnetopause and low latitude boundary layer region, from the Interball-Tall spacecraft. We interpret these distributions as a natural consequence of the formation of spiral magnetic flux tubes consisting of a mixture of alternating segments originating from the magnetosheath or interplanetary plasma and from the low latitude boundary layer or magnetospheric plasma. We further analyze a case of Interball LLBL crossing on the dusk flank of geomagnetosphere under southward magnetosheath magnetic field, previously categorized as an interval of highly structured LLBL. These conditions of highly structured LLBL include reconnection signatures. Observed ion velocity distributions with LLBL are quite variable. D-shaped distributions that are associated with the open reconnected flux tube are observed at the boundaries of LLBL transients and sometimes within the LLBL transients. In most cases the ion velocity distributions consist of two magnetosheath-type components with different velocities parallel to the magnetic field, or of three components one of which has nearly zero Vpar. The shapes of ion velocity distributions and their evolution with decreasing number density in LLBL indicate that most of LLBL is located on closed magnetic field lines. These observations strongly favor multiple reconnection between magnetosheath and magnetospheric flux tubes, creating long spiral flux tube islands at the magnetopause. We report evidence for the simultaneous occurrence of magnetic reconnection at multiple points across the magnetopause, as has been proposed and found to occur in magnetopause simulations. The evidence is in the form of highly structured distributions of ions in velocity parallel to the local magnetic field direction, within the magnetopause and low latitude boundary layer region, from the Interball-Tail spacecraft. We interpret these distributions as a natural consequence of the formation of spiral. We further analyze a case of Interball LLBL crossing on the dusk flank of geomagnetosphere under southward magnetosheath magnetic field, previously categorized as an interval of highly structured LLBL. These conditions of highly structured LLBL include reconnection signatures. Observed ion velocity distributions with LLBL are quite variable. D-shaped distributions that are associated with the open reconnected flux tube are observed at the boundaries of LLBL transients and sometimes within the LLBL transients. In most cases the ion velocity distributions consist of two magnetosheath-type components with different velocities parallel to the magnetic field, or of three components one of which has nearly zero Vpar.

Vaisberg, O. L.↗

Terminal Area Procedures for Paired Runways

Parallel runway operations have been found to increase capacity within the National Airspace but poor visibility conditions reduce the use of these operations. The NextGen and SESAR Programs have identified the capacity benefits from increased use of closely-space parallel runway. Previous research examined the concepts and procedures related to parallel runways however, there has been no investigation of the procedures associated with the strategic and tactical pairing of aircraft for these operations. This simulation study developed and examined the pilot and controller procedures and information requirements for creating aircraft pairs for parallel runway operations. The goal was to achieve aircraft pairing with a temporal separation of 15s (+/- 10s error) at a coupling point that was about 12 nmi from the runway threshold. Two variables were explored for the pilot participants: two levels of flight deck automation (current-day flight deck automation and auto speed control future automation) as well as two flight deck displays that assisted in pilot conformance monitoring. The controllers were also provided with automation to help create and maintain aircraft pairs. Results show the operations in this study were acceptable and safe. Subjective workload, when using the pairing procedures and tools, was generally low for both controllers and pilots, and situation awareness was typically moderate to high. Pilot workload was influenced by display type and automation condition. Further research on pairing and off-nominal conditions is required however, this investigation identified promising findings about the feasibility of closely-spaced parallel runway operations.

Lozito, Sandra↗

Onset of Reconnection in the near Magnetotail: PIC Simulations

Using 2.5-dimensional particle-in-cell (PIC) simulations of magnetotail dynamics, we investigate the onset of reconnection in two-dimensional tail configurations with finite Bz. Reconnection onset is preceded by a driven phase, during which magnetic flux is added to the tail at the high-latitude boundaries, followed by a relaxation phase, during which the configuration continues to respond to the driving. We found a clear distinction between stable and unstable cases, dependent on deformation amplitude and ion/electron mass ratio. The threshold appears consistent with electron tearing. The evolution prior to onset, as well as the evolution of stable cases, are largely independent of the mass ratio, governed by integral flux tube entropy conservation as imposed in MHD (magnetohydrodynamics). This suggests that ballooning instability in the tail should not be expected prior to the onset of tearing and reconnection. The onset time and other onset properties depend on the mass ratio, consistent with expectations for electron tearing. At onset,we found electron anisotropies T⊥⁄ T∥ (bottom tail divided by parallel tail) equals 1.1-1.3, raising growth rates and wavenumbers. Our simulations have provided a quantitative onset criterion that is easily evaluated in MHD simulations, provided the spatial resolution is sufficient. The evolution prior to onset and after the formation of a neutral line does not depend on the electron physics, which should permit an approximation by MHD simulations with appropriate dissipation terms.

PIC↗

Analysis and modeling of tungsten emission and net erosion in the DIII-D divertor using updated atomic data

Tungsten (W) is one of the leading candidate materials for plasma-facing components. However, its main drawback is its high radiative efficiency; if W penetrates the plasma, it can lead to core degradation or even collapse. Since eroded tungsten tends to ionize in the sheath and redeposit promptly, the net erosion flux that escapes prompt redeposition can differ significantly from the gross erosion. This work presents a modeling framework to estimate net erosion and photon emission from W coatings exposed to the lower divertor of DIII-D using the DiMES material exposure probe. The approach couples RustBCA for sputtering yields with a Monte Carlo transport code (LPTMC) that models redeposition and W emission. Computation is carried out with new atomic data, based on R-matrix and Mons calculations, leading to lower ionization probabilities and a twofold increase in net erosion estimates compared to calculations done with OPEN-ADAS atomic data. The model results are benchmarked against experimental measurements, showing quantitative agreement for erosion, although the trends in W emission are reproduced only qualitatively. The model is also used to assess whether W II emission can serve as a direct measurement of the net erosion of W in the lower divertor of DIII-D. Simulations show that this is not valid if the electron pressure is above ~120 Pa or if the toroidal length of the eroded material is smaller than the parallel-to-B distance traveled by impurity ions before steady-state conditions are reached. Finally, simulations suggest that when W is sputtered by carbon ions with high impact energies (≳300 eV) in DIII-D, W net erosion scales with W gross erosion and can be numerically approximated using W I flux alone as input.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Mahakala: A Python-based Modular Ray-tracing and Radiative Transfer Algorithm for Curved Spacetimes

We introduce Mahakala, a Python-based, modular, radiative ray-tracing code for curved spacetimes. We employ Google's JAX framework for accelerated automatic differentiation, which can efficiently compute Christoffel symbols directly from the metric, allowing the user to easily and quickly simulate photon trajectories through non-Kerr spacetimes. JAX also enables Mahakala to run in parallel on both CPUs and GPUs. Mahakala natively uses the Cartesian Kerr–Schild coordinate system, which avoids numerical issues caused by the pole in spherical coordinate systems. We demonstrate Mahakala's capabilities by simulating 1.3 mm wavelength images (the wavelength of Event Horizon Telescope observations) of general relativistic magnetohydrodynamic simulations of low-accretion rate supermassive black holes. The modular nature of Mahakala allows us to quantitatively explore how different regions of the flow influence different image features. We show that most of the emission seen in 1.3 mm images originates close to the black hole and peaks near the photon orbit. We also quantify the relative contribution of the disk, forward jet, and counterjet to 1.3 mm images.

79 ASTRONOMY AND ASTROPHYSICS↗

Parallelized domain decomposition for multi-dimensional Lagrangian random walk mass-transfer particle tracking schemes

Lagrangian particle tracking schemes allow a wide range of flow and transport processes to be simulated accurately, but a major challenge is numerically implementing the inter-particle interactions in an efficient manner. This article develops a multi-dimensional, parallelized domain decomposition (DDC) strategy for mass-transfer particle tracking (MTPT) methods in which particles exchange mass dynamically. We show that this can be efficiently parallelized by employing large numbers of CPU cores to accelerate run times. In order to validate the approach and our theoretical predictions we focus our efforts on a well-known benchmark problem with pure diffusion, where analytical solutions in any number of dimensions are well established. In this work, we investigate different procedures for “tiling” the domain in two and three dimensions (2-D and 3-D), as this type of formal DDC construction is currently limited to 1-D. An optimal tiling is prescribed based on physical problem parameters and the number of available CPU cores, as each tiling provides distinct results in both accuracy and run time. We further extend the most efficient technique to 3-D for comparison, leading to an analytical discussion of the effect of dimensionality on strategies for implementing DDC schemes. Increasing computational resources (cores) within the DDC method produces a trade-off between inter-node communication and on-node work. For an optimally subdivided diffusion problem, the 2-D parallelized algorithm achieves nearly perfect linear speedup in comparison with the serial run-up to around 2700 cores, reducing a 5 h simulation to 8 s, while the 3-D algorithm maintains appreciable speedup up to 1700 cores.

97 MATHEMATICS AND COMPUTING↗

Parallelization of Lower-Upper Symmetric Gauss-Seidel Method for Chemically Reacting Flow

Development of technologies for exploration of the solar system has revived an interest in computational simulation of chemically reacting flows since planetary probe vehicles exhibit non-equilibrium phenomena during the atmospheric entry of a planet or a moon as well as the reentry to the Earth. Stability in combustion is essential for new propulsion systems. Numerical solution of real-gas flows often increases computational work by an order-of-magnitude compared to perfect gas flow partly because of the increased complexity of equations to solve. Recently, as part of Project Columbia, NASA has integrated a cluster of interconnected SGI Altix systems to provide a ten-fold increase in current supercomputing capacity that includes an SGI Origin system. Both the new and existing machines are based on cache coherent non-uniform memory access architecture. Lower-Upper Symmetric Gauss-Seidel (LU-SGS) relaxation method has been implemented into both perfect and real gas flow codes including Real-Gas Aerodynamic Simulator (RGAS). However, the vectorized RGAS code runs inefficiently on cache-based shared-memory machines such as SGI system. Parallelization of a Gauss-Seidel method is nontrivial due to its sequential nature. The LU-SGS method has been vectorized on an oblique plane in INS3D-LU code that has been one of the base codes for NAS Parallel benchmarks. The oblique plane has been called a hyperplane by computer scientists. It is straightforward to parallelize a Gauss-Seidel method by partitioning the hyperplanes once they are formed. Another way of parallelization is to schedule processors like a pipeline using software. Both hyperplane and pipeline methods have been implemented using openMP directives. The present paper reports the performance of the parallelized RGAS code on SGI Origin and Altix systems.

Yoon, Seokkwan↗

Memory-Aware External Facelist Calculation: A Data-Parallel Atomic Hash Counting Approach

Unstructured volumetric meshes serve as fundamental data representations in various scientific simulations and analyses. They play a crucial role in representing complex computational domains and are essential for important numerical techniques, such as finite element analysis. Whenever such a mesh is read from a file, streamed in-situ, or generated by algorithms, scientific visualization libraries rely on calculating the external surface of a geometry, named “external facelist”, to produce a polygonal mesh for rendering. Consequently, external facelist calculation has become one of the most widely used algorithms in the scientific visualization domain, necessitating optimal performance. In this paper, we explore relevant work on external facelist calculation algorithms in two common visualization libraries, VTK and Viskores, assess their performance and memory constraints, and introduce a novel memory-aware external facelist calculation algorithm employing an atomic hash counting approach. This algorithm fully leverages Viskores' data-parallel primitive operations, facilitating its execution across diverse many-core architectures. Our algorithm features the lowest memory footprint on the GPU and the second-lowest on the CPU among all evaluated methods, and it also delivers the fastest performance on both CPU and GPU. It has been made available under an open-source license in the VTK and Viskores visualization systems.

Tsalikis, Spiros [Kitware] (ORCID:0000000151137195↗

Efficiently modeling neural networks on massively parallel computers

Neural networks are a very useful tool for analyzing and modeling complex real world systems. Applying neural network simulations to real world problems generally involves large amounts of data and massive amounts of computation. To efficiently handle the computational requirements of large problems, we have implemented at Los Alamos a highly efficient neural network compiler for serial computers, vector computers, vector parallel computers, and fine grain SIMD computers such as the CM-2 connection machine. This paper describes the mapping used by the compiler to implement feed-forward backpropagation neural networks for a SIMD (Single Instruction Multiple Data) architecture parallel computer. Thinking Machines Corporation has benchmarked our code at 1.3 billion interconnects per second (approximately 3 gigaflops) on a 64,000 processor CM-2 connection machine (Singer 1990). This mapping is applicable to other SIMD computers and can be implemented on MIMD computers such as the CM-5 connection machine. Our mapping has virtually no communications overhead with the exception of the communications required for a global summation across the processors (which has a sub-linear runtime growth on the order of O(log(number of processors)). We can efficiently model very large neural networks which have many neurons and interconnects and our mapping can extend to arbitrarily large networks (within memory limitations) by merging the memory space of separate processors with fast adjacent processor interprocessor communications. This paper will consider the simulation of only feed forward neural network although this method is extendable to recurrent networks.

Farber, Robert M.↗

Improving Fidelity of Launch Vehicle Liftoff Acoustic Simulations

Launch vehicles experience high acoustic loads during ignition and liftoff affected by the interaction of rocket plume generated acoustic waves with launch pad structures. Application of highly parallelized Computational Fluid Dynamics (CFD) analysis tools optimized for application on the NAS computer systems such as the Loci/CHEM program now enable simulation of time-accurate, turbulent, multi-species plume formation and interaction with launch pad geometry and capture the generation of acoustic noise at the source regions in the plume shear layers and impingement regions. These CFD solvers are robust in capturing the acoustic fluctuations, but they are too dissipative to accurately resolve the propagation of the acoustic waves throughout the launch environment domain along the vehicle. A hybrid Computational Fluid Dynamics and Computational Aero-Acoustics (CFD/CAA) modeling framework has been developed to improve such liftoff acoustic environment predictions. The framework combines the existing highly-scalable NASA production CFD code, Loci/CHEM, with a high-order accurate discontinuous Galerkin (DG) solver, Loci/THRUST, developed in the same computational framework. Loci/THRUST employs a low dissipation, high-order, unstructured DG method to accurately propagate acoustic waves away from the source regions across large distances. The DG solver is currently capable of solving up to 4th order solutions for non-linear, conservative acoustic field propagation. Higher order boundary conditions are implemented to accurately model the reflection and refraction of acoustic waves on launch pad components. The DG solver accepts generalized unstructured meshes, enabling efficient application of common mesh generation tools for CHEM and THRUST simulations. The DG solution is coupled with the CFD solution at interface boundaries placed near the CFD acoustic source regions. Both simulations are executed simultaneously with coordinated boundary condition data exchange.

Liever, Peter↗

MPACT Verification With Magnox Reactor Neutronics Progression Problems

MPACT is a state-of-the-art core simulator designed to perform high-fidelity analysis using whole-core, three-dimensional, pin-resolved neutron transport calculations on modern parallel computing hardware. MPACT was originally developed to model light water reactors, and its capabilities are being extended to simulate gas-cooled, graphite-moderated cores such as Magnox reactors. To verify MPACT’s performance in this new application, the code is being formally benchmarked using representative problems. Progression problems are a series of example models that increase in complexity designed to test a code’s performance. The progression problems include both beginning-of-cycle and depletion calculations. Reference solutions for each progression problem have been generated using Serpent 2, a continuous-energy Monte Carlo reactor physics burnup calculation code.Using the neutron multiplication eigenvalue ke as a metric, MPACT’s performance is assessed on each of the progression problems. Initial results showed that MPACT’s multi-group cross section libraries, originally developed for pressurized water reactor problems, were not sufficient to accurately solve Magnox problems. MPACT’s improved performance on the progression problems is demonstrated using this new optimized cross section library.

Luciano, Nicholas↗

Power combining in an array of microwave power rectifiers

This work analyzes the resultant efficiency degradation when identical rectifiers operate at different RF power levels as caused by the power beam taper. Both a closed-form analytical circuit model and a detailed computer-simulation model are used to obtain the output dc load line of the rectifier. The efficiency degradation is nearly identical with series and parallel combining, and the closed-form analytical model provides results which are similar to the detailed computer-simulation model.

Gutmann, R. J.↗

Weak, quasiparallel profiles of earth's bow shock - A comparison between numerical simulations and ISEE 3 observations on the far flank

Over 200 crossings of the distant downwind flanks of earth's magnetosonic bow shock by ISEE 3 included many cases of weak, or low Mach number, quasi-parallel shocks. A consistent feature of the magnetic field profiles was the presence of large amplitude, near periodic to irregular transverse oscillations downstream from even the weakest Q-parallel shocks. Large downstream perturbations with whistler-like features similar to those of the observations appear in 1D simulations when the Alfven Mach number M(A) is greater than 2.5 but not when M(A) = 2.1. The observed cases with downstream waves also occurred when M(A) is greater than about 2.5, suggesting the importance of the Alfven as opposed to magnetosonic Mach number in determining the signature of weak, Q-parallel shocks.

Greenstadt, E. W.↗

Development of PFLOTRAN Transport Capability for Use in the Waste Isolation Pilot Plant Performance Assessment - 20545

Waste Isolation Pilot Plant (WIPP) performance assessment (PA) calculations estimate the probability of radionuclide release from the repository to the land surface and across the land withdrawal boundary for a regulatory period of 10,000 years after facility closure. Simulations of flow and transport in the repository and the surrounding Salado Formation are foundational to the PA. Because proposed additional waste emplacement panels would result in an asymmetric repository layout, the US Department of Energy (DOE) is preparing to transition to use of a three-dimensional (3-D) model domain for simulation of flow and transport instead of the two-dimensional (2-D) flared grid domain currently used. DOE has charged Sandia National Laboratories with developing the capability necessary to simulate processes affecting flow and transport in the WIPP in PFLOTRAN, an open-source massively parallel multi-phase flow and reactive transport code. The new flow and transport capabilities developed in PFLOTRAN incorporate WIPP-specific process models and will replace the 2-D simulators (BRAGFLO and NUTS) that are currently utilized for Salado flow and transport calculations in WIPP PA. The focus of this paper is on the development of a new Nuclear Waste Transport (NWT) mode in PFLOTRAN that has all of the capabilities necessary for Salado transport simulations, including the ability to handle complete dry-out (100% gas saturation) of arbitrary cells in the model domain, radionuclide mass conservation at step changes in porosity associated with borehole intrusion, and the ability to calculate fluxes on a flared grid. The new PFLOTRAN transport capability and a suite of verification tests were designed around a list of functional requirements for WIPP PA calculations. (authors)

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Simulating Magnetic Reconnection in ‘Two Ribbon’ Type Solar Flares

Magnetic reconnection is an astrophysical process where neighboring magnetic field lines, facing anti-parallel, are reconfigured. This reconfiguration results in built up magnetic energy being explosively released as it is being converted to plasma kinetic and thermal energies. There are various kinds of simulations used to simulation reconnection; our work begins with Athena++, a magnetohydrodynamic (MHD) simulation code typically used for astrophysical problems, and a reconnection specific code file. The resistive MHD equations are solved with Riemann solvers. There was an initial test run without modifying the code to understand the dynamics of the simulation. We expand on the original reconnection problem file by implementing a radiative cooling term specific to the corona. The radiative cooling is theorized to have an effect on solar coronal plasma and magnetic reconnection dynamics. The cooling term will be tested with various parameters and compared to the case without cooling to study these dynamics. The condensation found in only the with cooling case emphasizes the importance of implementing this feature and will be later tested with a thermal conduction term. We want to determine the parameter regime where non-equilibrium cooling will be important for the reconnection dynamics.

79 ASTRONOMY AND ASTROPHYSICS↗