Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Intelligently Partitioned Phasor-EMT Hybrid Simulations of Large-Scale, High-IBR Power Systems

As the penetration level of power electronics-interfaced renewables such as photovoltaics (PV) and wind has surged in modern electric grids, new operational risks caused by the dynamics of those inverter-based resources (IBRs) are emerging in parallel. Lessons learned from various grid events include that the impact of IBRs on system-level grid stability will become prominent along with the increase of renewables and that the short-timescale dynamic impacts of IBRs on grid stability are not fully captured by current commercial dynamic simulation tools [1] [2]. For example, IBRs can be controlled to mitigate those destabilizing interactions, but conventional phasor-domain tools (e.g. PSS/E, PSLF) often cannot capture that; likewise, the existing electromagnetic transient (EMT) simulation tools (e.g. PSCAD, EMTP) can simulate detailed IBR controls, but for large power systems with many IBRs, slow simulation speeds severely impede the ability to study dynamic events [3] [4]. Massively paralleling simulations using high-performance computing (HPC) can help address this, especially now that cloud-based HPC capability is widely available, but today s EMT tools are not HPC-compatible, and parallelization of dynamic simulation solvers is not trivial because each region can dynamically affect the others. Thus, dynamic simulation of grids with very large numbers of IBRs potentially poses a barrier to the ongoing energy transition.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Non-local large-strain FFT-based formulation and its application to interface-dominated plasticity of nano-metallic laminates

Here, this paper presents a novel formulation and its robust numerical implementation of strain-gradient (SG) crystal plasticity within a large-strain (LS) elasto-viscoplastic (EVP) fast Fourier transform (FFT)-based micromechanical model. The resulting non-local SG-LS-EVPFFT formulation is used to model and understand the process of kink band formation during layer-parallel compression of nano-metallic laminates (NMLs). NMLs are layered composites with nanoscale thicknesses, thus requiring consideration of the interaction between dislocations and interfaces within the micromechanical model. The length-scale parameter of the SG model is calibrated by simulating a double pile-up and comparing predictions to analytical solution. This required new expressions for the defect energy, resulting in more accurate double pile-up predictions. The calibrated SG-LS-EVPFFT model is then used to simulate layer-parallel compression of copper-niobium NML. Formation of kink bands is predicted, and the model is used to rationalize the microscopic mechanisms enabling the formation process. It is found that accumulation of dislocations at interfaces leads to activation of layer-parallel slip, which in turn leads to kink band formation.

36 MATERIALS SCIENCE↗

Accelerating Neutrino Event Generation in MARLEY Using CUDA-Based RNG and GPU Parallelization

MARLEY is a simulation tool that helps scientists study how low-energy neutrinos interact with matter. To work properly, MARLEY uses random numbers thousands of times in each simulation. These random numbers are important for modeling things like how neutrinos collide with atoms and what particles they produce. Right now, MARLEY runs on a regular computer processor (CPU) and uses a built-in random number generator called the Mersenne Twister. This setup works, but it can be slow, especially when trying to simulate many events. This research focuses on making MARLEY run faster by moving the random number generation and some of the repetitive calculations from the CPU to a graphics processing unit (GPU), which can handle many tasks at the same time. We use CUDA (a tool for programming NVIDIA GPUs) and cuRAND (a GPU-based random number library) to test faster alternatives to the current random number system. We compare different GPU-based generators, like curand_mtgp32, xorwow, and philox, to see which ones are the quickest and still give reliable results. Early tests show that using the GPU can make MARLEY simulations much faster. This project not only helps improve current simulation performance but also moves closer to a full simulation chain where all stages can run on modern GPU hardware.

Dunkley, Kimieka [Florida A-M]↗

HPC for Optimizing Process Parameters to Control Material Evolution in Seamless Induction Hardening of Wind Turbine Main Shaft Bearings

Work proposed in this project focused on understanding the effect of martensitic transformation in the steel on the potential for cracking during seamless induction hardening (SIH) as a function of process conditions to allow the process to optimally scale up. Large-scale, three-dimensional phase-field simulations of martensitic transformation were performed using MEUMAPPS-SS (Microstructure Evolution Using Massively Parallel Phase-field Simulations – Solid State) code developed at Oak Ridge National Laboratory. The simulations were guided by location-specific thermal history generated by experimental measurements of time-temperature history generated at The Timken Company. The simulations were able to capture the morphological evolution of the martensite variants in an Fe-1.0C-1.5Cr steel based on the Nishiyama-Wasserman (NW) orientation relationship. The simulations were also able to quantify the stress-state at the interface between impinging martensite variants. The simulations indicated that the magnitude of the various stress and strain components were dependent on the sizes of the impinging plates with a reduction in these quantities with reduced plate size in agreement with experimental findings. The results obtained from the simulations will be used to guide the optimization of the alloy thermal conditions to eliminate quench cracking during SIH of bearing steels.

99 GENERAL AND MISCELLANEOUS↗

A parallel and performance portable implementation of a full-field crystal plasticity model

We have developed a parallel implementation of an Elasto-Viscoplastic Fast Fourier Transform-based (EVPFFT) micromechanical solver to enable computationally efficient crystal plasticity modeling for polycrystalline materials. Our primary focus lies in achieving performance portability, allowing a single EVPFFT implementation to run optimally on various homogeneous architectures, including multi-core Central Processing Units (CPUs), as well as on heterogeneous computer architectures comprising multi-core CPUs and Graphics Processing Units (GPUs) from different vendors. To accomplish this goal, we have leveraged MATAR, a C++ software library that simplifies the creation and utilization of multidimensional dense or sparse matrix and array data structures. These data structures are designed to be portable across diverse architectures through the use of Kokkos, a performance-portable library. Additionally, we have employed the Message Passing Interface (MPI) to efficiently distribute the computational workload among processors. The heFFTe (Highly Efficient FFT for Exascale) library is used to facilitate the performance portability of the fast Fourier transforms (FFTs) computation. The computational performance of EVPFFT is evaluated and presented in terms of parallel scalability and simulation runtime on different high-performance computing (HPC) architectures. As a result, the utility of the developed framework to efficiently simulate the micro-mechanical fields in polycrystalline microstructures in engineering applications is discussed.

36 MATERIALS SCIENCE↗

Additive Manufactured Composite Phase-Change Material for Thermal Energy Storage Applications

Phase-change materials play a critical role in industrial energy storage applications to drive efficiency improvements, thermal energy management, and carbon emissions reductions. Recently, it has been shown that rapid solidification of alloys with metastable immiscibility in the liquid phase has the potential to form unique microstructures in which a low-melting phase is uniformly distributed in a high-melting matrix. This feature can be exploited using additive manufacturing to produce components with complex geometries containing such unique phase-change microstructures. Phase-field simulations utilizing high-performance computing were used to provide a detailed description of the evolution of the active phase during service in terms of their morphology and composition in different polycrystalline matrix grain morphologies that are typically produced during additive manufacturing. Phase field simulations were performed using, MEUMAPPS-SL (Microstructure Evolution Using Massively Parallel Phase-field Simulations – Solid Liquid) code that was developed in-house by the Oak Ridge National Laboratory. The simulations utilized the capabilities of the Kestrel supercomputer at the National Renewable Energy Laboratory. The simulation results were compared with experimental results generated at Siemens Energy, Inc. The results indicate that the kinetics of liquid spreading along grain boundaries is largely determined by the mobility of the triple line along the intersection of the grain boundary liquid and the grain boundary plane.

25 ENERGY STORAGE↗

Reduction of blob-filament radial propagation by parallel variation of flows: Analysis of a gyrokinetic simulation

Data from the XGC1 gyrokinetic simulation is analyzed to understand the three-dimensional spatial structure and the radial propagation of blob-filaments generated by quasisteady turbulence in the tokamak edge pedestal and scrape-off layer plasma. Spontaneous toroidal flows vary in the poloidal direction and shear the filaments within a flux surface resulting in a structure that varies in the parallel direction. Here, this parallel structure allows the curvature and grad-B induced polarization charge density to be shorted out via parallel electron motion. As a result, it is found that the blob-filament radial velocity is significantly reduced from estimates which neglect parallel electron kinetics, broadly consistent with experimental observations. Conditions for when this charge shorting effect tends to dominate blob dynamics are derived and compared with the simulation.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

MARTINI-Compatible Coarse-Grained Model for the Mesoscale Simulation of Peptoids

Peptoids (poly-N-substituted glycines) are a class of synthetic polymers that are regioisomers of peptides (poly-C-substituted glycines), in which the point of side-chain connectivity is shifted from the backbone C to the N atom. Peptoids have found diverse applications as peptidomimetic drugs, protein mimetic polymers, surfactants, and catalysts. Computational modeling is valuable in the understanding and design of peptoid-based nanomaterials. In this work, we report the bottom-up parameterization of coarse-grained peptoid force fields based on the MARTINI peptide force field against all-atom peptoid simulation data. Our parameterization pipeline iteratively refits coarse-grained bonded interactions using iterative Boltzmann inversion and nonbonded interactions by matching the potential of mean force for chain extension. We assure good sampling of the amide bond cis/trans isomerizations in the all-atom simulation data using parallel bias metadynamics. We develop coarse-grained models for two representative peptoids—polysarcosine (poly(N-methyl glycine)) and poly(N-((4-bromophenyl)ethyl)glycine)—and show their structural and thermodynamic properties to be in excellent accord with all-atom calculations but up to 25-fold more efficient and compatible with MARTINI force fields. Here, this work establishes a new rigorously parameterized coarse-grained peptoid force field for the understanding and design of peptoid nanomaterials at length and time scales inaccessible to all-atom calculations.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Identifying the Growth Phase of Magnetic Reconnection Using Pressure‐Strain Interaction

Abstract Magnetic reconnection often initiates abruptly and then rapidly progresses to a nonlinear quasi‐steady state. While satellites frequently detect reconnection events, ascertaining whether the system has achieved steady‐state or is still evolving in time remains challenging. Here, we propose that the relatively rapid opening of the reconnection separatrices within the electron diffusion region serves as an indicator of the growth phase of reconnection. The opening of the separatrices is produced by electron flows diverging away from the neutral line downstream of the X‐line and flowing around a dipolarization front. This flow pattern leads to characteristic spatial structures in the electron pressure‐strain interaction that could be a useful indicator for the growth phase of a reconnection event. We employ two‐dimensional particle‐in‐cell numerical simulations of anti‐parallel magnetic reconnection to validate this prediction. We find that the signature discussed here, alongside traditional reconnection indicators, can serve as a marker of the growth phase. This signature is potentially accessible using multi‐spacecraft single‐point measurements, such as with NASA's Magnetospheric Multiscale satellites in Earth's magnetotail. Applications to other settings where reconnection occurs are also discussed.

Barbhuiya, M. Hasan [Department of Physics and Ast↗

Modeling of small tungsten dust grains in EAST tokamak with NDS-BOUT ++

In order to investigate the transport of small dusts as well as their evolution property along their trajectories, the NDS module is developed under the BOUT++ framework, a highly desirable C++ code package to perform parallel plasma fluid simulations with an arbitrary number of equations in three-dimensional curvilinear coordinates. Due to the severe dust ablation in fusion plasmas, the dust size would decrease from micrometer to nanometer, resulting in impurities. Small dusts in the simulations here are specified as tungsten spheres with the radii on or below the order of submicrometer. The Rayleigh limit is included in the charging process when the dust is ablated to the droplet phase. The simulation results from the NDS module show that a 200 nm radius spherical tungsten dust originated from upper divertor region of EAST Tokamak is ablated completely due to the intense heating from the incoming plasma inside the core region, well consistent with the CCD footage of EAST shot # 81459. Furthermore it is found that the magnetic field dominates the dust transport when the dust radius is below 100 nm during the ablation along the trajectory. Our simulations predict that a 10 nm radius spherical tungsten dust injected from the inner midplane is well constrained by the magnetic field, and it reaches the inner divertor target with a velocity on the order of km/s.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Magnetic island formation and rotation braking induced by low-Z impurity penetration in an EAST plasma

Abstract Recent observations of the successive formations of the 4 / 1 , 3 / 1 , and 2 / 1 magnetic islands as well as the subsequent braking of the 2 / 1 mode during a low- Z impurity penetration process in EAST experiments are well reproduced in our 3 D resistive MHD simulations. The enhanced parallel current perturbation induced by impurity radiation predominately contributes to the tearing mode growth, and the 2 / 1 island rotation is mainly damped by the impurity accumulation as results of the influence from high n modes.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Online and Scalable Data Compression Pipeline with Guarantees on Quantities of Interest

Data compression is becoming critical for data-intensive scientific applications. Scientists require compression techniques that accurately preserve derived quantities of interest (QoIs). Prior work has shown that a pipeline can be built to guarantee error on the primary data (PD) within user-defined bounds and achieve near-floating point QoI errors. In this paper, we present novel computational approaches for accelerating the pipeline and demonstrate results that enable concurrent execution of compression in parallel with the simulation nodes. This allows compression, including the writing of the required compression data, for the previous time step to be completed while the simulation proceeds with the current time step. Overall, the approach presented in this paper results in a 6–8 times improvement in computational overhead compared to previous work. These results were obtained using data generated by a large-scale fusion code called XGC, which produces hundreds of terabytes of data in a single day.

Banerjee, Tania↗

Evaluating Trade-offs in Potential Exascale Interconnect Technologies

This report details work to study trade-offs in topology and network bandwidth for potential interconnects in the exascale (2021-2022) timeframe. The work was done using multiple interconnect models across two parallel discrete event simulators. Results from each independent simulator are shown and discussed and the areas of agreement and disagreement are explored.

97 MATHEMATICS AND COMPUTING↗

Biologically Inspired Interception on an Unmanned System

Borrowing from nature, neural-inspired interception algorithms were implemented onboard a vehicle. To maximize success, work was conducted in parallel within a simulated environment and on physical hardware. The intercept vehicle used only optical imaging to detect and track the target. A successful outcome is the proof-of-concept demonstration of a neural-inspired algorithm autonomously guiding a vehicle to intercept a moving target. This work tried to establish the key parameters for the intercept algorithm (sensors and vehicle) and expand the knowledge and capabilities of implementing neural-inspired algorithms in simulation and on hardware.

42 ENGINEERING↗

Understanding Phase and Interfacial Effects of Spall Fracture in Additively Manufactured Ti-5Al-5V-5Mo-3Cr

Additive manufactured Ti-5Al-5V-5Mo-3Cr (Ti-5553) is being considered as an AM repair material for engineering applications because of its superior strength properties compared to other titanium alloys. Here, we describe the failure mechanisms observed through computed tomography, electron backscatter diffraction (EBSD), and scanning electron microscopy (SEM) of spall damage as a result of tensile failure in as-built and annealed Ti-5553. We also investigate the phase stability in native powder, as-built and annealed Ti-5553 through diamond anvil cell (DAC) and ramp compression experiments. We then explore the effect of tensile loading on a sample containing an interface between a Ti-6Al-V4 (Ti-64) baseplate and additively manufactured Ti-5553 layer. Post-mortem materials characterization showed spallation occurred in regions of initial porosity and the interface provides a nucleation site for spall damage below the spall strength of Ti-5553. Preliminary peridynamics modeling of the dynamic experiments is described. Finally, we discuss further development of Stochastic Parallel PARticle Kinteic Simulator (SPPARKS) Monte Carlo (MC) capabilities to include the integration of alpha (α)-phase and microstructural simulations for this multiphase titanium alloy.

36 MATERIALS SCIENCE↗

Studies of Quark Transport and Hadronization in Nuclei

In this project, we conducted the first measurement of di‑hadron azimuthal correlations in deep inelastic scattering (DIS) off nuclei using the CLAS detector at Jefferson Lab. Using 5 GeV electron‑beam data collected on deuterium, carbon, iron, and lead targets, we extracted di‑pion correlation functions over a broad kinematic range. The results show a monotonic broadening of the correlation peak with increasing nuclear mass, along with pronounced dependencies on the pions’ kinematics. Separately, we implemented an algorithm based on the Kalman filter that achieved the first complete alignment of the CLAS12 central tracking system. In parallel, we developed simulations, algorithms, and performance studies that informed the conceptual designs of the forward hadronic calorimeter Insert and the Zero Degree Calorimeter, both of which are now included in the ePIC detector baseline for the forthcoming Electron Ion Collider.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Fast Muon Capture Monitoring in Mu2e with the CAPHRI Detector

The Mu2e experiment at Fermilab will search for the charged lepton flavor violating (CLFV) process of a neutrinoless muon-to-electron conversion in the field of an aluminum nucleus. Reaching the experiment’s target sensitivity requires precise normalization of the physics signal through accurate monitoring of the muon capture rate on the stopping target. For this purpose, the Calorimeter Precise High-Resolution Intensity detector (CAPHRI) has been developed. The detector is composed of four LYSO crystals installed in the upstream disk of the Mu2e calorimeter and read out with the standard calorimeter readout. CAPHRI measures the muon capture rate by detecting the characteristic 1.8~MeV gamma emission line of the $^{27}Al(\mu^−, \nu n \gamma) ^{26}Mg$ nuclear reaction. The fast, precise response enables injection-by-injection monitoring of proton beam intensity fluctuations. We report on the commissioning and performance characterization of CAPHRI. The response of each channel is calibrated at two SiPM overvoltages using both the intrinsic self-emission of the LYSO crystals and cosmic ray signals. In parallel, Monte Carlo simulations are used to evaluate the detector acceptance and the expected signal-to-background ratio under realistic running conditions. Preliminary results show a crystal light yield consistent with expectations and a channel inter-calibration at the 2--4% level. Simulation studies indicate that the detector acceptance and background rejection satisfy the requirements for physics operations, with about 1000 detected events per beam injection at a beam power of 1.5~kW. These results demonstrate that CAPHRI is an effective tool for beam monitoring and signal normalization in Mu2e.

Ciccarella, V. [Frascati; U. Rome La Sapienza (mai↗