Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,711 records · Page 95

Breaking CFD Bottlenecks in Gas-Turbine Flow-Path Design

New ideas are forthcoming to break existing bottlenecks in using CFD during design. CAD-based automated grid generation. Multi-disciplinary use of embedded, overset grids to eliminate complex gridding problems. Use of time-averaged detached-eddy simulations as norm instead of "steady" RANS to include effects of self-excited unsteadiness. Combined GPU/Core parallel computing to provide over an order of magnitude increase in performance/price ratio. Gas-turbine applications are shown here but these ideas can be used for other Air Force, Navy, and NASA applications.

Davis, Roger L.↗

Evidence of Multiple Reconnection Lines at the Magnetopause from Cusp Observations

Recent global hybrid simulations investigated the formation of flux transfer events (FTEs) and their convection and interaction with the cusp. Based on these simulations, we have analyzed several Polar cusp crossings in the Northern Hemisphere to search for the signature of such FTEs in the energy distribution of downward precipitating ions: precipitating ion beams at different energies parallel to the ambient magnetic field and overlapping in time. Overlapping ion distributions in the cusp are usually attributed to a combination of variable ion acceleration during the magnetopause crossing together with the time-of-flight effect from the entry point to the observing satellite. Most "step up" ion cusp structures (steps in the ion energy dispersions) only overlap for the populations with large pitch angles and not for the parallel streaming populations. Such cusp structures are the signatures predicted by the pulsed reconnection model, where the reconnection rate at the magnetopause decreased to zero, physically separating convecting flux tubes and their parallel streaming ions. However, several Polar cusp events discussed in this study also show an energy overlap for parallel-streaming precipitating ions. This condition might be caused by reopening an already reconnected field line, forming a magnetic island (flux rope) at the magnetopause similar to that reported in global MHD and Hybrid simulations

flux transfer events↗

Development of a Robust and Efficient Parallel Solver for Unsteady Turbomachinery Flows

The traditional design and analysis practice for advanced propulsion systems relies heavily on expensive full-scale prototype development and testing. Over the past decade, use of high-fidelity analysis and design tools such as CFD early in the product development cycle has been identified as one way to alleviate testing costs and to develop these devices better, faster and cheaper. In the design of advanced propulsion systems, CFD plays a major role in defining the required performance over the entire flight regime, as well as in testing the sensitivity of the design to the different modes of operation. Increased emphasis is being placed on developing and applying CFD models to simulate the flow field environments and performance of advanced propulsion systems. This necessitates the development of next generation computational tools which can be used effectively and reliably in a design environment. The turbomachinery simulation capability presented here is being developed in a computational tool called Loci-STREAM [1]. It integrates proven numerical methods for generalized grids and state-of-the-art physical models in a novel rule-based programming framework called Loci [2] which allows: (a) seamless integration of multidisciplinary physics in a unified manner, and (b) automatic handling of massively parallel computing. The objective is to be able to routinely simulate problems involving complex geometries requiring large unstructured grids and complex multidisciplinary physics. An immediate application of interest is simulation of unsteady flows in rocket turbopumps, particularly in cryogenic liquid rocket engines. The key components of the overall methodology presented in this paper are the following: (a) high fidelity unsteady simulation capability based on Detached Eddy Simulation (DES) in conjunction with second-order temporal discretization, (b) compliance with Geometric Conservation Law (GCL) in order to maintain conservative property on moving meshes for second-order time-stepping scheme, (c) a novel cloud-of-points interpolation method (based on a fast parallel kd-tree search algorithm) for interfaces between turbomachinery components in relative motion which is demonstrated to be highly scalable, and (d) demonstrated accuracy and parallel scalability on large grids (approx 250 million cells) in full turbomachinery geometries.

West, Jeff↗

Hybrid particle-in-cell simulations of electromagnetic coupling and waves from streaming burst debris

Various systems can be modeled as a point-like explosion of ionized debris into a magnetized, collisionless background plasma-including astrophysical examples, active experiments in space, and laser-driven laboratory experiments. Debris streaming from the explosion parallel to the magnetic field may drive multiple resonant and non-resonant ion-ion beam instabilities, some of which can efficiently couple the debris energy to the background and may even support the formation of shocks. We present a large-scale hybrid (kinetic ions + fluid electrons) particle-in-cell simulation, extending hundreds of ion inertial lengths from a 3D explosion, that resolves these instabilities. We show that the character of these instabilities differs notably from the 1D equivalent by the presence of unique transverse structure. Additional 2D simulations explore how the debris beam length, width, density, and speed affect debris-background coupling, with implications for the generation of quasi-parallel shocks.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Scalable line and plane relaxation in a parallel structured multigrid solver

The efficient solution of sparse, linear systems that arise through the discretization of partial differential equations remains a key challenge for a range of high performance scientific simulations. One approach for reducing data movement and improving performance is by exposing and exploiting structure in a problem through the use of robust structured multilevel solvers. By choosing coarsening that preserves the structure of the problem, these methods maintain efficient structured computation and communication throughout the multigrid hierarchy. However, when coarsening is not permitted to be dependent on the operator, anisotropy must be addressed by the smoother — producing error compatible for coarse-grid correction with structured coarsening. Here, the components required in a scalable parallel structured solver are described with a focus on memory and communication efficiency of robust smoothers. While the implementation of communication and memory reduction techniques in smoothers integrated in a complete 3D solver present a significant engineering challenge, a novel approach is proposed that addresses these challenges systematically through a change to the solver’s execution model. Enabled by user-level threading paired with a set of data and communication abstractions, this approach permits seamless aggregation of communication in plane smoothers — directly reusing code for a 2D distributed multilevel cycle. Results show an effective reduction in communication costs for coarse-grid problems, and result in a speedup of 8.7x in smoothing routines shown in Fig. 12 using this approach. This produces a significant improvement to strong scalability while maintaining favorable weak scaling behavior. Finally, a parallel scaling study using a series of refined meshes is included that demonstrates the effectiveness of this approach in an application of interest.

97 MATHEMATICS AND COMPUTING↗

Core-edge integrated predictive studies of ST40 and NSTX plasmas with the scrape-off layer box model

The ability to model the interplay between the core and edge of tokamak plasmas is crucial to designing both the plasma operating scenario of a fusion pilot plant and the design of the tokamak itself. Scrape-off-layer (SOL) models that are tailored to integrated scenario modeling need to have fast turn-around time and minimal computational burden to enable wide parameter-space coverage for design scoping. The SOL 0-D Box model is a reduced SOL model based on global power and particle balance that captures the essential physics of SOL transport with little computational cost. The usage of the 0-D Box model in core-edge coupled simulations has been demonstrated in both interpretive and predictive modes on a variety of devices. This paper presents a sensitivity study of the 0-D Box model to the input SOL heat-flux width for an ST40 plasma. This study demonstrates that accurate prediction of this width is crucial to predicting global performance parameters of a plasma scenario, such as energy confinement time and flux consumption. We also present an extension of the Box model to 1-D to allow for parallel variation of plasma parameters along the magnetic field lines. The 1-D Box model is then compared with SOLPS-ITER simulations of an NSTX plasma. Advantages and limitations of the Box model are discussed, and future directions are outlined.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Optimizing Error-Bounded Lossy Compression for Scientific Data by Dynamic Spline Interpolation

Today's scientific simulations are producing vast volumes of data that cannot be stored and transferred efficiently because of limited storage capacity, parallel I/O bandwidth, and network bandwidth. The situation is getting worse over time because of the ever-increasing gap between relatively slow data transfer speed and fast-growing computation power in modern supercomputers. Error-bounded lossy compression is becoming one of the most critical techniques for resolving the big scientific data issue, in that it can significantly reduce the scientific data volume while guaranteeing that the reconstructed data is valid for users because of its compression-error-bounding feature. In this paper, we present a novel error-bounded lossy compressor based on a state-of-the-art prediction-based compression framework. Our solution exhibits substantially better compression quality than all of the existing error-bounded lossy compressors, with comparable compression speed. Specifically, our contribution is threefold. (1) We provide an in-depth analysis of why the best-existing prediction-based lossy compressor can only minimally improve the compression quality. (2) We propose a dynamic spline interpolation approach with a series of optimization strategies that can significantly improve the data prediction accuracy, substantially improving the compression quality in turn. (3) We perform a thorough evaluation using six real-world scientific simulation datasets across different science domains to evaluate our solution vs. all other related works. Experiments show that the compression ratio of our solution is higher than that of the second-best lossy compressor by 20%similar to 460% with the same error bound in most of the cases.

Zhao, Kai↗

Stability-preserving Lossy Compression for Large-scale Partial Differential Equations

Checkpoint/Restart (C/R) strategies are vital for fault tolerance in PDE-based scientific simulations, yet traditional checkpointing incurs significant I/O overhead. Lossy compression offers a scalable solution by reducing checkpoint data size, but conventional methods often lack control over physical invariants (e.g., energy), leading to instability such as oscillations or divergence in Partial Differential Equations (PDE) systems. This paper introduces a stability-preserving compression approach tailored for PDE simulations by explicitly controlling kinetic and potential energy perturbations to ensure stable restarts. Extensive experiments conducted across diverse PDE configurations demonstrate that our method maintains numerical stability with minimal error magnification—even across multiple checkpoint-restart cycles—outperforming state-of-the-art lossy compressors. Parallel evaluations on the Frontier supercomputer show up to 8.4× improvement in checkpoint write performance and 6.3× in read performance, while maintaining relative L2 errors ∼ 2e-6 throughout continued simulation. These results provide practical guidance for balancing compression accuracy, stability, and computational efficiency in large-scale PDE applications.

Gong, Qian [ORNL] (ORCID:0000000235704142)↗

Correlation Function Approach for Diffusion in Confined Geometries

This paper describes a formalism for extracting spatially varying transport coefficients from simulations of a molecular fluid in a nano channel. This approach is applied to self-diffusion of a Lennard-Jones fluid confined between two parallel surfaces. A numerical grid is laid over the domain confining the fluid, and fluid properties are projected onto the grid cells. The time correlation functions between properties in different grid cells are calculated and can be used as the basis for a fitting procedure for extracting spatially varying diffusion coefficients from the simulation. Results for the Lennard-Jones system show that transport behavior varies sharply near the liquid-solid boundary and that the changes depend on the details of the liquid-solid interaction. A quantitative difference between the reduced and detailed models is discussed. It is found that the difference could be associated with assumptions about the form of the transport equations at molecular scales in lieu of problems with the method itself. The study suggests that this approach to fitting molecular simulations to continuum equations may guide the development of appropriate coarse-grained equations to model transport phenomena at nanometer scales.

Nanoscale flow, Transport, molecular simulation↗

Intrinsic Toroidal Rotation Driven by Turbulent and Neoclassical Processes in Tokamak Plasmas from Global Gyrokinetic Simulations

Gyrokinetic tokamak plasmas can exhibit intrinsic toroidal rotation driven by the residual stress. While most studies have attributed the residual stress to the parallel-momentum flux from the turbulent E × B motion, the parallel-momentum flux from the drift-orbit motion (denoted $Π^D_\parallel$) and the E × B-momentum flux from the E × B motion (denoted $Π_{E×B}$) are often neglected. Here, we use the global total-f gyrokinetic code XGC to study the residual stress in the core and the edge of a DIII-D H-mode plasma. Numerical results show that both $Π^D_\parallel$ and $Π_{E×B}$ make up a significant portion of the residual stress. In particular, $Π^D_\parallel$ in the core is higher than the collisional neoclassical level in the presence of turbulence, while in the edge it represents an outflux of countercurrent momentum even without turbulence. Using a recently developed “orbit-flux” formulation, we show that the higher-than-neoclassical-level $Π^D_\parallel$ in the core is driven by turbulence, while the outflux of countercurrent momentum from the edge is mainly due to collisional ion orbit loss. In conclusion, these results suggest that $Π^D_\parallel$ and $Π_{E×B}$ can be important for the study of intrinsic toroidal rotation.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Observation and Numerical Simulation of Cold Ions Energized by EMIC Waves

This is the first report of significant energization (up to 7,000 eV) of low-energy He + ions, which occurred simultaneously with H-band electromagnetic ion cyclotron (EMIC) wave activity, in a direction mostly perpendicular to the ambient magnetic field. The event was detected by the Arase satellite in the dayside plasmatrough region off the magnetic equator on 15 May 2019. The peak energy of the He + flux enhancements is mostly above 1,000 eV. At some interval, the He + ions are energized up to ~7,000 eV. The H-band waves are excited in a frequency band between the local crossover and helium gyrofrequencies and are close to a linear polarization state with weakly left-handed or right-handed polarization. The normal angle of the waves exhibits significant variation between 0° and 80°, indicating a non-parallel propagation. Here, we run a hybrid code with parameters estimated from the Arase observations to examine the He + energization. The simulations show that cold He + ions are energized up to more than 1,000 eV, similar to the spacecraft observations. From the analysis of the simulated wave fields and cold plasma motions, we found that the ratio of the wave frequency to He + gyrofrequency is a primary factor for transverse energization of cold He + ions. As a consequence of the numerical analysis, we suggest that the significant transverse energization of He + ions observed by Arase is attributed to H-band EMIC waves excited near the local helium gyrofrequency.

79 ASTRONOMY AND ASTROPHYSICS↗

Multi-dimensional high order essentially non-oscillatory finite difference methods in generalized coordinates

The nonlinear stability of compact schemes for shock calculations is investigated. In recent years compact schemes were used in various numerical simulations including direct numerical simulation of turbulence. However to apply them to problems containing shocks, one has to resolve the problem of spurious numerical oscillation and nonlinear instability. A framework to apply nonlinear limiting to a local mean is introduced. The resulting scheme can be proven total variation (1D) or maximum norm (multi D) stable and produces nice numerical results in the test cases. The result is summarized in the preprint entitled 'Nonlinearly Stable Compact Schemes for Shock Calculations', which was submitted to SIAM Journal on Numerical Analysis. Research was continued on issues related to two and three dimensional essentially non-oscillatory (ENO) schemes. The main research topics include: parallel implementation of ENO schemes on Connection Machines; boundary conditions; shock interaction with hydrogen bubbles, a preparation for the full combustion simulation; and direct numerical simulation of compressible sheared turbulence.

Shu, Chi-Wang↗

Visualization Co-Processing of a CFD Simulation

OVERFLOW, a widely used CFD simulation code, is combined with a visualization system, pV3, to experiment with an environment for simulation/visualization co-processing on a SGI Origin 2000 computer(O2K) system. The shared memory version of the solver is used with the O2K 'pfa' preprocessor invoked to automatically discover parallelism in the source code. No other explicit parallelism is enabled. In order to study the scaling and performance of the visualization co-processing system, sample runs are made with different processor groups in the range of 1 to 254 processors. The data exchange between the visualization system and the simulation system is rapid enough for user interactivity when the problem size is small. This shared memory version of OVERFLOW, with minimal parallelization, does not scale well to an increasing number of available processors. The visualization task takes about 18 to 30% of the total processing time and does not appear to be a major contributor to the poor scaling. Improper load balancing and inter-processor communication overhead are contributors to this poor performance. Work is in progress which is aimed at obtaining improved parallel performance of the solver and removing the limitations of serial data transfer to pV3 by examining various parallelization/communication strategies, including the use of the explicit message passing.

Vaziri, Arsi↗

Relative alignment between magnetic fields and molecular gas structure in molecular clouds

Here, we compare the structure of synthetic dust polarization with synthetic molecular line emission from radiative transfer calculations using a three-dimensional, turbulent collapsing-cloud magnetohydrodynamics simulation. The histogram of relative orientation (HRO) technique and the projected Rayleigh statistic (PRS) are considered. In our trans-Alfvénic (more strongly magnetized) simulation, there is a transition to perpendicular alignment at densities above ~4 × 10 3 cm –3 . This transition is recovered in most of our synthetic observations of optically thin molecular tracers; however, for 12 CO it does not occur and the PRS remains in parallel alignment across the whole observer space. We calculate the physical depth of the optical depth τ = 1 surface and find that for 12 CO it is largely located in front of the cloud mid-plane, suggesting that 12 CO is too optically thick and instead mainly probes low-volume density gas. In our super-Alfvénic simulation, the magnetic field becomes significantly more tangled, and all observed tracers tend towards no preference for perpendicular or parallel alignment. An observable difference in alignment between optically thin and optically thick tracers may indicate the presence of a dynamically important magnetic field, though there is some degeneracy with viewing angle. We convolve our data with a Gaussian beam and compare it with HRO results of the Vela C molecular cloud. We find good agreement between these results and our sub-Alfvénic simulations when viewed with the magnetic field in the plane of the sky (especially when sensitivity limitations are considered), though the observations are also consistent with an intermediately inclined magnetic field.

79 ASTRONOMY AND ASTROPHYSICS↗

Scale Dependence of Cirrus Horizontal Heterogeneity Effects on TOA Measurements: MODIS Brightness Temperatures in the Thermal Infrared - Part I

This paper presents a study on the impact of cirrus cloud heterogeneities on MODIS simulated thermal infrared (TIR) brightness temperatures (BTs) at the top of the atmosphere (TOA) as a function of spatial resolution from 50 meters to 10 kilometers. A realistic 3-D (three-dimensional) cirrus field is generated by the 3DCLOUD model (average optical thickness of 1.4, cloudtop and base altitudes at 10 and 12 kilometers, respectively, consisting of aggregate column crystals of D (sub eff) equals 20 microns), and 3-D thermal infrared radiative transfer (RT) is simulated with the 3DMCPOL (3-D Monte Carlo Polarized) code. According to previous studies, differences between 3-D BT computed from a heterogenous pixel and 1-D (one-dimensional) RT computed from a homogeneous pixel are considered dependent at nadir on two effects: (i) the optical thickness horizontal heterogeneity leading to the plane-parallel homogeneous bias (PPHB); and the (ii) horizontal radiative transport (HRT) leading to the independent pixel approximation error (IPAE). A single but realistic cirrus case is simulated and, as expected, the PPHB mainly impacts the low-spatial resolution results (above approximately 250 meters), with averaged values of up to 5-7 K (thousand), while the IPAE mainly impacts the high-spatial resolution results (below approximately 250 meters) with average values of up to 1-2 K (thousand). A sensitivity study has been performed in order to extend these results to various cirrus optical thicknesses and heterogeneities by sampling the cirrus in several ranges of parameters. For four optical thickness classes and four optical heterogeneity classes, we have found that, for nadir observations, the spatial resolution at which the combination of PPHB and HRT effects is the smallest, falls between 100 and 250 meters. These spatial resolutions thus appear to be the best choice to retrieve cirrus optical properties with the smallest cloud heterogeneity-related total bias in the thermal infrared. For off-nadir observations, the average total effect is increased and the minimum is shifted to coarser spatial resolutions.

TOA↗

Impact of plasma density/collisionality on divertor heat flux width

Both ASDEX-Upgrade (AUG) data and the generalized HD (GHD) model showed that the scrape-off width broadens as the density/collisionality increases [1, 2]. A series of BOUT++ transport simulations are performed to study the physics of the scaling characteristics of the divertor heat flux width vs density/collisionality via a plasma density scan with either fixed pressure profile or fixed temperature profile inside separatrix. Additionally, the simulations show that even in the drift dominated regime, the divertor heat flux width can be broadened due to the transition of the SOL residence time from the parallel particle flow time to the enhanced parallel conduction time as the collisionality/density increases as posited in the GHD model. In addition, the heat flux width is found to be proportional to the square root of ion mass for low collisionality while it has a weakly dependence on ion mass for high collisionality. Furthermore, our simulations show that as the density increases, the radial electric field (E r ) well shallows, which potentially weakens E r × B flow shear stabilization of turbulence at high density.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Solidification and crystallographic texture modeling of laser powder bed fusion Ti-6Al-4V using finite difference-monte carlo method

Laser powder bed fusion (LPBF) additive manufacturing makes near-net-shaped parts with reduced material cost and time, rising as a promising technology to fabricate Ti-6Al-4V, a widely used titanium alloy in aerospace and medical industries. However, LPBF Ti-6Al-4V parts produced with 67° rotation between layers, a scan strategy commonly used to reduce microstructure and property inhomogeneity, have varying grain morphologies and weak crystallographic textures that change depending on processing parameters. Here, this study predicts LPBF Ti-6Al-4V solidification at three energy levels using a finite difference-Monte Carlo method and validates the simulations with large-area electron backscatter diffraction (EBSD) scans. The developed model accurately shows that a <001> texture forms at low energy and a <111> texture occurs at higher energies parallel to the build direction but with a lower strength than the textures observed from EBSD. A validated and well-established method of combining spatial correlation and general spherical harmonics representation of texture is developed to calculate a difference score between simulations and experiments. The quantitative comparison enables effective fine-tuning of nucleation density (N 0 ) input, which shows a nonlinear relationship with increasing energy level. Future improvements in texture prediction code and a more comprehensive study of N 0 with different energy levels will further advance the optimization of LPBF Ti-6Al-4V components. These developments contribute a novel understanding of crystallographic texture formation in LPBF Ti-6Al-4V, the development of robust model validation and calibration pipeline methodologies, and provide a platform for mechanical property prediction and process parameter optimization.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Progress of the NASA/USGS Lunar Regolith Simulant Project

Beginning in 2004 personnel at MSFC began serious efforts to develop a new generation of lunar simulants. The first two products were a replication of the previous JSC-1 simulant under a contract to Orbitec and a major workshop in 2005 on future simulant development. Beginning in 2006 the project refocused its efforts and approached simulant development in a new and more comprehensive manner, examining new approaches in simulant development and ways to more accurately compare simulants to actual lunar materials. This led to a multi-year effort with five major tasks running in parallel. The five tasks are Requirements, Lunar Analysis, Process Development, Feed Stocks, and Standards. Major progress has been made in all five areas. A substantial draft of a formal requirements document now exists and has been largely stable since 2007. It does evolve as specific details of the standards and Lunar Analysis efforts proceed. Lunar Analysis has turned out to be vastly more difficult than anticipated. After great effort to mine existing published and gray literature, the team has realized the necessity of making new measurements of the Apollo samples, an effort that is currently in progress. Process development is substantially ahead of expectations in 2006. It is now practical to synthesize glasses of appropriate composition and purity. It is also possible to make agglutinate particles in significant quantities. A series of minerals commonly found on the Moon has been synthesized. Separation of mineral constituents from starting rock material is also proceeding. Customized grinding and mixing processes have been developed and tested are now being documented. Identification and development of appropriate feedstocks has been both easier and more difficult than anticipated. The Stillwater Mining Company, operating in the Stillwater layered mafic intrusive complex of Montana, has been an amazing resource for the project, but finding adequate sources for some of the components remains a difficult problem. For example the ratio of clino- to ortho-pyroxenes in the Stillwater is not an exact match for lunar materials. One of the sources being examined as an alternative pyroxene source is the Bushveld Complex in South Africa. Standards have been a major success for the project. The Figure of Merit algorithms have been created, tested, and are being considered for an ISO standard. Agreement has been reached in the community about how to make many of the critical measurements. There remains much work to do: (1) driving down the cost of simulants remains a major obstacle; (2) documentation and cost data analysis have not kept up with progress; (3) educating users in the complexity of the lunar regolith and the use of simulants remains a major task. In summary the project has made enormous progress and is successfully placing simulant development and use on a rigorous, scientifically defensible, engineering basis.

Rickman, Doug↗