Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,765 records · Page 98

Massively parallel processor

A brief description is given of the Massively Parallel Processor (MPP). Major applications of the MPP are in the area of image processing (where the operands are often very small integers) from very high spatial resolution passive image sensors, signal processing of radar data, and numerical modeling simulations of climate. The system can be programmed in assembly language or a high level language. Information on background, status, architecture, programming, hardware reliability, applications, and the MPP's development as a national resource for parallel algorithm research are presented in outline form.

Source record↗

Particle Acceleration and Associated Emission from Relativistic Shocks

Five talks consist of a research program consisting of numerical simulations and theoretical development designed to provide an understanding of the emission from accelerated particles in relativistic shocks. The goal of this lecture is to discuss the particle acceleration, magnetic field generation, and radiation along with the microphysics of the shock process in a self-consistent manner. The discussion involves the collisionless shocks that produce emission from gamma-ray bursts and their afterglows, and producing emission from supernova remnants and AGN relativistic jets. Recent particle-in-cell simulation studies have shown that the Weibel (mixed mode two-stream filamentation) instability is responsible for particle (electron, positron, and ion) acceleration and magnetic field generation in relativistic collisionless shocks. 3-D RPIC code parallelized with MPI has been used to investigate the dynamics of collisionless shocks in electron-ion and electron-positron plasmas with and without initial ambient magnetic fields. In this lecture we will present brief tutorials of RPIC simulations and RMHD simulations, a brief summary of recent RPIC simulations, mechanisms of particle acceleration in relativistic shocks, and calculation of synchrotron radiation by tracing particles. We will discuss on emission from the collisionless shocks, which will be calculated during the simulation by tracing particle acceleration self-consistently in the inhomogeneous magnetic fields generated in the shocks. In particular, we will discuss the differences between standard synchrotron radiation and the jitter radiation that arises in turbulent magnetic fields.

Nishkawa, Ken-Ichi↗

Parallel-in-Time Solution of Allen-Cahn Equations by Integrating Operator Learning into the Parareal Method

While recent advances in deep learning have shown promising efficiency gains in solving time-dependent partial differential equations (PDEs), matching the accuracy of conventional numerical solvers still remains a challenge. One strategy to improve the accuracy of deep learning-based solutions for time-dependent PDEs is to use the learned model as the coarse propagator in the Parareal method and a traditional numerical method as the fine solver. However, successful integration of deep learning into the Parareal method requires consistency between the coarse and fine solvers, particularly for PDEs exhibiting rapid changes such as sharp transitions. Here, to ensure this consistency, we propose using convolutional neural networks (CNNs) to learn the fully discrete time-stepping operator defined by the same numerical scheme employed as the fine solver. We demonstrate the effectiveness of the proposed method in solving the classical and mass-conservative Allen–Cahn (AC) equations. Through iterative updates in the Parareal algorithm, our approach achieves a significant computational speedup compared to traditional fine solvers while converging to high-accuracy solutions. Our results highlight that the proposed hybrid Parareal algorithm effectively accelerates simulations, particularly when implemented on multiple GPUs, and converges to the desired accuracy in only a few iterations. Another advantage of our method is that the CNN model is trained on trajectory-based data generated from random initial conditions, such that the trained model can be used to solve the AC equations with various initial conditions without retraining. This work demonstrates the potential of integrating neural network methods into parallel-in-time frameworks for efficient and accurate simulations of time-dependent PDEs.

97 MATHEMATICS AND COMPUTING↗

Coarse-grained simulation of colloidal self-assembly, cation exchange, and rheology in Na/Ca smectite clay gels

Knowledge Gap: The aggregation of clay minerals—layered silicate nanoparticles—strongly impacts fluid flow, solute migration, and solid mechanics in soils, sediments, and sedimentary rocks. Experimental and computational characterization of clay aggregation is inhibited by the delicate water-mediated nature of clay colloidal interactions and by the range of spatial scales involved, from 1 nm thick platelets to flocs with dimensions up to micrometers or more. Simulations: Using a new coarse-grained molecular dynamics (CGMD) approach, we predicted the microstructure, dynamics, and rheology of hydrated smectite (more precisely, montmorillonite) clay gels containing up to 2,000 clay platelets on length scales up to 0.1 μm. Further, simulations investigated the impact of simulation time, platelet diameters (6 to 25nm), and the ratio of Na to Ca exchangeable cations on the assembly of tactoids (i.e., stacks of parallel clay platelets) and larger aggregates (i.e., assemblages of tactoids). We analyzed structural features including tactoid size and size distribution, basal spacing, counterion distribution in the electrical double layer, clay association modes, and the rheological properties of smectite gels. Findings: Our results demonstrate new potential to characterize and understand clay aggregation in dilute suspensions and gels on a scale of thousands of particles with explicit representation of counterion clouds and with accuracy approaching that of all-atom molecular dynamics (MD) simulations. For example, our simulations predict the strong impact of Na/Ca ratio on clay tactoid formation and the shear-thinning rheology of clay gels.

42 ENGINEERING↗

Simulation of Sweep-Jet Flow Control, Single Jet and Full Vertical Tail

This work is a simulation technology demonstrator, of sweep jets used to suppress boundary layer separation and increase maximum achievable load coefficients. A sweep jet is a discrete Coanda jet that oscillates in the plane parallel to an aerodynamic surface. It injects mass and momentum in the approximate stream wise direction. It also generate turbulent eddies at the oscillation frequency, which are typically large relative to boundary layer turbulence, and which augmenting mixing across the boundary layer to attack flow separation. Simulations of a fluidic oscillator, the sweep jet emerging from the oscillator, and the suppression of boundary layer separation by an array of sweep jets are performed. Simulation results are compared to data from a dedicated CFD validation experiment of a single oscillator and its sweep jet, and from a study of a full-scale Boeing 757 vertical tail augmented with an array of sweep jets.2, 20 A critical step in the work is the development of realistic time-dependent sweep-jet in flow boundary conditions, derived from the results of the single-oscillator simulations, which create the sweep jets in the full-tail simulations. Simulations were performed using the Over flow CFD solver, with high-order spatial discretization and a range of turbulence modeling. Good results were obtained for all flows simulated, when suitable turbulence modeling was used.

flow control↗

Use of animal models for space flight physiology studies, with special focus on the immune system

Animal models have been used to study the effects of space flight on physiological systems. The animal models have been used because of the limited availability of human subjects for studies to be carried out in space as well as because of the need to carry out experiments requiring samples and experimental conditions that cannot be performed using humans. Experiments have been carried out in space using a variety of species, and included developmental biology studies. These species included rats, mice, non-human primates, fish, invertebrates, amphibians and insects. The species were chosen because they best fit the experimental conditions required for the experiments. Experiments with animals have also been carried out utilizing ground-based models that simulate some of the effects of exposure to space flight conditions. Most of the animal studies have generated results that parallel the effects of space flight on human physiological systems. Systems studied have included the neurovestibular system, the musculoskeletal system, the immune system, the neurological system, the hematological system, and the cardiovascular system. Hindlimb unloading, a ground-based model of some of the effects of space flight on the immune system, has been used to study the effects of space flight conditions on physiological parameters. For the immune system, exposure to hindlimb unloading has been shown to results in alterations of the immune system similar to those observed after space flight. This has permitted the development of experiments that demonstrated compromised resistance to infection in rodents maintained in the hindlimb unloading model as well as the beginning of studies to develop countermeasures to ameliorate or prevent such occurrences. Although there are limitations to the use of animal models for the effects of space flight on physiological systems, the animal models should prove very valuable in designing countermeasures for exploration class missions of the future.

Review↗

Parallelized solvers for heat conduction formulations

Based on multilevel partitioning, this paper develops a structural parallelizable solution methodology that enables a significant reduction in computational effort and memory requirements for very large scale linear and nonlinear steady and transient thermal (heat conduction) models. Due to the generality of the formulation of the scheme, both finite element and finite difference simulations can be treated. Diverse model topologies can thus be handled, including both simply and multiply connected (branched/perforated) geometries. To verify the methodology, analytical and numerical benchmark trends are verified in both sequential and parallel computer environments.

Padovan, Joe↗

Ion anisotropies in the magnetosheath

One- and two-dimensional initial value hybrid computer simulations are used with a magnetosheath parameter model to study the consequences of the growth of the mirror and ion cyclotron anisotropy instabilities. Magnetosheath observations have demonstrated inverse correlations between the proton and helium ion temperature anisotropies and the proton parallel beta. Using the maximum growth rate as a fitting pararamter, linear Vlasov instability theory for the proton cyclotron and helium cyclotron anisotropy instabilities reproduces both observed correlations. Furthermore, results from the asymptotic states of simulations of the ion cyclotron instabilities qualitatively reproduce both observations.

Gary, S. P.↗

Elevating SolTrace's Capabilities for the Next Generation of Concentrating Solar Analysis

SolTrace is an open-source Monte Carlo ray tracing software developed at NREL. SolTrace can characterize concentrating solar thermal (CST) collector optical performance and is CST technology agnostic. Shown in Fig. 1, SolTrace is a foundational tool in NREL's CST system and component modeling suite. SolTrace's generic surface elements can flexibly model novel collector and receiver designs to predict spatial and temporal flux distributions - critical to understand for CST component design, performance prediction, and system integration. Since its initial development, SolTrace has over 1,650 references on Google Scholar, over 9,800 downloads since 2017, and has served the CST research and development community as a benchmark of 3rd party verification. SolTrace provides users with many options for defining surface shape and boundaries. However, SolTrace provides limited documentation which can result in a steep learning curve for new users. Additionally, SolTrace lacks the computational performance required to evaluate optical performance of a CST system over the course of a year and/or iteratively over design parameters in a timely manner. To address this, we are working towards a new release of SolTrace that enables increased computational throughput by implementing ray tracing acceleration structures and enabling GPU parallelization. Additionally, we are working to improve SolTrace's usability, accessibility, and maintainability by (1) automating solar position time-dependent simulation processes, (2) creating general CST collector templates of grouped elements, (3) updating the user interface to better visualize model inputs and outputs, and (4) creating a user support network through forums, "how to" videos, and documentation.

14 SOLAR ENERGY↗

A view toward future fluid dynamics computing

Advances in computational fluid dynamics are paced by simulation methodology and computer resources. Examples of three-dimensional fluid dynamic simulations are presented to illustrate recent developments in equation modeling and numerical methods and to point out the need for increased computer power. Electronic technology dictates that to fill this need, computers will be based on parallel processing principles. The identification of parallelism in three dimensions is illustrated by examining an implicit, approximate-factorization approach to the Navier-Stokes equations. Finally, two computer concepts aimed at satisfying the demands of the three-dimensional Reynolds averaged Navier-Stokes simulations are discussed.

Bailey, F. R.↗

The Simplified Aircraft-Based Paired Approach With the ALAS Alerting Algorithm

This paper presents the results of an investigation of a proposed concept for closely spaced parallel runways called the Simplified Aircraft-based Paired Approach (SAPA). This procedure depends upon a new alerting algorithm called the Adjacent Landing Alerting System (ALAS). This study used both low fidelity and high fidelity simulations to validate the SAPA procedure and test the performance of the new alerting algorithm. The low fidelity simulation enabled a determination of minimum approach distance for the worst case over millions of scenarios. The high fidelity simulation enabled an accurate determination of timings and minimum approach distance in the presence of realistic trajectories, communication latencies, and total system error for 108 test cases. The SAPA procedure and the ALAS alerting algorithm were applied to the 750-ft parallel spacing (e.g., SFO 28L/28R) approach problem. With the SAPA procedure as defined in this paper, this study concludes that a 750-ft application does not appear to be feasible, but preliminary results for 1000-ft parallel runways look promising.

Perry, Raleigh B.↗

Large-Eddy Simulations of Idealized Atmospheric Boundary Layers Using Nalu-Wind

Accurate prediction of wind-plant performance relies, in part, on properly characterizing the turbulent atmospheric boundary layer (ABL) flow in which wind turbines operate. Large-eddy simulation (LES) is a powerful tool for simulating ABLs because it resolves the largest, most energetic scales of three-dimensional turbulent motions. Yet LES predictions are well known to depend on modeling choices such as grid resolution, numerical discretization schemes, and closures for unresolved scales of turbulence. Here, we evaluate how these choices influence predictions of ABL winds using Nalu-Wind, a wind-specific fork of the open-source, generalized, unstructured, massively parallel flow solver NaluCFD/Nalu.

17 WIND ENERGY↗

Cross-field electron diffusion due to the coupling of drift-driven microinstabilities

In this paper, the nonlinear interaction between kinetic instabilities driven by multiple ion beams and magnetized electrons is investigated. Electron diffusion across magnetic field lines is enhanced by the coupling of plasma instabilities. Here, a two-dimensional collisionless particle-in-cell simulation is performed accounting for singly and doubly charged ions in a cross-field configuration. Consistent with prior linear kinetic theory analysis and observations from coherent Thomson scattering experiments, the present simulations identify an ion-ion two-stream instability due to multiply charged ions (flowing in the direction parallel to the applied electric field) which coexists with the electron cyclotron drift instability (propagating perpendicular to the applied electric field and parallel to the ExB drift). Small-scale fluctuations due to the coupling of these naturally driven kinetic modes are found to be a mechanism that can enhance cross-field electron transport and contribute to the broadening of the ion velocity distribution functions.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Toward a 2D Local Implementation of Quantum Low-Density Parity-Check Codes

Geometric locality is an important theoretical and practical factor for quantum low-density parity-check (qLDPC) codes that affects code performance and ease of physical realization. For device architectures restricted to two-dimensional (2D) local gates, naively implementing the high-rate codes suitable for low-overhead fault-tolerant quantum computing incurs prohibitive overhead. In this work, we present an error-correction protocol built on a bilayer architecture that aims to reduce operational overheads when restricted to 2D local gates by measuring some generators less frequently than others. We investigate the family of bivariate-bicycle qLDPC codes and show that they are well suited for a parallel syndrome-measurement scheme using fast routing with local operations and classical communication (LOCC). Through circuit-level simulations, we find that in some parameter regimes, bivariate-bicycle codes implemented with this protocol have logical error rates comparable to the surface code while using fewer physical qubits. Published by the American Physical Society 2025

Berthusen, Noah (ORCID:0000000275862786)↗

Artificial Intelligence for Multiphysics Nuclear Design Optimization with Additive Manufacturing

The geometric flexibility of additively manufactured metals and ceramics generates a very large and open design space that requires advanced modeling and simulation tools for physics simulations and the rigorous definition of design problems. This effort deploys artificial intelligence (AI) and machine learning (ML) algorithms to understand the design space, evaluate potential designs, and more efficiently generate optimized results. The Transformational Challenge Reactor (TCR) program is leveraging advances in several scientific areas—including materials, manufacturing, sensors and control systems, data analytics, and high-fidelity modeling and simulation—to accelerate the design, manufacturing, qualification, and deployment of advanced nuclear energy systems. Through a manufacturing-informed design approach, the TCR program seeks to integrate digital data for rapid nuclear innovation; accelerate the adoption of advances in manufacturing, materials, and computational sciences for nuclear applications; and dramatically reduce deployment costs and timelines for new nuclear reactor technologies. This report documents efforts under the TCR program to leverage advanced modeling and simulation techniques driven by AI/ML algorithms on high-performance computing (HPC) systems to yield more optimized TCR core designs. A multiphysics ML surrogate model was developed to run on the HPC architectures. The surrogate model is trained on high-fidelity simulation data of coupled neutronics and thermofluidics and is used to quickly evaluate thousands of candidate core designs in parallel, which drives the evolution of the cooling channel shapes to minimize temperature peaking and material stress. Outcomes from these activities provide design information and feedback into the core design efforts.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Climate Model Output Rewriter

The Climate Model Output Rewriter (CMOR) software was first developed by LLNL’s PCMDI program in early 2000s and was formally released with v1.0 (July 2006), v2.0 (January 2011), and v3.1(June 2016). CMOR is used to produce Climate and Forecast Convention (http://cfconventions.org/) CF-compliant netCDF files, in the standard format required to satisfy the World Climate Research Program (WCRP) Coupled Model Intercomparison Project (CMIP). The software has been used across multiple phases of the Earth System Modeling (ESM) project CMIP (CMIP3, CMIP5, CMIP6, and planned use in CMIP7) along with numerous parallel projects focused on preparation observations for use in model evaluation (obs4MIPs) and forcing datasets (input4MIPs) to guide ESM simulations to meet strict experimental protocols. More information can be obtained from the CMOR website and code repositories: https://cmor.llnl.gov/; https://github.com/pcmdi/cmor; https://github.com/PCMDI/cmor3_documentation The ESM variable definitions used as input for CMOR can also be viewed in code repositories: https://github.com/PCMDI/cmip3-cmor-tables/; https://github.com/PCMDI/cmip5-cmor-tables/; https://github.com/PCMDI/cmip6-cmor-tables/

Mauzey, ChristopherF↗

Abort separation of the shuttle

A sensitivity analysis of the factors which affect a successful abort maneuver following a space shuttle launching is presented. Wind tunnel tests were conducted using optimum simulation techniques and data acquisition procedures. Static stability, dynamic stability, and local loads were investigated. It is concluded that parallel abort separation of the space shuttle components is possible at both high and low dynamic pressures. Successful separation is dependent upon configuration, Mach number, rocket exhaust impingement and relative position and attitude of the stages.

Decker, J. P.↗

The use of Ada in distributed simulations

The increasing need for detailed information about systems of continually growing complexity enhances steadily the demands regarding the employed models. The present investigation is concerned with work related to the development of high-performance computer hardware intended for the support of the real-time simulation of jet engines. The hardware is structured in the form of a network of communicating microprocessors running in parallel. The need for a higher-order language capability for programming such a network has led to the research considered in this study. Attention is given to the hardware which is being developed, an abstract model, programming language considerations, research considerations, research objectives, Ada tasks, Ada packages, the Ada model, the mapping of the model to the hardware, a precompiler example, and the advantages of Ada.

Collins, W. R.↗