Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “parallel simulation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 433 records · Page 24

Re-forming supercritical quasi-parallel shocks. II - Mechanism for wave generation and front re-formation

This paper continues the study of Thomas et al. (1990) in which hybrid simulations of quasi-parallel shocks were performed in one and two spatial dimensions. To identify the wave generation processes, the electromagnetic structure of the shock is examined by performing a number of one-dimensional hybrid simulations of quasi-parallel shocks for various upstream conditions. In addition, numerical experiments were carried out in which the backstreaming ions were removed from calculations to show their fundamental importance in reformation process. The calculations show that the waves are excited before ions can propagate far enough upstream to generate resonant modes. At some later times, the waves are regenerated at the leading edge of the interface, with properties like those of their initial interactions.

Winske, D.↗

The Cirrus Parcel Model Comparison Project

The cirrus Parcel Model Comparison Project involves the systematic comparison of current models of ice crystal nucleation and growth for specified, typical, cirrus cloud environments. In Phase 1 of the project reported here, simulated cirrus cloud microphysical properties are compared for situations of "warm" (-40 C) and "cold" (-60 C) cirrus subject to updrafts of 4, 20 and 100 centimeters per second, respectively. Five models are participating in the project. These models employ explicit microphysical schemes wherein the size distribution of each class of particles (aerosols and ice crystals) is resolved into bins. Simulations are made including both homogeneous and heterogeneous ice nucleation mechanisms. A single initial aerosol population of sulfuric acid particles is prescribed for all simulations. To isolate the treatment of the homogeneous freezing (of haze drops) nucleation process, the heterogeneous nucleation mechanism is disabled for a second parallel set of simulations. Qualitative agreement is found amongst the models for the homogeneous-nucleation-only simulations, e.g., the number density of nucleated ice crystals increases with the strength of the prescribed updraft. However, non-negligible quantitative differences are found. Systematic bias exists between results of a model based on a modified classical theory approach and models using an effective freezing temperature approach to the treatment of nucleation. Each approach is constrained by critical freezing data from laboratory studies. This information is necessary, but not sufficient, to construct consistent formulae for the two approaches. Large haze particles may deviate considerably from equilibrium size in moderate to strong updrafts (20-100 centimeters per second) at -60 C when the commonly invoked equilibrium assumption is lifted. The resulting difference in particle-size-dependent solution concentration of haze particles may significantly affect the ice nucleation rate during the initial nucleation interval. The uptake rate for water vapor excess by ice crystals is another key component regulating the total number of nucleated ice crystals. This rate, the product of ice number concentration and ice crystal diffusional growth rate, partially controls the peak nucleation rate achieved in an air parcel and the duration of the active nucleation time period.

Lin, Ruei-Fong↗

GCSS Cirrus Parcel Model Comparison Project

The Cirrus Parcel Model Comparison Project, a project of GCSS Working Group on Cirrus Cloud Systems (WG2), involves the systematic comparison of current models of ice crystal nucleation and growth for specified, typical, cirrus cloud environments. The goal of this project is to document and understand the factors resulting in significant inter-model differences. The intent is to foment research leading to model improvement and validation. In Phase 1 of the project reported here, simulated cirrus cloud microphysical properties are compared for situations of "warm" (-40 C) and "cold" (-60 C) cirrus subject to updrafts of 4, 20 and 100 cm/s, respectively. Five models participated. These models employ explicit microphysical schemes wherein the size distribution of each class of particles (aerosols and ice crystals) is resolved into bins. Simulations are made including both homogeneous and heterogeneous ice nucleation mechanisms. A single initial aerosol population of sulfuric acid particles is prescribed for all simulations. To isolate the treatment of the homogeneous freezing (of haze drops) nucleation process, the heterogeneous nucleation mechanism is disabled for a second parallel set of simulations. Qualitative agreement is found for the homogeneous-nucleation-only simulations, e.g., the number density of nucleated ice crystals increases with the strength of the prescribed updraft. However, non-negligible quantitative differences are found. Detailed analysis reveals that the homogeneous nucleation formulation, aerosol size, ice crystal growth rate (particularly the deposition coefficient), and water vapor uptake rate are critical components that lead to differences in predicted microphysics. Systematic bias exists between results based on a modified classical theory approach and models using an effective freezing temperature approach to the treatment of nucleation. Each approach is constrained by critical freezing data from laboratory studies, but each includes assumptions that can only be justified by further laboratory data. Consequently, it is not yet clear if the two approaches can be made consistent. Large haze particles may deviate considerably from equilibrium size in moderate to strong updrafts (20-100 cm/s) at -60 C when the commonly invoked equilibrium assumption is lifted. The resulting difference in particle-size-dependent solution concentration of haze particles may significantly affect the ice nucleation rate during the initial nucleation interval. The uptake rate for water vapor excess by ice crystals is another key component regulating the total number of nucleated ice crystals. This rate, the product of ice number concentration and ice crystal diffusional growth rate, which is sensitive to the deposition coefficient when ice particles are small, partially controls the peak nucleation rate achieved in an air parcel and the duration of the active nucleation time period. The effects of heterogeneous nucleation are most pronounced in weak updraft situations. Vapor competition by the nucleated (heterogeneous) ice crystals limits the achieved ice supersaturation and thus suppresses the contribution of homogeneous nucleation. Correspondingly, ice crystal number density is markedly reduced. Definitive laboratory and atmospheric benchmark data are needed for the heterogeneous nucleation process. Inter-model differences are correspondingly greater than in the case of the homogeneous nucleation process acting alone.

Lin, Ruei-Fong↗

Cirrus Parcel Model Comparison Project: The Critical Components to Simulate Cirrus Initiation Explicitly - Phase 1

The Cirrus Parcel Model Comparison Project, a project of the GCSS (GEWEX Cloud System Studies) Working Group on Cirrus Cloud Systems, involves the systematic comparison of current models of ice crystal nucleation and growth for specified, typical, cirrus cloud environments. In Phase I of the project reported here, simulated cirrus cloud microphysical properties are compared for situations of "warm" (40 C) and "cold" (-60 C) cirrus, both subject to updrafts of 4, 20 and 100 centimeters per second. Five models participated. The various models employ explicit microphysical schemes wherein the size distribution of each class of particles (aerosols and ice crystals) is resolved into bins or treated separately. Simulations are made including both the homogeneous and heterogeneous ice nucleation mechanisms. A single initial aerosol population of sulfuric acid particles is prescribed for all simulations. To isolate the treatment of the homogeneous freezing (of haze droplets) nucleation process, the heterogeneous nucleation mechanism is disabled for a second parallel set of simulations. Qualitative agreement is found for the homogeneous-nucleation- only simulations, e.g., the number density of nucleated ice crystals increases with the strength of the prescribed updraft. However, significant quantitative differences are found. Detailed analysis reveals that the homogeneous nucleation rate, haze particle solution concentration, and water vapor uptake rate by ice crystal growth (particularly as controlled by the deposition coefficient) are critical components that lead to differences in predicted microphysics. Systematic bias exists between results based on a modified classical theory approach and models using an effective freezing temperature approach to the treatment of nucleation. Each approach is constrained by critical freezing data from laboratory studies, but each includes assumptions that can only be justified by further laboratory research. Consequently, it is not yet clear if the two approaches can be made consistent. Large haze particles may deviate considerably from equilibrium size in moderate to strong updrafts (20-100 centimeters per second) at -60 C when the commonly invoked equilibrium assumption is lifted. The resulting difference in particle-size- dependent solution concentration of haze particles may significantly affect the ice particle formation rate during the initial nucleation interval. The uptake rate for water vapor excess by ice crystals is another key component regulating the total number of nucleated ice crystals. This rate, the product of particle number concentration and ice crystal diffusional growth rate, which is particularly sensitive to the deposition coefficient when ice particles are small, modulates the peak particle formation rate achieved in an air parcel and the duration of the active nucleation time period. The effects of heterogeneous nucleation are most pronounced in weak updraft situations. Vapor competition by the heterogeneously nucleated ice crystals may limit the achieved ice supersaturation and thus suppresses the contribution of homogeneous nucleation. Correspondingly, ice crystal number density is markedly reduced. Definitive laboratory and atmospheric benchmark data are needed for the heterogeneous nucleation process. Inter-model differences are correspondingly greater than in the case of the homogeneous nucleation process acting alone.

Lin, Ruei-Fong↗

Enhanced Light Outcoupling from OLEDs Fabricated on Novel Low-Cost Patterned Plastic Substrates of Varying Periodicity

OLEDs continue to make strides in display applications, but their commercial utilization in solid-state lighting (SSL) is lagging. An ongoing challenge, in particular for manufacturing, is the need for enhanced efficiency and hence the necessity to increase in an inexpensive approach the extraction of the light generated inside the OLED into the forward (viewing) hemisphere. In conventional OLEDs fabricated on a transparent flat anode coated on glass, the external quantum efficiency (EQE) is only ~20%. About 50% of the light is lost to internal waveguiding in the high refractive index (RI) organic + ITO anode layers and to surface plasmon polaritons (SPPs) at the organic/metal cathode interface. Another ~30% of the light is externally waveguided in the substrate to its edges. While extraction of the externally waveguided light is commonly addressed by adding a microlens array (MLA) or a scattering layer at the substrate’s air-side, light outcoupling increases by only ~1.6-1.7x (vs up to 2.5x in improving from ~20% to ~50%). The use of a hemispherical lens or an index matching fluid (IMF) at the substrate/photodetector (PD) interface increases the outcoupling by at least 2x; these approaches however, are not viable industrially, and even a MLA is sometimes undesirable due to its non-planar, scattering structure. In multi-stack tandem OLEDs, where the metal cathode is far from the emitting zone(s), the impact of photons loss to SPPs decreases. Our project addressed the ~50% loss to the internally waveguided light and SPPs. We evaluated OLEDs fabricated on patterned or planarized plastic substrates manufactured in a cost-effective approach compatible with a roll-to-roll (R2R) process. The OLEDs were either (i) patterned to various degrees depending on the pitch a and height or depth h of the pattern features or (ii) planar, with a pattern buried under a flat high RI planarization layer. We demonstrated that the outcoupling from green patterned OLEDs reaches ~50% by mitigating plasmon–related loss and internal waveguiding, even without the addition of a MLA, a hemispherical lens, or IMF. Simulations conducted in parallel with the experimental effort demonstrated how diffraction by conformally corrugated OLEDs increases the outcoupling to >60%. Structures with varying pitch values were also simulated indicating that combining domains of varying pitch could increase outcoupling to 55-60%. Experimentally, we additionally assessed the role a and h in determining not only the OLED efficiencies, but also their structural properties, i.e., the uniformity and conformality throughout the OLED stack. As planar OLEDs are preferred over corrugated devices, we studied different patterns in plastic substrates that were planarized by a high RI formulation. Planar green OLEDs on such structures showed enhanced efficiencies with EQEs larger than 60% with the addition of an IMF (to extract the substrate mode) at the substrate/Si PD interface. White OLEDs showed EQEs of 45.5%. Plastic substrates are currently less attractive than glass substrates due to drawbacks such as permeability to water vapor and oxygen, and in some cases thermal instability. Plastic substrates however, are flexible and easy to handle unlike thin flexible glass, and once transparent thin barrier films are available, they will become more attractive; they are already of interest in medical applications. Importantly, as it is easy to generate various patterns in different plastic materials, they provide excellent means for assessing and optimizing enhancing extracting structures. Such structures can also be transferred to glass substrates with some process modifications. The technical effectiveness and economic feasibility of the project lie in the patterning of the extracting plastic substrates in an approach that is scalable to R2R manufacturing. R2R processes are of drastically lower-cost than batch or single-unit fabrication. The patterned plastic can be a part of an integrated substrate either plastic or glass, which includes also a MLA or a planar layer with embedded scattering particles, as well as a conductive metal mesh/electrode design. SSL is environmentally-friendly and as OLED SSL becomes more efficient it will reduce electricity consumption, and hence lighting cost, as well as produce less expensive attractive lighting fixtures. Our university-industry collaboration is hence of major benefit to the public as it demonstrates the feasibility of manufacturing optimized extracting substrates for highly efficient OLEDs for SSL in a future R2R process, which would drastically reduce the manufacturing cost and increase production in the USA. Moreover, newly developed methods by our team allow low-cost roll manufactured substrates to be transferred to flexible or rigid glass substrates, which solves the plastic substrate barrier issues, and when combined with device encapsulation will increase the OLEDs’ environmental stability.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Flight Simulation

Advanced Rotorcraft Technology, (ART), Mountain View, CA, developed FLIGHTLAB, a design/analysis tool and pilot simulation system for rotary wing aircraft, based on a research project conducted by ART, the Army Aerodynamics Directorate and Ames Research Center on using parallel processing in simulation models. The experiments showed that a new software architecture could provide real time simulations of high performance rotorcraft that closely matched actual flight at substantially lower cost than the mainframe computers previously needed for the simulations.

Source record↗

Parallelized modelling and solution scheme for hierarchically scaled simulations

This two-part paper presents the results of a benchmarked analytical-numerical investigation into the operational characteristics of a unified parallel processing strategy for implicit fluid mechanics formulations. This hierarchical poly tree (HPT) strategy is based on multilevel substructural decomposition. The Tree morphology is chosen to minimize memory, communications and computational effort. The methodology is general enough to apply to existing finite difference (FD), finite element (FEM), finite volume (FV) or spectral element (SE) based computer programs without an extensive rewrite of code. In addition to finding large reductions in memory, communications, and computational effort associated with a parallel computing environment, substantial reductions are generated in the sequential mode of application. Such improvements grow with increasing problem size. Along with a theoretical development of general 2-D and 3-D HPT, several techniques for expanding the problem size that the current generation of computers are capable of solving, are presented and discussed. Among these techniques are several interpolative reduction methods. It was found that by combining several of these techniques that a relatively small interpolative reduction resulted in substantial performance gains. Several other unique features/benefits are discussed in this paper. Along with Part 1's theoretical development, Part 2 presents a numerical approach to the HPT along with four prototype CFD applications. These demonstrate the potential of the HPT strategy.

Padovan, Joe↗

International Conference on Numerical Methods in Fluid Dynamics, 11th, Williamsburg, VA, June 27-July 1, 1988, Proceedings

Recent advances in computational fluid dynamics (CFD) are discussed in reviews and reports. Topics addressed include CFD models in plasma dynamics, parallel computation for simulation studies, CFD for hypersonic airbreathing aircraft, multigrid methods for the steady incompressible Navier-Stokes equations, upwind differencing techniques, TV stable schemes for shock-interacting flows, Euler models of hypersonic vortex flows, parallel multilevel adaptive methods, and vortex methods for slightly viscous three-dimensional flows. Consideration is given to the accuracy of node-based solutions on irregular meshes, multigrid calculations for cascades, a finite-volume-element method for planar cavity flow, parallel heterogeneous mesh refinement for advection-diffusion equations, the convergence of the spectral-viscosity method for nonlinear conservation laws, and numerical simulations of Taylor vortices in a spherical gap.

Dwoyer, D. L.↗

Massively parallel axisymmetric fluid model for streamer discharges

A highly parallelizable fluid plasma simulation tool based upon the first-order drift-diffusion equations is discussed. Atmospheric pressure plasmas have densities and gradients that require small element sizes in order to accurately simulate the plasm resulting in computational meshes on the order of millions to tens of millions of elements for realistic size plasma reactors. To enable simulations of this nature, parallel computing is required and must be optimized for the particular problem. Here, a finite-volume, electrostatic drift-diffusion implementation for low-temperature plasma is discussed. The implementation is built upon the Message Passing Interface (MPI) library in C++ using Object Oriented Programming. The underlying numerical method is outlined in detail and benchmarked against simple streamer formation from other streamer codes. Electron densities, electric field, and propagation speeds are compared with the reference case and show good agreement. Convergence studies are also performed showing a minimal space step of approximately 4 μm required to reduce relative error to below 1% during early streamer simulation times and even finer space steps are required for longer times. Additionally, strong and weak scaling of the implementation are studied and demonstrate the excellent performance behavior of the implementation up to 100 million elements on 1024 processors. Lastly, different advection schemes are compared for the simple streamer problem to analyze the influence of numerical diffusion on the resulting quantities of interest.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Coupling Surface Flow with High-performance Subsurface Reactive Flow and Transport Code PFLOTRAN

Water exchange between the surface and subsurface is important for both water resource management and environmental protection. In this paper, we develop coupled surface and subsurface flow simulation capability in a parallel subsurface flow and reactive transport code PFLOTRAN. We sequentially couple the diffusion wave-based surface flow with the subsurface flow governedby the Richards equation in PFLOTRAN. These two flow domains are linked with a boundary condition switching method that ensures continuity of pressure and flux at the surface-subsurface interface. We verify the coupled code against other existing hydrologic models and observation data using a number of numerical experiments. The coupled hydrological model exhibits good performance in strong parallel scaling tests. The new coupled surface and subsurface simulator significantly advance community simulation capability towards improving integrated hydrologic and biogeochemical understanding of complex systems such as watersheds and river corridors. Keywords: Surface flow, Integrated hydrological modeling, Boundary condition switching, Parallel computing

Wu, Runjian↗

FULL RANGE TUNE SCAN STUDIES USING GRAPHICS PROCESSING UNITS WITH CUDA IN EIC BEAM-BEAM SIMULATIONS

The hadron beam in the Electron-Ion Collider (EIC) suffers high order betatron and synchro-betatron resonances. In this paper, we present a weak-strong full range (0.0 ~ 0.5) fractional tune scan with a step size as small as 0.001. Multiple Graphics Processing Units (GPUs) are used to speed up the simulation. A code parallelized with MPI and CUDA is implemented. The good tune region from weak-strong scan is further checked by the self-consistent strong-strong simulation. This study provides beam dynamics guidance in choosing proper working points for the future EIC.

43 PARTICLE ACCELERATORS↗

High-fidelity wind farm simulation methodology with experimental validation

The complexity and associated uncertainties involved with atmospheric-turbine-wake interactions produce challenges for accurate wind farm predictions of generator power and other important quantities of interest (QoIs), even with state-of-the-art high-fidelity atmospheric and turbine models. A comprehensive computational study was undertaken with consideration of simulation methodology, parameter selection, and mesh refinement on atmospheric, turbine, and wake QoIs to identify capability gaps in the validation process. For neutral atmospheric boundary layer conditions, the massively parallel large eddy simulation (LES) code Nalu-Wind was used to produce high-fidelity computations for experimental validation using high-quality meteorological, turbine, and wake measurement data collected at the Department of Energy/Sandia National Laboratories Scaled Wind Farm Technology (SWiFT) facility located at Texas Tech University’s National Wind Institute. The wake analysis showed the simulated lidar model implemented in Nalu-Wind was successful at capturing wake profile trends observed in the experimental lidar data.

17 WIND ENERGY↗

A-SST Initial Specification

The U.S. Army Research Office (ARO), in partnership with IARPA, are investigating innovative, efficient, and scalable computer architectures that are capable of executing next-generation large scale data-analytic applications. These applications are increasingly sparse, unstructured, non-local, and heterogeneous. Under the Advanced Graphic Intelligence Logical computing Environment (AGILE) program, Performer teams will be asked to design computer architectures to meet the future needs of the DoD and the Intelligence Community (IC). This design effort will require flexible, scalable, and detailed simulation to assess the performance, efficiency, and validity of their designs. To support AGILE, Sandia National Labs will be providing the AGILE-enhanced Structural Simulation Toolkit (A-SST). This toolkit is a computer architecture simulation framework designed to support fast, parallel, and multi-scale simulation of novel architectures. This document describes the A-SST framework, some of its library of simulation models, and how it may be used by AGILE Performers.

97 MATHEMATICS AND COMPUTING↗

Predictive Tools for Customizing Heat Treatment of Additively Manufactured Aerospace Components

Laser-bed powder fusion (LBPF) additive manufacturing is increasingly being used to produce components of complex geometries using the Ni-base superalloy Inconel 718. The composition and the microstructure of the alloy are currently well optimized for wrought components made using conventional manufacturing processes such as rolling, forging, extrusion, etc. The attractive mechanical properties of the alloy result from the underlying austenitic matrix with fine equiaxed grains, and a high density and uniform distribution of the precipitation hardening phase, γ". Heat treatment steps such as homogenization, solutioning and aging are well documented for the wrought alloy. However, when the same wrought alloy compositions are used for the additive manufacturing (AM) processes, the asprocessed microstructure is significantly different, because of the different thermal history associated with LBPF, including rapid solidification and multiple temperature excursions that lead to multiple re-melting and reheating in the solid state. Rapid solidification introduces potential non-equilibrium effects at the moving solid-liquid interfaces that impact the extent of solute segregation, as well as the morphology of the dendritic grains that form. In order to recover the target mechanical properties, AM components have to undergo post-process heat treatments. However, such heat treatments have to be custom designed for the AM process and the component geometry because of the expected vast differences in the microstructure at various locations of a component with complex geometry. The homogenization and precipitation steps should be optimized for the component so that target mechanical properties can be obtained throughout the part. The objective of this research is to utilize High Performance Computing in phase field simulations of microstructure evolution during post-processing of AM components. The physics-based modeling will be beneficial in reducing the experimental effort required for heat treatment process selection, optimization, and certification, thus leading to a significant reduction in energy consumption for AM and post-processing heat treatment. The optimization study will help identify heat treatments steps that are critical for development of a final desired microstructure with the minimum energy input. This combined with shortening of the production cycle (time-to-market) by reducing the number of failed parts (property targets), and reduction in the number of iterations for process optimization, will enable 30-40% savings in the energy costs. Phase field simulations of the degree of homogenization and the effect of local matrix composition on the nucleation and growth of competing precipitating phases were performed using the Microstructure Evolution Using Massively Parallel Phase Field Simulations code developed in-house at the Oak Ridge National Laboratory. The simulations were able to successfully capture the kinetics of nucleation and growth, and morphologies of various precipitating phases as a function of local matrix compositions and composition gradients characteristic of local microstructures arising from location-dependent variations in the thermal conditions. Future work will involve extending the simulations to a length scale consisting of multiple dendrites, so that the effect of homogenization on the coarsening of the dendrites can be simulated and used as an additional input to the optimization of the heat treatment process.

36 MATERIALS SCIENCE↗

Performance Evaluation of Three Distributed Computing Environments for Scientific Applications

We present performance results for three distributed computing environments using the three simulated CFD applications in the NAS Parallel Benchmark suite. These environments are the DCF cluster, the LACE cluster, and an Intel iPSC/860 machine. The DCF is a prototypic cluster of loosely coupled SGI R3000 machines connected by Ethernet. The LACE cluster is a tightly coupled cluster of 32 IBM RS6000/560 machines connected by Ethernet as well as by either FDDI or an IBM Allnode switch. Results of several parallel algorithms for the three simulated applications are presented and analyzed based on the interplay between the communication requirements of an algorithm and the characteristics of the communication network of a distributed system.

Fatoohi, Rod↗

Accelerating Simulation for High-Fidelity PV Inverter System Reliability Assessment with High-Performance Computing

The overall cost of photovoltaic (PV) systems has shown a downward trend during the last decade; however, PV inverter failures account for the highest cost of operation and maintenance. To address this, reliability tools with powerful computation and better accuracy are required for the lifetime prediction and degradation evaluation of PV inverters. This paper proposes an event-driven parallel computing-based simulator. The proposed simulator applies high-performance computing techniques and other accessory optimization techniques-including cluster merging, adaptive model updates, and steady-state identification-to make reliability assessments for PV inverters under given input mission profiles and operating conditions with high efficiency and high fidelity. The main idea of the simulator and its workflow are introduced. Then, a demo PV inverter system simulator is implemented, and the speedup of the total simulations of the switching model reaches 123.03 times.

high-performance computing↗

A Massively Parallel Bayesian Approach to Planetary Protection Trajectory Analysis and Design

The NASA Planetary Protection Office has levied a requirement that the upper stage of future planetary launches have a less than 10(exp -4) chance of impacting Mars within 50 years after launch. A brute-force approach requires a decade of computer time to demonstrate compliance. By using a Bayesian approach and taking advantage of the demonstrated reliability of the upper stage, the required number of fifty-year propagations can be massively reduced. By spreading the remaining embarrassingly parallel Monte Carlo simulations across multiple computers, compliance can be demonstrated in a reasonable time frame. The method used is described here.

parallel computing↗