Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “direct solver”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 271 records · Page 15

A Perspective on Quantum Computing Applications in Quantum Chemistry Using 25-100 Logical Qubits

The intersection of quantum computing and quantum chemistry represents a promising frontier for achieving quantum utility in domains of both scientific and societal relevance. Owing to the exponential growth of classical resource requirements for simulating quantum systems, quantum chemistry has long been recognized as a natural candidate for quantum computation. This perspective focuses on identifying scientifically meaningful use cases where early fault-tolerant quantum computers, which are considered to be equipped with approximately 25-100 logical qubits, could deliver tangible impact. While recent advances in classical computing have pushed the boundaries of tractable simulations to unprecedented scales, this logical-qubit regime represents the first window where quantum devices can pursue qualitatively distinct strategies, such as polynomial-scaling phase estimation, direct simulation of quantum dynamics, and active-space embedding, that remain challenging for classical solvers, such as multireference charge-transfer and conical-intersection states central to photochemistry and materials design. We highlight near-term opportunities in algorithm and software design, discuss representative chemical problems suited for quantum acceleration, and propose strategic roadmaps and collaborative pathways for advancing practical quantum utility in quantum chemistry.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Modeling Shed Vorticity from Coaxial Blade Interactions

Coaxial counter-rotating rotors operate in a flowfield different from single rotors. Aerodynamic interactions such as blade crossing and shed vorticity result in potential sources of noise and impulsive blade loads. In previous research, the authors simulated two trains of airfoils traveling in opposite directions for specified speeds, airfoil thickness and vertical separation distances, using the compressible Navier-Stokes solver OVERFLOW. Previously, the effects of circulation, thickness, and compressibility were explored. This work continues the previous research by exploring downwash and shed vorticity effects. These phenomena are explored by simulating two trains of eight airfoils vertically separated traveling in opposite directions. The effects of downwash are simulated by introducing a vertical flow. Vorticity shed from the upper train of airfoils is shown to interact with the lower train, affecting the loading on the lower airfoils. Furthermore, viscid and inviscid calculations are performed to further understand the behavior of shed vorticity.

Natasha L Schatzman↗

Application of Aeroelastic Solvers Based on Navier Stokes Equations

The propulsion element of the NASA Advanced Subsonic Technology (AST) initiative is directed towards increasing the overall efficiency of current aircraft engines. This effort requires an increase in the efficiency of various components, such as fans, compressors, turbines etc. Improvement in engine efficiency can be accomplished through the use of lighter materials, larger diameter fans and/or higher-pressure ratio compressors. However, each of these has the potential to result in aeroelastic problems such as flutter or forced response. To address the aeroelastic problems, the Structural Dynamics Branch of NASA Glenn has been involved in the development of numerical capabilities for analyzing the aeroelastic stability characteristics and forced response of wide chord fans, multi-stage compressors and turbines. In order to design an engine to safely perform a set of desired tasks, accurate information of the stresses on the blade during the entire cycle of blade motion is required. This requirement in turn demands that accurate knowledge of steady and unsteady blade loading is available. To obtain the steady and unsteady aerodynamic forces for the complex flows around the engine components, for the flow regimes encountered by the rotor, an advanced compressible Navier-Stokes solver is required. A finite volume based Navier-Stokes solver has been developed at Mississippi State University (MSU) for solving the flow field around multistage rotors. The focus of the current research effort, under NASA Cooperative Agreement NCC3- 596 was on developing an aeroelastic analysis code (entitled TURBO-AE) based on the Navier-Stokes solver developed by MSU. The TURBO-AE code has been developed for flutter analysis of turbomachine components and delivered to NASA and its industry partners. The code has been verified. validated and is being applied by NASA Glenn and by aircraft engine manufacturers to analyze the aeroelastic stability characteristics of modem fans, compressors and turbines.

Keith, Theo G., Jr.↗

Comparison of Some RANS Solvers

We will take a look at solving the Reynolds-averaged Navier-Stokes (RANS) equations that are encountered in the context of wind farm performance simulations and optimizations. We will compare some of the more popular ways to solve these equations with a focus on using iterative solvers for the linear solve. We will compare their performance and reliability to a direct solve as we scale the problem both by adding more and parallel resources and by increasing the size of the domain, both in two and three dimensions. There are many strategies that can be applied to solving the RANS equations, some are very efficient, while others are very insensitive to, for example, the Reynolds number. The first contender we will consider is the Pressure-Convection-Diffusion (PCD) preconditioner. Early results suggest that PCD is indeed a very efficient solver, in particular in two dimensions, as long as the Reynold's number remains small. Next we will try to reorder our degrees of freedom such that we can use GMRES with ILU for our linear solve. Another popular choice we will consider for solving the RANS equations is SIMPLE (and its derivatives). For all of our implementations we make use of either FEniCS or Firedrake, basing our work on both existing implementations of some of these solvers while also writing new extensions for others.

CFD↗

Multi-variance replica exchange SGMCMC for inverse and forward problems via Bayesian PINN

Physics-informed neural network (PINN) has been successfully applied in solving a variety of nonlinear non-convex forward and inverse problems. However, the training is challenging because of the non-convex loss functions and the multiple optima in the Bayesian inverse problem. In this work, we propose a multi-variance replica exchange stochastic gradient Langevin dynamics method to tackle the challenge of the multiple local optima in the optimization and the challenge of the multiple modal posterior distribution in the inverse problem. Replica exchange methods are capable of escaping from the local traps and accelerating the convergence; two chains with different temperatures are designed where the low temperature chain aims for the local convergence, and the target of the high temperature chain is to travel globally and explore the whole loss function entropy landscape. However, it may not be efficient to solve mathematical inversion problems by using the vanilla replica method directly since the method doubles the computational cost in evaluating the forward solvers (likelihood functions) in the two chains. To address this issue, we propose to make different assumptions on the energy function estimation and this facilities one to use solvers of different fidelities in the likelihood function evaluation. More precisely, one can use a solver with low fidelity in the high temperature chain while using a solver with high fidelity in the low temperature chain. Our proposed method significantly lowers the computational cost in the high temperature chain, meanwhile preserving the accuracy and converging very fast. Here we give an unbiased estimate of the swapping rate and give an estimation of the discretization error of the scheme. To verify our idea, we design and solve four inverse problems which have multiple modes. The proposed method is also employed to train the Bayesian PINN to solve the forward and inverse problems; faster and more accurate convergence has been observed when compared to the stochastic gradient Langevin dynamics (SGLD) method and vanilla replica exchange methods.

97 MATHEMATICS AND COMPUTING↗

Postbuckling of long orthotropic plates in combined shear and compression

The nonlinear large-deflection partial differential equations of von karman for orthotropic plates loaded in combined shear and compression are converted into a set of first-order nonlinear ordinary differential equations by assuming trigonometric functions in one direction. These equations are solved numerically using a two point boundary problem solver which makes use of Newton's method. Results are obtained which determine the postbuckling behavior of rectangular plates with loading up to about three times the buckling load. Both isotropic and orthotropic composite plates are considered. Results show that orthotropic plates may behave quite differently than isotropic plates and that in-plane boundary conditions are important for plates loaded in shear.

Stein, M.↗

Gust Acoustics Computation with a Space-Time CE/SE Parallel 3D Solver

The benchmark Problem 2 in Category 3 of the Third Computational Aero-Acoustics (CAA) Workshop is solved using the space-time conservation element and solution element (CE/SE) method. This problem concerns the unsteady response of an isolated finite-span swept flat-plate airfoil bounded by two parallel walls to an incident gust. The acoustic field generated by the interaction of the gust with the flat-plate airfoil is computed by solving the 3D (three-dimensional) Euler equations in the time domain using a parallel version of a 3D CE/SE solver. The effect of the gust orientation on the far-field directivity is studied. Numerical solutions are presented and compared with analytical solutions, showing a reasonable agreement.

Wang, X. Y.↗

Direct Numerical Simulation of Turbulent Condensation in Clouds

In this brief, we investigate the turbulent condensation of a population of droplets by means of a direct numerical simulation. To that end, a coupled Navier-Stokes/Lagrangian solver is used where each particle is tracked and its growth by water vapor condensation is monitored exactly. The main goals of the study are to find out whether turbulence broadens the droplet size distribution, as observed in in situ measurements. The second issue is to understand if and for how long a correlation between the droplet radius and the local supersaturation exists for the purpose of modeling sub-grid scale microphysics in cloud-resolving codes. This brief is organized as follows. In Section 2 the governing equations are presented, including the droplet condensation model. The implementation of the forcing procedure is described in Section 3. The simulation results are presented in Section 4 together with a sketch of a simple stochastic model for turbulent condensation. Conclusions and the main outcomes of the study are given in Section 5.

Shariff, K.↗

Aerodynamic and Acoustic Interactions Associated with Inboard Propeller-Wing Configurations

A series of aerodynamic performance and acoustic measurements have been made on a range of inboard propeller-wing interaction configurations in the NASA Langley Low Speed Aeroacoustic Wind Tunnel (LSAWT). The results presented in this paper are part of a more expansive testing campaign encompassing both single propeller-wing and multipropeller-wing interactions, the former of which is discussed in the present work. The primary testing parameters of interest to this study are the axial and vertical positioning of the wing relative to the propeller slipstream under a constant propeller advance ratio. A multi-faceted computational effort was also employed in an effort to identify reflection and scattering effects imposed by both the wing geometry as well as the primary components of the facility test setup. This effort consisted of aerodynamic predictions using high-fidelity computational fluid dynamics (CFD), acoustic predictions using an impermeable Ffowcs Williams and Hawkings (FW-H) solver, and acoustic scattering predictions. Acoustic measurements reveal variations in the acoustic directivity behavior of the propeller blade passage frequency for even modest variations in wing position. CFD-based acoustic predictions reveal discrepancies relative to the experimental data, which is believed to be due to complex acoustic scattering behavior within the test section. Initial attempts at modeling the scattered acoustic field showed functional dependency of the acoustic amplitude variations on the wing position relative to the propeller disk, however discrepancies with experimental data remain.

Nikolas S. Zawodny↗

HTR solver: An open-source exascale-oriented task-based multi-GPU high-order code for hypersonic aerothermodynamics

In this study, the open-source Hypersonics Task-based Research (HTR) solver for hypersonic aerothermodynamics is described. The physical formulation of the code includes thermochemical effects induced by high temperatures (vibrational excitation and chemical dissociation). The HTR solver uses high-order TENO-based spatial discretization on structured grids and efficient time integrators for stiff systems, is highly scalable in GPU-based supercomputers as a result of its implementation in the Regent/Legion stack, and is designed for direct numerical simulations of canonical hypersonic flows at high Reynolds numbers. Additionally, the performance of the HTR solver is tested with benchmark cases including inviscid vortex advection, low- and high-speed laminar boundary layers, inviscid one-dimensional compressible flows in shock tubes, supersonic turbulent channel flows, and hypersonic transitional boundary layers of both calorically perfect gases and dissociating air.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Parallel 3D Multi-Stage Simulation of a Turbofan Engine

A 3D multistage simulation of each component of a modern GE Turbofan engine has been made. An axisymmetric view of this engine is presented in the document. This includes a fan, booster rig, high pressure compressor rig, high pressure turbine rig and a low pressure turbine rig. In the near future, all components will be run in a single calculation for a solution of 49 blade rows. The simulation exploits the use of parallel computations by using two levels of parallelism. Each blade row is run in parallel and each blade row grid is decomposed into several domains and run in parallel. 20 processors are used for the 4 blade row analysis. The average passage approach developed by John Adamczyk at NASA Lewis Research Center has been further developed and parallelized. This is APNASA Version A. It is a Navier-Stokes solver using a 4-stage explicit Runge-Kutta time marching scheme with variable time steps and residual smoothing for convergence acceleration. It has an implicit K-E turbulence model which uses an ADI solver to factor the matrix. Between 50 and 100 explicit time steps are solved before a blade row body force is calculated and exchanged with the other blade rows. This outer iteration has been coined a "flip." Efforts have been made to make the solver linearly scaleable with the number of blade rows. Enough flips are run (between 50 and 200) so the solution in the entire machine is not changing. The K-E equations are generally solved every other explicit time step. One of the key requirements in the development of the parallel code was to make the parallel solution exactly (bit for bit) match the serial solution. This has helped isolate many small parallel bugs and guarantee the parallelization was done correctly. The domain decomposition is done only in the axial direction since the number of points axially is much larger than the other two directions. This code uses MPI for message passing. The parallel speed up of the solver portion (no 1/0 or body force calculation) for a grid which has 227 points axially.

Turner, Mark G.↗

PETSc/TAO Users Manual (Rev. 3.19)

This manual describes the use of the Portable, Extensible Toolkit for Scientific Computation (PETSc) and the Toolkit for Advanced Optimization (TAO) for the numerical solution of partial differential equations and related problems on high-performance computers. PETSc/TAO is a suite of data structures and routines that provide the building blocks for the implementation of large-scale application codes on parallel (and serial) computers. PETSc uses the MPI standard for all distributed memory communication. PETSc/TAO includes a large suite of parallel linear solvers, nonlinear solvers, time integrators, and opti mization that may be used in application codes written in Fortran, C, C++, and Python (via petsc4py; see Getting Started). PETSc provides many of the mechanisms needed within parallel application codes, such as parallel matrix and vector assembly routines. The library is organized hierarchically, enabling users to employ the level of abstraction that is most appropriate for a particular problem. By using techniques of object-oriented programming, PETSc provides enormous flexibility for users. PETSc is a sophisticated set of software tools; as such, for some users it initially has a much steeper learning curve than packages such as MATLAB or a simple subroutine library. In particular, for individuals without some computer science background, experience programming in C, C++, python, or Fortran and experience using a debugger such as gdb or lldb, it may require a significant amount of time to take full advantage of the features that enable efficient software use. However, the power of the PETSc design and the algorithms it incorporates may make the efficient implementation of many application codes simpler than “rolling them” yourself. For many tasks a package such as MATLAB is often the best tool; PETSc is not intended for the classes of problems for which effective MATLAB code can be written. There are several packages, built on PETSc, that may satisfy your needs without requiring directly using PETSc. We recommend reviewing these packages functionality before starting to code directly with PETSc. PETSc can be used to provide a “MPI parallel linear solver” in an otherwise sequential, or OpenMP parallel code. This approach cannot provide extremely large improvements in the application time by utilizing large numbers of MPI processes but can still improve the performance. Certainly all parts of a previously sequential code need not be parallelized but the matrix generation portion must be parallelized to expect true scalability to large numbers of MPI processes. See PCMPI for details on how to utilize the PETSc MPI linear solver server. Since PETSc is under continued development, small changes in usage and calling sequences of routines will occur. PETSc has been supported for twenty-five years; see mailing list information on our website for information on contacting support.

97 MATHEMATICS AND COMPUTING↗

Gas-Kinetic Theory Based Flux Splitting Method for Ideal Magnetohydrodynamics

A gas-kinetic solver is developed for the ideal magnetohydrodynamics (MHD) equations. The new scheme is based on the direct splitting of the flux function of the MHD equations with the inclusion of "particle" collisions in the transport process. Consequently, the artificial dissipation in the new scheme is much reduced in comparison with the MHD Flux Vector Splitting Scheme. At the same time, the new scheme is compared with the well-developed Roe-type MHD solver. It is concluded that the kinetic MHD scheme is more robust and efficient than the Roe- type method, and the accuracy is competitive. In this paper the general principle of splitting the macroscopic flux function based on the gas-kinetic theory is presented. The flux construction strategy may shed some light on the possible modification of AUSM- and CUSP-type schemes for the compressible Euler equations, as well as to the development of new schemes for a non-strictly hyperbolic system.

Xu, Kun↗

Assessment of the Fluid Dynamics Boundary Condition in Ablating or Blowing Flows

Improved models of ablative thermal protection systems have enabled the treatment of materials and fluid behavior in a coupled manner. This paper reports a new approach to modeling the interface between fluid and material, with attention to the conservation of species mass flux and energy on the fluid side of the interface. The general equation is presented and is shown to recover the traditional uncoupled fluid/materials response interface. Including the chemical reaction terms on the CFD side of the interface makes the heat flux exchange independent of the thermodynamic reference state and, therefore, a measurable quantity. Doing so allows the material response solver to take as input the surface heat flux rather than a film coefficient. Removing the film coefficient approximation enables more direct solution of vehicle thermal response but requires consistency in the wall state. The mixing of the shock layer and pyrolysis gas is then computed with finite rate chemistry within the fluid solver. The boundary conditions described have been implemented in the DPLR v4.05.1 code. Char removal is captured using finite rate chemistry in DPLR’s gas surface interaction module. Aspects of coupling these solutions to material response are discussed.

Ablation↗

Assessment of the Fluid Dynamics Boundary Condition in Ablating or Blowing Flows

Improved models of ablative thermal protection systems have enabled the treatment of materials and fluid behavior in a coupled manner. This paper reports a new approach to modeling the interface between fluid and material, with attention to the conservation of species mass flux and energy on the fluid side of the interface. The general equation is presented and is shown to recover the traditional uncoupled fluid/materials response interface. Including the chemical reaction terms on the CFD side of the interface makes the heat flux exchange independent of the thermodynamic reference state and, therefore, a measurable quantity. Doing so allows the material response solver to take as input the surface heat flux rather than a film coefficient. Removing the film coefficient approximation enables more direct solution of vehicle thermal response but requires consistency in the wall state. The mixing of the shock layer and pyrolysis gas is then computed with finite rate chemistry within the fluid solver. The boundary conditions described have been implemented in the DPLR v4.05.1 code. Char removal is captured using finite rate chemistry in DPLR’s gas surface interaction module. Aspects of coupling these solutions to material response are discussed.

Ablation↗

On the Convergence of Overlapping Schwarz Decomposition for Nonlinear Optimal Control

Here, we study the convergence properties of an overlapping Schwarz decomposition algorithm for solving nonlinear optimal control problems (OCPs). The algorithm decomposes the time domain into a set of overlapping subdomains, and solves all subproblems defined over subdomains in parallel. The convergence is attained by updating primal-dual information at the boundaries of overlapping subdomains. We show that the algorithm exhibits local linear convergence, and that the convergence rate improves exponentially with the overlap size. We also establish global convergence results for a general quadratic programming, which enables the application of the Schwarz scheme inside second-order optimization algorithms (e.g., sequential quadratic programming). The theoretical foundation of our convergence analysis is a sensitivity result of nonlinear OCPs, which we call "exponential decay of sensitivity" (EDS). Intuitively, EDS states that the impact of perturbations at domain boundaries (i.e., initial and terminal time) on the solution decays exponentially as one moves into the domain. Here, we expand a previous analysis available in the literature by showing that EDS holds for both primal and dual solutions of nonlinear OCPs, under uniform second-order sufficient condition, controllability condition, and boundedness condition. We conduct experiments with a quadrotor motion planning problem and a partial differential equations (PDE) control problem to validate our theory, and show that the approach is significantly more efficient than alternating direction method of multipliers and as efficient as the centralized interior-point solver.

42 ENGINEERING↗

KiT-RT: An Extendable Framework for Radiative Transfer and Therapy

Here, in this article, we present Kinetic Transport Solver for Radiation Therapy (KiT-RT), an open-source C++-based framework for solving kinetic equations in therapy applications available at https://github.com/CSMMLab/KiT-RT . This software framework aims to provide a collection of classical deterministic solvers for unstructured meshes that allow for easy extendability. Therefore, KiT-RT is a convenient base to test new numerical methods in various applications and compare them against conventional solvers. The implementation includes spherical harmonics, minimal entropy, neural minimal entropy, and discrete ordinates methods. Solution characteristics and efficiency are presented through several test cases ranging from radiation transport to electron radiation therapy. Due to the variety of included numerical methods and easy extendability, the presented open-source code is attractive for both developers, who want a basis to build their numerical solvers, and users or application engineers, who want to gain experimental insights without directly interfering with the codebase.

97 MATHEMATICS AND COMPUTING↗

Hierarchial parallel computer architecture defined by computational multidisciplinary mechanics

The goal is to develop an architecture for parallel processors enabling optimal handling of multi-disciplinary computation of fluid-solid simulations employing finite element and difference schemes. The goals, philosphical and modeling directions, static and dynamic poly trees, example problems, interpolative reduction, the impact on solvers are shown in viewgraph form.

Padovan, Joe↗