Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “conjugate gradient”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

ORNL_AISD_NiPt_108atoms

This dataset describes the nickel-platinum (NiPt) solid solution binary alloy, where the two constituent elements nickel (Ni) and platinum (Pt) are randomly placed on the face centered cubic (FCC) crystal structure, with the lattice constant of 3.840 angstroms. The dataset comprises data for crystal structures with 108 atoms with 1,900 configurations. The data set was generated for concentrations ranging from 0at% of Pt to 100at% of Pt in the NiPt binary system, with increasing the concentration of Pt in the system every 5at%. For each one of the chemical compositions, 100 random configurations were generated, each with a different random seed. Each of the output files contains the mass, type, atomic coordinates, energy per atom, and forces in x, y, and z directions respectively. For each atomic configuration, the output was collected every 150 steps during the minimization stage and every 1000 steps during the replica exchange stage. Large-scale Atomic/Molecular Massively Parallel Simulator (LAMMPS) [1], which is a molecular dynamics code, was used to generate data for NiPt alloy. The simulation used the interatomic potential for NiPt binary system 'MEAM_LAMMPS_KimSeolJi_2017_PtNi__MO_020840179467_001' [3] from the OpenKIM library (Open Knowledgebase of Interatomic Models) [2]. This potential was developed based on the second nearest-neighbor modified embedded-atom method (2NN MEAM). The simulation process begins with the generation of the random NiPt structure and follows with the short minimization and replica exchange simulation. The minimization procedure adjusts atomic coordinates and performs energy minimization, which typically leads to a local potential energy minimum. The method used for the minimization was the conjugate gradient algorithm. A short replica exchange (parallel tempering) simulation involves four replicas (ensembles) of a system and follows the minimization stage. Multiple snapshots of the configuration were collected during the minimization and replica exchange stages. NiPt alloy is interesting due to its magnetic and charge transfer properties [4]. The data is provided in a compressed zipped folders atoms108.zip. The zipped folder contains the data structured in the following way: - Ni_ground_state.cfg --> atomic configuration for the pure nickel - Pt_ground_state.cfg --> atomic configuration for the pure platinum - Pt#_filtered --> folders containing atomic configurations for #at% concentration of platinum. The folder contains 100 atomic configurations, each saved in a subfolder - Each subfolder named config* is associated with a specific atomic configuration. Each of these subfolders contains files with .cfg format, corresponding to outputs for each atomic configuration The total number of atomic configurations contained in atoms108.zip is 66,132. This dataset is an extension to the dataset ORNL_AISD_NiPt [5] that has been previously released with crystal structures of 256 atoms, 864 atoms, and 2,048 atoms, with the same methodology for data collection. References [1] https://www.lammps.org/ [2] https://openkim.org/ [3] https://openkim.org/id/MEAM_LAMMPS_KimSeolJi_2017_PtNi__MO_020840179467_001 [4] El-Gendy, Ahmed A. and Hampel, Silke and Büchner, Bernd and Klingeler, Rüdiger, Tuneable magnetic properties of carbon-shielded NiPt-nanoalloys, RSC Adv., volume 6, issue 57, pages 52427-52433, 2016, The Royal Society of Chemistry, doi:10.1039/C6RA05910D [5] M. Karabin, M. Lupo Pasini, and M. Eisenbach. ORNL_AISD_NiPt. United States: N. p., 2023. Web. doi:10.13139/OLCF/1958172.

36 MATERIALS SCIENCE↗

Program Verification for Extreme-Scale Applications (Final Scientific/Technical Report)

The project seeks to develop tools and techniques to help software developers verify the correctness and accuracy of their code. The target domain is scientific software of the kind widely used and developed in the Department of Energy research community, with a particular focus on "extreme scale" programs - those that are expected to involve possibly millions of parallel threads of execution. The report covers the University of Delaware contribution to the collaborative project. The project had a number of successful outcomes, especially regarding the development and extension of the CIVL software verification framework. CIVL is a verification tool for C or Fortran programs that use MPI, OpenMP, CUDA, and/or Pthreads for parallelization. CIVL went through vast improvements and extensions, and was successfully applied to a number of challenging codes. It found a subtle bug in the Devito PDE framework. It was able to verify the functional correctness of a conjugate gradient solver using a novel probabilistic technique with vanishingly small chance of error. CIVL was used very successfully in verification competitions, and the CIVL solutions to the competition challenges were presented and published. CIVL was also used successfully by other researchers on a computational chemistry kernel.

97 MATHEMATICS AND COMPUTING↗

US Department of Energy, Office of Science High Performance Computing Facility Operational Assessment 2019 Oak Ridge Leadership Computing Facility

Oak Ridge National Laboratory's (ORNL's) Leadership Computing Facility (OLCF) continues to surpass its operational target goals: supporting users; delivering fast, reliable computational ecosystems; creating innovative solutions for high performance computing (HPC) needs; and managing risks, safety, and security associated with operating some of the most powerful computers in the world. The results can be seen in the cutting-edge science conducted by users and the praise from the research community. Calendar year (CY) 2019 was a big year as OLCF staff ran five world-class resources (the leadershipclass computers Titan and Summit, the large analysis cluster called Eos, and the massive parallel filesystems called Atlas and Alpine)) and also began power and cooling upgrades for a 2021 exascale system called Frontier. While continuing exceptional operation of Titan, Eos, and Rhea, the OLCF released the Summit supercomputer for production on January 1, 2019. Summit debuted as the most capable and efficient system in its class and has been recognized as the most powerful system in the world for its performance on both the high performance linpack (HPL) and conjugate gradient (HPCG) benchmark applications since June 2018 according to TOP500. Summit represents the culmination of a multiyear effort between the OLCF, IBM, NVIDIA, and Mellanox to deliver a system that is unmatched for modeling, simulation, data analysis, and learning. To hit the ground running with science-ready applications on day one, application teams worked closely with the OLCF through the Center for Accelerated Application Readiness (CAAR) program for years in advance of the Summit deployment. CY 2019 was filled with outstanding results and accomplishments: a very high rating from users on overall satisfaction for the sixth year in a row; a tremendous amount of core-hours delivered to researchers from two leadership-class systems; and success in delivering on the allocation split of roughly 60%, 30%, and 10% of core-hours offered for the Innovative and Novel Computational Impact on Theory and Experiment (INCITE), Advanced Scientific Computing Research Leadership Computing Challenge (ALCC), and Director's Discretionary (DD) programs, respectively (see Operational Performance section). These accomplishments, coupled with the high utilization rates (overall and capability usage), represent the fulfillment of the promise of both leadership-class machines: efficient facilitation of leadership-class computational applications. Table ES.1 presents a summary of the 2019 OLCF metric targets and the associated results. More information can be found in the Operational Performance section for each OLCF resource. The scientific accomplishments of OLCF users are a strong indication of long-term operational success, with publications this year in such notable journals and publications as Nature, Nature Physics, Nature Plants, Physical Review X, Journal of the American Physical Society, Cell, Nano Letters, and Trends in Biotechnology. Crucial domain-specific discoveries facilitated by resources at the OLCF are described in the High Performance Computing Facility Operational Assessment 2019 Oak Ridge Leadership Computing Facility (OAR) Strategic Results section. For example, researchers used Summit to pinpoint and understand the production of proteins from genetic information, including mutations and the functional expression of disease (Section 8.2).

97 MATHEMATICS AND COMPUTING↗

Numerical Experiments and Identification of Areas for Further Consideration (B639388 Subcontract Quarter II Report)

The second quarter of the project was spent carrying out numerical experiments on a variety of test matrices, using a combination of precisions, with the intent to study the numerical properties (namely, accuracy and convergence behavior) of the Conjugate Gradient (CG) method. We summarize our findings in the remainder of the document. Other activities include attending biweekly xSDK meetings. We are also currently collaborating with Steven Thomas (NREL) on the numerical stability analysis of “low-synch” Gram-Schmidt routines, for potential application within high-performance GMRES variants. This work is in progress. The subsequent quarter will be spent delving into the numerical analysis in order to provide theoretical explanation for the behavior observed in our experiments.

97 MATHEMATICS AND COMPUTING↗

Preliminary Theoretical Analysis of Mixed Precision Krylov Subspace Methods (Q3 Report)

The third quarter of the project was spent performing theoretical finite precision analysis of Krylov subspace method variants that use mixed precision. Our focus here is on the Conjugate Gradient (CG) method and the Lanczos method. We have performed an analysis of maximum attainable accuracy for the classical CG method in which 3 precisions are used: a working precision ε, a precision ε IP for the inner product computations, and a precision ε MV for the matrix-vector products. Our results show that performing inner product computations in lower precision does not affect the attainable accuracy. Further, we have performed a complete error analysis of the s-step Lanczos algorithm. In this case, we show that the numerical behavior of the algorithm can be significantly improved by using extra precision in a small part of the computation. We summarize the main theorems in the remainder of the document. Other activities include attending biweekly xSDK meetings. The subsequent quarter will be spent finalizing these results into technical reports and/or manuscripts for submission to journals, as well as identifying opportunities for future work.

97 MATHEMATICS AND COMPUTING↗

Replicated Computational Results (RCR) Report for "Adaptive Precision Block-Jacobi for High Performance Preconditioning in the Ginkgo Linear Algebra Software''

In, a practical implementation of a novel adaptive precision block-Jacobi preconditioner is introduced. In particular, the authors present a heavily-tuned GPU implementation of the adaptive precision block-Jacobi preconditioner within the Ginkgo numerical linear algebra library. The performance of the methodology and implementation is demonstrated using the proposed preconditioning scheme within Ginkgo’s high-performance Conjugate Gradient (CG) implementation on an NVIDIA Volta GPU. In this report, we replicate a subset of the computational results presented in. The focus is generating results from Fig. 9 to evaluate the performance of using Ginkgo’s CG solver integrated with either the full or the adaptive precision block-Jacobi preconditioner applied to a variety of test cases

97 MATHEMATICS AND COMPUTING↗

Optimization-based algorithms for nonlinear mechanics and frictional contact

An optimization-based strategy for solving nonlinear mechanics problems is proposed. In contrast to typical nonlinear equation solver algorithms that aim to find zeros in the residual force function, we minimize an energy (or energy-like) function to encourage solutions which are locally stable equilibria. These smooth and potentially non-convex objective functions are minimized using a preconditioned conjugate-gradient trust-region algorithm. Contact is formulated as an inequality constrained minimization problem, and is solved with an augmented Lagrangian algorithm. Friction is included in the approach via a regularized quasi-potential energy, and other dissipative behavior is included through the use of variational constitutive updates. Finally, to accelerate convergence rates for the Lagrange multipliers, we propose a novel multiplier update algorithm utilizing the Fischer-Burmeister function, and demonstrate super-linear solver convergence for some applications.

42 ENGINEERING↗

Report on hypre performance on AMD GPUs

In this report, the performance of the algebraic multigrid solver used as a preconditioner for conjugate gradient is investigated on 2 nodes of Spock, an early access system at ORNL with 4 MI-100 GPUs and a 64-core Rome CPU per node. We compare GPU and CPU performance for three different diffusion problems using increasing problem sizes.

97 MATHEMATICS AND COMPUTING↗

A discrete variable approximation to minimum weight panel designs subject to a supersonic flutter-speed constraint.

A numerical technique is presented that determines the minimum-weight thickness distribution for a one-dimensional, simply supported panel subjected to a parallel supersonic flow on one side. An aeroelastic eigenvalue, which characterizes flutter speed, is held constant. The governing differential equations are approximated by sets of difference equations adjoined to the weight function via a penalty function. A conjugate gradient method then solves the resulting sequence of unconstrained minimization problems. Numerical results are obtained without imposing a minimum thickness constraint; the effect of constant inplane tensile and compressive stresses is also investigated.

Pierson, B. L.↗

Iterative explicit guidance for low thrust spacecraft.

A retargeting procedure is developed for use as a nonlinear low thrust guidance scheme. The selection of a control program composed of a sequence of inertially fixed thrust-acceleration vectors permits all trajectory computations to be made with closed form expressions, and allows the controls to be represented by constant parameters, thrust-acceleration vectors and thrusting times. By requiring each trajectory to be time optimal, the guidance problem is transformed into a parameter optimization problem which is solved by the conjugate gradient method. The scheme is applied to a low thrust capture mission, and the results of computer simulations are presented.

Jacobson, R. A.↗

The use of minimum order state observers in digital flight-control systems.

This paper deals with the problem of selecting the 'arbitrary' design parameters of digital state observers when they are being used as a part of a digital flight-control system. A cost index is developed which indicates the output noise caused by input quantization due to analog-to-digital conversion. The cost index assumes that the input quantization error is uniformly distributed over the least-significant-bit of the conversion. Formulas relating the cost index to the observer design parameters are presented. The cost index is minimized with respect to the design parameters using a conjugate gradient algorithm. An example of the theory is presented in which a digital observer is designed so that a satisfactory digital flight-control system is obtained starting from an unacceptable one.

Montgomery, R. C.↗

Elasto-plastic bending of cracked plates, including the effects of crack closure

A capability for solving elasto-plastic plate bending problems is developed using assumptions consistent with Kirchhoff plate theory. Both bending and extensional modes of deformation are admitted with the two modes becoming coupled as yielding proceeds. Equilibrium solutions are obtained numerically by determination of the stationary point of a functional which is analogous to the potential strain energy. The stationary value of the functional for each load increment is efficiently obtained through use of the conjugate gradient. This technique is applied to the problem of a large centrally through cracked plate subject to remote circular bending. Comparison is drawn between two cases of the bending problem. The first neglects the possibility of crack face interference with bending, and the second includes a kinematic prohibition against the crack face from passing through the symmetry plane. Results are reported which isolate the effects of elastoplastic flow and crack closure.

Jones, D. P.↗

Sparse matrix methods based on orthogonality and conjugacy

A matrix having a high percentage of zero elements is called spares. In the solution of systems of linear equations or linear least squares problems involving large sparse matrices, significant saving of computer cost can be achieved by taking advantage of the sparsity. The conjugate gradient algorithm and a set of related algorithms are described.

Lawson, C. L.↗

Gust alleviation for a STOL transport by using elevator, spoilers, and flaps

Control laws were developed to investigate methods of alleviating the response of a STOL transport to gusty air. The transport considered in the study had triple-slotted, externally blown jet flaps and a large T-tail. The control devices used were the elevator, spoilers, and flaps. A hybrid computing system was used to simulate linearized longitudinal dynamics of the aircraft and to implement a conjugate gradient optimal search algorithm. The aircraft was simulated in the low-speed approach condition only. Feedback control matrices were found which minimized the average of a quadratic functional involving passenger compartment accelerations, pitch angle and rate, flight path angle and speed variations. The optimization was performed for artificially designed gust inputs in the form of predetermined rectangular waveforms. Results were obtained for elevator, spoilers, and flaps acting singly and in combination. Additional results were obtained for unit sinusoidal gust inputs by using the gain matrices computed for the artificial test gusts. Various sensor configurations were also investigated.

Lallman, F. J.↗

Control system design using frequency domain models and parameter optimization, with application to supersonic inlet controls

A technique is described for designing feedback control systems using frequency domain models, a quadratic cost function, and a parameter optimization computer program. FORTRAN listings for the computer program are included. The approach is applied to the design of shock position controllers for a supersonic inlet. Deterministic or random system disturbances, and the presence of random measurement noise are considered. The cost function minimization is formulated in the time domain, but the problem solution is obtained using a frequency domain system description. A scaled and constrained conjugate gradient algorithm is used for the minimization. The approach to a supersonic inlet included the calculations of the optimal proportional-plus integral (PI) and proportional-plus-integral-plus-derivative controllers. A single-loop PI controller was the most desirable of the designs considered.

Seidel, R. C.↗

Optimal control of a variable spin speed CMG system for space vehicles

Many future NASA programs require very high accurate pointing stability. These pointing requirements are well beyond anything attempted to date. This paper suggests a control system which has the capability of meeting these requirements. An optimal control law for the suggested system is specified. However, since no direct method of solution is known for this complicated system, a computation technique using successive approximations is used to develop the required solution. The method of calculus of variations is applied for estimating the changes of index of performance as well as those constraints of inequality of state variables and terminal conditions. Thus, an algorithm is obtained by the steepest descent method and/or conjugate gradient method. Numerical examples are given to show the optimal controls.

Liu, T. C.↗

Transfer-function-parameter estimation from frequency response data: A FORTRAN program

A FORTRAN computer program designed to fit a linear transfer function model to given frequency response magnitude and phase data is presented. A conjugate gradient search is used that minimizes the integral of the absolute value of the error squared between the model and the data. The search is constrained to insure model stability. A scaling of the model parameters by their own magnitude aids search convergence. Efficient computer algorithms result in a small and fast program suitable for a minicomputer. A sample problem with different model structures and parameter estimates is reported.

Seidel, R. C.↗

Transfer-function parameters

Computer program fits linear-factored form transfer function to given frequency-response data. Program is based on conjugate-gradient search procedure that minimizes error between given frequency-response data and frequency response of transfer function that is supplied by user.

Seidel, R. C.↗