Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Performance optimizations”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Perform - A performance optimizing computer program for dynamic systems subject to transient loadings

A description and applications of a computer capability for determining the ultimate optimal behavior of a dynamically loaded structural-mechanical system are presented. This capability provides characteristics of the theoretically best, or limiting, design concept according to response criteria dictated by design requirements. Equations of motion of the system in first or second order form include incompletely specified elements whose characteristics are determined in the optimization of one or more performance indices subject to the response criteria in the form of constraints. The system is subject to deterministic transient inputs, and the computer capability is designed to operate with a large linear programming on-the-shelf software package which performs the desired optimization. The report contains user-oriented program documentation in engineering, problem-oriented form. Applications cover a wide variety of dynamics problems including those associated with such diverse configurations as a missile-silo system, impacting freight cars, and an aircraft ride control system.

Pilkey, W. D.↗

Accelerating detector simulations with Celeritas: profiling and performance optimizations

Celeritas is a GPU-optimized MC particle transport code designed to meet the growing computational demands of next-generation HEP experiments. It provides efficient simulation of EM physics processes in complex geometries with magnetic fields, detector hit scoring, and seamless integration into Geant4-driven applications to offload EM physics to GPUs. Recent efforts have focused on performance optimizations and expanding profiling capabilities. This paper presents some key advancements, including the integration of the Perfetto system profiling tool for detailed performance analysis and the development of track-sorting methods to improve computational efficiency.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Accelerating detector simulations with Celeritas: Profiling and performance optimizations

Celeritas is a GPU-optimized Monte Carlo (MC) particle transport code designed to meet the growing computational demands of next-generation high energy physics (HEP) experiments. It provides efficient simulation of electromagnetic (EM) physics processes in complex geometries with magnetic fields, detector hit scoring, and seamless integration into Geant4-driven applications to offload EM physics to GPUs. Recent efforts have focused on performance optimizations and expanding profiling capabilities. This paper presents some key advancements, including the integration of the Perfetto system profiling tool for detailed performance analysis and the development of track-sorting methods to improve computational efficiency.

Lund, Amanda [Argonne National Laboratory (ANL)]↗

Performance optimization of helicopter rotor blades

As part of a center-wide activity at NASA Langley Research Center to develop multidisciplinary design procedures by accounting for discipline interactions, a performance design optimization procedure is developed. The procedure optimizes the aerodynamic performance of rotor blades by selecting the point of taper initiation, root chord, taper ratio, and maximum twist which minimize hover horsepower while not degrading forward flight performance. The procedure uses HOVT (a strip theory momentum analysis) to compute the horse power required for hover and the comprehensive helicopter analysis program CAMRAD to compute the horsepower required for forward flight and maneuver. The optimization algorithm consists of the general purpose optimization program CONMIN and approximate analyses. Sensitivity analyses consisting of derivatives of the objective function and constraints are carried out by forward finite differences. The procedure is applied to a test problem which is an analytical model of a wind tunnel model of a utility rotor blade.

Walsh, Joanne L.↗

Strategy to Develop a Control Scheme for Core Thermal Performance Optimization

This report outlines a three-year research plan for optimizing reactor core thermal performance through use of a Digital Twin (DT) model and a set of neutron detectors to correspondingly adapt the reactor control strategy. It identifies reactor features and achievable neutron measurements that factor into this optimization task. It considers spatial effects that are important and how they can be managed by a real-time algorithm that controls reactivity actuators. The report presents a set of tasks, along with corresponding methods, for accomplishing this objective. The final task in this set is to validate the combined optimization procedure in the Purdue University’s PUR-1 reactor, for which we have an agreement of understanding. Additionally, we describe an alternative approach for overcoming some of the limitations inherent in the above approach, which includes the computational costs of high-fidelity simulations and achievable integration of neutron measurements into the DT. The proposed algorithms lower the technology readiness level (TRL) of the overall approach as some additional analysis and the development of a new neutron detector concept are necessary. In particular, a sensor measuring the gradient of the neutron flux is required, and the prototype is expected to be validated in the PUR-1 reactor. The potential benefits and the opportunity to significantly push forward the state-of-the-art make this approach worthy of further exploration in the upcoming years.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Method and Apparatus for Performance Optimization Through Physical Perturbation of Task Elements

The invention is an apparatus and method of biofeedback training for attaining a physiological state optimally consistent with the successful performance of a task, wherein the probability of successfully completing the task is made is inversely proportional to a physiological difference value, computed as the absolute value of the difference between at least one physiological signal optimally consistent with the successful performance of the task and at least one corresponding measured physiological signal of a trainee performing the task. The probability of successfully completing the task is made inversely proportional to the physiological difference value by making one or more measurable physical attributes of the environment in which the task is performed, and upon which completion of the task depends, vary in inverse proportion to the physiological difference value.

Prinzel, Lawrence J., III↗

The 15-meter antenna performance optimization using an interdisciplinary approach

A 15-meter diameter deployable antenna has been built and is being used as an experimental test system with which to develop interdisciplinary controls, structures, and electromagnetics technology for large space antennas. The program objective is to study interdisciplinary issues important in optimizing large space antenna performance for a variety of potential users. The 15-meter antenna utilizes a hoop column structural concept with a gold-plated molybdenum mesh reflector. One feature of the design is the use of adjustable control cables to improve the paraboloid reflector shape. Manual adjustment of the cords after initial deployment improved surface smoothness relative to the build accuracy from 0.140 in. RMS to 0.070 in. Preliminary structural dynamics tests and near-field electromagnetic tests were made. The antenna is now being modified for further testing. Modifications include addition of a precise motorized control cord adjustment system to make the reflector surface smoother and an adaptive feed for electronic compensation of reflector surface distortions. Although the previous test results show good agreement between calculated and measured values, additional work is needed to study modelling limits for each discipline, evaluate the potential of adaptive feed compensation, and study closed-loop control performance in a dynamic environment.

Grantham, William L.↗

A study of methods to predict and measure the transmission of sound through the walls of light aircraft. Numerical method for analyzing the optimal performance of active noise controllers

An optimal active noise controller is formulated and analyzed for three different active noise control problems. The first problem formulated is the active control of enclosed or partially enclosed harmonic sound fields where the noise source strengths and enclosure boundary description are known. The enclosure boundary is described by either pressure, velocity, or impedance boundary conditions. The second problem formulated is the active control of the free field power radiated from a distributed noise source with a known time harmonic surface velocity. The third problem formulated is the active control of enclosed or partially enclosed harmonic sound field where the noise source strengths of enclosure boundary description may not be known. All three formulations are derived using an indirect boundary element technique. Formulation and verification of an indirect boundary element method is presented. The active noise controller formulations for enclosures are capable of analyzing systems with generalized enclosure shapes, point noise sources, and/or locally reacting impedance boundary conditions. For each formulation, representative results of optimal active noise controller case studies are presented, and some general conclusions are drawn.

Mollo, Christopher G.↗

Performance optimization of an MHD generator with physical constraints

A method to optimize the Faraday MHD generator performance under a prescribed set of electrical and magnet constraints is described. The results of generator performance calculations using this technique are presented for a very large MHD/steam plant. The differences between the maximum power and maximum net power generators are described. The sensitivity of the generator performance to the various operational parameters are presented.

Pian, C. C. P.↗

Data-Driven Performance Optimization of Gamma Spectrometers With Many Channels

In gamma spectrometers with variable spectroscopic performance across many channels (e.g., many pixels or voxels), a tradeoff exists between including data from successively worse-performing readout channels and increasing efficiency. Brute-force calculation of the optimal set of included channels is exponentially infeasible as the number of channels grows, and approximate methods are required. In this work, we present a data-driven framework for attempting to find near-optimal sets of included detector channels. The framework leverages non-negative matrix factorization (NMF) to learn the behavior of gamma spectra across the detector and clusters similarly-performing detector channels together. Performance comparisons are then made between spectra with channel clusters removed, which is more feasible than brute force. The framework is general and can be applied to arbitrary, user-defined performance metrics depending on the application. We apply this framework to optimizing gamma spectra measured by H3D M400 CdZnTe (CZT) spectrometers, which exhibit variable performance across their crystal volumes. In particular, we show several examples optimizing various performance metrics for uranium and plutonium gamma spectra in non-destructive assay (NDA) for nuclear safeguards, and explore trends in performance versus parameters such as clustering algorithm type. We also compare the NMF + clustering pipeline to several non-machine-learning (ML) algorithms, including several greedy algorithms. Although, we find that the NMF + clustering pipeline tends to find the best-performing set of detector voxels, significantly improving over the unoptimized spectra, but that a greedy accumulation of spectra segmented by detector depth can, in some cases, give similar performance improvements in much less computation time.

Energy resolution↗

Performance Optimization Methods for a Memory-Bound, Unstructured-Grid CFD Application on Massively Parallel GPU Platforms

Computational performance of the FUN3D unstructured-grid computational fluid dynamics (CFD) application on massively parallel GPU environments is memory-bound and highly dependent upon efficient reads from and atomic updates to the irregular cell-, edge-, and node-based data structures. In this talk, we present recent efforts into optimizing select performance-critical kernels on NVIDIA Tesla V100 and A100 GPUs and AMD CDNA MI100 GPUs. A novel use of L2 cache residency controls and asynchronous loads into on-chip shared memory are explored on the A100 GPU for the sparse iterative solver, which is dominated by mixed-precision, sparse matrix vector multiplication. Demonstrations show that these methods improve global memory bandwidth utilization by 13.5% on the A100 GPU. Several techniques are also presented that use registers and/or shared memory to facilitate array transposition and aggregation which combine to reduce the frequency and increase the cache efficiency of floating-point atomic updates to the irregular data structures. These methods are demonstrated to improve the kernel throughput by nearly 500% on select kernels on the AMD MI100 over atomic updates directly to global memory. Overall, both V100 and A100 GPUs outperformed the MI100 GPU on kernels dominated by double-precision atomic updates; however, the techniques demonstrated here reduced the performance gap and improved the MI100 performance.

GPU CPU unstructured CFD memory↗

Advanced Modeling of Beam Physics and Performance Optimization for Nuclear Physics Colliders

High energy colliders provide a critical tool in nuclear physics study by probing the fundamental structure and dynamics of matter. To maximize the potential of scientific discovery in nuclear physics study, it is important to optimize the parameters of these colliders to attain the best performance. The performance of a collider is typically measured by its integrated luminosity of colliding beams since the probability of a new event is proportional to the integrated luminosity. However, the achievable luminosity is limited by the electromagnetic interactions (beam-beam effects) of two colliding beams at higher energy, and the interplay between the space-charge effects and the beam-beam effects at lower energy. To achieve the best performance of a collider means to attain the highest luminosity of the collider with optimized collider parameters. Optimizing the collider’s machine parameters is both computationally and experimentally expensive. A fast and robust computational framework including beam-beam and space-charge effects will be critical to attaining the best performance of the collider. In this project, we will study the beam dynamics challenges, specifically the interplay of the space-charge and the beam-beam effects, and the machine tuning models for maximizing the performance of RHIC experiments. We will develop an advanced modeling framework based on first-principles physical simulations, lattice models and the state-of-the-art machine learning methods and apply this framework to performance improvement of the RHIC in operation. We will build data manipulation packages to connect the simulation data and the experimental data with the framework, develop a self-consistent hybrid model of space-charge and beam-beam effects, study underlying physics mechanisms, build surrogate models using the labeled data, integrate the models into the advanced modeling framework, and apply the framework to RHIC luminosity (STAR and sPHENIX) optimization. The success of this project would substantially improve the performance of existing and future colliders and increase the opportunity for scientific discovery.

43 PARTICLE ACCELERATORS↗

Task Parallelism to Optimize Performance of Environmental Modeling Software

Climate modeling is an integral part of environmental research, from studying rare phenomena to predicting future climate trends. The need for more accurate models is only growing, but as climate modeling capabilities advance, existing workflows require optimization to recoup performance. A solution comes in the form of task parallelism, a novel programming capability that provides an opportunity for optimization at execution time by allowing tasks to be executed in parallel, reducing runtime significantly. Using Parsl, an intuitive and scalable parallel scripting library for Python, we implement task parallelism within support software to aid in the continuous advancement of climate modeling technology.

54 ENVIRONMENTAL SCIENCES↗

Design and Performance Optimizations of Advanced Erosion-Resistant Low Conductivity Thermal Barrier Coatings for Rotorcraft Engines

Thermal barrier coatings will be more aggressively designed to protect gas turbine engine hot-section components in order to meet future rotorcraft engine higher fuel efficiency and lower emission goals. For thermal barrier coatings designed for rotorcraft turbine airfoil applications, further improved erosion and impact resistance are crucial for engine performance and durability, because the rotorcraft are often operated in the most severe sand erosive environments. Advanced low thermal conductivity and erosion-resistant thermal barrier coatings are being developed, with the current emphasis being placed on thermal barrier coating toughness improvements using multicomponent alloying and processing optimization approaches. The performance of the advanced thermal barrier coatings has been evaluated in a high temperature erosion burner rig and a laser heat-flux rig to simulate engine erosion and thermal gradient environments. The results have shown that the coating composition and architecture optimizations can effectively improve the erosion and impact resistance of the coating systems, while maintaining low thermal conductivity and cyclic oxidation durability

Zhu, Dongming↗

Design, Modeling and Performance Optimization of a Novel Rotary Piezoelectric Motor

This work has demonstrated a proof of concept for a torsional inchworm type motor. The prototype motor has shown that piezoelectric stack actuators can be used for rotary inchworm motor. The discrete linear motion of piezoelectric stacks can be converted into rotary stepping motion. The stacks with its high force and displacement output are suitable actuators for use in piezoelectric motor. The designed motor is capable of delivering high torque and speed. Critical issues involving the design and operation of piezoelectric motors were studied. The tolerance between the contact shoes and the rotor has proved to be very critical to the performance of the motor. Based on the prototype motor, a waveform optimization scheme was proposed and implemented to improve the performance of the motor. The motor was successfully modeled in MATLAB. The model closely represents the behavior of the prototype motor. Using the motor model, the input waveforms were successfully optimized to improve the performance of the motor in term of speed, torque, power and precision. These optimized waveforms drastically improve the speed of the motor at different frequencies and loading conditions experimentally. The optimized waveforms also increase the level of precision of the motor. The use of the optimized waveform is a break-away from the traditional use of sinusoidal and square waves as the driving signals. This waveform optimization scheme can be applied to any inchworm motors to improve their performance. The prototype motor in this dissertation as a proof of concept was designed to be robust and large. Future motor can be designed much smaller and more efficient with lessons learned from the prototype motor.

Duong, Khanh A.↗