Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Performance optimizations”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

BISON fuel performance modeling optimization for experiment X447 and X447A using axial swelling and cladding strain measurements

With the recent need to qualify new reactor designs such as the Versatile Test Reactor (VTR), fuel performance calculations need to be performed to determine safety criteria of the proposed designs. In order to validate the fuel performance results obtained by a fuel performance code, BISON, for new reactor designs, legacy fuel from EBR-II and FFTF MFF with Post -Irradiation Examination (PIE) data need to be used as validation cases to benchmark models. Here in this work, BISON has been paired with the Fuels Irradiation & Physics Database (FIPD) and IFR Materials Information System (IMIS) to supply PIE data for comparison with simulations of EBR-II experiments X447/X447A. X447/X447A were assessed by implementing models for Fuel Cladding Chemical Interaction (FCCI) within BISON and optimizing the friction coefficient between the fuel surface and the cladding, the anisotropic swelling factor, and the HT9 first thermal creep scalar (which scales the first term in the HT9 creep equation) to best match the PIE axial fuel swelling height and cladding profilometry for all pins in X447/X447A. The optimal values were found using a generic algorithm developed to select different values for the three parameters until end criteria was met and error couldn’t be reduced further. The BISON-simulated cladding profilometry was evaluated using Standard Error of the Estimate (SEE) to account for the profile shape of the cladding profilometry. Optimal values for the friction coefficient, anisotropic fuel swelling factor, and HT9 first thermal creep scalar were found to best fit the BISON simulation results to the PIE measurements found in IMIS and FIPD. Improvements to current models are suggested to account for the underprediction of fuel swelling at low burnups and the overprediction of fuel swelling at higher burnups observed for the axial fuel swelling height. Although two pins in EBR-II X447/X447A (DP70 and DP75) were known to fail due to FCCI, none of the pins simulated in BISON reached a cumulative damage fraction (CDF) above 0.008 with FCCI correlations coupled in the BISON simulations. The error estimate generated for all pins in X447/X447A using optimal values was 209 µm, which is deemed acceptable.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Scalable molecular dynamics on CPU and GPU architectures with NAMD

NAMD is a molecular dynamics program designed for high-performance simulations of very large biological objects on CPU- and GPU-based architectures. NAMD offers scalable performance on petascale parallel supercomputers consisting of hundreds of thousands of cores, as well as on inexpensive commodity clusters commonly found in academic environments. It is written in C++ and leans on Charm++ parallel objects for optimal performance on low-latency architectures. NAMD is a versatile, multipurpose code that gathers state-of-the-art algorithms to carry out simulations in apt thermodynamic ensembles, using the widely popular CHARMM, AMBER, OPLS, and GROMOS biomolecular force fields. Here, we review the main features of NAMD that allow both equilibrium and enhanced-sampling molecular dynamics simulations with numerical efficiency. We describe the underlying concepts utilized by NAMD and their implementation, most notably for handling long-range electrostatics; controlling the temperature, pressure, and pH; applying external potentials on tailored grids; leveraging massively parallel resources in multiple-copy simulations; and hybrid quantum-mechanical/molecular-mechanical descriptions. We detail the variety of options offered by NAMD for enhanced-sampling simulations aimed at determining free-energy differences of either alchemical or geometrical transformations and outline their applicability to specific problems. Last, we discuss the roadmap for the development of NAMD and our current efforts toward achieving optimal performance on GPU-based architectures, for pushing back the limitations that have prevented biologically realistic billion-atom objects to be fruitfully simulated, and for making large-scale simulations less expensive and easier to set up, run, and analyze. NAMD is distributed free of charge with its source code at www.ks.uiuc.edu.

high-performance computing↗

Turbine scale and siting considerations in wind plant layout optimization and implications for capacity density

Improvements in wind energy technology, reduced costs, and ambitious clean energy goals have led to projections of high wind contribution in coming years. Developing methodologies to design wind plants with a variety of siting constraints and turbine sizes helps enable high wind penetration, and gain a better understanding of how wind plants are sensitive to setback constraints and turbine design. In this paper, we present a two-step optimization method to simultaneously determine the optimal number of turbines and their locations in a wind plant domain divided into many small, discrete parcels. We present the optimized performance metrics of a wind plant optimized with different turbine sizes and ratings, and with different siting restrictions within the wind plant. Our results indicate that taller and larger turbines are more sensitive to increasing siting constraints. We also compare the optimal wind plant layouts and performance for wind plants optimized for minimum COE and maximum profit. Wind plants optimized for profit had 130%-190% of the capacity of plants optimized for COE, which demonstrates that the optimal results are greatly affected by the objective function, which should be carefully considered. Finally, in this paper we demonstrate the effect of increasing siting constraints on wind plant capacity density, and how the results change when different land areas are used to calculate capacity density. When using the entire wind plant boundary area to determine capacity density, increasing siting constraints decreases the capacity density. However, when we only use the available area (the area left after removing the siting constraints) to calculate the capacity density, increasing the siting constraints increases capacity density. This is a critical insight because of how capacity density is typically defined and used in research, and has important implications for assessment of technical potential and capacity expansion modeling, as well as future wind deployment potential.

17 WIND ENERGY↗

On-Chip Batteries as Distributed Energy Sources in Heterogeneous 2.5D/3D Integrated Circuits

Energy efficiency in digital systems faces challenges due to the constraints imposed by small-scale transistors. Moreover, the growing demand for portable consumer electronics necessitates the use of compact energy sources. To address these challenges, heterogeneous 3D IC technology has emerged as a promising solution for the former. Regarding the latter, we propose the concept of distributed batteries within a heterogeneous 3D IC. This approach involves utilizing multiple smaller batteries with different specifications among different modules of 3D ICs. This approach optimizes performance and overcomes limitations associated with both 3D ICs and conventional power delivery methods. Distributed batteries play a vital role in effectively managing the heat generated by energy sources and modules within a 3D IC. Furthermore, they contribute to achieving a uniform distribution of heat throughout the entire structure, which ultimately ensures the optimal performance of the batteries and modules. The simulation results indicate a 40 percent enhancement in achieving a more even distribution of generated heat. Additionally, the proposed distributed battery techniques improve power delivery, enhance reliability, and enable optimized voltage regulation while improving efficiency. In addition to the primary benefits, alternative configurations of the proposed approach can offer extra energy storage capacity and act as efficient electromagnetic shields, resulting in an impressive reduction of external electromagnetic noises by 60 dB.

47 OTHER INSTRUMENTATION↗

Spike-Free Adaptive Sliding Mode Control: Application to Permanent Magnet Synchronous Motors

A new methodology for adaptive sliding mode control (ASMC) has been widely used to improve the control performance in various systems. This method exhibits several advantages, including low sliding mode control (SMC) chattering, no knowledge of the system disturbance bound, and no overestimation of the control gain. Despite its advantages, this method can be hampered by the spike phenomenon, slow control gain convergence, and difficulty in achieving optimal performance under varying disturbances. Consequently, this article proposes a spike-free ASMC method with a disturbance observer (DOB) to address these problems. Previous ASMC methods have been analyzed via simulations to verify the aforementioned problems. Here, this analysis highlights the need for disturbance compensation and improvements in the SMC gain adaptation law. Therefore, a DOB is designed to mitigate the spike phenomenon by compensating for disturbances. Subsequently, an SMC gain adaptation law based on disturbance error estimation is designed to eliminate the spike phenomenon completely. The proposed adaptation law makes the SMC gain to converge to a slightly higher value than the disturbance estimation error. Consequently, the proposed method not only eliminates the spike phenomenon, but also ensures optimal performance under varying disturbances. The performance of the proposed method is experimen tally validated through a comparative study.

42 ENGINEERING↗

The Simons Observatory: Design, Optimization, and Performance of Low-Frequency Detectors

The Simons Observatory (SO) is a cosmic microwave background (CMB) experiment located in the Atacama Desert in Chile that will make precise temperature and polarization measurements over six spectral bands ranging from 27 to 285 GHz. Three small aperture telescopes (SATs) and one large aperture telescope (LAT) will house ~60,000 detectors and cover angular scales between one arcminute and tens of degrees. We present the performance of the dichroic, low-frequency (LF) lenslet-coupled sinuous antenna transition-edge sensor (TES) bolometer arrays with bands centered at 27 and 39 GHz. The LF focal plane will primarily characterize Galactic synchrotron emission as a critical part of foreground subtraction from CMB data. We will discuss the design, optimization, and current testing status of these pixels.

79 ASTRONOMY AND ASTROPHYSICS↗

pnnl/EZBattery

Cell performance optimization is important for improving the system efficiency of a redox flow battery. To gain better insights into key controlling factors Of system efficiency, this work first proposed a theoretical model for a unit cell by extending a two-dimensional analytic model to a full cell. The model is then used for cell performance optimization after validating it with experimental and numerical modeling data.

Bao, Jie↗

Towards 5G-Enabled Operational Technology for Process Monitoring and Network Slicing

Cyber-Physical Systems (CPS) are deployed to monitor physical processes in critical cyber-enabled services like power generation. However, CPS ecosystems are typically designed without robust security. While it is important to ensure optimal performance of the Operational Technology (OT) environments, security cannot be overlooked. To modernize traditional OT services, 5G technology is being integrated. 5G technology offers low latency and high availability, making it a suitable infrastructure for managing and monitoring physical processes. How-ever, integrating 5G mechanisms into large-scale OT networks introduces new implementation and performance challenges. Therefore, this paper presents a 5G-enabled CPS architecture (5G-CPS) that describes the necessary components, services, and communication protocols and conducts feasibility study to integrate 5G technology in industrial control system networks to understand the performance merits. The 5G-CPS architecture aims to minimize implementation and operational challenges associated with integrating 5G technology into constrained OT.

Aguayo, Jared M.↗

Evaluating adaptive and predictive power management strategies for optimizing visualization performance on supercomputers

Power is becoming an increasingly scarce resource on the next generation of supercomputers, and should be used wisely to improve overall performance. One strategy for improving power usage is hardware overprovisioning, i.e., systems with more nodes than can be run at full power simultaneously without exceeding the system-wide power limit. With this study, we compare two strategies for allocating power throughout an overprovisioned system – adaptation and prediction – in the context of visualization workloads. While adaptation has been suitable for workloads with more regular execution behaviors, it may not be as suitable on visualization workloads, since they can have variable execution behaviors. This study considers a total of 104 experiments, which vary the rendering workload, power budget, allocation strategy, and node concurrency, including tests processing data sets up to 1 billion cells and using up to 18,432 cores across 512 nodes. Overall, we find that prediction is a superior strategy for this use case, improving performance up to 27% compared to an adaptive strategy.

97 MATHEMATICS AND COMPUTING↗

Etching-Chemistry-Driven Ruthenium Doping on Ti 3 C 2 T x MXene for Optimizing Electrochemical Performance

We demonstrate that the etching chemistry used during MXene synthesis from Ti 3 AlC 2 MAX phase significantly influences surface functionalization and structural vacancies, which in turn affect ruthenium (Ru) ion interactions. Using hydrofluoric acid (HF) and ammonium bifluoride (NH 4 HF 2 ) as etchants, we obtained MXene surfaces with distinct functional groups and Ti vacancies that impact Ru ion interactions and electrochemical performance. Both MXene variants (labeled MX(H) and MX(N), respectively) exhibited negative zeta potentials in their pristine state, but upon the addition of Ru the zeta potential for MX(H) reached 12.9 mV while that for MX(N) remained negative at −6.4 mV. This adsorption resulted in a 14.4-fold increase in the specific capacitance of MX(H)/Ru compared to pristine MX(H), whereas MX(N)/Ru exhibited only a 4.4-fold increase over its pristine counterpart. X-ray diffraction analysis identified the formation of ammonium titanium oxide fluoride, (NH 4 ) 3 TiOF 5 , on MX(N), which likely contributed to its reduced Ru adsorption. X-ray photoelectron spectroscopy suggested the presence of Ti vacancies in both MXene variants; however, their behavior toward Ru accommodation differed markedly, with MX(H) showing the most obvious shift in the Ti 2p peak in the XPS survey spectrum, while MX(N) showed the most obvious shift in the C 1s peak. Electron paramagnetic resonance spectroscopy further demonstrated a distinct alteration in the spectral signatures of MX(H) upon Ru addition, in contrast to the negligible changes in MX(N), indicating effective passivation of the Ti defect sites in MX(H) via vacancy-assisted Ru doping. Cyclic voltammetry showed that Ru-incorporated MX(H) nanocomposites exhibit more efficient redox-active sites, as reflected in their higher capacitance values. These findings highlight the pivotal role of MXene surface chemistry in controlling cation adsorption, providing valuable insights for the rational design of high-performance electrodes.

2D surface engineering↗

Criticality analysis of nuclear binding energy neural networks

Machine learning methods, in particular deep learning methods such as artificial neural networks (ANNs) with many layers, have become widespread and useful tools in nuclear physics. However, these ANNs are typically treated as ‘black boxes’, with their architecture (width, depth, and weight/bias initialization) and the training algorithm and parameters chosen empirically by optimizing learning based on limited exploration. We test a non-empirical approach to understanding and optimizing nuclear physics ANNs by adapting a criticality analysis based on renormalization group flows in terms of the hyperparameters for weight/bias initialization, training rates, and the ratio of depth to width. This treatment utilizes the statistical properties of neural network initialization to find a generating functional for network outputs at any layer, allowing for a path integral formulation of the ANN outputs as a Euclidean statistical field theory. We use a prototypical example to test the applicability of this approach: a simple ANN for nuclear binding energies. We find that with training using a stochastic gradient descent optimizer, the predicted criticality behavior is realized, and optimal performance is found with critical tuning. However, the use of an adaptive learning algorithm leads to somewhat superior results without concern for tuning and thus obscures the analysis. Nevertheless, the criticality analysis offers a way to look within the black box of ANNs, which is a first step towards potential improvements in network performance beyond using adaptive optimizers.

artificial neural network↗

Adaptive Time Step Control for Multirate Infinitesimal Methods

Multirate methods have been used for decades to temporally evolve initial-value problems in which different components evolve on distinct time scales, and thus use of different step sizes for these components can result in increased computational efficiency. Generally, such methods select these different step sizes based on experimentation or stability considerations. For problems that evolve on a single time scale, adaptivity approaches that strive to control local temporal error are widely used to achieve numerical results of a desired accuracy with minimal computational effort, while alleviating the need for manual experimentation with different time step sizes. However, there is a notable gap in the publication record on the development of adaptive time step controllers for multirate methods. In this paper, we extend the single-rate controller work of Gustafsson [ACM Trans. Math. Software, 20 (1994), pp. 496-517] to the multirate method setting. Specifically, we develop controllers based on polynomial approximations to the principal error functions for both the "fast" and "slow" time scales within multirate infinitesimal (MRI) methods. We additionally investigate a variety of approaches for estimating the errors arising from each time scale within MRI methods. We then numerically evaluate the proposed multirate controllers and error estimation strategies on a range of multirate test problems, comparing their performance against an estimated optimal performance. Through this work, we combine the most performant of these approaches to arrive at a set of multirate adaptive time step controllers that robustly achieve desired solution accuracy with minimal computational effort.

97 MATHEMATICS AND COMPUTING↗

Insights from Optimizing HPL Performance on Exascale Systems: A Comparative Analysis of Panel Factorization

High performance LINPACK (HPL) remains the primary benchmark for evaluating supercomputing performance. It includes many parts with substantial internal complexity, and its performance is affected by a large number of parameters that interact in ways that are difficult to predict on large-scale heterogeneous supercomputer systems. We present a comprehensive performance analysis of HPL on Frontier, the world’s first exascale supercomputer, which achieved HPL performance of 1.35 exaflops. Through empirical parameter tuning, detailed modeling, and comparative evaluation, we uncover critical performance insights, share lessons learned, and outline best practices for effective parameter tuning on exascale systems. We introduce and evaluate two novel PDFACT strategies: a dedicated-thread (DT) variant and a GPU-based variant (GPUPDFACT) implementation using HIP cooperative groups, demonstrating that GPU-based factorization outperforms conventional CPU-based PDFACT on Frontier’s architecture. Our findings establish key performance factors for HPL on exascale systems and offer valuable guidance for future high-performance computing and benchmarking efforts.

Lu, Hao [ORNL] (ORCID:000000018941870X)↗