Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “PERFORMANCE”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,081 records · Page 60

Crossover as Determinant for Safety and Performance Tradeoffs in Proton Exchange Membrane Water Electrolyzers

Hydrogen (H2) crossover is a pressing challenge constraining safe and efficient operation of proton exchange membrane water electrolyzers (PEMWEs) especially amongst strides to employ thinner membranes, which enables improved energy efficiency, and elevated cathode pressures, that reduces the energy burden on downstream compressors. Here, we develop a microstructure-aware multicomponent reactive-transport framework that resolves dissolved and gaseous H2 transport pathways and mechanistically links electrode architecture to crossover related safety and performance. We show that operability is co-governed by the cathode catalyst layer (CCL) and the anode porous transport layer (APTL) which sets the H2 crossover flux and the egress capacity respectively. Elevated Pt/C ratio in the CCL suppresses crossover flux by up to 23% while a higher APTL porosity lowers H2 in O2 fraction by 0.6% in the anode effluent. We condense the findings into (cathode pressure-current density) maps overlaid with safety limits and performance targets and ultimately define two safety-performance unified metrics to gauge the size and quality of the operating window. Given the push towards higher pressure and deeper turndown for renewable integration, this study provides mechanistic design guidance to prevent crossover-induced safety risks while preserving the desired performance.

Electrolysis↗

Exploring the effects of molecular beam epitaxy growth characteristics on the temperature performance of state-of-the-art terahertz quantum cascade lasers

This study conducts a comparative analysis, using non-equilibrium Green’s functions (NEGF), of two state-of-the-art two-well (TW) Terahertz Quantum Cascade Lasers (THz QCLs) supporting clean 3-level systems. The devices have nearly identical parameters and the NEGF calculations with an abrupt-interface roughness height of 0.12 nm predict a maximum operating temperature (T max ) of ~ 250 K for both devices. However, experimentally, one device reaches a T max of ~ 250 K and the other a T max of only ~ 134 K. Both devices were fabricated and measured under identical conditions in the same laboratory, with high quality processes as verified by reference devices. The main difference between the two devices is that they were grown in different MBE reactors. Our NEGF-based analysis considered all parameters related to MBE growth, including the maximum estimated variation in aluminum content, growth rate, doping density, background doping, and abrupt-interface roughness height. From our NEGF calculations it is evident that the sole parameter to which a drastic drop in T max could be attributed is the abrupt-interface roughness height. We can also learn from the simulations that both devices exhibit high-quality interfaces, with one having an abrupt-interface roughness height of approximately an atomic layer and the other approximately a monolayer. However, these small differences in interface sharpness are the cause of the large performance discrepancy. This underscores the sensitivity of device performance to interface roughness and emphasizes its strategic role in achieving higher operating temperatures for THz QCLs. We suggest Atom Probe Tomography (APT) as a path to analyze and measure the (graded)-interfaces roughness (IFR) parameters for THz QCLs, and subsequently as a design tool for higher performance THz QCLs, as was done for mid-IR QCLs. Our study not only addresses challenges faced by other groups in reproducing the record T max of ~ 250 K and ~ 261 K but also proposes a systematic pathway for further improving the temperature performance of THz QCLs beyond the state-of-the-art.

77 NANOSCIENCE AND NANOTECHNOLOGY↗

Core performance predictions in projected SPARC first-campaign plasmas with nonlinear CGYRO

This work characterizes the core transport physics of SPARC early-campaign plasmas using the PORTALS-CGYRO framework. Empirical modeling of SPARC plasmas with L-mode confinement indicates an ample window of breakeven (Q > 1) without the need of H-mode operation. Extensive modeling of multi-channel (electron energy, ion energy, and electron particle) flux-matched conditions with the nonlinear CGYRO code for turbulent transport coupled to the macroscopic plasma evolution using PORTALS reveals that the maximum fusion performance to be attained will be highly dependent on the near-edge pressure. Stiff core transport conditions are found, particularly when fusion gain approaches unity, and predicted density peaking is found to be in line with empirical databases of particle source-free H-modes. Impurity optimization is identified as a potential avenue to increase fusion performance while enabling core-edge integration. Extensive validation of the quasilinear TGLF model builds confidence in reduced-model predictions. The implications of projecting L-mode performance to high-performance and burning-plasma devices is discussed, together with the importance of predicting edge conditions.

Rodriguez-Fernandez, P. (ORCID:0000000273611131)↗

Magnetized liner inertial fusion platform development to assess performance scaling with drive parameters

Magnetized liner inertial fusion (MagLIF) experiments have demonstrated fusion-relevant ion temperatures up to 3.1 keV and thermonuclear production of up to 1.1 × 1013 deuterium–deuterium neutrons. This performance was enabled through platform development that provided increases in applied magnetic field, coupled preheat energy, and drive current. Advanced coil designs with internal reinforcement enabled an increase from 10 to 20 T. An improved laser pulse shape, beam smoothing, and thinner laser entrance foils increased preheat energy coupling from less than 1 to 2.3 kJ. A redesign of the final transmission line and load region increased peak load current from 16 to 20 MA. The wider range of input parameters was leveraged to study target performance trends with preheat energy, applied magnetic field, and peak load current. Ion temperature and neutron yield generally followed trends in two-dimensional clean Lasnex calculations. Stagnation performance improved with peak load current when other input parameters were also increased such that convergence was maintained. This dataset suggests that reducing convergence to less than 30 would improve predictability of target performance. Lasnex was used to identify a simulation-optimized scaling path, which suggests 10+ kJ of fusion yield is possible on the Z facility with achievable input parameters. This path also indicates >10 MJ could be generated through volume burn on a future facility with a path to high yield (>200 MJ) using cryogenic dense fuel layers. The newly developed MagLIF platform enables exploration of both this simulation optimized scaling path and a recently developed similarity-scaling path.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Enhanced plasma performance in C-2W advanced beam-driven field-reversed configuration experiments

TAE Technologies’ fifth-generation fusion device, C-2W (also called ‘Norman’), is the world’s largest compact-toroid device and has made significant progress in field-reversed configuration (FRC) plasma performance. C-2W produces record breaking, macroscopically stable, high-temperature advanced beam-driven FRC plasmas, dominated by injected fast particles and sustained in steady state, which is primarily limited by neutral-beam (NB) pulse duration. The NB power supply system has recently been upgraded to extend the pulse length from 30 ms to 40 ms, which allows for a longer plasma lifetime and thus better characterization and further enhancement of FRC performance. An active plasma control system is routinely used in C-2W to produce consistent FRC performance as well as for reliable machine operations using magnet coils, edge-biasing electrodes, gas injection and tunable-energy NBs. Google’s machine learning framework for experimental optimization has also been routinely used to enhance plasma performance. D

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Performance prediction applying different reduced turbulence models to the SMART tokamak

The SMall Aspect Ratio Tokamak (SMART) is currently being commissioned at the University of Seville and will be able to compare the performance of positive and negative triangularity plasmas at low aspect ratio. Predictive simulations have been performed for different machine scenarios and heating schemes using the TRANSP code. The objectives of these simulations are to predict the parameters expected in positive triangularity plasmas, to guide diagnostic development, and to validate transport models. Several reduced turbulence models have been used to predict electron and ion temperatures for the operational phase 2. All models provide similar results from approximately mid-radius to the separatrix but important discrepancies are found in the core region. These positive triangularity results are compared with experiments from a similar size machine like GLOBUS-M2. The multi-mode model (MMM) shows the best agreement. Simulations with different boundary conditions have been performed and no strong differences have been observed between them. The impact of neutral beam injection (NBI) on the predicted profiles has also been addressed. Rotation reduces turbulence levels so higher temperatures are achieved when included in the simulations. Studying the different contributions to the thermal diffusivities, it is observed that electron temperature gradient (ETG) turbulence dominates at the plasma core while micro-tearing modes (MTM) dominate at the edge in the electron channel. In the ion channel, the neoclassical contribution is dominant at the core and at the very edge while the Weiland component, which includes ion temperature gradient mode (ITG), trapped electron mode (TEM), kinetic ballooning mode (KBM), peeling mode (PM) and collisionless and collision dominated magnetohydrodynamic (MHD) modes governs the mid-radius region. For phase 3, two plasmas with different electron densities have been studied. The case with lower density matches well a specific discharge of GLOBUS-M2. The higher density plasma shows high performance with β N ≈ 3.8.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Enhancing electrocatalytic performance of RuO 2 -based catalysts: mechanistic insights, strategic approaches, and recent advances

Abstract Electrochemical water splitting presents the ultimate potential of hydrogen and oxygen production; however, regulating the rate and efficiency of water splitting is highly dependent on the accessibility of extremely efficient electrode materials for slow performance kinetics and large overpotential of both oxygen evolution reaction (OER) and hydrogen evolution reaction (HER). Ruthenium oxide (RuO 2 ) based materials display high performance for OER and HER because of their capacity to bind oxygen, eminent catalytic activity, low cost compared to other precious metals, and stability in a wide pH range. However, there is still much space to promote the OER and HER activity and stability of RuO 2 to fulfill the necessity for practical applications in water splitting. Different researchers applied multiple approaches that boosted the catalytic performance of RuO 2 -based electrocatalysts toward overall water splitting. Herein, this review provides a comprehensive overview of recent advancements in RuO 2 -based materials in the field of water electrolysis for the generation of alternative energies. It gives a general description of water splitting in acidic and alkaline settings, including reaction mechanisms as well as common evaluation elements for the catalytic function of the materials. Most of the reviews reported based on RuO 2 materials are only focused on OER performance, but this review highlighted comprehensive ideas on different strategies like morphology design, electronic structure, electrolytes, and compositions for optimizing both electrocatalytic HER and OER functioning of RuO 2 -based electrocatalysts.

KC, Binod Raj (ORCID:0009000885806906)↗

SYCL for Performance Portability: Application Experience with Coupled Cluster Formalism in Quantum Chemistry on Exascale Systems

The exascale computing has brought unprecedented heterogeneity in node architectures, with systems such as Frontier and Aurora featuring diverse GPU accelerators, network connectivity among others. Ensuring performance portability across these platforms is a key challenge. To address this, we employ the SYCL programming model to develop portable, high-performance quantum chemistry workloads. As a representative application, we focus on the non-iterative Triples component of the coupled-cluster CCSD(T) method, a key driver in quantum chemistry. In this work, we report on our experience deploying SYCL-based implementations using both DPC++ and AdaptiveCPP across two flagship exascale platforms: OLCF Frontier with AMD MI250X GPUs and ALCF Aurora with Intel GPUs. Our results demonstrate that SYCL enables efficient, single-source implementations that scale to thousands of nodes, delivering performance on par with vendor-optimized HIP solutions. We highlight key insights into runtime behavior, kernel portability, and scaling characteristics, showing that SYCL offers a viable path for performance-portable computing.

Bagusetty, Abhishek [Argonne National Laboratory (↗

Visual Analytics of Performance of Quantum Computing Systems and Circuit Optimization

Driven by potential exponential speedups in business, security, and scientific scenarios, interest in quantum computing is surging. This interest feeds the development of quantum computing hardware, but several challenges arise in optimizing application performance for hardware metrics (e.g., qubit coherence and gate fidelity). In this work, we describe a visual analytics approach for analyzing the performance properties of quantum devices and quantum circuit optimization. Our approach allows users to explore spatial and temporal patterns in quantum device performance data and it computes similarities and variances in key performance metrics. Detailed analysis of the error properties characterizing individual qubits is also supported. We also describe a method for visualizing the optimization of quantum circuits. The resulting visualization tool allows researchers to design more efficient quantum algorithms and applications by increasing the interpretability of quantum computations.

Chae, Junghoon↗

A Performance-Portable MultiGPU Implementation of 3D Euler Equations using ProtoX and IRIS

Computational scientists often face challenges when developing and optimizing code for high-performance computing (HPC), especially when trying to leverage GPUs. Given the heterogeneity of the nodes that comprise many modern HPC facilities, considerable demand exists for performance portable solutions for the core computational kernels used in many scientific computing libraries. In this work, we demonstrate a fourth-order finite volume method–based implementation of the Euler equations, which are an integral part of computational fluid dynamics. Our performance-portable multiGPU implementation for Euler equations uses ProtoX to generate kernels and IRIS for portability. ProtoX is a domain-specific language that uses a structured-grid partial differential equation library called Proto as its front end and the SPIRAL code generation system as its back end to generate optimized kernels for different architectures. Optimized kernels generated by ProtoX are orchestrated through the IRIS intelligent runtime system to provide portability. Two levels of optimizations within the IRIS runtime— directed acyclic graph fusion and task fusion—are explored to efficiently utilize computing resources in a multiGPU environment. Performance improvement through these optimizations is showcased by comparing the base ProtoX-IRIS implementation on AMD GPUs (Frontier node) and on NVIDIA GPUs (NVIDIA DGX-1).

Mankad, Het↗

Integrating ORNL’s HPC and Neutron Facilities with a Performance-Portable CPU/GPU Ecosystem

We explore the development of a performance-portable CPU/GPU ecosystem to integrate two of the US Department of Energy’s (DOE’s) largest scientific instruments, the Oak Ridge Leadership Computing facility and the Spallation Neutron Source (SNS), both of which are housed at Oak Ridge National Laboratory. We select a relevant data reduction workflow use-case to obtain the differential scattering cross-section from data collected by SNS’s CORELLI and TOPAZ instruments. We compare the current CPU-only production implementation using the Garnet Python multiprocess package based on the Mantid C++ framework against our proposed CPU/GPU implementation that uses the LLVM-based, just-in-time Julia scientific language and the JACC.jl performance-portable package. Two proxy apps were developed: (i) an app for extracting relevant Mantid kernels (MDNorm) in C++ and (ii) the Julia MiniVATES.jl miniapp. We present performance results for NVIDIA A100 and AMD MI100 GPUs and AMD EPYC 7513 and 7662 CPUs. The results provide insights for future generations of data reduction software that can embrace performance portability for an integrated research infrastructure across DOE’s experimental and computational facilities.

Hahn, Steven↗

Characterizing and Modeling the Influence of Geometry on the Performance of Superconducting Nanowire Cryotrons

The scaling of superconducting nanowire detectors to larger arrays is often limited by room-temperature-readout cabling. Cryogenic integrated circuits constructed from nanowire cryotrons, or nanocryotrons, can address this limitation by performing signal processing on chip. In this study, we characterize key performance metrics of the nanocryotron to elucidate its potential as a logical element in cryogenic integrated circuits and develop an electro-thermal model to connect material parameters with device performance. We find that the performance of the nanocryotron depends on the device geometry, and trade-offs are associated with optimizing the gain, jitter, and energy dissipation. Here, we demonstrate that nanocryotrons fabricated on niobium nitride can achieve a grey zone less than 210 nA wide for a 5 ns long input pulse corresponding to a maximum achievable gain of 48 dB, an energy dissipation of less than 20 aJ per operation, and a jitter of less than 60 ps.

Superconductor↗

Preliminary Study on Fine-Grained Power and Energy Measurements on Grace Hopper GH200 with Open-Source Performance Tools

The increasing adoption of tightly integrated, heterogeneous architectures, combined with the slowdown of Moore’s law, has made application power and energy-driven optimizations critical to efficiently use high-performance computing systems. This paper introduces a newly developed open-source toolkit that seamlessly integrates the Linux real-time hardware monitoring program hwmon with the Performance Application Programming Interface and the Score-P performance measurement system, thereby enabling fine-grained power and energy measurements for high-performance computing applications. Our primary target platform is the Wombat test bed, which is a system based on the NVIDIA GH200 superchip. The toolkit can capture transient power peaks with high temporal resolution (50 ms) and, thanks to Score-P integration, can map power metrics to specific code regions, thereby providing actionable information on power-intensive operations and inefficiencies. The toolkit also provides a holistic view of both the power and the energy consumption of the entire GH200 superchip by covering all major components: the Grace CPU, the Hopper GPU, and the I/O subsystem. Experiments that use Locally Self-consistent Multiple Scattering, which is an application for first-principles calculations of materials developed at Oak Ridge National Laboratory, have demonstrated the tool’s ability to identify transient power spikes and uncover opportunities for energy-aware optimizations. Additionally, we introduce a Python-based utility for converting Open Trace Format 2 traces to Parquet format, thus enabling advanced data analysis for numerical integration methods applied to power data for accurate energy profiling.

Hernandez Mendoza, Oscar [ORNL] (ORCID:00000002538↗

Mojo: MLIR-based Performance-Portable HPC Science Kernels on GPUs for the Python Ecosystem

We explore the performance and portability of the novel Mojo language for scientific computing workloads on GPUs. As the first language based on the LLVM’s Multi-Level Intermediate Representation (MLIR) compiler infrastructure, Mojo aims to close performance and productivity gaps by combining Python’s interoperability and CUDA-like syntax for compile-time portable GPU programming. We target four scientific workloads: a seven-point stencil (memory-bound), BabelStream (memory-bound), miniBUDE (compute-bound), and Hartree–Fock (compute-bound with atomic operations); and compare their performance against vendor baselines on NVIDIA H100 and AMD MI300A GPUs. We show that Mojo’s performance is competitive with CUDA and HIP for memory-bound kernels, whereas gaps exist on AMD GPUs for atomic operations and for fast-math compute-bound kernels on both AMD and NVIDIA GPUs. Although the learning curve and programming requirements are still fairly low-level, Mojo can close significant gaps in the fragmented Python ecosystem in the convergence of scientific computing and AI.

Godoy, William [ORNL] (ORCID:0000000225905178)↗

Spectral Efficiency and Tandem Performance Calculators [SWR-24-92]

The Spectral Efficiency and Tandem Performance Calculators allow users to test material types for use in tandem photovoltaic devices to better understand if a given material could work well. The Spectral Efficiency (SE) calculator takes a set of current vs voltage (I-V) and quantum efficiency (QE) data along with a spectrum input and plots the spectral efficiency of the cells represented by the datasets for the spectrum provided. The Tandem Performance calculator uses the calculated spectral efficiency of top and bottom cells and the transmission data of the top cell to calculate performance metrics for a tandem device in either a two or four terminal architecture. The SE and Tandem Performance calculators are based on the following publication: Yu, Z., Leilaeioun, M. & Holman, Z. Selecting tandem partners for silicon solar cells. Nat Energy 1, 16137 (2016).

Warren, Emily↗

MUPPET: An automated OpenMP mutation testing framework for performance optimization

MUPPET is a tool for OpenMP programs that identifies program modifications, called mutations, aimed at improving program performance. Existing performance optimization techniques, including profiling-based and auto-tuning techniques, fail to indicate program modifications at the source level thus preventing their portability across compilers. MUPPET aims to help HPC developers reason about performance defects and missed opportunities to improve performance at the source code level.

Parasyris, Konstantinos↗

Exogenous erythropoietin increases hematological status, fat oxidation, and aerobic performance in males following prolonged strenuous training

Abstract This study investigated the effects of EPO on hemoglobin (Hgb) and hematocrit (Hct), time trial (TT) performance, substrate oxidation, and skeletal muscle phenotype throughout 28 days of strenuous exercise. Eight males completed this longitudinal controlled exercise and feeding study using EPO (50 IU/kg body mass) 3×/week for 28 days. Hgb, Hct, and TT performance were assessed PRE and on Days 7, 14, 21, and 27 of EPO. Rested/fasted muscle obtained PRE and POST EPO were analyzed for gene expression, protein signaling, fiber type, and capillarization. Substrate oxidation and glucose turnover were assessed during 90‐min of treadmill load carriage (LC; 30% body mass; 55 ± 5% V̇O 2 peak) exercise using indirect calorimetry, and 6‐6‐[ 2 H 2 ]‐glucose PRE and POST. Hgb and Hct increased, and TT performance improved on Days 21 and 27 compared to PRE ( p < 0.05). Energy expenditure, fat oxidation, and metabolic clearance rate during LC increased ( p < 0.05) from PRE to POST. Myofiber type, protein markers of mitochondrial biogenesis, and capillarization were unchanged PRE to POST. Transcriptional regulation of mitochondrial activity and fat metabolism increased from PRE to POST ( p < 0.05). These data indicate EPO administration during 28 days of strenuous exercise can enhance aerobic performance through improved oxygen carrying capacity, whole‐body and skeletal muscle fat metabolism.

Drummer, Devin J.↗

Performance of Quantum Dot Coatings for Luminescent Solar Concentrating Windows: Cooperative Research and Development Final Report, CRADA Number CRD-16-00640

The two major outcomes from this DOE supported collaboration between UbiQD and NREL were: 1) an expert analysis/modeling of the expected performance and cost of luminescent solar concentrating windows with quantum dot coatings, and 2) critical R&D characterization of materials and device performance analysis using equipment and processes that are too expensive for UbiQD to perform in-house. The latter also included third-party validation of the device performance with an NREL-certified conversion efficiency that was ultimately published. (see ACS Energy Lett. 2018 and recently ACS Appl. Energy Mater. 2020).

14 SOLAR ENERGY↗