Benchmarking the performance of quantum computing software for quantum circuit creation, manipulation and compilation
Not Available
SEARCH · Engineering Papers
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
Not Available
Computing the GW quasiparticle band structure and Bethe-Salpeter equation (BSE) absorption spectra for materials with spin-orbit coupling have commonly been done by treating GW corrections and spin-orbit coupling (SOC) as separate perturbations to density-functional theory. However, accurate treatment of materials with strong spin-orbit coupling (such as many topological materials of recent interest, and thermoelectrics) often requires a nonperturbative approach using spinor wave functions in the Kohn-Sham equation and GW/BSE. Such calculations have only recently become available, in particular for the BSE. Here, we have implemented this approach in the plane-wave pseudopotential GW/BSE code BerkeleyGW, which is highly parallelized and widely used in the electronic-structure community. We present reference results for quasiparticle band structures and optical absorption spectra of solids with different strengths of spin-orbit coupling, including Si, Ge, GaAs, GaSb, CdSe, Au, and Bi 2 Se 3 . The calculated quasiparticle band gaps of these systems are found to agree with experiment to within a few tens of meV. SOC splittings are found to be generally in better agreement with experiment, including quasiparticle corrections to band energies. The absorption spectrum of GaAs is not significantly impacted by the inclusion of spin-orbit coupling due to its relatively small value (0.2 eV) in the Λ direction, while the absorption spectrum of GaSb calculated with the spinor GW/BSE captures the large spin-orbit splitting of peaks in the spectrum. For the prototypical topological insulator Bi 2 Se 3 , we find a drastic change in the low-energy band structure compared to that of DFT, with the spinorial treatment of the GW approximation correctly capturing the parabolic nature of the valence and conduction bands after including off-diagonal self-energy matrix elements. We present the detailed methodology, approach to spatial symmetries for spinors, comparison against other codes, and performance compared to spinless GW/BSE calculations and perturbative approaches to SOC. This work aims to spur further development of spinor GW/BSE methodology in excited-state research software and enables a more accurate and detailed exploration of electronic and optical properties of materials containing elements with large atomic numbers.
Laser beam welding (LBW) is affected by the extreme temperatures, reduced pressure, and reduced gravity present in space environments. Gravity and pressure especially influence its melt pool and solidification dynamics. A compact, modular vacuum chamber adaptable to flight platforms from parabolic to orbital currently hosts an experiment to investigate the combined influence of reduced gravity and pressure on LBW. A swappable cartridge contains a rotating platen on which customizable workpieces can be welded under vacuum, greatly increasing experimental throughput. Instrumentation includes weld and thermal cameras observing the process, thermocouples placed on workpieces, accelerometers, and vacuum sensors. Experimental data gathered during the welding process will be combined with post-flight nondestructive evaluation, metallography, and mechanical testing to provide validation datasets for computational modeling. Phase I of this effort involves a parabolic flight campaign in low gravity while an anticipated Phase II would proceed to in-space demonstration to access extended duration microgravity.
An abstract system of benchmark characteristics that makes it possible, in the beginning of the design stage, to design with benchmark performance in mind is presented. The benchmark characteristics for a set of commonly used benchmarks are then shown. The benchmark set used includes some benchmarks from the Systems Performance Evaluation Cooperative (SPEC). The SPEC programs are industry-standard applications that use specific inputs. Processor, memory-system, and operating-system characteristics are addressed.
This presentation discusses the objective of the research which is to update n+ 63,65 Cu cross section evaluations with recently measured data to resolve discrepancies in benchmark performance. It also discusses the models used which are the R -matrix analysis and the analysis of angular distribution coefficients. Additionally, the validation methods that were used are discussed, including the Rez shielding benchmark and the ICSBEP criticality benchmarks. In conclusion, he n+ 63,65 Cu cross sections have been updated via R-matrix analysis up to 100 keV. An increased average capture cross section and the adoption of experimentally based Legendre coefficients lead to improved performance in reactivity benchmarks. In the fast region, the adoption of the JENDL-4.0 cross sections above 4.0 MeV improves the performance in shielding benchmarks. Ultimately, the n 63,65 Cu ENDF files will be submitted to ENDF/B-VIII.1.
The Commercial Aviation Safety Team found the majority of recent international commercial aviation accidents attributable to loss of control inflight involved flight crew loss of airplane state awareness (ASA), and distraction was involved in all of them. Research on attention-related human performance limiting states (AHPLS) such as channelized attention, diverted attention, startle/surprise, and confirmation bias, has been recommended in a Safety Enhancement (SE) entitled "Training for Attention Management." To accomplish the detection of such cognitive and psychophysiological states, a broad suite of sensors was implemented to simultaneously measure their physiological markers during a high fidelity flight simulation human subject study. Twenty-four pilot participants were asked to wear the sensors while they performed benchmark tasks and motion-based flight scenarios designed to induce AHPLS. Pattern classification was employed to predict the occurrence of AHPLS during flight simulation also designed to induce those states. Classifier training data were collected during performance of the benchmark tasks. Multimodal classification was performed, using pre-processed electroencephalography, galvanic skin response, electrocardiogram, and respiration signals as input features. A combination of one, some or all modalities were used. Extreme gradient boosting, random forest and two support vector machine classifiers were implemented. The best accuracy for each modality-classifier combination is reported. Results using a select set of features and using the full set of available features are presented. Further, results are presented for training one classifier with the combined features and for training multiple classifiers with features from each modality separately. Using the select set of features and combined training, multistate prediction accuracy averaged 0.64 +/- 0.14 across thirteen participants and was significantly higher than that for the separate training case. These results support the goal of demonstrating simultaneous real-time classification of multiple states using multiple sensing modalities in high fidelity flight simulators. This detection is intended to support and inform training methods under development to mitigate the loss of ASA and thus reduce accidents and incidents.
In this paper we present a new GPU-oriented mesh optimization method based on high order finite elements. Our approach relies on node movement with fixed topology, through the Target-Matrix Optimization Paradigm (TMOP) and uses a global nonlinear solve over the whole computational mesh, i.e., all mesh nodes are moved together. A key property of the method is that the mesh optimization process is recast in terms of finite element operations, which allows us to utilize recent advances in the field of GPU-accelerated high order finite element algorithms. For example, we reduce data motion by using tensor factorization and matrix-free methods, which have superior performance characteristics compared to traditional full finite element matrix assembly and offer advantages for GPU based HPC hardware. Furthermore, we describe the major mathematical components of the method along with their efficient GPU-oriented implementation. In addition, we propose an easily reproducible mesh optimization test that can serve as a performance benchmark for the mesh optimization community.
The Critical Unresolved Region Integral Experiment (CURIE) critical experiment was performed at the National Criticality Experiments Research Center (NCERC) at the Device Assembly Facility (DAF) at the Nevada Nuclear Security Sites (NNSS) in 2020. The objective of CURIE was to improve the quality of integral nuclear data in the uranium-235 ( 235 U) unresolved resonance region (URR) by performing benchmark integral experiments that were sensitive to the URR energy ranges. The CURIE experiment was evaluated for the International Criticality Safety Benchmark Evaluation Project (ICSBEP) handbook and the benchmark evaluation was accepted in 2022. The URR is a region within the intermediate neutron energy range (the intermediate energy ranges from 0.7 eV to 100 keV). The observed resonance structure in neutron cross sections is due to discrete energy levels in the nucleus and are characterized by resonance parameters. In the URR region, the resonance parameters are only partially resolved as the resolution of the experiment techniques becomes comparable to the average width of the resonances themselves, and the resonances are so close to one another that the structure cannot be determined empirically. The ENDF/B-VIII.0 and JEFF-3.3 nuclear data libraries define the URR as beginning at 2.25 keV and continuing until 25 keV. There are minimal intermediate neutron energy benchmarks available in the ICSBEP benchmark handbook, and aside from CURIE there are none that are highly sensitive in the URR energy region. There is a current need for additional experiments sensitive to the URR energy region. There is a new proposed subgroup for the Organization for Economic Co-operation and Development and Nuclear Energy Agency (OECD NEA) Working Party on International Nuclear Data Evaluation Cooperation (WPEC), so any new models or information will need experimental validation and testing. The NEA Working Party on Nuclear Criticality Safety (WPNCS) recent experimental needs and priority list includes intermediate energy 235 U and 238 U experiments, as does the recent Integral Experiments to Address Nuclear Criticality Safety Needs Meeting (May 2023). A variant on the original CURIE experiment called the Multiple Critical Unresolved Region Integral Experiment (MCURIE) is proposed to provide further investigation of the intermediate and URR energy region for uranium. MCURIE will utilize existing fuels at NCERC but use alternative moderators and reflectors to modify the neutron absorption, scattering, and fission spectra of the experiments, allowing for precise targeting of nuclear data sensitivities. The overall goal of MCURIE is to develop a framework for designing and performing integral benchmark experiments with high sensitivities in the intermediate and URR energy regions for nuclear data validation.
Many-body perturbation theory is a powerful method to simulate electronic excitations in molecules and materials starting from the output of density functional theory calculations. By implementing the theory efficiently so as to run at scale on the latest leadership high-performance computing systems it is possible to extend the scope of GW calculations. Here, we present a GPU acceleration study of the full-frequency GW method as implemented in the WEST code. Excellent performance is achieved through the use of (i) optimized GPU libraries, e.g., cuFFT and cuBLAS, (ii) a hierarchical parallelization strategy that minimizes CPU-CPU, CPU-GPU, and GPU-GPU data transfer operations, (iii) nonblocking MPI communications that overlap with GPU computations, and (iv) mixed precision in selected portions of the code. A series of performance benchmarks has been carried out on leadership high-performance computing systems, showing a substantial speedup of the GPU-accelerated version of WEST with respect to its CPU version. Good strong and weak scaling is demonstrated using up to 25 920 GPUs. Finally, we showcase the capability of the GPU version of WEST for large-scale, full-frequency GW calculations of realistic systems, e.g., a nanostructure, an interface, and a defect, comprising up to 10 368 valence electrons.
High-performance coatings for Concentrating Solar Power (CSP) receivers are subjected to remarkable environmental stressors during normal operations. Applied to the receiver tubes, these coatings serve to maximize the solar absorptivity of the receiver, transferring as much heat as possible from the solar collectors into the heat-transfer fluid (HTF). The lifecycle of these coatings is not well-defined, and the harsh operational conditions make them difficult to test. NREL has designed, built, and tested an apparatus to expose these samples to design levels of environmental stress and well beyond, into accelerated and destructive conditions. The chamber is actively cooled, monitored, and has the capability to supply humidification for cycling tests, allowing us to test multi-modal degradation and failure conditions at high temperature, high flux, and high humidity conditions. These conditions can catalyze high-temperature oxidation, mechanical degradation, and other modes of absorptivity loss seen in selective solar receiver coatings. The experimental data can feed lifecycle models for expensive and necessarily resilient materials, offering insights to aid maintenance schedules, technoeconomic analysis, and material industry performance benchmarks. This presentation will demonstrate the apparatus design and performance, as well as initial results for aging on a selective receiver coating.
Low temperature double-layer capacitor operation enabled by: - Base acetonitrile / TEATFB salt formulation - Addition of low melting point formates, esters and cyclic ethers center dot Key electrolyte design factors: - Volume of co-solvent - Concentration of salt center dot Capacity increased through higher capacity electrodes: - Zeolite templated carbons - Asymmetric cell designs center dot Continuing efforts - Improve asymmetric cell performance at low temperature - Cycle life testing Motivation center dot Benchmark performance of commercial cells center dot Approaches for designing low temperature systems - Symmetric cells (activated carbon electrodes) - Symmetric cells (zeolite templated carbon electrodes) - Asymmetric cells (lithium titanate/activated carbon electrodes) center dot Experimental results center dot Summary
High Performance Fortran (HPF), the high-level language for parallel Fortran programming, is based on Fortran 90. HALF was defined by an informal standards committee known as the High Performance Fortran Forum (HPFF) in 1993, and modeled on TMC's CM Fortran language. Several HPF features have since been incorporated into the draft ANSI/ISO Fortran 95, the next formal revision of the Fortran standard. HPF allows users to write a single parallel program that can execute on a serial machine, a shared-memory parallel machine, or a distributed-memory parallel machine. HPF eliminates the complex, error-prone task of explicitly specifying how, where, and when to pass messages between processors on distributed-memory machines, or when to synchronize processors on shared-memory machines. HPF is designed in a way that allows the programmer to code an application at a high level, and then selectively optimize portions of the code by dropping into message-passing or calling tuned library routines as 'extrinsics'. Compilers supporting High Performance Fortran features first appeared in late 1994 and early 1995 from Applied Parallel Research (APR) Digital Equipment Corporation, and The Portland Group (PGI). IBM introduced an HPF compiler for the IBM RS/6000 SP/2 in April of 1996. Over the past two years, these implementations have shown steady improvement in terms of both features and performance. The performance of various hardware/ programming model (HPF and MPI (message passing interface)) combinations will be compared, based on latest NAS (NASA Advanced Supercomputing) Parallel Benchmark (NPB) results, thus providing a cross-machine and cross-model comparison. Specifically, HPF based NPB results will be compared with MPI based NPB results to provide perspective on performance currently obtainable using HPF versus MPI or versus hand-tuned implementations such as those supplied by the hardware vendors. In addition we would also present NPB (Version 1.0) performance results for the following systems: DEC Alpha Server 8400 5/440, Fujitsu VPP Series (VX, VPP300, and VPP700), HP/Convex Exemplar SPP2000, IBM RS/6000 SP P2SC node (120 MHz) NEC SX-4/32, SGI/CRAY T3E, SGI Origin2000.
High-strength aluminum alloys for elevated temperature applications are desirable to replace heavier and more expensive titanium alloys. However, most aluminum alloys lose a large fraction of their strength at temperatures above approximately 200°C. ORNL has designed DuAlumin-3D, an alloy with nominal composition Al-9Ce-4Ni-0.5Mn-1Zr (wt.%), which utilizes the high cooling rates in additive manufacturing (AM) to achieve a refined microstructure, and thermally stable mechanical properties. DuAlumin-3D was fabricated by laser powder bed fusion and tested for its tensile mechanical properties across a range of temperature, and for its room temperature high-cycle fatigue resistance. The alloy was tested in both the as-printed and heat treated conditions, and both parallel and perpendicular to the AM build direction. The alloy was found to have anisotropic mechanical behavior in the as-printed state, but the anisotropy significantly decreased (both for tensile and fatigue properties) following heat treatment. The tensile properties significantly out-performed benchmark wrought 2219-T61 across a wide temperature range. The room temperature fatigue performance was approximately similar to 2219-T61.
In buildings, performance assessment often focuses on energy use with metrics such as energy use intensity (EUI) used to benchmark performance. However, energy performance of a building is fundamentally determined by the control system that engages the energy-using systems. There are two aspects of control that are of particular importance: (1) the ability to regulate process variables to their setpoints; and (2) whether the setpoints are at the right levels and/or following desired profiles. Most buildings do not reach their energy efficiency potential due to deficiencies in control performance and operators do not have access to metrics that can illuminate these deficiencies. Here this paper addresses this problem by providing novel techniques that combine these two aspects of control performance into a single standardized score on the scale of 0-10. The concept of a standardized control scores enables all systems in a building to be compared on the same scale and also for scores to be rolled up to different levels in the building and system hierarchy for system-wide analysis. The paper presents the theory for the method, describes a prototype tool for displaying scores, and presents results from application to a large building in Minneapolis.
As the research community studying proton exchange membrane water electrolysis (PEMWE) grows, it is important to develop methods that achieve transparent, reproducible research. Reproducibility of performance and durability is a challenge facing the PEMWE field, as the published literature includes a wide spread of results obtained with nominally similar materials. Prior round-robin performance benchmarking efforts [1] have identified inadequate cell conditioning as a major source of variation in apparent cell performance. Inadequate pre-treatment and conditioning can lead to instability in initial performance and adds ambiguity to durability measurements. However, excessively prolonged conditioning procedures limit the throughput of testing and the pace of research. Pretreatment and conditioning methods vary significantly across the research literature, but little systematic investigation is available into the mechanisms of these procedures or how procedural differences may impact results. This presentation will discuss investigations into the effects of membrane pre-treatment and operating procedures during cell conditioning on initial performance and catalyst-specific accelerated stress tests, with the aim of recommending procedures to enable clear, reproducible, and high-throughput research. Methods investigated include the use of hydrogen peroxide, acids, and hydration at elevated temperature for membrane pretreatment, and conditioning procedures such as current or voltage holds and cycling. The investigations cover both the impacts of these procedures on cell performance and stability as well as underlying mechanisms and processes taking place in the cell materials.
This report serves as an introduction, tutorial, and benchmark specification for out-of-pile tests on metallic fuel. It introduces a new user to the EBR-II legacy fuel performance test program and the fast reactor fuel performance databases built to preserve the records. It then details the information stored in each database and how to find it. A benchmark specification is included for a small set of out-of-pile tests on U-10Zr fuel to function as a tutorial demonstrating how the legacy fuel performance data sets stored in the FIPD and OPTD databases can be used together to benchmark fuel performance models for steady-state and transient performance.
Data used by the publication "High-throughput calculations of charged point defect properties with semi-local density functional theory - performance benchmarks for materials screening applications." This work presented an in-depth benchmark analysis of automated, semi-local point defect calculations with a-posteriori corrections, compared to 245 “gold standard” hybrid calculations previously published. We considered three different a-posteriori correction sets for semi-local calculations, implemented in a fully automated workflow, and consider the qualitative and quantitative differences for four different categories of defect information: thermodynamic transition levels, formation energies, fermi levels, and dopability limits. We highlighted the type of qualitative information about point defect properties that can be extracted from high-throughput calculations based on semi-local DFT methods, while also demonstrating the limits of quantitative accuracy that can be achieved by these approaches.
Project Voyager radio metric data are used to evaluate the orbit determination abilities of several data strategies during spacecraft interplanetary cruise. Benchmark performance is established with an operational data strategy of conventional coherent doppler, coherent range, and explicitly differenced range data from two intercontinental baselines to ameliorate the low declination singularity of the doppler data. Employing a Voyager operations trajectory as a reference, the performance of the operational data strategy is compared to the performances of data strategies using differential VLBI delay data (spacecraft delay minus quasar delay) in combinations with the aforementioned conventional data types. The comparison of strategy performances indicates that high accuracy cruise orbit determination can be achieved with a data strategy employing differential VLBI delay data, where the quantity of coherent radio metric data has been greatly reduced.