Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Performance benchmark”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

The FTIO Benchmark

We introduce a new benchmark for measuring the performance of parallel input/ouput. This benchmark has flexible initialization. size. and scaling properties that allows it to satisfy seven criteria for practical parallel I/O benchmarks. We obtained performance results while running on the a SGI Origin2OOO computer with various numbers of processors: with 4 processors. the performance was 68.9 Mflop/s with 0.52 of the time spent on I/O, with 8 processors the performance was 139.3 Mflop/s with 0.50 of the time spent on I/O, with 16 processors the performance was 173.6 Mflop/s with 0.43 of the time spent on I/O. and with 32 processors the performance was 259.1 Mflop/s with 0.47 of the time spent on I/O.

Fagerstrom, Frederick C.↗

Evaluating the ProtoDUNE-SP Detector Performance to Measure a 6 GeV/c Positive Kaon Inelastic Cross Section on Argon

The ProtoDUNE Single-Phase Liquid Argon Time Projection Chamber \\ (ProtoDUNE-SP LArTPC) is a prototype for the Deep Underground Neutrino Experiment (DUNE), a future long-baseline neutrino oscillation experiment. Based at the CERN Neutrino Platform, ProtoDUNE-SP LArTPC collected data from a charged test beam in the fall of 2018. It then took data of cosmic-ray muons from November 2018 to the summer of 2020. The main goals of the prototype were to measure parameters related to charged particle passage in argon and evaluate the performance of the detector to inform future DUNE Far Detector development. The test beam provided kaons, pions, muons, protons, and electrons to the detector. These particles represent common final state particles in neutrino interactions, therefore providing information to DUNE on modeling charged particles in argon for its neutrino physics program. In addition to neutrino physics, DUNE has proposed an analysis using the DUNE Far Detector module to set limits for proton decay through the decay channel $p\rightarrow K^++\bar{\nu}$. This measurement would require information on kaons in argon, providing ProtoDUNE-SP LArTPC another opportunity to aid DUNE. This thesis describes the calibration of ProtoDUNE-SP and its detector performance, which serves as a benchmark for the performance of the DUNE Far Detector modules. A specific calibration highlighted is the evaluation of the liquid argon purity in the detector. These measurements use cosmic-ray muons reconstructed in the detector that are calibrated and matched to data from scintillator strips external to the TPC, known as the Cosmic Ray Tagger (CRT). The thesis will discuss the algorithms to match the data between the ProtoDUNE-SP LArTPC and the CRT and discuss the liquid argon purity measurements using one of the algorithms. Data sets of calibrated tracks measured the liquid argon contamination as consistently below 100 ppt oxygen equivalent. After these discussions on the ProtoDUNE-SP LArTPC detector, the thesis will present an evaluation of the inclusive inelastic, sometimes referred to as a reactive, cross section on argon of kaons from the ProtoDUNE-SP test beam with an average momentum of 6 GeV/c.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Development of a Pulsed Slowing-Down-Time Benchmark of Neutron Thermalization in Graphite

Graphite is a classic neutronic material that has been used as both a reactor reflector and moderator in various nuclear reactor systems. The ability to accurately predict the slowing down and thermalization of neutrons in graphite can have significant implications on the safety and operation of such reactor systems. In reactors, the neutron thermalization process is quantified using the thermal scattering law (TSL) and related cross sections for a given moderator. An ideal approach to assess the validity of TSL data is using benchmark measurement based on the pulsed Slowing-Down-Time technique and its comparison with the appropriate graphite nuclear library. In this work, experimental measurement and computational Monte Carlo simulations were performed to benchmark the slowing down characteristics and thermalization of neutrons in nuclear (reactor-grade) graphite. Given the density of graphite, various graphite libraries were selected for the benchmark analysis.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Evaluation of Opportunistic Contact Graph Routing in Random Mobility Environments

Routing in networks where nodes move randomly is particularly challenging due their potentially unpredictable, and rapidly changing topology. Several routing algorithms have been presented in the literature to address the needs of such networks, most of them implementing variants of controlled network flooding in the hope of successful data delivery. In this note, we compare the results of previous routing algorithms with Opportunistic Contact Graph Routing (OCGR), an enhanced version of Contact Graph Routing (CGR) that is suitable for networks where contacts cannot always be scheduled ahead of time. To perform the benchmark, we simulate a network of nodes moving in a certain space according to the Random Waypoint Mobility Model, and then take measurements of bundle delivery probabilty and overhead ratio as metrics of performance and cost respectively. Through this exercise, we demonstrate that the performance of OCGR is highly dependent on the type of network under consideration (e.g. very sparse vs. densely connected) and the assumed mobility model.

Burleigh, Scott↗

Rapid Evaluation Framework for the CMIP7 Assessment Fast Track

As Earth system models (ESMs) grow in complexity and in volume of output data, there is an increasing need for rapid, comprehensive evaluation of their scientific performance. The upcoming Assessment Fast Track for the Seventh Phase of the Coupled Model Intercomparison Project (CMIP7) will require expeditious response for model analyses designed to inform and drive integrated Earth system assessments. To meet this challenge, the Rapid Evaluation Framework (REF), a community-driven platform for benchmarking and performance assessment of ESMs, was designed and developed. The initial implementation of the REF, constructed to meet the near-term needs of the CMIP7 Assessment Fast Track, builds upon four disparate community evaluation and benchmarking tools that are coupled together using the Coordinated Model Evaluation Capabilities (CMEC) framework. The REF runs within a containerized workflow for portability and reproducibility and is aimed at generating and organizing diagnostics covering a variety of model variables. The REF leverages well documented observational datasets to provide assessments of model fidelity across a collection of diagnostics. All diagnostics were identified and selected with community involvement and consultation. Operational integration with the Earth System Grid Federation (ESGF) will permit automated execution of the REF for selected diagnostics as soon as model output data are published on ESGF by the originating modeling centers. The REF is designed to be portable across a range of current computational platforms to facilitate use by modeling centers for assessing the evolution of model versions or gauging the relative performance of CMIP simulations before being published on ESGF. When integrated into production simulation workflows, results from the REF provide immediate quantitative feedback that allows model developers and scientists to quickly identify model biases and performance issues. After the REF is released to the community, its subsequent development and support will be prioritized by an international consortium of scientists and engineers, enabling a broader impact across Earth science disciplines. For instance, the REF will facilitate improvements to models and will enhance confidence in model projections through process-based selection of models based on their performance with respect to observations. Production of reproducible diagnostics and community-based assessments are key features of the REF. Furthermore, providing interoperability with existing evaluation packages assures that contributions from previous community efforts will be available for use in future model intercomparison projects.

Hoffman, Forrest [ORNL] (ORCID:0000000158024134)↗

Congestion Avoidance Testbed Experiments

DARTnet provides an excellent environment for executing networking experiments. Since the network is private and spans the continental United States, it gives researchers a great opportunity to test network behavior under controlled conditions. However, this opportunity is not available very often, and therefore a support environment for such testing is lacking. To help remedy this situation, part of SRI's effort in this project was devoted to advancing the state of the art in the techniques used for benchmarking network performance. The second objective of SRI's effort in this project was to advance networking technology in the area of traffic control, and to test our ideas on DARTnet, using the tools we developed to improve benchmarking networks. Networks are becoming more common and are being used by more and more people. The applications, such as multimedia conferencing and distributed simulations, are also placing greater demand on the resources the networks provide. Hence, new mechanisms for traffic control must be created to enable their networks to serve the needs of their users. SRI's objective, therefore, was to investigate a new queueing and scheduling approach that will help to meet the needs of a large, diverse user population in a "fair" way.

Denny, Barbara A.↗

Evaluating Operators in Deep Neural Networks for Improving Performance Portability of SYCL

SYCL is a portable programming model for heterogeneous computing, so it is important to obtain reasonable performance portability of SYCL. Towards the goal of better understanding and improving performance portability of SYCL for machine learning workloads, we have been developing benchmarks for basic operators in deep neural networks (DNNs). These operators could be offloaded to heterogeneous computing devices such as graphics processing units (GPUs) to speed up computation. In this work, we introduce the benchmarks, evaluate the performance of the operators on GPU-based systems, and describe the causes of the performance gap between the SYCL and Compute Unified Device Architecture (CUDA) kernels. We find that the causes are related to the utilization of the texture cache for read-only data, optimization of the memory accesses with strength reduction, shared local memory accesses, and register usage per thread. We hope that the efforts of developing benchmarks for studying performance portability will stimulate discussion and interactions within the community.

97 MATHEMATICS AND COMPUTING↗

TEAMER Technical Support of Ramboll's for Numerical Modeling of WECs to Support OES Task 10: Cooperative Research and Development (Final Report)

National Technology & Engineering Solutions of Sandia, LLC (NTESS) in collaboration with the National Renewable Energy Laboratory (NREL), and with guidance from Ramboll, will perform fluid dynamics simulations to support of The Ocean Energy Systems (OES) Energy Technology Collaboration Program Task 10 Wave Energy Converters (WEC) Modelling Verification and Validation effort. Specific numerical simulations include the fixed device wave impingement studies to benchmark the performance of simulation techniques against physical testing results.

16 TIDAL AND WAVE POWER↗

Benchmarking of Solar Irradiance Nowcast Performance Derived from All-Sky Imagers

Fluctuations of the incoming solar irradiance impact the power generation from photovoltaic and concentrating solar thermal power plants. Accurate solar nowcasting becomes necessary to detect these sudden changes of generated power and to provide the desired information for optimal exploitation of solar systems. In the framework of the International Energy Agency's Photovoltaic Power Systems Program Task 16, a benchmarking exercise has been conducted relying on a bouquet of solar nowcasting methodologies by all-sky imagers (ASI). In this paper, four ASI systems nowcast the Global Horizontal Irradiance (GHI) with a time forecast ranging from 1 to 20 min during 28 days with variable cloud conditions spanning from September to November 2019 in southern Spain. All ASIs demonstrated their ability to accurately nowcast GHI, with RMSE ranging from 6.9% to 18.1%. Under cloudy conditions, all ASIs' nowcasts outperform the persistence models. Under clear skies, three ASIs are better than persistence. Discrepancies in the observed nowcasting performance become larger at increasing forecast horizons. The findings of this study highlight the feasibility of ASIs to reliably nowcast GHI at different sky conditions, time intervals and horizons. Such nowcasts can be used either to estimate solar power at distant times or detect sudden GHI fluctuations.

all-sky imagers↗

Benchmarking highly entangled states on a 60-atom analogue quantum simulator

Abstract Quantum systems have entered a competitive regime in which classical computers must make approximations to represent highly entangled quantum states 1,2 . However, in this beyond-classically-exact regime, fidelity comparisons between quantum and classical systems have so far been limited to digital quantum devices 2–5 , and it remains unsolved how to estimate the actual entanglement content of experiments 6 . Here, we perform fidelity benchmarking and mixed-state entanglement estimation with a 60-atom analogue Rydberg quantum simulator, reaching a high-entanglement entropy regime in which exact classical simulation becomes impractical. Our benchmarking protocol involves extrapolation from comparisons against an approximate classical algorithm, introduced here, with varying entanglement limits. We then develop and demonstrate an estimator of the experimental mixed-state entanglement 6 , finding our experiment is competitive with state-of-the-art digital quantum devices performing random circuit evolution 2–5 . Finally, we compare the experimental fidelity against that achieved by various approximate classical algorithms, and find that only the algorithm we introduce is able to keep pace with the experiment on the classical hardware we use. Our results enable a new model for evaluating the ability of both analogue and digital quantum devices to generate entanglement in the beyond-classically-exact regime, and highlight the evolving divide between quantum and classical systems.

Science & Technology - Other Topics↗

An Analysis of NASA Technology Transfer

A review of previous technology transfer metrics, recommendations, and measurements is presented within the paper. A quantitative and qualitative analysis of NASA's technology transfer efforts is performed. As a relative indicator, NASA's intellectual property performance is benchmarked against a database of over 100 universities. Successful technology transfer (commercial sales, production savings, etc.) cases were tracked backwards through their history to identify the key critical elements that lead to success. Results of this research indicate that although NASA's performance is not measured well by quantitative values (intellectual property stream data), it has a net positive impact on the private sector economy. Policy recommendations are made regarding technology transfer within the context of the documented technology transfer policies since the framing of the Constitution. In the second thrust of this study, researchers at NASA Langley Research Center were surveyed to determine their awareness of, attitude toward, and perception about technology transfer. Results indicate that although researchers believe technology transfer to be a mission of the Agency, they should not be held accountable or responsible for its performance. In addition, the researchers are not well educated about the mechanisms to perform, or policies regarding, technology transfer.

Bush, Lance B.↗

PDF4LHC21: Update on the benchmarking of the CT, MSHT and NNPDF global PDF fits

There have been recent updates to the three global PDF fits (CT, MSHT and NNPDF), all adding large amounts of data from the LHC, and this has resulted in significant changes to the global PDFs. Given the impact that the new PDFs will have on physics comparisons at the LHC, it is crucial to perform a benchmarking among the PDFs, similar in spirit to that which was carried out for PDF4LHC15, widely used for LHC physics. In this article we detail a benchmarking comparison of three global PDF sets - CT18, MSHT20 and NNPDF3.1 - and their similarities and differences that have been observed. The end result of this study will be a new PDF4LHC21 ensemble of combined PDFs suitable for a wide range of LHC applications.

Cridge, Thomas↗

Benchmarking Operators in Deep Neural Networks for Improving Performance Portability of SYCL

SYCL is a portable programming model for heterogeneous computing, so it is important to obtain reasonable performance portability of SYCL. Towards the goal of better understanding and improving performance portability of SYCL for machine learning workloads, we have been developing benchmarks for basic operators in deep neural networks (DNNs). These operators could be offloaded to heterogeneous computing devices such as graphics processing units (GPUs) to speed up computation. In this paper, we introduce the benchmarks, evaluate the performance of the operators on GPU-based systems, and describe the causes of the performance gap between the SYCL and Compute Unified Device Architecture (CUDA) kernels. We find that the causes are related to the utilization of the texture cache for read-only data, optimization of the memory accesses with strength reduction, use of local memory, and register usage per thread. We hope that the efforts of developing benchmarks for studying performance portability will stimulate discussion and interactions within the community.

Jin, Zheming [ORNL] (ORCID:000000027197780X)↗

Accelerating Thermochemical Equilibrium Calculations for Nuclear Reactor Applications

Thermochemical properties play a key role in modeling and simulation of several key phenomena in nuclear reactors. There has been an increasing interest in incorporating CALPHAD-based formulations in multiphysics simulations including for Molten Salt Reactors where knowledge of phase evolution of the salt and the chemical potentials of various elements are of utmost importance in source term analyses and redox control. However, the size of such simulations is often limited by the high computational cost of full thermodynamic equilibrium calculations. This work discusses the current efforts aimed at accelerating thermochemical equilibrium calculations for multiphysics simulations performed using the open-source finite element / finite volume code Multiphysics Object Oriented Simulation Environment (MOOSE) [1]. While several methods have been proposed for accelerating phase equilibrium calculations [2], most focus on relatively small systems and often rely on a- priori knowledge of the state-space of the system. Nuclear materials, however, are often multi-component systems owing to the evolution of composition under irradiation and an approach based on a-priori mapping of phase diagram is often not enough. This work is aimed at demonstrating an on-the-fly surrogate modeling framework that uses active learning to reduce the number of full equilibrium calculations that must be performed. By combining with efficient coupling approaches, the surrogate framework helps in reducing the computational cost of thermodynamic equilibrium informed multiphysics simulations of nuclear materials. The performance is benchmarked against full coupling with the thermochemistry library Thermochimica [3]. This work uses a machine learning based approach for constructing surrogate models to predict the stable phases in a multicomponent system. The surrogates were constructed using neural networks and Gaussian process classification. In this work, we compare the relative performance of the two methods. We also demonstrate the use of caching previous calculations by interpolating the values from nearest neighbors. References [1] Lindsay, A.D., et al. "2.0 – MOOSE: Enabling massively parallel multiphysics simulation", SoftwareX, 20 (2022): 101202. [2] Roos, W.A. and Zietsman J.H. "Accelerating complex chemical equilibrium calculations – A Review", Calphad, 77 (2022): 102380. [3] Piro, M.H.A., et al. "The thermochemistry library Thermochimica", Computational Materials Science, 67 (2013): 266-272.

36 MATERIALS SCIENCE↗

KOMPASS-II: Compaction of Crushed salt for Safe Containment – Phase 2

Long-term stable sealing elements are a basic component in the safety concept for a possible repository for heat-emitting radioactive waste in rock salt. The sealing elements will be part of the closure concept for drifts and shafts. They will be made from a welldefinied crushed salt in employ a specific manufacturing process. The use of crushed salt as geotechnical barrier as required by the German Site Selection Act from 2017 /STA 17/ represents a paradigm change in the safety function of crushed salt, since this material was formerly only considered as stabilizing backfill for the host rock. The demonstration of the long-term stability and impermeability of crushed salt is crucial for its use as a geotechnical barrier. The KOMPASS-II project, is a follow-up of the KOMPASS-I project and continues the work with focus on improving the understanding of the thermal-hydraulic-mechanical (THM) coupled processes in crushed salt compaction with the objective to enhance the scientific competence for using crushed salt for the long-term isolation of high-level nuclear waste within rock salt repositories. The project strives for an adequate characterization of the compaction process and the essential influencing parameters, as well as a robust and reliable long-term prognosis using validated constitutive models. For this purpose, experimental studies on long-term compaction tests are combined with microstructural investigations and numerical modeling. The long-term compaction tests in this project focused on the effect of mean stress, deviatoric stress and temperature on the compaction behavior of crushed salt. A laboratory benchmark was performed identifying a variability in compaction behavior. Microstructural investigations were executed with the objective to characterize the influence of pre-compaction procedure, humidity content and grain size/grain size distribution on the overall compaction process of crushed salt with respect to the deformation mechanisms. The created database was used for benchmark calculations aiming for improvement and optimization of a large number of constitutive models available for crushed salt. The models were calibrated, and the improvement process was made visible applying the virtual demonstrator.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Creating Benchmark Data for Artificial Intelligence and Machine Learning Space Biology Research

To identify an appropriate AI/ML approach for a specific problem, the best practice is to measure algorithm performance through the benchmarking process. A scientific benchmark consists of an AI-ready dataset and a reference implementation on a specific scientific question. The NASA Science Mission Directorate (SMD) has started the “Benchmark Initiative for AI/ML to create scientific benchmark datasets in three applications: 1) scientific benchmarking, which finds the best algorithm for a specific problem; 2) application benchmarking, which measures algorithm performance against a set of parameters; and 3) system benchmarking, which evaluates performance of hardware and software architecture. Currently, there are no standardized datasets available to benchmark AI/ML algorithms in the domain of space biology. In this work, we constructed two AI/ML-ready biological datasets from experiments in space-flown mice: cellular imaging and RNA-seq. First, radiation-exposed immune cells harbor DNA damage foci that can be fluorescently marked to visualize the amount of damage following exposure to ionizing radiation. However, such large datasets are difficult to analyze visually, due to imaging inconsistencies and human bias, and classical image processing approaches can fail on imaging artifacts. AI/ML are therefore exciting alternative, providing the speed of machines and the accuracy of humans. We have made this dataset available at https://registry.opendata.aws/bps_microscopy/. Second, high-throughput nucleic acid sequencing (DNA-seq, RNA-seq) has become widespread in biomedical research due to the growing availability and affordability of these assays. However, most sequencing datasets suffer from high dimensionality and low sample count. In this work, we used a generative adversarial network to synthesize a standardized, AI-ready, publicly available benchmark dataset for space biology RNA-seq data with sufficient space-flown and ground control mouse liver samples from NASA GeneLab. This dataset is available at https://registry.opendata.aws/bps_rnaseq/. These datasets are now fully open the Space Biology community to test their favorite AI/ML approaches.

James Casaletto↗

Precision-controlled ultrafast electron microscope platforms. A case study: Multiple-order coherent phonon dynamics in 1T-TaSe2 probed at 50 fs–10 fm scales

We report on the first detailed beam tests attesting the fundamental principle behind the development of high-current-efficiency ultrafast electron microscope systems where a radio frequency (RF) cavity is incorporated as a condenser lens in the beam delivery system. To allow for the experiment to be carried out with a sufficient resolution to probe the performance at the emittance floor, a new cascade loop RF controller system is developed to reduce the RF noise floor. Temporal resolution at 50 fs in full-width-at-half-maximum and detection sensitivity better than 1% are demonstrated on exfoliated 1T-TaSe2 system under a moderate repetition rate. To benchmark the performance, multi-terahertz edge-mode coherent phonon excitation is employed as the standard candle. The high temporal resolution and the significant visibility to very low dynamical contrast in diffraction signals via high-precision phase-space manipulation give strong support to the working principle for the new high-brightness femtosecond electron microscope systems.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Exxon Mobil-NETL Testing of DAC Materials

This joint initiative is aimed at extending our understanding of real DAC testing conditions at the NETL DAC testbed facility. Of particular focus will be: • Small Scale testing of powdered and formulated materials (supplied by ExxonMobil) to evaluate various performance metrics under DAC process cycles • Pilot scale testing of formulated materials at larger scales ExxonMobil will work with NETL to shake down equipment, validate testing methods, and define best practices for data analysis. Three tasks are proposed: 1. Validation of multi-cycle test data on powdered and formulated materials to benchmark various performance metrics of these materials. 2. Steam regeneration of powdered samples at small scale to define baseline performance under commercially relevant conditions. 3. Pilot scale testing of larger formulated materials to test commercially relevant samples under actual process cycles.

36 MATERIALS SCIENCE↗