Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Benchmark data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Impact of recent ENDF nuclear data update, high initial enrichment and high burnup fuel on critical experiments applicability determination via the integral index c k for burnup credit validation

In 2012, NUREG/CR-7109 reported on the validation of burnup credit calculations involving major and minor actinides and major fission products which was investigated for pressurized and boiling water reactor (PWR and BWR) fuel enrichments up to 5 wt% 235 U and assembly-average burnups up to 60 GWd/MTU. Recently, there has been interest in increasing the maximum enrichment used in PWR fuel as high as 8 wt% 235 U and correspondingly increasing the maximum assembly-average burnups to approximately 75 GWd/MTU. These proposed increases in enrichment and burnup necessitate reinvestigation of the validation basis for k eff calculations for this expanded application space. Additionally, the 2012 study was performed by using the Evaluated Nuclear Data File (ENDF)/B-VII.0 nuclear data with the SCALE 6 covariance library, and the effects of using the newly released ENDF/B-VII.1 and ENDF/B-VIII.0 nuclear data and covariance libraries should be evaluated. In this work, published in NUREG/CR-7309 in 2025, the validation assessment was performed consistently with NUREG/CR-7109: modeling irradiated fuel assemblies in the Generic Burnup Credit (GBC)-32 cask defined in NUREG/CR-6747. The TSUNAMI-3D sequence was used to generate sensitivity data for the application model, and the data were compared with sensitivity data from select benchmark models. The integral parameter c k is the metric of similarity used in this study and is consistent with NUREG/CR-7109, where a c k value in excess of 0.8 indicates sufficient similarity for use in validation. A new set of benchmark experiments with sensitivity data has been assembled for this effort. The number of experiments with available sensitivity data is now 2,104, compared to 474 in NUREG/CR-7109. This increase was facilitated by the efforts of the Nuclear Energy Agency to generate sensitivity data for a majority of the experiments in the International Criticality Safety Benchmark Evaluation Project (ICSBEP) Handbook to supplement the data available in the Oak Ridge National Laboratory (ORNL) Verified, Archived Library of Inputs and Data (VALID). The complete set of benchmarks considered here includes experiments for low-enriched uranium (LEU), intermediate enriched uranium (IEU), and a mixture of uranium and plutonium (MIX) from the ICSBEP Handbook and VALID, as well as ORNL models of the Haut Taux de Combustion (HTC) experiments and other potentially relevant models not included in VALID. The updated similarity study shows that none of the extended burnup and higher enrichment combinations considered show a significant decrease in the number of potentially applicable experiments, meaning sufficient critical experiments exist for the validation of BUC criticality safety calculations, with initial enrichments up to 8 wt% 235 U and burnups up to 80 GWd/MTU. Additionally, both the ENDF/B-VII.1 and ENDF/B-VIII.0 nuclear data libraries can be used for validation since the number of critical experiments applicable for validation increases for most cases with the most recent nuclear data compared to the previous one. As in previous BUC validation studies, the French HTC experiments are the most similar in a majority of the application cases studied, especially from representative discharge burnups ranging from 40 to 80 GWd/MTU. In conclusion, these results match the conclusions presented in NUREG/CR-7109 regarding validation of the primary actinides in BUC analyses.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Ice Accretions and Full-Scale Iced Aerodynamic Performance Data for a Two-Dimensional NACA 23012 Airfoil

This report documents the data collected during the large wind tunnel campaigns conducted as part of the SUNSET project (StUdies oN Scaling EffecTs due to ice) also known as the Ice-Accretion Aerodynamics Simulation study: a joint effort by NASA, the Office National d'Etudes et Recherches Aérospatiales (ONERA), and the University of Illinois. These data form a benchmark database of full-scale ice accretions and corresponding ice-contaminated aerodynamic performance data for a two-dimensional (2D) NACA 23012 airfoil. The wider research effort also included an analysis of ice-contaminated aerodynamics that categorized ice accretions by aerodynamic effects and an investigation of subscale, low- Reynolds-number ice-contaminated aerodynamics for the NACA 23012 airfoil. The low-Reynolds-number investigation included an analysis of the geometric fidelity needed to reliably assess aerodynamic effects of airfoil icing using artificial ice shapes. Included herein are records of the ice accreted during campaigns in NASA Glenn Research Center's Icing Research Tunnel (IRT). Two different 2D NACA 23012 airfoil models were used during these campaigns; an 18-in. (45.7-cm) chord (subscale) model and a 72-in. (182.9-cm) chord (full-scale) model. The aircraft icing conditions used during these campaigns were selected from the Federal Aviation Administration's (FAA's) Code of Federal Regulations (CFR) Part 25 Appendix C icing envelopes. The records include the test conditions, photographs of the ice accreted, tracings of the ice, and ice depth measurements. Model coordinates and pressure tap locations are also presented. Also included herein are the data recorded during a wind tunnel campaign conducted in the F1 Subsonic Pressurized Wind Tunnel of ONERA. The F1 tunnel is a pressured, high- Reynolds-number facility that could accommodate the full-scale (72-in. (182.9-cm) chord) 2D NACA 23012 model. Molds were made of the ice accreted during selected test runs of the full-scale model in the IRT. From these molds, castings were made that closely replicated the features of the accreted ice. The castings were then mounted on the full-scale model in the F1 tunnel, and aerodynamic performance measurements were made using model surface pressure taps, the facility force balance system, and a large wake rake designed specifically for these tests. Tests were run over a range of Reynolds and Mach numbers. For each run, the model was rotated over a range of angles-of-attack that included airfoil stall. The benchmark data collected during these campaigns were, and continue to be, used for various purposes. The full-scale data form a unique, ice-accretion and associated aerodynamic performance dataset that can be used as a reference when addressing concerns regarding the use of subscale ice-accretion data to assess full-scale icing effects. Further, the data may be used in the development or enhancement of both ice-accretion prediction codes and computational fluid dynamic codes when applied to study the effects of icing. Finally, as was done in the wider study, the data may be used to help determine the level of geometric fidelity needed for artificial ice used to assess aerodynamic degradation due to aircraft icing. The structured, multifaceted approach used in this research effort provides a unique perspective on the aerodynamic effects of aircraft icing. The data presented in this report are available in electronic form upon formal approval by proper NASA and ONERA authorities.

Addy, Harold E., Jr.↗

Interannual differences of Geosat altimeter heights and sea level - The importance of a datum

Sea surface height data from the Geosat altimeter are compared with island sea level data from 18 gages in the western and central tropical Pacific during December 1986 to November 1987. Care was taken to ensure that the two data sets are referenced to the same mean surface. This was done by requiring that both data sets have a zero mean sea level over the period April 1985 to April 1986. When the annual means are computed at each station in the later time period, it is found that the annual mean sea surface height values have drifted away from the corresponding sea level values by as much as 16 cm. Further, the pattern of differences that develop between the two data sets is not random but is spatially coherent with a strong east-west gradient. These observations illustrate the necessity for sea surface height data to be referred to a well-defined zero point, a datum, in order to reliably monitor interannual changes in the sea surface topography. Until these differences can be eliminated, it will be necessary to use tide gage data as benchmarks for the altimeter sea surface height data.

Wyrtki, Klaus↗

Bayes Error Rate Estimation Using Classifier Ensembles

The Bayes error rate gives a statistical lower bound on the error achievable for a given classification problem and the associated choice of features. By reliably estimating th is rate, one can assess the usefulness of the feature set that is being used for classification. Moreover, by comparing the accuracy achieved by a given classifier with the Bayes rate, one can quantify how effective that classifier is. Classical approaches for estimating or finding bounds for the Bayes error, in general, yield rather weak results for small sample sizes; unless the problem has some simple characteristics, such as Gaussian class-conditional likelihoods. This article shows how the outputs of a classifier ensemble can be used to provide reliable and easily obtainable estimates of the Bayes error with negligible extra computation. Three methods of varying sophistication are described. First, we present a framework that estimates the Bayes error when multiple classifiers, each providing an estimate of the a posteriori class probabilities, a recombined through averaging. Second, we bolster this approach by adding an information theoretic measure of output correlation to the estimate. Finally, we discuss a more general method that just looks at the class labels indicated by ensem ble members and provides error estimates based on the disagreements among classifiers. The methods are illustrated for artificial data, a difficult four-class problem involving underwater acoustic data, and two problems from the Problem benchmarks. For data sets with known Bayes error, the combiner-based methods introduced in this article outperform existing methods. The estimates obtained by the proposed methods also seem quite reliable for the real-life data sets for which the true Bayes rates are unknown.

Tumer, Kagan↗

Benchmark Wall Heat Flux Data for a GO2/GH2 Single Element Combustor

Wall heat flux measurements in a 1.5 in. diameter circular cross-section rocket chamber for a uni-element shear coaxial injector element operating on gaseous oxygen (GOz)/gaseous hydrogen (GH,) propellants are presented. The wall heat flux measurements were made using arrays of Gardon type heat flux gauges and coaxial thermocouple instrumentation. Wall heat flux measurements were made for two cases. For the first case, GOZ/GHz oxidizer-rich (O/F=l65) and fuel-rich preburners (O/F=1.09) integrated with the main chamber were utilized to provide vitiated hot fuel and oxidizer to the study shear coaxial injector element. For the second case, the preburners were removed and ambient temperature gaseous oxygen/gaseous hydrogen propellants were supplied to the study injector. Experiments were conducted at four chamber pressures of 750, 600, 450 and 300psia for each case. The overall mixture ratio for the preburner case was 6.6, whereas for the ambient propellant case, the mixture ratio was 6.0. Total propellant flow was nominally 0.27-0.29 Ibm/s for the 750 psia case with flowrates scaled down linearly for lower chamber pressures. The axial heat flux profile results for both the preburner and ambient propellant cases show peak heat flux levels a t axial locations between 2.0 and 3.0 in. from the injector face. The maximum heat flux level was about two times greater for the preburner case. This is attributed to the higher injector fuel-to-oxidizer momentum flux ratio that promotes mixing and higher initial propellant temperature for the preburner case which results in a shorter reaction zone. The axial heat flux profiles were also scaled with respect to the chamber pressure to the power 0.8. The results at the four chamber pressures for both cases collapsed to a single profile indicating that at least to first approximation, the basic fluid dynamic structures in the flow field are pressure independent as long as the chamber/njector/nozzle geometry and injection velocities remain the same.

Marshall, William M.↗

Nuclear Data Adjustment for Nonlinear Applications in the OECD/NEA WPNCS SG14 Benchmark -- A Bayesian Inverse UQ-based Approach for Data Assimilation

The Organization for Economic Cooperation and Development (OECD) Working Party on Nuclear Criticality Safety (WPNCS) proposed a benchmark exercise to assess the performance of current nuclear data adjustment techniques applied to nonlinear applications and experiments with low correlation to applications. This work introduces Bayesian Inverse Uncertainty Quantification (IUQ) as a method for nuclear data adjustments in this benchmark, and compares IUQ to the more traditional methods of Generalized Linear Least Squares (GLLS) and Monte Carlo Bayes (MOCABA). Posterior predictions from IUQ showed agreement with GLLS and MOCABA for linear applications. When comparing GLLS, MOCABA, and IUQ posterior predictions to computed model responses using adjusted parameters, we observe that GLLS predictions fail to replicate computed response distributions for nonlinear applications, while MOCABA shows near agreement, and IUQ uses computed model responses directly. We also discuss observations on why experiments with low correlation to applications can be informative to nuclear data adjustments and identify some properties useful in selecting experiments for inclusion in nuclear data adjustment. Performance in this benchmark indicates potential for Bayesian IUQ in nuclear data adjustments.

FOS: Computer and information sciences↗

Nuclear Data Adjustment for Nonlinear Applications in the OECD/NEA WPNCS SG14 Benchmark—A Bayesian Inverse UQ-Based Approach for Data Assimilation

The Organisation for Economic Co-operation and Development Working Party on Nuclear Criticality Safety has proposed a benchmark exercise to assess the performance of current nuclear data adjustment techniques applied to nonlinear applications and experiments with low correlation to applications. This work introduces Bayesian inverse uncertainty quantification (IUQ) employing scientific machine learning surrogate models as a method for nuclear data adjustments in this benchmark, and compares IUQ to the more traditional methods of generalized linear least squares (GLLS) and Monte Carlo Bayes (MOCABA). Posterior predictions from IUQ showed agreement with GLLS and MOCABA for linear applications. Here, when comparing GLLS, MOCABA, and IUQ posterior predictions to computed model responses using adjusted parameters, we observe that the GLLS predictions failed to replicate the computed response distributions for nonlinear applications, while MOCABA showed near agreement, and IUQ used the computed model responses directly. We also discuss observations on why experiments with low correlation to applications can be informative to nuclear data adjustments and identify some properties useful in selecting experiments for inclusion in nuclear data adjustment. Performance in this benchmark indicates potential for Bayesian IUQ in nuclear data adjustments.

Bayesian calibration↗

Informing Robust Functional Relationship Benchmarks: An Evaluation of the Temperature Sensitivity of Ecosystem Respiration Across the Arctic-Boreal Region

During land model development, simulated carbon dynamics are often benchmarked against observational data sets to evaluate model performance. Functional relationship benchmarks are the relationship between a driving variable (e.g., temperature) and a response variable (e.g., ecosystem respiration) and are a promising tool for assessing model performance by evaluating modeled sensitivities to changing environmental conditions. However, observed functional relationships can be influenced by choices made during data collection and throughout the benchmarking process, impacting the inferred skill of land models. To avoid misrepresenting a model's true performance, it is necessary to systematically evaluate best practices when constructing functional relationship benchmarks. We developed a set of guidelines for constructing functional relationship benchmarks, considering the choice of data set, number of daily observations, temporal extent, and temporal resolution across Alaska and Canada over a 20-year period from 2001 to 2020. The temperature sensitivity of ecosystem respiration from observations, evaluated through an apparent Q 10 , is highly variable both spatially and as a result of the data processing approach applied in the benchmark formation. When benchmarking 13 models from the Warming Permafrost Model Intercomparison Project (WrPMIP), the range in inferred model skill is substantially impacted by the choices applied in constructing functional relationship benchmarks. The inferred performance of a given model is most sensitive to the number of daily observations and temporal extent, followed by choice of benchmark data set and temporal averaging. Results from this analysis can guide the development of consistent and robust functional relationships for future model evaluation studies.

Poe, Jeralyn [Northern Arizona University, Flagsta↗

On-Demand Column Joining for High Energy Physics

As the Large Hadron Collider (LHC) transitions into the High-Luminosity LHC (HL-LHC) era, the volume of data to be processed is expected to increase significantly. The CMS Experiment currently utilizes various data formats, including AOD, MiniAOD, and NanoAOD, each with different levels of detail and storage requirements. This paper addresses the challenges of data duplication and storage inefficiencies in high-energy physics (HEP) analyses by proposing an on-demand column-joining solution. This approach aims to reduce data duplication by enabling the dynamic combination of NanoAOD data with auxiliary information from larger data tiers, such as MiniAOD. The proposed solution leverages Trino, a high-performance distributed SQL query engine, to perform efficient and scalable data joins. Benchmarks using CMS OpenData demonstrate the feasibility of this approach, showing that it can handle large datasets with low latency. Integration with the scikit-hep ecosystem and the coffea analysis framework is also discussed, highlighting the potential for seamless end-to-end data processing and analysis. Ongoing and future work focuses on expanding benchmarks, integrating ServiceX for data transformation, and exploring the use of native object storage solutions.

Manganelli, Nicholas [Northeastern U.]↗

Benchmark Tracking System for Performance Monitoring

Benchmarking is essential for high-performance software development, particularly for monitoring performance across code iterations. This project focused on enhancing the benchmarking process for Lamellar, an asynchronous runtime for High-Performance Computing (HPC) systems developed at Pacific Northwest National Laboratory. Prior to this work, benchmark results were difficult to track and compare across code versions, presenting significant challenges in identifying performance regressions and long-term trends. The primary objective was to establish a systematic, reproducible approach for measuring performance and detecting regressions following code commits. Our methodology involved three key components: standardizing benchmark outputs, implementing data versioning, and developing analysis tools. We standardized the benchmark output format to JSON Line records containing specific fields (execution time, hardware specifications, and environmental variables). To address data management challenges, we evaluated several options and eventually chose a git repository dedicated to benchmark data. We developed a suite of Python tools that processed benchmark results, enriched them with metadata, and facilitated search in the repository. The resulting system enables more efficient filtering and comparison of performance metrics across commit histories, hardware configurations, and benchmark variants through a unified query interface. Our implementation reduces computational overhead by first checking for existing results through configuration matching before initiating new benchmark runs, thereby conserving resources. The system has been validated by Lamellar developers. It organizes results by benchmark type and build configurations for efficient retrieval. Future developments include a planned Large Language Model interface for predicting benchmark performance, incorporating the criterion package for statistical analysis, which will enable automated detection of statistically significant performance changes, and integration with continuous integration pipelines. Despite these enhancements being reserved for future work, this project has successfully provided the Lamellar development team with a framework for maintaining consistent performance standards and identifying optimization opportunities across workloads and hardware environments.

97 MATHEMATICS AND COMPUTING↗

Dynamic object management for distributed data structures

In distributed-memory multiprocessors, remote memory accesses incur larger delays than local accesses. Hence, insightful allocation and access of distributed data can yield substantial performance gains. The authors argue for the use of dynamic data management policies encapsulated within individual distributed data structures. Distributed data structures offer performance, flexibility, abstraction, and system independence. This approach is supported by data from a trace-driven simulation study of parallel scientific benchmarks. Experimental data on memory locality, message count, message volume, and communication delay suggest that data-structure-specific data management is superior to a single, system-imposed policy.

Totty, Brian K.↗

Technology Transfer Plan: LuNaMaps Project

The main contribution of this project is the combined knowledge of terrain relative navigation experts and lunar scientists who are familiar with both the lunar orbital imagery and the instruments that collected the data as well as how a TRN system utilizes map data. This knowledge comes in the form of published technical papers, benchmark map data sets, and software tools that can help others automate the process of creating the necessary maps for their own landing sites in the future. This document represents the project's plans to share all the lessons learned, processes developed, and applicable software tools with the public.

optical navigation↗

Tutorial on LuNaMaps Developed Tools andProcesses for Mapping the Lunar Surface

The main contribution of this project is the combined knowledge of terrain relative navigation experts and lunar scientists who are familiar with both the lunar orbital imagery and the instruments that collected the data as well as how a TRN system utilizes map data. This knowledge comes in the form of published technical papers, benchmark map data sets, and software tools that can help others automate the process of creating the necessary maps for their own landing sites in the future. This presentation provides a brief overview of the tools and processes developed by the project.

optical navigation↗

Status Report on Characterization of High Burnup Fuel with Advanced Nondestructive Pulsed Neutron PIE

Characterizing irradiated or spent nuclear fuels with pulsed neutron techniques provides microstructural data such as phase fractions as well as crystallographic data, e.g. lattice parameters, from diffraction analysis. Diffraction characterization is complemented by spatially resolved mapping of isotope densities from energy-resolved neutron imaging, in particular neutron absorption resonance imaging, and overall bulk isotope assay with better sensitivity for minority isotopes from neutron absorption resonance spectroscopy without spatial resolution. Furthermore, after characterization at ambient condition, heating of irradiated or spent fuel will allow to characterize differences of e.g. lattice thermal expansion or phase transition temperature and kinetics compared to fresh fuel as well as enable the study of disappearance of irradiation defects. This data enables benchmarking of predictions of properties of irradiated fuels for which otherwise experimental data is sparse. The effort described here strives to characterize a section cut from a high-burnup fuel. Volumes smaller than entire fuel pellets or rodlets as proposed here, e.g. sections cut from a fuel pellet, to pave the way to characterize entire pellets or rodlets in the future.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

High Reynolds Number Effects on Multi-Hole Probes and Hot Wire Anemometers

The paper reports on the results from an experimental investigation of the response of multi-hole and hot wire probes at high flow Reynolds numbers (Re approx. 10(exp 6)). The limited results available in literature for 5-hole probes are restricted to Re approx. 10(exp 4). The experiment aims to investigate the probe response (in terms of dimensionless pressure ratios, characterizing pitch, and yaw angles and the total and static pressures) at high Re values and to gauge their effect on the calculated velocity vector. Hot wire calibrations were also undertaken with a parametric variation of the flow pressure, velocity and temperature. Different correction and calibration schemes are sought to be tested against the acquired data set. The data is in the analysis stage at the present time. The test provided good benchmark quality data that can be used to test future calibration and testing methods.

Ramachandran, N.↗

Comparison of URANS and LES predictions for the open phase of the OECD NEA CSNI fluid structure interaction CFD benchmark

The OECD NEA CSNI WGAMA CFD Task Group ran a benchmark in 2020 and 2021 to assess the predictive capabilities of coupled fluid structure interaction (FSI) CFD analysis methods. This paper presents the predictions made for the open phase of the benchmark using URANS and LES turbulence modelling approaches, and a comparison of the results to the experimental data. The benchmark comprised a channel containing two inline cylinders in cross-flow. The cylinders were fixed at one end, free at the other, and had measured resonant frequencies and damping properties. The URANS modelling used ANSYS Fluent 2-way coupled to ANSYS Mechanical. The LES modelling used Nek5000, 1-way coupled to Diablo. Comparisons with cross-channel velocity profiles are presented, both for the mean flow and its RMS. Comparisons are also made to the frequency spectra for point measurements of fluid velocity and pressure, and for the accelerations of the free end of each cylinder. URANS predicts the average velocity profiles relatively well, and is able to predict the velocity and acceleration spectra at the shedding frequency. However, the frequency content at the 4th harmonic of the shedding frequency is low in the URANS flow fields, and so does not excite accelerations at the resonant frequency of the cylinders. LES makes better predictions of the average profiles, and the velocity spectra agree well at both the shedding frequency and at higher frequencies. In conclusion, the 1-way coupled LES results show good agreement for acceleration spectra.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Comparison of secondary flows predicted by a viscous code and an inviscid code with experimental data for a turning duct

A comparison of the secondary flows computed by the viscous Kreskovsky-Briley-McDonald code and the inviscid Denton code with benchmark experimental data for turning duct is presented. The viscous code is a fully parabolized space-marching Navier-Stokes solver while the inviscid code is a time-marching Euler solver. The experimental data were collected by Taylor, Whitelaw, and Yianneskis with a laser Doppler velocimeter system in a 90 deg turning duct of square cross-section. The agreement between the viscous and inviscid computations was generally very good for the streamwise primary velocity and the radial secondary velocity, except at the walls, where slip conditions were specified for the inviscid code. The agreement between both the computations and the experimental data was not as close, especially at the 60.0 deg and 77.5 deg angular positions within the duct. This disagreement was attributed to incomplete modelling of the vortex development near the suction surface.

Schwab, J. R.↗

Comparison of secondary flows predicted by a viscous code and an inviscid code with experimental data for a turning duct

A comparison of the secondary flows computed by the viscous Kreskovsky-Briley-McDonald code and the inviscid Denton code with benchmark experimental data for turning duct is presented. The viscous code is a fully parabolized space-marching Navier-Stokes solver while the inviscid code is a time-marching Euler solver. The experimental data were collected by Taylor, Whitelaw, and Yianneskis with a laser Doppler velocimeter system in a 90 deg turning duct of square cross-section. The agreement between the viscous and inviscid computations was generally very good for the streamwise primary velocity and the radial secondary velocity, except at the walls, where slip conditions were specified for the inviscid code. The agreement between both the computations and the experimental data was not as close, especially at the 60.0 deg and 77.5 deg angular positions within the duct. This disagreement was attributed to incomplete modeling of the vortex development near the suction surface.

Schwab, J. R.↗