Engineering PapersSearch

SEARCH · Engineering Papers

Results for “Evaluation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

New procedure for evaluation of U(3) coupling and recoupling coefficients

A simple method to calculate Wigner coupling coefficients and Racah recoupling coefficients for U(3) in two group–subgroup chains is presented. While the canonical U(3) coupling and recoupling coefficients are applicable to any system that respects U(3) symmetry, the U(3) coupling coefficients are more specific to nuclear structure studies. This new procedure precludes the use of binomial coefficients and alternating sums which were used in the 1973 formulation of Draayer and Akiyama, and in so doing provides a faster and more accurate determination of any and all required results. The resolution of the outer multiplicity is based on the null space concept of the U(3) generators proposed by Alex et al., whereas the inner multiplicity in the angular momentum subgroup chain is obtained from the dimension of the null space of the SO(3) raising operator. It is anticipated that a C++ library will ultimately be available for determining generic coupling and recoupling coefficients associated with both the canonical and the physical group–subgroup chains of U(3).

Cross-Coupling Reaction

Freeze It or Leave It? Evaluating the Role of Cryo-Electron Microscopy in Battery Research

Cryogenic electron microscopy (cryo-EM) continues to gain prominence in materials science, particularly in battery research where it has enabled high-resolution, multimodal characterization of electrode materials and interfaces that otherwise degrade quickly under electron beam irradiation. But as anyone who has attempted cryo-EM techniques knows, freezing comes at a cost; cryo-EM experiments are time-consuming, highly sensitive, and carry an increased risk of artifacts due to issues such as frost contamination. Thus, when planning new characterization of battery materials or other beam-sensitive samples, it is critical to consider whether (and which) cryo-EM techniques are appropriate, based on study goals and an understanding of electron beam-sample interactions. Here we review such considerations for battery materials to elucidate the questions of when, why, and how to freeze to achieve high-quality characterization.

25 ENERGY STORAGE

Consistent $\overline{ν}$ evaluation for minor U isotopes with $\tt{CGMF}$

Following several successful prompt $\overline{ν}$ evaluations using $\tt{CGMF}$, including consistent evaluations for minor Pu isotopes, we detail in this report our efforts to perform a consistent $\overline{ν}$ evaluation for minor U isotopes during FY25. Although we have not yet produced a finalized evaluation, we present the progress that we have made towards such an evaluation for 232,233,234,236,237,239 U prompt $\overline{ν}$. Our milestone explicitly calls out evaluations for 233 U, 234 U, and 236 U, however, to better constrain the model with reliable experimental $\overline{ν}$ data, we also include 235 U and 238 U in the evaluation procedure. Then, we additionally produce evaluations for 232 U, 237 U and 239 U $\overline{ν}$ as a byproduct. Elsewhere, we will report our efforts on a stand-alone 233 U $\overline{ν}$ evaluation. This report is organized in the following manner. In Sec. 2, we briefly outline the updates to CGMF that were needed to be able to calculate all of these minor U fission reactions. The experimental data overview is given in Sec. 3. The evaluation methodology and results are presented in Secs. 4 and 5, respectively. Finally, we conclude and outline work for FY26 in Sec. 6.

07 ISOTOPE AND RADIATION SOURCES

Re-evaluating the prompt fission neutron spectrum of spontaneously fissioning 252 Cf

The prompt fission neutron spectrum (PFNS) of spontaneously fissioning 252 Cf is a Neutron Data Standards observable. Nearly all fission spectra of actinides were measured relative to it, using efficiencies derived from it, or analyzed with simulations validated by it. The current Standards evaluation was published by W. Mannhart in 1987. It could not be updated because the evaluation input, experimental mean values and covariances, were lost. First, we attempt to reproduce it. However, Mannhart’s evaluation can only be reproduced within its one-σ uncertainties as some of its aspects (e.g., experimental covariances, rejected data points) remain unknown. Therefore, a new evaluation is presented: We revisit all existing experimental 252 Cf(sf) PFNS data, including those published after the release of the current Standards evaluation, and re-estimate associated covariances. The newly evaluated 252 Cf(sf) PFNS differs distinctly from Mannhart’s below 300 keV and extends it to lower and higher outgoing neutron energies (500 eV–25 MeV). The new evaluated uncertainties are larger from 3–9 MeV and smaller otherwise. Spectrum averaged cross sections of importance to the International Reactor Dosimetry and Fusion File community calculated with the new spectrum are close to those calculated with Mannhart’s evaluation and agree with experimental values well within their uncertainties.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

ICSBEP evaluation lessons learned and good practice

This document describes lessons learned during my nearly 20 years of experience in International Criticality Safety Benchmark Evaluation Project (ICSBEP) meetings and in the development of several evaluations employing the IPEN/MB-01 reactor. The evaluations cover a wide range of possible configurations at the reactor core of this facility. Most of the evaluations are related to critical configurations, although two of them concern subcritical measurements. Examples of the evaluations submitted and approved for ICSBEP publication are the critical configurations employed in the standard core; critical configurations employing heavy reflectors composed of stainless steel, nickel, and carbon steel; and critical configurations employing the standard fuel rods and fuel rods composed of UO 2 -Gd 2 O 3 , among several others. Additionally, this document addresses good practices for the development of ICSBEP evaluations; the intention is to help beginners who are going to start evaluations for ICSBEP. The sections of the ICSBEP evaluation are described and illustrated with several examples.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Evaluation of HVAC & refrigeration system fault behaviors and impacts: A systematic review

Achieving the goals of green buildings critically depends on the fault-free operation of heating, ventilation and air conditioning and refrigeration (HVAC&R) systems. However, faults frequently occur in these systems, causing a range of negative consequences, including increased energy consumption, diminished operational performance, compromised indoor environmental quality, higher operational costs, and shortened system lifespan. The evaluation of fault behaviors and impacts plays a critical role in revealing fault characteristics and consequently supports many research areas, including the design of the high-performance equipment, development of fault detection and diagnostics (FDD) and robust control approaches, as well as the enhancement of maintenance decision-making activities. This paper systematically reviews 112 research publications that reported the analysis and evaluation of fault behaviors and impacts in HVAC&R systems over the past thirty years. Here, we designed a review approach to address five crucial research questions, namely: 1) the objectives of analysis and evaluation of fault behaviors and impacts, 2) data sources, 3) equipment/system types and fault types, 4) evaluation methods including evaluation measures and associated metrics, and 5) challenges and future directions in the research on evaluating of fault behaviors and impacts. In-depth discussions on these questions help bridge the gap between the evaluation of fault behaviors and impacts and their practical applications, such as the development of high-performance systems, fault models, FDD methods, and maintenance decision-making tools within the HVAC&R FDD domain.

Chen, Yimin [Oak Ridge National Laboratory (ORNL),

Consistent $\overline{ν}$ evaluation for minor Pu isotopes with CGMF

This report follows up on the evaluation work done in 2023 where we developed a procedure and performed the first consistent evaluation of the average prompt neutron multiplicity for minor Pu isotopes, including 238,240-242 Pu, using CGMF. Details on the necessary updates to the release version of CGMF, the experimental data and experimental uncertainty quantification that went into the evaluation, and this first evaluation effort are documented in and will not be repeated in this report. Instead, we discuss here the further investigations into the CGMF model space to improve the evaluation results from FY23. Additionally, we note that these evaluation results have been transformed into the ENDF format for mean values and covariances, and validation with critical assemblies has been performed; that work is documented elsewhere. In this short report, we discuss the updated evaluation efforts performed during FY24 in Sec. 2, show results from CGMF for other prompt fission observables in Sec. 3, and then briefly conclude in Sec. 4.

07 ISOTOPE AND RADIATION SOURCES

Status of the International Criticality Safety Benchmark Evaluation Project

The International Criticality Safety Benchmark Evaluation Project (ICSBEP) has continued its work generating evaluations of new and historical benchmark experiments since the last update to the nuclear criticality safety (NCS) community at the 12th International Conference on Nuclear Criticality Conference held in 2023. One additional version of the ICSBEP Handbook has been published since that update, and the Technical Review Group (TRG) held two in-person meetings to review and approve additional benchmarks. The 2022 and 2023 editions of the handbook were combined into one release (published in November 2024) and contained 13 new evaluations with 46 different configurations and two major revisions to existing evaluations. The 2024 version of the handbook, currently under publication review, will contain two new evaluations with 15 new configurations and one major revision to HEU-MET-FAST-028, the evaluation of Flattop with a uranium core. The ICSBEP TRG met again in person in April 2025 to review benchmarks for the 2025 ICSBEP Handbook and final comment resolution is currently ongoing. Many of the new benchmarks represent contemporaneous experiments that have been specifically optimized to provide validation cases relevant to the NCS community. One major area of focus for new critical experiments is to target the sparsely populated intermediate energy (or resonance) region. Another focus of many of the new benchmarks is to provide experiments sensitive to different materials, such as chlorine, hafnium, tantalum, titanium, molybdenum, chromium, and polymethyl methacrylate (PMMA, or Lucite). The ICSBEP continues to deliver high-quality, peer reviewed evaluations of integral experiments relevant to the nuclear data community.

HEU-MET-FAST-028

Resolved Resonance Evaluation for Neutron Interactions with 103 Rh up to 8 keV

A neutron cross-section evaluation for the n + 103 Rh reaction in the resolved resonance region was carried out in the energy range 10−5 eV to 8 keV encompassing thermal energy at 0.0253 eV. The scope of this work is to generate resonance parameters and resonance parameter covariances based on the Reich-Moore reduced R-matrix formalism using the code SAMMY. Some features of the new evaluation are the inclusion of high-resolution capture data in the SAMMY evaluation process and the extension of the resolved resonance range from 4 to 8 keV. Furthermore, the evaluation employs more accurate resonance parameter representation by exploring the use of the LRF = 7 ENDF feature and also the use of the LCOMP = 2 compact format for resonance parameter covariance representation. Included in the SAMMY evaluation are transmission data, capture cross-section data, and neutron scattering length information. Thermal cross-section values listed in the literature, as well as capture resonance integrals, were also incorporated into the evaluation process.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS

Guideline for Characterizing and Evaluating a Candidate Project Site for Solar Thermal Applications

This document presents a structured procedure for characterizing and evaluating candidate project sites for concentrating solar power (CSP) and solar heat for industrial processes (SHIP) applications. The objective is to provide project developers, researchers, and other stakeholders with a consistent, technology-agnostic framework for early-stage site assessment, enabling informed decision-making prior to significant investment in project development. Site selection is a critical factor in project success or failure for both CSP and SHIP projects. Key factors such as solar resource availability, land characteristics, environmental and regulatory constraints, infrastructure availability, and community context are determined by the choice of project site and can materially impact project performance, cost, schedule, and overall viability. This procedure is designed to systematically evaluate these factors, identify potential fatal flaws, and prioritize the most favorable candidate sites for further development. The process begins with rapid screening-level evaluation, using publicly available data to assess solar resource, land availability and suitability, zoning and land-use compatibility, and exclusion zones such as protected lands or sensitive habitats. Sites that meet the minimum screening criteria advance to a more detailed characterization. Subsequent sections of this report provide guidance for a next-level assessment of the most important technical and environmental parameters, including: 1) Solar resource quality, variability, and uncertainty using multiyear datasets and, where appropriate, on-site measurement campaigns; 2) Meteorological conditions such as wind, temperature, extreme weather events, and soiling impacts; 3) Land characteristics including slope, shading, and geotechnical conditions; and 4) Environmental and regulatory considerations, including permitting processes, endangered species, cultural resources, and visual impacts. The procedure also addresses infrastructure and integration considerations, including: 1) Grid interconnection requirements for CSP power generation projects; 2) Electrical and operational integration for SHIP facilities; 3) Water availability, quality, and permitting constraints, which are particularly critical for CSP in arid regions; and 4) Site access, construction logistics, and availability of workforce and supporting services. Recognizing the importance of social and economic context, the procedure includes evaluation of community engagement factors, such as stakeholder sentiment, proximity to sensitive visual receptors, workforce development opportunities, and local economic incentives. The outputs of these assessments are synthesized in a cost and risk evaluation, translating site characteristics into expected impacts on capital cost, operating cost, schedule, and technical risk. This is complemented by screening-level performance modeling, including 8760 simulations and long-term projections, to quantify expected energy or thermal output, assess variability thereof, and support comparison between candidate sites. Finally, the procedure provides high-level guidance on a structured go/no-go decision framework, categorizing sites based on identified risks and constraints, and outlining a clear path forward to feasibility studies and front-end engineering design for viable projects. By standardizing the site characterization process across both CSP and SHIP applications, this guideline aims to: 1) Improve consistency and transparency in early-stage project evaluation; 2) Reduce development risk and avoid investment in nonviable project sites; 3) Support collaboration between developers, researchers, and public agencies; and 4) Accelerate successful deployment of concentrating solar technologies for both power generation and industrial process heat.

14 SOLAR ENERGY

Evaluation of Saccadic Component Measure on Smooth Pursuit Tests

ABSTRACT Introduction Despite the advancement of eye-tracking technology for smooth pursuit (SP) eye movement evaluation, qualitative observation offers much information that is not captured by computers; hence, both objective and qualitative information should be utilized to evaluate SP. This study examined the consistency among our clinicians when evaluating SP using normal (N), grossly normal (GN), mildly abnormal (MA), and abnormal (AB) as classifications. We then evaluated the effect of combining GN and MA into a single subclinical (SUBC) category. We also evaluated the computerized percent saccade (PS) metric by determining its sensitivity and specificity in classifying SP. Materials and Methods Retrospective horizontal and vertical SP test videos and numerical data for 70 participants were obtained from the Neuro Kinetics Neuro-Otologic Test Center and de-identified. From this, eye-tracking videos, time plots of eye-tracking positional data, and tables of SP eye-tracking performance data were generated for 0.1, 0.3, and 0.5 Hz in both horizontal and vertical planes, totaling 6 tests per subject. Three clinicians rated each subject’s SP performance as N, GN, MA, or AB for a total of 6 ratings (3 frequencies, horizontal and vertical). This process was repeated using N, SUBC, and AB as rating categories. Clinicians also provided an overall SP rating for each plane as follows: AB if the results were abnormal for 2 or more frequencies tested. Alternatively, if fewer than 2 frequencies presented with a rating of AB, then an overall rating of MA, GN, or N was determined at the respective clinician’s discretion. Results When the 3 clinicians were tasked with classifying SP videos using 4 clinical categories, fair overall agreement was demonstrated. However, when MA and GN categories were combined into an SUBC category, the overall agreement for the 3 clinicians improved slightly for both horizontal SP (HSP) and vertical SP (VSP). This pattern of agreement did not differ considerably when comparing HSP versus VSP, and good consistency and reliability was observed across clinicians. Again, inter-rater consistency was smaller for VSP versus HSP despite the reduction in clinical categories. Cut-off values were generated for the PS metric and demonstrated good specificity and sensitivity when they were exceeded for 2 or more frequencies in a particular plane when evaluating a subject’s SP test. Conclusions

General & Internal Medicine

EVALUATION OF HRA METHODOLOGIES FOR APPLICATION IN SDP WORK

This study critically evaluates human reliability analysis (HRA) methodologies applicable to regulatory probabilistic safety assessment (PSA) model, with a particular focus on their role in supporting the significance determination process (SDP) in nuclear safety assessment. Firstly, three widely utilized HRA methods – IDHEAS-ECA, SPAR-H, and ASEP/THERP – were qualitatively and quantitatively assessed. Qualitative assessments were conducted using attributes from the NEA/CSNI/R(2015)1 report, while quantitative evaluations employed regression and correlation analyses to compare predicted human error probabilities (HEPs) against empirical data. Results reveal distinct strengths, for example, IDHEAS-ECA’s robust predictive accuracy and K-HRA’s alignment with operational practices. In addition, dependency analysis and recovery analysis were critically evaluated. For dependency analysis, the methods’ handling of inter-task dependencies and their impact on HEPs were examined, while recovery analysis highlighted strategies for mitigating failure events. Furthermore, strategies were proposed to evaluate performance-shaping factors under conditions of reduced human performance, such as stress, fatigue, or cognitive overload, addressing specific challenges faced in SDP evaluations. Human errors from KINS’s operational performance information system event reports were evaluated as a case study. This study identifies gaps and provides actionable insights to ensure their validity and applicability in SDP HRA applications. This paper is a part of research conducted by KINS, and it should be noted that this result does not represent the regulatory position of KINS.

99 - GENERAL AND MISCELLANEOUS

Validating automated resonance evaluation with synthetic data

The integrity and precision of nuclear data are crucial for a broad spectrum of applications, from national security and nuclear reactor design to medical diagnostics, where the associated uncertainties can significantly impact outcomes. A substantial portion of uncertainty in nuclear data originates from the subjective biases in the evaluation process, a crucial phase in the nuclear data production pipeline. Recent advancements indicate that automation of certain routines can mitigate these biases, thereby standardizing the evaluation process and enhancing reproducibility. This research aims to provide a methodology, framework, and metrics for the validation of automated nuclear data evaluation software leveraging high-quality synthetic data that closely mimic real experimental observables. An introduced error metric provides a scale and intuitive measure of the evaluation quality by quantifying the estimate’s accuracy and performance across the specified energy range. Synthetic data provides access to experimental observables and underlying resonance parameters, enabling comparison of different evaluations. The methodology is demonstrated using Ta-181 isotope data in the resolved resonance region. The Automated Resonance Identification Subroutine (ARIS), which operates without prior resonance information, was used to test and showcase the framework’s capabilities utilizing the proposed error metrics. The results demonstrate the effectiveness of the proposed approach and framework for optimizing software parameters and testing hypotheses through “what-if” controlled experiments, such as modifying assumptions about experimental conditions or average resonance parameters.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Heuristic Evaluation Methods Applied to a Predictive Maintenance Chatbot

The need for an accessible iterative approach for evaluating prospective artificial intelligence (AI)/ML based technologies in the nuclear industry is needed, given the nature of algorithms and rapid advancements. This paper explores existing heuristic design principles for user-centered design and evaluates them based on their relevancy and usefulness for evaluating AI/ ML based technologies. Researchers at the Idaho National Laboratory (INL) have developed a machine learning software application called VIsualization for PrEdictive maintenance Recommendation (VIPER), which is used to help users understand and engage with the tool to learn more about work orders, data used, predictive maintenance, and machine learning (ML) algorithms. Early user research studies used to access VIPER’s technology readiness level have occurred; however, there is room for further improvement of the software through heuristic evaluations along with other methods and user testing. This work describes the applicability of heuristic evaluation methods and cognitive walkthroughs to help ensure human readiness for prospective AI/ ML based applications, using VIPER as a candidate use case. This work supports industry in ensuring that prospective AI/ML based technologies are usable and useful for plant personnel at nuclear power plants, ultimately leading to their safe, reliable, and efficient use.

99 - GENERAL AND MISCELLANEOUS

The Technical, Economic, Risk, and Adoption Assessment for Evaluating Work Reduction Opportunities in the Nuclear Industry

Automation and cost-saving initiatives, such as process automation with advances in artificial intelligence, are gaining traction in modernization efforts across the nuclear industry. As these innovations are increasingly adopted, it becomes crucial to evaluate their impacts comprehensively. Various technical and economic attributes, along with risk and human readiness factors, must be achieved to ensure that innovative projects enabling automation and modernization are successful. However, no systematic or integrated framework exists that allows plants to evaluate these innovative projects. To address this gap, the Technical, Economic, Risk, and Adoption (TERA) assessment offers a structured method to evaluate innovative technologies, ensuring solutions meet both operational and safety standards. The TERA framework integrates the disparate perspectives to assess modernization opportunities in nuclear operations. It combines qualitative and quantitative models to evaluate the relationship between performance and business impacts, while also enabling continuous re-evaluation during project development. This approach helps plant owners identify high-priority opportunities, optimize cost savings, and minimize risks, ensuring projects remain on track to achieve desired returns. By providing a comprehensive, systematic methodology, TERA enables informed, data-driven decisions that support successful modernization efforts, enhancing efficiency, safety, and cost savings across nuclear operations. This paper explains the TERA framework and its benefits for the nuclear industry.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Heuristic Evaluation Methods Applied to a Predictive Maintenance Chatbot

The need for an accessible iterative approach for evaluating prospective artificial intelligence (AI)/ML based technologies in the nuclear industry is needed, given the nature of algorithms and rapid advancements. This paper explores existing heuristic design principles for user-centered design and evaluates them based on their relevancy and usefulness for evaluating AI/ ML based technologies. Researchers at the Idaho National Laboratory (INL) have developed a machine learning software application called VIsualization for PrEdictive maintenance Recommendation (VIPER), which is used to help users understand and engage with the tool to learn more about work orders, data used, predictive maintenance, and machine learning (ML) algorithms. Early user research studies used to access VIPER?s technology readiness level have occurred; however, there is room for further improvement of the software through heuristic evaluations along with other methods and user testing. This work describes the applicability of heuristic evaluation methods and cognitive walkthroughs to help ensure human readiness for prospective AI/ ML based applications, using VIPER as a candidate use case. This work supports industry in ensuring that prospective AI/ML based technologies are usable and useful for plant personnel at nuclear power plants, ultimately leading to their safe, reliable, and efficient use. PowerPoint for conference that was reviewed in PRS and LRS PRS/CON-25-05379 and INL/CON-25-82946

99 - GENERAL AND MISCELLANEOUS

A Benchmarking Framework for Evaluating Large Language Model Capabilities in Nuclear Reactor Safety Applications

Large language models (LLMs) are increasingly capable of answering technical questions, synthesizing domain knowledge, and supporting engineering workflows. For nuclear science and engineering, these capabilities require careful, domain-specific evaluation before they can be credibly incorporated into safety-related activities, regulatory review, or technical decision support. This paper presents preliminary results from benchmarking framework for evaluating LLM capabilities in nuclear contexts. The framework is organized into three evaluation categories: nuclear fundamentals, general dual-use knowledge, and plant specific knowledge. These categories are intended to distinguish general nuclear engineering competence from broader technical reasoning and more context-dependent nuclear knowledge. Initial evaluations focus on nuclear fundamentals using questions representative of the knowledge expected of a nuclear professional engineer. Results indicate that contemporary frontier models perform at a high level and substantially exceed the performance of older model generations, with some models approaching saturation of the current benchmark. These findings suggest both the rapid improvement of LLM capabilities in specialized technical domains and the need for more discriminating evaluation methods. The paper presents the benchmark structure, preliminary model-comparison results, and ongoing work. This work supports development of verifiable, responsible, and safety-conscious methods for assessing AI systems in nuclear engineering applications.

21 - SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLAN

Applying queueing theory to evaluate wait-time-savings of triage algorithms

Abstract In the past decade, artificial intelligence (AI) algorithms have made promising impacts in many areas of healthcare. One application is AI-enabled prioritization software known as computer-aided triage and notification (CADt). This type of software as a medical device is intended to prioritize reviews of radiological images with time-sensitive findings, thus shortening the waiting time for patients with these findings. While many CADt devices have been deployed into clinical workflows and have been shown to improve patient treatment and clinical outcomes, quantitative methods to evaluate the wait-time-savings from their deployment are not yet available. In this paper, we apply queueing theory methods to evaluate the wait-time-savings of a CADt by calculating the average waiting time per patient image without and with a CADt device being deployed. We study two workflow models with one or multiple radiologists (servers) for a range of AI diagnostic performances, radiologist’s reading rates, and patient image (customer) arrival rates. To evaluate the time-saving performance of a CADt, we use the difference in the mean waiting time between the diseased patient images in the with-CADt scenario and that in the without-CADt scenario as our performance metric. As part of this effort, we have developed and also share a software tool to simulate the radiology workflow around medical image interpretation, to verify theoretical results, and to provide confidence intervals for the performance metric we defined. We show quantitatively that a CADt triage device is more effective in a busy, short-staffed reading setting, which is consistent with our clinical intuition and simulation results. Although this work is motivated by the need for evaluating CADt devices, the evaluation methodology presented in this paper can be applied to assess the time-saving performance of other types of algorithms that prioritize a subset of customers based on binary outputs.

Thompson, Yee Lam Elim (ORCID:0000000196537707)