Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “performance data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

A Qualitative Strategy for Fusion of Physics into Empirical Models for Process Anomaly Detection

To facilitate the automated online monitoring of power plants, a systematic and qualitative strategy for anomaly detection is presented. This strategy is essential to provide credible reasoning on why and when an empirical versus hybrid (i.e., physics-supported) approach should be used and to determine the ideal mix of these two approaches for a defined anomaly detection scope. Empirical methods are usually based on pattern, statistical, and causal inference. Hybrid methods include the use of physics models to train and test data methods, reduce data dimensionality, reduce data-model complexity, augment data, and reduce empirical uncertainty; hybrid methods also include the use of data to tune physics models. The presented strategy is driven by key decision points related to data relevance, simple modeling feasibility, data inference, physics-modeling value, data dimensionality, physics knowledge, method of validation, performance, data availability, and suitability for training and testing, cause-effect, entropy inference, and model fitting. The strategy is demonstrated through a pilot use case for the application of anomaly detection to capture a valve packing leak at the high-pressure coolant injection system of a nuclear power plant.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Geospatial Data Platform for All

Spatiotemporal data has evolved in scale due to augmented use in cross-domain applications. Simultaneously, there is substantial growth in the availability of Geographic Information Systems (GIS) data provided by the United States Geological Survey (USGS) along with other federal, state, county, or local agencies through open-data portals and public access APIs. However, data availability does not equate with accessibility. Large-scale analyses and applications require robust, performant data management with co-location of data storage and computing. The insufficiency of data management infrastructure compels researchers to adopt ad hoc project- specific GIS data storage solutions (e.g., copying data to High-Performance computer file systems). As an ad hoc storage strategy does not scale, it hampers cross-domain analyses causing difficulty in data reuse and utilizing existing code bases. Furthermore, GIS data is complex and requires expertise to analyze and manipulate due to its intricate data structures and data-specific projection transformations. Despite the challenges, we recognize that derived GIS data products, e.g., satellite or LIDAR-based images, can be used in downstream applications such as AI by domain, but non-GIS experts. To address the data needs and overcome the challenges, we are working towards a GIS Data Platform focused on efficient data storage, data discovery and access, and an API to enable common workflows. We propose a knowledge-graph (KG) approach for data discovery, whereby datasets are semantically linked to higher- level constructs such as projects and research areas. The semantic data links enable researchers to explore datasets in a top-down approach by specifying relevant and meaningful terms (assists in finding hidden data). An advantage is that the nodes and edges in a knowledge graph create built-in semantic documentation. Deeper spatiotemporal connections between data sources can be encoded via Graph Neural Networks (GNN) (Zhang et al., 2021). The KG approach can be extended to integrate the data itself in a Virtual KG (VKG). Our work will derive inspiration from large-scale VKG efforts that have been undertaken or are currently underway as part of the OpenStreetMap project (Ding et al., 2021). For DOE Data Days, we share the proposed geospatial data platform hybrid (cloud/on-prem) architecture, our work-to-date on storing, retrieving, and transforming LiDAR and raster data relevant to two important NREL use-cases, including the Renewable Energy Potential (reV) Model, and present our proposal for a KG based data discovery engine.

data platform↗

Photovoltaic Data Acquisition (PVDAQ) Public Datasets

The NREL PVDAQ is a large-scale time-series database containing system metadata and performance data from a variety of experimental PV sites and commercial public PV sites. The datasets are used to perform on-going performance and degradation analysis. Some of the sets can exhibit common elements that effect PV performance (e.g. soiling). The dataset consists of a series of files devoted to each of the systems and an associated set of metadata information that explains details about the system hardware and the site geo-location. Some system datasets also include environmental sensors that cover irradiance, temperatures, wind speeds, and precipitation at the site.

Array↗

Assessing high fidelity multi-component models to facilitate safeguards at Gas Centrifuge Enrichment Plants

We report that the International Atomic Energy Agency (IAEA) inspectors routinely carry out environmental sampling (ES) as a verification method. Collection of environmental swipe samples at various locations in Gas Centrifuge Enrichment Facilities (GCEPs) is an important process in detecting misuse of a declared facility and possibly the existence of undeclared nuclear material. These samples are measured for isotopic composition in uranium containing particles by Thermal Ionization Mass Spectrometry (TIMS) or Inductively Coupled Plasma Mass Spectrometry (ICP-MS). Even though ES is highly effective in detecting the absolute value of enrichments and their deviations from the declared values, it cannot explain the cause of those changes. Several potential explanations can serve as possibilities for particles detected above or other than the declared enrichment. These include normal and non-malicious events such as the design of the enrichment cascades, unintentional failures of the machines, or overshoot during startup of a cascade. It can also be the result of deliberate misuse by the facility operators. The primary objective of this work is to understand how these factors affect the enrichments produced by a cascade and quantify anticipated multi-isotopic concentrations for each case. The following methodology is employed to determine signatures at a particular facility. 1) Utilize a new two-dimensional multi-component diffusion code to obtain centrifuge performance data and use that information to design and perform cascade analysis. Compare and contrast the results with previous 1-D radially averaged solutions from the Pancake code. 2) Design a GCEP cascade with the production goal of 19.75% 235 U for each set of machine data above. Investigate two cascade scenarios that include enrichment of natural uranium (NU) feed to 19.75% 235U in a single cascade compared to a two-step process of NU to 5% and then 5% to 19.75%. 3) Simulate the intentional vs. unintentional off-normal scenarios in the cascades to assess the differences in isotopic concentrations. A non-ideal squared-off cascade model developed at the University of Virginia is used to calculate flow rates and isotopic concentrations of the process gas. The analysis is performed using the Rome machine model operated at 600 m/s rotor speed. The upper and lower bounds of normal and abnormal enrichments in a typical facility are used in conjunction with ES results to understand the root causes of such observations.

98 NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL P↗

Evaluation of Saccadic Component Measure on Smooth Pursuit Tests

ABSTRACT Introduction Despite the advancement of eye-tracking technology for smooth pursuit (SP) eye movement evaluation, qualitative observation offers much information that is not captured by computers; hence, both objective and qualitative information should be utilized to evaluate SP. This study examined the consistency among our clinicians when evaluating SP using normal (N), grossly normal (GN), mildly abnormal (MA), and abnormal (AB) as classifications. We then evaluated the effect of combining GN and MA into a single subclinical (SUBC) category. We also evaluated the computerized percent saccade (PS) metric by determining its sensitivity and specificity in classifying SP. Materials and Methods Retrospective horizontal and vertical SP test videos and numerical data for 70 participants were obtained from the Neuro Kinetics Neuro-Otologic Test Center and de-identified. From this, eye-tracking videos, time plots of eye-tracking positional data, and tables of SP eye-tracking performance data were generated for 0.1, 0.3, and 0.5 Hz in both horizontal and vertical planes, totaling 6 tests per subject. Three clinicians rated each subject’s SP performance as N, GN, MA, or AB for a total of 6 ratings (3 frequencies, horizontal and vertical). This process was repeated using N, SUBC, and AB as rating categories. Clinicians also provided an overall SP rating for each plane as follows: AB if the results were abnormal for 2 or more frequencies tested. Alternatively, if fewer than 2 frequencies presented with a rating of AB, then an overall rating of MA, GN, or N was determined at the respective clinician’s discretion. Results When the 3 clinicians were tasked with classifying SP videos using 4 clinical categories, fair overall agreement was demonstrated. However, when MA and GN categories were combined into an SUBC category, the overall agreement for the 3 clinicians improved slightly for both horizontal SP (HSP) and vertical SP (VSP). This pattern of agreement did not differ considerably when comparing HSP versus VSP, and good consistency and reliability was observed across clinicians. Again, inter-rater consistency was smaller for VSP versus HSP despite the reduction in clinical categories. Cut-off values were generated for the PS metric and demonstrated good specificity and sensitivity when they were exceeded for 2 or more frequencies in a particular plane when evaluating a subject’s SP test. Conclusions

General & Internal Medicine↗

POWER DATA PIPELINE

SF-25-081 Utility software for creating high-performance data pipelines to extract, load, and transform raw electric power systems measurements. For use with anomaly detection models training workflows. The software supports the project: Adaptive Cybersecurity for DER: A Game-Theoretic and Machine Learning approach for Real-Time Threat Detection and Mitigation

Plathottam, Silby Jose [Argonne National Laborator↗

Reading Between the Lines: Measuring the Effects of Linguistic-Based Indicators of Deception on Experts’ Identification and Categorization of Disinformation

There is currently very limited research into how experts analyze and assess potentially fraudulent content in their expertise areas, and most research within the disinformation space involves very limited text samples (e.g., news headlines). The overarching goal of the present study was to explore how an individual’s psychological profile and the linguistic features in text might influence an expert’s ability to discern disinformation/fraudulent content in academic journal articles. At a high level, the current design tasked experts with reading journal articles from their area of expertise and indicating if they thought an article was deceptive or not. Half the articles they read were journal papers that had been retracted due to academic fraud. Demographic and psychological inventory data collected on the participants was combined with performance data to generate insights about individual expert susceptibility to deception. Our data show that our population of experts were unable to reliably detect deception in formal technical writing. Several psychological dimensions such as comfort with uncertainty and intellectual humility may provide some protection against deception. This work informs our understanding of expert susceptibility to potentially fraudulent content within official, technical information and can be used to inform future mitigative efforts and provide a building block for future disinformation work.

99 GENERAL AND MISCELLANEOUS↗

ComStock Measure Scenario Documentation: Laboratory-Informed Modeling of Standard Performance Heat Pump Rooftop Units

This measure scenario replaces gas and electric resistance RTUs in the U.S. commercial building stock with standard efficiency commercial off the shelf heat pump rooftop units. This study uses performance data informed by NREL laboratory testing of a standard efficiency 7.5-ton heat pump RTU. This is the key distinction between this measure scenario and a similar ComStock measure scenario - Standard Performance Heat Pump Rooftop Units - that uses published manufacturer data tables to inform performance. These two scenarios are compared in this report.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Adaptive elasticity policies for staging-based in situ visualization

In situ processing aims to alleviate the growing gap between computation and I/O capabilities by performing data processing close to the data source. In situ processing is widely used to process data generated by multiple data sources, including observation data from edge devices or scientific observational facilities and the simulation data generated by scientific computation on a high-performance computing (HPC) platform. For a scientific workflow that is run on an HPC platform and composed of a simulation program and an in situ data analytics or visualization (abbreviated as ana/vis) task, there is an implicit assumption that the computing resources assigned to the workflow keep static during the workflow execution. However, with the converging trend between the HPC and cloud computing platform, running the in situ ana/vis task in an elastic way is promising to decrease its overhead and improve its resource utilization rate. Resource elasticity represents the ability to change resource configurations such as the number of computing nodes/processes during workflow execution. An elastic job may dynamically adjust resource configurations; it may use a few resources at the beginning and more resources toward the end of the job when interesting data appear. However, it is hard to predict a priori how many computing nodes/processes need to be added/removed during the workflow execution to adapt to changing workflow needs. How to efficiently guide elasticity operations, such as growing or shrinking the number of processes used for in situ analysis during workflow execution, is an open-ended research question. In this article, we present adaptive elasticity policies that adopt workflow runtime information collected during workflow execution to predict how to trigger the addition/removal of processes in order to minimize in situ processing overhead. Taking in situ visualization tasks as an example, we integrate the presented elasticity policies into a staging-based elastic workflow and evaluate its efficiency in multiple elasticity scenarios. Compared with the situation without elasticity or with a static elasticity policy that uses a fixed number of processes for each rescaling operation, the adaptive elasticity policy can save overhead in finding a proper resource configuration and improve resource utilization efficiency. Furthermore, one experiment illustrates that the adaptive elasticity policy saves 41% of core-hours compared with the situation without the resource elasticity.

97 MATHEMATICS AND COMPUTING↗

Availability and Performance Loss Factors for U.S. PV Fleet Systems

In the PV Fleet Performance Data Initiative, we partner with photovoltaic (PV) fleet owners to collect time-series PV production data and publish aggregated, anonymized results. This report is an update of our previous publications, specifically a FY 2021 performance index publication and a FY 2022 fleet degradation analysis. In this analysis, we have increased our data participants and system totals by around 10% to 8.5 GW and 24,000 separate inverter data channels. Four major analysis topics are considered in this report: Performance Index (PI) trends, PV system availability, soiling losses, and PV system degradation. Performance Index and inverter availability are assessed on a larger set of data from our FY 2021 report: 1,128 systems compared with 200 systems from before. The increased number of systems is due to an improved data quality methodology, as well as introducing new systems to the analysis. Overall results are similar to previously published values - overall inverter availability is low in the first six months of system performance before reaching steady-state by the end of the first year. Excluding this six-month startup period, system-level aggregated data shows a median (P50) system availability of 0.99 and a lower 10th percentile (P90) value of 0.95 (Figure ES-1). A dependence on system size is also demonstrated, with worse inverter availability results for larger PV systems. Causes of this effect are under investigation, but may be impacted by inverter size, which also show lower availability for larger inverter sizes. This report also investigates PI, correcting for degradation, soiling, snow, and availability. Following these corrections, the median system PI over its entire lifetime is 0.95. PI values reported here are approximately 3% lower than what we presented in our previous FY 2021 report. Soiling loss is assessed in a comprehensive way for the first time in this report. Results are presented using the COmbined Degradation and Soiling (CODS) method, as implemented in RdTools (v3.0.0a4). Soiling values are presented for 255 systems, which indicated irradiance-weighted soiling loss greater than 1%. The values have been published in an updated NREL soiling map at nrel.gov/pv/soiling.html. Finally, we investigated system degradation using three different data analysis techniques: conventional RdTools (year-on-year (YOY)), CODS, and Performance Loss Rate (PLR) analysis. Overall degradation results are consistent with our previous publications. Rerunning conventional RdTools on our updated fleet shows that some data partners have systematically fallen below the median system degradation rate (change over time) of -0.75 %/year. A comparison with PLR analysis, which looks at change in annual PI over time, shows that median system degradation is consistent with -0.5% to -0.75% per year change. However, at the P90 value, system degradation is substantially faster. These two results are consistent and indicate that resulting degradation statistics depend to a great degree on the population of PV systems making up the analysis cohort and whether soiling impacts the systems. The use of CODS for degradation analysis provides a different method for degradation assessment, which explicitly excludes the impact of recoverable soiling on degradation analysis. Excluding soiling effects yields an annual system degradation around -0.5% per year on average. This indicates that a portion of system performance loss may be attributed to periodic soiling that is not fully recovered. This report provides PV system owners/operators with background and methods to analyze PV system performance, give guidance for expected cohort performance, and performance loss values for use in pro-forma financial models, which guide new-build system design and bankability reports.

14 SOLAR ENERGY↗

Relating Aerial Infrared Thermography Defects to Photovoltaic Performance: Preprint

In this research, we examine the relationship between aerial IR defect analysis and photovoltaic (PV) performance data for twelve utility- and commercial-scale solar sites in the United States. To do this, we fuse the site diagram geoJSON's, aerial infrared thermography (aIRT) defect analyses, and associated inverter time series, allowing for a direct comparison between site defects and time series data. Defect analyses were provided by Zeitview, under its Solar Insights platform. Following the data fusion process, we look at the relationship between system performance and aIRT defects. We investigate the relationship between degradation and hotspot defects, as well as the relationship between AC power data and offline strings and misaligned modules. In general, system degradation was not affected by long-term or balance-of-system (BoS) defects as they occurred infrequently in the data set. However, for one system, a near statistically significant relationship (p-value=0.057) was found when comparing the degradation of inverter blocks with several multi-hotspot defects to all other inverter blocks without this particular defect. There was strong alignment when comparing short-term recoverable module defects such as stuck trackers and offline strings to time series data. In general, we found that when an inverter block has more than 80% of modules flagged for one of these defects, its AC power time data is flat-lined and the inverter block is not producing.

aerial inspection↗

Visual Analytics of Performance of Quantum Computing Systems and Circuit Optimization

Driven by potential exponential speedups in business, security, and scientific scenarios, interest in quantum computing is surging. This interest feeds the development of quantum computing hardware, but several challenges arise in optimizing application performance for hardware metrics (e.g., qubit coherence and gate fidelity). In this work, we describe a visual analytics approach for analyzing the performance properties of quantum devices and quantum circuit optimization. Our approach allows users to explore spatial and temporal patterns in quantum device performance data and it computes similarities and variances in key performance metrics. Detailed analysis of the error properties characterizing individual qubits is also supported. We also describe a method for visualizing the optimization of quantum circuits. The resulting visualization tool allows researchers to design more efficient quantum algorithms and applications by increasing the interpretability of quantum computations.

Chae, Junghoon↗

Improving the Accessibility and Usability of Geothermal Information with Data Lakes and Data Pipelines on the Geothermal Data Repository: Preprint

The Geothermal Data Repository (GDR) provides universal access to data and information resulting from research and development activities funded by the Department of Energy (DOE). The GDR has extended this universal access to big data through integration with data lakes developed by the Open Energy Data Initiative (OEDI). Previously, large datasets such as seismic waveform or distributed acoustic sensing (DAS) data could only be accessed by institutions with high performance data storage and compute capabilities, effectively limiting the accessibility of big data to national labs, larger universities, and major corporations. Moreover, the time and resources needed to transport big data and configure them can produce additional barriers to use. Many of the standard formats used for structured data models (also known as content models) are incapable of handling big data and can introduce additional usability problems, often requiring data to be reformatted prior to use. This paper will explore how recent integrations between the GDR and the OEDI data lake have improved the accessibility and usability of geothermal data in a big way, making the data available to a broader audience, and enabling collaborative analysis and innovation across the greater geothermal industry.

access↗

Improving the Accessibility and Usability of Geothermal Information with Data Lakes and Data Pipelines on the Geothermal Data Repository

The Geothermal Data Repository (GDR) provides universal access to data and information resulting from research and development activities funded by the Department of Energy (DOE). The GDR has extended this universal access to big data through integration with data lakes developed by the Open Energy Data Initiative (OEDI). Previously, large datasets such as seismic waveform or distributed acoustic sensing (DAS) data could only be accessed by institutions with high performance data storage and compute capabilities, effectively limiting the accessibility of big data to national labs, larger universities, and major corporations. Moreover, the time and resources needed to transport big data and configure them can produce additional barriers to use. Many of the standard formats used for structured data models (also known as content models) are incapable of handling big data and can introduce additional usability problems, often requiring data to be reformatted prior to use. This paper will explore how recent integrations between the GDR and the OEDI data lake have improved the accessibility and usability of geothermal data in a big way, making the data available to a broader audience, and enabling collaborative analysis and innovation across the greater geothermal industry.

access↗

ROADRUNNER uranium nitride MiniFuel: Experimental design, fabrication and pre-irradiation baseline characterization for accelerated burnup testing

Uranium nitride (UN) is a promising fuel candidate for advanced reactor systems owing to its high uranium density and thermal conductivity; however, its qualification remains constrained by the scarcity of well-controlled irradiation performance data. Here, to address this limitation, the ROADRUNNER (Research On ADvancing the peRformance of UraNium Nitrides in Extreme enviRonments) campaign employs the MiniFuel platform in the High Flux Isotope Reactor (HFIR) to enable accelerated burnup irradiation testing under tightly controlled and largely isothermal conditions. This paper presents the experimental design, fuel fabrication, and pre-irradiation baseline characterization of the ROADRUNNER UN MiniFuel campaign. Thirty-six UN minidisc specimens were fabricated with systematically varied as-fabricated density (86–96% of theoretical density), carbon impurity content (961–5240 ppm), oxygen content (≤ ∼2000 ppm), and grain size (2.5–24 μm). The irradiation matrix spans nominal fuel temperatures of 873 K, 1173 K, and 1473 K and target burnups of 3.75%, 6.0%, and 7.5% fissions per initial metal atom (FIMA). Neutronic and thermal analyses were performed to define specimen-specific burnup accumulation and temperature histories, establishing the boundary conditions for subsequent in-pile behavior. Comprehensive pre-irradiation characterization—including dimensional metrology, density verification, impurity analysis, X-ray diffraction, Raman spectroscopy, scanning electron microscopy, X-ray computed tomography, and confocal profilometry—provides a detailed baseline for post-irradiation examination. Pre-irradiation data were further used to generate predictive estimates of fission gas release and swelling using existing empirical correlations. This quantitative comparison reveals substantial inter-model divergence at intermediate and elevated temperatures that exceeds propagated input uncertainties, highlighting structural gaps in the historical irradiation database. The ROADRUNNER irradiation campaign is currently underway in HFIR, with initial firs cycle completed in late 2025 and remaining targets scheduled through 2027. The experimental design and baseline dataset presented here establish the framework needed to interpret forthcoming post-irradiation measurements and to provide discriminating data for the validation and refinement of physics-based UN fuel performance models.

Lopes, Denise Adorno [Oak Ridge National Laborator↗

PV Reliability and Resilience in Challenging Climates

Challenging climates for Photovoltaics are usually based on climate classification. However, extreme weather events such as high wind, flooding, large hail, extreme snow etc. have become more ubiquitous globally. To study the impact of extraordinary weather events on PV reliability we used two of the largest databases in the USA. First, the National Oceanic and Atmospheric Administration (NOAA) database on extreme weather and secondly, the PV Fleet Data Initiative where we have collected high-resolution PV performance data of more than 8 gigawatts or about 6-7% of all commercial and utility systems in the USA. We analyzed almost 200 systems between 2008-20022 that were immediately impacted by these weather events. The immediate impact (outages) was determined to be about 1% of or a median of approximately 3 days of annual lost production. However, the risk these events pose is exemplified by a long tail where 0.4 % of all systems lost more than 2 weeks annual production. We also found a threshold for high wind (90 km/hr) and hail (25mm), above which we observed significantly higher degradation implying long-term damage to the systems. In addition, we are using satellite imagery to quantify visible damage to PV plants. Finally, we share module, design and installation lessons from some observed case studies to improve extreme weather resilience for PV power systems.

degradation↗

Multiresolution classification of turbulence features in image data through machine learning

During large-scale simulations, intermediate data products such as image databases have become popular due to their low relative storage cost and fast in-situ analysis. Serving as a form of data reduction, these image databases have become more acceptable to perform data analysis on. In this work, we present an image-space detection and classification system for extracting vortices at multiple scales through wavelet-based filtering. A custom image-space descriptor is used to encode a large variety of vortex-types and a machine learning system is trained for fast classification of vortex regions. By combining a radial-based histogram descriptor, a bag of visual words feature descriptor, and a support vector machine, our results show that we are able to detect and classify vortex features at various sizes at multiple scales. Once trained, our framework enables the fast extraction of vortices on new, unknown image datasets for flow analysis.

97 MATHEMATICS AND COMPUTING↗