Engineering PapersSearch

SEARCH · Engineering Papers

Results for “validation data sets”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

WUS324: Multiscale Full Waveform Inversion Approaching Convergence Improves Waveform Fits While Imaging Seismic Structure of the Western United States

Abstract We report a new model of radially anisotropic crustal and upper mantle structure of the western United States (WUS324) obtained from full waveform inversion of earthquake data. We ran three multiscale inversion stages beyond model WUS256 (Rodgers et al., 2022, https://doi.org/10.1029/2022jb024549 ) allowing them to approach convergence to fit a larger data set to a shorter minimum period of 16 s. WUS324 is based on 324 total iterations from its starting model, significantly more (16 times) than previous studies. Waveform misfit reductions are 66%–70% for both the inversion data and an independent validation data set providing confidence in the predictive power of the model. WUS324 provides much better fits and reveals shear wavespeed, v S , structure of this large region with more detail than previous waveform tomography models. We show representative images demonstrating the resolution of diverse seismic structure across this highly heterogeneous region including oceanic lithosphere, subducting slabs and continental magmatism.

58 GEOSCIENCES

Numerical Simulation of a Natural Convection–Driven Air-Cooled Reactor Cavity Cooling System Experiment

Ensuring the efficient removal of decay heat from the reactor vessel is essential for the safety of advanced reactor technologies. Several Generation-IV concepts incorporate variations in the reactor vessel cooling systems to achieve this objective. High-temperature gas-cooled reactors utilize a reactor cavity cooling system (RCCS), a passive ex-vessel system designed to operate without active components or external power during accident conditions. The RCCS removes decay heat primarily through radiative and convective heat transfer mechanisms. Here, this study presents a comprehensive validation of a computational fluid dynamics Reynolds-averaged Navier-Stokes model for the University of Wisconsin-Madison air-cooled RCCS facility. Validation was conducted for both high- and low-power natural convection cases under a uniform heating profile. Near-wall resolution was found to be critical for accurately modeling natural convection in the RCCS; employing an all-𝑦 + wall treatment resulted in wall temperature discrepancies exceeding 50 °⁢𝐶 compared to a wall-resolved mesh. Thermal-hydraulic behaviors under natural and forced convection conditions were compared within the heated cavity and RCCS. A turbulence model sensitivity analysis indicated that low-Reynolds number k-ɛ, k-ω shear stress transport (SST), and Reynolds stress transport models produce similar wall temperature predictions. A buoyancy modeling sensitivity study revealed that the Boussinesq approximation significantly underpredicted thermal-hydraulic behavior in the RCCS. Based on these findings, modeling recommendations are provided. The validated data set along with identified sensitivities refine the modeling of natural convection in the RCCS. The information produced by this study supports RCCS design, optimization, and safety evaluations, enabling the calibration and verification of reduced-order thermal-hydraulic models.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Factors That Influence Variability in Stress-Drop Measurements Using Spectral Decomposition and Spectral-Ratio Methods for the 2019 Ridgecrest Earthquake Sequence

Stress drop is a fundamental parameter related to earthquake source physics, but is hard to measure accurately. To better understand how different factors influence stress-drop measurements, we compare two different methods using the Ridgecrest stress-drop validation data set: spectral decomposition (SD) and spectral ratio (SR), each with different processing options. Here, we also examine the influence of spectral complexity on source parameter measurement. Applying the SD method, we find that frequency bandwidth and time-window length could influence spectral magnitude calibration, while depth-dependent attenuation is important to correctly map stress-drop variations. For the SR method, we find that the selected source model has limited influence on the measurements; however, the Boatwright model tends to produce smaller standard deviation and larger magnitude dependence than the Brune model. Variance reduction threshold, frequency bandwidth, and time-window length, if chosen within an appropriate parameter range, have limited influence on source parameter measurement. For both methods, wave type, attenuation correction, and spectral complexity strongly influence the result. The scale factor that quantifies the magnitude dependence of stress drop show large variations with different processing options, and earthquakes with complex source spectra deviating from the Brune-type source models tend to have larger scale factor than earthquakes without complexity. Based on these detailed comparisons, we make a few specific suggestions for data processing workflows that could help future studies of source parameters and interpretations.

58 GEOSCIENCES

AK112: Full Waveform Inversion Tomography of Alaska Improves Waveform Fits While Imaging Crustal, Mantle, and Slab Structure

We report a full waveform inversion tomography model of Alaska and the surrounding regions, inferring radially anisotropic shear and isotropic compressional wavespeeds by fitting complete waveforms from 120 regional earthquakes. Our multiscale approach inverted time–frequency phase misfits (maximum period of 100 s), starting with a minimum period of 40 s and ending at 20 s in 7 stages and 112 total iterations. The model (AK112) was evaluated by computing the misfits for 36 independent validation events. We find that misfit reductions were large and equal (∼55%) for both the inversion and validation data sets, providing confidence in the model. AK112 also provides much better waveform fits compared to other reported models for the region, including an isotropic version of itself, highlighting the importance of anisotropy. The model resolves known crustal, upper mantle, and slab structure to depths of 100 km with new detail: sedimentary basins in the Alaskan Shelf, Cook Inlet, and Colville basins, among others; discontinuous lithospheric structure across major terrane boundaries; and subducting slab geometry and back‐arc volcanic sources. In addition to tectonic interpretations, the model enables full waveform simulations for long‐period earthquake ground motions and source characterization (e.g., moment tensor and finite‐fault inversion).

Rodgers, Arthur [Lawrence Livermore National Labor

Oak Ridge National Laboratory EAGLE-I TM : Modeling Electric Utility County Customers for Situational Awareness

During natural hazard events (hurricanes, wildfires, earthquakes, etc.) and recent man-made events (e.g., cyber attacks), the exchange of near real-time, spatially refined data within the response community is critical. The EAGLE-I$^{TM}$ platform is one tool that facilitates this data for decision makers within the energy sector. While much information can be collected and integrated into the system directly, other pertinent data must be augmented by other derived data products to enhance the information and allow for a consistent evaluation of on-the-ground conditions. One such data set that requires the addition of other derived data is the electric utility customer outage data that is aggregated to the county level within the EAGLE-I application. Without a county customer data set, outages can only be compared on total counts, which gives greater importance to higher population outages. Including an electric utility customer data set at the county level allows for these outage counts to be converted to percent outages and brings a consistent classification of outages and equal importance to all outages. To achieve this, several available data sets were combined and spatial disaggregation techniques were employed to model customer estimates at the county scale. This paper presents the approach to produce this data for the United States and lessons learned from working with these disparate data sets. Data validation is provided, where possible, and limitations of the model and possible improvements are discussed.

24 POWER TRANSMISSION AND DISTRIBUTION

Statistical Validation of Multiple Related Data Sets—Case Study Using Interstellar Boundary Explorer Satellite Data

Abstract Space scientists often face the question of whether data collected by different instruments are measurements of the same source population. This paper proposes a statistical validation method for evaluating the agreement between such related data sets. It offers a detailed case study focused on validating a new data set from the Interstellar Boundary Explorer (IBEX) mission, which serves as a practical how-to guide for similar analyses. Since 2008, the IBEX satellite has been gathering data on heliospheric energetic neutral atoms (ENAs) while being exposed to various sources of background noise, such as cosmic rays and solar energetic particles. The IBEX mission initially released only a qualified triple-coincidence (qABC) data product, which was designed to provide observations of ENAs free of background contamination. Further measurements revealed that the qABC data were in fact susceptible to contamination, having relatively low ENA counts and high background rates. To mitigate this issue, the mission team recently considered releasing a certain qualified double-coincidence (qBC) data product, which has roughly twice the detection rate of the qABC data product. This paper presents a simulation-based validation of the new qBC data product against the already-released qABC data product. The results show that the qBCs can plausibly be said to be measuring the same source population as the qABCs up to an average absolute deviation of 3.6%. Visual diagnostics provide additional confirmation of source rate coherence across data products. The framework introduced here is general and can be applied to other validation problems both within and outside the field of space physics.

79 ASTRONOMY AND ASTROPHYSICS

Automated Nanocrystal Synthesis: Lessons from 25 Years of Robots, Microfluidics, and Machine Learning

Here, this perspective highlights the evolution of techniques for automating the synthesis of colloidal nanocrystals. Over the past 25 years, microfluidic reactors and robotic workflows have been developed to enhance the reproducibility of nanocrystal synthesis, facilitate rapid screening of reaction conditions, optimize material properties, and perform multistep syntheses of high-quality nanoparticles with complex heterostructures. Modern automated systems are now valued for their ability to generate robust data sets for validating physical models, supporting chemical mechanisms, training machine learning models, and for directing autonomous experimentation. We discuss the early challenges and limitations of these technologies and present key lessons for effectively utilizing automated and ML-guided tools to accelerate nanocrystal discovery for the next 25 years.

Nanocrystals

Orbital-Radar v1.0.0: a tool to transform suborbital radar observations to synthetic EarthCARE cloud radar data

The Earth Cloud, Aerosol and Radiation Explorer (EarthCARE) satellite developed by the European Space Agency (ESA) and the Japan Aerospace Exploration Agency (JAXA) launched in May 2024 carries a novel 94 GHz cloud profiling radar (CPR) with Doppler capability. This work describes the open-source instrument simulator Orbital-Radar, which transforms high-resolution radar data from field observations or forward simulations of numerical models to CPR primary measurements and uncertainties. The transformation accounts for sampling geometry and surface effects. We demonstrate Orbital-Radar's ability to provide realistic CPR views of typical cloud and precipitation scenes. The presented case studies show small-scale convection, marine stratus clouds, and Arctic mixed-phase cloud cases. These results provide valuable insights into the capabilities and challenges of the EarthCARE CPR mission and its advantages over the CloudSat CPR. Finally, Orbital-Radar allows for evaluating kilometre-scale numerical weather prediction models with EarthCARE CPR observations. So, Orbital-Radar can generate calibration and validation (Cal/Val) data sets already pre-launch. Nevertheless, an evaluation of synthetic CPR output data to accurate EarthCARE CPR data is missing.

54 ENVIRONMENTAL SCIENCES

Data for A Hybrid Biophysical-Machine Learning Framework for Diurnal Surface Energy Flux Estimation Using Proximal Sensing

Thermal infrared-based remote sensing of surface energy fluxes has traditionally relied on high spatial resolution satellite data with revisit frequencies on the order of weeks. In this study, we evaluate a biophysics-based analytical surface energy balance model for predicting latent energy (LE) and sensible heat (H) fluxes using proximal sensing observations. The Surface Temperature Initiated Closure (STIC1.2) model has been extensively validated across a wide range of spatial and temporal scales using various satellite-derived thermal infrared data sets. Here we extend this validation by applying STIC at sub-hourly temporal resolution over multiple growing seasons for four distinct agricultural systems. We further develop and evaluate novel STIC variants that incorporate machine learning (ML) techniques to eliminate the need for surface energy balance observations, specifically net radiation and soil heat flux, thereby enhancing model applicability in data-sparse settings. The integration of a ML component to estimate surface available energy is shown to have strong predictive performance for both LE (R2 = 0.81–0.94) and H (R2 = 0.46–0.72) across all agricultural systems examined here, demonstrating the potential of hybrid biophysical-machine learning approaches for surface energy balance modeling with minimal data requirements. This study concludes with a novel application of explainable machine learning (exML) to diagnose sources of model error. This exML framework attributes residual prediction errors to both model input variables and environmental drivers not explicitly included in the simulation experiments. This approach provides a new pathway for improving model design and integrating previously overlooked yet influential variables into future model iterations.

AI/ML

A Lake Biogeochemistry Model for Global Methane Emissions: Model Development, Site‐Level Validation, and Global Applicability

Abstract Lakes are important sentinels of climate change and may contribute over 30% of natural methane (CH 4 ) emissions; however, no earth system model (ESM) has represented lake CH 4 dynamics. To fill this gap, we refined a process‐based lake biogeochemical model to simulate global lake CH 4 emissions, including representation of lake bathymetry, oxic methane production (OMP), the effect of water level on ebullition, new non‐linear CH 4 oxidation kinetics, and the coupling of sediment carbon pools with in‐lake primary production and terrigenous carbon loadings. We compiled a lake CH 4 data set for model validation. The model shows promising performance in capturing the seasonal and inter‐annual variabilities of CH 4 emissions at 10 representative lakes for different lake types and the variations in mean annual CH 4 emissions among 106 lakes across the globe. The model reproduces the variations of the observed surface CH 4 diffusion and ebullition along the gradients of lake latitude, depth, and surface area. The results suggest that OMP could play an important role in surface CH 4 diffusion, and its relative importance is higher in less productive and/or deeper lakes. The model performance is improved for capturing CH 4 outgassing events in non‐floodplain lakes and the seasonal variability of CH 4 ebullition in floodplain lakes by representing the effect of water level on ebullition. The model can be integrated into ESMs to constrain global lake CH 4 emissions and climate‐CH 4 feedback.

54 ENVIRONMENTAL SCIENCES

A database and meta-analysis on the performance of exploding pusher implosions conducted at OMEGA

A database of 222 exploding pusher implosions conducted at the OMEGA Laser Facility is presented. The dataset consists of glass-shell capsules filled with varying pressures of D 2 , T 2 , and 3 He, which were imploded using square laser pulses with intensities ranging from 1 to 1 × 10 15 W/cm 2 . The database includes measurements of bang times, ion temperatures, and yields from the DD, D 3 He, and DT fusion reactions. A semi-analytic exploding pusher model is introduced, which effectively captures the observed trends in the data. This model predicts that the measurements scale according to a power-law relation based on the initial capsule and laser conditions. A generalized power-law scaling relation is directly fit to each dataset, providing a useful interpolation of the entire database. Overall, the database provides a valuable resource to estimating bang times, temperatures, and yields for the design of future experiments. Additionally, it provides a diverse set of data for validating more advanced implosion physics models.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Archetype-based Redshift Estimation for the Dark Energy Spectroscopic Instrument Survey

We present a computationally efficient galaxy archetype-based redshift estimation and spectral classification method for the Dark Energy Survey Instrument (DESI) survey. The DESI survey currently relies on a redshift fitter and spectral classifier using a linear combination of principal component analysis–derived templates, which is very efficient in processing large volumes of DESI spectra within a short time frame. However, this method occasionally yields unphysical model fits for galaxies and fails to adequately absorb calibration errors that may still be occasionally visible in the reduced spectra. Our proposed approach improves upon this existing method by refitting the spectra with carefully generated physical galaxy archetypes combined with additional terms designed to absorb data reduction defects and provide more physical models to the DESI spectra. We test our method on an extensive data set derived from the survey validation (SV) and Year 1 (Y1) data of DESI. Our findings indicate that the new method delivers marginally better redshift success for SV tiles while reducing catastrophic redshift failure by 10%–30%. At the same time, results from millions of targets from the main survey show that our model has relatively higher redshift success and purity rates (0.5%–0.8% higher) for galaxy targets while having similar success for QSOs. These improvements also demonstrate that the main DESI redshift pipeline is generally robust. Additionally, it reduces the false-positive redshift estimation by 5%–40% for sky fibers. We also discuss the generic nature of our method and how it can be extended to other large spectroscopic surveys, along with possible future improvements.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

Genomic prediction of regional-scale performance in switchgrass ( Panicum virgatum ) by accounting for genotype-by-environment variation and yield surrogate traits

Switchgrass is a potential crop for bioenergy or carbon capture schemes, but further yield improvements through selective breeding are needed to encourage commercialization. To identify promising switchgrass germplasm for future breeding efforts, we conducted multisite and multitrait genomic prediction with a diversity panel of 630 genotypes from 4 switchgrass subpopulations (Gulf, Midwest, Coastal, and Texas), which were measured for spaced plant biomass yield across 10 sites. Our study focused on the use of genomic prediction to share information among traits and environments. Specifically, we evaluated the predictive ability of cross-validation (CV) schemes using only genetic data and the training set (cross-validation 1: CV1), a subset of the sites (cross-validation 2: CV2), and/or with 2 yield surrogates (flowering time and fall plant height). We found that genotype-by-environment interactions were largely due to the north–south distribution of sites. The genetic correlations between the yield surrogates and the biomass yield were generally positive (mean height r = 0.85; mean flowering time r = 0.45) and did not vary due to subpopulation or growing region (North, Middle, or South). Genomic prediction models had CV predictive abilities of –0.02 for individuals using only genetic data (CV1), but 0.55, 0.69, 0.76, 0.81, and 0.84 for individuals with biomass performance data from 1, 2, 3, 4, and 5 sites included in the training data (CV2), respectively. To simulate a resource-limited breeding program, we determined the predictive ability of models provided with the following: 1 site observation of flowering time (0.39); 1 site observation of flowering time and fall height (0.51); 1 site observation of fall height (0.52); 1 site observation of biomass (0.55); and 5 site observations of biomass yield (0.84). The ability to share information at a regional scale is very encouraging, but further research is required to accurately translate spaced plant biomass to commercial-scale sward biomass performance.

09 BIOMASS FUELS

Machine learning inversion from scattering for mechanically driven polymers

A machine learning inversion method is developed for analyzing scattering functions of mechanically driven polymers and extracting the corresponding feature parameters, which include energy parameters and conformation variables. The polymer is modeled as a chain of fixed-length bonds constrained by bending energy, and it is subject to external forces such as stretching and shear. We generate a data set consisting of random combinations of energy parameters, including bending modulus, stretching and shear force, along with Monte Carlo-calculated scattering functions and conformation variables such as end-to-end distance, radius of gyration and off-diagonal component of the gyration tensor. The effects of the energy parameters on the polymer are captured by the scattering function, and principal component analysis ensures the feasibility of the machine learning inversion. Finally, we train a Gaussian process regressor using part of the data set as a training set and validate the trained regressor for inversion using the rest of the data. The regressor successfully extracts the feature parameters.

Gaussian process regressors

A Hybrid Biophysical‐Machine Learning Framework for Diurnal Surface Energy Flux Estimation Using Proximal Sensing

Thermal infrared-based remote sensing of surface energy fluxes has traditionally relied on high spatial resolution satellite data with revisit frequencies on the order of weeks. In this study, we evaluate a biophysics-based analytical surface energy balance model for predicting latent energy ( LE ) and sensible heat ( H ) fluxes using proximal sensing observations. The Surface Temperature Initiated Closure (STIC1.2) model has been extensively validated across a wide range of spatial and temporal scales using various satellite-derived thermal infrared data sets. Here we extend this validation by applying STIC at sub-hourly temporal resolution over multiple growing seasons for four distinct agricultural systems. We further develop and evaluate novel STIC variants that incorporate machine learning (ML) techniques to eliminate the need for surface energy balance observations, specifically net radiation and soil heat flux, thereby enhancing model applicability in data-sparse settings. The integration of a ML component to estimate surface available energy is shown to have strong predictive performance for both LE (R 2 = 0.81–0.94) and H (R 2 = 0.46–0.72) across all agricultural systems examined here, demonstrating the potential of hybrid biophysical-machine learning approaches for surface energy balance modeling with minimal data requirements. This study concludes with a novel application of explainable machine learning (exML) to diagnose sources of model error. This exML framework attributes residual prediction errors to both model input variables and environmental drivers not explicitly included in the simulation experiments. This approach provides a new pathway for improving model design and integrating previously overlooked yet influential variables into future model iterations.

evapotranspiration

Active multi-mode data analysis to improve fault diagnosis in AHUs

Faults in heating, ventilation and air conditioning systems can lead to increased energy consumption, occupant comfort issues, and reduced equipment lifetime. Commercial fault detection and diagnosis (FDD) tools has been increasingly deployed in U.S. commercial buildings. While they are helping to achieve energy efficiency and operational reliability, there remain gaps in their fault diagnostic capabilities. The diagnostic results often contain multiple distinct candidate root causes (CRCs) or offer no insight into CRCs. This study developed a novel active rule-based multi-mode data analysis method to enhance diagnostic resolution by applying proven rule sets and additional new rules to data from multiple known operational modes. The proposed method was demonstrated using enhanced air handling unit performance assessment rule sets and validated with the simulated data of two air handling units. New metrics, namely, reduced number of CRCs and improvement ratio, were developed to quantify the improvement of fault diagnostic resolution. The validation results showed that the proposed method effectively reduced the number of CRCs in contrast to analyzing data solely for a single mode of operation. It achieved a median improvement ratio of 80% in 19 test cases.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

The ARM Precipitation Best Estimate (PrecipBE) Value-Added Product Report

The U.S. Department of Energy Atmospheric Radiation Measurement (ARM) User Facility’s Precipitation Best-Estimate (PrecipBE) Value-Added Product integrates multiple precipitation datastreams, accounting for data quality and instrument limitations, to deliver comprehensive per-precipitation event properties alongside ancillary ARM data set data. PrecipBE bundles all valid surface rainfall samples into artificial intelligence (AI)-ready tabular and time-series formats, reporting bundle means and uncertainty ranges. This per-event structure provides an insightful and easy-to-use resource for researchers analyzing precipitation characteristics.

54 ENVIRONMENTAL SCIENCES

Experimental Investigation of Vapor Formation in Liquid CO2 Flow Through a Converging-Diverging Nozzle

Carbon dioxide is an attractive working fluid for many cycles, including for pumped thermal energy storage (PTES). A challenge with some proposed sCO2 PTES cycles is the operation of sCO2 machinery outside the typical bounds of experience, with local phase change from the liquid state being particularly unknown. Presently, there is insufficient data in the literature regarding multiphase CO2 to adequately design a multiphase-tolerant turbine, so generation of foundational data is required. This experimental study investigates the flow characteristics of sub-sonic liquid CO2 undergoing expansion and phase change in a converging-diverging nozzle. The nozzle is instrumented to measure static pressure, unsteady pressure, temperature, and density. The static pressure transducers are located at 27 axial locations to accurately characterize the pressure profile in the nozzle. High-accuracy RTDs are located at the entrance and exit of the nozzle, and three dynamic pressure transducers are strategically located to capture any unsteady phenomena. During testing, values of mass flow and nozzle inlet pressure are swept to vary the pressure drop and fluid properties. The measured total pressure drop in the nozzle is compared to a homogenous model and the Lockhart-Martinelli correlation method, with the latter predicting loss quite closely. The resulting data set is valuable for validating multiphase numerical models in a simple geometry before implementation of these models in turbomachinery design.

25 ENERGY STORAGE