Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Model error”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 397 records · Page 22

Calibration, measurement, and characterization of soil moisture dynamics in a central Amazonian tropical forest

Soil moisture plays a key role in hydrological, biogeochemical, and energy budgets of terrestrial ecosystems. Accurate soil moisture measurements in remote ecosystems such as the Amazon are difficult and limited because of logistical constraints. Time domain reflectometry (TDR) sensors are widely used to monitor soil moisture and require calibration to convert the TDR's dielectric permittivity measurement (K a ) to volumetric water content (θ v ). In this study, our objectives were to develop a field-based calibration of TDR sensors in an old-growth upland forest in the central Amazon, to evaluate the performance of the calibration, and then to apply the calibration to determine the dynamics of soil moisture content within a 14.2 -m-deep vertical soil profile. Depth-specific TDR calibration using local soils in a controlled laboratory setting yielded a novel K a –θ v third-degree polynomial calibration. The sensors were later installed to their specific calibration depth in a 14.2-m pit. The widely used K a –θ v relationship (Topp model) underestimated the site-specific θ v by 22–42%, indicating significant error in the model when applied to these well-structured, clay-rich tropical forest soils. The calibrated wet- and dry-season θ v data showed a variety of depth and temporal variations highlighting the importance of soil textural differentiation, root uptake depths, as well as event to seasonal precipitation effects. Data such as these are greatly needed for improving our understanding of ecohydrological processes within tropical forests and for improving models of these systems in the face of changing environmental conditions.

54 ENVIRONMENTAL SCIENCES↗

Wind turbine gearbox fault prognosis using high-frequency SCADA data

Condition-based maintenance using routinely collected Supervisory Control and Data Acquisition (SCADA) data is a promising strategy to reduce downtime and costs associated with wind farm operations and maintenance. New approaches are continuously being developed to improve the condition monitoring for wind turbines. Development of normal behaviour models is a popular approach in studies using SCADA data. This paper first presents a data-driven framework to apply normal behaviour models using an artificial neural network approach for wind turbine gearbox prognostics. A one-class support vector machine classifier, combining different error parameters, is used to analyse the normal behaviour model error to develop a robust threshold to distinguish anomalous wind turbine operation. A detailed sensitivity study is then conducted to evaluate the potential of using high-frequency SCADA data for wind turbine gearbox prognostics. The results based on operational data from one wind turbine show that, compared to the conventionally used 10-min averaged SCADA data, the use of high-frequency data is valuable as it leads to improved prognostic predictions. High-frequency data provides more insights into the dynamics of the condition of the wind turbine components and can aid in earlier detection of faults.

17 WIND ENERGY↗

Comparing Compressed and Full-Modeling analyses with FOLPS: implications for DESI 2024 and beyond

The Dark Energy Spectroscopic Instrument (DESI) will provide unprecedented information about the large-scale structure of our Universe. In this work, we study the robustness of the theoretical modelling of the power spectrum of F OLPS , a novel effective field theory-based package for evaluating the redshift space power spectrum in the presence of massive neutrinos. We perform this validation by fitting the AbacusSummit high-accuracy N -body simulations for Luminous Red Galaxies, Emission Line Galaxies and Quasar tracers, calibrated to describe DESI observations. We quantify the potential systematic error budget of F OLPS finding that the modelling errors are fully sub-dominant for the DESI statistical precision within the studied range of scales. Additionally, we study two complementary approaches to fit and analyse the power spectrum data, one based on direct Full-Modelling fits and the other on the ShapeFit compression variables, both resulting in very good agreement in precision and accuracy. In each of these approaches, we study a set of potential systematic errors induced by several assumptions, such as the choice of template cosmology, the effect of prior choice in the nuisance parameters of the model, or the range of scales used in the analysis. Furthermore, we show how opening up the parameter space beyond the vanilla ΛCDM model affects the DESI observables. These studies include the addition of massive neutrinos, spatial curvature, and dark energy equation of state. We also examine how relaxing the usual Cosmic Microwave Background and Big Bang Nucleosynthesis priors on the primordial spectral index and the baryonic matter abundance, respectively, impacts the inference on the rest of the parameters of interest. This paper pathways towards performing a robust and reliable analysis of the shape of the power spectrum of DESI galaxy and quasar clustering using F OLPS .

79 ASTRONOMY AND ASTROPHYSICS↗

A probability density function model describing height estimation uncertainty due to image pixel intensity noise in digital fringe projection measurements

Digital fringe projection is a surface-profiling technique used for highly accurate non-contact measurements. As with any measurement technique, a variety of sources degrade to the measurement accuracy of the method. Here, this paper presents an analytically-derived probability density function that explicitly models the surface height measurement error due to inevitable phase measurement error, and it includes the specific case of pixel noise inducing the phase measurement error that ultimately leads to the height estimation error. The accuracy of the model was validated through Monte-Carlo simulations of resultant height distributions subject to arbitrarily correlated pixel intensity noise and experimental digital fringe projection measurements where the pixel-by-pixel height uncertainty estimations were compared to the predictions of the derived model.

42 ENGINEERING↗

Novel Data Driven Noise Emulation Framework using Deep Neural Network for Generating Synthetic PMU Measurements

Sensors play a critical role in supporting day-to-day grid operations and they are essential to operator’s decision-making process. Furthermore, sensors and sensor behaviors need to be emulated with grid simulations to perform modeling studies and to design cutting edge power systems applications. Ensuring the accurate behavior of these applications requires accurate emulation of sensors and pertinent signals. However, most grid simulators and modeling tools assume either zero error scenarios or simplistic noise models that may not always correlate to real-world sensors. To address the above issue, this work presents an initial study on the noise characteristics of phasor measurement units (PMUs), along with models for recreating their unique noise signatures. The proposed methods (both analytical and machine-learning-based) provide a substantial increase in a sensor’s model fidelity, a feature that can be leveraged by an end-user application to yield more accurate system representations. The proposed methods were then applied to micro PMU data from the EPFL microgrid campus to extract sensor noise profiles. This data was used to train a deep learning model, which was tested to emulate the noise characteristics present in actual signals. Based on the observed results and the employed data-driven methodology, the proposed methods may be adapted to replicate the behavior of other grid sensors and power new applications capable of detecting sensor degradation and eventual device failures in near real-time.

PMU, noise emulation, synthetic measurements, deep↗

Better calibration of cloud parameterizations and subgrid effects increases the fidelity of the E3SM Atmosphere Model version 1

Abstract. Realistic simulation of the Earth's mean-state climate remains a major challenge, and yet it is crucial for predicting the climate system in transition. Deficiencies in models' process representations, propagation of errors from one process to another, and associated compensating errors can often confound the interpretation and improvement of model simulations. These errors and biases can also lead to unrealistic climate projections and incorrect attribution of the physical mechanisms governing past and future climate change. Here we show that a significantly improved global atmospheric simulation can be achieved by focusing on the realism of process assumptions in cloud calibration and subgrid effects using the Energy Exascale Earth System Model (E3SM) Atmosphere Model version 1 (EAMv1). The calibration of clouds and subgrid effects informed by our understanding of physical mechanisms leads to significant improvements in clouds and precipitation climatology, reducing common and long-standing biases across cloud regimes in the model. The improved cloud fidelity in turn reduces biases in other aspects of the system. Furthermore, even though the recalibration does not change the global mean aerosol and total anthropogenic effective radiative forcings (ERFs), the sensitivity of clouds, precipitation, and surface temperature to aerosol perturbations is significantly reduced. This suggests that it is possible to achieve improvements to the historical evolution of surface temperature over EAMv1 and that precise knowledge of global mean ERFs is not enough to constrain historical or future climate change. Cloud feedbacks are also significantly reduced in the recalibrated model, suggesting that there would be a lower climate sensitivity when it is run as part of the fully coupled E3SM. This study also compares results from incremental changes to cloud microphysics, turbulent mixing, deep convection, and subgrid effects to understand how assumptions in the representation of these processes affect different aspects of the simulated atmosphere as well as its response to forcings. We conclude that the spectral composition and geographical distribution of the ERFs and cloud feedback, as well as the fidelity of the simulated base climate state, are important for constraining the climate in the past and future.

54 ENVIRONMENTAL SCIENCES↗

Estimating Subhourly Inverter Clipping Loss From Satellite-Derived Irradiance Data

Photovoltaic system production simulations are conventionally run using hourly weather datasets. Hourly simulations are sufficiently accurate to predict the majority of long-term system behavior but cannot resolve high-frequency effects like inverter clipping caused by short-duration irradiance variability. Direct modeling of this subhourly clipping error is only possible for the few locations with high-resolution irradiance datasets. This paper describes a method of predicting the magnitude of this error using a machine learning regressor ensemble model, comprised of a random forest and an XGBoost model, and 30-minute satellite irradiance data. The method predicts a correction for each 30-minute interval with the potential to roll up into 60-minute corrections to match an hourly energy model. The model is trained and validated at locations where the error can be directly simulated from 1-minute ground data. The validation shows low bias at most ground station locations. The model is also applied to gridded satellite irradiance to produce a heatmap of the estimated clipping error across the United States. Finally, the relative importance of each predictor satellite variable is retrieved from the model and discussed.

41 EE - Solar Energy Technologies Office (EE-4S)↗

Multiscale modeling high-order methods and data-driven modeling

Projection-based reduced-order models (ROMs) comprise a promising set of data-driven approaches for accelerating the simulation of high-fidelity numerical simulations. Standard projection-based ROM approaches, however, suffer from several drawbacks when applied to the complex nonlinear dynamical systems commonly encountered in science and engineering. These limitations include a lack of stability, accuracy, and sharp a posteriori error estimators. This work addresses these limitations by leveraging multiscale modeling, least-squares principles, and machine learning to develop novel reduced-order modeling approaches, along with data-driven a posteriori error estimators, for dynamical systems. Theoretical and numerical results demonstrate that the two ROM approaches developed in this work - namely the windowed least-squares method and the Adjoint Petrov - Galerkin method - yield substantial improvements over state-of-the-art approaches. Additionally, numerical results demonstrate the capability of the a posteriori error models developed in this work.

97 MATHEMATICS AND COMPUTING↗

Optimizing qubit control pulses for state preparation

In the burgeoning field of quantum computing, the precise design and optimization of quantum pulses are essential for enhancing qubit operation fidelity. This study focuses on refining the pulse engineering techniques for superconducting qubits, employing a detailed analysis of square and Gaussian pulse envelopes under various approximation schemes. We evaluated the effects of coherent errors induced by naive pulse designs. Furthermore, we identified the sources of these errors in the Hamiltonian model’s approximation level. We mitigated these errors through adjustments to the external driving frequency and pulse durations, thus implementing a pulse scheme with stroboscopic error reduction. Our results demonstrate that these refined pulse strategies improve performance and reduce coherent errors. Moreover, the techniques developed herein are applicable across different quantum architectures, such as ion-trap, atomic, and photonic systems.

Chirp modulation↗

An Analysis of the Spatial Variations in the Relationship Between Built Environment and Severe Crashes

Traffic crashes significantly contribute to global fatalities, particularly in urban areas, highlighting the need to evaluate the relationship between urban environments and traffic safety. This study extends former spatial modeling frameworks by drawing paths between global models, including spatial lag (SLM), and spatial error (SEM), and local models, including geographically weighted regression (GWR), multi-scale geographically weighted regression (MGWR), and multi-scale geographically weighted regression with spatially lagged dependent variable (MGWRL). Utilizing the proposed framework, this study analyzes severe traffic crashes in relation to urban built environments using various spatial regression models within Leon County, Florida. According to the results, SLM outperforms OLS, SEM, and GWR models. Local models with lagged dependent variables outperform both the global and generic versions of the local models in all performance measures, whereas MGWR and MGWRL outperform GWR and GWRL. Local models performed better than global models, showing spatial non-stationarity; so, the relationship between the dependent and independent variables varies over space. The better performance of models with lagged dependent variables signifies that the spatial distribution of severe crashes is correlated. Finally, the better performance of multi-scale local models than classical local models indicates varying influences of independent variables with different bandwidths. According to the MGWRL model, census block groups close to the urban area with higher population, higher education level, and lower car ownership rates have lower crash rates. On the contrary, motor vehicle percentage for commuting is found to have a negative association with severe crash rate, which suggests the locality of the mentioned associations.

Alisan, Onur (ORCID:0000000193113984)↗

Likelihood-based signal and noise analysis for docking of models into cryo-EM maps

Fast, reliable docking of models into cryo-EM maps requires understanding of the errors in the maps and the models. Likelihood-based approaches to errors have proven to be powerful and adaptable in experimental structural biology, finding applications in both crystallography and cryo-EM. Indeed, previous crystallographic work on the errors in structural models is directly applicable to likelihood targets in cryo-EM. Likelihood targets in Fourier space are derived here to characterize, based on the comparison of half-maps, the direction- and resolution-dependent variation in the strength of both signal and noise in the data. Because the signal depends on local features, the signal and noise are analysed in local regions of the cryo-EM reconstruction. The likelihood analysis extends to prediction of the signal that will be achieved in any docking calculation for a model of specified quality and completeness. A related calculation generalizes a previous measure of the information gained by making the cryo-EM reconstruction.

59 BASIC BIOLOGICAL SCIENCES↗

A Fast Time-Stepping Strategy for Dynamical Systems Equipped with a Surrogate Model

Simulation of complex dynamical systems arising in many applications is computationally challenging due to their size and complexity. Model order reduction, machine learning, and other types of surrogate modeling techniques offer cheaper and simpler ways to describe the dynamics of these systems but are inexact and introduce additional approximation errors. In order to overcome the computational difficulties of the full complex models, on one hand, and the limitations of surrogate models, on the other, this work proposes a new accelerated time-stepping strategy that combines information from both. This approach is based on the multirate infinitesimal general-structure additive Runge--Kutta framework. The inexpensive surrogate model is integrated with a small time step to guide the solution trajectory, and the full model is treated with a large time step to occasionally correct for the surrogate model error and ensure convergence. Here, we provide a theoretical error analysis, and several numerical experiments, to show that this approach can be significantly more efficient than using only the full or only the surrogate model for the integration.

Surrogate models↗

Exact and Model Exchange-Correlation Potentials for Open-Shell Systems

The conventional approaches to the inverse density functional theory problem typically assume nondegeneracy of the Kohn–Sham (KS) eigenvalues, greatly hindering their use in open-shell systems. Here, we present a generalization of the inverse density functional theory problem that can seamlessly admit degenerate KS eigenvalues. Additionally, we allow for fractional occupancy of the Kohn–Sham orbitals to also handle noninteracting ensemble-v-representable densities, as opposed to just noninteracting pure-v-representable densities. We present the exact exchange-correlation (XC) potentials for six open-shell systems–four atoms (Li, C, N, and O) and two molecules (CN and CH 2 )–using accurate ground-state densities from configuration interaction calculations. We compare these exact XC potentials with model XC potentials obtained using nonlocal (B3LYP, SCAN0) and local/semilocal (SCAN, PBE, PW92) XC functionals. Although the relative errors in the densities obtained from these DFT functionals are of $O$(10 –3 to 10 –2 ), the relative errors in the model XC potentials remain substantially large–$O$(10 –1 to 10 0 ).

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Tula: Optimizing Time, Cost, and Generalization in Distributed Large-Batch Training

Distributed training increases the number of batches processed per iteration either by scaling-out (adding more nodes) or scaling-up (increasing the batch-size). However, the largest configuration does not necessarily yield the best performance. Horizontal scaling introduces additional communication overhead, while vertical scaling is constrained by computation cost and device memory limits. Thus, simply increasing the batch-size leads to diminishing returns: training time and cost decrease initially but eventually plateaus, creating a knee-point in the time/cost vs. batch-size pareto curve. The optimal batch-size therefore depends on the underlying model, data and available compute resources. Large batches also suffer from worse model quality due to the well-known “generalization gap”. In this paper, we present Tula, an online service that automatically optimizes time, cost, and convergence quality for large-batch training of convolutional models. It combines parallel-systems modeling with statistical performance prediction to identify the optimal batchsize. Tula predicts training time and cost within 7.5−14% error across multiple models, and achieves up to 20× overall speedup and improves test accuracy by ≈9% on average over standard large-batch training on various vision tasks, thus successfully mitigating the generalization gap and accelerating training at the same time.

Tyagi, Sahil [ORNL] (ORCID:0009000783144745)↗

Benchmarking the performance of uncertainty quantification methods for neural network-based interatomic potentials

Machine-learned interatomic potentials (ML-IAPs) continue to gain popularity as accurate, computationally efficient replacements for traditional, physics-based interatomic potentials and expensive ab initio methods. Uncertainty quantification (UQ) of ML-IAPs is a growing area of research as UQ is critical in many applications of IAPs, such as developing curated datasets, active learning-based data augmentation, self-improving models, and estimating the uncertainty of molecular dynamics simulations. In this paper, we construct and benchmark a series of different neural network potentials (NNPs) with varying network architectures to determine the performance of these models with respect to both the mean and uncertainty calibration error. Each NNP method is specifically designed to predict either epistemic or aleatoric uncertainty with particular focus on the differences in behavior between the epistemic and aleatoric uncertainty estimates. We benchmark these methods using multiple datasets common in the ML-IAP literature. The results show that the aleatoric uncertainty from single-shot model architectures is a competitive alternative to ensemble-based epistemic uncertainty predictions in regions of sufficient data-density. However, in regions where the representative data is sparse, aleatoric uncertainty models tend to overpredict and epistemic methods tend to underpredict the actual model error. We conclude that the type of UQ is crucial when discussing performance of probabilistic model results as different methods have different performance characteristics depending on the regime in which they are evaluated. Therefore, the type of UQ method should be carefully evaluated against both the data characteristics and requirements for the intended application.

97 MATHEMATICS AND COMPUTING↗

Exploring Autoencoder-based Error-bounded Compression for Scientific Data

Error-bounded lossy compression is becoming an indispensable technique for the success of today's scientific projects with vast volumes of data produced during the simulations or instrument data acquisitions. Not only can it significantly reduce data size, but it also can control the compression errors based on user-specified error bounds. Autoencoder (AE) models have been widely used in image compression, but few AE-based compression approaches support error-bounding features, which are highly required by scientific applications. To address this issue, we explore using convolutional autoencoders to improve error-bounded lossy compression for scientific data, with the following three key contributions. (1) We provide an in-depth investigation of the characteristics of various autoencoder models and develop an error-bounded autoencoder-based framework in terms of the SZ model. (2) We optimize the compression quality for main stages in our designed AE-based error-bounded compression framework, fine-tuning the block sizes and latent sizes and also optimizing the compression efficiency of latent vectors. (3) We evaluate our proposed solution using five real-world scientific datasets and comparing them with six other related works. Experiments show that our solution exhibits a very competitive compression quality from among all the compressors in our tests. In absolute terms, it can obtain a much better compression quality (100%similar to 800% improvement in compression ratio with the same data distortion) compared with SZ2.1 and ZFP in cases with a high compression ratio.

Liu, Jinyang↗

Verification of the 3-Region Advanced Test Reactor MCNP Model

The verification of the 3-region homogenized fuel Advanced Test Reactor MCNP model. The 3-region model was compared to the 19-plate model found in the 94-CIC report. Flux tallies, energy deposition tallies, and quarter core mesh tallies were used to compare the two models. The 3-region model needed updating in order to make good comparisons between the models. The percent error from the flux and energy deposition tallies data shows that experiment positions inside the flux trap have higher errors than positions outside the fuel ring. The standard deviation data obtained from the mesh tallies shows that the two models agree within two standard deviations throughout the reactor. It is concluded that the model works adequately for what it is used for.

99 GENERAL AND MISCELLANEOUS↗

Verification of the 3-Region Advanced Test Reactor MCNP Model

The verification of the 3-region homogenized fuel Advanced Test Reactor MCNP model. The 3-region model was compared to the 19-plate model found in the 94-CIC report. Flux tallies, energy deposition tallies, and quarter core mesh tallies were used to compare the two models. The 3-region model needed updating in order to make good comparisons between the models. The percent error from the flux and energy deposition tallies data shows that experiment positions inside the flux trap have higher errors than positions outside the fuel ring. The standard deviation data obtained from the mesh tallies shows that the two models agree within two standard deviations throughout the reactor. It is concluded that the model works adequately for what it is used for.

99 GENERAL AND MISCELLANEOUS↗