Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Statistical accuracy”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Sampling Size Optimization for Bioburden Density Estimation in Planetary Protection

Planetary protection (PP) is a discipline that focuses on minimizing the biological contamination of spacecraft to ensure compliance with international policy. Precise estimation of bioburden - the total number of microbes in or on spacecraft hardware – and the bioburden density are of utmost importance for PP. Such estimation is the way concordance with requirements is demonstrated, and it is critical for quantifying the potential risk of inadvertently contaminating other planetary bodies. Although a suite of molecular techniques have been used to thoroughly characterize and profile the microbiome of various cleanroom environments and spacecraft, the gold standard remains the physical enumeration of microbes via culturing of samples directly taken from spacecraft and associated surfaces. However, due to technical, budgetary, and programmatic constraints, only a manageable portion (around 10%) of the entire spacecraft surface is directly sampled with cotton swabs or wipes. To generate the bioburden current best estimate (CBE) for components not directly verifiable, the accepted approach is to apply a NASA-defined bioburden estimate based on the components’ manufacturing or assembly environment. This approach utilizes a prespecified bioburden density estimation that applies a maximum value across the total surface area of the specified component. For hardware components that underwent similar assembly processes, an implied bioburden is adopted for all components, based on a direct verification of a representative component within the same lot. Once all components have a CBE, the bioburden estimates are generated. In previous publication [ 1], we have shown that statistical risks quantifying the accuracy of the estimates for sampled, prespecified, and implied components can be derived and ranked. For mean squared error (MSE) function, the risks are available analytically and hence a cost function can be obtained to optimize the risks with respect to the sampling area and sampling cost. Since the sampling area and sampling cost are two complimentary variables, their sum will have a well-defined minimum. This paper presents the multivariate optimization of the integrated risk of an empirical Bayes estimator to determine the optimal sampling schedule for a given number of components. It is assumed that given a number of components, N, the bioburden density for each component can either be sampled, implied, or prespecified. The multivariate optimization searches through different options to sample, imply or prespecify the bioburden density for a component, and account for the component’s surface area and cost of sampling. The idea of the optimization is based on the observation that the statistical risk of using an estimator is a monotonically decreasing function of the sampled area. The larger the sampled area, the lower the risk of using the estimator as the estimator becomes more and more accurate as the sampling area increases. On the other hand, the cost of sampling is monotonically increasing as the sampled surface grows. This makes the risk and total cost of sampling complimentary variables which can be counterbalanced to achieve an optimal overall value with respect to the sampled surface. In this paper, the integrated risk has been used to quantify the accuracy of the estimator. This risk has been selected because it depends on neither the true value of the parameter nor on the collected data. The cost of each sample was also available to obtain the total cost of sampling of N components. The paper will present the results based on computer-simulated data as well as the data collected during the InSight mission. The computer-simulated data have N components with randomly generated total areas and each component assigned to one of the three categories according to the method of estimating of bioburden density: sampled, implied, or prespecified. The cost of sampling is also available. The cost of sampling is estimated based on a cost model provided by the planetary protection group at JPL. For this paper, the overall cost was assumed to be a linear function of exposure. The optimization process finds the allocation of the components to the three categories that minimizes the tradeoff between integrated risk and total cost. For the InSight data, a set of components is selected representing all three categories, and optimization is performed to determine if the performed allocation was optimal or if a better allocation could have been obtained. To the best of our knowledge, this work is the first attempt not only perform an accurate estimation of bioburden density but also do it in an optimal way.

97 - MATHEMATICS AND COMPUTING↗

Perspectives for Measuring Neutrino Cross-Sections at Short Baseline Detectors (ICARUS and DUNE-PRISM-like Detector) and Their Implications for Precision Physics

Neutrino physics is rapidly advancing with current and future experiments promising high-statistics data, enabling unprecedented accuracy in measuring Standard Model (SM) parameters and providing crucial insights into neutrino-nucleus interactions. The ICARUS experiment, utilizing LAr TPC technology, is pivotal in this growth. As the far detector of the SBN program at Fermilab, ICARUS is set to investigate the neutrino anomalies observed in LSND and MiniBooNE and prepare for future long-baseline experiments like DUNE. Positioned along the BNB and 5.75°off-axis from the NuMI beamline, ICARUS benefits from numerous neutrino interactions, allowing for precise neutrino-argon nucleus cross-section measurements, which will benchmark future DUNE measurements. This work presents the prospects and progress in measuring the cross-section of the NuMI νµ CC inclusive channel in the ICARUS experiment. High-statistics data will also enable precise determinations of fundamental parameters like the weak mixing angle and the neutrino charge radius (NCR), crucial for testing the SM and exploring potential new physics. This work includes studies of future near detectors, such as DUNE-PRISM, which are sensitive to radiative corrections in neutrino-electron scattering. The DUNE-PRISM configuration allows for the analysis of different neutrino energy spectra, promising accurate NCR measurements with controlled systematic uncertainties.

Moreno-Granados, Guadalupe↗

A Regression Method for Comparing Launch Zone Ranges Derived From Different Sources

In modern fighter aircraft air-to-air missile launch acceptable ranges (LARs) are provided to the pilot for his launch decision. These LARs are calculated in the avionic software, using very simple algebraic equations and algorithms, that attempt to model complex missile behavior. Frequently these algorithms are revised to reflect inaccuracies discovered, when comparisons are made with missile six degree-of-freedom (6-DOF) simulations. The process of modifying this software has become expensive. In the past, these simple LAR generators were modified based only on subjective judgmental approach, with little regard to the missile range design limits. With reduced budgets a more rational methodology should be taken to assess whether or not simple fighter LAR generators need to be modified. Any LAR generator can produce large amounts of range (minimum and maximum) data over the board engagement limits of modem fighter and air-to-air missiles. This infinite population lends itself to some statistical methodology to access simple LAR accuracy and the need for software modification. The methodology should produce statistical parameters for use in the decision making process of modifying a simple fighter LAR generator. This paper presents a regression method, based on a missile range design limit philosophy. Acceptance criteria from the regression method are given that provide simple statistical parameters to aid in this decision making process.

Flight Testing↗

SLS Integrated Modal Test Uncertainty Quantification using the Hybrid Parametric Variation Method

Uncertainty in structural loading during launch is a significant concern in the development of spacecraft and launch vehicles. Small variations in launch vehicle and payload mode shapes and their interaction can result in significant variation in system loads. In many cases involving large aerospace systems it is difficult, not economical, or impossible to perform a system modal test. However, it is still vital to obtain test results that can be compared with analytical predictions to validate models. Instead, the “Building Block Approach” is used in which system components are tested individually. Component models are correlated and updated to agree as best they can with test results. The Space Launch System consists of a number of components that are assembled into a launch vehicle. Finite element models of the components are developed, reduced to Hurty/Craig-Bampton models and assembled to represent different phases of flight. The only opportunity to obtain modal test data from an assembled Space Launch System will be during the Integrated Modal Test. There is always uncertainty in every model, which flows into uncertainty in predicted system results. Uncertainty Quantification is used to determine statistical bounds on prediction accuracy based on model uncertainty. For the Space Launch System, model uncertainty is at the Hurty/Craig-Bampton component level. Uncertainty in the Hurty/Craig-Bampton components is quantified using the hybrid parametric variation approach that combines parametric and nonparametric uncertainty. Uncertainty in model form is one of the biggest contributors to uncertainty in complex built-up structures. This type of uncertainty cannot be represented by variations infinite element model input parameters and thus cannot be included in a parametric approach. However, model-form uncertainty can be modeled using a nonparametric approach based on random matrix theory. The hybrid parametric variation method requires the selection of dispersion values for the Hurty/Craig-Bampton fixed-interface eigenvalues, and the Hurty/Craig-Bampton stiffness matrices. Component test/analysis frequency error is used to identify the fixed-interface eigenvalue dispersions, while test/analysis cross-orthogonality is used to identify stiffness dispersion values. The hybrid parametric variation uncertainty quantification approach is applied to the Space Launch System Integrated Modal Test configuration. Monte Carlo analysis is performed, and statistics are determined for modal correlation metrics, frequency response from Integrated Modal Test shakers to selected accelerometers, as well as other metrics for determining how well target modes are excited and identified. If the predicted uncertainty envelopes future Integrated Modal Test results, then there will be increased confidence in the utility of the component-based hybrid parametric variation uncertainty quantification approach.

Uncertainty Quantification↗

Geospatial Method for Computing Supplemental Multi-Decadal U.S. Coastal Land-Use and Land-Cover Classification Products, Using Landsat Data and C-CAP Products

This paper discusses the development and implementation of a geospatial data processing method and multi-decadal Landsat time series for computing general coastal U.S. land-use and land-cover (LULC) classifications and change products consisting of seven classes (water, barren, upland herbaceous, non-woody wetland, woody upland, woody wetland, and urban). Use of this approach extends the observational period of the NOAA-generated Coastal Change and Analysis Program (C-CAP) products by almost two decades, assuming the availability of one cloud free Landsat scene from any season for each targeted year. The Mobile Bay region in Alabama was used as a study area to develop, demonstrate, and validate the method that was applied to derive LULC products for nine dates at approximate five year intervals across a 34-year time span, using single dates of data for each classification in which forests were either leaf-on, leaf-off, or mixed senescent conditions. Classifications were computed and refined using decision rules in conjunction with unsupervised classification of Landsat data and C-CAP value-added products. Each classification's overall accuracy was assessed by comparing stratified random locations to available reference data, including higher spatial resolution satellite and aerial imagery, field survey data, and raw Landsat RGBs. Overall classification accuracies ranged from 83 to 91% with overall Kappa statistics ranging from 0.78 to 0.89. The accuracies are comparable to those from similar, generalized LULC products derived from C-CAP data. The Landsat MSS-based LULC product accuracies are similar to those from Landsat TM or ETM+ data. Accurate classifications were computed for all nine dates, yielding effective results regardless of season. This classification method yielded products that were used to compute LULC change products via additive GIS overlay techniques.

Spruce, J. P.↗

Stochastic Simulation of Daily Suspended Sediment Concentration Using Multivariate Copulas

Estimation of daily suspended sediment concentration (SSC) is required for water resources and environment management. In this paper, a copula-based stochastic method was proposed for daily SSC simulation. Here, the multivariate copula function, constructed based on a bivariate copula and two bivariate conditional probability distributions, was used to model the temporal and cross dependence structures in daily SSCs. Then, the daily SSCs were generated by sampling from the multivariate conditional distribution. As a result, synthetic long-term SSCs data beyond the limited observation period can be provided for water resources managers, which plays a critical role in accurately estimating frequency and magnitude of extreme SSCs events. The proposed method was under rigorous examination by applying to a case study at Pingshan station in the Jinsha River Basin, China. Results showed that the generated daily SSC sequences not only had a high degree of accuracy in preserving the statistical characteristics of the daily SSC observations, but also captured both the temporal correlation and the cross-correlation between the daily streamflow and daily SSC. Specifically, the average daily relative error values corresponding to mean, standard deviation, skewness, lag-1 temporal correlation, and cross correlation were 0.87%, 4.24%, 7.52%, 0.51% and 2.02%, respectively. The multivariate copula framework proposed here can accurately and efficiently generate long-term daily SSC data for water resources management such as frequency analysis and risk assessment of extreme SSC events.

54 ENVIRONMENTAL SCIENCES↗

Demonstrating the viability of Lagrangian in situ reduction on supercomputers

Performing exploratory analysis and visualization of large-scale time-varying computational science applications is challenging due to inaccuracies that arise from under-resolved data. In recent years, Lagrangian representations of the vector field computed using in situ processing are being increasingly researched and have emerged as a potential solution to enable exploration. However, prior works have offered limited estimates of the encumbrance on the simulation code as they consider “theoretical” in situ environments. Further, the effectiveness of this approach varies based on the nature of the vector field, benefitting from an in-depth investigation for each application area. With this study, an extended version of Sane et al. (2021), we contribute an evaluation of Lagrangian analysis viability and efficacy for simulation codes executing at scale on a supercomputer. We investigated previously unexplored cosmology and seismology applications as well as conducted a performance benchmarking study by using a hydrodynamics mini-application targeting exascale computing. Here, to inform encumbrance, we integrated in situ infrastructure with simulation codes, and evaluated Lagrangian in situ reduction in representative homogeneous and heterogeneous HPC environments. To inform post hoc accuracy, we conducted a statistical analysis across a range of spatiotemporal configurations as well as a qualitative evaluation. Additionally, our study contributes cost estimates for distributed-memory post hoc reconstruction. In all, we demonstrate viability for each application — data reduction to less than 1% of the total data via Lagrangian representations, while maintaining accurate reconstruction and requiring under 10% of total execution time in over 90% of our experiments.

97 MATHEMATICS AND COMPUTING↗

Future Climate Projections for South Florida: Improving the Accuracy of Air Temperature and Precipitation Extremes With a Hybrid Statistical Bias Correction Technique

Projecting future climate variables is essential for comprehending the potential impacts on hydroclimatic hazards like floods and droughts. Evaluating these impacts is challenging due to the coarse spatial resolution of global climate models (GCMs); therefore, bias correction is widely used. Here, we applied two statistical methods—standard empirical quantile mapping (EQM) and a hybrid approach, EQM with linear correction (EQM-LIN)—to bias correct precipitation and air temperature simulated by nine GCMs. We used historical observations from 20 weather stations across South Florida to project future climate under three shared socioeconomic pathways (SSPs). Compared to the EQM, the hybrid EQM-LIN method improved R 2 of daily quantiles by up to 30% over the historical period and improved MAE up to 70% in months that contain most extreme values. Projected extreme precipitation at the weather stations showed that, compared to the EQM-LIN, the EQM method underestimates the high quantiles by up to 26% in SSP585. The projected changes in annual maximum precipitation from historical period (1985–2014) to near future (2040–2069) and far future (2070–2100) were between 2% and 16% across the study area. Projected future precipitation suggested a slight decrease during summer but an increase in fall. This, along with rising summer temperatures, suggested that South Florida can experience rapid oscillations from warmer summers and increased flooding in fall under future climate. Additionally, our comparative analyses with globally and nationally downscaled studies showed that such coarse scale studies do not represent the climatic extremes well, particularly for high quantile precipitation.

54 ENVIRONMENTAL SCIENCES↗

Projected income data under different shared socioeconomic pathways for Washington state

Abstract High-resolution income projections under different Shared Socioeconomic Pathways (SSPs) are essential for the climate change research communities to devise climate change adaptation and mitigation strategies. To generate income projections for Washington state, we obtain state-level GDP per capita projections and convert them into projected annual household income. The resulting state-level income projections are subsequently downscaled to the census block-level based on the Longitudinal Origin-Destination Employment Statistics (LODES) dataset. For accuracy assessment, we downscale historical income data from state- level to block- and block group-level and compare the downscaled results against the actual income data from LODES. County-level accuracy assessment is also conducted based on American Community Survey. The results demonstrate a good agreement (Average R 2 of 0.67, 0.8, and 0.99 for block-, block group-, and county-level, respectively) between the downscaled income data and the reference data, thereby validating the methodology employed. Our approach is applicable to other states for income projections, which can be utilized by a broader audience, including those involved in demographic analysis, economic research, and urban planning.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Predictive understanding of the surface tension and velocity of sound in ionic liquids using machine learning

Knowledge of the physical properties of ionic liquids (ILs), such as the surface tension and speed of sound, is important for both industrial and research applications. Unfortunately, technical challenges and costs limit exhaustive experimental screening efforts of ILs for these critical properties. Previous work has demonstrated that the use of quantum-mechanics-based thermochemical property prediction tools, such as the conductor-like screening model for real solvents, when combined with machine learning (ML) approaches, may provide an alternative pathway to guide the rapid screening and design of ILs for desired physiochemical properties. However, the question of which machine-learning approaches are most appropriate remains. In the present study, we examine how different ML architectures, ranging from tree-based approaches to feed-forward artificial neural networks, perform in generating nonlinear multivariate quantitative structure–property relationship models for the prediction of the temperature- and pressure-dependent surface tension of and speed of sound in ILs over a wide range of surface tensions (16.9–76.2 mN/m) and speeds of sound (1009.7–1992 m/s). The ML models are further interrogated using the powerful interpretation method, shapley additive explanations. We find that several different ML models provide high accuracy, according to traditional statistical metrics. The decision tree-based approaches appear to be the most accurate and precise, with extreme gradient-boosting trees and gradient-boosting trees being the best performers. However, our results also indicate that the promise of using machine-learning to gain deep insights into the underlying physics driving structure–property relationships in ILs may still be somewhat premature.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Reducing Communication Overhead in Federated Learning for Network Anomaly Detection with Adaptive Client Selection

Communication overhead in federated learning (FL) poses a significant challenge for network anomaly detection systems, where the myriad of client configurations and network conditions can severely impact system efficiency and detection accuracy. While existing approaches attempt to address this through individual optimization techniques, they often fail to maintain the delicate balance between reduced overhead and detection performance. This paper presents an adaptive FL framework that dynamically combines batch size optimization, client selection, and asynchronous updates to achieve efficient anomaly detection. Through extensive profiling and experimental analysis on two distinct datasets-UNSW-NBIS for general network traffic and ROAD for automotive networks-our framework reduces communication overhead by 97.6%; (from 700.0s to 16.8s) compared to synchronous baseline approaches while maintaining comparable detection accuracy (95.10%; vs. 95.12%;). Statistical validation using Mann-Whitney U test confirms significant improvements (p < 0.05) over existing FL approaches across both datasets, demonstrating the framework's adaptability to different network security contexts. Detailed profiling analysis reveals the efficiency gains through dramatic reductions in GPU operations and memory transfers while maintaining robust detection performance under varying client conditions.

Marfo, William [University of Texas at El Paso]↗

Single-shot picosecond pump coherent Rayleigh scattering thermometry

A single-shot coherent Rayleigh scattering (CRS) technique capable of measurement times less than 10 ns is presented. Here, the use of a mode-locked picosecond pump laser yields repeatable electrostrictive forcing compared to previous CRS experiments using unseeded nanosecond pump pulses, which are beset by shot-to-shot variations. Quantitative measurements are achieved by dispersing the CRS signal onto an EMCCD sensor using a virtually imaged phased array and comparing the experimental spectra to an existing kinetic model with a least-squares fitting routine. Measurements are demonstrated at ambient and low-pressure (2 Torr), low-temperature (100 K) conditions where the CRS measurement is within the collisionless regime. Single-shot statistics indicated precision and accuracy within 4% at ambient conditions and within 8% at low density and temperature conditions. This single-shot CRS technique is a powerful diagnostic tool with potential for multi-parameter measurements in complex flow environments, including high-speed aerodynamic ground test facilities.

Senior, William Charles Bowman [Sandia National La↗

Regional Land Use Mapping: the Phoenix Pilot Project

The Phoenix Pilot Program has been designed to make effective use of past experience in making land use maps and collecting land use information. Conclusions reached from the project are: (1) Land use maps and accompanying statistical information of reasonable accuracy and quality can be compiled at a scale of 1:250,000 from orbital imagery. (2) Orbital imagery used in conjunction with other sources of information when available can significantly enhance the collection and analysis of land use information. (3) Orbital imagery combined with modern computer technology will help resolve the problem of obtaining land use data quickly and on a regular basis, which will greatly enhance the usefulness of such data in regional planning, land management, and other applied programs. (4) Agreement on a framework or scheme of land use classification for use with orbital imagery will be necessary for effective use of land use data.

Anderson, J. R.↗

Generation of pseudo-random numbers

Practical methods for generating acceptable random numbers from a variety of probability distributions which are frequently encountered in engineering applications are described. The speed, accuracy, and guarantee of statistical randomness of the various methods are discussed.

Howell, L. W.↗

Performance of the spectropolarimeter for the Space Telescope faint object spectrograph

The design and preliminary test results for the spectropolarimeter for the Faint Object Spectrograph (FOS) for the Space Telescope are presented. The mechanical design and optical specifications of the spectropolarimeter are described noting that a Wollaston prism with an internal wedge angle of 20 deg is fixed behind each of two rotatable waveplate retarders of different retardations. Either waveplate/prism combination can be positioned at either of the two FOS entrance ports. Magnesium fluoride is chosen as the birefringent crystal for the polarizing elements to allow linear and circular polarization measurements down to Lyman-alpha at 1216 A. Mechanical stability and repeatability were determined by operational testing to give polarization-position angles of + or - 0.5 deg, corresponding to degree-of-polarization measurements of + or - 0.1 percent. Faint-object accuracy, dependent on photoelectron statistics and hence on observation time, is calculated to be one percent in each 100-A-wide spectral band for a 20-min observation of an AO star with V = 15th magnitude.

Allen, R. G.↗

Classifying northern forests using Thematic Mapper Simulator data

Thematic Mapper Simulator data were collected over a 23,200 hectare forested area near Baxter State Park in north-central Maine. Photointerpreted ground reference information was used to drive a stratified random sampling procedure for waveband discriminant analyses and to generate training statistics and test pixel accuracies. Stepwise discriminant analyses indicated that the following bands best differentiated the thirteen level II - III cover types (in order of entry): near infrared (0.77 to 0.90 micron), blue (0.46 0.52 micron), first middle infrared (1.53 to 1.73 microns), second middle infrared (2.06 to 2.33 microsn), red (0.63 to 0.69 micron), thermal (10.32 to 12.33 microns). Classification accuracies peaked at 58 percent for thirteen level II-III land-cover classes and at 65 percent for ten level II classes.

Nelson, R. F.↗

Landsat Thematic Mapper studies of land cover spatial variability related to hydrology

Past accomplishments involving remote sensing based land-cover analysis for hydrologic applications are reviewed. Ongoing research in exploiting the increased spatial, radiometric, and spectral capabilities afforded by the TM on Landsats 4 and 5 is considered. Specific studies to compare MSS and TM for urbanizing watersheds, wetlands, and floodplain mapping situations show that only a modest improvement in classification accuracy is achieved via statistical per pixel multispectral classifiers. The limitations of current approaches to multispectral classification are illustrated. The objectives, background, and progress in the development of an alternative analysis approach for defining inputs to urban hydrologic models using TM are discussed.

Wharton, S.↗

Experimental characterization of the perceptron laser rangefinder

In this report, we characterize experimentally a scanning laser rangefinder that employs active sensing to acquire three-dimensional images. We present experimental techniques applicable to a wide variety of laser scanners, and document the results of applying them to a device manufactured by Perceptron. Nominally, the sensor acquires data over a 60 deg x 60 deg field of view in 256 x 256 pixel images at 2 Hz. It digitizes both range and reflectance pixels to 12 bits, providing a maximum range of 40 m and a depth resolution of 1 cm. We present methods and results from experiments to measure geometric parameters including the field of view, angular scanning increments, and minimum sensing distance. We characterize qualitatively problems caused by implementation flaws, including internal reflections and range drift over time, and problems caused by inherent limitations of the rangefinding technology, including sensitivity to ambient light and surface material. We characterize statistically the precision and accuracy of the range measurements. We conclude that the performance of the Perceptron scanner does not compare favorably with the nominal performance, that scanner modifications are required, and that further experimentation must be conducted.

Kweon, I. S.↗