Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Statistical forecasting”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Soil Moisture Active Passive Mission L4_SM Data Product Assessment (Version 2 Validated Release)

During the post-launch SMAP calibration and validation (Cal/Val) phase there are two objectives for each science data product team: 1) calibrate, verify, and improve the performance of the science algorithm, and 2) validate the accuracy of the science data product as specified in the science requirements and according to the Cal/Val schedule. This report provides an assessment of the SMAP Level 4 Surface and Root Zone Soil Moisture Passive (L4_SM) product specifically for the product's public Version 2 validated release scheduled for 29 April 2016. The assessment of the Version 2 L4_SM data product includes comparisons of SMAP L4_SM soil moisture estimates with in situ soil moisture observations from core validation sites and sparse networks. The assessment further includes a global evaluation of the internal diagnostics from the ensemble-based data assimilation system that is used to generate the L4_SM product. This evaluation focuses on the statistics of the observation-minus-forecast (O-F) residuals and the analysis increments. Together, the core validation site comparisons and the statistics of the assimilation diagnostics are considered primary validation methodologies for the L4_SM product. Comparisons against in situ measurements from regional-scale sparse networks are considered a secondary validation methodology because such in situ measurements are subject to up-scaling errors from the point-scale to the grid cell scale of the data product. Based on the limited set of core validation sites, the wide geographic range of the sparse network sites, and the global assessment of the assimilation diagnostics, the assessment presented here meets the criteria established by the Committee on Earth Observing Satellites for Stage 2 validation and supports the validated release of the data. An analysis of the time average surface and root zone soil moisture shows that the global pattern of arid and humid regions are captured by the L4_SM estimates. Results from the core validation site comparisons indicate that "Version 2" of the L4_SM data product meets the self-imposed L4_SM accuracy requirement, which is formulated in terms of the ubRMSE: the RMSE (Root Mean Square Error) after removal of the long-term mean difference. The overall ubRMSE of the 3-hourly L4_SM surface soil moisture at the 9 km scale is 0.035 cubic meters per cubic meter requirement. The corresponding ubRMSE for L4_SM root zone soil moisture is 0.024 cubic meters per cubic meter requirement. Both of these metrics are comfortably below the 0.04 cubic meters per cubic meter requirement. The L4_SM estimates are an improvement over estimates from a model-only SMAP Nature Run version 4 (NRv4), which demonstrates the beneficial impact of the SMAP brightness temperature data. L4_SM surface soil moisture estimates are consistently more skillful than NRv4 estimates, although not by a statistically significant margin. The lack of statistical significance is not surprising given the limited data record available to date. Root zone soil moisture estimates from L4_SM and NRv4 have similar skill. Results from comparisons of the L4_SM product to in situ measurements from nearly 400 sparse network sites corroborate the core validation site results. The instantaneous soil moisture and soil temperature analysis increments are within a reasonable range and result in spatially smooth soil moisture analyses. The O-F residuals exhibit only small biases on the order of 1-3 degrees Kelvin between the (re-scaled) SMAP brightness temperature observations and the L4_SM model forecast, which indicates that the assimilation system is largely unbiased. The spatially averaged time series standard deviation of the O-F residuals is 5.9 degrees Kelvin, which reduces to 4.0 degrees Kelvin for the observation-minus-analysis (O-A) residuals, reflecting the impact of the SMAP observations on the L4_SM system. Averaged globally, the time series standard deviation of the normalized O-F residuals is close to unity, which would suggest that the magnitude of the modeled errors approximately reflects that of the actual errors. The assessment report also notes several limitations of the "Version 2" L4_SM data product and science algorithm calibration that will be addressed in future releases. Regionally, the time series standard deviation of the normalized O-F residuals deviates considerably from unity, which indicates that the L4_SM assimilation algorithm either over- or under-estimates the actual errors that are present in the system. Planned improvements include revised land model parameters, revised error parameters for the land model and the assimilated SMAP observations, and revised surface meteorological forcing data for the operational period and underlying climatological data. Moreover, a refined analysis of the impact of SMAP observations will be facilitated by the construction of additional variants of the model-only reference data. Nevertheless, the “Version 2” validated release of the L4_SM product is sufficiently mature and of adequate quality for distribution to and use by the larger science and application communities.

SMAP L4_SM↗

Deep Learning Experiments for Tropical Cyclone Intensity Forecasts

Reducing tropical cyclone (TC) intensity forecast errors is a challenging task that has interested the operational forecasting and research community for decades. To address this, we developed a deep learning (DL)-based multilayer perceptron (MLP) TC intensity prediction model. The model was trained using the global Statistical Hurricane Intensity Prediction Scheme (SHIPS) predictors to forecast the change in TC maximum wind speed for the Atlantic basin. In the first experiment, a 24-h forecast period was considered. To overcome sample size limitations, we adopted a leave one year out (LOYO) testing scheme, where a model is trained using data from all years except one and then evaluated on the year that is left out. When tested on 2010–18 operational data using the LOYO scheme, the MLP outperformed other statistical–dynamical models by 9%–20%. Additional independent tests in 2019 and 2020 were conducted to simulate real-time operational forecasts, where the MLP model again outperformed the statistical–dynamical models by 5%–22% and achieved comparable results as HWFI. The MLP model also correctly predicted more rapid intensification events than all the four operational TC intensity models compared. In the second experiment, we developed a lightweight MLP for 6-h intensity predictions. When coupled with a synthetic TC track model, the lightweight MLP generated realistic TC intensity distribution in the Atlantic basin. Therefore, the MLP-based approach has the potential to improve operational TC intensity forecasts, and will also be a viable option for generating synthetic TCs for climate studies.

58 GEOSCIENCES↗

Technical Report Series on Global Modeling and Data Assimilation: Soil Moisture Active Passive (SMAP) Project Assessment Report for the Beta-Release L4_SM Data Product - Volume 40

During the post-launch SMAP calibration and validation (Cal/Val) phase there are two objectives for each science data product team: 1) calibrate, verify, and improve the performance of the science algorithm, and 2) validate the accuracy of the science data product as specified in the science requirements and according to the Cal/Val schedule. This report provides an assessment of the SMAP Level 4 Surface and Root Zone Soil Moisture Passive (L4_SM) product specifically for the product's public beta release scheduled for 30 October 2015. The primary objective of the beta release is to allow users to familiarize themselves with the data product before the validated product becomes available. The beta release also allows users to conduct their own assessment of the data and to provide feedback to the L4_SM science data product team. The assessment of the L4_SM data product includes comparisons of SMAP L4_SM soil moisture estimates with in situ soil moisture observations from core validation sites and sparse networks. The assessment further includes a global evaluation of the internal diagnostics from the ensemble-based data assimilation system that is used to generate the L4_SM product. This evaluation focuses on the statistics of the observation-minus-forecast (O-F) residuals and the analysis increments. Together, the core validation site comparisons and the statistics of the assimilation diagnostics are considered primary validation methodologies for the L4_SM product. Comparisons against in situ measurements from regional-scale sparse networks are considered a secondary validation methodology because such in situ measurements are subject to upscaling errors from the point-scale to the grid cell scale of the data product. Based on the limited set of core validation sites, the assessment presented here meets the criteria established by the Committee on Earth Observing Satellites for Stage 1 validation and supports the beta release of the data. The validation against sparse network measurements and the evaluation of the assimilation diagnostics address Stage 2 validation criteria by expanding the assessment to regional and global scales.

ubRMSE↗

Understanding Differences in California Climate Projections Produced by Dynamical and Statistical Downscaling

Abstract We compare historical and end‐of‐century temperature and precipitation patterns over California from one dynamically downscaled simulation using the Weather Research and Forecast (WRF) model and two simulations statistically downscaled using Localized Constructed Analogs (LOCA). We uniquely separate causes of differences between dynamically and statistically based future climate projections into differences in historical climate (gridded observations versus regional climate model output) and differences in how these downscaling techniques explicitly handle future climate changes (numerical modeling versus analogs). In these methods, solutions between different downscaling techniques differ more in the future compared to the historical period. Changes projected by LOCA are insensitive to the choice of driving data. Only through dynamical downscaling can we simulate physically consistent regional springtime warming patterns across the Sierra Nevada, while the statistical simulations inherit an unphysical signal from their parent Global Climate Model (GCM) or gridded data. The results of our study clarify why these different techniques produce different outcomes and may also provide guidance on which downscaled products to use for certain impact analyses in California and perhaps other Mediterranean regimes.

54 ENVIRONMENTAL SCIENCES↗

Future mission studies: Preliminary comparisons of solar flux models

The results of comparisons of the solar flux models are presented. (The wavelength lambda = 10.7 cm radio flux is the best indicator of the strength of the ionizing radiations such as solar ultraviolet and x-ray emissions that directly affect the atmospheric density thereby changing the orbit lifetime of satellites. Thus, accurate forecasting of solar flux F sub 10.7 is crucial for orbit determination of spacecrafts.) The measured solar flux recorded by National Oceanic and Atmospheric Administration (NOAA) is compared against the forecasts made by Schatten, MSFC, and NOAA itself. The possibility of a combined linear, unbiased minimum-variance estimation that properly combines all three models into one that minimizes the variance is also discussed. All the physics inherent in each model are combined. This is considered to be the dead-end statistical approach to solar flux forecasting before any nonlinear chaotic approach.

Ashrafi, S.↗

A New Ensemble Canonical Correlation Prediction Scheme for Seasonal Precipitation

Department of Mathematical Sciences, University of Alberta, Edmonton, Canada This paper describes the fundamental theory of the ensemble canonical correlation (ECC) algorithm for the seasonal climate forecasting. The algorithm is a statistical regression sch eme based on maximal correlation between the predictor and predictand. The prediction error is estimated by a spectral method using the basis of empirical orthogonal functions. The ECC algorithm treats the predictors and predictands as continuous fields and is an improvement from the traditional canonical correlation prediction. The improvements include the use of area-factor, estimation of prediction error, and the optimal ensemble of multiple forecasts. The ECC is applied to the seasonal forecasting over various parts of the world. The example presented here is for the North America precipitation. The predictor is the sea surface temperature (SST) from different ocean basins. The Climate Prediction Center's reconstructed SST (1951-1999) is used as the predictor's historical data. The optimally interpolated global monthly precipitation is used as the predictand?s historical data. Our forecast experiments show that the ECC algorithm renders very high skill and the optimal ensemble is very important to the high value.

Kim, Kyu-Myong↗

Skillful Seasonal Forecasts of Land Carbon Uptake in Northern Mid- and High Latitudes

Here we present a first look at the Gross Primary Production (GPP) forecast skill levels achievable with a state-of-the-art subseasonal-to-seasonal (S2S) forecast system. Using NASA’s retrospective S2S ensemble forecast in conjunction with a terrestrial biosphere model, and using an independent, remote sensing-based dataset for validation, we demonstrate an ability to accurately forecast spring-summer carbon uptake at multi-month leads. Averaged across mid-and high latitudes of the Northern Hemisphere land, the GPP forecast initialized on January 1 produces statistically significant skill through summer. The skill achieved, however, is spatially variable, with some regions appearing to extract skill from accurate forecasts of snowpack removal and others extracting skill from the initialization of carbon and nitrogen states. Our results reveal some heretofore unexplored facets of climate predictability and provide a look at what might be possible with future S2S forecast systems that are fully integrated with biogeochemical cycles.

Eunjee Lee↗

Redshift evolution and covariances for joint lensing and clustering studies with DESI Y1

ABSTRACT Galaxy–galaxy lensing (GGL) and clustering measurements from the Dark Energy Spectroscopic Instrument Year 1 (DESI Y1) data set promise to yield unprecedented combined-probe tests of cosmology and the galaxy–halo connection. In such analyses, it is essential to identify and characterize all relevant statistical and systematic errors. We forecast the covariances of DESI Y1 GGL + clustering measurements and the systematic bias due to redshift evolution in the lens samples. Focusing on the projected clustering and GGL correlations, we compute a Gaussian analytical covariance, using a suite of N-body and lognormal simulations to characterize the effect of the survey footprint. Using the DESI one percent survey data, we measure the evolution of galaxy bias parameters for the DESI luminous red galaxy (LRG) and bright galaxy survey (BGS) samples. We find mild evolution in the LRGs in $0.4 < z < 0.8$, subdominant to the expected statistical errors. For BGS, we find less evolution for brighter absolute magnitude cuts, at the cost of reduced sample size. We find that for a redshift bin width $\Delta z = 0.1$, evolution effects on DESI Y1 GGL is negligible across all scales, all fiducial selection cuts, all fiducial redshift bins. Galaxy clustering is more sensitive to evolution due to the bias squared scaling. Nevertheless the redshift evolution effect is insignificant for clustering above the 1-halo scale of $0.1h^{-1}$ Mpc. For studies that wish to reliably access smaller scales, additional treatment of redshift evolution is likely needed. This study serves as a reference for GGL and clustering studies using the DESI Y1 sample.

79 ASTRONOMY AND ASTROPHYSICS↗

A hybrid data-driven and model-based approach for computationally efficient stochastic unit commitment and economic dispatch under wind and solar uncertainty

Stochastic unit commitment (UC) and economic dispatch (ED) are imperative in dealing with uncertainty in renewable forecast for power system operation and planning such that the overall expected production cost is minimized over the planning horizon. However, accurate calculation of the expected production cost requires assessment of a very large number of different scenarios of uncertain renewable resources, such as solar and wind, which is practically infeasible to simulate in real time. This article proposes a hybrid datadriven and physics-based model-predictive paradigm to efficiently solve for stochastic unit commitment and economic dispatch considering uncertainty in wind and solar power forecasts. Here, the novelty of the approach lies in decoupling the production cost estimation from the unit commitment and economic dispatch optimization problems under uncertainty without compromising on the fidelity of the solutions. A data-driven machine learning model is first developed to predict the mean optimal production cost. A physics-based inverse problem is then solved to get the stochastic UC and ED profiles from the expected cost. The presented approach considers, for the first time, solar uncertainty in UC/ED determination and enables efficient and accurate propagation of wind and solar uncertainty to estimate the statistics of the production cost. The effectiveness of the developed approach is demonstrated systematically on a stylized RTS-GMLC single-node system. The overall framework predicts the expected cost 62.5% more accurately than the existing state-of-the-art, on unforeseen days during the entire year, and yields, for the first time, the associated physically consistent UC and ED profiles. The solutions are also shown to be flexible in providing adequate daily reserves to address any statistical deviations from probabilistic power forecasts. The computational time associated with the presented method is only about 10 s compared to over 24 h needed for a conventional stochastic UC/ED determination under uncertainty on an Intel Core i9 processor with 32 GB of RAM.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Applications of Principled Search Methods in Climate Influences and Mechanisms

Forest and grass fires cause economic losses in the billions of dollars in the U.S. alone. In addition, boreal forests constitute a large carbon store; it has been estimated that, were no burning to occur, an additional 7 gigatons of carbon would be sequestered in boreal soils each century. Effective wildfire suppression requires anticipation of locales and times for which wildfire is most probable, preferably with a two to four week forecast, so that limited resources can be efficiently deployed. The United States Forest Service (USFS), and other experts and agencies have developed several measures of fire risk combining physical principles and expert judgment, and have used them in automated procedures for forecasting fire risk. Forecasting accuracies for some fire risk indices in combination with climate and other variables have been estimated for specific locations, with the value of fire risk index variables assessed by their statistical significance in regressions. In other cases, the MAPSS forecasts [23, 241 for example, forecasting accuracy has been estimated only by simulated data. We describe alternative forecasting methods that predict fire probability by locale and time using statistical or machine learning procedures trained on historical data, and we give comparative assessments of their forecasting accuracy for one fire season year, April- October, 2003, for all U.S. Forest Service lands. Aside from providing an accuracy baseline for other forecasting methods, the results illustrate the interdependence between the statistical significance of prediction variables and the forecasting method used.

Glymour, Clark↗

Peak Wind Forecasts for the Launch-Critical Wind Towers on Kennedy Space Center/Cape Canaveral Air Force Station, Phase IV

This final report describes the development of a peak wind forecast tool to assist forecasters in determining the probability of violating launch commit criteria (LCC) at Kennedy Space Center (KSC) and Cape Canaveral Air Force Station (CCAFS). The peak winds arc an important forecast clement for both the Space Shuttle and Expendable Launch Vehicle (ELV) programs. The LCC define specific peak wind thresholds for each launch operation that cannot be exceeded in order to ensure the safety of the vehicle. The 45th Weather Squadron (45 WS) has found that peak winds are a challenging parameter to forecast, particularly in the cool season months of October through April. Based on the importance of forecasting peak winds, the 45 WS tasked the Applied Meteorology Unit (AMU) to update the statistics in the current peak-wind forecast tool to assist in forecasting LCC violations. The tool includes onshore and offshore flow climatologies of the 5-minute mean and peak winds and probability distributions of the peak winds as a function of the 5-minute mean wind speeds.

Crawford, Winifred↗

Review of Onsite Temperature and Solar Forecasting Models to Enable Better Building Design and Operations

Advanced building controls and energy optimization for new constructions and retrofits rely on accurate weather data. Traditionally, most studies utilize airport weather information as the decision inputs. However, most buildings are in environments that are quite different than those at the airport miles away. Tree cover, adjacent buildings, and micro-climate effects caused by the larger surrounding area can all yield deviations in air temperature, humidity, solar irradiance, and wind that are large enough to influence design and operation decisions. In order to overcome this challenge, there are many prior studies on developing weather forecasting algorithms from micro-to meso-scales. Additionally, this paper reviews and complies knowledge on common weather data resources, data processing methodologies and forecasting techniques of weather information. Commonly used statistical, machine learning and physical-based models are discussed and presented as two major categories: deterministic forecasting and probabilistic forecasting. Finally, evaluation metrics for forecasting errors are listed and discussed.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

The GEOS Ozone Data Assimilation System: Specification of Error Statistics

A global three-dimensional ozone data assimilation system has been developed at the Data Assimilation Office of the NASA/Goddard Space Flight Center. The Total Ozone Mapping Spectrometer (TOMS) total ozone and the Solar Backscatter Ultraviolet (SBUV) or (SBUV/2) partial ozone profile observations are assimilated. The assimilation, into an off-line ozone transport model, is done using the global Physical-space Statistical Analysis Scheme (PSAS). This system became operational in December 1999. A detailed description of the statistical analysis scheme, and in particular, the forecast and observation error covariance models is given. A new global anisotropic horizontal forecast error correlation model accounts for a varying distribution of observations with latitude. Correlations are largest in the zonal direction in the tropics where data is sparse. Forecast error variance model is proportional to the ozone field. The forecast error covariance parameters were determined by maximum likelihood estimation. The error covariance models are validated using x squared statistics. The analyzed ozone fields in the winter 1992 are validated against independent observations from ozone sondes and HALOE. There is better than 10% agreement between mean Halogen Occultation Experiment (HALOE) and analysis fields between 70 and 0.2 hPa. The global root-mean-square (RMS) difference between TOMS observed and forecast values is less than 4%. The global RMS difference between SBUV observed and analyzed ozone between 50 and 3 hPa is less than 15%.

Stajner, Ivanka↗

Systematic Effects in Galaxy-Galaxy Lensing with DESI

The Dark Energy Spectroscopic Instrument (DESI) survey will measure spectroscopic redshifts for millions of galaxies across roughly $14,000 \, \mathrm{deg}^2$ of the sky. Cross-correlating targets in the DESI survey with complementary imaging surveys allows us to measure and analyze shear distortions caused by gravitational lensing in unprecedented detail. In this work, we analyze a series of mock catalogs with ray-traced gravitational lensing and increasing sophistication to estimate systematic effects on galaxy-galaxy lensing estimators such as the tangential shear $\gamma_{\mathrm{t}}$ and the excess surface density $\Delta\Sigma$ . We employ mock catalogs tailored to the specific imaging surveys overlapping with the DESI survey: the Dark Energy Survey (DES), the Hyper Suprime-Cam (HSC) survey, and the Kilo-Degree Survey (KiDS). Among others, we find that fiber incompleteness can have significant effects on galaxy-galaxy lensing estimators but can be corrected effectively by up-weighting DESI targets with fibers by the inverse of the fiber assignment probability. Similarly, we show that intrinsic alignment and lens magnification are expected to be statistically significant given the precision forecasted for the DESI year-1 data set. Our study informs several analysis choices for upcoming cross-correlation studies of DESI with DES, HSC, and KiDS.

79 ASTRONOMY AND ASTROPHYSICS↗

A study of Stormsat interactive data analysis facility

Stormsat, a third generation geosynchronous satellite intended for the investigation and forecasting of severe local storms, tropical cyclones and other mesoscale phenomena, could be launched as early as 1982. The paper outlines the interactive data processing and analysis functions which must be performed and presents a possible configuration capable of achieving the demonstration forecast objectives. Important features of the data processing system include (1) interactive extraction of wind, temperature and moisture profiles, cloud structure and precipitation from Stormsat data, and (2) numerical and statistical modeling necessary to provide demonstration forecasts.

Hasler, A. F.↗

Predicting dangerous ocean waves with spaceborne synthetic aperture radar

It is pointed out that catastrophes, related to the occurrence of strong winds and large ocean waves, can consume more lives and property than most naval battles. The generation of waves by wind are considered, Pierson et al. (1955) have incorporated statistical concepts into a wave forecast model. The concept of an 'ocean wave spectrum' was introduced, with the wind acting independently on each Fourier component. However, even after 30 years of research and debate, the generation, propagation, and dissipation of the spectrum under arbitrary conditions continue to be controversial. It has now been found that spaceborne SAR has a surprising ability to precisely monitor spatially evolving wind and wave fields. Approaches to overcome certain weaknesses of the SAR method are discussed, taking into account the second Shuttle Imaging Radar experiment, and a possible long-term solution provided by Spectrasat. Spectrasat should be a low-altitude (200 to 250 km) satellite with active drag compensation.

Beal, R. C.↗

Lightning Initiation Forecasting: An Operational Dual-Polarimetric Radar Technique

The objective of this NASA MSFC and NOAA CSTAR funded study is to develop and test operational forecast algorithms for the prediction of lightning initiation utilizing the C-band dual-polarimetric radar, UAHuntsville's Advanced Radar for Meteorological and Operational Research (ARMOR). Although there is a rich research history of radar signatures associated with lightning initiation, few studies have utilized dual-polarimetric radar signatures (e.g., Z(sub dr) columns) and capabilities (e.g., fuzzy-logic particle identification [PID] of precipitation ice) in an operational algorithm for first flash forecasting. The specific goal of this study is to develop and test polarimetric techniques that enhance the performance of current operational radar reflectivity based first flash algorithms. Improving lightning watch and warning performance will positively impact personnel safety in both work and leisure environments. Advanced warnings can provide space shuttle launch managers time to respond appropriately to secure equipment and personnel, while they can also provide appropriate warnings for spectators and players of leisure sporting events to seek safe shelter. Through the analysis of eight case dates, consisting of 35 pulse-type thunderstorms and 20 non-thunderstorm case studies, lightning initiation forecast techniques were developed and tested. The hypothesis is that the additional dual-polarimetric information could potentially reduce false alarms while maintaining high probability of detection and increasing lead-time for the prediction of the first lightning flash relative to reflectivity-only based techniques. To test the hypothesis, various physically-based techniques using polarimetric variables and/or PID categories, which are strongly correlated to initial storm electrification (e.g., large precipitation ice production via drop freezing), were benchmarked against the operational reflectivity-only based approaches to find the best compromise between forecast skill and lead-time. Forecast skill is determined by statistical analysis of probability of detection (POD), false alarm ratio (FAR), Operational Utility Index (OUI), and critical success index (CSI).

Woodard, Crystal J.↗