Engineering PapersSearch

SEARCH · Engineering Papers

Results for “estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

CONUS-wide Projected Flood Frequency and Uncertainty Estimates, Version 1.0

This dataset presents a large-ensemble of CONUS-wide projected flood frequency and uncertainty estimates across ~2.7 million NHDPlusV2 river reaches over the CONUS. The framework producing this dataset leverages a multi-model, uncertainty-aware modeling framework that allows evaluating shifts in flood frequences at the stream reach level across the CONUS. CONUS-wide ensemble streamflow projections generated from hydrologic simulations driven by downscaled and bias-corrected Coupled Model Intercomparison Project Phase 6 (CMIP6) outputs are used to derive these flood frequency and uncertainty estimates over the period 1980 - 2099. A spatially consistent regional L-moment algorithm is applied across clusters defined by the US Hydrologic Unit Code Subregions (HUC4s and HUC8s) and NHDPlusV2 stream orders to estimate flood frequencies. The dataset also includes at-site based flood estimates that allow for the comparison between local and regional approach-based estimates, assess projected changes, and characterize their uncertainties. For more reliable estimation of rare flood frequencies such as 500 and 1000-year return periods, super-ensemble based estimates are also included in the dataset. This dataset is derived to support the "Impact-Informed Dam Safety Risk Assessment for Securing Hydropower Assests" project for the US Department of Energy (DOE) Hydropower and Hydrokinetic Office (H2O). For further details, refer to Kao et al. (2022), Ghimire et al. (2023), Ghimire et al. (2025), and Hosking and Wallis (1997).

Ghimire, Ganesh [ORNL] (ORCID:0000000242843941)

Preventive Power Outage Estimation Based on a Novel Scenario Clustering Strategy

The increasing occurrence of extreme weather events is challenging power grid operation. For extreme weather events, the system operator is responsible for estimating the power outages and scheduling the restoration resources. This paper proposes an outage evaluation framework to identify the possible unserved load profiles, vulnerable areas, and mobile energy adequacy. The outputs of an outage prediction model tool are used to generate numerous faulted line scenarios. Next, each scenario's nodal unserved load profile is obtained by solving a three-phase restoration model that considers repair crews and mobile energy resources (MERs). Then, a novel scenario clustering strategy is developed to cluster the unserved load profiles into multiple representative profiles which the system operator can focus on. Finally, case studies on a distribution system evaluate the damage caused by an extreme weather event and verify the effectiveness of the proposed scenario clustering strategy.

MATHEMATICS AND COMPUTING,POWER TRANSMISSION AND D

Data for Clumping Index Estimation With 30°-tilted Cameras in Row Crops: Evaluation of Methods and Segment Size Effects

The clumping index (CI) quantifies the spatial distribution of foliage elements and is essential for accurately estimating the plant area index (PAI), canopy radiative transfer, and photosynthesis. Traditionally, the finite-length averaging method (LX), the gap size distribution method (CC), and a combined approach of CC and LX (CLX) have been applied to instruments like TRAC and digital hemispherical photography to estimate CI. However, a comprehensive evaluation of these methods in row crops remains limited, especially regarding the influence of segment size on CI. Meanwhile, digital cameras offer a cost-effective and user-friendly solution for canopy measurements in row crops, yet their application in this context remains underexplored. In this study, we employed a new approach using a 30°-tilted digital camera to estimate CI in corn and soybean fields, applying the LX, CC, and CLX methods. We systematically assessed the performance of these three methods by combining field measurements in real-world fields with simulations using the LESS 3D radiative transfer model. Our results showed that CLX applied to the whole image and 45° segment offered accurate estimation of CI (bias within ±0.1, RMSE < 0.2) and PAI (bias within ±0.4, RMSE < 1) in real-world fields and LESS simulations. The accuracy of the LX method was highly sensitive to segment size, with the best performance observed at the 15° segment (PAI bias within ±0.4). In contrast, the CC method remained stable across different segment sizes, and its performance was generally comparable to that of LX, except at the 15° segment. Across view zenith angles, CI derived from CC generally showed a continuous increase, while those from LX and CLX followed a rising trend at small zenith angles but began to decline at 68°, likely due to an increasing proportion of no-gap segments. Seasonally, LX tended to show decreasing CI during early growth stages but increased as the canopy matured, whereas CC and CLX showed gradually increasing CI before plateauing at peak PAI. The 30°-tilted camera effectively captured CI variations across different angles and growth stages, making it a practical and robust instrument for row crop canopy structure analysis. Applying these CI methods to digital cameras offers a low-cost and accessible CI estimation alternative, improving canopy structure monitoring accuracy in row crops.

Modeling

Estimators and Fusers for Fiber Delay Estimation Using Environmental Measurements

The properties of deployed network fiber are affected by environmental factors due to their exposure to the elements. Particularly for quantum networks, the resultant delay variations may have significant impacts due to the extreme sensitivity of synchronization, coincidence counting, and other critical operations. In this paper, the delays of 15 km aerial-inground fiber connections are measured, and effects due to temperature, humidity and wind speed are analyzed over multiple periods spanning four seasons of a year. Machine learning methods are first utilized to reveal surprisingly pronounced effects of humidity on the delay, in addition to the expected temperature and its seasonal variations. Estimator and fusion methods are developed to estimate the delay using temperature, humidity and wind speed measurements, by utilizing smooth Gaussian Process Regression (GPR) and nonsmooth Ensemble of Trees (EOT) methods. Measurements from winter and summer periods are temporally fused using twelve different methods, and eight methods provide estimates for the delay throughout the year with median test errors under 1.28%. The results reveal distinct temperature-humidity trends across the seasons, and the ability of estimator and temporal fusion methods to exploit them for estimating the delay. These results constitute a case study of machine learning analytical results, wherein generalization equations explain the performance of various estimator and fuser methods.

Rao, Nageswara [ORNL] (ORCID:0000000234085941)

Diesel Fuel Consumption in Prominent U.S. Open-Pit Mines: Site-Level Estimates

This report presents a comprehensive framework for estimating diesel fuel consumption and prices at open-pit mines in the United States. The framework includes transparent methods for calculating site-level diesel energy use when direct reporting is unavailable, and a structured confidence evaluation for each method. The framework is demonstrated to estimate current diesel consumption at 21 open-pit mines in the United States. Initial findings support ongoing efforts to strengthen the competitiveness and security of the U.S. industrial base by supporting data-driven supply chain analysis and decision-making, improved transparency in mining sector energy use, and targeted deployment of energy innovation and cost-reduction strategies. Future updates to the dataset—coupled with expanded data transparency and method validation—will help ensure that the findings remain relevant as the sector continues to evolve.

02 PETROLEUM

Source shape estimation for neutron imaging systems using convolutional neural networks

Neutron imaging systems are important diagnostic tools for characterizing the physics of inertial confinement fusion reactions at the National Ignition Facility (NIF). In particular, neutron images give diagnostic information on the size, symmetry, and shape of the fusion hot spot and surrounding cold fuel. Images are formed via collection of neutron flux from the source using a system of aperture arrays and scintillator-based detectors. Currently, reconstruction of fusion source geometry from the collected neutron images is accomplished by solving a computationally intensive maximum likelihood estimation problem via expectation maximization. In contrast, it is often useful to have simple representations of the overall source geometry that can be computed quickly. In this work, we develop convolutional neural networks (CNNs) to reconstruct the outer contours of simple source geometries. We compare the performance of the CNN for penumbral and pinhole data and provide experimental demonstrations of our methods on both non-noisy and noisy data.

Machine learning, neutron imaging, source reconstr

Maximum a posteriori Ly α estimator (MAPLE): band power and covariance estimation of the 3D Ly α forest power spectrum

We present a novel maximum a posteriori estimator to jointly estimate band powers and the covariance of the three-dimensional power spectrum (P3D) of Ly $\alpha$ forest flux fluctuations, called MAPLE. Our Wiener-filter based algorithm reconstructs a window-deconvolved P3D in the presence of complex survey geometries typical for Ly $\alpha$ surveys that are sparsely sampled transverse to and densely sampled along the line of sight. We demonstrate our method on idealized Gaussian random fields with two selection functions: (i) a sparse sampling of 30 background sources per square degree designed to emulate the current Dark Energy Spectroscopic Instrument; (ii) a dense sampling of 900 background sources per square degree emulating the upcoming Prime Focus Spectrograph Galaxy Evolution Survey. Our proof-of-principle shows promise, especially since the algorithm can be extended to marginalize jointly over nuisance parameters and contaminants, i.e. offsets introduced by continuum fitting. Our code is implemented in JAX and is publicly available on GitHub.

79 ASTRONOMY AND ASTROPHYSICS

Data for A Hybrid Biophysical-Machine Learning Framework for Diurnal Surface Energy Flux Estimation Using Proximal Sensing

Thermal infrared-based remote sensing of surface energy fluxes has traditionally relied on high spatial resolution satellite data with revisit frequencies on the order of weeks. In this study, we evaluate a biophysics-based analytical surface energy balance model for predicting latent energy (LE) and sensible heat (H) fluxes using proximal sensing observations. The Surface Temperature Initiated Closure (STIC1.2) model has been extensively validated across a wide range of spatial and temporal scales using various satellite-derived thermal infrared data sets. Here we extend this validation by applying STIC at sub-hourly temporal resolution over multiple growing seasons for four distinct agricultural systems. We further develop and evaluate novel STIC variants that incorporate machine learning (ML) techniques to eliminate the need for surface energy balance observations, specifically net radiation and soil heat flux, thereby enhancing model applicability in data-sparse settings. The integration of a ML component to estimate surface available energy is shown to have strong predictive performance for both LE (R2 = 0.81–0.94) and H (R2 = 0.46–0.72) across all agricultural systems examined here, demonstrating the potential of hybrid biophysical-machine learning approaches for surface energy balance modeling with minimal data requirements. This study concludes with a novel application of explainable machine learning (exML) to diagnose sources of model error. This exML framework attributes residual prediction errors to both model input variables and environmental drivers not explicitly included in the simulation experiments. This approach provides a new pathway for improving model design and integrating previously overlooked yet influential variables into future model iterations.

AI/ML

Synthesis of ARM User Facility Surface Rainfall Datasets to Construct a Best Estimate Value Added Product (PrecipBE)

Surface precipitation measurements are essential for Earth system model (ESM) evaluation and understanding cloud processes. An ever-growing need for robust, temporally evolving, and easy-to-use statistical datasets provides motivation for a baseline ground-based precipitation properties data product. The U.S. Department of Energy Atmospheric Radiation Measurement (ARM) user facility operates an extensive suite of precipitation instruments with various sensitivities and operating mechanisms, which render the decision of which instrument to use based on one or more fixed thresholds challenging and prone to errors and bias. Using a long-term instrument inter-comparison from a unique per-precipitation event perspective, rather than instantaneous sample comparison, we demonstrate that ARM rainfall-measuring instruments are generally consistent with each other at the statistical level. Inter-instrument deviations at the single event level can be large, especially for specific rainfall event properties such as maximum precipitation rates. A machine-learning (ML) analysis using a random forest regressor indicates that in some cases, depending on instrument, local site climatology, and/or specific deployment configuration, certain atmospheric state variables influence the measured quantities in an unpredictable manner. Thus, a-priori weighting of different instruments does not necessarily lead to more accurate and less biased synthesis of instrument data. These results motivate the design of the ARM precipitation best-estimate (PrecipBE) value-added product, which incorporates all valid precipitation data while considering data quality and other instrument limitations. PrecipBE consists of time series and tabular statistics datasets in an easy-to-use and insightful per-precipitation event format. It provides a large set of precipitation event properties supplemented with ancillary data from ARM datasets that correspond to the detected precipitation events. We describe the PrecipBE algorithm and demonstrate its use via the examination of a single-day output as well as a long-term trend analysis of precipitation events at the ARM Southern Great Plains (SGP) site, covering more than 30 years of data. The trend analysis tentatively suggests a long-term temporal tendency for mainly shorter and less intense precipitation events at the SGP site, but a long-term increase in annual rainfall by more than 36 mm (5 %) per decade. This rainfall trend is catalyzed primarily by more extreme event properties of relatively rare, intense precipitation events, with event total and 1 min maximum precipitation rate at a 1 year timeframe increasing up to 5 mm and 9 mm h −1 (several percent) per decade, respectively. While the currently available PrecipBE datasets (at https://adc.arm.gov/discovery/, last access: 8 December 2025) cover rainfall from multiple ARM deployments up to March 2025, PrecipBE is planned to be expanded to include solid-phase precipitation and will soon become an operational product with a several-day lag from real-time. We invite the ARM user community to leverage this new product and welcome user feedback to further enhance the dataset.

Silber, Israel [Pacific Northwest National Laborat

Sampling Size Optimization for Bioburden Density Estimation in Planetary Protection

Planetary protection (PP) is a discipline that focuses on minimizing the biological contamination of spacecraft to ensure compliance with international policy. Precise estimation of bioburden - the total number of microbes in or on spacecraft hardware – and the bioburden density are of utmost importance for PP. Such estimation is the way concordance with requirements is demonstrated, and it is critical for quantifying the potential risk of inadvertently contaminating other planetary bodies. Although a suite of molecular techniques have been used to thoroughly characterize and profile the microbiome of various cleanroom environments and spacecraft, the gold standard remains the physical enumeration of microbes via culturing of samples directly taken from spacecraft and associated surfaces. However, due to technical, budgetary, and programmatic constraints, only a manageable portion (around 10%) of the entire spacecraft surface is directly sampled with cotton swabs or wipes. To generate the bioburden current best estimate (CBE) for components not directly verifiable, the accepted approach is to apply a NASA-defined bioburden estimate based on the components’ manufacturing or assembly environment. This approach utilizes a prespecified bioburden density estimation that applies a maximum value across the total surface area of the specified component. For hardware components that underwent similar assembly processes, an implied bioburden is adopted for all components, based on a direct verification of a representative component within the same lot. Once all components have a CBE, the bioburden estimates are generated. In previous publication [ 1], we have shown that statistical risks quantifying the accuracy of the estimates for sampled, prespecified, and implied components can be derived and ranked. For mean squared error (MSE) function, the risks are available analytically and hence a cost function can be obtained to optimize the risks with respect to the sampling area and sampling cost. Since the sampling area and sampling cost are two complimentary variables, their sum will have a well-defined minimum. This paper presents the multivariate optimization of the integrated risk of an empirical Bayes estimator to determine the optimal sampling schedule for a given number of components. It is assumed that given a number of components, N, the bioburden density for each component can either be sampled, implied, or prespecified. The multivariate optimization searches through different options to sample, imply or prespecify the bioburden density for a component, and account for the component’s surface area and cost of sampling. The idea of the optimization is based on the observation that the statistical risk of using an estimator is a monotonically decreasing function of the sampled area. The larger the sampled area, the lower the risk of using the estimator as the estimator becomes more and more accurate as the sampling area increases. On the other hand, the cost of sampling is monotonically increasing as the sampled surface grows. This makes the risk and total cost of sampling complimentary variables which can be counterbalanced to achieve an optimal overall value with respect to the sampled surface. In this paper, the integrated risk has been used to quantify the accuracy of the estimator. This risk has been selected because it depends on neither the true value of the parameter nor on the collected data. The cost of each sample was also available to obtain the total cost of sampling of N components. The paper will present the results based on computer-simulated data as well as the data collected during the InSight mission. The computer-simulated data have N components with randomly generated total areas and each component assigned to one of the three categories according to the method of estimating of bioburden density: sampled, implied, or prespecified. The cost of sampling is also available. The cost of sampling is estimated based on a cost model provided by the planetary protection group at JPL. For this paper, the overall cost was assumed to be a linear function of exposure. The optimization process finds the allocation of the components to the three categories that minimizes the tradeoff between integrated risk and total cost. For the InSight data, a set of components is selected representing all three categories, and optimization is performed to determine if the performed allocation was optimal or if a better allocation could have been obtained. To the best of our knowledge, this work is the first attempt not only perform an accurate estimation of bioburden density but also do it in an optimal way.

97 - MATHEMATICS AND COMPUTING

Comparison of removal and spatial mark‐resight models for estimating wild pig density

Density estimation is critical to effectively manage invasive species and elucidate areas of highest concern. For wild pigs (Sus scrofa), the ability to estimate density is complicated because of their variable home range sizes and social structure. Common methods for estimating density (e.g., mark-recapture) may be unsuitable in management applications because additional data needs to be collected before and after management. Removal models offer a suitable alternative to estimate density changes following management and can be applied broadly across areas where management of wild pigs is ongoing. We collected wild pig removal and camera trap data from 25 private properties ranging in size from approximately 0.5 km 2 to 95 km 2 across 3 ecoregions in South Carolina, USA, from 2020–2023. We compared factors affecting consistency and precision of property-level density estimates between removal and spatial mark-resight (SMR) models. In general, excluding 1 large outlier, density estimates from removal models were between 0.60 and 15.85 wild pigs/km 2 (median = 5.34) with a median coefficient of variation (CV) of 0.76 and 95% confidence intervals for the CV between 0.70 and 0.94. Similarly, excluding 1 large outlier, density estimates from SMR were between 0.22 and 30.97 wild pigs/km 2 (median = 5.48) with a median CV of 0.39 and 95% confidence intervals for the CV between 0.38 and 1.20. We found the precision of removal models was affected primarily by the number of wild pigs dispatched in the removal period (3 months) and the ecoregion in which they were removed. None of the covariates, including the number of recaptures (a corresponding measure of sample size), influenced precision of the SMR models, although recaptures did influence the density estimates. At the individual property level, density estimates from our 2 estimators were dissimilar from each other in approximately 80% of instances, although none of the covariates we examined influenced dissimilarity. Our results provide unique insight into how sample size affects density estimates using 2 common methods and into novel SMR models that incorporate both marked and unmarked detections. In addition, the density estimates in this study can be used as a reference for wild pig densities in common land cover types throughout the southeastern United States.

60 APPLIED LIFE SCIENCES

Evaluation of top-down and bottom-up global terrestrial respiration estimates and their mismatch with model simulations

Terrestrial respiration is one of the most poorly understood processes in the global carbon cycle, making respiration predictions uncertain. However, expanding observations and machine learning approaches have led to a proliferation of estimates. We compiled total ecosystem and heterotrophic respiration estimates derived from top-down atmospheric inversions and bottom-up upscaling of ecosystem observations and compared them with dynamic vegetation models (DGVM) simulations over the 1980-2020 period. Our analysis revealed a convergence in mean annual global total ecosystem respiration estimates between top-down 97.1 (± SD 6.8) PgC yr-1 and bottom-up 98.5 (+/-13.4) PgC yr-1, which were both significantly lower than the ensemble mean from DGVMs estimates 133.7 (±4.7) PgC yr-1. We also found similar temporal trends between top-down estimates with a mean of 0.075 (±0.05) PgC yr-2, and bottom-up estimates of 0.05 (±0.05) PgC yr-2, which were 5 to 7 times smaller than the ensemble mean trend of 0.34 PgC yr-2 simulated by DGVMs. Global heterotrophic respiration showed much less agreement, ranging from top-down estimates of 42.7 (±4.0) PgC yr-1 to bottom-up estimates of 51.5 (±4.0) PgC yr-1 and a significantly larger ensemble model mean estimate of 60.8 PgC yr-1 (±1.9). The temporal trends in observation-based bottom-up estimates of heterotrophic respiration of 0.03 PgC yr-2 were five times lower than the model ensemble mean 0.15 PgC yr-2. Large regional disagreements in heterotrophic respiration estimates and simulations were evident in tropical and boreal latitudes. Therefore, improved regional and heterotrophic respiration estimates are necessary to reduce uncertainties regarding the future vulnerability of soil carbon.

Ballantyne, Ashley

Raccoon density estimation from camera traps for raccoon rabies management

Abstract Density estimation for unmarked animals is particularly challenging, yet density estimates are often necessary for effective wildlife management. Raccoons ( Procyon lotor ) are the primary terrestrial wildlife reservoir for Lyssavirus rabies within the United States. The raccoon rabies variant (RRVV) is actively managed at landscape scales using oral rabies vaccination (ORV) within the eastern United States. To effectively manage RRVV, it is important to know the density of raccoons to appropriately scale the density of ORV baits distributed on the landscape. We compared methods to estimate raccoon densities from camera‐trap data versus more intensive capture‐mark‐recapture (CMR) estimates across 2 land cover types (upland pine and bottomland hardwood) in the southeastern United States during 2019 and 2020. We evaluated the effect of alternative camera configurations and durations of camera trapping on density estimates and used an N‐mixture model to estimate raccoon densities, including covariates on abundance and detection. We further compared different methods of scaling camera‐based counts, with the maximum number of raccoons seen on any given image within a day best explaining density. Camera‐trap density estimates were moderately correlated with CMR estimates ( r = 0.56). However, densities from camera‐trap data were more reliable when classifying category of density as an index used to inform management (83% correct when compared to CMR estimates), although the densities in our study fell into the 2 lowest density classes only. Using more cameras reduced bias and uncertainty around density estimates; however, if ≤6 camera traps were used at a site, a line transect approach proved less biased than a grid design. Camera trapping should be conducted for at least 3 weeks for more accurate estimates of raccoon population density in our study area (<5% bias). We show that camera‐trap data can be used to assign raccoon densities to management‐relevant density index bins, but more studies are needed to ensure reliability across a greater range of environmental conditions and raccoon densities.

Davis, Amy J.

A Comparison of Pre‐Construction and Operational Wake Loss Estimates for Land‐Based Wind Plants

The overall bias between pre‐construction energy yield assessment (EYA) estimates of wind plant energy production and the achieved operational production is improving in the wind industry, but uncertainty remains high for individual wind plants. Wake effects within wind plants are one of the largest sources of energy loss considered in the EYA process, and previous work shows wake loss estimates to be a major source of disagreement among wind energy consultants who perform EYAs. To better understand the accuracy of wake loss predictions, we compare overall operational wake loss estimates based on supervisory control and data acquisition data to pre‐construction estimates provided by six wind energy consultants for five land‐based wind plants in North America. By augmenting existing approaches for quantifying operational wake losses, we estimate wake losses during the period of record for which operational data are available as well as the expected long‐term wake losses, based on historical reanalysis weather data, to which the EYA estimates are compared. To account for power variations at different turbine locations caused by terrain‐induced wind resource heterogeneity, we correct the operational wake loss estimates using predicted freestream wind speed variations from the Wind Systems Engineering Reynolds‐averaged Navier–Stokes (RANS) tool. We identify long‐term corrected operational wake losses between 1.9% and 6.4% for the five plants, with a mean loss of 4%. For the project deemed most acceptable for operational wake loss assessment, which is located in the simplest terrain and isolated from neighboring plants, the mean EYA wake loss estimate is within 0.7 percentage points of the operational value of 6.4%. For most of the remaining plants, results suggest that wake losses are generally overpredicted by 2.6–6.3 percentage points. However, operational wake losses may be underestimated for many of these projects because of spatial wind resource variations not captured by the RANS model, external wake effects that are unaccounted for in the estimation process, and wind plant blockage effects. To better understand factors that contribute to the observed wake losses, we investigate operational wake losses as a function of wind direction and wind speed. As expected, wake losses are generally concentrated near wind directions that are aligned with rows of closely spaced turbines and at below‐rated wind speeds; however, for some projects, the energy produced by the wind plant exceeds the estimated potential energy of the plant without wake interactions for certain wind directions and wind speeds, suggesting inaccurate assumptions in the wake loss estimation method for those plants. Lastly, we compare predicted and operational wake losses for individual wind turbines, finding that even when overall wake losses are predicted accurately, large uncertainty exists at the turbine level.

17 WIND ENERGY

Advancements and opportunities to improve bottom–up estimates of global wetland methane emissions

Wetlands are the single largest natural source of atmospheric methane (CH 4 ), contributing approximately 30% of total surface CH 4 emissions, and they have been identified as the largest source of uncertainty in the global CH 4 budget based on the most recent Global Carbon Project CH 4 report. High uncertainties in the bottom–up estimates of wetland CH 4 emissions pose significant challenges for accurately understanding their spatiotemporal variations, and for the scientific community to monitor wetland CH 4 emissions from space. In fact, there are large disagreements between bottom–up estimates versus top–down estimates inferred from inversion of atmospheric CH 4 concentrations. To address these critical gaps, we review recent development, validation, and applications of bottom–up estimates of global wetland CH 4 emissions, as well as how they are used in top–down inversions. These bottom–up estimates, using (1) empirical biogeochemical modeling (e.g. WetCHARTs: 125–208 TgCH 4 yr -1 ); (2) process-based biogeochemical modeling (e.g. WETCHIMP: 190 ± 39 TgCH 4 yr -1 ); and (3) data-driven machine learning approach (e.g. UpCH4: 146 ± 43 TgCH 4 yr -1 ). Bottom–up estimates are subject to significant uncertainties (~80 Tg CH 4 yr -1 ), and the ranges of different estimates do not overlap, further amplifying the overall uncertainty when combining multiple data products. These substantial uncertainties highlight gaps in our understanding of wetland CH 4 biogeochemistry and wetland inundation dynamics. Major tropical and arctic wetland complexes are regional hotspots of CH 4 emissions. However, the scarcity of satellite data over the tropics and northern high latitudes offer limited information for top–down inversions to improve bottom–up estimates. Recent advances in surface measurements of CH 4 fluxes (e.g. FLUXNET-CH 4 ) across a wide range of ecosystems including bogs, fens, marshes, and forest swamps provide an unprecedented opportunity to improve existing bottom–up estimates of wetland CH 4 estimates. We suggest that continuous long-term surface measurements at representative wetlands, high fidelity wetland mapping, combined with an appropriate modeling framework, will be needed to significantly improve global estimates of wetland CH 4 emissions. There is also a pressing unmet need for fine-resolution and high-precision satellite CH 4 observations directed at wetlands.

54 ENVIRONMENTAL SCIENCES

The Poisson tensor completion non-parametric differential entropy estimator

We introduce the Poisson tensor completion (PTC) estimator, a non-parametric differential entropy estimator. The PTC estimator leverages inter-sample relationships to compute a low-rank Poisson tensor decomposition of the frequency histogram. Our crucial observation is that the histogram bins are an instance of a space partitioning of counts and thus can be identified with a spatial Poisson process. The Poisson tensor decomposition leads to a completion of the intensity measure over all bins—including those containing few to no samples—and leads to our proposed PTC differential entropy estimator. A Poisson tensor decomposition models the underlying distribution of the count data and guarantees non-negative estimated values and so can be safely used directly in entropy estimation. Our estimator is the first tensor-based estimator that exploits the underlying spatial Poisson process related to the histogram explicitly when estimating the probability density with low-rank tensor decompositions for the purpose of tensor completion. Furthermore, we demonstrate that our PTC estimator is a substantial improvement over standard histogram-based estimators for sub-Gaussian probability distributions because of the concentration of norm phenomenon.

42 ENGINEERING