Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “likelihood function”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Efficient high-dimensional variational data assimilation with machine-learned reduced-order models

Abstract. Data assimilation (DA) in geophysical sciences remains the cornerstone of robust forecasts from numerical models. Indeed, DA plays a crucial role in the quality of numerical weather prediction and is a crucial building block that has allowed dramatic improvements in weather forecasting over the past few decades. DA is commonly framed in a variational setting, where one solves an optimization problem within a Bayesian formulation using raw model forecasts as a prior and observations as likelihood. This leads to a DA objective function that needs to be minimized, where the decision variables are the initial conditions specified to the model. In traditional DA, the forward model is numerically and computationally expensive. Here we replace the forward model with a low-dimensional, data-driven, and differentiable emulator. Consequently, gradients of our DA objective function with respect to the decision variables are obtained rapidly via automatic differentiation. We demonstrate our approach by performing an emulator-assisted DA forecast of geopotential height. Our results indicate that emulator-assisted DA is faster than traditional equation-based DA forecasts by 4 orders of magnitude, allowing computations to be performed on a workstation rather than a dedicated high-performance computer. In addition, we describe accuracy benefits of emulator-assisted DA when compared to simply using the emulator for forecasting (i.e., without DA). Our overall formulation is denoted AIEADA (Artificial Intelligence Emulator-Assisted Data Assimilation).

58 GEOSCIENCES↗

Characterization of Cloud Water-Content Distribution

The development of realistic cloud parameterizations for climate models requires accurate characterizations of subgrid distributions of thermodynamic variables. To this end, a software tool was developed to characterize cloud water-content distributions in climate-model sub-grid scales. This software characterizes distributions of cloud water content with respect to cloud phase, cloud type, precipitation occurrence, and geo-location using CloudSat radar measurements. It uses a statistical method called maximum likelihood estimation to estimate the probability density function of the cloud water content.

Lee, Seungwon↗

Frequency Domain Quasi Maximum Likelihood Identification of Low Order Aeroservoelastic Models from Flight-Test Data

In this paper a quasi maximum likelihood method for estimating a low order model of a flexible vehicle has been developed and demonstrated. The quasi maximum likelihood method uses a large number of sensors to estimate the parameters and covariance in a method consistent with the maximum likelihood filter-error method. The cost function has been defined in a frequency domain to improve computational efficiency and to enable estimation of an unstable plant. The method has been demonstrated using flight data from the National Aeronautics and Space Administration X-56A Multi-Utility Technology Testbed flex wing tests, which include unstable flutter modes. The method was able to effectively estimate the frequency and damping of the dynamics of the aircraft and generate a transfer function that is useful for control system evaluations.

Jeffrey Ouellette↗

Deep Interacting Multiple Model Filtering

In this paper, a deep learning-based multiple model estimation framework is presented for the state estimation of hybrid dynamical systems from high dimensional observations such as camera images. A low dimensional vector which represents the measurement of the latent dynamical system and its corresponding variance are learned using a deep encoder neural network. An Interacting Multiple Model (IMM) filter is used to generate the latent state estimates and covariances using multiple dynamical models, which can be learned using backpropagation through time. The state estimates of the dynamical system and the corresponding covariance matrix are generated from the latent state estimates and covariance using a deep decoder neural network. The whole network is trained in an end-to-end manner using a loss function which minimizes the negative log-likelihood of the neural network parameters. Simulation results are presented using a 2D bouncing ball example and estimation error statistics are computed which demonstrates the accuracy and consistency of the estimation.

Ghananeel Rotithor↗

Two-stage model of radon-induced malignant lung tumors in rats: effects of cell killing

A two-stage stochastic model of carcinogenesis is used to analyze lung tumor incidence in 3750 rats exposed to varying regimens of radon carried on a constant-concentration uranium ore dust aerosol. New to this analysis is the parameterization of the model such that cell killing by the alpha particles could be included. The model contains parameters characterizing the rate of the first mutation, the net proliferation rate of initiated cells, the ratio of the rates of cell loss (cell killing plus differentiation) and cell division, and the lag time between the appearance of the first malignant cell and the tumor. Data analysis was by standard maximum likelihood estimation techniques. Results indicate that the rate of the first mutation is dependent on radon and consistent with in vitro rates measured experimentally, and that the rate of the second mutation is not dependent on radon. An initial sharp rise in the net proliferation rate of initiated cell was found with increasing exposure rate (denoted model I), which leads to an unrealistically high cell-killing coefficient. A second model (model II) was studied, in which the initial rise was attributed to promotion via a step function, implying that it is due not to radon but to the uranium ore dust. This model resulted in values for the cell-killing coefficient consistent with those found for in vitro cells. An "inverse dose-rate" effect is seen, i.e. an increase in the lifetime probability of tumor with a decrease in exposure rate. This is attributed in large part to promotion of intermediate lesions. Since model II is preferable on biological grounds (it yields a plausible cell-killing coefficient), such as uranium ore dust. This analysis presents evidence that a two-stage model describes the data adequately and generates hypotheses regarding the mechanism of radon-induced carcinogenesis.

NASA Program Radiation Health↗

Incorporating spatial context into statistical classification of multidimensional image data

Compound decision theory is employed to develop a general statistical model for classifying image data using spatial context. The classification algorithm developed from this model exploits the tendency of certain ground-cover classes to occur more frequently in some spatial contexts than in others. A key input to this contextural classifier is a quantitative characterization of this tendency: the context function. Several methods for estimating the context function are explored, and two complementary methods are recommended. The contextural classifier is shown to produce substantial improvements in classification accuracy compared to the accuracy produced by a non-contextural uniform-priors maximum likelihood classifier when these methods of estimating the context function are used. An approximate algorithm, which cuts computational requirements by over one-half, is presented. The search for an optimal implementation is furthered by an exploration of the relative merits of using spectral classes or information classes for classification and/or context function estimation.

Bauer, M. E.↗

A possible mass distribution of primordial black holes implied by LIGO-Virgo

The LIGO-Virgo Collaboration has so far detected around 90 black holes, some of which have masses larger than what were expected from the collapse of stars. The mass distribution of LIGO-Virgo black holes appears to have a peak at ~ 30 M ⊙ and two tails on the ends. By assuming that they all have a primordial origin, we analyze the GWTC-1 (O1&O2) and GWTC-2 (O3a) datasets by performing maximum likelihood estimation on a broken power law mass function f ( m ), with the result f ∝ m 1.2 for m < 35 M ⊙ and f ∝ m -4 for m > 35 M ⊙ . This appears to behave better than the popular log-normal mass function. Surprisingly, such a simple and unique distribution can be realized in our previously proposed mechanism of PBH formation, where the black holes are formed by vacuum bubbles that nucleate during inflation via quantum tunneling. Moreover, this mass distribution can also provide an explanation to supermassive black holes formed at high redshifts.

Astronomy & Astrophysics↗

Neural simulation-based inference of the neutron star equation of state directly from telescope spectra

Neutron stars provide a unique opportunity to study strongly interacting matter under extreme density conditions. The intricacies of matter inside neutron stars and their equation of state are not directly visible, but determine bulk properties, such as mass and radius, which affect the star's thermal X-ray emissions. However, the telescope spectra of these emissions are also affected by the stellar distance, hydrogen column, and effective surface temperature, which are not always well-constrained. Uncertainties on these nuisance parameters must be accounted for when making a robust estimation of the equation of state. In this study, we develop a novel methodology that, for the first time, can infer the full posterior distribution of both the equation of state and nuisance parameters directly from telescope observations. This method relies on the use of neural likelihood estimation, in which normalizing flows use samples of simulated telescope data to learn the likelihood of the neutron star spectra as a function of these parameters, coupled with Hamiltonian Monte Carlo methods to efficiently sample from the corresponding posterior distribution. Our approach surpasses the accuracy of previous methods, improves the interpretability of the results by providing access to the full posterior distribution, and naturally scales to a growing number of neutron star observations expected in the coming years.

79 ASTRONOMY AND ASTROPHYSICS↗

2D-FFTLog: efficient computation of real-space covariance matrices for galaxy clustering and weak lensing

ABSTRACT Accurate covariance matrices for two-point functions are critical for inferring cosmological parameters in likelihood analyses of large-scale structure surveys. Among various approaches to obtaining the covariance, analytic computation is much faster and less noisy than estimation from data or simulations. However, the transform of covariances from Fourier space to real space involves integrals with two Bessel integrals, which are numerically slow and easily affected by numerical uncertainties. Inaccurate covariances may lead to significant errors in the inference of the cosmological parameters. In this paper, we introduce a 2D-FFTLog algorithm for efficient, accurate, and numerically stable computation of non-Gaussian real-space covariances for both 3D and projected statistics. The 2D-FFTLog algorithm is easily extended to perform real-space bin-averaging. We apply the algorithm to the covariances for galaxy clustering and weak lensing for a Dark Energy Survey Year 3-like and a Rubin Observatory’s Legacy Survey of Space and Time Year 1-like survey, and demonstrate that for both surveys, our algorithm can produce numerically stable angular bin-averaged covariances with the flat sky approximation, which are sufficiently accurate for inferring cosmological parameters. The code CosmoCov for computing the real-space covariances with or without the flat-sky approximation is released along with this paper.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A system debugging model.

Consideration of the nature of the 'debugging' process applied to a new complex system during the initial period of its life. During this period failures and errors are corrected as they occur, with resulting improvement in the subsequent performance of the system. One mathematical idealization of this process leads to the assumption that system failure rate is decreasing with time. In practice, the debugging phase is considered completed when the failure rate reaches an equilibrium or constant value. Models are formulated for this phenomenon. Maximum likelihood estimates are obtained for relevant failure rate functions and for the end of the debugging period. A conservative upper confidence bound on the stable failure rate is obtained.

Barlow, R. E.↗

A random search algorithm for laboratory computers

The small laboratory computer is ideal for experimental control and data acquisition. Postexperimental data processing is often performed on large computers because of the availability of sophisticated programs, but costs and data compatibility are negative factors. Parameter optimization can be accomplished on the small computer, offering ease of programming, data compatibility, and low cost. A previously proposed random-search algorithm ('random creep') was found to be very slow in convergence. A method is proposed (the 'random leap' algorithm) which starts in a global search mode and automatically adjusts step size to speed convergence. A FORTRAN executive program for the random-leap algorithm is presented which calls a user-supplied function subroutine. An example of a function subroutine is given which calculates maximum-likelihood estimates of receiver operating-characteristic parameters from binary response data. Other applications in parameter estimation, generalized least squares, and matrix inversion are discussed.

Curry, R. E.↗

Estimating crop proportions from remotely sensed data

The classification/pixel-count method for estimating the proportion of wheat in each segment is theoretically biased even if all distributional assumptions are met. Alternative ways to estimate crop proportions are examined and their performance testing is considered. Topics covered include general linear functional estimates, the method of moments, and maximum likelihood estimators.

Feiveson, A. H.↗

CCD data processor for maximum likelihood feature classification

The paper describes an advanced technology development which utilizes a high speed analog/binary CCD correlator to perform the matrix multiplications necessary to implement onboard feature classification. The matrix manipulation module uses the maximum likelihood classification algorithm assuming a Gaussian probability density function. The module will process 16 element multispectral vectors at rates in excess of 500 thousand multispectral vector elements per second. System design considerations for the optimum use of this module are discussed, test results from initial device fabrication runs are presented, and the performance in typical processing applications is described

Benz, H. F.↗

The luminosity functions of the 1969 Perseid and Orionid meteor showers

Observations of the 1969 Perseid and Orionid meteor showers are presented and used to derive luminosity functions for the 288 Perseids and 56 Orionids detected. Visual counts were performed under very good to excellent seeing conditions at the times of peak activities, and the brightnesses of the meteors were estimated to the nearest magnitude by comparison with the magnitudes of known objects. Maximum likelihood estimates of the power law index of the luminosity function of 1.56 + or - 0.06 for the Perseids and of 1.85 + or - 0.1 for the Orionids are obtained which are lower than the values found by other investigators. Under the assumption that the luminosity of visual meteors is proportional to their mass, the luminosity function power law may also be used to characterize the mass function.

Krisciunas, K.↗

Prospects for Characterizing the Haziest Sub-Neptune Exoplanets with High-resolution Spectroscopy

Observations to characterize planets larger than Earth but smaller than Neptune have led to largely inconclusive interpretations at low spectral resolution due to hazes or clouds that obscure molecular features in their spectra. However, here we show that high-resolution spectroscopy (R ∼ 25,000–100,000) enables one to probe the regions in these atmospheres above the clouds where the cores of the strongest spectral lines are formed. We present models of transmission spectra for a suite of GJ 1214b–like planets with thick photochemical hazes covering 1–5 μm at a range of resolutions relevant to current and future ground-based spectrographs. Furthermore, we compare the utility of the cross-correlation function that is typically used with a more formal likelihood-based approach, finding that only the likelihood-based method is sensitive to the presence of haze opacity. We calculate the signal-to-noise ratio (S/N) of these spectra, including telluric contamination, Required to robustly detect a host of molecules such as CO, CO{sub 2}, H{sub 2}O, and CH{sub 4} and photochemical products like HCN as a function of wavelength range and spectral resolution. Spectra in the M band require the lowest S/N{sub res} to detect multiple molecules simultaneously. CH{sub 4} is only observable for the coolest models (T {sub eff} = 412 K) and only in the L band. We quantitatively assess how these requirements compare to what is achievable with current and future instruments, demonstrating that characterization of small cool worlds with ground-based high-resolution spectroscopy is well within reach.

79 ASTRONOMY AND ASTROPHYSICS↗

Predicting the Redshift 2 H-Alpha Luminosity Function Using [OIII] Emission Line Galaxies

Upcoming space-based surveys such as Euclid and WFIRST-AFTA plan to measure Baryonic Acoustic Oscillations (BAOs) in order to study dark energy. These surveys will use IR slitless grism spectroscopy to measure redshifts of a large number of galaxies over a significant redshift range. In this paper, we use the WFC3 Infrared Spectroscopic Parallel Survey (WISP) to estimate the expected number of H-alpha emitters observable by these future surveys. WISP is an ongoing Hubble Space Telescope slitless spectroscopic survey, covering the 0.8 - 1.65 micrometers wavelength range and allowing the detection of H-alpha emitters up to z approximately equal to 1.5 and [OIII] emitters to z approximately equal to 2.3. We derive the H-alpha-[OIII] bivariate line luminosity function for WISP galaxies at z approximately equal to 1 using a maximum likelihood estimator that properly accounts for uncertainties in line luminosity measurement, and demonstrate how it can be used to derive the H-alpha luminosity function from exclusively fitting [OIII] data. Using the z approximately equal to 2 [OIII] line luminosity function, and assuming that the relation between H-alpha and [OIII] luminosity does not change significantly over the redshift range, we predict the H-alpha number counts at z approximately equal to 2 - the upper end of the redshift range of interest for the future surveys. For the redshift range 0.7 less than z less than 2, we expect approximately 3000 galaxies per sq deg for a flux limit of 3 x 10(exp −16) ergs per sec per sq cm (the proposed depth of Euclid galaxy redshift survey) and approximately 20,000 galaxies per sq deg for a flux limit of approximately 10(exp −16) ergs per sec per sq cm (the baseline depth of WFIRST galaxy redshift survey).

Redshift↗

Structured Covariance Gaussian Networks for Orion Crew Module Aerodynamic Uncertainty Quantification

In this paper we propose a new approach for nonlinear regression and uncertainty quantification. The method is based on a pair of neural networks which parameterize mean and dense covariance functions of a multivariate Gaussian process, trained together to maximize the log-likelihood of observing the given data. The covariance matrix is made positive definite at every input by construction. We also propose a sampling approach that produces viable surrogate function realizations from the Gaussian process. We call the proposed model a Structured Covariance Gaussian Network (SCGN). We illustrate the use of SCGNs for learning an aerodynamic response surface with built-in uncertainty for the Orion crew module. We find that SCGN provides an efficient and systematic way to learn nonlinear functional relationships and dense covariances. We compare results to a baseline Gaussian process regressor and observe that the SCGN provides comparable uncertainty descriptions with improved scalability to dataset size. The sample functions generated by SCGN are fast to evaluate online and are therefore convenient for use in trajectory simulations. These results suggest that SCGN may be a viable computational method for aerodynamic uncertainty quantification.

machine learning↗

Structured Covariance Gaussian Networks for Orion Crew Module Aerodynamic Uncertainty Quantification

In this paper we propose a new approach for nonlinear regression and uncertainty quantification. The method is based on a pair of neural networks which parameterize mean and dense covariance functions of a multivariate Gaussian process, trained together to maximize the log-likelihood of observing the given data. The covariance matrix is made positive definite at every input by construction. We also propose a sampling approach that produces viable surrogate function realizations from the Gaussian process. We call the proposed model a Structured Covariance Gaussian Network (SCGN). We illustrate the use of SCGNs for learning an aerodynamic response surface with built-in uncertainty for the Orion crew module. We find that SCGN provides an efficient and systematic way to learn nonlinear functional relationships and dense covariances. We compare results to a baseline Gaussian process regressor and observe that the SCGN provides comparable uncertainty descriptions with improved scalability to dataset size. The sample functions generated by SCGN are fast to evaluate online and are therefore convenient for use in trajectory simulations. These results suggest that SCGN may be a viable computational method for aerodynamic uncertainty quantification.

machine learning↗