Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Gaussian mixture model”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Photometry of Outer Solar System Objects from the Dark Energy Survey. II. A Joint Analysis of Trans-Neptunian Absolute Magnitudes, Colors, Light Curves and Dynamics

For the 696 trans-Neptunian objects (TNOs) with absolute magnitudes 5.5 < H r < 8.2 detected in the Dark Energy Survey, we characterize the relationships between their dynamical state and physical properties—namely H r , indicating size; colors, indicating surface composition; and flux variation semiamplitude A, indicating asphericity and surface inhomogeneity. We seek “birth” physical distributions that can recreate these parameters in every dynamical class. We show that the observed colors of these TNOs are consistent with two Gaussian distributions in griz space, “near-infrared bright” (NIRB) and “near-infrared faint” (NIRF), presumably an inner and outer birth population, respectively. We find a model in which both the NIRB and NIRF H r and A distributions are independent of current dynamical states, supporting their assignment as birth populations. All objects are consistent with a common rolling p(H r ), but NIRF objects are significantly more variable. Cold classicals (CCs) are purely NIRF, while hot classical (HC), scattered, and detached TNOs are consistent with ≈ 70% NIRB and the resonance NIRB fractions show significant variation. The NIRB components of the HCs and of some resonances have broader inclination distributions than the NIRFs, i.e. their current dynamics retains information about birth location. We find evidence for radial stratification within the birth NIRB population, in that HC NIRBs are on average redder than detached or scattered NIRBs; a similar effect distinguishes CCs from other NIRFs. We estimate total object counts and masses of each class within our H r range. These results will strongly constrain models of the outer solar system.

79 ASTRONOMY AND ASTROPHYSICS↗

Evaluation of a locally homogeneous model of spray evaporation

Measurements were conducted on an evaporating spray in a stagnant environment. The spray was formed using an air-atomizing injector to yield a Sauter mean diameter of the order of 30 microns. The region where evaporation occurred extended approximately 1 m from the injector for the test conditions. Profiles of mean velocity, temperature, composition, and drop size distribution, as well as velocity fluctuations and Reynolds stress, were measured. The results are compared with a locally homogeneous two-phase flow model which implies no velocity difference and thermodynamic equilibrium between the phases. The flow was represented by a k-epsilon-g turbulence model employing a clipped Gaussian probability density function for mixture fraction fluctuations. The model provides a good representation of earlier single-phase jet measurements, but generally overestimates the rate of development of the spray. Using the model predictions to represent conditions along the centerline of the spray, drop life-history calculations were conducted which indicate that these discrepancies are due to slip and loss of thermodynamic equilibrium between the phases.

Shearer, A. J.↗

Investigating the effects of precise mass measurements of Ru and Pd isotopes on machine learning mass modeling

Atomic masses are a foundational quantity in our understanding of nuclear structure, astrophysics, and fundamental symmetries. The longstanding goal of creating a predictive global model for the binding energy of a nucleus remains a significant challenge, however, and prompts the need for precise measurements of atomic masses to serve as anchor points for model developments. We present precise mass measurements of neutron-rich Ru and Pd isotopes performed at the Californium Rare Isotope Breeder Upgrade facility at Argonne National Laboratory using the Canadian Penning Trap mass spectrometer. The masses of 108 Ru, 110 Ru, and 116 Pd were measured to a relative mass precision $\delta$⁢$m/m$ ≈ 10 -8 via the phase-imaging ion-cyclotron-resonance technique, and represent an improvement of approximately an order of magnitude over previous measurements. Further, these mass data were used in conjunction with the physically interpretable machine learning (PIML) model, which uses a mixture density neural network to model mass excesses via a mixture of Gaussian distributions. The effects of our new mass data on a Bayesian-updating of a PIML model are presented.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Low-Cost Sensor Performance Intercomparison, Correction Factor Development, and 2+ Years of Ambient PM2.5 Monitoring in Accra, Ghana

Particulate matter air pollution is a leading cause of global mortality, particularly in Asia and Africa. Addressing the high and wide-ranging air pollution levels requires ambient monitoring, but many low- and middle-income countries (LMICs) remain scarcely monitored. To address these data gaps, recent studies have utilized low-cost sensors. These sensors have varied performance, and little literature exists about sensor intercomparison in Africa. By colocating 2 QuantAQ Modulair-PM, 2 PurpleAir PA-II SD, and 16 Clarity Node-S Generation II monitors with a reference-grade Teledyne monitor in Accra, Ghana, we present the first intercomparisons of different brands of low-cost sensors in Africa, demonstrating that each type of low-cost sensor PM2.5 is strongly correlated with reference PM2.5, but biased high for ambient mixture of sources found in Accra. When compared to a reference monitor, the QuantAQ Modulair-PM has the lowest mean absolute error at 3.04 μg/m3, followed by PurpleAir PA-II (4.54 μg/m3) and Clarity Node-S (13.68 μg/m3). We also compare the usage of 4 statistical or machine learning models (Multiple Linear Regression, Random Forest, Gaussian Mixture Regression, and XGBoost) to correct low-cost sensors data, and find that XGBoost performs the best in testing (R2: 0.97, 0.94, 0.96; mean absolute error: 0.56, 0.80, and 0.68 μg/m3 for PurpleAir PA-II, Clarity Node-S, and Modulair-PM, respectively), but tree-based models do not perform well when correcting data outside the range of the colocation training. Therefore, we used Gaussian Mixture Regression to correct data from the network of 17 Clarity Node-S monitors deployed around Accra, Ghana, from 2018 to 2021. We find that the network daily average PM2.5 concentration in Accra is 23.4 μg/m3, which is 1.6 times the World Health Organization Daily PM2.5 guideline of 15 μg/m3. While this level is lower than those seen in some larger African cities (such as Kinshasa, Democratic Republic of the Congo), mitigation strategies should be developed soon to prevent further impairment to air quality as Accra, and Ghana as a whole, rapidly grow.

Humidity↗

Classification Experiments on Real-World Texture

Many papers have been published concerning the analysis of visual texture and yet, very few application domains use texture for image classification. A possible reason for this low transfer of the technology is the lack of experience and testing in real-world imagery. In this paper, we assess the performance of texture-based classification methods on a number of real-world images relevant to autonomous navigation on cross-country terrain and to autonomous geology. Texture analysis will form part of the closed loop that allows a robotic system to navigate autonomously. We have implemented two different classifiers on features extracted by Gabor filter banks. The first classifier models feature distributions for each texture class using a mixture of Gaussians. Classification is performed using Maximum Likelihood. The second classifier represents local statistics using marginal histograms of the features over a region centered on the pixel to be classified. We measure system performance by comparison to ground truth image labels.

image segmentation↗

Convective shells in the interior of Cepheid variable stars: Overshooting models based on hydrodynamic simulations

Context. Because Cepheid variable stars have long been used as a cosmic benchmark for scaling distances in our Galaxy and beyond, the accuracy of stellar evolution models for Cepheids have wide-reaching effects. However, our understanding of the dynamics in the interiors of these physically complex stars is limited. Aims. Our goal is to provide a detailed multi-dimensional picture of hydrodynamic convection and convective boundary mixing in the interior of Cepheids. Methods. Using the Modules for Experiments in Stellar Astrophysics (MESA), we studied the structure of intermediate-mass stars that cross the instability strip. Then, we performed two-dimensional hydrodynamic simulations of six stars with the fully compressible Multidimensional Stellar Implicit Code (MUSIC). Our simulations did not model the radial pulsations but focused on the interior structure of this family of stars. We developed and applied a new statistical analysis to examine convection and convective boundary mixing in the interior of these stellar simulations. Results. Based on a grid of MESA models, we demonstrated that a common structure for intermediate mass Cepheids includes an interior convective shell as well as a thin outer convective envelope. Using the extreme value theory approach to analyze our MUSIC simulation data, we found that overshooting above the convective shell fills the space between these convectively unstable layers. We developed a new statistical analysis that provides a clearer picture of how overshooting fills this layer; it also allowed us to formulate a detailed comparison between overshooting above and below the convective shell. Our analysis effectively decomposes the overshooting layer into two layers: a weak overshooting layer and a strong overshooting layer. Statistically, this is accomplished by decomposing the strongly non-Gaussian probability density function into a mixture of gamma distributions. Using our mixture model, we showed that the ratio of overshooting lengths above and below the convective shell depends directly on the radial extent of the convective shell as well as its depth in the star. We proposed a new form for the diffusion coefficient that addresses the need for overlapping overshooting layers between convective shells. We introduced the idea of a “super-mixing layer” where overshooting from both the convective shell and the convective envelope results in efficient mixing and could be viewed as merging the two adjacent convective zones.

79 ASTRONOMY AND ASTROPHYSICS↗

Mixture-Tuned, Clutter Matched Filter for Remote Detection of Subpixel Spectral Signals

Mapping localized spectral features in large images demands sensitive and robust detection algorithms. Two aspects of large images that can harm matched-filter detection performance are addressed simultaneously. First, multimodal backgrounds may thwart the typical Gaussian model. Second, outlier features can trigger false detections from large projections onto the target vector. Two state-of-the-art approaches are combined that independently address outlier false positives and multimodal backgrounds. The background clustering models multimodal backgrounds, and the mixture tuned matched filter (MT-MF) addresses outliers. Combining the two methods captures significant additional performance benefits. The resulting mixture tuned clutter matched filter (MT-CMF) shows effective performance on simulated and airborne datasets. The classical MNF transform was applied, followed by k-means clustering. Then, each cluster s mean, covariance, and the corresponding eigenvalues were estimated. This yields a cluster-specific matched filter estimate as well as a cluster- specific feasibility score to flag outlier false positives. The technology described is a proof of concept that may be employed in future target detection and mapping applications for remote imaging spectrometers. It is of most direct relevance to JPL proposals for airborne and orbital hyperspectral instruments. Applications include subpixel target detection in hyperspectral scenes for military surveillance. Earth science applications include mineralogical mapping, species discrimination for ecosystem health monitoring, and land use classification.

Thompson, David R.↗

AutoClass: A Bayesian Approach to Classification

We describe a Bayesian approach to the untutored discovery of classes in a set of cases, sometimes called finite mixture separation or clustering. The main difference between clustering and our approach is that we search for the "best" set of class descriptions rather than grouping the cases themselves. We describe our classes in terms of a probability distribution or density function, and the locally maximal posterior probability valued function parameters. We rate our classifications with an approximate joint probability of the data and functional form, marginalizing over the parameters. Approximation is necessitated by the computational complexity of the joint probability. Thus, we marginalize w.r.t. local maxima in the parameter space. We discuss the rationale behind our approach to classification. We give the mathematical development for the basic mixture model and describe the approximations needed for computational tractability. We instantiate the basic model with the discrete Dirichlet distribution and multivariant Gaussian density likelihoods. Then we show some results for both constructed and actual data.

Stutz, John↗

Deconvolution of mineral absorption bands - An improved approach

Although visible and near IR reflectance spectra contain absorption bands that are characteristic of the composition and structure of the absorbing species, deconvolving a complex spectrum is nontrivial. An improved approach to spectral deconvolution is presented that accurately represents absorption bands as discrete mathematical distributions and resolves composite absorption features into individual absorption bands. The frequently used Gaussian model of absorption bands is shown to be inappropriate for the Fe(2+) electronic transition absorptions in pyroxene spectra. A modified Gaussian model is derived using a power law relationship of energy to average bond length. The modified Gaussian model is shown to provide an objective and consistent tool for deconvolving individual absorption bands in the more complex orthopyroxene, clinopyroxene, pyroxene mixtures, and olivine spectra.

Sunshine, Jessica M.↗

Symmetric normal mixtures

We consider mixture density estimation under the symmetry constraint x = Az for an orthogonal matrix A. This distributional constraint implies a corresponding constraint on the mixture parameters. Focusing on the gaussian case, we derive an expectation-maximization (EM) algorithm to enforce the constraint and show results for modeling of image feature vectors.

symmetry constraint↗

Data-Efficient Methods for Determining Flory–Huggins χ Parameters in Multicomponent Polymer Formulations

Polymer formulations are essential in diverse applications including personal care products, coatings, paints, adhesives, and plastic materials. Designing these formulations requires navigating large, complex design spaces, where phase and self-assembly behavior critically impact performance. The Flory–Huggins χ parameter, which quantifies segmental miscibility, is widely used to parametrize the excess free energy of mixing in formulation models. In this work, we introduce two data-efficient, top-down methods for estimating χ parameters using the Random Phase Approximation (RPA): (i) Boundary Nonlinear Regression (Boundary-NLR), which fits theoretical spinodal boundaries to experimental phase boundaries, and (ii) Surrogate Model Inverse Parameter Estimation (SMIPE), which uses a Gaussian Process Classifier to fit sparse phase maps via a surrogate model. Both methods allow rapid parametrization of polymer field-theoretic models without the need for additional experiments. We evaluate these approaches on data sets involving polymer–solvent–nonsolvent ternary mixtures and block copolymer–solvent systems, demonstrating their robustness to experimental noise and their relevance for real-world formulation design.

copolymers↗

Continental Spatio-Temporal Data Analysis with Linear Spectral Mixture Model Using FOSS

This work demonstrates the development and implementation of a Fully Constrained Least Squares (FCLS) unmixing model developed in C++ programming language with OpenCV package and boost C++ libraries in the NASA Earth Exchange (NEX). Visualization of the results is supported by GRASS GIS and statistical analysis is carried in R in a Linux system environment. FCLS was first tested on computer simulated data with Gaussian noise of various signal-to-noise ratio, and Landsat data of an agricultural scenario and an urban environment using a set of global end members of substrate (soils, sediments, rocks, and non-photosynthetic vegetation), vegetation that includes green photosynthetic plants and dark objects which encompasses absorptive substrate materials, clear water, deep shadows, etc. For the agricultural scenario, a spectrally diverse collection of 11 scenes of Level 1 terrain corrected, cloud free Landsat-5 TM data of Fresno, California, USA were unmixed and the results were validated with the corresponding ground data. To study an urbanized landscape, a clear sky Landsat-5 TM data were unmixed and validated with coincident World View-2 abundance maps (of 2 m spatial resolution) for an area of San Francisco, California, USA. The results were evaluated using descriptive statistics, correlation coefficient, RMSE, probability of success, boxplot and bivariate distribution function. Finally, FCLS was used for sub-pixel land cover analysis of the monthly WELD (Wen-enabled Landsat data) repository from 2008 to 2011 of North America. The abundance maps in conjunction with DMSP-OLS nighttime lights data were used to extract the urban land cover features and analyze their spatial-temporal growth.

Landsat Satellites↗

Analysis of the Westland Data Set

The "Westland" set of empirical accelerometer helicopter data with seeded and labeled faults is analyzed with the aim of condition monitoring. The autoregressive (AR) coefficients from a simple linear model encapsulate a great deal of information in a relatively few measurements; and it has also been found that augmentation of these by harmonic and other parameters call improve classification significantly. Several techniques have been explored, among these restricted Coulomb energy (RCE) networks, learning vector quantization (LVQ), Gaussian mixture classifiers and decision trees. A problem with these approaches, and in common with many classification paradigms, is that augmentation of the feature dimension can degrade classification ability. Thus, we also introduce the Bayesian data reduction algorithm (BDRA), which imposes a Dirichlet prior oil training data and is thus able to quantify probability of error in all exact manner, such that features may be discarded or coarsened appropriately.

Wen, Fang↗

Amplification of 10 μm picosecond pulses in a compact optically pumped multiatmosphere CO 2 laser

Picosecond pulses, generated by a commercial mid-IR optical parametric source (∼0.5 ps, 1-5 μJ), have been regeneratively amplified to an energy of ≤2 mJ in a palm-size high-pressure CO 2 amplifier optically pumped by a Q-switched 2.8 μm Cr:Er:YSGG laser. The characteristics of the amplified pulses have been studied as a function of seed wavelength centered around 10.3 or 10.6 μm, and the pressure in the gain medium. The measured amplified pulse length is ∼5 ps. Numerical modeling suggests that pulses as short as ∼0.5 ps with multi-GW peak power can be generated in a mixture of CO 2 isotopes. Robustness of the output of a compact CO 2 regenerative amplifier to the seed and pump fluctuations measured at 3 Hz and a perfect Gaussian beam profile indicates the potential of such laser technology for producing high-power ultrashort long-wave infrared pulses at high repetition rates.

43 PARTICLE ACCELERATORS↗

Condition Monitoring for Helicopter Data

In this paper the classical "Westland" set of empirical accelerometer helicopter data is analyzed with the aim of condition monitoring for diagnostic purposes. The goal is to determine features for failure events from these data, via a proprietary signal processing toolbox, and to weigh these according to a variety of classification algorithms. As regards signal processing, it appears that the autoregressive (AR) coefficients from a simple linear model encapsulate a great deal of information in a relatively few measurements; it has also been found that augmentation of these by harmonic and other parameters can improve classification significantly. As regards classification, several techniques have been explored, among these restricted Coulomb energy (RCE) networks, learning vector quantization (LVQ), Gaussian mixture classifiers and decision trees. A problem with these approaches, and in common with many classification paradigms, is that augmentation of the feature dimension can degrade classification ability. Thus, we also introduce the Bayesian data reduction algorithm (BDRA), which imposes a Dirichlet prior on training data and is thus able to quantify probability of error in an exact manner, such that features may be discarded or coarsened appropriately.

Wen, Fang↗

Probabilistic Mixture Model-Based Spectral Unmixing

Spectral unmixing attempts to decompose a spectral ensemble into the constituent pure spectral signatures (called endmembers) along with the proportion of each endmember. This is essential for techniques like hyperspectral imaging (HSI) used in environment monitoring, geological exploration, etc. Several spectral unmixing approaches have been proposed, many of which are connected to hyperspectral imaging. However, most extant approaches assume highly diverse collections of mixtures and extremely low-loss spectroscopic measurements. Additionally, current non-Bayesian frameworks do not incorporate the uncertainty inherent in unmixing. We propose a probabilistic inference algorithm that explicitly incorporates noise and uncertainty, enabling us to unmix endmembers in collections of mixtures with limited diversity. We use a Bayesian mixture model to jointly extract endmember spectra and mixing parameters while explicitly modeling observation noise and the resulting inference uncertainties. We obtain approximate distributions over endmember coordinates for each set of observed spectra while remaining robust to inference biases from the lack of pure observations and the presence of non-isotropic Gaussian noise. As a direct impact of our methodology, access to reliable uncertainties on the unmixing solutions would enable robust solutions to noise, as well as informed decision-making for HSI applications and other unmixing problems.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Combining High-Throughput Experiments and Active Learning to Characterize Deep Eutectic Solvents

The high tunability of deep eutectic solvents (DESs) stems from the ease of changing their precursors and relative compositions. However, measuring the physicochemical properties across large composition and temperature ranges, necessary to properly design target-specific DESs, is tedious and error-prone and represents a bottleneck in the advancement and scalability of DES-based applications. As such, active learning (AL) methodologies based on Gaussian processes (GPs) were developed in this work to minimize the experimental effort necessary to characterize DESs. Owing to its importance for large-scale applications, the reduction of DES viscosity through the addition of a low-molecular-weight solvent was explored as a case study. A high-throughput experimental screening was initially performed on nine different ternary DESs. Then, GPs were successfully trained to predict DES viscosity from its composition and temperature, showcasing the ability of these stochastic, nonparametric models to accurately describe the physicochemical properties of complex mixtures. Finally, the ability of GPs to provide estimates of their own uncertainty was leveraged through an AL framework to minimize the number of data points necessary to obtain accurate viscosity modes. This led to a significant reduction in data requirements, with many systems requiring only five independent viscosity data points to be properly described.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Radar optimization for sea surface and geodetic measurements

The efficient estimation of geoid and sea state parameters is discussed, and the optimum processing structures, including maximum likelihood estimators, and their accuracy limits are given for a model. The model accounts for random surface reflectivity, sea height, and additive noise, and allows for arbitrary radar system parameters, based on the assumption the received signal is a sample function of a normal random process. The integral equation associated with the Gaussian signal in Gaussian noise inference problem was solved. It is shown that the optimum processing is generally a mixture of coherent and incoherent integrations which may be viewed as a weighted summation of received power of the match-filtered received data. When estimates are correlated, the strongest correlation appears between geoid and asymmetry estimates, and between wave height standard deviation and reflectivity estimates.

Harger, R. O.↗