Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “pattern classification”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

ERIM progress report on use of ERTS-1 data: Summary report of work on ten tasks

The author has identified the following significant results. Several of the tasks have produced significant results which are summarized: (1) Absolute water depth can be calculated from a ratio of signals from bands MSS 4 and MSS 5. (2) A 13 category terrain feature classification map of Yellowstone National Park has been produced using supervised pattern recognition techniques. (3) ERTS-1 data has been shown to provide a detection and monitoring capability for a number of water quality problems associated with off-shore ocean dumping sites and inland lakes. (4) A corrected ratio of bands MSS-5 and MSS-7 signals has been formed. (5) A concise format has been devised for storing the ratio signatures of geologic rock and mineral materials determined from laboratory reflectance spectra. (6) Results of work in information extraction demonstrate: signal variability exists among ERTS-1 detectors in any one spectral band that will impact users doing quantitative analysis on successive ERTS-1 images; a newly developed computer-aided procedure for correlating ERTS-1 pixels to ground features; the strong influence of atmospheric effects in ERTS-1 data; and area estimation accuracies are better using the ERIM proportion estimation algorithm than for conventional recognition techniques.

Thomson, F. J.↗

Remote sensing exploration for metallic mineral resources in central Baja California

Remote sensor data (primarily LANDSAT) was analyzed by photogeologic and computer-assisted enhancement techniques to evaluate the metallic mineral potential of Baja California. Overlays were prepared at 1:1,000,000 and 1:500,000 and included known geologic relationships and mineral occurrences, lineament, drainage and structural patterns, tonal anomalies, and enhancement results. Computer-assisted enhancement and classification of the test sites was performed using the IMAGE 100 system to identify subtle tonal anomalies thought related to mineralization using known sites as analysis guides. Mineral potential maps of Baja California were generated from these analyses and the ten highest priority targets visited. Preliminary assay results (atomic absorption analysis) for the samples recovered showed moderate to high geochemical anomalies for Copper (10 of 12 samples), Zinc (3 of 12 samples) and Lead (4 of 12 samples).

Baker, R. N.↗

Assessment of Extremes in Global Precipitation Products: How Reliable Are They?

Global gridded precipitation products have proven essential for many applications ranging from hydrological modeling and climate model validation to natural hazard risk assessment. They provide a global picture of how precipitation varies across time and space, specifically in regions where ground-based observations are scarce. While the application of global precipitation products has become widespread, there is limited knowledge on how well these products represent the magnitude and frequency of extreme precipitation—the key features in triggering flood hazards. Here, five global precipitation datasets (MSWEP, CFSR, CPC, PERSIANN-CDR, and WFDEI) are compared to each other and to surface observations. The spatial variability of relatively high precipitation events (tail heaviness) and the resulting discrepancy among datasets in the predicted precipitation return levels were evaluated for the time period 1979–2017. The analysis shows that 1) these products do not provide a consistent representation of the behavior of extremes as quantified by the tail heaviness, 2) there is strong spatial variability in the tail index, 3) the spatial patterns of the tail heaviness generally match the Köppen–Geiger climate classification, and 4) the predicted return levels for 100 and 1000 years differ significantly among the gridded products. More generally, our findings reveal shortcomings of global precipitation products in representing extremes and highlight that there is no single global product that performs best for all regions and climates.

Risk assessment↗

Dissociation of gas-phase anisole induced by low-energy electron interactions: understanding patterns of aromatic bond cleavage

Abstract In view of elucidating the fragmentation patterns of aromatic systems induced by low-energy electron interactions, dissociative electron attachment (DEA) to gas-phase anisole was performed. Anionic fragments resulting from this DEA process were detected by a quadrupole mass spectrometer, and ion yields of those fragments as a function of incident electron energy were rendered. Our study showed the formation of CH 3 − , HCC − , and OCH 3 − fragments, suggesting that various dissociation channels proceed out of DEA to anisole. We employed density functional theory to compute thermodynamic threshold energies for each potential dissociation channel. Those theoretical calculations supported the prediction that the CH 3 − and OCH 3 − fragments form via mechanisms of single-bond cleavage; the HCC − fragments may form through two-, three-, or four-body dissociation channels that entail hydrogen transfers and the cleavage of multiple aromatic bonds. The experimental resonance energies that form the CH 3 − , HCC − , and OCH 3 − fragments were 6.0 eV, 5.8 and 9.7 eV, and 9.8 eV, respectively. Given the classification of anisole as a monosubstituted aromatic species, our results explain generalizable patterns of electron-mediated dissociation in aromatic systems.

Finley, Jacob (ORCID:0009000606963352)↗

Fast Characterization of Inducible Regions of Atrial Fibrillation Models With Multi-Fidelity Gaussian Process Classification

Computational models of atrial fibrillation have successfully been used to predict optimal ablation sites. A critical step to assess the effect of an ablation pattern is to pace the model from different, potentially random, locations to determine whether arrhythmias can be induced in the atria. In this work, we propose to use multi-fidelity Gaussian process classification on Riemannian manifolds to efficiently determine the regions in the atria where arrhythmias are inducible. We build a probabilistic classifier that operates directly on the atrial surface. We take advantage of lower resolution models to explore the atrial surface and combine seamlessly with high-resolution models to identify regions of inducibility. We test our methodology in 9 different cases, with different levels of fibrosis and ablation treatments, totalling 1,800 high resolution and 900 low resolution simulations of atrial fibrillation. When trained with 40 samples, our multi-fidelity classifier that combines low and high resolution models, shows a balanced accuracy that is, on average, 5.7% higher than a nearest neighbor classifier. We hope that this new technique will allow faster and more precise clinical applications of computational models for atrial fibrillation. All data and code accompanying this manuscript will be made publicly available at: https://github.com/fsahli/AtrialMFclass.

59 BASIC BIOLOGICAL SCIENCES↗

Fast Query-Optimized Kernel-Machine Classification

A recently developed algorithm performs kernel-machine classification via incremental approximate nearest support vectors. The algorithm implements support-vector machines (SVMs) at speeds 10 to 100 times those attainable by use of conventional SVM algorithms. The algorithm offers potential benefits for classification of images, recognition of speech, recognition of handwriting, and diverse other applications in which there are requirements to discern patterns in large sets of data. SVMs constitute a subset of kernel machines (KMs), which have become popular as models for machine learning and, more specifically, for automated classification of input data on the basis of labeled training data. While similar in many ways to k-nearest-neighbors (k-NN) models and artificial neural networks (ANNs), SVMs tend to be more accurate. Using representations that scale only linearly in the numbers of training examples, while exploring nonlinear (kernelized) feature spaces that are exponentially larger than the original input dimensionality, KMs elegantly and practically overcome the classic curse of dimensionality. However, the price that one must pay for the power of KMs is that query-time complexity scales linearly with the number of training examples, making KMs often orders of magnitude more computationally expensive than are ANNs, decision trees, and other popular machine learning alternatives. The present algorithm treats an SVM classifier as a special form of a k-NN. The algorithm is based partly on an empirical observation that one can often achieve the same classification as that of an exact KM by using only small fraction of the nearest support vectors (SVs) of a query. The exact KM output is a weighted sum over the kernel values between the query and the SVs. In this algorithm, the KM output is approximated with a k-NN classifier, the output of which is a weighted sum only over the kernel values involving k selected SVs. Before query time, there are gathered statistics about how misleading the output of the k-NN model can be, relative to the outputs of the exact KM for a representative set of examples, for each possible k from 1 to the total number of SVs. From these statistics, there are derived upper and lower thresholds for each step k. These thresholds identify output levels for which the particular variant of the k-NN model already leans so strongly positively or negatively that a reversal in sign is unlikely, given the weaker SV neighbors still remaining. At query time, the partial output of each query is incrementally updated, stopping as soon as it exceeds the predetermined statistical thresholds of the current step. For an easy query, stopping can occur as early as step k = 1. For more difficult queries, stopping might not occur until nearly all SVs are touched. A key empirical observation is that this approach can tolerate very approximate nearest-neighbor orderings. In experiments, SVs and queries were projected to a subspace comprising the top few principal- component dimensions and neighbor orderings were computed in that subspace. This approach ensured that the overhead of the nearest-neighbor computations was insignificant, relative to that of the exact KM computation.

Mazzoni, Dominic↗

Classification of physiography from ERTS imagery

The potential application of optical data processing to ERTS imagery as a means for automatic identification of large-scale ground patterns was investigated. Spatial frequency distribution and orientational information were derived from ERTS-1 imagery of Kansas for each of 80 sample areas, each 37 km in diameter. The application of classification algorithms to this data reveals that a high degree of correlation exists between the physiography of a sample area and its frequency information. Specifically, the band of frequencies between 1.1 and 2.8 cycles/km appear to contain most of the information needed in distinguishing different physiographic regions.

Ulaby, F. T.↗

Geophysical phenomena classification by artificial neural networks

Space science information systems involve accessing vast data bases. There is a need for an automatic process by which properties of the whole data set can be assimilated and presented to the user. Where data are in the form of spectrograms, phenomena can be detected by pattern recognition techniques. Presented are the first results obtained by applying unsupervised Artificial Neural Networks (ANN's) to the classification of magnetospheric wave spectra. The networks used here were a simple unsupervised Hamming network run on a PC and a more sophisticated CALM network run on a Sparc workstation. The ANN's were compared in their geophysical data recognition performance. CALM networks offer such qualities as fast learning, superiority in generalizing, the ability to continuously adapt to changes in the pattern set, and the possibility to modularize the network to allow the inter-relation between phenomena and data sets. This work is the first step toward an information system interface being developed at Sussex, the Whole Information System Expert (WISE). Phenomena in the data are automatically identified and provided to the user in the form of a data occurrence morphology, the Whole Information System Data Occurrence Morphology (WISDOM), along with relationships to other parameters and phenomena.

Gough, M. P.↗

Accurate and Data‐Efficient Micro X‐ray Diffraction Phase Identification Using Multitask Learning: Application to Hydrothermal Fluids

Traditional analysis of highly distorted micro X‐ray diffraction (μ‐XRD) patterns from hydrothermal fluid environments is a time‐consuming process, often requiring substantial data preprocessing and labeled experimental data. Herein, the potential of deep learning with a multitask learning (MTL) architecture to overcome these limitations is demonstrated. MTL models are trained to identify phase information in μ‐XRD patterns, minimizing the need for labeled experimental data and masking preprocessing steps. Notably, MTL models show superior accuracy compared to binary classification convolutional neural networks. Additionally, introducing a tailored cross‐entropy loss function improves MTL model performance. Most significantly, MTL models tuned to analyze raw and unmasked XRD patterns achieve close performance to models analyzing preprocessed data, with minimal accuracy differences. This work indicates that advanced deep learning architectures like MTL can automate arduous data handling tasks, streamline the analysis of distorted XRD patterns, and reduce the reliance on labor‐intensive experimental datasets.

97 MATHEMATICS AND COMPUTING↗

AICCA: AI-Driven Cloud Classification Atlas

Clouds play an important role in the Earth’s energy budget, and their behavior is one of the largest uncertainties in future climate projections. Satellite observations should help in understanding cloud responses, but decades and petabytes of multispectral cloud imagery have to date received only limited use. This study describes a new analysis approach that reduces the dimensionality of satellite cloud observations by grouping them via a novel automated, unsupervised cloud classification technique based on a convolutional autoencoder, an artificial intelligence (AI) method good at identifying patterns in spatial data. Our technique combines a rotation-invariant autoencoder and hierarchical agglomerative clustering to generate cloud clusters that capture meaningful distinctions among cloud textures, using only raw multispectral imagery as input. Cloud classes are therefore defined based on spectral properties and spatial textures without reliance on location, time/season, derived physical properties, or pre-designated class definitions. We use this approach to generate a unique new cloud dataset, the AI-driven cloud classification atlas (AICCA), which clusters 22 years of ocean images from the Moderate Resolution Imaging Spectroradiometer (MODIS) on NASA’s Aqua and Terra instruments—198 million patches, each roughly 100 km × 100 km (128 × 128 pixels)—into 42 AI-generated cloud classes, a number determined via a newly-developed stability protocol that we use to maximize richness of information while ensuring stable groupings of patches. AICCA thereby translates 801 TB of satellite images into 54.2 GB of class labels and cloud top and optical properties, a reduction by a factor of 15,000. The 42 AICCA classes produce meaningful spatio-temporal and physical distinctions and capture a greater variety of cloud types than do the nine International Satellite Cloud Climatology Project (ISCCP) categories—for example, multiple textures in the stratocumulus decks along the West coasts of North and South America. We conclude that our methodology has explanatory power, capturing regionally unique cloud classes and providing rich but tractable information for global analysis. AICCA delivers the information from multi-spectral images in a compact form, enables data-driven diagnosis of patterns of cloud organization, provides insight into cloud evolution on timescales of hours to decades, and helps democratize climate research by facilitating access to core data.

97 MATHEMATICS AND COMPUTING↗

The infinite level normal forms for non-resonant double Hopf singularities

Here, in this paper, we explore hypernormal forms of vector fields that have non-resonant double Hopf singularities with a non-zero radial cubic part. Our primary focus is on investigating the infinite-level normal form classification of this type of singularities. We provide a normal form decomposition in terms of planar-rotating and planar-radial vector fields, which greatly facilitate the pattern recognition and analysis of the corresponding generalized homological maps. Notably, our paper represents the first instance of the normal form classification for general non-resonant double Hopf singularities without structural symmetry.

97 MATHEMATICS AND COMPUTING↗

ARMing the Edge: Designing Edge Computing–Capable Machine Learning Algorithms to Target ARM Doppler Lidar Processing

Abstract There is a need for long-term observations of cloud and precipitation fall speeds in validating and improving rainfall forecasts from climate models. To this end, the U.S. Department of Energy Atmospheric Radiation Measurement (ARM) user facility Southern Great Plains (SGP) site at Lamont, Oklahoma, hosts five ARM Doppler lidars that can measure cloud and aerosol properties. In particular, the ARM Doppler lidars record Doppler spectra that contain information about the fall speeds of cloud and precipitation particles. However, due to bandwidth and storage constraints, the Doppler spectra are not routinely stored. This calls for the automation of cloud and rain detection in ARM Doppler lidar data so that the spectral data in clouds can be selectively saved and further analyzed. During the ARMing the Edge field experiment, a Waggle node capable of performing machine learning applications in situ was deployed at the ARM SGP site for this purpose. In this paper, we develop and test four algorithms for the Waggle node to automatically classify ARM Doppler lidar data. We demonstrate that supervised learning using a ResNet50-based classifier will classify 97.6% of the clear-air images and 94.7% of cloudy images correctly, outperforming traditional peak detection methods. We also show that a convolutional autoencoder paired with k -means clustering identifies 10 clusters in the ARM Doppler lidar data. Three clusters correspond to mostly clear conditions with scattered high clouds, and seven others correspond to cloudy conditions with varying cloud-base heights.

54 ENVIRONMENTAL SCIENCES↗

A study of morphology, provenance, and movement of desert sand seas in Africa, Asia, and Australia

A description and classification of major types of sand seas on the basis of morphological pattern and lineation are discussed. The steps involved in analyzing the patterns of deposits on ERTS-1 imagery, where the visible forms are mostly dune complexes rather than individual dunes are outlined. After completion of thematic maps portraying the pattern and lineation of the sand bodies, data on directions and intensity of prevailing and other winds are plotted on corresponding bases, as a preliminary to determination of internal structures through ground truth.

Mckee, E. D.↗

Identifying priority ecosystem services in tidal wetland restoration

Classification systems can be an important tool for identifying and quantifying the importance of relationships, assessing spatial patterns in a standardized way, and forecasting alternative decision scenarios to characterize the potential benefits (e.g., ecosystem services) from ecosystem restoration that improve human health and well-being. We present a top-down approach that systematically leverages ecosystem services classification systems to identify potential services relevant for ecosystem restoration decisions. We demonstrate this approach using the U.S. Environmental Protection Agency’s National Ecosystem Service Classification System Plus (NESCS Plus) to identify those ecosystem services that are relevant to restoration of tidal wetlands. We selected tidal wetland management documents from federal agencies, state agencies, wetland conservation organizations, and land stewards across three regions of the continental United States (northern Gulf of Mexico, Mid-Atlantic, and Pacific Northwest) to examine regional and organizational differences in identified potential benefits of tidal wetland restoration activities and the potential user groups who may benefit. We used an automated document analysis to quantify the frequencies at which different wetland types were mentioned in the management documents along with their associated beneficiary groups and the ecological end products (EEPs) those beneficiaries care about, as defined by NESCS Plus. Results showed that a top combination across all three regions, all four organizations, and all four tidal wetland types was the EEP naturalness paired with the beneficiary people who care (existence). Overall, the Mid-Atlantic region and the land steward organizations mentioned ecosystem services more than the others, and EEPs were mentioned in combination with tidal wetlands as a high-level, more general category than the other more specific tidal wetland types. Certain regional and organizations differences were statistically significant. Those results may be useful in identifying ecosystem services-related goals for tidal wetland restoration. This approach for identifying and comparing ecosystem service priorities is broadly transferrable to other ecosystems or decision-making contexts.

54 ENVIRONMENTAL SCIENCES↗

Analysis of multispectral data using an unsupervised classification technique: Application to VAS

A statistical classification method based on clustering of multidimensional histograms was applied to several channels of the VAS multispectral imagery. The method automatically discriminates and classifies atmospheric ground features such as cloud types, atmospheric moisture patterns, ocean, or ground. Such a clustering method has the advantage of forming natural data groupings, without a priori classification. Clusters are not limited by straight lines or plane surfaces as is the case in threshold methods. The method was applied to simultaneous full resolution images from channels 8 (11.2 micron), 10 (6.7 micron), and 12 (3.9 micron). Twenty image segments of 64 by 64, 12 image segments of 128 by 128, and 4 image segments of 254 by 254 picture elements were analyzed. In addition, normal VISSR mode images at 1800, 1830, and 2000 GMT were used to identify the classes. The gray levels measured along a scan line and the result of the classification scheme (dashed curves) for the three channels investigated are shown. Each point of the image is affected to a class. Each class is identified by a center of gravity that is represented by a vector in the three dimensional space of gray levels.

Szejwach, G.↗

Classification by means of B-spline potential functions with applications to remote sensing

A method is presented for using B-splines as potential functions in the estimation of likelihood functions (probability density functions conditioned on pattern classes), or the resulting discriminant functions. The consistency of this technique is discussed. Experimental results of using the likelihood functions in the classification of remotely sensed data are given.

Bennett, J. O.↗

Unsupervised learning approaches to characterizing heterogeneous samples using X-ray single-particle imaging

One of the outstanding analytical problems in X-ray single-particle imaging (SPI) is the classification of structural heterogeneity, which is especially difficult given the low signal-to-noise ratios of individual patterns and the fact that even identical objects can yield patterns that vary greatly when orientation is taken into consideration. Proposed here are two methods which explicitly account for this orientation-induced variation and can robustly determine the structural landscape of a sample ensemble. The first, termed common-line principal component analysis (PCA), provides a rough classification which is essentially parameter free and can be run automatically on any SPI dataset. The second method, utilizing variation auto-encoders (VAEs), can generate 3D structures of the objects at any point in the structural landscape. Both these methods are implemented in combination with the noise-tolerant expand–maximize–compress (EMC) algorithm and its utility is demonstrated by applying it to an experimental dataset from gold nanoparticles with only a few thousand photons per pattern. Both discrete structural classes and continuous deformations are recovered. These developments diverge from previous approaches of extracting reproducible subsets of patterns from a dataset and open up the possibility of moving beyond the study of homogeneous sample sets to addressing open questions on topics such as nanocrystal growth and dynamics, as well as phase transitions which have not been externally triggered.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Adaptive statistical pattern classifiers for remotely sensed data

A technique for the adaptive estimation of nonstationary statistics necessary for Bayesian classification is developed. The basic approach to the adaptive estimation procedure consists of two steps: (1) an optimal stochastic approximation of the parameters of interest and (2) a projection of the parameters in time or position. A divergence criterion is developed to monitor algorithm performance. Comparative results of adaptive and nonadaptive classifier tests are presented for simulated four dimensional spectral scan data.

Gonzalez, R. C.↗