Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “data classification”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Automated vegetation classification using Thematic Mapper Simulation data

The present investigation is concerned with the results of a study of Thematic Mapper Simulation (TMS) data. One of the objectives of the study was related to an evaluation of the usefulness of the Thematic Mapper's (TM) improved spatial resolution and spectral coverage. The study was undertaken as part of a preparation for the efficient incorporation of Landsat 4 data into ongoing technology development in remote sensing. The study included an application of automated Landsat vegetation classification technology to TMS data. Results of comparing TMS data to Multispectral Scanner (MSS) data were found to indicate that all field definition, crop type discrimination, and subsequent proportion estimation may be greatly increased with the availability of TM data.

Nedelman, K. S.↗

Neuro-classification of multi-type Landsat Thematic Mapper data

Neural networks have been successful in image classification and have shown potential for classifying remotely sensed data. This paper presents classifications of multitype Landsat Thematic Mapper (TM) data using neural networks. The Landsat TM Image for March 23, 1987 with accompanying ground observation data for a study area In Miami County, Indiana, U.S.A. was utilized to assess recognition of crop residues. Principal components and spectral ratio transformations were performed on the TM data. In addition, a layer of the geographic information system (GIS) for the study site was incorporated to generate GIS-enhanced TM data. This paper discusses (1) the performance of neuro-classification on each type of data, (2) how neural networks recognized each type of data as a new image and (3) comparisons of the results for each type of data obtained using neural networks, maximum likelihood, and minimum distance classifiers.

Zhuang, Xin↗

Landsat digital data application to forest vegetation and land use classification in Minnesota

Landsat digital data were used to map eleven categories of land cover in north central Minnesota. The classification accuracy of these maps was found to be very low and they were not adequate for use by field level resource managers. A discussion of the advantages and disadvantages of various processing systems, different algorithms, and the problems in selecting training sets, is included.

Mead, R. A.↗

Machine boundary finding and sample classification of remotely sensed agricultural data

A method based on the use of spectral variations in combination with spatial variations is developed for automatic boundary finding and sample classification of remotely sensed multispectral data. Preliminary applications of the method to agricultural data show significant improvements in accuracy as compared to the use of spectral data alone.

Gupta, J. N.↗

Rock types present in lunar highland soils

Several investigators have studied soils from the lunar highlands with the objective of recognizing the parent rocks that have contributed significant amounts of material to these soils. Comparing only major element data, and thus avoiding the problems induced by individual classifications, these data appear to converge on a relatively limited number of rock types. The highland soils are derived from a suite of highly feldspathic rocks comprising anorthositic gabbros (or norites), high alumina basalts, troctolites, and less abundant gabbroic (or noritic) anorthosites, anorthosites, and KREEP basalts.

Reid, A. M.↗

Comparison of level I land cover classification accuracy for MSS and AVHRR data

The capabilities of the Advanced Very-High-Resolution Radiometer (AVHRR) for land-cover mapping were investigated by comparing the accuracy of land-cover information for the Washington, DC area derived from NOAA-7 AVHRR data with that from Landsat Multispectral Scanner Subsystem (MSS) data. Unsupervised level I land-cover classifications were performed for MSS and AVHRR data sets collected on July 11, 1981. A detailed accuracy assessment was conducted based on ground data delineated on 12 U.S. Geological Survey 7-5 min series topographic maps. These results produced overall land-cover classification accuracies of 71.9 and 76.8 per cent for AVHRR and MSS, respectively. While the accuracies for predominant categories were similar for both sensors, land-cover discrimination for less commonly occurring and/or spatially heterogeneous categories was improved with the MSS data set. The AVHRR, however, performed as well as or better than the MSS in classifying large homogeneous areas. The application of AVHRR data with its lower processing cost and more frequent worldwide coverage appears promising for regional land-cover mapping.

Gervin, J. C.↗

Prime agricultural land monitoring and assessment component of the California Integrated Remote Sensing System

The use of digital LANDSAT techniques for monitoring agricultural land use conversions was studied. Two study areas were investigated: one in Ventura County and the other in Fresno County (California). Ventura test site investigations included the use of three dates of LANDSAT data to improve classification performance beyond that previously obtained using single data techniques. The 9% improvement is considered highly significant. Also developed and demonstrated using Ventura County data is an automated cluster labeling procedure, considered a useful example of vertical data integration. Fresno County results for a single data LANDSAT classification paralleled those found in Ventura, demonstrating that the urban/rural fringe zone of most interest is a difficult environment to classify using LANDSAT data. A general raster to vector conversion program was developed to allow LANDSAT classification products to be transferred to an operational county level geographic information system in Fresno.

Estes, J. E.↗

LANDSAT-4 image data quality analysis

Classification performance from LANDSAT 4 TM and MSS data is evaluated using the SECHO computer program. The data accuracy is compared using forest, corn, soybeans, bare soil, grass, water, and urban areas as classes for investigation.

Anuta, P. E.↗

An unsupervised classification approach for analysis of Landsat data to monitor land reclamation in Belmont county, Ohio

Two unsupervised classification procedures for analyzing Landsat data used to monitor land reclamation in a surface mining area in east central Ohio are compared for agreement with data collected from the corresponding locations on the ground. One procedure is based on a traditional unsupervised-clustering/maximum-likelihood algorithm sequence that assumes spectral groupings in the Landsat data in n-dimensional space; the other is based on a nontraditional unsupervised-clustering/canonical-transformation/clustering algorithm sequence that not only assumes spectral groupings in n-dimensional space but also includes an additional feature-extraction technique. It is found that the nontraditional procedure provides an appreciable improvement in spectral groupings and apparently increases the level of accuracy in the classification of land cover categories.

Brumfield, J. O.↗

Crop identification technology assessment for remote sensing (CITARS). Volume 10: Interpretation of results

The CITARS was an experiment designed to quantitatively evaluate crop identification performance for corn and soybeans in various environments using a well-defined set of automatic data processing (ADP) techniques. Each technique was applied to data acquired to recognize and estimate proportions of corn and soybeans. The CITARS documentation summarizes, interprets, and discusses the crop identification performances obtained using (1) different ADP procedures; (2) a linear versus a quadratic classifier; (3) prior probability information derived from historic data; (4) local versus nonlocal recognition training statistics and the associated use of preprocessing; (5) multitemporal data; (6) classification bias and mixed pixels in proportion estimation; and (7) data with differnt site characteristics, including crop, soil, atmospheric effects, and stages of crop maturity.

Bizzell, R. M.↗

Crop identification and area estimation by computer-aided analysis of Landsat data

This report describes the results of a study involving the use of computer-aided analysis techniques applied to Landsat MSS data for identification and area estimation of winter wheat in Kansas and corn and soybeans in Indiana. Key elements of the approach included use of aerial photography for classifier training, stratification of Landsat data and extension of training statistics to areas without training data, and classification of a systematic sample of pixels from each county. Major results and conclusions are: (1) Landsat data was adequate for accurate identification and area estimation of winter wheat in Kansas, but corn and soybean estimates for Indiana were less accurate; (2) computer-aided analysis techniques can be effectively used to extract crop identification information from Landsat MSS data, and (3) systematic sampling of entire counties made possible by computer classification methods resulted in very precise area estimates at county as well as district and state levels.

Bauer, M. E.↗

Application of multivariate statistics to vestibular testing: discriminating between Meniere's disease and migraine associated dizziness

Meniere's disease (MD) and migraine associated dizziness (MAD) are two disorders that can have similar symptomatologies, but differ vastly in treatment. Vestibular testing is sometimes used to help differentiate between these disorders, but the inefficiency of a human interpreter analyzing a multitude of variables independently decreases its utility. Our hypothesis was that we could objectively discriminate between patients with MD and those with MAD using select variables from the vestibular test battery. Sinusoidal harmonic acceleration test variables were reduced to three vestibulo-ocular reflex physiologic parameters: gain, time constant, and asymmetry. A combination of these parameters plus a measurement of reduced vestibular response from caloric testing allowed us to achieve a joint classification rate of 91%, independent quadratic classification algorithm. Data from posturography were not useful for this type of differentiation. Overall, our classification function can be used as an unbiased assistant to discriminate between MD and MAD and gave us insight into the pathophysiologic differences between the two disorders.

NASA Discipline Neuroscience↗

Clustering Days with Similar Airport Weather Conditions

On any given day, traffic flow managers must often rely on past experience and intuition when developing traffic flow management initiatives that mitigate imbalances between the aircraft demand and the weather impacted airport capacity. The goal of this study was to build on recent efforts to apply data mining classification and clustering algorithms to vast archives of historical weather and air traffic data to identify patterns and past decisions that can ultimately inform day-of-operations decision-making. More specifically, this study identified similar weather impacted days at select U.S. airports, and analyzed the traffic management initiatives implemented on these representative days. The identification of the similar days was accomplished by applying a decision tree algorithm to the hourly Localized Aviation Model Output Statistics Program observations and the arrival delays for Newark Liberty International Airport. The branches from the trained decision tree were subsequently pruned to identify four weather conditions that resulted in medium to high delays for the arrivals scheduled to Newark in 2012. Using these weather conditions, four, daily airport-level Weather Impacted Traffic Index values were calculated using the Localized Aviation Model Output Statistics Program observations and the 2012 scheduled arrival counts from the FAAs Aviation System Performance Metric system. The four, daily Weather Impacted Traffic Index values for 2012 were subsequently clustered using an Expectation Maximization clustering algorithm, and nine unique types of weather days at Newark were identified. By far the most prominent type of day at Newark was a day associated with relatively good weather conditions, where there was little convective activity, winds were low, ceilings and visibility were high and there was little precipitation. Moderate levels of convective activity characterized the next most prominent type of day. Days with persistently high winds or low ceiling and visibility levels were relatively rare in 2012. Lastly, the frequency at which Ground Delay Programs, Ground Stops and Miles-in-Trail restrictions were implemented on each of the typical types of days at Newark were analyzed. Based on the results, it does appear as if the usage of Miles-in-Trail, Ground Delay Program and Ground Stop restrictions correlates well with the severity of the weather associated with each unique type of weather impacted day at Newark. Furthermore, the results demonstrate that it is feasible to use historical weather and air traffic archives to provide guidance on the types of traffic management restrictions to implement in response to the weather conditions impacting an airport.

traffic flow management↗

Clustering Days with Similar Airport Weather Conditions

On any given day, traffic flow managers must often rely on past experience and intuition when developing traffic flow management initiatives that mitigate imbalances between the aircraft demand and the weather impacted airport capacity. The goal of this study was to build on recent efforts to apply data mining classification and clustering algorithms to vast archives of historical weather and air traffic data to identify patterns and past decisions that can ultimately inform day-of-operations decision-making. More specifically, this study identified similar weather impacted days at select U.S. airports, and analyzed the traffic management initiatives implemented on these representative days. The identification of the similar days was accomplished by applying a decision tree algorithm to the hourly Localized Aviation Model Output Statistics Program observations and the arrival delays for Newark Liberty International Airport. The branches from the trained decision tree were subsequently pruned to identify four weather conditions that resulted in medium to high delays for the arrivals scheduled to Newark in 2012. Using these weather conditions, four, daily airport-level Weather Impacted Traffic Index values were calculated using the Localized Aviation Model Output Statistics Program observations and the 2012 scheduled arrival counts from the FAAs Aviation System Performance Metric system. The four, daily Weather Impacted Traffic Index values for 2012 were subsequently clustered using an Expectation Maximization clustering algorithm, and nine unique types of weather days at Newark were identified. By far the most prominent type of day at Newark was a day associated with relatively good weather conditions, where there was little convective activity, winds were low, ceilings and visibility were high and there was little precipitation. Moderate levels of convective activity characterized the next most prominent type of day. Days with persistently high winds or low ceiling and visibility levels were relatively rare in 2012. Lastly, the frequency at which Ground Delay Programs, Ground Stops and Miles-in-Trail restrictions were implemented on each of the typical types of days at Newark were analyzed. Based on the results, it does appear as if the usage of Miles-in-Trail, Ground Delay Program and Ground Stop restrictions correlates well with the severity of the weather associated with each unique type of weather impacted day at Newark. Furthermore, the results demonstrate that it is feasible to use historical weather and air traffic archives to provide guidance on the types of traffic management restrictions to implement in response to the weather conditions impacting an airport.

weather↗

Application of Bayesian Classification to Content-Based Data Management

The high volume of Earth Observing System data has proven to be challenging to manage for data centers and users alike. At the Goddard Earth Sciences Distributed Active Archive Center (GES DAAC), about 1 TB of new data are archived each day. Distribution to users is also about 1 TB/day. A substantial portion of this distribution is MODIS calibrated radiance data, which has a wide variety of uses. However, much of the data is not useful for a particular user's needs: for example, ocean color users typically need oceanic pixels that are free of cloud and sun-glint. The GES DAAC is using a simple Bayesian classification scheme to rapidly classify each pixel in the scene in order to support several experimental content-based data services for near-real-time MODIS calibrated radiance products (from Direct Readout stations). Content-based subsetting would allow distribution of, say, only clear pixels to the user if desired. Content-based subscriptions would distribute data to users only when they fit the user's usability criteria in their area of interest within the scene. Content-based cache management would retain more useful data on disk for easy online access. The classification may even be exploited in an automated quality assessment of the geolocation product. Though initially to be demonstrated at the GES DAAC, these techniques have applicability in other resource-limited environments, such as spaceborne data systems.

Lynnes, Christopher↗

Computer-aided analysis of Skylab multispectral scanner data in mountainous terrain for land use, forestry, water resource, and geologic applications

The author has identified the following significant results. One of the most significant results of this Skylab research involved the geometric correction and overlay of the Skylab multispectral scanner data with the LANDSAT multispectral scanner data, and also with a set of topographic data, including elevation, slope, and aspect. The Skylab S192 multispectral scanner data had distinct differences in noise level of the data in the various wavelength bands. Results of the temporal evaluation of the SL-2 and SL-3 photography were found to be particularly important for proper interpretation of the computer-aided analysis of the SL-2 and SL-3 multispectral scanner data. There was a quality problem involving the ringing effect introduced by digital filtering. The modified clustering technique was found valuable when working with multispectral scanner data involving many wavelength bands and covering large geographic areas. Analysis of the SL-2 scanner data involved classification of major cover types and also forest cover types. Comparison of the results obtained wth Skylab MSS data and LANDSAT MSS data indicated that the improved spectral resolution of the Skylab scanner system enabled a higher classification accuracy to be obtained for forest cover types, although the classification performance for major cover types was not significantly different.

Hoffer, R. M.↗

Classification of ERTS-1 MSS data by canonical analysis

The objective of canonical analysis is to obtain the maximum separability among a number of catergories. The application of canonical analysis was investigated using the merged MSS ERTS-1 data for one area viewed on two dates. The effect of threshold values on classification regions and confusion regions was investigated.

Lachowski, H. M.↗

A case study in the practical use of LANDSAT data

The use of computer aided classification of LANDSAT data in developing water quality plans for New Jersey watersheds is used to exemplify how a state natural resource management program benefits from satellite imagery. The transition of a research and development system into an operational remote sensing system to help decision makers is demonstrated. Nontechnial issues that can assist (or hinder) an agency in adopting a new technology are examined. The progress of LANDSAT use by state government from the earliest stage of curiosity through to incorporation in actual state planning methods is charted. Potential applications of LANDSAT data to real information needs and solutions to management problems are examined. The problems and mistakes that occurred in using LANDSAT data in the past are discussed as well as the ways by which these problems were overcome.

Cox, S.↗