Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “read classification”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Two papers on feed-forward networks

Connectionist feed-forward networks, trained with back-propagation, can be used both for nonlinear regression and for (discrete one-of-C) classification, depending on the form of training. This report contains two papers on feed-forward networks. The papers can be read independently. They are intended for the theoretically-aware practitioner or algorithm-designer; however, they also contain a review and comparison of several learning theories so they provide a perspective for the theoretician. The first paper works through Bayesian methods to complement back-propagation in the training of feed-forward networks. The second paper addresses a problem raised by the first: how to efficiently calculate second derivatives on feed-forward networks.

Buntine, Wray L.↗

Operational earth resources data handling system for the 1980's

Results of a recent study of data handling requirements for future operational earth observation systems are reported. Such systems in the 1980's may have 10-20 meter resolution and generate .2 tecabits of data per day, with peak rates of 0.8 gigabits per second based on the dominant requirements of agriculture. System alternatives are considered that will handle such data. Data relayed to the facility are recorded at high rates and then processed at lower speed by computer. Hardwired special digital logic computers may be used with an appropriate classification algorithm for crop recognition. Optical mass memories now in the prototype stage will handle the 10 tecabits and 800 megabit per second read-in rate required.

Van Vleck, E. M.↗

DMSP SSJ4 Data Restoration, Classification, and On-Line Data Access

Compress and clean raw data file for permanent storage We have identified various error conditions/types and developed algorithms to get rid of these errors/noises, including the more complicated noise in the newer data sets. (status = 100% complete). Internet access of compacted raw data. It is now possible to access the raw data via our web site, http://www.jhuapl.edu/Aurora/index.html. The software to read and plot the compacted raw data is also available from the same web site. The users can now download the raw data, read, plot, or manipulate the data as they wish on their own computer. The users are able to access the cleaned data sets. Internet access of the color spectrograms. This task has also been completed. It is now possible to access the spectrograms from the web site mentioned above. Improve the particle precipitation region classification. The algorithm for doing this task has been developed and implemented. As a result, the accuracies improved. Now the web site routinely distributes the results of applying the new algorithm to the cleaned data set. Mark the classification region on the spectrograms. The software to mark the classification region in the spectrograms has been completed. This is also available from our web site.

Wing, Simon↗

An interactive method for digitizing zone maps

A method is presented for digitizing maps that consist of zones, such as contour or climatic zone maps. A color-coded map is prepared by any convenient process. The map is then read into memory of an Image 100 computer by means of its table scanner, using colored filters. Zones are separated and stored in themes, using standard classification procedures. Thematic data are written on magnetic tape and these data, appropriately coded, are combined to make a digitized image on tape. Step-by-step procedures are given for digitization of crop moisture index maps with this procedure. In addition, a complete example of the digitization of a climatic zone map is given.

Giddings, L. E.↗

Robotic Rock Classification

This report describes a three-month research program undertook jointly by the Robotics Institute at Carnegie Mellon University and Ames Research Center as part of the Ames' Joint Research Initiative (JRI.) The work was conducted at the Ames Research Center by Mr. Liam Pedersen, a graduate student in the CMU Ph.D. program in Robotics under the supervision Dr. Ted Roush at the Space Science Division of the Ames Research Center from May 15 1999 to August 15, 1999. Dr. Martial Hebert is Mr. Pedersen's research adviser at CMU and is Principal Investigator of this Grant. The goal of this project is to investigate and implement methods suitable for a robotic rover to autonomously identify rocks and minerals in its vicinity, and to statistically characterize the local geological environment. Although primary sensors for these tasks are a reflection spectrometer and color camera, the goal is to create a framework under which data from multiple sensors, and multiple readings on the same object, can be combined in a principled manner. Furthermore, it is envisioned that knowledge of the local area, either a priori or gathered by the robot, will be used to improve classification accuracy. The key results obtained during this project are: The continuation of the development of a rock classifier; development of theoretical statistical methods; development of methods for evaluating and selecting sensors; and experimentation with data mining techniques on the Ames spectral library. The results of this work are being applied at CMU, in particular in the context of the Winter 99 Antarctica expedition in which the classification techniques will be used on the Nomad robot. Conversely, the software developed based on those techniques will continue to be made available to NASA Ames and the data collected from the Nomad experiments will also be made available.

Hebert, Martial↗

Identifying microbial functional guilds performing cryptic organotrophic and lithotrophic redox cycles in anaerobic granular biofilms

Granular biofilms used in anaerobic digester systems contain diverse microbial populations that interact to hydrolyze organic matter and produce methane within controlled environments. Prior research investigated the feasibility of utilizing granular biofilms obtained from an anaerobic digester to remove nitrate without the addition of exogenous electron donors. These granules possessed a unique structure of alternating light and dark iron sulfide and pyrite rich layers that potentially served as both an electron source and sink, linking carbon, nitrogen, sulfur, and iron cycles. To characterize the functional roles of diverse microbial populations enriched within these layered biofilms, we analyzed metagenomes obtained from three different granules. Comparisons between the functional gene content of forty metagenome assembled genomes (MAGs) identified phylogenetically cohesive functional guilds. Each of these functional MAG clusters was assigned to specific steps in anaerobic digestion (hydrolysis, acidogenesis, acetogenesis, and methanogenesis) and anaerobic respiration (denitrification and sulfate reduction). Comparisons with metagenomes derived from a variety of natural and engineered ecosystems confirmed that the enriched denitrifying bacteria were similar to populations typically found in wetlands and biological nitrogen removal systems. Analysis of read alignments to individual genes within the forty MAGs identified conserved genomic features that were representative of the functions that distinguished functional guilds. Overall, this research illustrates the utility of functional based classification of microorganisms for characterizing ecosystem functions and highlights the potential application of engineered ecosystems to serve as experimental models for complex natural ecosystems.

Ecosystem engineering↗

Land classification of south-central Iowa from computer enhanced images

The author has identified the following significant results. Enhanced LANDSAT imagery was most useful for land classification purposes, because these images could be photographically printed at large scales such as 1:63,360. The ability to see individual picture elements was no hindrance as long as general image patterns could be discerned. Low cost photographic processing systems for color printings have proved to be effective in the utilization of computer enhanced LANDSAT products for land classification purposes. The initial investment for this type of system was very low, ranging from $100 to $200 beyond a black and white photo lab. The technical expertise can be acquired from reading a color printing and processing manual.

Lucas, J. R.↗

Quantification of gas concentrations in NO/NO 2 /C 3 H 8 /NH 3 mixtures using machine learning

We employ machine learning to decode the composition of unknown gas mixtures from the output of an array of four electrochemical sensors. The sensors use metal oxide electrodes paired with a ceramic electrolyte, yttria-stabilized zirconia (YSZ), to produce voltage responses to the presence of gases in complex mixtures. The voltages from the sensor array serve as inputs to a machine learning pipeline which first carries out multi-class classification of mixtures into types based on which gases are present at non-zero concentrations, and subsequently predicts gas concentrations given the mixture type. Thus, our model is able to take a single reading from the sensor array in response to gas mixtures involving NO, NO 2 , C 3 H 8 , and NH 3 , and output a highly accurate prediction of which gases are present in the mixture, along with the concentrations of each constituent gas. Of note, our computational framework can be easily expanded to include additional gases and additional mixture types, allowing it to be used in numerous automotive, industrial and environmental monitoring settings.

47 OTHER INSTRUMENTATION↗

Statistical processing of Pioneer front film data, part 1

A program was constructed to read the data on impacts and positional information on Pioneer and to classify these events according to a number of different criteria. The program is flexible enough to permit the introduction of further criteria and additional classifications, should this appear desirable. Not all cards correspond to particle impacts on the Pioneer sensors, many are inserted only to supply Pioneer position information.

Wolf, H.↗

Using ensembles and distillation to optimize the deployment of deep learning models for the classification of electronic cancer pathology reports

One of the goals of the Surveillance, Epidemiology, and End Results (SEER) program is to estimate incidence, prevalence, and mortality of all cancers. To that end, cancer registries across the country maintain a massive database of cancer pathology reports which contain rich information to understand cancer trends. However, these reports are stored in the form of unstructured text, and human annotators are required to read and extract relevant information. In this article, we show that existing deep learning models for automating information extraction from cancer pathology reports can be significantly improved by using ensemble model distillation. We found that by training multiple predictive models and transferring their knowledge to a single, low-resource model, we can reduce the number of highly confident wrong predictions. Our results show that our implemented methods could save 1000s of manual annotation hours.

60 APPLIED LIFE SCIENCES↗

Simulating water dynamics related to pedogenesis across space and time: Implications for four-dimensional digital soil mapping

Digital soil mapping (DSM) relies on machine-learning and geostatistics to represent soil property observations across space. DSM techniques are powerful but often empirical, being limited to the quality and density of point samples. Water dynamics are closely related to soil variability, and the physics that govern water movement are well known. Hydrological properties can hence be simulated by physical models through space and time, unveiling key characteristics about soils. We propose the use of hydrologic models to map soils across the surface (2D), depth (1D), and time (1D)–which provides a 4D approach to digital soil mapping (4DSM). The Distributed Hydrology Soil Vegetation Model (DHSVM) was applied to a watershed currently under pasture. Moisture sensors and wells were installed at different depths in the watershed on summit, sideslope and toeslope positions to validate the model. DHSVM simulations of soil moisture distribution and depth to saturation were performed during the hydrological year (October 2008-September 2009). Clusters of similar pixels based on soil moisture values were determined using Dynamic Time Warping (DTW) to align temporal data and K-means. Clustering was performed both seasonally and for the entire year. Temporal patterns simulated by DHSVM matched measurements given by moisture sensors and wells. Seasonal clusters differed from the annual cluster. Distinct clusters were observed for each season and with depth, showing that spatiotemporal soil variability is lost when statically assessing soils. Spatiotemporal clusters corroborated field observations of fragipan occurrence not explicitly spatially mapped by Soil Survey Geographic Database (SSURGO). If a connection can be made between water and soils, static and dynamic soil variability can be predicted using physically based hydrologic models. Hydrologic models can benefit soil mapping by enabling reliable 4D simulation of water dynamics, which are fundamental to soil variability and soil classification and directly relate to biological, physical and chemical soil processes not captured by typical soil sampling protocols.

54 ENVIRONMENTAL SCIENCES↗

A catalog of selected compact radio sources for the construction of an extragalactic radio/optical reference frame (Argue et al. 1984): Documentation for the machine-readable version

This document describes the machine readable version of the Selected Compact Radio Source Catalog as it is currently being distributed from the international network of astronomical data centers. It is intended to enable users to read and process the computerized catalog. The catalog contains 233 strong, compact extragalactic radio sources having identified optical counterparts. The machine version contains the same data as the published catalog and includes source identifications, equatorial positions at J2000.0 and their mean errors, object classifications, visual magnitudes, redshift, 5-GHz flux densities, and comments.

Source record↗

Early evaluation of Thematic Mapper data for coastal process studies

Two sets of TM data taken over the ocean off the coast of the Southeastern U.S. Bight were studied for the applicability of TM data to marine environments. First, the results of applying TM and TMS data to determine chlorophyll concentration in the ocean are presented. Chlorophyll quantification in the range of 0.5 to 2.0 mg/cu m was achieved by taking the ratio of TM band-1/band-2. Second, the results of applying TM band-6 data to monitor sea surface temperature are described. A comparison of TM data with AVHRR data shows TM readings coincide with AVHRR data within a scatter of 0.5 deg C in most of the areas studied. Lastly, the results of a technique to map the water depths of coral reefs in the Great Bahama Bank are demonstrated. Depths from 0 to 20 meters were delineated using TM band-1. The classification accuracy and origins of anomalous depth points are discussed.

Kim, H. H.↗

Refinement of the “ Candidatus Accumulibacter” genus based on metagenomic analysis of biological nutrient removal (BNR) pilot-scale plants operated with reduced aeration

Members of the “Candidatus Accumulibacter” genus are widely studied as key polyphosphate-accumulating organisms (PAOs) in biological nutrient removal (BNR) facilities performing enhanced biological phosphorus removal (EBPR). This diverse lineage includes 18 “Ca. Accumulibacter” species, which have been proposed based on the phylogenetic divergence of the polyphosphate kinase 1 (ppk1) gene and genome-scale comparisons of metagenome-assembled genomes (MAGs). Phylogenetic classification based on the 16S rRNA genetic marker has been difficult to attain because most “Ca. Accumulibacter” MAGs are incomplete and often do not include the rRNA operon. Here, we investigate the “Ca. Accumulibacter” diversity in pilot-scale treatment trains performing BNR under low dissolved oxygen (DO) conditions using genome-resolved metagenomics. Using long-read sequencing, we recovered medium- and high-quality MAGs for 5 of the 18 “Ca. Accumulibacter” species, all with rRNA operons assembled, which allowed a reassessment of the 16S rRNA-based phylogeny of this genus and an analysis of phylogeny based on the 23S rRNA gene.

59 BASIC BIOLOGICAL SCIENCES↗

COMPASS-FME Terrestrial Ecosystem Manipulation to Probe the Effects of Storm Treatments (TEMPEST) Experiment Tree Inventory

This is the tree inventory (diameter, species, and live/dead status) data from the Terrestrial Ecosystem Manipulation to Probe the Effects of Storm Treatments (TEMPEST) experimental site. This manipulative, ecosystem-scale TEMPEST experiment is part of the COMPASS-FME (Coastal Observations, Mechanisms, and Predictions Across Systems and Scales: Field Measurements and Experiments; see https://compass.pnnl.gov/FME/COMPASSFME) project. It addresses the potential for freshwater and estuarine-water disturbance events to alter tree function, species composition, and ecosystem processes in a deciduous coastal forest in eastern Maryland, USA. The experiment uses a large-unit (2000 m2), un-replicated experimental design, with three 50 m × 40 m plots serving as control, freshwater, and estuarine-water treatments.This dataset includes:- An overall dataset README file.- The tree inventory data in both "wide" and "long" forms. These contain the same information but are structured differently, with the former more useful for human viewers and the latter more amenable for programmatic analyses.- A key to the species/genus codes used, which follow the U.S. Department of Agriculture's PLANTS schema (https://plants.usda.gov/).- A copy of the R code used to generate the wide- and long-form data files.All files are comma-separated value (CSV) and no special software is required to read them.

54 ENVIRONMENTAL SCIENCES↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (May to September 2017)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken roughly every month in the period May 18 to September 13 in 2017 at a location (Pit2) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Cores were taken with a hand-auger and separated into 5-20 cm segments based on soil horizonation down to 150 cm depth below surface. Each segment was subsampled for microbial analyses. Corresponding 16S rRNA gene amplicon data is available at the NCBI Single Read Archive (SRA) Database BioProject ID PRJNA626616, and soil geochemistry data at doi:10.15485/1631972. 40 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0142591. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores). This dataset includes a zip file of 6993 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs generated from the Wind River Basin (WRB). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.

54 ENVIRONMENTAL SCIENCES↗

Chemical Sensing of Unexploded Ordnance with the Mobile Underwater Survey System (MUDSS)

The ability to sense explosives residues in the marine environment is a critical tool for identification and classification of underwater unexploded ordnance (UXO). Trace explosives signatures of TNT and DNT have been extracted from multiple sediment samples adjacent to unexploded undersea ordnance at Halifax Harbor, Canada. The ordnance was hurled into the harbor during a massive explosion fifty years earlier, in 1945 after World War II had ended. Laboratory sediment extractions were made using the solid-phase microextraction (SPME) method in seawater, and detection using the Reversal Electron Attachment Detection (READ) technique and, in the case of DNT, a commercial gas-chromatography/mass spectrometer (GC/MS). Results show that, after more than 50 years in the environment, ordnance which appeared to be physically intact gave good explosives signatures at the parts-per-billion level, whereas ordnance which had been cracked open during the explosion gave no signatures at the 10 parts-per-trillion sensitivity level. These measurements appear to provide the first reported data of explosives signatures from undersea UXOs.

Darrach, M. R.↗

Signal detection theory and methods for evaluating human performance in decision tasks

Signal Detection Theory (SDT) can be used to assess decision making performance in tasks that are not commonly thought of as perceptual. SDT takes into account both the sensitivity and biases in responding when explaining the detection of external events. In the standard SDT tasks, stimuli are selected in order to reveal the sensory capabilities of the observer. SDT can also be used to describe performance when decisions must be made as to the classification of easily and reliably sensed stimuli. Numbers are stimuli that are minimally affected by sensory processing and can belong to meaningful categories that overlap. Multiple studies have shown that the task of categorizing numbers from overlapping normal distributions produces performance predictable by SDT. These findings are particularly interesting in view of the similarity between the task of the categorizing numbers and that of determining the status of a mechanical system based on numerical values that represent sensor readings. Examples of the use of SDT to evaluate performance in decision tasks are reviewed. The methods and assumptions of SDT are shown to be effective in the measurement, evaluation, and prediction of human performance in such tasks.

Obrien, Kevin↗