Engineering PapersSearch

SEARCH · Engineering Papers

Results for “data discover”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

DISCOVER-AQ: An Overview and Initial Comparisons of NO2 with OMI Observations

The first deployment of the Earth Venture -1 DISCOVER-AQ (Deriving Information on Surface conditions from Column and Vertically Resolved Observations Relevant to Air Quality) project was conducted during July 2011 in the Baltimore-Washington region. Two aircraft (a P-3B for in-situ sampling and a King Air for remote sensing) were used along with an extensive array of surface-based in-situ and remote sensing instrumentation. Fourteen flight days were accomplished by both aircraft and over 250 profiles of trace gases and aerosols were performed by the P-3B over surface air quality monitoring stations, which were specially outfitted with sunphotometers and Pandora UV/Vis spectrometers. The King Air flew with the High Spectral Resolution Lidar for aerosols and the ACAM UV/Vis spectrometer for trace gases. This suite of observations allows linkage of surface air quality with the vertical distributions of gases and aerosols, with remotely-sensed column amounts observed from the surface and from the King Air, and with satellite observations from Aura (OMI and TES), GOME-2, MODIS and GOES. The DISCOVER-AQ data will allow determination of under what conditions satellite retrievals are indicative of surface air quality, and they will be useful in planning new satellites. In addition to an overview of the project, a preliminary comparison of tropospheric column NO2 densities from the integration of in-situ P-3B observations, from the Pandoras and ACAM, and from the new Goddard OMI NO2 algorithm will be presented.

Pickering, Kenneth

System for Contributing and Discovering Derived Mission and Science Data

A system was developed to provide a new mechanism for members of the mission community to create and contribute new science data to the rest of the community. Mission tools have allowed members of the mission community to share first order data (data that is created by the mission s process in command and control of the spacecraft or the data that is captured by the craft itself, like images, science results, etc.). However, second and higher order data (data that is created after the fact by scientists and other members of the mission) was previously not widely disseminated, nor did it make its way into the mission planning process.

Wallick, Michael N.

Formaldehyde Column Density Measurements as a Suitable Pathway to Estimate Near-Surface Ozone Tendencies from Space

In support of future satellite missions that aim to address the current shortcomings in measuring air quality from space, NASA's Deriving Information on Surface Conditions from Column and Vertically Resolved Observations Relevant to Air Quality (DISCOVER-AQ) field campaign was designed to enable exploration of relationships between column measurements of trace species relevant to air quality at high spatial and temporal resolution. In the DISCOVER-AQ data set, a modest correlation (r2 = 0.45) between ozone (O3) and formaldehyde (CH2O) column densities was observed. Further analysis revealed regional variability in the O3-CH2O relationship, with Maryland having a strong relationship when data were viewed temporally and Houston having a strong relationship when data were viewed spatially. These differences in regional behavior are attributed to differences in volatile organic compound (VOC) emissions. In Maryland, biogenic VOCs were responsible for approx.28% of CH2O formation within the boundary layer column, causing CH2O to, in general, increase monotonically throughout the day. In Houston, persistent anthropogenic emissions dominated the local hydrocarbon environment, and no discernable diurnal trend in CH2O was observed. Box model simulations suggested that ambient CH2O mixing ratios have a weak diurnal trend (+/-20% throughout the day) due to photochemical effects, and that larger diurnal trends are associated with changes in hydrocarbon precursors. Finally, mathematical relationships were developed from first principles and were able to replicate the different behaviors seen in Maryland and Houston. While studies would be necessary to validate these results and determine the regional applicability of the O3-CH2O relationship, the results presented here provide compelling insight into the ability of future satellite missions to aid in monitoring near-surface air quality.

Surface Conditions from Column and Vertically Reso

ESIP Documentation Cluster Session: GCMD Keyword Update

The Global Change Master Directory (GCMD) Keywords are a hierarchical set of controlled Earth Science vocabularies that help ensure Earth science data and services are described in a consistent and comprehensive manner and allow for the precise searching of collection-level metadata and subsequent retrieval of data and services. Initiated over twenty years ago, the GCMD Keywords are periodically analyzed for relevancy and will continue to be refined and expanded in response to user needs. This talk explores the current status of the GCMD keywords, the value and usage that the keywords bring to different tools/agencies as it relates to data discovery, and how the keywords relate to SWEET (Semantic Web for Earth and Environmental Terminology) Ontologies.

data discover

Maximizing Spaceflight Biological Data with Omics Analytics: The NASA GeneLab Database

NASA’s GeneLab includes an open-access repository of some 250+ omics datasets generated by biological experiments relevant to spaceflight including simulated cosmic radiation and microgravity. In order to maximize the intelligibility of these data, particularly for users with limited bioinformatics background, GeneLab has become a knowledgebase platform converting raw genetic and proteomic signatures found in flight samples into biological and physiological meanings. A large community of more than 100 scientists has rallied behind GeneLab and organized into four Analysis Working Groups (AWGs: Animal, Plant, Microbe, and Multi-Omics). Together, the AWGs have gained scientific recognition worldwide by establishing a consortium in charge of adopting new complex standards for data analysis workflows and omics sample processing in a rapidly evolving field. We will demonstrate the usage of the repository with smart search capability, an online controlled-access toolshed "Galaxy" to process user data with vetted standard workflows, a workspace for data sharing and a data submission portal with ontology control for better metadata curation. The GeneLab visualization portal will also be demonstrated, showing how anyone without formal training in bioinformatics can now browse the space biology omics data to discover new biology and potential solutions to improve life in space.

Sylvain Vincent Costes

GeneLab: The NASA System Biology Platform for Space Omics Repository, Analysis and Visualization

NASA’s GeneLab includes an open-access repository of some 250+ omics datasets generated by biological experiments relevant to spaceflight including simulated cosmic radiation and microgravity. In order to maximize the intelligibility of these data, particularly for users with limited bioinformatics background, GeneLab has become a knowledgebase platform converting raw genetic and proteomic signatures found in flight samples into biological and physiological meanings. A large community of more than 100 scientists has rallied behind GeneLab and organized into four Analysis Working Groups (AWGs: Animal, Plant, Microbe, and Multi-Omics). Together, the AWGs have gained scientific recognition worldwide by establishing a consortium in charge of adopting new complex standards for data analysis workflows and omics sample processing in a rapidly evolving field. We will demonstrate the usage of the repository with smart search capability, an online controlled-access toolshed "Galaxy" to process user data with vetted standard workflows, a workspace for data sharing and a data submission portal with ontology control for better metadata curation. The GeneLab visualization portal will also be demonstrated, showing how anyone without formal training in bioinformatics can now browse the space biology omics data to discover new biology and potential solutions to improve life in space.

GeneLab

Identification of a precambrian rift through Missouri by digital image processing of geophysical and geological data

A newly discovered feature in the midcontinent - a gravity low that begins at a break in the midcontinent gravity high in SE Nebraska, extends across Missouri in a NW-SE direction, and intersects the Mississippi Valley graben to form the Pascola arch - is discussed. The anomaly varies from 120 to 160 km in width, extends approximately 700 km, and is best expressed in southern Missouri, where it has a Bouguer amplitude of about -34 mGal. It is noted that the magnitude of the anomaly cannot be explained on the basis of a thickened section of Paleozoic sedimentary rock. The gravity data and the sparse seismic refraction data for the region are found to be consistent with an increased crustal thickness beneath the gravity low. It is thought that the gravity anomaly is probably the present expression of a failed arm of a rifting event, perhaps one associated with the spreading that led to or preceded formation of the granite and rhyolite terrain of southern Missouri.

Guinness, E. A.

Discovering Jupiter. II

Data derived from Pioneer 10 and Pioneer 11 and other sources on the Jovian magnetosphere, the circum-Jovian radiation belts, and Jupiter's radio emission are presented at some length, descriptions are given of the principal Jovian satellites (Io, Europa, Ganymede, Callisto), and inferences are drawn on the origin of the planet and its place in the solar system. The inner, middle, and outer regions of the magnetosphere, the bow shock wave, and the particularly heavy intensity of the inner radiation belt region (within 1.44 million km of the planet) are discussed. All of the major satellites except Callisto lie immersed in the intense radiation belts. Jupiter's failure to become a stellar companion to the sun, Io's action in 'switching on' Jovian radio emission, and other Pioneer discoveries relating to asteroids, the solar system in general, and trans-Jovian space, are discussed

Source record

TPSAS-NF1676L-16833-DND

Semantic Infrastructure is central to realizing the first goal of the ASDC's Strategic Plan: expanding the ASDC's customer base by improving access to ASDC data. ASDC data comprises a widely heterogeneous set of complex products which presents two significant challenges in data access: Helping customers discover, among many available options, the most suitable data products for their purpose; and Guiding customers to easily and appropriately use products. Data products differ significantly in terms of how the data was collected and processed, even with similar subject matter. Understanding differences is critical to using data effectively. To reach a broader customer range, the ASDC must provide prospective users with enough information to quickly and meaningfully compare and evaluate data products. Data formats and structures also differ among products. Applications displaying and analyzing data need access to federated and semantically disambiguated data. Semantic technologies offer functionality for addressing this issue. Ontologies can provide robust, stable domain models serving as common schema for discovering, evaluating, comparing, and integrating data from disparate products. Reasoning engines and triple stores can leverage ontologies to support intelligent search applications allowing users to discover, query, retrieve, and easily reformat data from a broad spectrum of sources.

Beth Huffer

Identifying genomic data use with the Data Citation Explorer

Increases in sequencing capacity, combined with rapid accumulation of publications and associated data resources, have increased the complexity of maintaining associations between literature and genomic data. As the volume of literature and data have exceeded the capacity of manual curation, automated approaches to maintaining and confirming associations among these resources have become necessary. Here we present the Data Citation Explorer (DCE), which discovers literature incorporating genomic data that was not formally cited. This service provides advantages over manual curation methods including consistent resource coverage, metadata enrichment, documentation of new use cases, and identification of conflicting metadata. The service reduces labor costs associated with manual review, improves the quality of genome metadata maintained by the U.S. Department of Energy Joint Genome Institute (JGI), and increases the number of known publications that incorporate its data products. The DCE facilitates an understanding of JGI impact, improves credit attribution for data generators, and can encourage data sharing by allowing scientists to see how reuse amplifies the impact of their original studies.

59 BASIC BIOLOGICAL SCIENCES

Stewardship Best Practices for Improved Discovery and Reuse of Heterogeneous and Cross-Disciplinary Earth System Data

Some of the Earth system data products such as those from NASA airborne and field investigations (a.k.a. campaigns), are highly heterogeneous and cross-disciplinary, making the data extremely challenging to manage. For example, airborne and field campaign measurements tend to be sporadic over a period of time, with large gaps. Data products generated are of various processing levels and utilized for a wide range of inter- and cross-disciplinary research and applications. Data and derived products have been historically stored in a variety of domain-specific standard (and some non-standard) formats and in various locations such as NASA Distributed Active Archive Centers (DAACs), NASA airborne science facilities, field archives, or even individual scientists’ computer hard drives. As a result, airborne and field campaign data products have often been managed and represented differently, making it onerous for data users to find, access, and utilize campaign data. Some difficulties in discovering and accessing the campaign data originate from the incomplete data product and contextual metadata that may contain details relevant to the campaign (e.g. campaign acronym and instrument deployment locations), but tend to lack other significant information needed to understand conditions surrounding the data. Such details can be burdensome to locate after the conclusion of a campaign. Utilizing consistent terminology, essential for improved discovery and reuse, is also challenging due to the variety of involved disciplines. To help address the aforementioned challenges faced by many repositories and data managers handling airborne and field data, this presentation will describe stewardship practices developed by the Airborne Data Management Group (ADMG) within the Interagency Implementation and Advanced Concepts Team (IMPACT) under the NASA’s Earth Science Data systems (ESDS) Program.

best practices

Study of boundary-layer transition using transonic cone Preston tube data

Laminar layer Preston tube data on a sharp nose, ten degree cone obtained in the Ames 11 ft TWT and in flight tests are analyzed. During analyses of the laminar-boundary layer data, errors were discovered in both the wind tunnel and the flight data. A correction procedure for errors in the flight data is recommended which forces the flight data to exhibit some of the orderly characteristics of the wind tunnel data. From corrected wind tunnel data, a correlation is developed between Preston tube pressures and the corresponding values of theoretical laminar skin friction. Because of the uncertainty in correcting the flight data, a correlation for the unmodified data is developed, and, in addition, three other correlations are developed based on different correction procedures. Each of these correlations are used in conjunction with the wind tunnel correlation to define effective freestream unit Reynolds numbers for the 11 ft TWT over a Mach number range of 0.30 to 0.95. The maximum effective Reynolds numbers are approximately 6.5% higher than the normal values. These maximum values occur between freestream Mach numbers of 0.60 and 0.80. Smaller values are found outside this Mach number range. These results indicate wind tunnel noise affects the average laminar skin friction much less than it affects boundary layer transition. Data on the onset, extent, and end of boundary layer transition are summarized. Application of a procedure for studying the relative effects of varying nose radius on a ten degree cone at supercritical speeds indicates that increasing nose radius promotes boundary layer transition and separation of laminar boundary layers.

Reed, T. D.

Discovering Communicable Scientific Knowledge from Spatio-Temporal Data

This paper describes how we used regression rules to improve upon a result previously published in the Earth science literature. In such a scientific application of machine learning, it is crucially important for the learned models to be understandable and communicable. We recount how we selected a learning algorithm to maximize communicability, and then describe two visualization techniques that we developed to aid in understanding the model by exploiting the spatial nature of the data. We also report how evaluating the learned models across time let us discover an error in the data.

Schwabacher, Mark

Maps Suggest Transport and Source Processes of PM2.5 at 1 km x 1 km for the Whole San Joaquin Valley, Winter 2011 (Generalizations from DISCOVER-AQ)

We present interpreted data analysis using MAIAC (Multiangle implementation of Atmospheric Correction) retrievals and appropriate RAPid Update Cycle (RAP) meteorology to map respirable aerosol (PM2.5) for the period January and February, 2011. The San Joaquin Valley is one of the unhealthiest regions in the USA for PM2.5 and related morbidity. The methodology evaluated can be used for the entire moderate-resolution imaging spectrometer (MODIS, VIIRS) data record. Other difficult areas of the West: Riverside, CA, Salt Lake City, UT, and Doa Ana County, NM share similar difficulties and solutions. The maps of boundary layer depth for 1116 hr local time from RAP allows us to interpret aerosol optical thickness as a concentration of particles in a nearly well-mixed box capped by clean air. That mixing is demonstrated by DISCOVER-AQ data and afternoon samples from the airborne measurements, P3B (on-board) and B200 (HSRL2 lidar). This data and the PM2.5 gathered at the deployment sites allowed us to estimate and then evaluate consistency and daily variation of the AOT to PM2.5 relationship. Mixed-effects modeling allowed a refinement of that relation from day to day; RAP mixed layers explained the success of previous mixed-effects modeling. Compositional, size-distribution, and MODIS angle-of-regard effects seem to describe the need for residual daily correction beyond ML depth.We report on an extension method to the entire San Joaquin Valley for all days with MODIS imagery using the permanent PM2.5 stations, evaluated for representativeness. Resulting map movies show distinct sources, particularly Interstate-5 (at approx. 1km x 1km resolution) and the broader Bakersfield area. Accompanying winds suggest transport effects and variable pathways of pollution cleanout. Such estimates should allow morbiditymortality studies. They should be also useful for actual model assimilations, where composition and sources are uncertain. We conclude with a description of new work to extend these insights to similar regions, e.g. interior valleys of California, the Po Valley, the Mediterranean litoral, and the Ganges Plain.This work show generalizable use of remote sensing, a major goal of DISCOVER-AQ, Deriving Information on Surface Conditions from COlumn and VERtically Resolved Observations Relevant to Air Quality.

Chatfield, R.

XML Based Scientific Data Management Facility

The World Wide Web consortium has developed an Extensible Markup Language (XML) to support the building of better information management infrastructures. The scientific computing community realizing the benefits of HTML has designed markup languages for scientific data. In this paper, we propose a XML based scientific data management facility, XDMF. The project is motivated by the fact that even though a lot of scientific data is being generated, it is not being shared because of lack of standards and infrastructure support for discovering and transforming the data. The proposed data management facility can be used to discover the scientific data itself, the transformation functions, and also for applying the required transformations. We have built a prototype system of the proposed data management facility that can work on different platforms. We have implemented the system using Java, and Apache XSLT engine Xalan. To support remote data and transformation functions, we had to extend the XSLT specification and the Xalan package.

Mehrotra, Piyush

XML Based Scientific Data Management Facility

The World Wide Web consortium has developed an Extensible Markup Language (XML) to support the building of better information management infrastructures. The scientific computing community realizing the benefits of XML has designed markup languages for scientific data. In this paper, we propose a XML based scientific data management ,facility, XDMF. The project is motivated by the fact that even though a lot of scientific data is being generated, it is not being shared because of lack of standards and infrastructure support for discovering and transforming the data. The proposed data management facility can be used to discover the scientific data itself, the transformation functions, and also for applying the required transformations. We have built a prototype system of the proposed data management facility that can work on different platforms. We have implemented the system using Java, and Apache XSLT engine Xalan. To support remote data and transformation functions, we had to extend the XSLT specification and the Xalan package.

Mehrotra, P.

MAGSAT data processing: A report for investigators

The in-flight attitude and vector magnetometer data bias recovery techniques and results are described. The attitude bias recoveries are based on comparisons with a magnetic field model and are thought to be accurate to 20 arcsec. The vector magnetometer bias recoveries are based on comparisons with the scalar magnetometer data and are thought to be accurate to 3 nT or better. The MAGSAT position accuracy goals of 60 m radially and 300 m horizontally were achieved for all but the last 3 weeks of Magsat lifetime. This claim is supported by ephemeris overlap statistics and by comparisons with ephemerides computed with an independent orbit program using data from an independent tracking network. MAGSAT time determination accuracy is estimated at 1 ms. Several errors in prelaunch assumptions regarding data time tags, which escaped detection in prelaunch data tests, and were discovered and corrected postlaunch are described. Data formats and products, especially the Investigator-B tapes, which contain auxiliary parameters in addition to the basic magnetometer and ephemeris data, are described.

Langel, R. A.