Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “distributed databases”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

U.S. Naval Observatory VLBI Analysis Center

This report summarizes the activities of the VLBI Analysis Center at the United States Naval Observatory for the 2012 calendar year. Over the course of the year, Analysis Center personnel continued analysis and timely submission of IVS-R4 databases for distribution to the IVS. During the 2012 calendar year, the USNO VLBI Analysis Center produced two VLBI global solutions designated as usn2012a and usn2012b. Earth orientation parameters (EOP) based on this solution and updated by the latest diurnal (IVS-R1 and IVS-R4) experiments were routinely submitted to the IVS. Sinex files based upon the bi-weekly 24-hour experiments were also submitted to the IVS. During the 2012 calendar year, Analysis Center personnel continued a program to use the Very Long Baseline Array (VLBA) operated by the NRAO for the purpose of measuring UT1-UTC. Routine daily 1-hour duration Intensive observations were initiated using the VLBA antennas at Pie Town, NM and Mauna Kea, HI. High-speed network connections to these two antennas are now routinely used for electronic transfer of VLBI data over the Internet to a USNO point of presence. A total of 270 VLBA Intensive experiments were observed and electronically transferred to and processed at USNO in 2012.

Boboltz, David A.↗

Re-Organizing Earth Observation Data Storage to Support Temporal Analysis of Big Data

The Earth Observing System Data and Information System archives many datasets that are critical to understanding long-term variations in Earth science properties. Thus, some of these are large, multi-decadal datasets. Yet the challenge in long time series analysis comes less from the sheer volume than the data organization, which is typically one (or a small number of) time steps per file. The overhead of opening and inventorying complex, API-driven data formats such as Hierarchical Data Format introduces a small latency at each time step, which nonetheless adds up for datasets with O(10^6) single-timestep files. Several approaches to reorganizing the data can mitigate this overhead by an order of magnitude: pre-aggregating data along the time axis (time-chunking); storing the data in a highly distributed file system; or storing data in distributed columnar databases. Storing a second copy of the data incurs extra costs, so some selection criteria must be employed, which would be driven by expected or actual usage by the end user community, balanced against the extra cost.

data storage↗

Update transport - A new technique for update synchronization in replicated database systems

A fully distributed approach to update synchronization is presented where each site completely executes every update. This approach has several features-higher resiliency to different kinds of failures, higher parallelism, improved response to user requests, and low communication overhead. A fully distributed algorithm for concurrency control obtained by rehashing a previously published semidistributed algorithm into the fully distributed model of update execution is presented. A performance model of replicated database systems is presented and used to study the performance of the proposed algorithm and its semidistributed version. The results of the performance study reveal that the proposed approach can substantially improve the performance at the cost of moderate input/output overhead.

Singhal, Mukesh↗

System for Performing Single Query Searches of Heterogeneous and Dispersed Databases

The present invention is a distributed computer system of heterogeneous databases joined in an information grid and configured with an Application Programming Interface hardware which includes a search engine component for performing user-structured queries on multiple heterogeneous databases in real time. This invention reduces overhead associated with the impedance mismatch that commonly occurs in heterogeneous database queries.

Maluf, David A.↗

Reprocessing Microflare Data

The report concerns work on detecting and cataloging solar microflares using an automated. An accompanying figure represents the solar microflare distribution during the period of April 1991 to November 1992, the height of solar activity after the launch of CGRO. It also shows the distribution extending below the distribution obtained at GSFC by manual means. We have implemented significant refinements in the search algorithm. The algorithm in its simplest form searches for transient events and based upon the distribution of the signal among the different BATSE detectors, we can assign it to be of solar origin if the signal distribution conforms to what one expects from a burst or transient from that direction. One of the major problems in an earlier effort was to search for microflares and large flares simultaneously. The requirement for a dynamic range of almost 10(exp 4) resulted in ambiguous identifications at the low side of the distribution. We have since restricted the search to events with peak count rates under 2000/s. Larger events are easily identified in the manual search, so we have chosen not to duplicate that work. The second problem was that missing counts existed below channel 0 in the BATSE Large Area Detector (LAD) data. These have been recovered and are now included in the search process. This provides data below 20 keV, and as we get closer to the thermal part of the spectrum, it provides greater sensitivity. The third problem was that too many BATSE detectors were used in the search. Detectors with pointing directions far from the Sun, although detecting the event, had poorly known responses. Detectors greater than approximately 60 degrees off the Sun are no longer included in the search process. By reducing the systematic errors with the large off-axis detectors we can conduct more rigorous statistical tests of a candidate event to ascertain whether it originated from the solar direction. We have reprocessed the period in the early mission that covers solar maximum and constructed the microflare distribution shown in the figure. The results of the automated search start to deviate from the manual search results below about 1000/s. Not only do we now have this distribution but we have a database of solar microflares that was used to construct the distribution. This database contains the signal at higher energy channels as well as that in channel zero (and below). From this one can, using software at GSFC, construct a photon spectrum for some of the larger microflares. It can also be used in other solar studies, especially those that correlate the X-ray flux with emission at other wavelengths. With some additional effort we hope to integrate this database into the corresponding one residing at the Solar Data Analysis Center at GSFC. The entire CGRO mission's data can now be reprocessed to obtain the microflare distribution at all phases of the solar cycle. This work is in progress. The results of this work will be presented in forthcoming scientific workshops and conferences.

Ryan, James M.↗

Solar Microflare with BATSE

Our work on detecting and cataloging solar microflares using an automated method is illustrated in the accompanying figure. The figure represents the solar microflare distribution during the period of April 1991 to November 1992, the height of solar activity after the launch of The Compton Gamma Ray Observatory (CGRO). It also shows the distribution extending below the distribution obtained at Goddard Space Flight Center (GSFC) by manual means. We have implemented significant refinements in the search algorithm. The algorithm in its simplest form searches for transient events and based upon the distribution of the signal among the different Burst and Transient Source Experiment (BATSE) detectors, we can assign it to be of solar origin if the signal distribution conforms to what one expects from a burst or transient from that direction. One of the major problems in the earlier effort was to search for microflares and large flares simultaneously. The requirement for a dynamic range of almost 10 (exp 4) resulted in ambiguous identifications at the low side of the distribution. We have since restricted the search to events with peak count rates under 2000 s (exp -1). Larger events are easily identified in the manual search, so we have chosen not to duplicate that work. The second problem was that missing counts existed below channel 0 in the Burst and Transient Source Experiment Large Area Detector data (BATSE LAD). These have been recovered and are now included in the search process. This provides data below 20 keV, and as we get closer to the thermal part of the spectrum, it provides greater sensitivity. The third problem was that too many BATSE detector were used in the search. Detectors with pointing directions far from the Sun, although detecting the event, had poorly known responses. Detectors greater than approximately 60 deg. off the Sun are no longer included in the search process. By reducing the systematic errors with the large off-axis detectors we can conduct more rigorous statistical tests of a candidate event to ascertain whether it originated from the solar direction. We have reprocessed the period in the early mission that covers solar maximum and constructed the microflare distribution shown in the figure. The results of the automated search start to deviate from the manual search results below about 1000 s (exp -1). Not only do we now have this distribution but we have a database of solar microflares that was used to construct the distribution. This database contains the signal at higher energy channels as well as that in channel zero (and below). From this one can, using software at GSFC, construct a photon spectrum for some of the larger microflares. It can also be used in other solar studies, especially those that correlate the X-ray flux with emission at other wavelengths. With some additional effort we hope to integrate this database into the corresponding one residing at the Solar Data Analysis Center at GSFC. The entire CGRO mission's data can now be reprocessed to obtain the microflare distribution at all phases of the solar cycle. This work is in progress. The results of this work will be presented in forthcoming scientific workshops and conferences.

Ryan, James M.↗

Heterogeneous distributed query processing: The DAVID system

The objective of the Distributed Access View Integrated Database (DAVID) project is the development of an easy to use computer system with which NASA scientists, engineers and administrators can uniformly access distributed heterogeneous databases. Basically, DAVID will be a database management system that sits alongside already existing database and file management systems. Its function is to enable users to access the data in other languages and file systems without having to learn the data manipulation languages. Given here is an outline of a talk on the DAVID project and several charts.

Jacobs, Barry E.↗

Relationship and distribution of Salmonella enterica serovar I 4,[5],12:i:- strain sequences in the NCBI Pathogen Detection database

Background: Of the > 2600 Salmonella serovars, Salmonella enterica serovar I 4,[5],12:i:- (serovar I 4,[5],12:i:-) has emerged as one of the most common causes of human salmonellosis and the most frequent multidrug-resistant (MDR; resistance to ≥3 antimicrobial classes) nontyphoidal Salmonella serovar in the U.S. Serovar I 4,[5],12:i:- isolates have been described globally with resistance to ampicillin, streptomycin, sulfisoxazole, and tetracycline (R-type ASSuT) and an integrative and conjugative element with multi-metal tolerance named Salmonella Genomic Island 4 (SGI-4). Results: We analyzed 13,612 serovar I 4,[5],12:i:- strain sequences available in the NCBI Pathogen Detection database to determine global distribution, animal sources, presence of SGI-4, occurrence of R-type ASSuT, frequency of antimicrobial resistance (AMR), and potential transmission clusters. Genome sequences for serovar I 4,[5],12:i:- strains represented 30 countries from 5 continents (North America, Europe, Asia, Oceania, and South America), but sequences from the United States (59%) and the United Kingdom (28%) were dominant. The metal tolerance island SGI-4 and the R-type ASSuT were present in 71 and 55% of serovar I 4,[5],12:i:- strain sequences, respectively. Sixty-five percent of strain sequences were MDR which correlates to serovar I 4,[5],12:i:- being the most frequent MDR serovar. The distribution of serovar I 4,[5],12:i:- strain sequences in the NCBI Pathogen Detection database suggests that swine-associated strain sequences were the most frequent food-animal source and were significantly more likely to contain the metal tolerance island SGI-4 and genes for MDR compared to all other animal-associated isolate sequences. Conclusions: Our study illustrates how analysis of genomic sequences from the NCBI Pathogen Detection database can be utilized to identify the prevalence of genetic features such as antimicrobial resistance, metal tolerance, and virulence genes that may be responsible for the successful emergence of bacterial foodborne pathogens.

59 BASIC BIOLOGICAL SCIENCES↗

Aerosol Optical Depth Distribution in Extratropical Cyclones over the Northern Hemisphere Oceans

Using Moderate Resolution Imaging Spectroradiometer and an extratropical cyclone database,the climatological distribution of aerosol optical depth (AOD) in extratropical cyclones is explored based solely on observations. Cyclone-centered composites of aerosol optical depth are constructed for the Northern Hemisphere mid-latitude ocean regions, and their seasonal variations are examined. These composites are found to be qualitatively stable when the impact of clouds and surface insolation or brightness is tested. The larger AODs occur in spring and summer and are preferentially found in the warm frontal and in the post-cold frontal regions in all seasons. The fine mode aerosols dominate the cold sector AODs, but the coarse mode aerosols display large AODs in the warm sector. These differences between the aerosol modes are related to the varying source regions of the aerosols and could potentially have different impacts on cloud and precipitation within the cyclones.

MODIS↗

Inferring Cirrus Size Distributions Through Satellite Remote Sensing and Microphysical Databases

Since cirrus clouds have a substantial influence on the global energy balance that depends on their microphysical properties, climate models should strive to realistically characterize the cirrus ice particle size distribution (PSD), at least in a climatological sense. To date, the airborne in situ measurements of the cirrus PSD have contained large uncertainties due to errors in measuring small ice crystals (D<60 m). This paper presents a method to remotely estimate the concentration of the small ice crystals relative to the larger ones using the 11- and 12- m channels aboard several satellites. By understanding the underlying physics producing the emissivity difference between these channels, this emissivity difference can be used to infer the relative concentration of small ice crystals. This is facilitated by enlisting temperature-dependent characterizations of the PSD (i.e., PSD schemes) based on in situ measurements. An average cirrus emissivity relationship between 12 and 11 m is developed here using the Moderate Resolution Imaging Spectroradiometer (MODIS) satellite instrument and is used to retrieve the PSD based on six different PSD schemes. The PSDs from the measurement-based PSD schemes are compared with corresponding retrieved PSDs to evaluate differences in small ice crystal concentrations. The retrieved PSDs generally had lower concentrations of small ice particles, with total number concentration independent of temperature. In addition, the temperature dependence of the PSD effective diameter De and fall speed Vf for these retrieved PSD schemes exhibited less variability relative to the unmodified PSD schemes. The reduced variability in the retrieved De and Vf was attributed to the lower concentrations of small ice crystals in the retrieved PSD.

Mitchell, David↗

Areal Distribution of the Oxygen-Isotope Ratio in Greenland

Mean values of the oxygen-isotope ratio relative to standard mean ocean water reported for 46 sites on the Greenland ice sheet are compiled together with data on mean annual surface temperature, latitude, 6180 elevation, and mean annual shortest distance to the open ocean denoted by the 10% sea-ice concentration boundary. Stepwise regression analyses, with 6180 as the dependent variable, define two robust models. In the forward mode at the 99.9% confidence level, only temperature enters the model. In the backward mode at the 95% confidence level, only temperature, latitude, and distance to the open ocean remain in the model. Inversions of the models on the basis of 160 gridpoint locations 100 km apart in the area delimited by the surface equilibrium line produce four contoured distributions of 6"0. Two distributions are based on the bivariate model and two on the multivariate model. The second distribution for each model is obtained substituting mean annual surface-temperature values obtained from the Nimbus-7 Temperature Humidity Infrared Radiometer (THIR) database. All four distributions are considered valid, and differences between them are evaluated using contoured anomaly maps. It is suggested that the inversion of the multivariate model using THIR data provides the more reliable pattern for studies of atmospheric advection or for the derivation of ice-flow adjustments for 6180 series obtained from deep-core or ablation-zone sites.

Zwally, H. Jay↗

NASA Image eXchange (NIX)

This paper discusses the technical aspects of and the project background for the NASA Image exchange (NIX). NIX, which provides a single entry point to search selected image databases at the NASA Centers, is a meta-search engine (i.e., a search engine that communicates with other search engines). It uses these distributed digital image databases to access photographs, animations, and their associated descriptive information (meta-data). NIX is available for use at the following URL: http://nix.nasa.gov./NIX, which was sponsored by NASAs Scientific and Technical Information (STI) Program, currently serves images from seven NASA Centers. Plans are under way to link image databases from three additional NASA Centers. images and their associated meta-data, which are accessible by NIX, reside at the originating Centers, and NIX utilizes a virtual central site that communicates with each of these sites. Incorporated into the virtual central site are several protocols to support searches from a diverse collection of database engines. The searches are performed in parallel to ensure optimization of response times. To augment the search capability, browse functionality with pre-defined categories has been built into NIX, thereby ensuring dissemination of 'best-of-breed' imagery. As a final recourse, NIX offers access to a help desk via an on-line form to help locate images and information either within the scope of NIX or from available external sources.

vonOfenheim. William H. C.↗

Open database for GPD analyses

This article summarizes the main ideas behind creating an open database proposed for use in the exploration of generalized parton distributions (GPDs). This lightweight database is well suited for GPD phenomenology and is designed to store both experimental and lattice-QCD data. It can also aid in benchmarking GPD-related developments, such as GPD models. The database utilizes a new data format based on the YAML serialization language, enabling the storage of essential information for modern analyses, such as replica values. It includes interfaces for both Python and C++, allowing straightforward integration with analysis codes.

Burkert, V. D. [Thomas Jefferson National Accelera↗

High Performance EVA Glove Collaboration: Glove Injury Data Mining Effort

Human hands play a significant role during extravehicular activity (EVA) missions and Neutral Buoyancy Lab (NBL) training events, as they are needed for translating and performing tasks in the weightless environment. It is because of this high frequency usage that hand- and arm-related injuries and discomfort are known to occur during training in the NBL and while conducting EVAs. Hand-related injuries and discomforts have been occurring to crewmembers since the days of Apollo. While there have been numerous engineering changes to the glove design, hand-related issues still persist. The primary objectives of this study are therefore to: 1) document all known EVA glove-related injuries and the circumstances of these incidents, 2) determine likely risk factors, and 3) recommend ergonomic mitigations or design strategies that can be implemented in the current and future glove designs. METHODS: The investigator team conducted an initial set of literature reviews, data mining of Lifetime Surveillance of Astronaut Health (LSAH) databases, and data distribution analyses to understand the ergonomic issues related to glove-related injuries and discomforts. The investigation focused on the injuries and discomforts of U.S. crewmembers who had worn pressurized suits and experienced glove-related incidents during the 1980 to 2010 time frame, either during training or on-orbit EVA. In addition to data mining of the LSAH database, the other objective of the study was to find complimentary sources of information such as training experience, EVA experience, suit-related sizing data, and hand-arm anthropometric data to be tied to the injury data from LSAH. RESULTS: Past studies indicated that the hand was the most frequently injured part of the body during both EVA and NBL training. This study effort thus focused primarily on crew training data in the NBL between 2002 and 2010. Of the 87 recorded training incidents, 19 occurred to women and 68 to men. While crew ages ranged from thirties to fifties, the age category most affected was in the forties range. Incident rate calculations (incidents per 100 training runs) revealed that the 2002, 2003, and 2004 time periods registered the highest reported incident rate levels (3.4, 6.1, and 4.1 respectively) when compared to the following years (all ≤ 1.0). In addition to general hand-arm discomfort being the highest reported result from training, specific types of hand injuries or symptoms included erythema, fingernail delamination, abrasions, muscle soreness/fatigue, paresthesia, bruising, blanching, and edema. Specific body locations most affected by hand injuries included the metacarpophalangeal joints, fingernails, finger crotches, fingers in general, interphalangeal joints, and fingertips. Causes of injuries reported in the LSAH data were primarily attributed to the forces that the gloved hands were exposed to due to hand intensive tasks and/or poor glove sizing. DISCUSSION: Although the age data indicate that most injuries are reported by male crewmembers in their forties, that is also the dominant gender and age range of most EVA crew therefore it is not an unexpected finding. Age and gender analysis will continue as more details on the uninjured population is accrued. While there is a reasonable mechanism to link training quantity to injury, the results were inconsistent and point to the need for a consistent method of suit-related injury screening and documentation. For instance, the high-incident rate levels for the years 2002 to 2004 could be attributed to a comprehensive medical review of crewmembers post-NBL EVA training that occurred from July 19, 2002 to January 16, 2004. Furthermore, there could have been increased awareness from an investigation at the NBL. These investigations may have temporarily increased the fidelity of reported injuries and discomforts during these dates as compared to surrounding years, when injury signs and symptom were no longer actively being investigated but rather voluntarily reported. Data mining for possible mechanistic factors continues and includes more detailed training timelines, hand anthropometry, and suit sizing information. The limited published data looking at hand-arm anthropometry correlated hand-anthropometry metrics with injuries stemming from glove design and operation. Future work will include further evaluation of body sizing and fit in relation to hand injury incidents.

Reid, C. R.↗

Development of a Data Fusion Methodology for Lineload Aerodynamic Databases for a Launch Vehicle during Liftoff and Transition

The need for databases for the distributed loading on launch vehicles during the early portion of flight necessitates the use of expensive computational flows in regimes where wake effects dominate. While also being expensive, this is a regime that computational tools tend to historically have problems simulating accurately. To help tackle this problem, a method of data fusion to combine computational results to wind tunnel derived force and moment data is developed. Using this method, significant reduction in computational costs and increases in confidence of the final product is possible and has been used to generate several databases for the Space Launch System (SLS) at NASA. While the full details of database generation are not part of this work, the crucial method at its core is developed here. Two SLS geometries are used throughout the work to demonstrate the techniques. These are two of the larger geometries and represent both planned crewed missions to the Moon as well as potential cargo missions to deep space. The method uses principal component analysis (PCA) to generate a reduced ordered model (ROM) to help fill in the full parameter space. Other similar techniques are explored, but were not found to have a significant result on the predictions of the ROM. Because the full number of components are kept to generate the model, this lack of difference is expected. This method is then extended to ensure that predicted surfaces match trusted force and moment data derived from wind tunnel testing. This extension is done by setting up a constrained optimization problem in order to minimize the deviation from the surface resolved computational data while still integrating to the desired values. When generating the constrained optimization problem, a weighting factor to balance these competing needs is introduced. The work compares previously introduced weighting terms from similar work to the proposed terms and shows that the previously used terms do not have as desirable behavior in this flow regime. This method is then expanded by developing a technique to incorporate uncertainty quantification into the developed data fusion methodology. This expansion takes a two pronged approach. One examines transferring the uncertainties in the force and moment database and characterizes how those adjustments change the predicted lineloads. The second looks at model form error and looks how rebuilding the model using slightly different data changes the predictions. These two terms are then combined in order to create an uncertainty model that takes both effects into account. The limitations of the proposed methods is then discussed as well as possible techniques to address these shortcomings.

Launch Vehicles↗

Transportability, distributability and rehosting experience with a kernel operating system interface set

For the past two years, PRC has been transporting and installing a software engineering environment framework, the Automated Product control Environment (APCE), at a number of PRC and government sites on a variety of different hardware. The APCE was designed using a layered architecture which is based on a standardized set of interfaces to host system services. This interface set called the APCE Interface Set (AIS), was designed to support many of the same goals as the Common Ada Programming Support Environment (APSE) Interface Set (CAIS). The APCE was developed to provide support for the full software lifecycle. Specific requirements of the APCE design included: automation of labor intensive administrative and logistical tasks: freedom for project team members to use existing tools: maximum transportability for APCE programs, interoperability of APCE database data, and distributability of both processes and data: and maximum performance on a wide variety of operating systems. A brief description is given of the APCE and AIS, a comparison of the AIS and CAIS both in terms of functionality and of philosophy and approach and a presentation of PRC's experience in rehosting AIS and transporting APCE programs and project data. Conclusions are drawn from this experience with respect to both the CAIS efforts and Space Station plans.

Blumberg, F. C.↗

The infrared astronomical satellite asteroid database

The creation, maintenance, and distribution of the set of ground-based observational data being used to help calculate the diameters and albedos of asteroids observed by IRAS during the sky survey of February-November 1983 are discussed in a status report. The database comprises orbital elements, absolute magnitudes, UBV colors, thermal radiometry, light-curve parameters, and taxonomic classifications; supplementary files include eight-color photometry, spectrophotometry, polarimetry, JHK observations, proper elements, discovery data, occultation diameters, radar characteristics, and references. Distribution in machine-readable and hard-copy formats and maintenance of the database by incorporating new data as they are published are planned.

Tedesco, E. F.↗