Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “distributed clustering methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Testing the Archivas Cluster (Arc) for Ozone Monitoring Instrument (OMI) Scientific Data Storage

The Ozone Monitoring Instrument (OMI) launched on NASA's Aura Spacecraft, the third of the major platforms of the EOS program on July 15,2004. In addition to the long term archive and distribution of the data from OM1 through the Goddard Earth Science Distributed Active Archive Center (GESDAAC), we are evaluating other archive mechanisms that can archive the data in a more immediately available method where it can be used for futher data production and analysis. In 2004, Archivas, Inc. was selected by NASA s Small Business Innovative Research (SBIR) program for the development of their Archivas Cluster (ArC) product. Arc is an online disk based system utilizing self-management and automation on a Linux cluster. Its goal is to produce a low cost solution coupled with the ease of management. The OM1 project is an application partner of the SBIR program, and has deployed a small cluster (5TB) based on the beta Archwas software. We performed extensive testing of the unit using production OM1 data since launch. In 2005, Archivas, Inc. was funded in SBIR Phase II for further development, which will include testing scalability with the deployment of a larger (35TB) cluster at Goddard. We plan to include Arc in the OM1 Team Leader Computing Facility (TLCF) hosting OM1 data for direct access and analysis by the OMI Science Team. This presentation will include a brief technical description of the Archivas Cluster, a summary of the SBIR Phase I beta testing results, and an overview of the OMI ground data processing architecture including its interaction with the Phase II Archivas Cluster and hosting of OMI data for the scientists.

Tilmes, Curt↗

The CLASSY clustering algorithm: Description, evaluation, and comparison with the iterative self-organizing clustering system (ISOCLS)

A clustering method, CLASSY, was developed, which alternates maximum likelihood iteration with a procedure for splitting, combining, and eliminating the resulting statistics. The method maximizes the fit of a mixture of normal distributions to the observed first through fourth central moments of the data and produces an estimate of the proportions, means, and covariances in this mixture. The mathematical model which is the basic for CLASSY and the actual operation of the algorithm is described. Data comparing the performances of CLASSY and ISOCLS on simulated and actual LACIE data are presented.

Lennington, R. K.↗

Aerosol Models for the CALIPSO Lidar Inversion Algorithms

We use measurements and models to develop aerosol models for use in the inversion algorithms for the Cloud Aerosol Lidar and Imager Pathfinder Spaceborne Observations (CALIPSO). Radiance measurements and inversions of the AErosol RObotic NETwork (AERONET1, 2) are used to group global atmospheric aerosols using optical and microphysical parameters. This study uses more than 105 records of radiance measurements, aerosol size distributions, and complex refractive indices to generate the optical properties of the aerosol at more 200 sites worldwide. These properties together with the radiance measurements are then classified using classical clustering methods to group the sites according to the type of aerosol with the greatest frequency of occurrence at each site. Six significant clusters are identified: desert dust, biomass burning, urban industrial pollution, rural background, marine, and dirty pollution. Three of these are used in the CALIPSO aerosol models to characterize desert dust, biomass burning, and polluted continental aerosols. The CALIPSO aerosol model also uses the coarse mode of desert dust and the fine mode of biomass burning to build a polluted dust model. For marine aerosol, the CALIPSO aerosol model uses measurements from the SEAS experiment 3. In addition to categorizing the aerosol types, the cluster analysis provides all the column optical and microphysical properties for each cluster.

Omar, Ali H.↗

Determination of the Cosmic Distance Scale from Sunyaev-Zel'dovich Effect and Chandra X-ray Measurements of High Redshift Galaxy Clusters

We determine the distance to 38 clusters of galaxies in the redshift range 0.14 less than or equal to z less than or equal to 0.89 using X-ray data from Chandra and Sunyaev-Zeldovich Effect data from the Owens Valley Radio Observatory and the Berkeley-Illinois-Maryland Association interferometric arrays. The cluster plasma and dark matter distributions are analyzed using a hydrostatic equilibrium model that accounts for radial variations in density, temperature and abundance, and, the statistical and systematic errors of this method are quantified. The analysis is performed via a Markov chain Monte Carlo technique that provides simultaneous estimation of all model parameters. W

Bonamente, Massimiliano↗

Synthesis of a laterally displaced cluster feed for a reflector antenna with application to multiple beams and contoured patterns

Two methods are described for efficiently synthesizing the excitation coefficients of a laterally displaced cluster feed in a reflector antenna subject to beam distortion. Applications are presented for rotationally symmetric paraboloids excited by an equilateral triangular array of feed elements. The basic cluster is a central element surrounded by a hexagonal ring. The first method - termed the sequential current method - determines a set of excitation coefficients which minimizes the phase distortion in the 'effective' aperture distribution of the reflector. The second method - termed the gradient optimization method - is such that the distortion in the secondary power pattern is directly minimized in a min-L2 form by a gradient optimization algorithm regarded as a systematic computer iteration procedure. Application to the synthesis of contour patterns is included.

Galindo-Israel, V.↗

An interactive Barnes objective map analysis scheme for use with satellite and conventional data

The Barnes (1973) objective map analysis scheme is employed to develop an interactive analysis package for assessing the impact of satellite-derived data on analyses of conventional meteorological data sets. The method permits modification of the values of input parameters in the objective analysis within objectively determined, internally set limits. The effects of the manipulations are rapidly displayed, and methods are included for assimilating the spatially clustered characteristics of satellite data and the various horizontal resolutions of the data types. Data sets from the SESAME rawinsonde wind data with uniform spatial distribution, with the same data set plus satellite cloud motion data, and a data set from the atmospheric sounder radiometer on the GOES satellite were analyzed as examples. The scheme is demonstrated to recover details after two iterations through the data.

Koch, S. E.↗

Dust extinction and molecular gas in the dark cloud IC 5146

In this paper we describe a powerful method for mapping the distribution of dust through a molecular cloud using data obtained in large-scale, multiwavelength, infrared imaging surveys. This method combines direct measurements of near-infrared color excess and certain techniques of star counting to derive mean extinctions and map the dust column density distribution through a cloud at higher angular resolutions and greater optical depths than those achieved previously by optical star counting. We report the initial results of the application of this method to a dark cloud complex near the cluster IC 5146, where we have performed coordinated, near-infrared, JHK imaging and (13)CO, C(18)O, and CS millimeter-wave, molecular-line surveys of a large portion of the complex. More than 4000 stars were detected in our JHK survey of the cloud. Of these, all but about a dozen appear to be field stars not associated with the cloud. Star count maps at J band show a striking and detailed anticorrelation between the surface density of J-band sources and CO and CS molecular-line emission. We used the (H-K) colors and positions of nearly 1300 sources to directly measure and map the extinction and thus trace the dust column density through the cloud at an effective angular resolution of 1 min .5. We report an interesting correlation between the measured dispersion in our extinction determinations and the extinction. Modeling this relation indicates that effects of small-scale cloud structure dominate the uncertainties in our measurements. Moreover, we demonstrate that such observations can be used to place constraints on the nature of the spatial distribution of extinction on scales smaller than our resolution. In particular, we show that models in which the dust is distributed uniformly or in discrete high-extinction clumps on scales smaller than (1 min .5) are inconsistent with the observations. We have derived extinctions at the same positions and at the same angular resolution (1 min .7) as our molecular-line observations. This enabled a direct comparison of (13)CO, C(18)O, and CS integrated intensities and column densities with A(sub V) for more than 500 positions in the cloud, corresponding to a range in A(sub V) between 0 to 32 mag of extinction. We found the integrated intensities of (13)CO, C(18)O, and CS to be roughly linearly correlated with extinction over different ranges of extinction. However, for all three molecules we find the scatter in the observed relations to be larger than can be accounted for by instrumental error, suggesting that there are large intrinsic variations in the abundances or excitation of the molecules through the cloud. Mean abundances for all the molecules relative to hydrogen were directly derived from the data. The ratio of (13)CO to C(18)O abundances was found to be significantly higher than the terrestrial ratio in regions where extinction is less than 10 mag. In the same region, the dispersion in the abundance ratio is also found to be very large, suggesting that the abundances of one or both molecules are very unstable even at relatively large cloud optical depths. Beyond 10 mag of extinction the abundances of both species appear very stable with their ratio close to the terrestrial value.

Lada, Charles J.↗

A new solution-adaptive grid generation method for transonic airfoil flow calculations

The clustering algorithm is controlled by a second-order, ordinary differential equation which uses the airfoil surface density gradient as a forcing function. The solution to this differential equation produces a surface grid distribution which is automatically clustered in regions with large gradients. The interior grid points are established from this surface distribution by using an interpolation scheme which is fast and retains the desirable properties of the original grid generated from the standard elliptic equation approach.

Nakamura, S.↗

Groups of galaxies in the ROSAT north ecliptic pole survey

The X-ray properties of groups of galaxies are presented. Their distribution of luminosity and temperature appears to be associated with the extrapolation of these distributions from rich clusters of galaxies. The properties of the ensemble of groups of galaxies are almost totally unknown. Only a few X-ray observations of groups that were selected by optical methods were published so far. A sample of eight groups with 'z' inferior to 0.04, of which three have 'z' inferior to 0.03 was investigated. The temperature and the luminosity functions at one point were determined.

Henry, J. Patrick↗

The Effect of Approximating Some Molecular Integrals in Coupled-Cluster Calculations: Fundamental Frequencies and Rovibrational Spectroscopic Constants of Cyclopropenylidene

The singles and doubles coupled-cluster method that includes a perturbational estimate of connected triple excitations, denoted CCSD(T), has been used, in conjunction with approximate integral techniques, to compute highly accurate rovibrational spectroscopic constants of cyclopropenylidene, C3H2. The approximate integral technique was proposed in 1994 by Rendell and Lee in order to avoid disk storage and input/output bottlenecks, and today it will also significantly aid in the development of algorithms for distributed memory, massively parallel computer architectures. It is shown in this study that use of approximate integrals does not impact the accuracy of CCSD(T) calculations. In addition, the most accurate spectroscopic data yet for C3H2 is presented based on a CCSD(T)/cc-pVQZ quartic force field that is modified to include the effects of core-valence electron correlation. Cyclopropenylidene is of great astronomical and astrobiological interest because it is the smallest aromatic ringed compound to be positively identified in the interstellar medium, and is thus involved in the prebiotic processing of carbon and hydrogen. The singles and doubles coupled-cluster method that includes a perturbational estimate of

Lee, Timothy J.↗

3D Drop Size Distribution Extrapolation Algorithm Using a Single Disdrometer

Determining the Z-R relationship (where Z is the radar reflectivity factor and R is rainfall rate) from disdrometer data has been and is a common goal of cloud physicists and radar meteorology researchers. The usefulness of this quantity has traditionally been limited since radar represents a volume measurement, while a disdrometer corresponds to a point measurement. To solve that problem, a 3D-DSD (drop-size distribution) method of determining an equivalent 3D Z-R was developed at the University of Central Florida and tested at the Kennedy Space Center, FL. Unfortunately, that method required a minimum of three disdrometers clustered together within a microscale network (.1-km separation). Since most commercial disdrometers used by the radar meteorology/cloud physics community are high-cost instruments, three disdrometers located within a microscale area is generally not a practical strategy due to the limitations of these kinds of research budgets. A relatively simple modification to the 3D-DSD algorithm provides an estimate of the 3D-DSD and therefore, a 3D Z-R measurement using a single disdrometer. The basis of the horizontal extrapolation is mass conservation of a drop size increment, employing the mass conservation equation. For vertical extrapolation, convolution of a drop size increment using raindrop terminal velocity is used. Together, these two independent extrapolation techniques provide a complete 3DDSD estimate in a volume around and above a single disdrometer. The estimation error is lowest along a vertical plane intersecting the disdrometer position in the direction of wind advection. This work demonstrates that multiple sensors are not required for successful implementation of the 3D interpolation/extrapolation algorithm. This is a great benefit since it is seldom that multiple sensors in the required spatial arrangement are available for this type of analysis. The original software (developed at the University of Central Florida, 1998.- 2000) has also been modified to read standardized disdrometer data format (Joss-Waldvogel format). Other modifications to the software involve accounting for vertical ambient wind motion, as well as evaporation of the raindrop during its flight time.

Lane, John↗

Collaborative Clustering for Sensor Networks

Traditionally, nodes in a sensor network simply collect data and then pass it on to a centralized node that archives, distributes, and possibly analyzes the data. However, analysis at the individual nodes could enable faster detection of anomalies or other interesting events, as well as faster responses such as sending out alerts or increasing the data collection rate. There is an additional opportunity for increased performance if individual nodes can communicate directly with their neighbors. Previously, a method was developed by which machine learning classification algorithms could collaborate to achieve high performance autonomously (without requiring human intervention). This method worked for supervised learning algorithms, in which labeled data is used to train models. The learners collaborated by exchanging labels describing the data. The new advance enables clustering algorithms, which do not use labeled data, to also collaborate. This is achieved by defining a new language for collaboration that uses pair-wise constraints to encode useful information for other learners. These constraints specify that two items must, or cannot, be placed into the same cluster. Previous work has shown that clustering with these constraints (in isolation) already improves performance. In the problem formulation, each learner resides at a different node in the sensor network and makes observations (collects data) independently of the other learners. Each learner clusters its data and then selects a pair of items about which it is uncertain and uses them to query its neighbors. The resulting feedback (a must and cannot constraint from each neighbor) is combined by the learner into a consensus constraint, and it then reclusters its data while incorporating the new constraint. A strategy was also proposed for cleaning the resulting constraint sets, which may contain conflicting constraints; this improves performance significantly. This approach has been applied to collaborative clustering of seismic and infrasonic data collected by the Mount Erebus Volcano Observatory in Antarctica. Previous approaches to distributed clustering cannot readily be applied in a sensor network setting, because they assume that each node has the same view of the data set. A view is the set of features used to represent each object. When a single data set is partitioned across several computational nodes, distributed clustering works; all objects have the same view. But when the data is collected from different locations, using different sensors, a more flexible approach is needed. This approach instead operates in situations where the data collected at each node has a different view (e.g., seismic vs. infrasonic sensors), but they observe the same events. This enables them to exchange information about the likely cluster membership relations between objects, even if they do not use the same features to represent the objects.

Wagstaff. Loro :/↗

A method of using cluster analysis to study statistical dependence in multivariate data

A technique is presented that uses both cluster analysis and a Monte Carlo significance test of clusters to discover associations between variables in multidimensional data. The method is applied to an example of a noisy function in three-dimensional space, to a sample from a mixture of three bivariate normal distributions, and to the well-known Fisher's Iris data.

Borucki, W. J.↗

Porting a Hall MHD Code to a Graphic Processing Unit

We present our experience porting a Hall MHD code to a Graphics Processing Unit (GPU). The code is a 2nd order accurate MUSCL-Hancock scheme which makes use of an HLL Riemann solver to compute numerical fluxes and second-order finite differences to compute the Hall contribution to the electric field. The divergence of the magnetic field is controlled with Dedner?s hyperbolic divergence cleaning method. Preliminary benchmark tests indicate a speedup (relative to a single Nehalem core) of 58x for a double precision calculation. We discuss scaling issues which arise when distributing work across multiple GPUs in a CPU-GPU cluster.

Dorelli, John C.↗

The Gas Distribution in Galaxy Cluster Outer Regions

Aims. We present the analysis of a local (z = 0.04 - 0.2) sample of 31 galaxy clusters with the aim of measuring the density of the X-ray emitting gas in cluster outskirts. We compare our results with numerical simulations to set constraints on the azimuthal symmetry and gas clumping in the outer regions of galaxy clusters. Methods. We exploit the large field-of-view and low instrumental background of ROSAT/PSPC to trace the density of the intracluster gas out to the virial radius. We perform a stacking of the density profiles to detect a signal beyond r200 and measure the typical density and scatter in cluster outskirts. We also compute the azimuthal scatter of the profiles with respect to the mean value to look for deviations from spherical symmetry. Finally, we compare our average density and scatter profiles with the results of numerical simulations. Results. As opposed to some recent Suzaku results, and confirming previous evidence from ROSAT and Chandra, we observe a steepening of the density profiles beyond approximately r(sub 500). Comparing our density profiles with simulations, we find that non-radiative runs predict too steep density profiles, whereas runs including additional physics and/or treating gas clumping are in better agreement with the observed gas distribution. We report for the first time the high-confidence detection of a systematic difference between cool-core and non-cool core clusters beyond 0.3r(sub 200), which we explain by a different distribution of the gas in the two classes. Beyond r(sub 500), galaxy clusters deviate significantly from spherical symmetry, with only little differences between relaxed and disturbed systems. We find good agreement between the observed and predicted scatter profiles, but only when the 1% densest clumps are filtered out in the simulations. Conclusions. Comparing our results with numerical simulations, we find that non-radiative simulations fail to reproduce the gas distribution, even well outside cluster cores. Although their general behavior is in better agreement with the observations, simulations including cooling and star formation convert a large amount of gas into stars, which results in a low gas fraction with respect to the observations. Consequently, a detailed treatment of gas cooling, star formation, AGN feedback, and taking into account gas clumping is required to construct realistic models of cluster outer regions.

Eckert, D.↗

On Multi-Dimensional Unstructured Mesh Adaption

Anisotropic unstructured mesh adaption is developed for a truly multi-dimensional upwind fluctuation splitting scheme, as applied to scalar advection-diffusion. The adaption is performed locally using edge swapping, point insertion/deletion, and nodal displacements. Comparisons are made versus the current state of the art for aggressive anisotropic unstructured adaption, which is based on a posteriori error estimates. Demonstration of both schemes to model problems, with features representative of compressible gas dynamics, show the present method to be superior to the a posteriori adaption for linear advection. The performance of the two methods is more similar when applied to nonlinear advection, with a difference in the treatment of shocks. The a posteriori adaption can excessively cluster points to a shock, while the present multi-dimensional scheme tends to merely align with a shock, using fewer nodes. As a consequence of this alignment tendency, an implementation of eigenvalue limiting for the suppression of expansion shocks is developed for the multi-dimensional distribution scheme. The differences in the treatment of shocks by the adaption schemes, along with the inherently low levels of artificial dissipation in the fluctuation splitting solver, suggest the present method is a strong candidate for applications to compressible gas dynamics.

Wood, William A.↗

The Gas Distribution in the Outer Regions of Galaxy Clusters

Aims. We present our analysis of a local (z = 0.04 - 0.2) sample of 31 galaxy clusters with the aim of measuring the density of the X-ray emitting gas in cluster outskirts. We compare our results with numerical simulations to set constraints on the azimuthal symmetry and gas clumping in the outer regions of galaxy clusters. Methods. We have exploited the large field-of-view and low instrumental background of ROSAT/PSPC to trace the density of the intracluster gas out to the virial radius, We stacked the density profiles to detect a signal beyond T200 and measured the typical density and scatter in cluster outskirts. We also computed the azimuthal scatter of the profiles with respect to the mean value to look for deviations from spherical symmetry. Finally, we compared our average density and scatter profiles with the results of numerical simulations. Results. As opposed to some recent Suzaku results, and confirming previous evidence from ROSAT and Chandra, we observe a steepening of the density profiles beyond approximately r(sub 500). Comparing our density profiles with simulations, we find that non-radiative runs predict density profiles that are too steep, whereas runs including additional physics and/ or treating gas clumping agree better with the observed gas distribution. We report high-confidence detection of a systematic difference between cool-core and non cool-core clusters beyond approximately 0.3r(sub 200), which we explain by a different distribution of the gas in the two classes. Beyond approximately r(sub 500), galaxy clusters deviate significantly from spherical symmetry, with only small differences between relaxed and disturbed systems. We find good agreement between the observed and predicted scatter profiles, but only when the 1% densest clumps are filtered out in the ENZO simulations. Conclusions. Comparing our results with numerical simulations, we find that non-radiative simulations fail to reproduce the gas distribution, even well outside cluster cores. Although their general behavior agrees more closely with the observations, simulations including cooling and star formation convert a large amount of gas into stars, which results in a low gas fraction with respect to the observations. Consequently, a detailed treatment of gas cooling, star formation, AGN feedback, and consideration of gas clumping is required to construct realistic models of the outer regions of clusters.

Eckert, D.↗

On star formation in stellar systems. I - Photoionization effects in protoglobular clusters

The progressive ionization and subsequent dynamical evolution of nonhomogeneously distributed low-metal-abundance diffuse gas after star formation in globular clusters are investigated analytically, taking the gravitational acceleration due to the stars into account. The basic equations are derived; the underlying assumptions, input parameters, and solution methods are explained; and numerical results for three standard cases (ionization during star formation, ionization during expansion, and evolution resulting in a stable H II region at its equilibrium Stromgren radius) are presented in graphs and characterized in detail. The time scale of residual-gas loss in typical clusters is found to be about the same as the lifetime of a massive star on the main sequence.

Tenorio-Tagle, G.↗