Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “clustering statistics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Mimas: Preliminary Evidence For Amorphous Water Ice from VIMS

We have conducted a statistical clustering analysis (1,2) on a mosaic of VIMS data cubes obtained on February 13, 2010, for Saturn s satellite Mimas. Seven VIMS cubes were geometrically projected and re-sampled to a common spatial resolution. The clustering technique consists of a partitioning algorithm coupled to a criterion that prevents sub-optimal solutions and tests for the influence of random noise in the measurements. The clustering technique is agnostic about the meaning of the clusters, and scientific interpretation requires their a posteriori evaluation. The preliminary results yielded five clusters, demonstrating that spectral variability across Mimas surface is statistically significant. The ratios of the means calculated for each of the clusters show structure within the 1.6- micron water ice band, as well as the shape and the central wavelength of the strong ice band at 2 micron, that map spatially in patterns apparently related to the topography of Mimas, in particular certain regions in and around Herschel crater. The mean spectra of the five clusters, show similarities with laboratory spectra of amorphous and crystalline H2O ice (3) that are suggestive of the presence of an amorphous ice component in certain regions of Mimas, notably on the central peak of Herschel, on the crater floor, and in faults surrounding the crater. This may represent a mixture of both ice phases, or perhaps a layer of amorphous ice on a base of crystalline ice. Another possible occurrence of amorphous ice appears southwest of Herschel, close to the south pole.

Cruikshank, Dale P.↗

The composite sequential clustering technique for analysis of multispectral scanner data

The clustering technique consists of two parts: (1) a sequential statistical clustering which is essentially a sequential variance analysis, and (2) a generalized K-means clustering. In this composite clustering technique, the output of (1) is a set of initial clusters which are input to (2) for further improvement by an iterative scheme. This unsupervised composite technique was employed for automatic classification of two sets of remote multispectral earth resource observations. The classification accuracy by the unsupervised technique is found to be comparable to that by traditional supervised maximum likelihood classification techniques. The mathematical algorithms for the composite sequential clustering program and a detailed computer program description with job setup are given.

Su, M. Y.↗

ADDGALS: Simulated Sky Catalogs for Wide Field Galaxy Surveys

Abstract We present a method for creating simulated galaxy catalogs with realistic galaxy luminosities, broadband colors, and projected clustering over large cosmic volumes. The technique, denoted Addgals (Adding Density Dependent GAlaxies to Lightcone Simulations), uses an empirical approach to place galaxies within lightcone outputs of cosmological simulations. It can be applied to significantly lower-resolution simulations than those required for commonly used methods such as halo occupation distributions, subhalo abundance matching, and semi-analytic models, while still accurately reproducing projected galaxy clustering statistics down to scales of r ∼ 100 h −1 kpc . We show that Addgals catalogs reproduce several statistical properties of the galaxy distribution as measured by the Sloan Digital Sky Survey (SDSS) main galaxy sample, including galaxy number densities, observed magnitude and color distributions, as well as luminosity- and color-dependent clustering. We also compare to cluster–galaxy cross correlations, where we find significant discrepancies with measurements from SDSS that are likely linked to artificial subhalo disruption in the simulations. Applications of this model to simulations of deep wide-area photometric surveys, including modeling weak-lensing statistics, photometric redshifts, and galaxy cluster finding, are presented in DeRose et al., and an application to a full cosmology analysis of Dark Energy Survey (DES) Year 3 like data is presented in DeRose et al. We plan to publicly release a 10,313 square degree catalog constructed using Addgals with magnitudes appropriate for several existing and planned surveys, including SDSS, DES, VISTA, Wide-field Infrared Survey Explorer, and Rubin Observatory’s Legacy Survey of Space and Time.

79 ASTRONOMY AND ASTROPHYSICS↗

Unsupervised classification of earth resources data.

A new clustering technique is presented. It consists of two parts: (a) a sequential statistical clustering which is essentially a sequential variance analysis and (b) a generalized K-means clustering. In this composite clustering technique, the output of (a) is a set of initial clusters which are input to (b) for further improvement by an iterative scheme. This unsupervised composite technique was employed for automatic classification of two sets of remote multispectral earth resource observations. The classification accuracy by the unsupervised technique is found to be comparable to that by existing supervised maximum liklihood classification technique.

Su, M. Y.↗

A search for extended halos of hot gas in the Perseus, Virgo, and Coma Clusters

Observations of the Perseus cluster by the HEAO 1 satellite have revealed a faint X-ray halo extending at least 2.5 deg from the center and contributing between 5% and 20% to the total luminosity. This may be of nonthermal origin, but it also may be explained in terms of hot gas bound by the gravitational field of the cluster. Statistical uncertainties made it impossible to detect any such halo in the Coma cluster. Observations of the Virgo cluster confirmed the detection by the Ariel 5 satellite of a broad region of faint X-ray emission (core radius 60 arcmin). If the very extended X-ray emission from Virgo is due to hot intracluster gas, the density of this gas is lower than expected from a consideration of gas and galaxy densities in the Perseus cluster.

Ulmer, M. P.↗

DESI DR2 Reference Mocks: Clustering results from UCHUU ELGs and QSOs

High-redshift galaxy clustering provides a powerful probe of the growth of structure, testing models of dark matter, dark energy, and galaxy formation during the epoch when the Universe was rapidly evolving. Emission line galaxies (ELGs) and quasars (QSOs) are used as tracers of dark matter by the Dark Energy Spectroscopic Instrument (DESI) to probe this redshift regime. We present results from ELG and QSO mock catalogs created from the Uchuu N-body simulation and tuned to DESI Data Release 2 (DR2) clustering. Employing a modified subhalo abundance matching (SHAM) technique, we populate Uchuu halos and subhalos with QSOs between 0.8 < z < 2.1. For ELGs, we modify this method to select satellite galaxies with low velocities relative to their associated central halos, and populate a separate set of Uchuu halos and subhalos with ELGs between 0.8 < z < 1.6. In this paper, we reproduce the redshift evolution of number density and clustering statistics across the fitted range of scales. We also measure the large-scale clustering bias of both the data and mock samples. These results improve simulated lightcone construction from cosmological models and enhance our understanding of the galaxy-halo connection.

Vaisakh, R. [Southern Methodist U.] (ORCID:0009000↗

Galaxy-multiplet clustering from DESI DR2

We present an efficient estimator for higher-order galaxy clustering using small groups of nearby galaxies, or multiplets. Using the Luminous Red Galaxy (LRG) sample from the Dark Energy Spectroscopic Instrument (DESI) Data Release 2, we identify galaxy multiplets as discrete objects and measure their cross-correlations with the general galaxy field. Our results show that the multiplets exhibit stronger clustering bias as they trace more massive dark matter halos than individual galaxies. When comparing the observed clustering statistics with the mock catalogs generated from the N-body simulation AbacusSummit, we find that the mocks underpredict multiplet clustering despite reproducing the galaxy two-point auto-correlation reasonably well. This discrepancy indicates that the standard Halo Occupation Distribution (HOD) model is insufficient to describe the properties of galaxy multiplets, revealing the greater constraining power of this higher-order statistic on galaxy-halo connection and the possibility that multiplets are specific to additional assembly bias. We demonstrate that incorporating secondary biases into the HOD model improves agreement with the observed multiplet statistics, specifically by allowing galaxies to preferentially occupy halos in denser environments. Our results highlight the potential of utilizing multiplet clustering, beyond traditional two-point correlation measurements, to break degeneracies in models describing the galaxy-dark matter connection.

cosmology↗

Production of alternate realizations of DESI fiber assignment for unbiased clustering measurement in data and simulations

A critical requirement of spectroscopic large scale structure analyses is correcting for selection of which galaxies to observe from an isotropic target list. This selection is often limited by the hardware used to perform the survey which will impose angular constraints of simultaneously observable targets, requiring multiple passes to observe all of them. In SDSS this manifested solely as the collision of physical fibers and plugs placed in plates. In DESI, there is the additional constraint of the robotic positioner which controls each fiber being limited to a finite patrol radius. A number of approximate methods have previously been proposed to correct the galaxy clustering statistics for these effects, but these generally fail on small scales. To accurately correct the clustering we need to upweight pairs of galaxies based on the inverse probability that those pairs would be observed (Bianchi & Percival 2017). This paper details an implementation of that method to correct the Dark Energy Spectroscopic Instrument (DESI) survey for incompleteness. To calculate the required probabilities, we need a set of alternate realizations of DESI where we vary the relative priority of otherwise identical targets. These realizations take the form of alternate Merged Target Ledgers (AMTL), the files that link DESI observations and targets. We present the method used to generate these alternate realizations and how they are tracked forward in time using the real observational record and hardware status, propagating the survey as though the alternate orderings had been adopted. We detail the first applications of this method to the DESI One-Percent Survey (SV3) and the DESI year 1 data. We include evaluations of the pipeline outputs, estimation of survey completeness from this and other methods, and validation of the method using mock galaxy catalogs.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

An observational view of large scale structure

A summary of recent observations of galaxy clustering is presented, including a brief review of redshift maps and galaxy clustering statistics. Simple arguments are presented that argue the underlying mass fluctuations are most likely associated with a clustering scale no larger than that of individual galaxies. The acceleration of the local group from the comoving frame of the universe and its connection to the microwave dipole anisotropy are also discussed. A final topic for consideration is the existence of large voids and clusters, and whether they are consistent with Gaussian initial conditions. The extreme size and depth of the Bootes void, if real, do present a puzzle. Finally, future directions for observational study of large scale structure are discussed.

Davis, Marc↗

The impact of anisotropic redshift distributions on angular clustering

A leading way to constrain physical theories from cosmological observations is to test their predictions for the angular clustering statistics of matter tracers, a technique that is set to become ever more central with the next generation of large imaging surveys. Interpretation of this clustering requires knowledge of the projection kernel, or the redshift distribution of the sources, and the typical assumption is an isotropic redshift distribution for the objects. However, variations in the kernel are expected across the survey footprint due to photometric variations and residual observational systematic effects. Here, we develop the formalism for anisotropic projection and present several limiting cases that elucidate the key aspects. We quantify the impact of anisotropies in the redshift distribution on a general class of angular two-point statistics. In particular, we identify a mode-coupling effect that can add power to auto-correlations, including galaxy clustering and cosmic shear, and remove it from certain cross-correlations. If the projection anisotropy is primarily at large scales, the mode-coupling depends upon its variance as a function of redshift; furthermore, it is often of similar shape to the signal. In contrast, the cross-correlation of a field whose selection function is anisotropic with another one featuring no such variations — such as CMB lensing — is immune to these effects. We discuss explicitly several special cases of the general formalism including galaxy clustering, galaxy-galaxy lensing, cosmic shear and cross-correlations with CMB lensing, and publicly release a code to compute the biases.

79 ASTRONOMY AND ASTROPHYSICS↗

Covariance matrices for variance-suppressed simulations

ABSTRACT Cosmological N-body simulations provide numerical predictions of the structure of the Universe against which to compare data from ongoing and future surveys, but the growing volume of the Universe mapped by surveys requires correspondingly lower statistical uncertainties in simulations, usually achieved by increasing simulation sizes at the expense of computational power. It was recently proposed to reduce simulation variance without incurring additional computational costs by adopting fixed-amplitude initial conditions. This method has been demonstrated not to introduce bias in various statistics, including the two-point statistics of galaxy samples typically used for extracting cosmological parameters from galaxy redshift survey data, but requires us to revisit current methods for estimating covariance matrices of clustering statistics for simulations. In this work, we find that it is not trivial to construct covariance matrices analytically for fixed-amplitude simulations, but we demonstrate that ezmock (Effective Zel’dovich approximation mock catalogue), the most efficient method for constructing mock catalogues with accurate two- and three-point statistics, provides reasonable covariance matrix estimates for such simulations. We further examine how the variance suppression obtained by amplitude-fixing depends on three-point clustering, small-scale clustering, and galaxy bias, and propose intuitive explanations for the effects we observe based on the ezmock bias model.

79 ASTRONOMY AND ASTROPHYSICS↗

Land cover stratification using Landsat Thematic Mapper data in Sahelian and Sudanian woodland and wooded grassland

A standard methodology for thematic mapping of natural vegetation using remotely sensed imagery and digital image processing was modified to account for the spatial and spectral properties of semi-arid landscapes, and tested in study areas in the Sahelian and Sudanian zones, Mali. A principal components transformation of registered wet and dry season Landsat TM images produced a set of synthetic spectral channels differentiating vegetation cover between seasons, and allowed areas with annual grass growth to be distinguished from areas with woody cover. The transformed data were statistically clustered and clusters were assigned to vegetation type and density categories. In a separate step, the images were manually interpreted to differentiate broad soil classes. Four statistics were compared to evaluate the accuracy of the maps based on sample points from air photos. For the relatively detailed categories initially defined, map accuracies were substandard; however, when vegetation density classes were aggregated, overall accuracy was around 90 percent, and class accuracy was greater than 80 percent for most classes. This method is suitable for stratification and inventory of woody biomass at a regional scale in semi-arid woodland and wooded grassland.

Franklin, J.↗

Disentangling Structures in the Cluster of Galaxies Abell 133

A dynamical analysis of the structure of the cluster of galaxies Abell 133 will be presented using multi-wavelength data combined from multiple space and earth based observations. New and familiar statistical clustering techniques are used in combination in an attempt to gain a fully consistent picture of this interesting nearby cluster of galaxies. The type of analysis presented should be typical of cluster studies in the future, especially those to come from the surveys like the Sloan Digital Sky Survey and the 2DF.

Way, Michael J.↗

Precision redshift-space galaxy power spectra using Zel'dovich control variates

Numerical simulations in cosmology require trade-offs between volume, resolution and run-time that limit the volume of the Universe that can be simulated, leading to sample variance in predictions of ensemble-average quantities such as the power spectrum or correlation function(s). Sample variance is particularly acute at large scales, which is also where analytic techniques can be highly reliable. This provides an opportunity to combine analytic and numerical techniques in a principled way to improve the dynamic range and reliability of predictions for clustering statistics. In this paper we extend the technique of Zel'dovich control variates, previously demonstrated for 2-point functions in real space, to reduce the sample variance in measurements of 2-point statistics of biased tracers in redshift space. We demonstrate that with this technique, we can reduce the sample variance of these statistics down to their shot-noise limit out to k ~ 0.2 h Mpc -1 . This allows a better matching with perturbative models and improved predictions for the clustering of e.g. quasars, galaxies and neutral Hydrogen measured in spectroscopic redshift surveys at very modest computational expense. We discuss the implementation of ZCV, give some examples and provide forecasts for the efficacy of the method under various conditions.

79 ASTRONOMY AND ASTROPHYSICS↗

Modeling Redshift-space Clustering with Abundance Matching

Abstract We explore the degrees of freedom required to jointly fit projected and redshift-space clustering of galaxies selected in three bins of stellar mass from the Sloan Digital Sky Survey Main Galaxy Sample (SDSS MGS) using a subhalo abundance matching (SHAM) model. We employ emulators for relevant clustering statistics in order to facilitate our analysis, leading to large speed gains with minimal loss of accuracy. We are able to simultaneously fit the projected and redshift-space clustering of the two most massive galaxy samples that we consider with just two free parameters: scatter in stellar mass at fixed SHAM proxy, and the dependence of the SHAM proxy on dark matter halo concentration. We find some evidence for models that include velocity bias, but including orphan galaxies improves our fits to the lower-mass samples significantly. We also model the clustering signals of specific star formation rate (sSFR) selected samples using conditional abundance matching (CAM). We obtain acceptable fits to projected and redshift-space clustering as a function of sSFR and stellar mass using two CAM variants, although the fits are worse than for stellar-mass-selected samples alone. By incorporating nonunity correlations between the CAM proxy and sSFR, we are able to resolve previously identified discrepancies between CAM predictions and SDSS observations of the environmental dependence of quenching for isolated central galaxies.

79 ASTRONOMY AND ASTROPHYSICS↗

Seasonal- and Beta-Angle-Dependent Latitude Bias Variations in Natural Decays

Prior work has demonstrated pronounced statistical clustering of natural decays of medium-to-high-inclination orbital objects peaking approximately 30 degrees in Argument of Latitude ahead of nodal crossings. This effect is caused by the physical bulge in the Earth and the overlying atmosphere, that cyclically modifies effective altitude (and therefore density) faster than the trajectory's decay itself. While prior work has averaged seasonal and RAAN effects over all non-uniform atmosphere possibilities to support long-term characterization of the clustering of final entries in generating a pre-mission Expectation of Casualty, the current study characterizes seasonal and beta angle effects as potential influences on the near-term statistical risks of specific tactical decay scenarios, relative to the average. Such effects on the density profile along an orbit may be important considerations in any scenario where small orbital adjustments are used to optimize the timing and location of final entry trajectories. I.E., two identical spacecraft entering in different seasons and/or beta angles may have different minimum-risk scenarios for identical control capabilities and space weather conditions. Further, the early heating history of shallow trajectories is explored, examining the influence of dramatically different density profiles over the final orbit as the spacecraft either skims over or dives into the atmosphere.

Bacon, John B.↗

Mapping of terrain by computer clustering techniques using multispectral scanner data and using color aerial film

Two clustering techniques were used for terrain mapping by computer of test sites in Yellowstone National Park. One test was made with multispectral scanner data using a composite technique which consists of (1) a strictly sequential statistical clustering which is a sequential variance analysis, and (2) a generalized K-means clustering. In this composite technique, the output of (1) is a first approximation of the cluster centers. This is the input to (2) which consists of steps to improve the determination of cluster centers by iterative procedures. Another test was made using the three emulsion layers of color-infrared aerial film as a three-band spectrometer. Relative film densities were analyzed using a simple clustering technique in three-color space. Important advantages of the clustering technique over conventional supervised computer programs are (1) human intervention, preparation time, and manipulation of data are reduced, (2) the computer map, gives unbiased indication of where best to select the reference ground control data, (3) use of easy to obtain inexpensive film, and (4) the geometric distortions can be easily rectified by simple standard photogrammetric techniques.

Smedes, H. W.↗

Mitigating imaging systematics for DESI 2024 emission Line Galaxies and beyond

Emission Line Galaxies (ELGs) are one of the main tracers that the Dark Energy Spectroscopic Instrument (DESI) uses to probe the universe. However, they are afflicted by strong spurious correlations between target density and observing conditions known as imaging systematics. In this paper, we present the imaging systematics mitigation applied to the DESI Data Release 1 (DR1) large-scale structure catalogs used in the DESI 2024 cosmological analyses. We also explore extensions of the fiducial treatment. This includes a combined approach, through forward image simulations (Obiwan) in conjunction with neural network-based regression, to obtain an angular selection function that mitigates the imaging systematics observed in the DESI DR1 ELGs target density. We further derive a line of sight selection function from the forward model that removes the strong redshift dependence between imaging systematics and low redshift ELGs. Combining both angular and redshift-dependent systematics, we construct a three-dimensional selection function and assess the impact of all selection functions on clustering statistics. We quantify differences between these extended treatments and the fiducial treatment in terms of the measured 2-point statistics. We find that the results are generally consistent with the fiducial treatment and conclude that the differences are far less than the imaging systematics uncertainty included in DESI 2024 full-shape measurements. We extend our investigation to the ELGs at 0.6 < z < 0.8, i.e., beyond the redshift range (0.8 < z < 1.6) adopted for the DESI clustering catalog, and demonstrate that determining the full three-dimensional selection function is necessary in this redshift range. Our tests showed that all changes are consistent with statistical noise for BAO analyses indicating they are robust to even severe imaging systematics. Specific tests for the full-shape analysis will be presented in a companion paper.

79 ASTRONOMY AND ASTROPHYSICS↗