Engineering Papers⌕ Search

SEARCH · Engineering Papers

Results for “Astronomy data analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Methods for Incorporating Model Uncertainty into Exoplanet Atmospheric Analysis

A key goal of exoplanet spectroscopy is to measure atmospheric properties, such as abundances of chemical species, in order to connect them to our understanding of atmospheric physics and planet formation. In this new era of high-quality JWST data, it is paramount that these measurement methods are robust. When comparing atmospheric models to observations, multiple candidate models may produce reasonable fits to the data. Typically, conclusions are reached by selecting the best-performing model according to some metric. This ignores model uncertainty in favor of specific model assumptions, potentially leading to measured atmospheric properties that are overconfident and/or incorrect. In this paper, we compare three ensemble methods for addressing model uncertainty by combining posterior distributions from multiple analyses: Bayesian model averaging, a variant of Bayesian model averaging using leave-one-out predictive densities, and stacking of predictive distributions. We demonstrate these methods by fitting the Hubble Space Telescope (HST) + Spitzer transmission spectrum of the hot Jupiter HD 209458b using models with different cloud and haze prescriptions. All of our ensemble methods lead to uncertainties on retrieved parameters that are larger but more realistic and consistent with physical and chemical expectations. Since they have not typically accounted for model uncertainty, uncertainties of retrieved parameters from HST spectra have likely been underreported. We recommend stacking as the most robust model combination method. Our methods can be used to combine results from independent retrieval codes and from different models within one code. They are also widely applicable to other exoplanet analysis processes, such as combining results from different data reductions.

79 ASTRONOMY AND ASTROPHYSICS↗

The Early Data Release of the Dark Energy Spectroscopic Instrument

The Dark Energy Spectroscopic Instrument (DESI) completed its 5 month Survey Validation in 2021 May. Spectra of stellar and extragalactic targets from Survey Validation constitute the first major data sample from the DESI survey. This paper describes the public release of those spectra, the catalogs of derived properties, and the intermediate data products. In total, the public release includes good-quality spectral information from 466,447 objects targeted as part of the Milky Way Survey, 428,758 as part of the Bright Galaxy Survey, 227,318 as part of the Luminous Red Galaxy sample, 437,664 as part of the Emission Line Galaxy sample, and 76,079 as part of the Quasar sample. In addition, the release includes spectral information from 137,148 objects that expand the scope beyond the primary samples as part of a series of secondary programs. Here, we describe the spectral data, data quality, data products, Large-Scale Structure science catalogs, access to the data, and references that provide relevant background to using these spectra.

79 ASTRONOMY AND ASTROPHYSICS↗

Considerations for Optimizing the Photometric Classification of Supernovae from the Rubin Observatory

The Vera C. Rubin Observatory will increase the number of observed supernovae (SNe) by an order of magnitude; however, it is impossible to spectroscopically confirm the class for all SNe discovered. Thus, photometric classification is crucial, but its accuracy depends on the not-yet-finalized observing strategy of Rubin Observatory's Legacy Survey of Space and Time (LSST). We quantitatively analyze the impact of the LSST observing strategy on SNe classification using simulated multiband light curves from the Photometric LSST Astronomical Time-Series Classification Challenge (PLAsTiCC). First, we augment the simulated training set to be representative of the photometric redshift distribution per SNe class, the cadence of observations, and the flux uncertainty distribution of the test set. Then we build a classifier using the photometric transient classification library snmachine, based on wavelet features obtained from Gaussian process fits, yielding a similar performance to the winning PLAsTiCC entry. We study the classification performance for SNe with different properties within a single simulated observing strategy. We find that season length is important, with light curves of 150 days yielding the highest performance. Cadence also has an important impact on SNe classification; events with median inter-night gap <3.5 days yield higher classification performance. Interestingly, we find that large gaps (>10 days) in light-curve observations do not impact performance if sufficient observations are available on either side, due to the effectiveness of the Gaussian process interpolation. This analysis is the first exploration of the impact of observing strategy on photometric SN classification with LSST.

79 ASTRONOMY AND ASTROPHYSICS↗

Archetype-based Redshift Estimation for the Dark Energy Spectroscopic Instrument Survey

We present a computationally efficient galaxy archetype-based redshift estimation and spectral classification method for the Dark Energy Survey Instrument (DESI) survey. The DESI survey currently relies on a redshift fitter and spectral classifier using a linear combination of principal component analysis–derived templates, which is very efficient in processing large volumes of DESI spectra within a short time frame. However, this method occasionally yields unphysical model fits for galaxies and fails to adequately absorb calibration errors that may still be occasionally visible in the reduced spectra. Our proposed approach improves upon this existing method by refitting the spectra with carefully generated physical galaxy archetypes combined with additional terms designed to absorb data reduction defects and provide more physical models to the DESI spectra. We test our method on an extensive data set derived from the survey validation (SV) and Year 1 (Y1) data of DESI. Our findings indicate that the new method delivers marginally better redshift success for SV tiles while reducing catastrophic redshift failure by 10%–30%. At the same time, results from millions of targets from the main survey show that our model has relatively higher redshift success and purity rates (0.5%–0.8% higher) for galaxy targets while having similar success for QSOs. These improvements also demonstrate that the main DESI redshift pipeline is generally robust. Additionally, it reduces the false-positive redshift estimation by 5%–40% for sky fibers. We also discuss the generic nature of our method and how it can be extended to other large spectroscopic surveys, along with possible future improvements.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

The Astropy Project: Sustaining and Growing a Community-oriented Open-source Project and the Latest Major Release (v5.0) of the Core Package

The Astropy Project supports and fosters the development of open-source and openly developed Python packages that provide commonly needed functionality to the astronomical community. A key element of the Astropy Project is the core package astropy, which serves as the foundation for more specialized projects and packages. In this article, we summarize key features in the core package as of the recent major release, version 5.0, and provide major updates on the Project. We then discuss supporting a broader ecosystem of interoperable packages, including connections with several astronomical observatories and missions. We also revisit the future outlook of the Astropy Project and the current status of Learn Astropy. We conclude by raising and discussing the current and future challenges facing the Project.

79 ASTRONOMY AND ASTROPHYSICS↗

Impact of Rubin Observatory Cadence Choices on Supernovae Photometric Classification

The Vera C. Rubin Observatory's Legacy Survey of Space and Time (LSST) will discover an unprecedented number of supernovae (SNe), making spectroscopic classification for all the events infeasible. LSST will thus rely on photometric classification, whose accuracy depends on the not-yet-finalized LSST observing strategy. In this work, we analyze the impact of cadence choices on classification performance using simulated multiband light curves. First, we simulate SNe with an LSST baseline cadence, a nonrolling cadence, and a presto-color cadence, which observes each sky location three times per night instead of twice. Each simulated data set includes a spectroscopically confirmed training set, which we augment to be representative of the test set as part of the classification pipeline. Then we use the photometric transient classification library snmachine to build classifiers. We find that the active region of the rolling cadence used in the baseline observing strategy yields a 25% improvement in classification performance relative to the background region. This improvement in performance in the actively rolling region is also associated with an increase of up to a factor of 2.7 in the number of cosmologically useful Type Ia SNe relative to the background region. However, adding a third visit per night as implemented in presto-color degrades classification performance due to more irregularly sampled light curves. Overall, our results establish desiderata on the observing cadence related to classification of full SNe light curves, which in turn impacts photometric SNe cosmology with LSST.

79 ASTRONOMY AND ASTROPHYSICS↗

The Astronomy Commons Platform: A Deployable Cloud-based Analysis Platform for Astronomy

Abstract We present a scalable, cloud-based science platform solution designed to enable next-to-the-data analyses of terabyte-scale astronomical tabular data sets. The presented platform is built on Amazon Web Services (over Kubernetes and S3 abstraction layers), utilizes Apache Spark and the Astronomy eXtensions for Spark for parallel data analysis and manipulation, and provides the familiar JupyterHub web-accessible front end for user access. We outline the architecture of the analysis platform, provide implementation details and rationale for (and against) technology choices, verify scalability through strong and weak scaling tests, and demonstrate usability through an example science analysis of data from the Zwicky Transient Facility’s 1Bn+ light-curve catalog. Furthermore, we show how this system enables an end user to iteratively build analyses (in Python) that transparently scale processing with no need for end-user interaction. The system is designed to be deployable by astronomers with moderate cloud engineering knowledge, or (ideally) IT groups. Over the past 3 yr, it has been utilized to build science platforms for the DiRAC Institute, the ZTF partnership, the LSST Solar System Science Collaboration, and the LSST Interdisciplinary Network for Collaboration and Computing, as well as for numerous short-term events (with over 100 simultaneous users). In a live demo instance, the deployment scripts, source code, and cost calculators are accessible. 4 4 http://hub.astronomycommons.org/

79 ASTRONOMY AND ASTROPHYSICS↗

SDSS-IV MaStar: Data-driven Parameter Derivation for the MaStar Stellar Library

The Mapping Nearby Galaxies at Apache Point Observatory (MaNGA) Stellar Library (MaStar) is a large collection of high-quality empirical stellar spectra designed to cover all spectral types and ideal for use in the stellar population analysis of galaxies observed in the MaNGA survey. The library contains 59,266 spectra of 24,130 unique stars with spectral resolution R ~ 1800 and covering a wavelength range of 3622–10,354 Å. In this work, we derive five physical parameters for each spectrum in the library: effective temperature (T eff ), surface gravity ($\mathrm{log} g$), metallicity ([Fe/H]), microturbulent velocity ($\mathrm{log}({v}_{\mathrm{micro}})$), and alpha-element abundance ([α/Fe]). These parameters are derived with a flexible data-driven algorithm that uses a neural network model. We train a neural network using the subset of 1675 MaStar targets that have also been observed in the Apache Point Observatory Galactic Evolution Experiment (APOGEE), adopting the independently-derived APOGEE Stellar Parameter and Chemical Abundance Pipeline parameters for this reference set. For the regions of parameter space not well represented by the APOGEE training set (7000 ≤ T ≤ 30,000 K), we supplement with theoretical model spectra. We present our derived parameters along with an analysis of the uncertainties and comparisons to other analyses from the literature.

79 ASTRONOMY AND ASTROPHYSICS↗

FAST Drift Scan Survey for Ηι Intensity Mapping: I. Preliminary Data Analysis

This work presents the initial results of the drift-scan observation for the neutral hydrogen (Hi) intensity mapping survey with the Five-hundred-meter Aperture Spherical radio Telescope (FAST). The data analyzed in this work were collected in night observations from 2019 through 2021. The primary findings are based on 28 hr of drift-scan observation carried out over 7 nights in 2021, which covers 60 deg 2 sky area. Our main findings are, first, our calibration strategy can successfully correct both the temporal and bandpass gain variation over the 4 hr drift-scan observation. Second, the continuum maps of the surveyed region are made with frequency resolution of 28 kHz and pixel area of $_{2.95\,{\mathrm{arcmin}}^{2}}$. The pixel noise levels of the continuum maps are slightly higher than the forecast assuming $T$ sys = 20 K, which are 36.0 mK (for 10.0 s integration time) at the 1050–1150 MHz band, and 25.9 mK (for 16.7 s integration time) at the 1323–1450 MHz band, respectively. Third, the flux-weighted differential number count is consistent with the NRAO-VLA Sky Survey (NVSS) catalog down to the confusion limit ~7 mJy beam –1 . Finally, the continuum flux measurements of the sources are consistent with those found in the literature. The difference in the flux measurement of 81 isolated NVSS sources is about 6.3%. Our research offers a systematic analysis for the FAST Hi intensity mapping drift-scan survey and serves as a helpful resource for further cosmology and associated galaxies sciences with the FAST drift-scan survey.

79 ASTRONOMY AND ASTROPHYSICS↗

BatAnalysis - A Comprehensive Python Pipeline for Swift BAT Survey Analysis

The Swift Burst Alert Telescope (BAT) is a coded-aperture gamma-ray instrument with a large field of view that primarily operates in survey mode when it is not triggering on transient events. The survey data consist of 80-channel detector plane histograms that accumulate photon counts over periods of at least 5 minutes. These histograms are processed on the ground and are used to produce the survey data set between 14 and 195 keV. Survey data comprise >90% of all BAT data by volume and allow for the tracking of long-term light curves and spectral properties of cataloged and uncataloged hard X-ray sources. Until now, the survey data set has not been used to its full potential due to the complexity associated with its analysis and the lack of easily usable pipelines. Here, we introduce the BatAnalysis Python package, a wrapper for HEASoftpy, which provides a modern, open-source pipeline to process and analyze BAT survey data. BatAnalysis allows members of the community to use BAT survey data in more advanced analyses of astrophysical sources, including pulsars, pulsar wind nebula, active galactic nuclei, and other known/unknown transient events that may be detected in the hard X-ray band. We outline the steps taken by the Python code and exemplify its usefulness and accuracy by analyzing survey data of the Crab Nebula, NGC 2992, and a previously uncataloged MAXI transient. The BatAnalysis package allows for ~18 yr of BAT survey data to be used in a systematic way to study a large variety of astrophysical sources.

79 ASTRONOMY AND ASTROPHYSICS↗

SpecDis: Value Added Distance Catalog for 4 Million Stars from DESI Year-1 Data

We present the SpecDis value-added stellar distance catalog accompanying DESI Data Release 1. SpecDis trains a feed-forward neural network (NN) with Gaia parallaxes and gets the distance estimates. To build up an unbiased training sample, we do not apply selections on parallax error or signal-to-noise (S/N) of the stellar spectra, and instead, we incorporate parallax error into the loss function. Moreover, we employ principal component analysis to reduce the noise and dimensionality of stellar spectra. Validated by independent external samples of member stars with precise distances from globular clusters, dwarf galaxies, stellar streams, combined with blue horizontal branch stars, we demonstrate that our distance measurements show no significant bias up to 100 kpc, and are much more precise than Gaia parallax beyond 7 kpc. The median distance uncertainties are 23%, 19%, 11%, and 7% for S/N < 20, 20 ≤ S/N < 60, 60 ≤ S/N < 100, and S/N ≥ 100. Selecting stars with ${\mathrm{log}}\,g\lt 3.8$ and distance uncertainties smaller than 25%, we have more than 74,000 giant candidates within 50 kpc of the Galactic center and 1500 candidates beyond this distance. Additionally, we develop a Gaussian mixture model to identify unresolvable equal-mass binaries by modeling the discrepancy between the NN-predicted and the geometric absolute magnitudes from Gaia parallaxes and identify 120,000 equal-mass binary candidates. Our final catalog provides distances and distance uncertainties for >4 million stars, offering a valuable resource for Galactic astronomy.

astronomy data analysis↗

Deep Learning of Dark Energy Spectroscopic Instrument Mock Spectra to Find Damped Lyα Systems

We have updated and applied a convolutional neural network (CNN) machine-learning model to discover and characterize damped Ly α systems (DLAs) based on Dark Energy Spectroscopic Instrument (DESI) mock spectra. We have optimized the training process and constructed a CNN model that yields a DLA classification accuracy above 99% for spectra that have signal-to-noise ratios (S/N) above 5 per pixel. The classification accuracy is the rate of correct classifications. This accuracy remains above 97% for lower S/N ≈1 spectra. This CNN model provides estimations for redshift and H i column density with standard deviations of 0.002 and 0.17 dex for spectra with S/N above 3 pixel -1 . Also, this DLA finder is able to identify overlapping DLAs and sub-DLAs. Further, the impact of different DLA catalogs on the measurement of baryon acoustic oscillations (BAO) is investigated. The cosmological fitting parameter result for BAO has less than 0.61% difference compared to analysis of the mock results with perfect knowledge of DLAs. This difference is lower than the statistical error for the first year estimated from the mock spectra: above 1.7%. We also compared the performances of the CNN and Gaussian Process (GP) models. Our improved CNN model has moderately 14% higher purity and 7% higher completeness than an older version of the GP code, for S/N > 3. Both codes provide good DLA redshift estimates, but the GP produces a better column density estimate by 24% less standard deviation. A credible DLA catalog for the DESI main survey can be provided by combining these two algorithms.

79 ASTRONOMY AND ASTROPHYSICS↗

A Simulation-based Method for Correcting Mode Coupling in CMB Angular Power Spectra

Modern cosmic microwave background (CMB) analysis pipelines regularly employ complex time-domain filters, beam models, masking, and other techniques during the production of sky maps and their corresponding angular power spectra. However, these processes can generate couplings between multipoles from the same spectrum and from different spectra, in addition to the typical power attenuation. Within the context of pseudo-C ℓ based, MASTER-style analyses, the net effect of the time-domain filtering is commonly approximated by a multiplicative transfer function, F ℓ , that can fail to capture mode mixing and is dependent on the spectrum of the signal. To address these shortcomings, we have developed a simulation-based spectral correction approach that constructs a two-dimensional transfer matrix, ${J}_{{\ell }{\ell }^{\prime} }$, which contains information about mode mixing in addition to mode attenuation. We demonstrate the application of this approach on data from the first flight of the Spider balloon-borne CMB experiment.

79 ASTRONOMY AND ASTROPHYSICS↗

Scaler Rates from the Pierre Auger Observatory: A New Proxy of Solar Activity

The modulation of low-energy galactic cosmic rays reflects interplanetary magnetic field variations and can provide useful information on solar activity. An array of ground-surface detectors can reveal the secondary particles, which originate from the interaction of cosmic rays with the atmosphere. In this work, we present an investigation of the low-threshold rate (scaler) time series recorded in 16 yr of operation by the Pierre Auger Observatory surface detectors in Malargüe, Argentina. Through an advanced spectral analysis, we detected highly statistically significant variations in the time series with periods ranging from the decadal to the daily scale. We investigate their origin, revealing a direct connection with solar variability. Thanks to their intrinsic very low noise level, the Auger scalers allow a thorough and detailed investigation of the galactic cosmic-ray flux variations in the heliosphere at different timescales and can, therefore, be considered a new proxy of solar variability.

79 ASTRONOMY AND ASTROPHYSICS↗

Mitigation of the Brighter-fatter Effect in the LSST Camera

Abstract Thick, fully depleted charge-coupled devices are known to exhibit nonlinear behavior at high signal levels due to the dynamic behavior of charges collecting in the potential wells of pixels, called the brighter-fatter effect (BFE). The effect results in distorted images of bright calibration stars, creating a flux-dependent point-spread function that if left unmitigated, could make up a large fraction of the error budget in Stage IV weak-lensing (WL) surveys such as the Legacy Survey of Space and Time (LSST). In this paper, we analyze image measurements of flat fields and artificial stars taken at different illumination levels with the LSST Camera (LSSTCam) at SLAC National Accelerator Laboratory in order to quantify this effect in the LSSTCam before and after a previously introduced correction technique. We observe that the BFE evolves anisotropically as a function of flux due to higher-order BFEs, which violates the fundamental assumption of this correction method. We then introduce a new method based on a physically motivated model to account for these higher-order terms in the correction, and then we test the modified correction on both data sets. We find that the new method corrects the effect in flat fields better than it corrects the effect in artificial stars, which we suggest is the result of sub-pixel physics not included in this correction model. We use these results to define a new metric for the full-well capacity of our sensors and advise image processing strategies to further limit the impact of the effect on LSST WL science pathways.

47 OTHER INSTRUMENTATION↗

Clustering of DESI galaxies split by thermal Sunyaev-Zeldovich effect

The thermal Sunyaev-Zeldovich (tSZ) effect is associated with galaxy clusters - extremely large and dense structures tracing the dark matter with a higher bias than isolated galaxies. We propose to use the tSZ data to separate galaxies from redshift surveys into distinct subpopulations corresponding to different densities and biases independently of the redshift survey systematics. Leveraging the information from different environments, as in density-split and density-marked clustering, is known to tighten the constraints on cosmological parameters, like $\Omega_m$, $\sigma_8$ and neutrino mass. We use data from the Dark Energy Spectroscopic Instrument (DESI) and the Atacama Cosmology Telescope (ACT) in their region of overlap to demonstrate informative tSZ splitting of Luminous Red Galaxies (LRGs). We discover a significant increase in the large-scale clustering of DESI LRGs corresponding to detections starting from 1-2 sigma in the ACT DR6 + Planck tSZ Compton-$y$ map, below the cluster candidate threshold (4 sigma). We also find that such galaxies have higher line-of-sight coordinate (and velocity) dispersions and a higher number of close neighbors than both the full sample and near-zero tSZ regions. We produce simple simulations of tSZ maps that are intrinsically consistent with galaxy catalogs and do not include systematic effects, and find a similar pattern of large-scale clustering enhancement with tSZ effect significance. Moreover, we observe that this relative bias pattern remains largely unchanged with variations in the galaxy-halo connection model in our simulations. This is promising for future cosmological inference from tSZ-split clustering with semi-analytical models. Thus, we demonstrate that valuable cosmological information is present in the lower signal-to-noise regions of the thermal Sunyaev-Zeldovich map, extending far beyond the individual cluster candidates.

Astronomy data analysis↗

Detecting Long-period Variability in the SDSS Stripe 82 Standards Catalog

We report the results of a search for long-period (100 < P < 600 days) periodic variability in the SDSS Stripe 82 standards catalog. The SDSS coverage of Stripe 82 enables such a search because there are on average 20 observations per band in ugriz bands for about one million sources, collected over about 6 yr, with a faint limit of r ~ 22 mag and precisely calibrated 1%–2% photometry. We calculated the periods of variable source candidates in this sample using the Lomb–Scargle periodogram and considered the three highest periodogram peaks in each of the gri filters as relevant. Only those sources with gri periods consistent within 0.1% were later studied. We use the Kuiper statistic to ensure uniform distribution of data points in phased light curves. We present five sources with the spectra consistent with quasar spectra and plausible periodic variability. This SDSS-based search bodes well for future sensitive large-area surveys, such as the Rubin Observatory Legacy Survey of Space and Time, which, due to its larger sky coverage (about a factor of 60) and improved sensitivity (~2 mag), will be more powerful for finding such sources.

79 ASTRONOMY AND ASTROPHYSICS↗

Detecting and Characterizing Mg ii Absorption in DESI Survey Validation Quasar Spectra

Abstract We present findings of the detection of Magnesium II (Mg ii , λ = 2796, 2803 Å) absorbers from the early data release of the Dark Energy Spectroscopic Instrument (DESI). DESI is projected to obtain spectroscopy of approximately 3 million quasars (QSOs), of which over 99% are anticipated to be at redshifts greater than z > 0.3, such that DESI would be able to observe an associated or intervening Mg ii absorber illuminated by the background QSO. We have developed an autonomous supplementary spectral pipeline that detects these systems through an initial line-fitting process and then confirms the line properties using a Markov Chain Monte Carlo sampler. Based upon a visual inspection of the resulting systems, we estimate that this sample has a purity greater than 99%. We have also investigated the completeness of our sample in regard to both the signal-to-noise properties of the input spectra and the rest-frame equivalent width ( W 0 ) of the absorber systems. From a parent catalog containing 83,207 quasars, we detect a total of 23,921 Mg ii absorption systems following a series of quality cuts. Extrapolating from this occurrence rate of 28.8% implies a catalog at the completion of the five-year DESI survey that will contain over eight hundred thousand Mg ii absorbers. The cataloging of these systems will enable significant further research because they carry information regarding circumgalactic medium environments, the distribution of intervening galaxies, and the growth of metallicity across the redshift range 0.3 ≤ z < 2.5.

79 ASTRONOMY AND ASTROPHYSICS↗