Search arXivSearch

arXiv · 1910.05255

Quasar and galaxy classification in Gaia Data Release 2

Abstract

We construct a supervised classifier based on Gaussian Mixture Models to probabilistically classify objects in Gaia data release 2 (GDR2) using only photometric and astrometric data in that release. The model is trained empirically to classify objects into three classes -- star, quasar, galaxy -- for G<=14.5 mag down to the Gaia magnitude limit of G=21.0 mag. Galaxies and quasars are identified for the training set by a cross-match to objects with spectroscopic classifications from the Sloan Digital Sky Survey. Stars are defined directly from GDR2. When allowing for the expectation that quasars are 500 times rarer than stars, and galaxies 7500 times rarer than stars (the class imbalance problem), samples classified with a threshold probability of 0.5 are predicted to have purities of 0.43 for quasars and 0.28 for galaxies, and completenesses of 0.58 and 0.72 respectively. The purities can be increased up to 0.60 by adopting a higher threshold. Not accounting for this expected low frequency of extragalactic objects (the class prior) would give both erroneously optimistic performance predictions and severely impure samples. Applying our model to all 1.20 billion objects in GDR2 with the required features, we classify 2.3 million objects as quasars and 0.37 million objects as galaxies (with individual probabilities above 0.5). The small number of galaxies is due to the strong bias of the satellite detection algorithm and on-ground data selection against extended objects. We infer the true number of quasars and galaxies -- as these classes are defined by our training set -- to be 690,000 and 110,000 respectively (+/- 50%). The aim of this work is to see how well extragalactic objects can be classified using only GDR2 data. Better classifications should be possible with the low resolution spectroscopy (BP/RP) planned for GDR3.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Coryn A. L. Bailer-Jones, Morgan Fouesneau, Rene Andrae. 2019-10-11. Quasar and galaxy classification in Gaia Data Release 2. https://doi.org/10.1093/mnras%2Fstz2947

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The Entangling of Supernova Feedback Impacts with Coarsening Simulation Resolution

It is often understood that supernova (SN) feedback in galaxies is responsible for regulating star formation (SF) and generating gaseous outflows. However, a detailed look at the small-scale effects of SNe on the interstellar medium (ISM) in simulations shows that the macroscopic processes of SF suppression and outflow generation proceed in distinct channels. We demonstrate this finding in two independent simulations of isolated dwarf galaxies with very high (m_gas ~ Msun) numerical resolution, LYRA and RIGEL. Our findings suggest that the macroscopic effect of a given SN on the galaxy is best predicted by its local density. Outflows are driven by SNe in diffuse regions expanding to their cooling radii on large (~kpc) scales, while dense SF regions are disrupted in a localized (~pc) manner. However, these separate feedback channels are only distinguishable at very high resolutions capable of following mass scales \lesssim 10^2 \msun. When averaging on coarser scales, ISM densities are greatly mis-estimated, and variations between different SF and SNe-affected regions are severely washed out. It therefore cannot be __self-consistently__ determined, from coarse-resolution information __alone__, (1) whether a SN tends to contribute to outflows or direct SF suppression, and (2) the rate of SF in a given region. In particular, commonly used parameters in coarse-resolution (subgrid) models, such as the SN cooling radius and SF density threshold, may require more detailed treatments informed by high-resolution studies.

astro-ph.GA

Computational advances and challenges in simulations of turbulence and star formation

We review recent advances in the numerical modeling of turbulent flows and star formation. An overview of the most widely used simulation codes and their core capabilities is provided. We then examine methods for achieving the highest-resolution magnetohydrodynamical turbulence simulations to date, highlighting challenges related to numerical viscosity and resistivity. State-of-the-art approaches to modeling gravity and star formation are discussed in detail, including implementations of star particles and feedback from jets, winds, heating, ionization, and supernovae. We review the latest techniques for radiation hydrodynamics, including ray tracing, Monte Carlo, and moment methods, with comparisons between the flux-limited diffusion, moment-1, and variable Eddington tensor methods. The final chapter summarizes advances in cosmic-ray transport schemes, emphasizing their growing importance for connecting small-scale star formation physics with galaxy-scale evolution.

astro-ph.GA

How significant is the lensing interpretation of GW231123?

GW231123 is one of the most unusual gravitational-wave (GW) events, with exceptionally large inferred masses and near-extremal spins, offering an opportunity to test whether propagation effects contribute to these properties. We therefore examine whether the data support wave-optics microlensing embedded in a strong-lensing galaxy, whose detection becomes increasingly likely as observations accumulate, whether this interpretation can explain these properties, and how significant the preference remains under detector noise and waveform systematics. We compare six hypotheses: unlensed, isolated point mass, and embedded point-mass (EPM) and binary-lens (EB) effective models in Type-I (minimum) and Type-II (saddle) macro images. The EB Type-I model is most favored. For the most accurate waveform model NRSur7dq4, it gives $\log_{10}B^{\rm EB-I}_{\rm U}=2.60$, versus $0.89$ for Type II, indicating sensitivity to macro-image geometry. Within Type I, however, the binary improves over the point mass by only $\log_{10}B^{\rm EB-I}_{\rm EPM-I}=0.16$ and $Δ\ln\mathcal{L}_{\max}=0.56$, providing no clear evidence for structure beyond a single effective perturber. Moreover, under embedded lensing, waveform-template discrepancies and inferred masses and spins are reduced. However, real O4a backgrounds from numerical-relativity injections show that the apparent lensing evidence is sensitive to waveform systematics and realistic detector noise: although the commonly used waveform IMRPhenomXPHM gives the largest Bayes factor, $\log_{10}B^{\rm EB-I}_{\rm U}=4.52$, it is less exceptional relative to its own background, with a false-alarm probability of $6.5$--$8\%$, whereas NRSur7dq4 gives only $2$--$3\%$. Thus, waveform systematics can amplify apparent lensing evidence, but GW231123 remains an intriguing lensing candidate.

astro-ph.GA