Search arXivSearch

arXiv · 2202.12776

Detecting gravitational lenses using machine learning: exploring interpretability and sensitivity to rare lensing configurations

Abstract

Forthcoming large imaging surveys such as Euclid and the Vera Rubin Observatory Legacy Survey of Space and Time are expected to find more than $10^5$ strong gravitational lens systems, including many rare and exotic populations such as compound lenses, but these $10^5$ systems will be interspersed among much larger catalogues of $\sim10^9$ galaxies. This volume of data is too much for visual inspection by volunteers alone to be feasible and gravitational lenses will only appear in a small fraction of these data which could cause a large amount of false positives. Machine learning is the obvious alternative but the algorithms' internal workings are not obviously interpretable, so their selection functions are opaque and it is not clear whether they would select against important rare populations. We design, build, and train several Convolutional Neural Networks (CNNs) to identify strong gravitational lenses using VIS, Y, J, and H bands of simulated data, with F1 scores between 0.83 and 0.91 on 100,000 test set images. We demonstrate for the first time that such CNNs do not select against compound lenses, obtaining recall scores as high as 76\% for compound arcs and 52\% for double rings. We verify this performance using Hubble Space Telescope (HST) and Hyper Suprime-Cam (HSC) data of all known compound lens systems. Finally, we explore for the first time the interpretability of these CNNs using Deep Dream, Guided Grad-CAM, and by exploring the kernels of the convolutional layers, to illuminate why CNNs succeed in compound lens selection.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Joshua Wilde, Stephen Serjeant, Jane M. Bromley, Hugh Dickinson, Leon V. E. Koopmans, R. Benton Metcalf. 2022-02-25. Detecting gravitational lenses using machine learning: exploring interpretability and sensitivity to rare lensing configurations. https://doi.org/10.1093/mnras%2Fstac562

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Spatially Resolved Physical Properties of Young Star Clusters and Star-forming Clumps in the Brightest z>6 Galaxy, the Strongly Lensed Cosmic Spear at z=6.2

We present spatially resolved analysis of stellar populations in the brightest $z>6$ galaxy known to date (AB mag 23), the strongly lensed MACS0308$-$zD1 (dubbed the ``Cosmic Spear'') at $z_{\rm spec}=6.2$. New JWST NIRCam imaging and high-resolution NIRSpec IFU spectroscopy span the rest-frame ultraviolet to optical. The NIRCam imaging reveals bright star-forming clumps and a tail consisting of three distinct, extremely compact star clusters that are multiply-imaged by gravitational lensing. The star clusters have delensed effective radii of $R_{\rm{eff}} \lesssim 8$ pc, stellar masses of $M_{*} \sim 10^{6}-10^{7}\,M_{\odot}$, and high stellar mass surface densities of $Σ_{*} \gtrsim 2\times 10^{4}\,M_{\odot}~\rm{pc}^{-2}$. While their stellar populations are very young ($\sim 6-11$ Myr), their dynamical ages exceed unity, consistent with the clusters being gravitationally bound systems. Placing the star clusters in the size vs.~stellar mass density plane, we find they occupy a region similar to other high-redshift star clusters within galaxies observed recently with JWST, being significantly more massive and denser than local star clusters. Spatially resolved analysis of the brightest clump reveals a compact, intensely star-forming core. The ionizing photon production efficiency ($ξ_{\rm{ion}}$) is slightly suppressed in this central region, potentially indicating a locally elevated Lyman continuum escape fraction facilitated by feedback-driven channels.

astro-ph.GA

Predicting Supermassive Black Hole-Host Mass Offsets from Broadband Photometry Across Cosmological Simulations with Forecasts for LSST

The possibility of over-massive black holes suggested by James Webb Space Telescope photometric discoveries of 'little red dots', may disfavor light supermassive black hole (SMBH) seeds. However, what should constitute the mass (range) of 'heavy' seeds remains relatively unconstrained. Moreover, Vera C Rubin Observatory's Legacy Survey of Space and Time will photometrically characterize galaxies without direct black hole mass measurements. We forward-model the SIMBA, IllustrisTNG, and EAGLE cosmological simulations into the photometric bands of LSST to train an ensemble machine learning classifier. Our framework achieves $91\%$--$94\%$ accuracy across SIMBA and IllustrisTNG in distinguishing between over-massive and under-massive SMBH growth regimes under LSST magnitude limits, using only broadband photometry. Furthermore, cross-simulation transfer experiments (training on one cosmological simulation and evaluating on another using rank-normalized features) achieve $83\%$--$89\%$ accuracy. This suggests the relative photometric ordering of growth regimes is largely preserved even across fundamentally different sub-grid SMBH feedback prescriptions. Signal decomposition shows our classification is driven by host galaxy colors ($82\%$--$87\%$ accuracy) and, relatedly, the accretion-state's spectral energy distribution shape as opposed to an inversion of our forward model's analytical luminosity prescription. Given that the evaluated simulations employ heavy seed prescriptions ($\geq 10^{4}~M_\odot$), our methodology establishes a validated baseline for classifying post-seeding growth regimes.

astro-ph.GA

Cross Subtype Transferability of Machine Learning Photometric Redshift Relations in Low Redshift Seyfert AGN

Photometric redshift estimation for active galactic nuclei (AGN) is complicated by the combined effects of host-galaxy light, nuclear emission, dust attenuation, and broadband spectral diversity. We investigate whether machine learning photo-z relations trained on one low-redshift Seyfert subtype remain valid when transferred to another, and whether probabilistic subtype classification can be used to identify sources for which a specialised regressor is reliable. Using spectroscopically selected Seyfert I and Seyfert II samples from SDSS, matched to AllWISE photometry over 0 < z_spec <= 0.6, we constructed a common 45-feature representation from SDSS ugriz and WISE W1-W4 data. Random Forest and XGBoost regressors were evaluated within each subtype, followed by controlled cross-subtype transfer tests, redshift and sample size-matched experiments, feature ablations, and an independent classifier-gated regression test. The subtype specific models achieved strong within-sample performance, with the Seyfert II model reaching R2 = 0.965 and sigma_NMAD = 0.0169. However, transfer between Seyfert I and Seyfert II produced a clear and asymmetric degradation in accuracy that persisted after matching the samples and restricting the photometric inputs. A probabilistic Seyfert classifier further identified subsets for which the Seyfert II regressor was more reliable, while extrapolation beyond the redshift range represented in training produced systematic underestimation. These results demonstrate that AGN photo-z performance depends strongly on the population and redshift domain represented in the training data, supporting subtype-aware calibration and applicability-based source selection.

astro-ph.GA