arXiv · 2402.15551
Improving Photometric Redshift Estimates with Training Sample Augmentation
Abstract
Large imaging surveys will rely on photometric redshifts (photo-z's), which are typically estimated through machine learning methods. Currently planned spectroscopic surveys will not be deep enough to produce a representative training sample for LSST, so we seek methods to improve the photo-z estimates that arise from non-representative training samples. Spectroscopic training samples for photo-z's are biased towards redder, brighter galaxies, which also tend to be at lower redshift than the typical galaxy observed by LSST, leading to poor photo-z estimates with outlier fractions nearly 4 times larger than for a representative training sample. In this paper, we apply the concept of training sample augmentation, where we augment simulated non-representative training samples with simulated galaxies possessing otherwise unrepresented features. When we select simulated galaxies with (g-z) color, i-band magnitude and redshift outside the range of the original training sample, we are able to reduce the outlier fraction of the photo-z estimates for simulated LSST data by nearly 50% and the normalized median absolute deviation (NMAD) by 56%. When compared to a fully representative training sample, augmentation can recover nearly 70% of the degradation in the outlier fraction and 80% of the degradation in NMAD. Training sample augmentation is a simple and effective way to improve training samples for photo-z's without requiring additional spectroscopic samples.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Irene Moskowitz, Eric Gawiser, John Franklin Crenshaw, Brett H. Andrews, Alex I. Malz, Samuel Schmidt, The LSST Dark Energy Science Collaboration. 2024-02-23. Improving Photometric Redshift Estimates with Training Sample Augmentation. https://doi.org/10.3847/2041-8213%2Fad4039
Cite the original work for its findings. Save a collection to share your selection of sources.