Data-driven, interpretable photometric redshifts trained on heterogeneous and unrepresentative data
arXiv:1612.00847 · doi:10.3847/1538-4357/aa6332
Abstract
We present a new method for inferring photometric redshifts in deep galaxy and quasar surveys, based on a data driven model of latent spectral energy distributions (SEDs) and a physical model of photometric fluxes as a function of redshift. This conceptually novel approach combines the advantages of both machine-learning and template-fitting methods by building template SEDs directly from the training data. This is made computationally tractable with Gaussian Processes operating in flux--redshift space, encoding the physics of redshift and the projection of galaxy SEDs onto photometric band passes. This method alleviates the need of acquiring representative training data or constructing detailed galaxy SED models; it requires only that the photometric band passes and calibrations be known or have parameterized unknowns. The training data can consist of a combination of spectroscopic and deep many-band photometric data, which do not need to entirely spatially overlap with the target survey of interest or even involve the same photometric bands. We showcase the method on the -magnitude-selected, spectroscopically-confirmed galaxies in the COSMOS field. The model is trained on the deepest bands (from SUBARU and HST) and photometric redshifts are derived using the shallower SDSS optical bands only. We demonstrate that we obtain accurate redshift point estimates and probability distributions despite the training and target sets having very different redshift distributions, noise properties, and even photometric bands. Our model can also be used to predict missing photometric fluxes, or to simulate populations of galaxies with realistic fluxes and redshifts, for example. This method opens a new era in which photometric redshifts for large photometric surveys are derived using a flexible yet physical model of the data trained on all available surveys (spectroscopic and photometric).
16 pages, 8 figures, to be submitted to ApJ
References in corpus (6)
- EAZY: A Fast, Public Photometric Redshift Code
- K-corrections and filter transformations in the ultraviolet, optical, and near infrared
- The Zurich Extragalactic Bayesian Redshift Analyzer (ZEBRA) and its first application: COSMOS
- Photometric redshift analysis in the Dark Energy Survey Science Verification data
- Reconstructing Redshift Distributions with Cross-Correlations: Tests and an Optimized Recipe
- Galaxy And Mass Assembly (GAMA): Curation and reanalysis of 16.6k redshifts in the G10/COSMOS region
Cited by in corpus (8)
- Non-parametric Star Formation History Reconstruction with Gaussian Processes I: Counting Major Episodes of Star Formation
- Surveying the reach and maturity of machine learning and artificial intelligence in astronomy
- Galaxy clustering in the DESI Legacy Survey and its imprint on the CMB
- Photometric Redshifts with the LSST: Evaluating Survey Observing Strategies
- Photometric redshifts with machine learning, lights and shadows on a complex data science use case
- The PAU Survey: narrowband photometric redshifts using Gaussian processes
- PhotoRedshift-MML: a multimodal machine learning method for estimating photometric redshifts of quasars
- Estimating Spectra from Photometry