Pith. sign in

REVIEW 3 cited by

Data-driven, interpretable photometric redshifts trained on heterogeneous and unrepresentative data

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1612.00847 v1 pith:WHP3HTSZ submitted 2016-12-02 astro-ph.CO

classification astro-ph.CO
keywords photometricdatamodelredshiftsbandsmethodredshifttraining
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We present a new method for inferring photometric redshifts in deep galaxy and quasar surveys, based on a data driven model of latent spectral energy distributions (SEDs) and a physical model of photometric fluxes as a function of redshift. This conceptually novel approach combines the advantages of both machine-learning and template-fitting methods by building template SEDs directly from the training data. This is made computationally tractable with Gaussian Processes operating in flux--redshift space, encoding the physics of redshift and the projection of galaxy SEDs onto photometric band passes. This method alleviates the need of acquiring representative training data or constructing detailed galaxy SED models; it requires only that the photometric band passes and calibrations be known or have parameterized unknowns. The training data can consist of a combination of spectroscopic and deep many-band photometric data, which do not need to entirely spatially overlap with the target survey of interest or even involve the same photometric bands. We showcase the method on the $i$-magnitude-selected, spectroscopically-confirmed galaxies in the COSMOS field. The model is trained on the deepest bands (from SUBARU and HST) and photometric redshifts are derived using the shallower SDSS optical bands only. We demonstrate that we obtain accurate redshift point estimates and probability distributions despite the training and target sets having very different redshift distributions, noise properties, and even photometric bands. Our model can also be used to predict missing photometric fluxes, or to simulate populations of galaxies with realistic fluxes and redshifts, for example. This method opens a new era in which photometric redshifts for large photometric surveys are derived using a flexible yet physical model of the data trained on all available surveys (spectroscopic and photometric).

Discussion (0). Sign in to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. pop-cosmos: Disentangling galaxy properties from observables using data-driven approaches

    astro-ph.GA 2026-06 unverdicted novelty 6.0 of 10

    A beta-VAE analysis of pop-cosmos models finds that five latent dimensions capture the rest-frame optical SED, corresponding to stellar mass, recent star formation, dust, and two gas ionization states.

  2. pop-cosmos: Disentangling galaxy properties from observables using data-driven approaches

    astro-ph.GA 2026-06 conditional novelty 6.0 of 10

    Rest-frame optical galaxy SEDs from a 16-parameter SPS model are captured by five disentangled VAE latents (mass, young stars, dust, soft/hard ionization); metallicity and age are not independent drivers.

  3. pop-cosmos: Galaxy size evolution across structural and star-formation classifications in COSMOS-Web

    astro-ph.GA 2026-06 unverdicted novelty 5.0 of 10

    Galaxy size-mass relations exhibit double power-law breaks at different pivot masses for quiescent versus bulge-dominated samples, coinciding with AGN activity scales.

Pith tools