Pith. sign in

REVIEW 1 cited by

Towards out-of-distribution generalization in large-scale astronomical surveys: robust networks learn similar representations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.18007 v1 pith:VUJ2SCTZ submitted 2023-11-29 astro-ph.IM astro-ph.GAcs.LG

classification astro-ph.IMastro-ph.GAcs.LG
keywords representationsgeneralizationsimilarityastronomicaldatalayermodelsnetworks
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

The generalization of machine learning (ML) models to out-of-distribution (OOD) examples remains a key challenge in extracting information from upcoming astronomical surveys. Interpretability approaches are a natural way to gain insights into the OOD generalization problem. We use Centered Kernel Alignment (CKA), a similarity measure metric of neural network representations, to examine the relationship between representation similarity and performance of pre-trained Convolutional Neural Networks (CNNs) on the CAMELS Multifield Dataset. We find that when models are robust to a distribution shift, they produce substantially different representations across their layers on OOD data. However, when they fail to generalize, these representations change less from layer to layer on OOD data. We discuss the potential application of similarity representation in guiding model design, training strategy, and mitigating the OOD problem by incorporating CKA as an inductive bias during training.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Reproducibility of machine learning analyses of 21 cm reionization maps

    astro-ph.CO 2024-12 conditional novelty 6.0 of 10

    Convolutional networks trained on 21 cm reionization images often memorize simulation boxes rather than physics, yielding high same-box test scores but poor performance on unseen simulations.

Pith tools