Pith. sign in

REVIEW 1 cited by

Provable concept learning for interpretable predictions using variational autoencoders

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2204.00492 v3 pith:QX4VXVZO submitted 2022-04-01 cs.LG stat.ME

classification cs.LGstat.ME
keywords conceptsclapexplanationsground-truthinterpretableclassifierdatasetspreviously
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In safety-critical applications, practitioners are reluctant to trust neural networks when no interpretable explanations are available. Many attempts to provide such explanations revolve around pixel-based attributions or use previously known concepts. In this paper we aim to provide explanations by provably identifying \emph{high-level, previously unknown ground-truth concepts}. To this end, we propose a probabilistic modeling framework to derive (C)oncept (L)earning and (P)rediction (CLAP) -- a VAE-based classifier that uses visually interpretable concepts as predictors for a simple classifier. Assuming a generative model for the ground-truth concepts, we prove that CLAP is able to identify them while attaining optimal classification accuracy. Our experiments on synthetic datasets verify that CLAP identifies distinct ground-truth concepts on synthetic datasets and yields promising results on the medical Chest X-Ray dataset.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Logit Distance Bounds Representational Similarity

    cs.LG 2026-02 conditional novelty 6.0 of 10

    Logit distance between two softmax models bounds their internal representational dissimilarity, while KL divergence does not, making logit-matching the right objective for preserving linear representational structure ...

Pith tools