Pith. sign in

REVIEW 1 cited by

Beyond Accuracy: Ensuring Correct Predictions With Correct Rationales

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2411.00132 v2 pith:Q3FKWDSU submitted 2024-10-31 cs.LG cs.AIcs.CV

classification cs.LGcs.AIcs.CV
keywords correctmodelsrationalesaccuracymodelpredictionpredictionsdataset
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large pretrained foundation models demonstrate exceptional performance and, in some high-stakes applications, even surpass human experts. However, most of these models are currently evaluated primarily on prediction accuracy, overlooking the validity of the rationales behind their accurate predictions. For the safe deployment of foundation models, there is a pressing need to ensure double-correct predictions, i.e., correct prediction backed by correct rationales. To achieve this, we propose a two-phase scheme: First, we curate a new dataset that offers structured rationales for visual recognition tasks. Second, we propose a rationale-informed optimization method to guide the model in disentangling and localizing visual evidence for each rationale, without requiring manual annotations. Extensive experiments and ablation studies demonstrate that our model outperforms state-of-the-art models by up to 10.1% in prediction accuracy across a wide range of tasks. Furthermore, our method significantly improves the model's rationale correctness, improving localization by 7.5% and disentanglement by 36.5%. Our dataset, source code, and pretrained weights: https://github.com/deep-real/DCP

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Interpretable Failure Detection with Human-Level Concepts

    cs.CV 2025-02 conditional novelty 6.0 of 10

    ORCA ranks concept activations from CLIP and uses the rank-weighted agreement of the top-K concepts with the predicted category as its confidence score, improving failure-detection FPR on several benchmarks.

Pith tools