Pith. sign in

REVIEW 3 cited by

Correctness is not Faithfulness in RAG Attributions

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.18004 v1 pith:HUXG33WM submitted 2024-12-23 cs.CL

classification cs.CL
keywords citationcorrectnessfaithfulnessdocumentstrustanswersattributedattribution
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Retrieving relevant context is a common approach to reduce hallucinations and enhance answer reliability. Explicitly citing source documents allows users to verify generated responses and increases trust. Prior work largely evaluates citation correctness - whether cited documents support the corresponding statements. But citation correctness alone is insufficient. To establish trust in attributed answers, we must examine both citation correctness and citation faithfulness. In this work, we first disentangle the notions of citation correctness and faithfulness, which have been applied inconsistently in previous studies. Faithfulness ensures that the model's reliance on cited documents is genuine, reflecting actual reference use rather than superficial alignment with prior beliefs, which we call post-rationalization. We design an experiment that reveals the prevalent issue of post-rationalization, which undermines reliable attribution and may result in misplaced trust. Our findings suggest that current attributed answers often lack citation faithfulness (up to 57 percent of the citations), highlighting the need to evaluate correctness and faithfulness for trustworthy attribution in language models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

    cs.LG 2026-07 conditional novelty 6.0 of 10

    A provenance-constrained ledger runtime improves multimodal agent accuracy and trajectory faithfulness by binding claims to tool evidence and restricting repair to typed, non-amplifying operators.

  2. Evaluating and Guarding Citation Faithfulness in Agentic Scientific Synthesis

    cs.AI 2026-07 conditional novelty 6.0 of 10

    Citation-faithfulness metrics for AI science agents are verifier-dependent (3–18% on identical outputs), and a split-conformal guard provides a finite-sample catch-rate guarantee anchored on human gold.

  3. Tracing Facts or just Copies? A critical investigation of the Competitions of Mechanisms in Large Language Models

    cs.CL 2025-07 conditional novelty 5.0 of 10

    Attention heads in GPT-2 and Pythia-6.9B that promote factual output act by general copy suppression rather than selective counterfactual suppression, with domain-dependent effects that sharpen in larger models.

Pith tools