A protocol-level identifiability audit over finite behavioral policy classes shows that base-only observation under-identifies the selective-response estimand, full support identifies it, and base accuracy diverges from intervention-response fidelity in two instruction-tuned models.
InProceedings of the 2021 ACM Conference on Fairness, Accountability, and Trans- parency, pages 375–385, New York, NY , USA
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Beyond Local Accuracy: A Protocol-Level Identifiability Audit for Controlled LLM Reasoning Evaluation
A protocol-level identifiability audit over finite behavioral policy classes shows that base-only observation under-identifies the selective-response estimand, full support identifies it, and base accuracy diverges from intervention-response fidelity in two instruction-tuned models.