Pith. sign in

Paper Citation Record · LEDGER

Interpreting Language Reward Models via Contrastive Explanations

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 2 inbound Pith citation observations for arXiv:2411.16502.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16502 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 2 of 2 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:09.068412Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T05:17:05.898300Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eb42c8bb-e6f5-46c5-bd0c-475a44b6d185 · inbound

Multi-Domain Explainability of Preferences cites this paper.

Multi-Domain Explainability of Preferences Interpreting Language Reward Models via Contrastive Explanations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:09.068412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:05:09.068412Z digest=sha256:29b567833bfaa30f220eb5b92045f526b106fa500f8a6a62e762261cb29f2c12

Observation a250044e-9fae-47e7-a728-f8814d8e2bf9 · inbound

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling cites this paper.

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling Interpreting Language Reward Models via Contrastive Explanations

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:17:05.901461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T05:16:22.274580Z digest=sha256:e4c4aa37271259b987322ef70a6dcd11e8bba6bff3c6e8a1410ad8bc3ddc904b