Pith. sign in

Paper Citation Record · LEDGER

Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

As of 6 August 2026, this Paper Citation Record lists 2 of 2 outbound references and 7 inbound Pith citation observations for arXiv:2512.00920.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.00920 v5

Coverage vector

measured 2 of 2 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T18:27:04.488695Z

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T07:01:23.147046Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-30T16:35:12.226032Z

Reference resolution

2 of 2 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 17d2b84f-01b3-4f9d-97ac-de21864e110d · outbound

This paper cites A statistically significant result (i.e., a small p-value) implies that we can reject the null hypothesis that the data follows a normal distribution.

Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios A statistically significant result (i.e., a small p-value) implies that we can reject the null hypothesis that the data follows a normal distribution

Reference 1

Resolution
malformed identifier
raw_fallback, observed 2026-05-21T18:30:29.415116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T18:27:04.488695Z digest=sha256:f4ab61585c3effcab49ed5334ab5d72a1d2c030b18394ea63e1dc524b263eea1

Observation 5af36abe-227a-4eb6-aec3-8edc72c88c93 · outbound

This paper cites sum of RM effect sizes.

Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios sum of RM effect sizes

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-05-21T18:30:29.412078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T18:27:04.488695Z digest=sha256:786626b193e2826f627410d32207e79533fc8d2f106d6d6a99cdb90d16f66ac2

Pith citing papers

Observation 99793f72-c83c-429a-811f-da018396c9a2 · inbound

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks cites this paper.

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-16T07:00:42.718997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T07:00:18.807922Z digest=sha256:a2d7a95c94d985c1873c7eb3b895cb16d0f774a35bb131d3834783bfc61d8ddf

Observation 49b12e28-c744-4554-b048-91303911f464 · inbound

BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models cites this paper.

BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 124

Resolution
verified exact
local_arxiv, observed 2026-05-13T06:07:22.415442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T06:03:32.270553Z digest=sha256:3ef1f21ea768a84880d80671b3bbc0f6365866ce6d2e2de1b32308a83b23d808

Observation 41d56ea7-68fb-4b13-a056-de672b69187d · inbound

The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Management cites this paper.

The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Management Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:23.607615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T05:22:22.228384Z digest=sha256:7bfb2eb2b7624304a4f9bd12f66150e649f803849f94818a66a01ce812c14eb1

Observation 9414ad72-b38a-42c6-a92d-18a6fdd235aa · inbound

The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Management cites this paper.

The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Management Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:35:12.227344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T16:34:29.680105Z digest=sha256:cc1fc618532d5fbeb47ad8f9d14415bc7fc652b4d5d874d0c82a711ed43ab10f

Observation 4c02bab2-08be-4125-b66d-970e92c09a43 · inbound

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG cites this paper.

Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T13:03:26.570848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T12:55:50.485412Z digest=sha256:d01c80d9083e571e343e513049d30cfd1c32ba81a6c1280443732175a45378e8

Observation 80b71b26-31f0-4f58-8a9d-88712b151971 · inbound

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems cites this paper.

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:52:36.027022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T18:45:33.322389Z digest=sha256:a5446bbb61d86a68ddfa1c2901520b863c11ccecf9b1af9b56185ecdbf39baea

Observation 0e1aaed2-53ac-4a13-8bc1-9161665f35a8 · inbound

Phantom Guardrails: When Self-Improving Agent Harnesses Fix Failures That Never Happened cites this paper.

Phantom Guardrails: When Self-Improving Agent Harnesses Fix Failures That Never Happened Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T07:01:23.147046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:01:23.147046Z digest=sha256:b1c7be675b776b1b3bd2d2566ffb9f970afb5e1928d34523ddffc4569fab25cf