Pith. sign in

Paper Citation Record · LEDGER

OffsetBias: Leveraging Debiased Data for Tuning Evaluators

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2407.06551.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.06551 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:10:45.605360Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T17:35:43.819146Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b8114625-c371-4b9f-ba32-9776f8bf6193 · inbound

Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs cites this paper.

Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-17T16:18:01.653875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T16:18:01.560780Z digest=sha256:c6d86918053afa51b03df844d0ce24e3e21f91887b63ae8530402e64b666c22f

Observation 6e23e892-f64e-402d-9d1a-76d508738a51 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 113

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:43.821753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:7ce052f85789cb2b1bf6a942477e15f26b223e4f44ad2fab379ff51add279a13

Observation d10d3d48-12cd-4bb2-a8f7-0f6ab458808a · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 179

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:13.149064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:dd28095c80316a36012d6f74738b89fdbb253ed88e949e5fc116a59f340cecb5

Observation 07db7b63-fc3d-4db5-8548-104e0b6028c1 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:22:16.601136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:45285fb233fb235d62daa0c9757b88209348e554fefb0185859491e1f51fa430

Observation 56114032-d189-4cae-8ef5-06c0cdd39060 · inbound

Helpful Agent Meets Deceptive Judge: Understanding Vulnerabilities in Agentic Workflows cites this paper.

Helpful Agent Meets Deceptive Judge: Understanding Vulnerabilities in Agentic Workflows OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:10:45.605360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:10:45.605360Z digest=sha256:7b89beff9162a469441aa83f03c6f0398c63cfa3d1f3ece2dace5d9b26552ba1

Observation a27bbd21-8242-4092-b6a8-dd026c14e7e7 · inbound

RewardAnything: Generalizable Principle-Following Reward Models cites this paper.

RewardAnything: Generalizable Principle-Following Reward Models OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.819184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.819184Z digest=sha256:4239a10f1ed7e2b3c598a9b9880a000a64ef740e4550fb1ac12405442d49a27b

Observation a6e2eded-492c-4857-b642-92f634971cc9 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 287

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:47.252751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:47.252751Z digest=sha256:7bba5b8f40f93078b779b568932cef5f26cff5b3b9a1595e62d149625de2fff4

Observation c125dec4-740c-4de0-b143-a7f6f052dd4f · inbound

An Empirical Study of LLM-as-a-Judge: How Design Choices Impact Evaluation Reliability cites this paper.

An Empirical Study of LLM-as-a-Judge: How Design Choices Impact Evaluation Reliability OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:36.029108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:36.029108Z digest=sha256:a248af558b4e5d3f308a6d20c9134c845d46c3cab2e67535a9bb4221e181a931

Observation 0a394f64-376d-46e7-b9ed-4fdaf1a71a9c · inbound

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling cites this paper.

Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:17:05.937026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T05:16:22.274580Z digest=sha256:9629f0991022b52134e37df883e94b7fb44f2f4ff35baa0294fb7842c4b41204

Observation 0448b508-fee8-48e2-aa94-56507906638f · inbound

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization cites this paper.

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.550549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T12:53:45.767341Z digest=sha256:53e2ee7d3f13dc399e4b371033941320c6e13c4bae74527a9ade5081892ece55

Observation 5ef0d219-294c-4580-92d3-33a2f1d989a2 · inbound

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation cites this paper.

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:44.046679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:44.046679Z digest=sha256:c864bd8b3ebf5e73c476a6742329a3986abde30d4586034c17d7466708c3b54e

Observation d2012570-b487-402f-af39-a05be86e5a6b · inbound

SafeScreen: A Safety-First Screening Framework for Personalized Video Retrieval for Vulnerable Users cites this paper.

SafeScreen: A Safety-First Screening Framework for Personalized Video Retrieval for Vulnerable Users OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:25:30.934598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T11:23:43.022505Z digest=sha256:28f5b8ac28412e17fb16b276427aedb8ebe859f82c0ec0d3addece64c95468a9

Observation f60d5783-df58-4dea-be63-361f7b5252de · inbound

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines cites this paper.

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:41:14.032476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T08:14:18.535385Z digest=sha256:f684764b4088eab3600ad518e842948a52c777a321fc1032aa74ab3b877ec03a

Observation e7a42db8-9698-428b-85c8-67fa49688360 · inbound

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning cites this paper.

U-Define: Designing User Workflows for Hard and Soft Constraints in LLM-Based Planning OffsetBias: Leveraging Debiased Data for Tuning Evaluators

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:39.061387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:19:55.849451Z digest=sha256:5d12bd0d5b75ff887609cf23aae20da116dad2c43a93baeb3e6b333a958d54ec