Pith. sign in

Paper Citation Record · LEDGER

RewardBench: Evaluating Reward Models for Language Modeling

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2403.13787.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.13787 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:52:50.233278Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4837305e-d645-41f6-abde-f621aa5cc087 · inbound

Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing cites this paper.

Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing RewardBench: Evaluating Reward Models for Language Modeling

Reference 119

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T06:58:36.864176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-16T06:58:36.684583Z digest=sha256:858cc9307d45e1e4a0525a8625ed973748bc0fee864f0907ac7a6fc29f66e6f2

Observation f60d7cd4-f05a-43a8-bbe0-7b35578a78e3 · inbound

Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs cites this paper.

Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs RewardBench: Evaluating Reward Models for Language Modeling

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T16:18:01.640159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T16:18:01.560780Z digest=sha256:88becd3f771b17c156c2a17dc3f37c785a32492a6cc379dec22937a71f7f14ae

Observation be6c4a43-db3f-4f82-a220-d8163efb1b52 · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods RewardBench: Evaluating Reward Models for Language Modeling

Reference 120

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:37.061394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:ab8d3ced3697cef0a237b3fd117338a3cf037cb7e326ab9d6b971bd516f69aa2

Observation 936ed788-f2ff-4077-8299-61906f93abfb · inbound

Qwen2.5 Technical Report cites this paper.

Qwen2.5 Technical Report RewardBench: Evaluating Reward Models for Language Modeling

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:25:27.895174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T06:25:00.376073Z digest=sha256:e0fc88feb789539d734488c806082ffbc7f8f2d57bd9db62d631b6309387655f

Observation c045ae24-dc61-44d1-badd-76157314aa2a · inbound

Improving Video Generation with Human Feedback cites this paper.

Improving Video Generation with Human Feedback RewardBench: Evaluating Reward Models for Language Modeling

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T15:30:02.708122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T15:30:02.578430Z digest=sha256:0e57ce94833104c6d499906f632a1bc0398756941a4205ba8bd7675267c72e33

Observation e7d46cc0-4d97-4312-9f6c-32857585d627 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation RewardBench: Evaluating Reward Models for Language Modeling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T11:22:16.836111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:56b15d8a2869dc0d688a21c56ff0234804ad8dc3916a05721bf45b5748a51b36

Observation 7a841bc7-63da-41a7-ba87-bb019948e034 · inbound

Exploring the Secondary Risks of Large Language Models cites this paper.

Exploring the Secondary Risks of Large Language Models RewardBench: Evaluating Reward Models for Language Modeling

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:42:13.993428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T09:40:58.067398Z digest=sha256:1241ecd48adbbf8e1d862b6d6a2b7a2e54bdb84e2e8eb290c9129e21e5e653fe

Observation 6ecefc4e-1d42-41d6-8b40-73afba4c166b · inbound

TabArena: A Living Benchmark for Machine Learning on Tabular Data cites this paper.

TabArena: A Living Benchmark for Machine Learning on Tabular Data RewardBench: Evaluating Reward Models for Language Modeling

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T08:42:12.655284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T08:41:35.789878Z digest=sha256:3ad6dcf5bd05940d24c7d743273ba42b52a565ac4ebc3c12954844d8e181fc5b

Observation 02e09a74-85eb-419c-b489-28852c0dc2f3 · inbound

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap cites this paper.

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap RewardBench: Evaluating Reward Models for Language Modeling

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.617954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T23:46:24.208438Z digest=sha256:ce662de3a5aa3f6afbc455042cc84911c5240c4f8c6a386c1bc65d58ede6da55

Observation b3ec0fdf-f59d-440d-9450-2477156f23d6 · inbound

Controlling Multimodal LLMs via Reward-guided Decoding cites this paper.

Controlling Multimodal LLMs via Reward-guided Decoding RewardBench: Evaluating Reward Models for Language Modeling

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T19:52:50.233278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:52:50.233278Z digest=sha256:bbe8a97331f829bed6003b3810ec3161a4f3c6b9184bcf21267dccf945396973

Observation fe8d6d21-146a-42de-9908-d1c58720b217 · inbound

Hermes 4 Technical Report cites this paper.

Hermes 4 Technical Report RewardBench: Evaluating Reward Models for Language Modeling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T16:32:54.405593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:32:54.405593Z digest=sha256:0c95101272b3e81e6a8e60aab9999d9e3b61d1b82e2cb9493251b9e1d70bb4eb

Observation 4b4df1d4-e9d8-435f-ae1b-2cbf8d1d07bf · inbound

Counterfactual Reward Model Training for Bias Mitigation in Multimodal Reinforcement Learning cites this paper.

Counterfactual Reward Model Training for Bias Mitigation in Multimodal Reinforcement Learning RewardBench: Evaluating Reward Models for Language Modeling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T15:44:38.975171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:44:38.975171Z digest=sha256:6c750e44a71e6c00460d7cfc47e1ad3fd7febd610acf29e7091c98682a8dcd79

Observation cf22cafc-b38e-42bd-a6d0-9e4780abe1ed · inbound

HEAL: A Hypothesis-Based Preference-Aware Analysis Framework cites this paper.

HEAL: A Hypothesis-Based Preference-Aware Analysis Framework RewardBench: Evaluating Reward Models for Language Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:25:35.346495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:25:35.346495Z digest=sha256:58cb427ca4aec294c329daa9961386e6884c2cda040215d78e09c945f98510d4

Observation a3727717-6844-4be7-9d66-1c15b74f0283 · inbound

Evalet: Evaluating Large Language Models through Functional Fragmentation cites this paper.

Evalet: Evaluating Large Language Models through Functional Fragmentation RewardBench: Evaluating Reward Models for Language Modeling

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:01:40.069981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:57:25.259866Z digest=sha256:8d166ebccdb66e2ca345f8a673ec357747b24098f3ab692167a91a524f6ba88c

Observation 1e925bf3-af47-4214-bfe5-5b02f232054a · inbound

Adaptive Margin RLHF via Preference over Preferences cites this paper.

Adaptive Margin RLHF via Preference over Preferences RewardBench: Evaluating Reward Models for Language Modeling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T14:52:20.583374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:52:20.583374Z digest=sha256:6da257c8814e0ceeb56c9868a1823b35b21afaca3904eb1c6c28ee8d926e8546

Observation 45ff3ac9-20ca-4966-b64f-d54c27368c1f · inbound

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images cites this paper.

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images RewardBench: Evaluating Reward Models for Language Modeling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:16:14.370848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T10:13:08.112072Z digest=sha256:55d74ce68d64344b58638af342b27fb88a55c3de4f157eed0e073b7b020c453c

Observation 2ec61eb6-808a-4173-9cb2-004f4a719cde · inbound

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling cites this paper.

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling RewardBench: Evaluating Reward Models for Language Modeling

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T03:20:49.320809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T03:15:57.744706Z digest=sha256:417b146afc86b8d1e4f74b47177b4b1d0350c131fc37df197c0eb3a98f4b3ed0

Observation bb48b874-e55b-4f2c-8572-c3f6a9ed6cb4 · inbound

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning cites this paper.

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning RewardBench: Evaluating Reward Models for Language Modeling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T07:43:24.665067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:43:24.665067Z digest=sha256:98c07febff4ac63b4d1b01815634987ad3bf7d75431f90de79ee445d4502881c

Observation 7edbe698-b7a3-458e-9e3d-ec620a6efa60 · inbound

SAM 3D: 3Dfy Anything in Images cites this paper.

SAM 3D: 3Dfy Anything in Images RewardBench: Evaluating Reward Models for Language Modeling

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:47:13.371686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T11:47:11.554134Z digest=sha256:0db56591dfb551f8c373b873686d1a1e2bc0cf9f24cf3659912a588353294458

Observation fb3c57c4-7e89-44c6-8440-31393d531545 · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data RewardBench: Evaluating Reward Models for Language Modeling

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:22.075845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:22.075845Z digest=sha256:96f5e6d50ee07686593f01af28dee5a54199251d9b30187e30660bf47f486b7e

Observation ad7b4d95-86db-494a-8959-c9d62fcc80f7 · inbound

AI Can Learn Scientific Taste cites this paper.

AI Can Learn Scientific Taste RewardBench: Evaluating Reward Models for Language Modeling

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T18:14:50.265339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:14:50.265339Z digest=sha256:c785974f946637981ab8d55508f3162999969bd04d306b23b7fb69419889d63d

Observation 9e046eb5-bcad-4b7c-a6db-61232767e668 · inbound

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems cites this paper.

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems RewardBench: Evaluating Reward Models for Language Modeling

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:26:02.367040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:56:00.449776Z digest=sha256:74106ea29da1fba93a498db2ab994334e181a69f63c531093dcb8797530c0151

Observation 79d4d7a1-b5a7-4336-ad5c-715649dda6db · inbound

Four-Axis Decision Alignment for Long-Horizon Enterprise AI Agents cites this paper.

Four-Axis Decision Alignment for Long-Horizon Enterprise AI Agents RewardBench: Evaluating Reward Models for Language Modeling

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:06:05.342606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T02:19:57.341402Z digest=sha256:11fbd78e0125cd3fcbd2a79c4ad50447856f042babde0f65ef6d5ad9d9760d27

Observation ed37f525-9e2d-4ef8-9279-cf37a138580b · inbound

Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model cites this paper.

Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model RewardBench: Evaluating Reward Models for Language Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.161906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T22:29:53.202779Z digest=sha256:9292ed7f2b50b8ca991856caeee70962220e0286b06421702748f1582b1fcade

Observation fb242eb0-85d6-4c4d-ab0b-1890bbbbcbea · inbound

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines cites this paper.

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines RewardBench: Evaluating Reward Models for Language Modeling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:41:14.089601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T08:14:18.535385Z digest=sha256:dcb3cf8cb9d8ca114b59bbf235eeb6ccd97a19244f6d946e1a90c44d1d839d05

Observation b2bcb6ca-2584-4584-beb6-d2a9e5679d51 · inbound

Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning cites this paper.

Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning RewardBench: Evaluating Reward Models for Language Modeling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:00:35.104015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T19:14:22.162163Z digest=sha256:b774a6c98352fac8688446741994334cee0f0d616527f3491681a0ca7f0bbe8a

Observation 15ae18c7-113b-450f-b49e-deaa22192fbb · inbound

Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning cites this paper.

Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning RewardBench: Evaluating Reward Models for Language Modeling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:57.990159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T02:10:46.725354Z digest=sha256:3b0aa7c386a5f16b1ba190f751da3e1fa2e3f7fef249c748f81553f1217f2e2e

Observation e458652d-1cae-4cf4-ac8c-a05762a10c21 · inbound

Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone cites this paper.

Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone RewardBench: Evaluating Reward Models for Language Modeling

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:40:43.244784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T18:11:42.971428Z digest=sha256:adbbf2bf3d90bb7253da80c1422a3bf7152831e2248cd09070c54167a658158b

Observation f3255c4e-82e1-4cae-b4c5-c13fa899701c · inbound

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders cites this paper.

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders RewardBench: Evaluating Reward Models for Language Modeling

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T22:49:10.154459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T22:48:54.238767Z digest=sha256:5955f8a6e62ec0b5a0d7cf0602ff5fc11faacb5b82df054b8aec19f4749b6ca2

Observation fc8c028b-4999-46e6-98cb-c0ec40f0b283 · inbound

Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR cites this paper.

Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR RewardBench: Evaluating Reward Models for Language Modeling

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T05:03:03.609176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T05:02:35.271960Z digest=sha256:fe956072825891bccd0fd376fac158a26ff27db6958815cd7fe40fd209fd6259

Observation 54af2848-6ec6-4303-abc2-74d4261a815c · inbound

Pairwise Reference Alignment as a Model-Level Ordinal Observable cites this paper.

Pairwise Reference Alignment as a Model-Level Ordinal Observable RewardBench: Evaluating Reward Models for Language Modeling

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T19:25:59.706899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:46:11.450812Z digest=sha256:b91fd4922e63157ef3061a2fc57803d8b5e80da6451a955c4308e6a36e6c71d0

Observation 8f55d6e4-a7a1-4529-8c44-ed14ca1e86b8 · inbound

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges cites this paper.

PReMISE: Policy Rubrics as Measurement Specifications for LLM Judges RewardBench: Evaluating Reward Models for Language Modeling

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.796432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:33:05.343496Z digest=sha256:2aa2981c550e32bba9ce22f6e9f3ea001ab002d67374610bc07218730a0583d2

Observation 363a698c-d10d-4f83-a758-774e45d93061 · inbound

EST-PRM: Stress-Testing Process Reward Models Before They Become Load-Bearing cites this paper.

EST-PRM: Stress-Testing Process Reward Models Before They Become Load-Bearing RewardBench: Evaluating Reward Models for Language Modeling

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:22:34.486732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T19:18:55.398976Z digest=sha256:5ca6bed46017c07947b8cf4b67c103783f9e14b80b7105b946429654d0bd2d74

Observation add9a932-09bc-4dfd-9171-9ca2dde3b2a5 · inbound

A Finite-Calibration Regime Map for LLM Judge Panels cites this paper.

A Finite-Calibration Regime Map for LLM Judge Panels RewardBench: Evaluating Reward Models for Language Modeling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:13.126347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T17:34:28.258222Z digest=sha256:6c516144d22c1ae432304f02a7c5db6781ce3e8857a1fef8259f537c753edcfd

Observation 288c43ac-8ee4-4bce-89b5-c9943aeebd86 · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization RewardBench: Evaluating Reward Models for Language Modeling

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:37:30.532714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:f95be4f04e49c29b1977a0d3610e9b6a49546f40388a789eb9fa03902a208731

Observation 074aa224-c373-46aa-ae16-23f30a1aab78 · inbound

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups cites this paper.

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups RewardBench: Evaluating Reward Models for Language Modeling

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-06-26T12:59:29.168628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T12:58:09.227648Z digest=sha256:8089fcfdf82e14ad7b034a44e1e2b0f0558798afc15cbd50736cf073e4092d98

Observation 29dd4dd1-dfcd-457f-a273-5deadf5ffba2 · inbound

Addressing Over-Refusal in LLMs with Competing Rewards cites this paper.

Addressing Over-Refusal in LLMs with Competing Rewards RewardBench: Evaluating Reward Models for Language Modeling

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:55:35.489468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-01T06:59:12.695984Z digest=sha256:24dd7ae28401f1190e14e95c66a3bea8043b8e92196be59f22c4d449c8338502

Observation 157aae37-3740-48b6-9780-a27d40860516 · inbound

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents cites this paper.

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents RewardBench: Evaluating Reward Models for Language Modeling

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:05:28.936347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-01T05:59:13.631078Z digest=sha256:7678378b1b5f9763db944b27d2838477ffb707e93d12883c5949b09e7e47cf0a

Observation 50d8a596-f9d8-4d93-b405-c6eefd3e9c9d · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay RewardBench: Evaluating Reward Models for Language Modeling

Reference 143

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:d6dc3f9a50f809228998faacf667fd97dfad20f73960983910ef63f04a31a7d6

Observation 4d5d7f8e-46e4-4991-8960-a65454de3acb · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay RewardBench: Evaluating Reward Models for Language Modeling

Reference 144

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:48.152223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:48.152223Z digest=sha256:faee3b87fdfacfaca686aeddfb4ea9a97be3c93819f5970688fc2a15112ae724

Observation 46da7b85-3aed-42ef-a3a3-0aba1766dafe · inbound

When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability cites this paper.

When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability RewardBench: Evaluating Reward Models for Language Modeling

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:56:50.525489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T05:47:32.670216Z digest=sha256:0d93868254295b29d0a49bffcaaac9ee106bf6358aa070b3d77aa24705c8545d

Observation 96b1c4e1-4c5b-43a2-a34c-d67a402cb7ea · inbound

SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning cites this paper.

SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning RewardBench: Evaluating Reward Models for Language Modeling

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-14T07:59:56.441098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T07:59:56.441098Z digest=sha256:3757eb1f42638400d16b6a9bd1481fca7588b0dd0952ced201f8e61b24586a4d

Observation 32b72cbd-74b6-441e-9202-a0963abf23a3 · inbound

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias cites this paper.

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias RewardBench: Evaluating Reward Models for Language Modeling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-14T02:33:34.084111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T02:33:34.084111Z digest=sha256:1111d2bbac05cbef1b48689a06f320405c2af0e78da72829545d4d4ca1a9d744

Observation 2ee9dde9-c887-4ede-b44e-8cfc9300ca8b · inbound

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary cites this paper.

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary RewardBench: Evaluating Reward Models for Language Modeling

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T15:08:37.823880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:08:37.823880Z digest=sha256:e907ae00d076570b2a9de6c33af370d110279414305a126ae5976ec86e644ac5

Observation 41b013a8-943d-46c9-b0db-4fe04a4b6371 · inbound

Evaluating and Guarding Citation Faithfulness in Agentic Scientific Synthesis cites this paper.

Evaluating and Guarding Citation Faithfulness in Agentic Scientific Synthesis RewardBench: Evaluating Reward Models for Language Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T07:46:11.984755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:46:11.984755Z digest=sha256:1dcd3f6076bd9f3c3db7aa9276976286041c8845b937737c80825afca062a84c

Observation 2f1d217c-8280-4bb1-8336-802bfeaecb64 · inbound

Test-Time Scaling via Error Localization cites this paper.

Test-Time Scaling via Error Localization RewardBench: Evaluating Reward Models for Language Modeling

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-01T07:28:29.938151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:28:29.938151Z digest=sha256:936b55ccda3f4a985d2826cc3059cd2e8c0615b391b9fbce8524577cc5ab940c

Observation d10d2cd2-a4c2-4cb7-88a7-3f85aa373f0c · inbound

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback cites this paper.

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback RewardBench: Evaluating Reward Models for Language Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T03:14:11.464266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:14:11.464266Z digest=sha256:6751c36c85a0365c003c04517cf3b1ef91133a117cd60047ada95a18c1bcf807

Observation 6d6e2847-5870-4ffc-acac-8822098b0191 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges RewardBench: Evaluating Reward Models for Language Modeling

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:25.686463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:25.686463Z digest=sha256:d1e43716336571cd382c5021bb11effaf71b6b45f75097e422e7ec4fc6d15340