Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2406.10216.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:55:53.556785Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f5786330-897c-49a8-9ffd-bc257326029b · inbound
Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 447d25aa-7459-4144-8226-1b23bbab127d · inbound
Test-Time Alignment via Hypothesis Reweighting Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c2be624d-9675-487d-92de-65d134ad6429 · inbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17375a3c-39b5-4c5f-93b5-1113f5dbca4a · inbound
VL-GenRM: Enhancing Vision-Language Verification via Vision Experts and Iterative Training Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b248e687-6a40-4976-a20b-c4b0a2463933 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 107
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca0fdb99-e30d-4830-9981-4e1aa7a27de4 · inbound
HEAL: A Hypothesis-Based Preference-Aware Analysis Framework Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a94515-8b5d-404a-a12a-94383c9bb9c5 · inbound
HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a77023ae-d679-40fd-88d7-c47287cfdbbd · inbound
DynaCF: Mitigating Shortcut Learning in Reward Models via Dynamic Counterfactual Sensitivity Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 49586af3-8237-4e3c-805f-e53f310e390b · inbound
Addressing Over-Refusal in LLMs with Competing Rewards Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.