Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2410.08193.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:22:59.901678Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T01:20:52.147219Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 513ae582-aa7c-4aba-a346-c1826eabce63 · inbound
Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8b48618-9ec6-41d2-b9c0-7b2731102926 · inbound
UCD: Unlearning in LLMs via Contrastive Decoding GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation affda21e-2a61-4a2e-a931-ae8f8241e175 · inbound
From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bf46a25-5c75-4d75-ae39-4efc3705e560 · inbound
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e914c68-0a1e-45aa-85aa-b74a332cbd70 · inbound
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e491323c-cac4-4d1f-9eb0-fb71d021fbfb · inbound
A Survey on Training-free Alignment of Large Language Models GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d2e5f6e-7f7c-4fb8-a578-6658cb4e8d47 · inbound
Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1272e9f4-5e14-47ea-9842-b35fa6d61fee · inbound
Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31c699c7-3224-44c8-b4ca-1fc5a30c9f91 · inbound
Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a88186b-a2c7-42e4-8f82-912266d671c2 · inbound
Confidence-Aware Alignment Makes Reasoning LLMs More Reliable GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5b03e13-64cd-4850-af1b-d7723e7ba66f · inbound
Common-agency Games for Multi-Objective Test-Time Alignment GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 237
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 536ccdae-99c9-4ac8-a60e-d67c75ff6468 · inbound
Inference-Time Policy Alignment for Fair Reinforcement Learning GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fefa55d-22c4-41d3-947e-649ac10207e5 · inbound
Aligning Large Vision-Language Models at Test Time: A Trajectory-Guided Structured Sampling Approach GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.