Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2402.01694.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:21:28.390947Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T06:56:44.378731Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c3258aed-6092-456c-8e05-fe4a1ca085de · inbound
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time ARGS: Alignment as Reward-Guided Search
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eabd63f-8d6d-4fa7-a945-faa52e58f4af · inbound
BiasFilter: An Inference-Time Debiasing Framework for Large Language Models ARGS: Alignment as Reward-Guided Search
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ae2857e-1b3d-4a3a-bc91-16c3ace39994 · inbound
Disentangled Safety Adapters Enable Efficient Guardrails and Flexible Inference-Time Alignment ARGS: Alignment as Reward-Guided Search
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55bc5fd5-0f8d-40b5-8ddd-d4c87946db01 · inbound
LLM-ML Teaming: Integrated Symbolic Decoding and Gradient Search for Valid and Stable Generative Feature Transformation ARGS: Alignment as Reward-Guided Search
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3223f1b-673c-43ab-ae62-693610b0990b · inbound
Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis ARGS: Alignment as Reward-Guided Search
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 041852b0-1b03-4c76-b9c6-1bf402807f70 · inbound
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling ARGS: Alignment as Reward-Guided Search
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6a67e93-2552-4e32-97b0-f01c49869cd0 · inbound
Bradley-Terry and Multi-Objective Reward Modeling Are Complementary ARGS: Alignment as Reward-Guided Search
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50e73708-3b7d-4fc0-90bc-a99fa3c9ceb4 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities ARGS: Alignment as Reward-Guided Search
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b960119-1126-497f-8a76-8dfbd9da11f7 · inbound
A Survey on Training-free Alignment of Large Language Models ARGS: Alignment as Reward-Guided Search
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f9ff3c-6405-491a-bbab-d6af7d01fcbe · inbound
Virtual Agent Economies ARGS: Alignment as Reward-Guided Search
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94dfb535-c47f-463f-9aa0-cb9d3c44c5c4 · inbound
T-POP: Test-Time Personalization with Online Preference Feedback ARGS: Alignment as Reward-Guided Search
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8888d26-c4bb-474f-8d6e-a3042504f5fd · inbound
Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards ARGS: Alignment as Reward-Guided Search
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80061e44-3d5b-4f76-975c-ae0241157149 · inbound
Representation-Based Exploration for Language Models: From Test-Time to Post-Training ARGS: Alignment as Reward-Guided Search
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce5fda3-b4e0-4e8f-b7b2-94d90910bdd8 · inbound
Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective ARGS: Alignment as Reward-Guided Search
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63eda618-2b8d-4180-ae03-67eaa70e4fea · inbound
Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning ARGS: Alignment as Reward-Guided Search
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b261bb4-3bc0-405d-8574-37ed6cabbf66 · inbound
Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control ARGS: Alignment as Reward-Guided Search
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 68bba25e-0882-401e-95c8-7161af716b3f · inbound
Pref-CTRL: Preference Driven LLM Alignment using Representation Editing ARGS: Alignment as Reward-Guided Search
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fdb043b-ea26-481c-b498-e96acf9505b8 · inbound
Training-Free Cultural Alignment of Large Language Models via Persona Disagreement ARGS: Alignment as Reward-Guided Search
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e3afd58-137b-463b-ad8d-35702c7b44c6 · inbound
Spectral Souping: A Unified Framework for Online Preference Alignment ARGS: Alignment as Reward-Guided Search
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b6cdbd5-0343-4c1f-ae14-6ef2bd6f0982 · inbound
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning ARGS: Alignment as Reward-Guided Search
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3bbb5d27-7c5d-49d2-b8c2-2b40751759b5 · inbound
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning ARGS: Alignment as Reward-Guided Search
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2daf7f82-38aa-4714-815e-050469ebd7b7 · inbound
Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control ARGS: Alignment as Reward-Guided Search
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2aec05c3-0fba-4c6c-9285-d088541337c5 · inbound
Safe Inference-Time Alignment via Lagrangian Reward Augmentation ARGS: Alignment as Reward-Guided Search
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71560dde-6950-480f-8a10-0553bd785f97 · inbound
IFHierBench: Hierarchical Instruction Following for Large Language Models ARGS: Alignment as Reward-Guided Search
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.