Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2402.01694.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:31:57.817590Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T06:56:44.378731Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 898702c6-e737-4651-8b91-939f6f63ebeb · inbound
RED: Unleashing Token-Level Rewards from Holistic Feedback via Reward Redistribution ARGS: Alignment as Reward-Guided Search
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 965e2f08-0769-4155-9b02-14dad57a1b4e · inbound
Dynamic Rewarding with Prompt Optimization Enables Tuning-free Self-Alignment of Language Models ARGS: Alignment as Reward-Guided Search
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69102108-2525-434e-b73e-5e9fd28869b8 · inbound
Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review ARGS: Alignment as Reward-Guided Search
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 124f1ebe-35b4-4a99-8e30-73cfad6e3a45 · inbound
On Almost Surely Safe Alignment of Large Language Models at Inference-Time ARGS: Alignment as Reward-Guided Search
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 572fb1eb-4d61-48a5-b9af-2e6be92ca496 · inbound
Persona-judge: Personalized Alignment of Large Language Models via Token-level Self-judgment ARGS: Alignment as Reward-Guided Search
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64c21c09-81fb-4ae8-ba5e-d5da690a9317 · inbound
CEC-Zero: Chinese Error Correction Solution Based on LLM ARGS: Alignment as Reward-Guided Search
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3f9c18b-c540-4974-8201-5feb17b9aa2d · inbound
Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment ARGS: Alignment as Reward-Guided Search
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f7ffe76-90bb-44b7-ad54-10d5e720510a · inbound
MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO ARGS: Alignment as Reward-Guided Search
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3258aed-6092-456c-8e05-fe4a1ca085de · inbound
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time ARGS: Alignment as Reward-Guided Search
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eabd63f-8d6d-4fa7-a945-faa52e58f4af · inbound
BiasFilter: An Inference-Time Debiasing Framework for Large Language Models ARGS: Alignment as Reward-Guided Search
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ae2857e-1b3d-4a3a-bc91-16c3ace39994 · inbound
Disentangled Safety Adapters Enable Efficient Guardrails and Flexible Inference-Time Alignment ARGS: Alignment as Reward-Guided Search
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 55bc5fd5-0f8d-40b5-8ddd-d4c87946db01 · inbound
LLM-ML Teaming: Integrated Symbolic Decoding and Gradient Search for Valid and Stable Generative Feature Transformation ARGS: Alignment as Reward-Guided Search
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba331d66-217c-4371-a38e-96c144afed5f · inbound
Relic: Enhancing Reward Model Generalization for Low-Resource Indic Languages with Few-Shot Examples ARGS: Alignment as Reward-Guided Search
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db04d10-d604-4cc7-a8b4-31080f448bea · inbound
Aligning Frozen LLMs by Reinforcement Learning: An Iterative Reweight-then-Optimize Approach ARGS: Alignment as Reward-Guided Search
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3223f1b-673c-43ab-ae62-693610b0990b · inbound
Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis ARGS: Alignment as Reward-Guided Search
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 041852b0-1b03-4c76-b9c6-1bf402807f70 · inbound
Teach a Reward Model to Correct Itself: Reward Guided Adversarial Failure Discovery for Robust Reward Modeling ARGS: Alignment as Reward-Guided Search
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e6a67e93-2552-4e32-97b0-f01c49869cd0 · inbound
Bradley-Terry and Multi-Objective Reward Modeling Are Complementary ARGS: Alignment as Reward-Guided Search
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50e73708-3b7d-4fc0-90bc-a99fa3c9ceb4 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities ARGS: Alignment as Reward-Guided Search
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b960119-1126-497f-8a76-8dfbd9da11f7 · inbound
A Survey on Training-free Alignment of Large Language Models ARGS: Alignment as Reward-Guided Search
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f9ff3c-6405-491a-bbab-d6af7d01fcbe · inbound
Virtual Agent Economies ARGS: Alignment as Reward-Guided Search
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94dfb535-c47f-463f-9aa0-cb9d3c44c5c4 · inbound
T-POP: Test-Time Personalization with Online Preference Feedback ARGS: Alignment as Reward-Guided Search
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8888d26-c4bb-474f-8d6e-a3042504f5fd · inbound
Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards ARGS: Alignment as Reward-Guided Search
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80061e44-3d5b-4f76-975c-ae0241157149 · inbound
Representation-Based Exploration for Language Models: From Test-Time to Post-Training ARGS: Alignment as Reward-Guided Search
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce5fda3-b4e0-4e8f-b7b2-94d90910bdd8 · inbound
Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective ARGS: Alignment as Reward-Guided Search
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63eda618-2b8d-4180-ae03-67eaa70e4fea · inbound
Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning ARGS: Alignment as Reward-Guided Search
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8b261bb4-3bc0-405d-8574-37ed6cabbf66 · inbound
Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control ARGS: Alignment as Reward-Guided Search
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 68bba25e-0882-401e-95c8-7161af716b3f · inbound
Pref-CTRL: Preference Driven LLM Alignment using Representation Editing ARGS: Alignment as Reward-Guided Search
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0fdb043b-ea26-481c-b498-e96acf9505b8 · inbound
Training-Free Cultural Alignment of Large Language Models via Persona Disagreement ARGS: Alignment as Reward-Guided Search
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6e3afd58-137b-463b-ad8d-35702c7b44c6 · inbound
Spectral Souping: A Unified Framework for Online Preference Alignment ARGS: Alignment as Reward-Guided Search
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7b6cdbd5-0343-4c1f-ae14-6ef2bd6f0982 · inbound
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning ARGS: Alignment as Reward-Guided Search
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3bbb5d27-7c5d-49d2-b8c2-2b40751759b5 · inbound
OPPO: Bayesian Value Recursion for Token-Level Credit Assignment in LLM Reasoning ARGS: Alignment as Reward-Guided Search
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2daf7f82-38aa-4714-815e-050469ebd7b7 · inbound
Activation Steering of Video Generation Models via Reduced-Order Linear Optimal Control ARGS: Alignment as Reward-Guided Search
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2aec05c3-0fba-4c6c-9285-d088541337c5 · inbound
Safe Inference-Time Alignment via Lagrangian Reward Augmentation ARGS: Alignment as Reward-Guided Search
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71560dde-6950-480f-8a10-0553bd785f97 · inbound
IFHierBench: Hierarchical Instruction Following for Large Language Models ARGS: Alignment as Reward-Guided Search
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.