Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2306.17492.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:45:56.960790Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-24T03:23:49.545047Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 40d25fbf-7ebd-45be-9e3e-a9900ab27456 · inbound
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs Preference Ranking Optimization for Human Alignment
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f5fef7b2-412f-455c-ba81-f5e01666bb5b · inbound
A Comprehensive Overview of Large Language Models Preference Ranking Optimization for Human Alignment
Reference 171
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 329d016a-54f8-49f1-b1c0-90de4cbce338 · inbound
Scaling Relationship on Learning Mathematical Reasoning with Large Language Models Preference Ranking Optimization for Human Alignment
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19e54d70-74d9-4a2a-b89b-e4968bd647a5 · inbound
Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations Preference Ranking Optimization for Human Alignment
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 534d41e2-0769-4bdc-bcd2-c9068ffc6a9e · inbound
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models Preference Ranking Optimization for Human Alignment
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c6015d5c-045a-42b9-9395-26cdf9f6bc42 · inbound
A Survey on Knowledge Distillation of Large Language Models Preference Ranking Optimization for Human Alignment
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ab3de7c-cef6-40c9-9e6c-860760496a9e · inbound
ORPO: Monolithic Preference Optimization without Reference Model Preference Ranking Optimization for Human Alignment
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b0894764-c8cd-48c4-9789-8e388b060228 · inbound
UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types Preference Ranking Optimization for Human Alignment
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a0eac52-e50a-4f56-8924-1a3c45f46d9c · inbound
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Preference Ranking Optimization for Human Alignment
Reference 140
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29a73b90-9297-4977-b2fb-e39741917e83 · inbound
LPOI: Listwise Preference Optimization for Vision Language Models Preference Ranking Optimization for Human Alignment
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5879515a-38ee-40e5-bfec-08e0afebe51f · inbound
From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Preference Ranking Optimization for Human Alignment
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 487224f8-3174-4d64-a19a-5037895ca1d1 · inbound
HAEPO: History-Aggregated Exploratory Policy Optimization Preference Ranking Optimization for Human Alignment
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b76bbd58-df78-4e81-b59a-ae9dd490350d · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Preference Ranking Optimization for Human Alignment
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a78797e-5a87-438a-8e17-98d56cd6abd1 · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Preference Ranking Optimization for Human Alignment
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 141c671e-9ff5-4bd7-9acc-a92dc05504ba · inbound
(Towards) Scalable Reliable Automated Evaluation with Large Language Models Preference Ranking Optimization for Human Alignment
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.