Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:38:10.381600Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2507.07725.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:38:10.381600Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 757181e0-3525-40f8-bb06-84aa17af40a3 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization In: Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96c56c30-3ddb-42ec-a8d2-3f4ce638959a · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization the method of paired comparisons
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a3dfb02-53da-46da-a6b8-b9cb4713ac29 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edf792c8-05a7-415b-8318-369e1b8a0856 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f0bc812-d474-4d61-bbfa-f4c4e9543a29 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Direct Preference Knowledge Distillation for Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f567fc3-240d-42c3-9de7-06984d22f8d7 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Rho-1: Not All Tokens Are What You Need
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0468e90-10fe-4539-9ed8-a4cd3ff64abb · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fd08862-22af-4123-86d5-2c1132d1fd50 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization SimPO: Simple Preference Optimization with a Reference-Free Reward
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 542a403c-c130-4b1b-b348-5626121a0de7 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e5aa232-9d66-42ed-861c-1a5db1eb57e1 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Proximal Policy Optimization Algorithms
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2a4ea7f-6a3c-4fa5-9f01-286a080192e0 · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2aaae1ee-68f1-41d5-b2ac-a08da058766e · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization , " * write output.state after.block = add.period write
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 252f0a5b-4b4f-4d3e-a159-bab57d9c5aef · outbound
Not All Preferences are What You Need for Post-Training: Selective Alignment Strategy for Preference Optimization write newline
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.