Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-12T03:50:30.002103Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2607.03248.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-12T03:50:30.002103Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
16 of 16 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7685b021-cb95-4531-945e-73576eea0429 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Qwen Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 312ae219-22d4-4818-8bcc-956f263e0ab9 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c0c7a06-27ad-4677-982b-92c9c40f6844 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22df411c-19d2-46b3-ae7c-360413ce0342 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences The Llama 3 Herd of Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 472a3426-3be9-4046-b703-010870a5e3a7 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences H., Ghandeharioun, A., Ferguson, C., Lapedriza, A., Jones, N., Gu, S., and Picard, R
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4575764d-2ac8-457d-b570-17abf0bea412 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challenges
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3da6bf0-28f8-4f54-92b3-d184b6060e54 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences DeepSeek-V3 Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d92ddab-da28-442b-8778-3481b1b90723 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences WebGPT: Browser-assisted question-answering with human feedback
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cffa97fa-0313-4b24-b93e-43b2cb96443b · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Proximal Policy Optimization Algorithms
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88dbfaf9-80ff-4b5d-b8ef-b564808f0418 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dd035b0-6c76-423b-94de-9fbaef639182 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Qwen3 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1470d54a-8ab3-4f29-b8b4-1ffcaa06c8cf · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences SLiC-HF: Sequence Likelihood Calibration with Human Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1f31224-f330-4bc4-904c-092c4b5b66f5 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f847fc5d-9843-4ae7-a7bd-e7d97b8e041c · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77a050f8-c2a6-4c74-a8a0-476e51c04128 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences Proof of Corollary 4.10 Proof
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a23be41-4f6f-4aa9-9774-ae2abc36f540 · outbound
Unbiased Alignment for Large Language Models with Noisy Preferences For reward model training, we train the model for 3 epochs with learning rate 1e-5
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.