Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:08:52.401981Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2502.03699.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:08:52.401981Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:26:50.289327Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T15:26:55.832369Z
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 76fe5319-e7fa-4d24-af88-a2e61e644c49 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d84378-cb7f-4400-888a-ea88aea6d08b · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective w. current
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 54c53931-f9ee-4c86-9d6a-57b05da0f170 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective As the number (N) of retrieved responses increases, the retrieval recall increases
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 42601634-a9db-45d0-a53f-55ca2c825252 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective in-batch negatives
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8e4bfde4-bb0d-4683-a010-175cf8146e34 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective For AlpacaEval2, we report the result with both opensource LLM evaluator alpaca eval llama3 70b fn and GPT4 evaluator alpaca eval gpt4 turbo fn
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c1f5c3d6-53c7-4d0b-b26c-70b1f98de59e · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f899329-eeeb-4975-a10f-9638b1848321 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective KTO: Model Alignment as Prospect Theoretic Optimization
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c124081d-487d-4d44-9c0f-09880e7e44e4 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Direct Language Model Alignment from Online AI Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 863b84ed-5a22-4b9b-8a23-d8a3220b1db1 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Poly-encoders: Transformer Architectures and Pre-training Strategies for Fast and Accurate Multi-sentence Scoring
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb7a7fe1-8661-4954-9ac5-d6a5b9cc41c6 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Mistral 7B
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba1bd6fa-dce3-4982-a340-cdfec74cdcc9 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Dense Passage Retrieval for Open-Domain Question Answering
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 385d2405-60e4-462c-b6a5-1595e6ad2afd · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective RewardBench: Evaluating Reward Models for Language Modeling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 829c2c00-7db5-49f3-a51a-747a545cc2d7 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective From Matching to Generation: A Survey on Generative Information Retrieval
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bde4c4a1-91de-42ed-ac10-3695cbf60e06 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective SimPO: Simple Preference Optimization with a Reference-Free Reward
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ed1b2e7-df5c-42be-a1f5-9480b05bf95e · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Passage Re-ranking with BERT
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d353fc4-1597-4cc9-9813-e0681c393432 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Document Ranking with a Pretrained Sequence-to-Sequence Model
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec8409b5-0173-49ee-9f53-106c5931d72c · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Disentangling Length from Quality in Direct Preference Optimization
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 855c5902-532f-4aab-85f1-7ec604c04db5 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective RocketQA: An Optimized Training Approach to Dense Passage Retrieval for Open-Domain Question Answering
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caeada24-616c-401f-9508-afbefbb737ad · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Proximal Policy Optimization Algorithms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 780ebc08-f491-4ed6-803c-edc08895bff9 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective BOND: Aligning LLMs with Best-of-N Distillation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10b2bfe1-2e97-43ff-84ea-926ff5b48a99 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f63c5adc-abad-494f-8605-abccd89c4418 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 818504ca-51be-4dbd-bcb2-03e2a091b13b · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Aligning Large Language Models with Human: A Survey
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f7339df-a957-4429-9187-a0870b3f8513 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2968a6cf-ff46-4227-a412-69d962d63f3f · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective RRHF: Rank Responses to Align Language Models with Human Feedback without tears
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d911d71a-bd4d-4443-9239-5ab309c3d01e · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Curriculum learning for dense retrieval distillation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ff55ae22-d549-4c47-bb36-1e8685b45293 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective SLiC-HF: Sequence Likelihood Calibration with Human Feedback
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d544be6-869f-4a5c-a40a-fa75e9e1658a · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective We consider two base models: Mistral-7b-base and Mistral-7b-it
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 74668662-f023-4d21-8c85-9e044b489e73 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective We generate 4/6/8/10 responses with the LLM and score the responses with the off-the-shelf reward model (Dong et al., 2024)
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d11b961d-6358-4b2e-bf97-52626eef8fbe · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
Reference 1952
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c721f883-1ae1-43dc-9f9e-aa39314a6c33 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdb5b3ed-d9ff-4cd1-9f16-1847d43020e8 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Evaluating Large Language Models Trained on Code
Reference 2010
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d32c006a-4cd4-4c32-9c59-53719ade133b · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Training Verifiers to Solve Math Word Problems
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7679d1e-e52f-465b-9e9c-008b54927d1f · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 813c0465-420d-40b7-b806-df77bd1c0e38 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Representation Learning with Contrastive Predictive Coding
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3d3ef6e-1b92-4c03-a865-6301d853366a · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective A Survey of Large Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b1e524f-e854-4e74-b820-d5a26e47e936 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective LiPO: Listwise Preference Optimization through Learning-to-Rank
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f51955d9-5e2b-4572-bf71-eefcf0854e34 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective RLHF Workflow: From Reward Modeling to Online RLHF
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a89aed-cdc5-4984-ab27-839566a6f2c2 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7120ae5c-e864-4839-b264-e5eee69331c8 · outbound
LLM Alignment as Retriever Optimization: An Information Retrieval Perspective MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717a137d-4795-4d92-aa45-2f3bd6b12ce5 · inbound
An Empirical Study on Reinforcement Learning for Reasoning-Search Interleaved LLM Agents LLM Alignment as Retriever Optimization: An Information Retrieval Perspective
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.