Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:59:01.287449Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2604.02091.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:59:01.287449Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e319dbd0-c789-41db-bc1c-27ef7513cfc1 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d64b3c60-9a7d-49f3-ae14-76661596701f · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91744e54-eac7-42f6-9732-e298f256bb8f · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Attention in Large Language Models Yields Efficient Zero-Shot Re-Rankers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8db992d-2cfa-45c7-a029-a14ef9c750f6 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db6996dc-f1c9-4f85-91ba-6b71ff16808e · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2181c2b-5f0b-4fdb-b041-619478788764 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d32d157-23e2-4cac-9cd5-e7149cc2f6b1 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Retrieval-Augmented Generation for Large Language Models: A Survey
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aaa5580-a00e-4da4-9507-541ad90ce0bc · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning DeepRAG: Thinking to Retrieve Step by Step for Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2631c07a-d806-4ac4-a165-0921c27eab26 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fc202cb-85db-4836-96c9-7cbe4707b7b6 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dd75243-53c7-4e5b-b69b-6975849bc3b5 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Leveraging Passage Retrieval with Generative Models for Open Domain Question Answering
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffb2735e-d440-47bf-b948-c8d9734e8cd0 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66bf8b90-d21b-4c01-96a1-555573eadf98 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Adaptive-RAG: Learning to Adapt Retrieval-Augmented Large Language Models through Question Complexity
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57fb572e-0ade-40f1-816e-2b929b1b04d1 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 521d7dc9-1e27-43e7-9b4e-769c3123052e · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2bf93c0-5b4f-4fb6-a5ab-ed82c5e3b010 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e44e9ec-47e5-49ce-a5be-7cbe6e32a579 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f16c8cb4-a615-4083-a195-b28e8a44f26d · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1e5531d-bd7e-4d62-94ec-5dc2274fa37f · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Gonzalez, Hao Zhang, and Ion Stoica
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3259635c-63d2-4752-a955-e0eda334a52f · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 98d2654a-bef3-4f9b-8e0b-6b9a6bb4e4c9 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt\
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebb46ea2-80a8-420e-84cd-430ba0764a6c · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e371d8d8-0baa-409c-b2a1-d63aa8fb4861 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e2699b-0ae9-4ccd-aa75-7094988120bc · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 531003a4-cd1b-4e7e-b626-afd8c859b197 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0284e23-730a-41ba-b72c-ad9ed24d3480 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning AmbigQA: Answering Ambiguous Open-domain Questions
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5405dbae-be00-47ae-8d3f-70f46d10a6bb · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdc17a1c-5daa-4bfc-9034-c5d2ced7933a · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning WebGPT: Browser-assisted question-answering with human feedback
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac505bf9-f590-4b74-bf48-8f4cc96274ec · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Passage Re-ranking with BERT
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d073a16f-e56f-4ac4-9403-90d940f5c569 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Multi-Stage Document Ranking with BERT
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0a56791-c706-46fb-9e00-47c56d7b762a · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc6cbe7-d292-40c1-b696-2169ba8a52e1 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning RankZephyr: Effective and Robust Zero-Shot Listwise Reranking is a Breeze!
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02c0f16a-ee4a-4503-8588-f657852e92d7 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f7e9cc9-11f8-4523-99e4-1e16235a9300 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning FIRST: Faster Improved Listwise Reranking with Single Token Decoding
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4d113d1-f455-44fa-9f62-b9ee1c9951e2 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c898c14-49db-4b9f-a007-3f6ebbba6adf · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 830f22c7-3991-4bf5-b515-0992798518c1 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea2c8636-a071-4899-90b7-3e6e9d9c83b1 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Retrieval Augmentation Reduces Hallucination in Conversation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e646110-e67d-49df-b641-55a4bea42bd9 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5c6cba6-58a2-4529-a767-ee34a853a36c · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning DynamicRAG: Leveraging Outputs of Large Language Model as Feedback for Dynamic Reranking in Retrieval-Augmented Generation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d80fef6f-11d2-40af-894b-de80f63b820d · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking Agents
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f6c0135-5b84-4663-8c3f-8bd3a42d0358 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Qwen2 Technical Report
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4875af82-6871-46ed-b30f-54e2222d9230 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 778fc0f8-4e93-4bec-ade4-4b16e201a1d8 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5323cb47-3a6a-45ac-89b6-538237a64e13 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75fd3e11-57cf-4541-b5d2-b577384457e7 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning PromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt Optimization
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4954a0d-e10d-4e80-8280-6129688e4989 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning C-Pack: Packed Resources For General Chinese Embeddings
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5475d41e-0df9-47d3-b6e8-2ba75bcf861b · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7776387-9543-4468-a19e-5fa2955252a2 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4db4453b-27ea-4b04-b281-ffbc8d386f01 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f89c5535-6543-42aa-be52-24ed84fa887e · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd0f40a5-8536-44cf-a258-89ff43212f08 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e5b552c-b100-400b-a8f4-79f2f8db55a4 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning REARANK: Reasoning Re-ranking Agent via Reinforcement Learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86ab9ee2-68d9-4402-a0e9-ca2c83ba7f37 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0e94330a-cf02-47e1-baa5-9192a790102e · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c5d59e8-6cc1-41d4-8909-9c5cacc56aa0 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 954e8775-b5a0-49a2-ae9d-60be74f45334 · outbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning INTERS: Unlocking the Power of Large Language Models in Search with Instruction Tuning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.