Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:39:31.713492Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 100 of 297 outbound references and 0 inbound Pith citation observations for arXiv:2607.29559.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:39:31.713492Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 297 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5bee8ff9-5c58-4934-b338-19479b3eb5f2 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Structure and Interpretation of Computer Programs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b30f37c-da2c-481c-bfa0-4e9f6ecc15f4 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Visual Information Extraction with Lixto
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f188bd46-0683-490f-8df6-c71ba6d0e7b3 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Brachman and James G
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4b422d4-937d-47da-9294-400f8b6dd384 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Complexity results for nonmonotonic logics
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1791ea39-2bd3-4a9a-845a-4e9d3e1396b7 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback International conference on machine learning , pages=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2b936d4-1aac-468b-ad57-d588e4de3995 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Hypertree Decompositions and Tractable Queries
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 874628da-b357-4d2b-b021-20f4df54aff4 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Levesque
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 718dda53-a904-439c-8cec-b0e522001c5b · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Levesque
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5d6fc0-6329-4d30-ac5b-612dbc22a701 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback On the compilability and expressive power of propositional planning formalisms
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c012139-7e05-4380-8366-8b1ecf6f9a68 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6857fca7-2dc5-41bd-ac20-fec6fefca1b9 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback The Knowledge Engineering Review , volume =
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5145e282-dee6-4f7a-aaf6-794fbe7f9b2d · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Artificial Intelligence , volume =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28349626-1ec0-4800-bc66-552eef11895c · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Logics of programs: axiomatics and descriptive power
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00802faa-6e1c-4d96-ab0b-ac437ee5a7a2 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Clarkson
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a2c6029-d150-4ed9-b5f9-13ef503b7f71 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback A More Perfect Union
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa38e6a1-b867-410e-96e5-aaf08eedfcb9 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback The fountain of youth
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a46ff4-018c-4ca9-8cfc-eb50925f6c5c · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b35fef3-5325-4b20-855b-3503a655c8b2 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Proceedings of the 20th International Colloquium on Automata, Languages and Programming , series =
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0696e14-2f17-45cd-9113-515c1ccab64e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ca4f76-3181-44dc-9806-784ed28f8ec7 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Anisi , title =
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8027626-9377-4990-a60d-30ef340c0f38 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Journal of Artificial Intelligence Research , author =
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6cf07bab-36b6-42d1-aaa7-7cf783553107 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f9dd01d-2f57-4723-82e6-86a18ceda1b4 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Reinforcement Learning , author =
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5295577c-5a5c-4fd0-a3fd-26657c5e5361 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98c2686b-c84a-4893-9817-5c9a2595a88e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6b68dfec-b62d-4cd2-a103-c5f50a2c0958 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Divide and Conquer: Provably Unveiling the Pareto Front with Multi-Objective Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bbc5ba6b-ab98-4543-aa82-cbf85e72c605 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Fairness in Preference-based Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7311e445-1e27-4c5f-ba38-dddb4f2487ce · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4467589-f136-4aa9-88bd-8a31f7fa049c · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Advances in Neural Information Processing Systems , volume=
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9295a79a-5eeb-4002-8212-51dcb5f837c8 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Promptable Behaviors: Personalizing Multi-Objective Rewards from Human Preferences
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 14452a8c-e7fb-4342-8515-23dc6816b271 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Proceedings of the 41st International Conference on Machine Learning , pages =
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 996524c1-8261-4f57-9c33-36780dec384a · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Fine-tuning language models to find agreement among humans with diverse preferences
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bd86fb5-4b35-472d-a65a-1b88e422cdaa · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b756a39e-1aa6-4116-b690-3c3671133fb9 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a6c6a0f-78f6-4c38-84fa-8fdfeed44829 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback 2024 , pages =
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e0eed1b-71b7-41d5-b112-ebdecbe953e8 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8994e582-dfc2-45a5-bb47-dd15029abe7b · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Yang, Diyi and Vosoughi, Soroush , month = oct, year =
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dc6aa27-f5cb-4155-ad93-1d27d700f73e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f051224-659f-489d-b2b7-27447a6814e0 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Hassenzahl, Marc , month = jul, year =
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 766bef44-b4e9-43b8-849b-f56bf7b7ddce · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback IEEE Access , author =
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d63e0ac1-be9c-4eed-8225-f0d62014ec83 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Advice to
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 763a2c3b-bcc3-4a93-b075-69576f7a0dc9 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Applied Intelligence , author =
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eb6dbf2-6d28-47ca-992f-176b78bfe290 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Frontiers in Computer Science , author =
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f962405-36d3-43db-af86-be6858457f23 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback International Journal of Social Robotics , author =
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e68047-3b0c-4162-a5c2-ee5573d5470f · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Analysis of
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08808c1b-9707-47da-bdc9-b8130d669f69 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Multimodal Technologies and Interaction , author =
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb74602b-fb75-4018-a1c1-72a8ffa45802 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Trends in Cognitive Sciences , author =
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c759fdb9-7b23-49b2-ba9f-5eaa6cf3527d · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback iScience , author =
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50facedf-7392-42dc-814d-530e0afdf03e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Journal of Artificial Intelligence Research , author =
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63418822-c77a-42ba-a1df-f2bd9c35d25e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Transformers are
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2e3a93c-482e-4e7c-a128-769964af40e0 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Evaluating Agents using Social Choice Theory
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 013cc6e2-fc5f-43b4-b00a-c9d1ca894a43 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Developing, Evaluating and Scaling Learning Agents in Multi-Agent Environments
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 82229c54-26b0-4761-9832-e67686709713 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Autonomous Agents and Multi-Agent Systems , author =
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1d8f45c5-c3be-4bd4-aee9-03c779c17589 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Journal of Artificial Intelligence Research , author =
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a9e4b278-bcf0-49c4-b9de-dde5656552a6 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Journal of Artificial Intelligence Research , author =
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31449dda-d39a-46d5-b32e-a85bc33eebba · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Neural Architecture Search: Insights from 1000 Papers
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5f807f2-f37b-4114-ab1a-3798a692f2aa · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Barto, Andrew , year =
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19a06b44-d4d5-48c1-8b96-26123e292310 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Barto, Andrew G
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9894208b-911b-4749-8fe2-e5180db8bb08 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Royal Society Open Science , author =
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 21b12909-a796-4a19-a9a3-5e000d6edab9 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback A Review of Cooperation in Multi-agent Learning
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63bd049d-e006-4ea7-84f5-e410c8be53f5 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback , month = oct, year =
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6498bee1-836a-4139-ad8a-e68c6e9b103c · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Mediated
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c2cf02a-971c-473f-8edf-f60a5a12e900 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Autonomous Agents and Multi-Agent Systems , author =
Reference 65
Source-reported events for the cited work
correction dated 2024-06-07. Source: crossref record 10.1007/s10458-024-09654-9->10.1007/s10458-024-09649-6:correction, observed 2026-07-11T03:00:36.460696+00:00. This notice travels one citation hop only.
Observation 9817d800-dc4a-4ad2-8ebc-132e96241b54 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Cognition , author =
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2eee9dd-060e-4613-97fd-9fdbe0974ca5 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Everett, Richard and Weidinger, Laura and Isaac, William S
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ba10fd-b2c5-4b05-b302-883f468d5c08 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Cooperative
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ee80319-2cc3-4ffc-9cbd-c32069438a0a · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback IEEE Robotics and Automation Letters , author =
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2141b031-3153-4c47-9e74-12baf8384c26 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Current Robotics Reports , author =
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 944e8b32-f415-499c-ab9b-dde16231b2a5 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Social Cognitive and Affective Neuroscience , author =
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0a0f1814-5932-41ea-8f1e-0aee02a71c29 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Trends in Cognitive Sciences , author =
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a6865620-a6b8-4435-8907-a69076c020d0 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Computational Intelligence and Neuroscience , author =
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 55af0887-b944-430a-82f4-8788f7724e23 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Adapting a
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc90fbd5-be26-4c4f-8889-a23294a9334e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback IEEE Transactions on Affective Computing , author =
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1fd4bd7-c2b8-46bc-b200-c7178fe8802a · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3abd1ed-5039-4116-9bfd-306a0b9298a2 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Metin and Yemez, Yucel , month = aug, year =
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e52079-457f-4d75-b913-1c032f5ac3ab · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Metin and Yemez, Yucel , month = aug, year =
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97dbda7e-315a-49f6-8356-7dd4b3812788 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e32b8c5-426e-40cc-91f6-084daaee667c · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Human-to-
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eee5a56b-dfad-48f7-a4fd-52b13a7b474e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback IEEE Robotics and Automation Letters , author =
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80d2053f-f062-430e-813e-56aed56d01c7 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Efficient
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2362b2b-3b77-45c1-96b3-14db60f9121c · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Modeling
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 398d1d3a-0f5b-4192-8465-e7616806cbf7 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Machine Learning , author =
Reference 84
Source-reported events for the cited work
correction dated 2022-12-28. Source: crossref record 10.1007/s10994-022-06298-2->10.1007/s10994-022-06273-x:correction, observed 2026-07-11T02:56:29.207212+00:00. This notice travels one citation hop only.
Observation aac3b22b-e924-4590-a1fa-6a6bdd1207d3 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfc7ef3b-3178-45c8-924a-2f082fa68de0 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Emotional
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f42a059-110c-4e68-a3ef-896b353f0183 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Islam, Usman and Willis, Richard and Sunehag, Peter , month = dec, year =
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58d7e322-5396-48c8-a347-3319934c83fd · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Gemp, Ian and McWilliams, Brian and Duéñez-Guzmán, Edgar A
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c916df2-d1b0-4243-9d6e-f782a21519ef · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10f9f5b4-8f97-4e32-b704-623d55bcffed · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57260140-fee5-4958-bad2-e8a50ee4028e · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 480e41c0-4ef8-4e12-acaa-0ace24f19522 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Autonomous Agents and Multi-Agent Systems , author =
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5a6ac41-2cdf-45aa-890d-506b7065f227 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Bradley and Sadigh, Dorsa , month = apr, year =
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ffd6a80-e9ba-40a8-8103-9a3c3a09fd14 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback and Fidler, Sanja and Torralba, Antonio , month = may, year =
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99fc0100-028f-4dc4-b98e-931f6000a011 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Promptable
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86f42e63-cc98-44bb-a265-3c4581a50234 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55496758-f7d2-46bc-a06f-36b1c5706af4 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a0f04da-3b3b-41a8-b155-4ce8e945408b · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3936488-a1cb-4a90-912f-d7bf327a38e6 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Unresolved cited work
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e997296-835a-4343-b4d5-2f0a78e07893 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Cognitive Systems Research , author =
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0f55f6cf-6310-4e71-bc2c-05897c0bd980 · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Artificial Intelligence Review , author =
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f746c6a-e146-45f4-8560-67a2c894362d · outbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback ACM Transactions on Intelligent Systems and Technology , author =
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
No inbound Pith citation observations are available.