Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T11:56:53.066373Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 3 inbound Pith citation observations for arXiv:1908.08342.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T11:56:53.066373Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:02:40.822995Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-10T05:30:23.456663Z
52 of 52 outbound references displayed
External citation measurements
108
pith, observed 2026-08-10T05:30:23.456663Z
Observation 97ef9603-8df4-4bbc-801c-230e41a264fb · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Concrete Problems in AI Safety
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fca53784-e075-4603-95eb-3bc58eb17670 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Roijers, Peter Vamplew, Shimon Whiteson, and Richard Dazeley
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e29e6484-9007-4478-9c02-51aee9c00e77 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Adaptive weighted sum method for multiobjective optimization: a new method for pareto front generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 61f67dff-e0c2-4159-9024-e604e16842b8 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective optimization using genetic algorithms: A tutorial
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 43cac4b4-c847-47c1-b8d9-450d5a4d664a · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Sequential approximate multiobjective optimization using computational intelligence
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 036cb003-056d-4a84-9cf8-0063c8caa124 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation On min-norm and min-max methods of multi-objective optimization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 540d8384-3721-4d13-8ccd-3df0cc2c28e7 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Dynamic preferences in multi-criteria reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4d5784b7-209b-4953-8a97-e3a9a4783d67 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Learning all optimal policies with multiple criteria
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bb2a0f21-afc2-4288-92fb-e3ad7b0d4f41 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-Objective Deep Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b589f4e-4b71-472d-b4bf-63a82c09bc04 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Dynamic Programming
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ec5d1b04-e29e-4e39-88d0-08d17b04ba24 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Hindsight experience replay
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4febe6f2-c1bf-47a5-81a8-5aa9ba6c6ee9 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Modern homotopy methods in optimization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14b194a-4c3d-4535-81df-66472e260cfa · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Roijers, Tom Lenaerts, Ann Nowé, and Denis Steckelmacher
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 04e44619-49c7-43d2-8ed4-6274fe94e091 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Empirical evaluation methods for multiobjective reinforcement learning algorithms
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 13798454-20f3-40d4-a6e7-214c9c52134c · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multiobjective reinforcement learning: A comprehensive overview
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 44155fb4-dc3c-4237-aa59-60f21f1f21c3 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation The steering approach for multi-criteria reinforcement learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f96fc766-e30b-4f3b-9b1e-fda761680b35 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Kephart, David Levine, Freeman L
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 08a171c2-236f-4eef-8574-cf608612f3e9 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Drugan, and Ann Nowé
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b9a95802-83bd-4d67-9ccb-5e896214377f · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective reinforcement learning with continu- ous pareto frontier approximation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6eb876e7-22a6-4ea5-83b9-aa6182406405 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Manifold-based multi-objective policy search with sample reuse
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0fb281ef-9971-4042-8fd3-494e061c276f · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Parallel reinforcement learning for weighted multi-criteria model with adaptive margin
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b8e525ba-1402-4a51-ad84-3124b557be9c · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective reinforcement learning for acquiring all pareto optimal policies simultaneously - method of determining scalarization weights
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2cc3fb91-b85c-4cf9-aa8e-8b917cc3bdc9 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective fitted q-iteration: Pareto frontier approximation in one single run
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 864a7839-e25b-455c-9153-e016bf5a9993 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Tree-based fitted q-iteration for multi- objective markov decision problems
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 771284cd-6819-4f77-917d-1e48b20d2779 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Preference elicitation in combinatorial auctions
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2cbcf60e-c1ad-434b-959c-abd0c90cb950 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation A POMDP formulation of preference elicitation problems
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c9b3baff-61b3-46b7-a1b5-a7fdb790f07d · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Survey of preference elicitation methods
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6fede5a3-f0ff-45cb-978c-c6e5ddde259c · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Ng and Stuart J
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 21513b33-8442-4274-b6a9-613df94bfd04 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 57e6c74e-639e-403b-8c52-d96f417ff1e0 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Generative adversarial imitation learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ece7dc7d-7b41-45dc-8893-9e7480650ac0 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Learning an agent’s utility function by observing behavior
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation da78da7b-030d-47d3-ad0e-9ad67d475697 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2300fa20-8c1d-4fcc-99ee-2f20caac3983 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Linear operator theory in engineering and science
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 44e7c2b4-06b6-49e6-bbc4-2401d8240d7b · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Rusu, Joel Veness, Marc G
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b4642eb0-effd-478e-a960-f3d7e38de072 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Williams
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fcd0fe89-6a1e-40f7-89bc-39484dbb44e5 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d90ebf58-5692-4913-a2fb-54b6e1e2fb07 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Super Mario Bros for OpenAI Gym
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6c051386-6c0e-423d-8262-32f29e3ab58a · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 55ec15d9-1de6-4ba3-824d-74e674ebeb64 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation An introduction to metric spaces and fixed point theory , volume 53
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8b692673-271d-4975-b4ca-314fe904f370 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Bertsekas
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 09e96095-5683-4f31-869a-6b6fd7bfc207 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Abstract dynamic programming
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ae4ed655-411f-4a5f-b5a4-2ddab26ad216 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Bellemare, Will Dabney, and Rémi Munos
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 53faeaeb-8ad9-41f2-9e1a-5c9ffb3c2dbd · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation abc54813-b7b9-4f31-a20c-61254d0d0439 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9685e182-42af-4feb-aadb-17ca3c16f4de · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Continuous control with deep reinforcement learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44b7b07d-b4bb-44df-a302-5c30148d8cb7 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Lillicrap, Ilya Sutskever, and Sergey Levine
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 96f47eb8-f3d4-48f7-9303-ceffb0850709 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Discrete Sequential Prediction of Continuous Actions for Deep RL
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fb1e13f-327b-4261-be39-20ab4bf3c5d1 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Deep reinforcement learning with double q-learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6af8a5c1-7d74-4831-bcee-da4373d59237 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Prioritized Experience Replay
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09fe7a03-3e45-4254-98c2-4a39f857887c · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2472494a-0850-477a-9351-a95d7de76bca · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b9507c95-d276-4b46-82c7-954cd185ffc5 · outbound
A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation distance
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 63bef60f-fe8e-48a0-a249-9dda6b632686 · inbound
Multi-Objective Reinforcement Learning for Automated Resilient Cyber Defence A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c03d6e4-220e-41eb-b9ad-f2e3900461cb · inbound
Reinforcement Learning for Multi-Objective Multi-Echelon Supply Chain Optimisation A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 77951548-8b8f-4916-baf2-c4370b614546 · inbound
Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.