Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:37:31.607767Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 1 inbound Pith citation observation for arXiv:2412.10917.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:37:31.607767Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T00:30:33.390291Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-12T00:30:33.977487Z
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e8b4a402-f74d-4c99-b096-e5f8f7c13658 · outbound
Adaptive Reward Design for Reinforcement Learning Control synthesis from linear temporal logic specifications using model-free reinforcement learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5005d4f1-7527-460b-a175-e0e51bcb431c · outbound
Adaptive Reward Design for Reinforcement Learning OpenAI Gym
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50838048-cec8-4079-b4fe-6631b96abc7d · outbound
Adaptive Reward Design for Reinforcement Learning Overcoming exploration: Deep reinforcement learning for continuous control in cluttered environments from temporal logic specifications
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6028a52c-44e4-4ea1-9500-054d80fe0e5a · outbound
Adaptive Reward Design for Reinforcement Learning Learning minimally-violating continuous control for infeasible linear temporal logic specifications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a6937f62-cfca-4933-a5f8-f260ab2aaa7d · outbound
Adaptive Reward Design for Reinforcement Learning Ltl and beyond: Formal languages for reward function specification in reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ccd9e90b-0801-42ff-9ef7-04b9a0d11362 · outbound
Adaptive Reward Design for Reinforcement Learning Foundations for restraining bolts: Reinforcement learning with ltlf/ldlf restraining specifications
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3560acac-53cd-4c1e-93b4-015dc9c65932 · outbound
Adaptive Reward Design for Reinforcement Learning From language to goals: Inverse reinforcement learning for vision-based instruction following
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4c33d730-ac7d-4357-8f29-f7408f49d443 · outbound
Adaptive Reward Design for Reinforcement Learning Reinforcement learning for temporal logic control synthesis with probabilistic satisfaction guarantees
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 28d912a4-17f9-4163-96ae-cfc3421f9ada · outbound
Adaptive Reward Design for Reinforcement Learning Deep reinforcement learning with temporal logics
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6c8af446-f2ec-41a3-947d-e3028e860199 · outbound
Adaptive Reward Design for Reinforcement Learning Reward machines: Exploiting reward function structure in reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6a025cf4-3809-4f04-952d-b6f39dbbf243 · outbound
Adaptive Reward Design for Reinforcement Learning Temporal-logic-based reward shaping for continuing reinforcement learning tasks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 55ee23ac-cc8d-4ced-833f-b07335c6ef3b · outbound
Adaptive Reward Design for Reinforcement Learning A composable specification language for reinforcement learning tasks
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2e1da156-c139-4568-972f-3140e4b28b46 · outbound
Adaptive Reward Design for Reinforcement Learning Compositional reinforcement learning from logical specifications
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 990b3a30-54ab-490c-8476-c80992c7bf68 · outbound
Adaptive Reward Design for Reinforcement Learning Model checking of safety properties
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c8d86e92-4313-4978-9209-df85a676621b · outbound
Adaptive Reward Design for Reinforcement Learning Probabilistic planning with formal performance guarantees for mobile service robots
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e8bef7a2-3962-4188-bddf-15a9f20ee938 · outbound
Adaptive Reward Design for Reinforcement Learning Reinforcement learning with temporal logic rewards
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4c2b5c61-fb91-46b8-9fda-e4254a406a0f · outbound
Adaptive Reward Design for Reinforcement Learning Continuous control with deep reinforcement learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8fad458-da6e-45ce-b701-5681eb729509 · outbound
Adaptive Reward Design for Reinforcement Learning Human-level control through deep reinforcement learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8b3ba6-fda5-4c02-a686-379596ee1bf7 · outbound
Adaptive Reward Design for Reinforcement Learning Asynchronous methods for deep reinforcement learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c0e50507-434f-4da3-b512-3c9c80236a2b · outbound
Adaptive Reward Design for Reinforcement Learning Algorithms for inverse reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 08ec08ef-dc8d-42a8-80d9-42aabc55c140 · outbound
Adaptive Reward Design for Reinforcement Learning Policy invariance under reward transformations: Theory and application to reward shaping
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 38404a0c-da5b-459f-87bb-38bf87f9a166 · outbound
Adaptive Reward Design for Reinforcement Learning The temporal semantics of concurrent programs
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b79e5b12-0680-40e0-9c64-691dae414074 · outbound
Adaptive Reward Design for Reinforcement Learning Stable-baselines3: Reliable reinforcement learning implementations
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 74e00a7c-6247-4e1e-9cee-cc5afc606738 · outbound
Adaptive Reward Design for Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699c7cc4-085b-4d80-aa89-8ded6c4ed214 · outbound
Adaptive Reward Design for Reinforcement Learning Deep reinforcement learning with double q-learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f1af1b0f-3017-46e2-9fee-6bf33ef907e7 · outbound
Adaptive Reward Design for Reinforcement Learning A survey of preference-based reinforcement learning methods
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7def0c9b-206f-4f04-bbc5-6b2628a79826 · inbound
A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning Adaptive Reward Design for Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.