Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:04:00.844121Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2506.00962.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:04:00.844121Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c17865b5-f16c-417e-a61e-14af8991c27a · outbound
Reinforcement Learning with Random Time Horizons M., and Sun, W
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3d79a602-1722-42b5-8dcc-109581632a6f · outbound
Reinforcement Learning with Random Time Horizons An optimal control perspective on diffusion-based generative modeling
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11848eba-7792-4a3a-8349-4313dd1673b4 · outbound
Reinforcement Learning with Random Time Horizons Steady state analysis of episodic reinforcement learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d63d11e-ac53-4719-9996-b3efa119e559 · outbound
Reinforcement Learning with Random Time Horizons Finite-Sample Analysis of the Monte Carlo Exploring Starts Algorithm for Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6def04ef-1c65-4891-bbd3-dd906e61c7cd · outbound
Reinforcement Learning with Random Time Horizons Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb57821f-0633-4c4a-bde4-23972733b118 · outbound
Reinforcement Learning with Random Time Horizons Finite state Markovian decision processes
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1d87063d-636e-4e13-9170-a787eb8f0b5f · outbound
Reinforcement Learning with Random Time Horizons and Sch \"u tte, C
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0438410c-f251-4002-a688-757897fd52a0 · outbound
Reinforcement Learning with Random Time Horizons Variational characterization of free energy: Theory and algorithms
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d7fc37ce-0105-45ca-a973-6348e1cdcb16 · outbound
Reinforcement Learning with Random Time Horizons Fr\'{e}chet derivatives of expected functionals of solutions to stochastic differential equations
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 40fdb2c6-5c4c-4d16-97fb-d3442c12ae7c · outbound
Reinforcement Learning with Random Time Horizons Continuous control with deep reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e76d062e-9886-463a-884d-8ed878b4972b · outbound
Reinforcement Learning with Random Time Horizons Online reinforcement learning with uncertain episode lengths
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff03ed89-d27d-491b-a17e-dde9b5b8bc02 · outbound
Reinforcement Learning with Random Time Horizons Playing Atari with Deep Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb4803f5-2282-4267-bdea-77254ea5029f · outbound
Reinforcement Learning with Random Time Horizons and Thomas, P
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f83b191f-1f2c-4bbc-aedc-fded12b0ad2d · outbound
Reinforcement Learning with Random Time Horizons and Richter, L
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 285f59f2-a65c-4643-9515-7d892ef44a37 · outbound
Reinforcement Learning with Random Time Horizons Stochastic control foundations of autonomous behavior
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fea4c62f-45d4-416b-aed7-d4b9cf0e2e99 · outbound
Reinforcement Learning with Random Time Horizons Continuous-time stochastic control and optimization with financial applications, volume 61
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 620359c4-2a76-4563-93ca-4493dc8d1dee · outbound
Reinforcement Learning with Random Time Horizons Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef440227-290e-422e-bb6f-112ee4e62c7e · outbound
Reinforcement Learning with Random Time Horizons and Ribera Borrell, E
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ea9911ac-36ec-477e-8832-9d4d67f16afd · outbound
Reinforcement Learning with Random Time Horizons Improving control based importance sampling strategies for metastable diffusions via adapted metadynamics
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f08cd45a-5b85-4937-8c03-6786bf0f78e3 · outbound
Reinforcement Learning with Random Time Horizons Trust region policy optimization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 48911a4b-217a-475a-967a-300416e5dee8 · outbound
Reinforcement Learning with Random Time Horizons Proximal Policy Optimization Algorithms
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26949471-6f45-4a5c-b5a4-1cd2faf1e4a7 · outbound
Reinforcement Learning with Random Time Horizons Overcoming the timescale barrier in molecular dynamics: Transfer operators, variational principles and machine learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0d10f094-e17d-4716-a164-1eab4003f514 · outbound
Reinforcement Learning with Random Time Horizons Deterministic policy gradient algorithms
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 512c57e3-5cc5-4688-a4b3-e7b465b16bc0 · outbound
Reinforcement Learning with Random Time Horizons Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7646e382-a65e-4fb0-aff2-83722a1e2c91 · outbound
Reinforcement Learning with Random Time Horizons Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11121e87-018f-471b-a95d-7e2d24782b50 · outbound
Reinforcement Learning with Random Time Horizons S., McAllester, D., Singh, S., and Mansour, Y
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08383ad2-08e0-4ec9-99b2-ab3491473147 · outbound
Reinforcement Learning with Random Time Horizons U., De Cola, G., Deleu, T., Goulão, M., Kallinteris, A., Krimmel, M., KG, A., Perez-Vicente, R., Pierré, A., Schulhoff, S., Tai, J
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 49aa1aca-32ea-4a76-8e4a-7c88dab3e0fc · outbound
Reinforcement Learning with Random Time Horizons Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eb154fd5-b6ac-41cd-bb99-d20188244c60 · outbound
Reinforcement Learning with Random Time Horizons Unifying task specification in reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 066d052c-99c8-4bcb-a21f-1bb88685fb30 · outbound
Reinforcement Learning with Random Time Horizons Global convergence of policy gradient methods to (almost) locally optimal policies
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53440e3e-290b-41e0-bdcc-7bd52c7f6e3f · outbound
Reinforcement Learning with Random Time Horizons Actor-critic method for high dimensional static H amilton-- J acobi-- B ellman partial differential equations based on neural networks
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.