Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T14:35:04.201252Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2509.20869.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T14:35:04.201252Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 406c6330-e16c-4254-bc3d-d4e3e05296a3 · outbound
Model-Based Reinforcement Learning under Random Observation Delays write newline
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 72534a24-05bf-4b51-8beb-8a1f3a953b5f · outbound
Model-Based Reinforcement Learning under Random Observation Delays A cerebellar-based solution to the nondeterministic time delay problem in robotic control
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e01764b1-1e9a-4854-9f58-fb98943a31bb · outbound
Model-Based Reinforcement Learning under Random Observation Delays Closed-loop control with delayed information
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ef87f7b6-4d46-48e4-b0af-fdcf08c586fb · outbound
Model-Based Reinforcement Learning under Random Observation Delays Update with out-of-sequence measurements in tracking: exact solution
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7bc3da22-b67c-43f6-8632-964cb02ecbcb · outbound
Model-Based Reinforcement Learning under Random Observation Delays Reinforcement learning with random delays
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b381dd62-2fac-4a1d-9bd0-317180f5b5bc · outbound
Model-Based Reinforcement Learning under Random Observation Delays Bayesian filtering: From kalman filters to particle filters, and beyond
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fa3ed8f9-681a-4d5f-b352-ee0340607649 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Acting in Delayed Environments with Non-Stationary Markov Policies
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8be2a045-383f-418a-b260-d0a7cfb37400 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Communication delay in uav missions: A controller gain analysis to improve flight stability
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4e40442a-2c12-40dc-b7c9-dada1b99b725 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Recurrent world models facilitate policy evolution
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7bbcee7c-c512-4c9e-b4d7-ff8d076b055b · outbound
Model-Based Reinforcement Learning under Random Observation Delays Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fa203662-7aeb-4f51-b27d-28624a6e409c · outbound
Model-Based Reinforcement Learning under Random Observation Delays Learning latent dynamics for planning from pixels
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0b8b7500-655a-40b8-a17b-88c3cde4c777 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Mastering Atari with Discrete World Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c1e3cbb5-a00d-4b7b-9008-c9d334e38497 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Mastering Diverse Domains through World Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 56d1e66a-7a89-4d5f-b2e1-fd29dd86f723 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Mastering diverse control tasks through world models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b0bb3a76-b285-49d5-9221-c98dfe250bd5 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Temporal Difference Learning for Model Predictive Control
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2a7b6924-cb74-4409-8502-f3f51bf66c04 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Deep variational reinforcement learning for pomdps
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 84e91d7f-f57f-4b85-a5d4-51d70d9e83c3 · outbound
Model-Based Reinforcement Learning under Random Observation Delays When to trust your model: Model-based policy optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d05b673c-0738-4295-b6e8-4cbec606981b · outbound
Model-Based Reinforcement Learning under Random Observation Delays Planning and acting in partially observable stochastic domains
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 483fdeb5-9396-4933-90e4-86a7b701633c · outbound
Model-Based Reinforcement Learning under Random Observation Delays Reinforcement Learning from Delayed Observations via World Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ed57e875-7fdb-474e-a745-8451d699d7a0 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Markov decision processes with delays and asynchronous cost collection
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 60b78f21-3d45-49f6-a1a0-919c31b3491b · outbound
Model-Based Reinforcement Learning under Random Observation Delays Belief projection-based reinforcement learning for environments with delayed feedback
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9c8707cc-49fd-46a9-9bd7-ebfc7d22173d · outbound
Model-Based Reinforcement Learning under Random Observation Delays A partially observable markov decision process with lagged information
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b9cdf566-2d78-41cc-b6b1-aed62d3fca77 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Stochastic latent actor-critic: Deep reinforcement learning with a latent variable model
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fe66291f-0461-4119-825e-f72ce8efdeba · outbound
Model-Based Reinforcement Learning under Random Observation Delays Learning a belief representation for delayed reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ec3815fe-7798-4bc1-984f-757f1ab4db28 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Delayed reinforcement learning by imitation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bc6f3d72-ec48-4834-80fe-b06233236230 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Particle filter recurrent neural networks
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2b80e932-e12c-4774-8eac-8d5495f6be05 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Discriminative Particle Filter Reinforcement Learning for Complex Partial Observations
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 525a8490-4bcf-464d-89dd-97fffccd9277 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Setting up a reinforcement learning task with a real-world robot
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bd448093-7fd6-4800-8892-8cf16a760a76 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Transformers are Sample-Efficient World Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9c266cee-3908-4181-af7d-56d628e7cb3a · outbound
Model-Based Reinforcement Learning under Random Observation Delays Control delay in reinforcement learning for real-time dynamic systems: A memoryless approach
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e38cf315-f0b9-4ad4-a4d1-91e9bcfd6fcc · outbound
Model-Based Reinforcement Learning under Random Observation Delays Mujoco: A physics engine for model-based control
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 47ac839e-fa66-4f7c-9f3a-b135ebf7259a · outbound
Model-Based Reinforcement Learning under Random Observation Delays Tree Search-Based Policy Optimization under Stochastic Execution Delay
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5860a670-62eb-4c3a-95b7-e0412edd8011 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Planning and learning in environments with delayed feedback
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9bed5266-616d-44c8-aa8f-a47cb8b664df · outbound
Model-Based Reinforcement Learning under Random Observation Delays Addressing signal delay in deep reinforcement learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1d332628-6ea9-4567-8e3d-35524f4700cd · outbound
Model-Based Reinforcement Learning under Random Observation Delays Variational delayed policy optimization
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a142d5f8-0ee1-4fae-a18e-213b632459bf · outbound
Model-Based Reinforcement Learning under Random Observation Delays Boosting Reinforcement Learning with Strongly Delayed Feedback Through Auxiliary Short Delays
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b6fcf3af-f5b0-4693-8cf9-56201cff736d · outbound
Model-Based Reinforcement Learning under Random Observation Delays Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 972d2d93-31b7-41dc-ab81-7e9785bf1c06 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Storm: Efficient stochastic transformer based world models for reinforcement learning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3ea9858e-dd34-4c34-8177-f92a7c2eaa79 · outbound
Model-Based Reinforcement Learning under Random Observation Delays @esa (Ref
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fb5634f8-0d14-436a-a41e-bad2bc352143 · outbound
Model-Based Reinforcement Learning under Random Observation Delays Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4e0e15c9-4e47-4cb7-9b6e-aeddc37031be · outbound
Model-Based Reinforcement Learning under Random Observation Delays TD-MPC2: Scalable, Robust World Models for Continuous Control
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
No inbound Pith citation observations are available.