Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:00:50.297031Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2505.23857.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:00:50.297031Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee0928d1-5450-4824-8cb1-c4cca182d8af · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Outracing champion gran turismo drivers with deep reinforcement learning,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 787d3297-e4d8-4e5d-9163-85c5022bfe9a · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Perceiving the world: Question-guided reinforcement learning for text-based games,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3700d2d0-d447-40bb-8dd9-84e69ec476c8 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep reinforcement learning in health- care and bio-medical applications,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1f0340e-21e3-4e01-a9cd-875eeb1b51de · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep reinforcement learning in radiation therapy planning optimization: A comprehensive review,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8796b3e-719e-4147-8566-b3f304acd6ff · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Magnetic control of tokamak plasmas through deep reinforcement learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 612b66ac-8085-491b-ab01-7c23a58ebf16 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Reinforcement learning for decision-making and control in power systems,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25134737-fe53-4046-8510-6efc80983b3d · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Reinforcement learning with long short-term memory,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c813bee-709a-4a04-8910-346eda6952fc · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep recurrent q-learning for partially observable mdps,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f9a1636e-1115-4150-b901-67b8c5bc996d · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Attention is all you need,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 93c61bfe-d5f4-4e00-8e33-ebfe93463c0d · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Decision transformer: Reinforcement learning via sequence modeling,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1def55a1-4992-4de2-a9af-c0b7aab7be50 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Online decision transformer,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 06446c79-5413-42ee-a72e-54df07bcc829 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Trajectory transformer: Model-based reinforcement learning with long-term dependencies,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 06ed75b3-7397-4c66-b1a0-905fdb64cad9 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Optimal control of markov processes with incomplete state information,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a416ac10-d75e-4701-9591-b4d99b42f385 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Multi- agent rollout and policy iteration for pomdp with application to multi- robot repair problems,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 980d06a0-93f1-46b0-8080-a46c11894433 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Controlling contact-rich manipulation under partial observability,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1ac0fda2-bb84-4d20-b8f2-ca37bfe00c90 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Magic: Learning macro-actions for online pomdp planning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1ac94790-f526-4394-b40d-b91cc5ab4816 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Optimizing active surveil- lance for prostate cancer using partially observable markov decision processes,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82fe0bda-cff5-44d2-a44b-57e113b991b6 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Diagnostic policies optimization for chronic diseases based on pomdp model,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a9cf74e5-acff-4bfd-9039-d452f373a715 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Planning and acting in partially observable stochastic domains,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03faa789-ee6f-412e-8ece-635e2626ae33 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Finding structure in time,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c026bac9-5a73-4688-b88d-2ccd7dc3ea67 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Deep Recurrent Q-Learning for Partially Observable MDPs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f246ff6-32ab-4de7-b26f-cc04966336d6 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Long short-term memory,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1158ee86-f78e-4150-b2d9-c65ef2d035bf · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Human-level control through deep reinforcement learning,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd299bc2-33a4-4dbd-b520-42e05fef0675 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Recurrent determin- istic policy gradient method for bipedal locomotion on rough terrain challenge,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 25cf2b11-0e80-4300-be83-112202c4cea7 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Recurrent soft actor critic reinforcement learning for demand response problems,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8dc248fd-84b5-4920-a38b-e5e51796b581 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Memory-based deep reinforcement learning for pomdps,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5e61561-e27a-45bb-82a2-7cdd380b3d24 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Addressing function ap- proximation error in actor-critic methods,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d87b4ba8-4666-4b32-9376-d4908df20c88 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Recurrent model-free RL can be a strong baseline for many POMDPs,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8aa4f39-2325-4c08-a84a-14c8350b98b1 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Ode-based recurrent model-free reinforcement learning for pomdps,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5885b47e-04a4-4382-b76e-44fe516dc020 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Efficient recurrent off-policy rl requires a context-encoder-specific learning rate,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 94bc0942-4711-4676-ab77-67e683241ee3 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability The moving horizon estimation concept,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0eb34a92-e2a9-43e5-bfe7-c90ad9d335c2 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bade8861-cc21-45c7-b86a-e1fed2ee2fb5 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Xception: Deep learning with depthwise separable con- volutions,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a175d284-c13c-4de4-a929-76aa000de7b5 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Thin mobilenet: An enhanced mobilenet architecture,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7864e1a8-ca95-400d-81fe-c7983382935d · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Gymnasium,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5462c8ff-cf8e-477c-852e-295f4576d4e3 · outbound
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability Bert: Pre-training of deep bidirectional transformers for language understanding,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.