Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:01:22.868400Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.11153.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:01:22.868400Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1868bd64-3169-486a-ac63-7bef0f6ee600 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Reinforcement learning in game industry—review, prospects and challenges,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5b36d62b-6933-4947-ad82-0c273075a833 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Artificial intelligence, machine learning and deep learning in advanced robotics, a review,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afb15a91-ee65-49f1-81ff-92498e50dd95 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Playing Atari with Deep Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0d16ee2-0918-4a41-9bc0-b48eeb8b0c50 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Weakly coupled deep q-networks,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2ca25150-c058-45a0-950c-fe3891c86abf · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Tuning apex dqn: A reinforcement learning based deep q-network algorithm,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0e111d3c-b700-45ee-8af1-df7aa143c6ae · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Safe reinforcement learning via shielding under partial observability,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c0c88722-5cad-4be5-be71-14d9dad8da2e · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Integrated task and motion planning for safe legged navigation in partially observable environments,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2a3f48f8-7f06-47b2-a650-a797f9fb2ba3 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep reinforcement learning for multiagent systems: A review of challenges, solutions, and applications,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95640671-922e-4ad1-b63e-a3f8ee8426f4 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes A definition of continual reinforcement learning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d9b6466-e0f7-4ed8-a408-370001b6650b · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep reinforcement learning unleashing the power of ai in decision-making,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b554e27a-280c-45c4-89ba-5f6b8de6cda1 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Exploration in deep reinforcement learning: A survey,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca5fec4b-57dc-47e4-bf84-fa36857fee63 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Memory gym: Partially observable challenges to memory-based agents,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0657da13-58f0-4684-8e76-e35f23d2df8e · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep rein- forcement learning: A survey,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5f247310-6da1-4525-b271-3e4b7100aa90 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Partially Observable Markov Decision Processes (POMDPs) and Robotics
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 179ec413-502f-4197-b9b0-e231b42dee0c · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Recurrent neural networks,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5508d405-6ba2-44ff-a703-bd4ae0ce2130 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Long short-term memory,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c3462ee-db52-4081-a500-2e68d82f52db · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Gate-variants of gated recurrent unit (gru) neural networks,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 70c446ac-4926-4f2c-9ad1-327e144ea182 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Recurrent prediction model for partially observable mdps,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ef5001db-3314-4c8c-b216-d89e3637eeda · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Visualizing transformers for nlp: a brief survey,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c37841f9-3099-4022-80b6-56f081478351 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Transformers in vision: A survey,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc20339b-6200-4c54-9e2d-9b4618c91561 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Windows deep transformer q-networks: an extended variance reduction architecture for partially observable reinforcement learning,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c391dc60-aab2-49e3-89ab-0fe39343235a · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep Transformer Q-Networks for Partially Observable Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3888cc66-5364-4345-93f6-9de049e8a60e · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep Recurrent Q-Learning for Partially Observable MDPs
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72a74ba4-3b51-48b8-be89-0bedcb1171fc · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes On improving deep reinforcement learning for pomdps,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 46461b7f-00a1-4fad-b3cd-85734e91a628 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Learning to Communicate to Solve Riddles with Deep Distributed Recurrent Q-Networks
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf1657e5-c0f3-4a7c-b00d-852e039ba788 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep reinforcement learning with bidirectional recurrent neural networks for dynamic spectrum access,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ae84f17b-566d-4557-b30f-08976440d674 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes On transforming reinforcement learning with transformers: The development trajectory,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a562b3df-f1db-4c05-8ff3-3066d67b3fb7 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Transformer in reinforcement learning for decision-making: A survey,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bc4f5edf-f3a8-418b-9dd3-a8428fb3d894 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Attention Is All You Need
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc739bf0-97d2-4002-b7ab-733b2d16428b · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Decision Transformer: Reinforcement Learning via Sequence Modeling
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cb33299-03b4-4e57-a132-5d852f3352db · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Deep Attention Recurrent Q-Network
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bdde2bf2-25da-4f13-825a-1a06159bdf88 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Towards Interpretable Reinforcement Learning Using Attention Augmented Agents
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15e4264b-f93a-4ea3-9c95-268ffa9b797b · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Efficient Transformers in Reinforcement Learning using Actor-Learner Distillation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8c00b89-1879-40dc-97da-39507316d0b6 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Gated Linear Attention Transformers with Hardware-Efficient Training
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95495ef3-5df6-49fb-82f3-4c1d5f2e8526 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Offline reinforcement learning as one big sequence modeling problem,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cf540214-cd83-474a-abbf-f3ea2833c154 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Online Decision Transformer
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d2e5ca5-fe82-49a4-ab95-acd8ef48894f · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Structured State Space Models for In-Context Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9266ec07-5a79-46c4-9b8d-d0c3732ee6af · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Mastering Memory Tasks with World Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53be7564-2bce-4acd-a07d-f7e7840a8e45 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Mastering atari with discrete world models,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 95412e92-4c5f-4e87-961d-c47b4c45e61f · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Transformers are RNNs: Fast Autoregressive Transformers with Linear Attention
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb5fdc0-918f-403d-803b-cb9169442c2c · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes RWKV: Reinventing RNNs for the Transformer Era
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b344d7-2a02-4d2c-8cbe-5b654b71387a · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Resurrecting Recurrent Neural Networks for Long Sequences
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16c939f0-da6f-4dd8-88c1-87f85629d068 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Efficiently modeling long sequences with structured state spaces,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fec57d69-189a-4b98-b57d-14c3d8f1049d · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Simplified State Space Layers for Sequence Modeling
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28bbaeba-5c66-4e6f-9f28-eae008db609f · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30d1836c-1d8b-4175-bf0b-aa61e69e536f · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Efficiently Modeling Long Sequences with Structured State Spaces
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97ea078e-db2f-4e07-a3f1-83beccc4c463 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Longformer: The Long-Document Transformer
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ca39f0-2267-4829-b533-4e45e76c6fca · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Transformer-XL: Attentive language models beyond a fixed-length context,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1e28d407-380d-484e-9aca-ba9d95119dfe · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Stabilizing Transformers for Reinforcement Learning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfcfe6a5-0e3c-4837-889d-77b8519d1689 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Recurrent Memory Transformer
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abedf9b8-7a34-40b0-b995-1881e58d5461 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Reformer: The Efficient Transformer
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdddb050-6020-4a30-b036-b7c46fca5a9f · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Linear transformers are secretly fast weight programmers,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c186d155-232d-4e77-ac6e-696a6e491d05 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Investigating the Role of Feed-Forward Networks in Transformers Using Parallel Attention and Feed-Forward Net Design
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2057806e-73f7-49fb-943d-4dc4d76b0a5e · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Rethinking transformers in solving pomdps,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a7a99566-99ec-4cd8-8365-799994a5d2cb · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Rethinking Attention with Performers
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d42377e-36e6-426d-a376-46d3fb013090 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Pomdp robot domains,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 85361ea8-fcb7-41ea-82e3-c8634bc66ce6 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Learning policies for partially observable environments: Scaling up,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1f01bb95-6aeb-41f4-8076-af9d547f5e1c · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes gym-gridverse: Gridworld domains for fully and partially observable reinforcement learning,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 083fdfd6-0e13-46c0-a7d8-14ce4a391f09 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Solving large pomdps using real time dynamic programming,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1933d1b8-cf1c-4040-9f44-594c111672b1 · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes On Improving Deep Reinforcement Learning for POMDPs
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 926b6d1d-17ff-4c89-b562-24cb62e6ba2a · outbound
Bi-directional Recurrence Improves Transformer in Partially Observable Markov Decision Processes Mastering Atari with Discrete World Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.