Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:43:32.011996Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2509.26000.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T13:43:32.011996Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d1895590-a8ae-41da-a642-52a358684f77 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Reinforcement learning for HVAC control in intelligent buildings: A technical and conceptual review
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce86ea8-b16e-4a46-ac6c-6a0c3c6cb978 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access MicroPPO: Safe power flow management in decentralized micro-grids with proximal policy optimization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bb75d7a-ee11-48a0-a78e-4610715bd47f · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Deep reinforcement learning solutions for energy microgrids management
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52133c76-122a-4296-a6ce-e07a8dc78929 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Deep reinforcement learning framework for autonomous driving.Electronic Imaging, 2017:70–76, 01 2017
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9718af9b-a604-44be-b832-bafe48757bc4 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Deep reinforcement learning for robotics: A survey of real-world successes
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97e8ab1a-4cbd-431d-a0b1-0ef4da00ab7b · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Planning and acting in partially observable stochastic domains.Artificial intelligence, 101(1-2):99–134, 1998
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cf44693-a62c-464e-b551-f53110d02e18 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Deep recurrent Q-learning for partially observable MDPs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9029117-8f2c-4fc5-8ab4-e44ec99a76fc · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Learning deep neural network policies with continuous memory states
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d87faf64-147a-4028-b1cf-33390191b749 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Asymmetric Actor Critic for Image-Based Robot Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89f14fee-3b2c-4d77-bcdd-3934d7075e4e · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Unbiased asymmetric reinforcement learning under partial observability
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4473a43d-2153-4e3b-acc8-4ff7754032d2 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access On Improving Deep Reinforcement Learning for POMDPs
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39256772-0beb-4366-8859-be6985cac341 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Recurrent policy gradients
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68af1ac4-061f-427c-a9f2-2485903e3fda · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Reinforcement learning with long short-term memory
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e44a7f0-07de-4c95-b98c-eb67bccf8d3d · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Recurrent natural policy gradient for POMDPs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e961f44d-52cf-4de8-99ab-6c98522680bb · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Bridging State and History Representations: Understanding Self-Predictive RL
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13f3e427-d2ce-4bff-a5a2-6c6914a49a09 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Approximate information state for approximate planning and reinforcement learning in partially observed systems
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f78ed1fe-7a1a-4aab-87f0-3d21cb5b36a6 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Data-driven planning via imitation learning.The International Journal of Robotics Research, 37(13-14):1632–1672, 2018
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83bbfdd3-b98b-4170-83d9-57e9e99b6eee · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Robust asymmetric learning in POMDPs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d643a3bb-2e4c-49ea-a50e-808bcfaac549 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Informed POMDP: Leveraging additional information in model-based RL.Reinforcement Learning Journal, 2024
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69114514-c576-442a-b49e-0d5c1a44bd22 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access The wasserstein believer: Learning belief updates for partially observable environments through reliable latent space models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8be63fca-ba09-4327-bf7f-7fbaeab8df18 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Hu, James Springer, Oleh Rybkin, and Dinesh Jayaraman
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55edacc7-bfc8-4618-829b-275b50df21ec · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Neural policy gradient methods: Global optimality and rates of convergence
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0442c94-7ca7-42a4-a4ae-0edf1834d53a · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access A theoretical justification for asymmetric actor-critic algorithms
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf3f0c19-26e5-4169-998c-7d190c0e4778 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access A hilbert space embedding for distributions
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd101662-1b6f-4c81-b239-024beef257c0 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Measuring statistical dependence with hilbert-schmidt norms
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebde8a40-81ea-41f8-85cf-df3cd3d1741c · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access A measure-theoretic approach to kernel conditional mean embeddings
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e15a151e-9ea8-4e62-942d-c5b01d3735be · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access On overfitting and asymptotic bias in batch reinforcement learning with partial observability
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a1de3f-9aba-4d51-a38d-fa71277dc52e · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Solving large POMDPs using real time dynamic programming
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ba80b6-33b5-4c57-8a09-c92cff800ce1 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access gym-pomdps: Gym environments from POMDP files.https://github
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f264b523-9d61-4d30-8bb0-4f1527b2c5ba · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access POMDP Robot Domains
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77eeb484-d2fb-4b49-b37e-f5cf836fc87b · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Multi-agent reinforcement learning with directed exploration and selective memory reuse
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2adb9da-7b1f-47a2-bde3-b735d2dcdc22 · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access gym-gridverse: Gridworld domains for fully and partially observable settings.https://github.com/abaisero/gym-gridverse, 2021
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9554b710-679a-4493-8cd2-6f2e36c9363f · outbound
Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access asym-porl: Asymmetric methods for partially observable reinforcement learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.