Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:1910.07207.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:52:14.128108Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T19:40:05.986446Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 67f86b76-e5df-433a-9d5b-4d957e1c4d7f · inbound
AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers Soft Actor-Critic for Discrete Action Settings
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea992a80-c1bd-4a3e-9791-c4d206ad2e72 · inbound
Supervised Learning-enhanced Multi-Group Actor Critic for Live Stream Allocation in Feed Soft Actor-Critic for Discrete Action Settings
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6202d432-773c-4208-affb-58f1ba6d27dc · inbound
Dream to Drive with Predictive Individual World Model Soft Actor-Critic for Discrete Action Settings
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebce658c-c843-4fdb-a824-5f93eff0308e · inbound
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning Soft Actor-Critic for Discrete Action Settings
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 297a5d88-0833-4fa6-9b98-c16ed13bda4e · inbound
Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning Soft Actor-Critic for Discrete Action Settings
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e975619-5ca9-4778-966f-6b1b82faa63c · inbound
SLAC: Safe and Efficient Real-Robot Reinforcement Learning via Unsupervised Simulation Pre-Training Soft Actor-Critic for Discrete Action Settings
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c81fe141-3c17-4314-a79f-5993e0643127 · inbound
Hierarchical Learning-Enhanced MPC for Safe Crowd Navigation with Heterogeneous Constraints Soft Actor-Critic for Discrete Action Settings
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b6c6c6b-ad19-4b2e-8a59-d6bf083e2595 · inbound
Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Soft Actor-Critic for Discrete Action Settings
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a9a86bb-cec3-4ae3-a691-68ba02f6f199 · inbound
Multi-Agent Reinforcement Learning for Inverse Design in Photonic Integrated Circuits Soft Actor-Critic for Discrete Action Settings
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af61d2fc-ea81-4b3b-a899-4dd8e4a604de · inbound
Learning To Communicate Over An Unknown Shared Network Soft Actor-Critic for Discrete Action Settings
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9de1f464-3e53-4755-8767-6d82f0270712 · inbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Soft Actor-Critic for Discrete Action Settings
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50fb1378-d1f6-4cb1-9e71-21cdd1439884 · inbound
Relative Entropy Pathwise Policy Optimization Soft Actor-Critic for Discrete Action Settings
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8a37a432-c78c-4615-a09a-5be87034d878 · inbound
Personalized Exercise Recommendation with Semantically-Grounded Knowledge Tracing Soft Actor-Critic for Discrete Action Settings
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a43ee0-7ece-4498-97a2-fe9732ceeb15 · inbound
DOA: A Degeneracy Optimization Agent with Adaptive Pose Compensation Capability based on Deep Reinforcement Learning Soft Actor-Critic for Discrete Action Settings
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ea9f523-b5a7-47d8-9b2a-551b021f2ce8 · inbound
Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives Soft Actor-Critic for Discrete Action Settings
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fd073855-f9c5-42f0-b93a-14c5a3161913 · inbound
Generalizable Pareto-Optimal Offloading with Reinforcement Learning in Mobile Edge Computing Soft Actor-Critic for Discrete Action Settings
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cd3e863-f66f-4f17-a3a8-b8df328c222a · inbound
MacroNav: Multi-Task Context Representation Learning Enables Efficient Navigation in Unknown Environments Soft Actor-Critic for Discrete Action Settings
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 16501943-0d04-4c7d-ad33-cb1a8e2a7d5e · inbound
R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability Soft Actor-Critic for Discrete Action Settings
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b59e7892-1632-4b36-9df7-20360aefeeee · inbound
Emotion Entanglement and Bayesian Inference for Multi-Dimensional Emotion Understanding Soft Actor-Critic for Discrete Action Settings
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f2a1f31-9eaf-4298-bd9e-47f1e69b87a8 · inbound
Don't Forget the Critic: Value-Based Data Rehearsal for Multi-Cyclic Continual Reinforcement Learning Soft Actor-Critic for Discrete Action Settings
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4ce9cad1-9684-465b-a774-bee3b9e4f106 · inbound
Retry Policy Gradients in Continuous Action Spaces Soft Actor-Critic for Discrete Action Settings
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation caf878c9-6fa1-4d62-9296-9b5c8914068b · inbound
Your GFlowNet Secretly Learns an Optimal Transport Plan Soft Actor-Critic for Discrete Action Settings
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 21e4be6e-28a0-418b-a2a5-a72dab9f0226 · inbound
Back to the Familiar Future: Failure Recovery for VLA Policies via Pre-Imagined Milestone Selection Soft Actor-Critic for Discrete Action Settings
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bfa06eaf-6f94-443d-af2f-9c75b1f544cf · inbound
Event-Driven Reinforcement Learning Enables Long-Horizon Control in Semiconductor Fabrication Soft Actor-Critic for Discrete Action Settings
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1ac1056f-70e0-4e87-92f9-c9aba62a0250 · inbound
Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning Soft Actor-Critic for Discrete Action Settings
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 18408309-491c-402f-9e55-7c8bf0147e3e · inbound
FactorLibrary: From Polynomials to Circuits via Recursive Subgoals Soft Actor-Critic for Discrete Action Settings
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cc54dc00-345d-4228-9582-bdbcec8f327d · inbound
ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning Soft Actor-Critic for Discrete Action Settings
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 406d8430-2d0c-4105-9d0e-a070b86eef4b · inbound
Practical Graph Optimisation and AI-Driven Models for Active Directory Security Hardening Soft Actor-Critic for Discrete Action Settings
Reference 345
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d729623-6732-4a19-b31b-0a77d438c5f2 · inbound
Deep Reinforcement Learning: From First Principles to Reasoning Models Soft Actor-Critic for Discrete Action Settings
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 641735f2-9b57-48e6-98be-efa0f519e538 · inbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Soft Actor-Critic for Discrete Action Settings
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.