Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2006.14171.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:07:48.369414Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:39:41.546939Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 3793f95a-11f8-46cf-86c2-36f4bc144295 · inbound
Dynamic Collaborative Material Distribution System for Intelligent Robots In Smart Manufacturing A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53be11a9-015f-4c00-9348-8634fb2cef0c · inbound
Data-Driven Policy Mapping for Safe RL-based Energy Management Systems A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 116
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07e50d5f-b39a-42f3-8389-29f39ca7fc18 · inbound
Novel Multi-Agent Action Masked Deep Reinforcement Learning for General Industrial Assembly Lines Balancing Problems A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0ad9eb3-73f8-444f-8359-c85db362407c · inbound
Learning to Assemble the Soma Cube with Legal-Action Masked DQN and Safe ZYZ Regrasp on a Doosan M0609 A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe0e8f53-bd8b-43a7-9692-8648a2b0ea01 · inbound
Towards Scalable O-RAN Resource Management: Graph-Augmented Proximal Policy Optimization A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b57996-e29b-451e-8ec2-88e49c4e411e · inbound
TARMM: Scaling Delay-Critical Edge AI Offloading in 5G O-RAN via Temporal Graph Mobility Management A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a71177f-d0ee-4528-9be5-39d1d5221667 · inbound
Your Loss is My Gain: Low Stake Attacks on Liquid Staking Pools A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f5e053d-38c2-49ea-9cfe-528aecb8588a · inbound
TuniQ: Autotuning Compilation Passes for Quantum Workloads at Scale for Effectiveness and Efficiency A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a89b502-5468-4592-97a4-a75c77dbc6b8 · inbound
Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f968414-1195-4db5-977e-e03a6be706bb · inbound
Learning Selective Merge Policies for Deadline-Constrained Coded Caching via Deep Reinforcement Learning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 67e32243-3306-42fd-971d-72c4e23371c6 · inbound
AlphaTransit: Learning to Design City-scale Transit Routes A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7293490f-2098-479d-9857-e34000b2ab40 · inbound
Bellman-Taylor Score Decoding for Markov Decision Processes with State-Dependent Feasible Action Sets A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a3538820-553e-4d4c-ae3c-5905c9db9c56 · inbound
Deep RL for Fast Long-Horizon Operations Scheduling on NASA's Carruthers Geocorona Observatory Mission A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 934b253b-b71e-4296-9722-fe10141e3034 · inbound
Optimal Reward Shaping: Autonomous Car Parking Case Study A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.