Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T05:28:28.298787Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2606.31691.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T05:28:28.298787Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 577b2d4a-0e01-445a-8c35-2f5370713f3c · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Deepmimic: Example-guided deep reinforcement learning of physics-based char- acter skills
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c0851825-7d3c-4dc7-9249-86fdc1a84903 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Towards robust motion control in multi- source uncertain scenarios by robust policy iteration,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3bb1bd68-d373-4717-82e6-a2c714e9802e · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Isaac gym: High performance gpu-based physics simulation for robot learning,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 85d90521-0a46-4d7d-bbbb-db3f667cb576 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Learning to walk in minutes using massively parallel deep reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation beb051b3-c625-4441-a55e-36d2e95d0834 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Parallelq- learning: Scaling off-policy reinforcement learning under massively parallel simulation,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 18ef4f72-b86d-457f-aae2-034b3470a634 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Randomized ensembled double q-learning: Learning fast without a model,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d6a03610-57c2-453d-a00d-3c9d83c53571 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Understanding and preventing capacity loss in reinforcement learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 56a943d7-4848-4e59-bf97-2187a60888ca · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion The primacy bias in deep reinforcement learning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 199f3d32-e9c7-47d1-9c84-61554d50c5a1 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Deep reinforcement learning with plasticity injection,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6df9832f-17f5-4571-b5d5-08501d113738 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Human-level control through deep reinforcement learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9425a985-eb4f-4405-abcf-d5f72282856b · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Continuous control with deep reinforcement learning,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7038c482-182f-4e7e-95a5-3658d64bb3d4 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Off-policy deep reinforcement learning without exploration
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 67368d14-222b-4f99-80af-5d10b4fa672b · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Addressing function approxi- mation error in actor-critic methods
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6a6b1675-58f3-4067-98cf-fdb1d26e33a2 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Smoothed action value functions for learning gaussian policies,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e5b99d32-f6ad-4d9a-98c4-4dea8920deb0 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Stabilizing off-policy q-learning via bootstrapping error reduction,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 082f1e61-89ec-4554-ae27-ec28c0b8afc8 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Demonstrating MuJoCo playground,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b9387db1-df13-429d-b33a-ff9f31a8295d · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Humanoid- Bench: Simulated humanoid benchmark for whole-body locomotion and manipulation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 49629b77-0c2b-4599-b51f-aede4b038750 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Proximal Policy Optimization Algorithms
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ec7a3f50-326f-420d-a0ad-1a081c8816ec · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fa85eaec-3c3b-4835-b975-a38a3c06821e · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion A distributional per- spective on reinforcement learning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 777d2734-968a-4372-809d-a0cc91583f46 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5a9aa663-3fb2-45b1-97e9-8c7f442239a9 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Sferrazza, C., Huang, D.-M., Lin, X., Lee, Y ., and Abbeel, P
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fee1a2df-bc03-4776-b843-a95095394f45 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Understanding plasticity in neural networks,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f9cc123d-ef33-4701-ad96-21cbcefbcbea · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Dropout q-functions for doubly efficient reinforcement learning,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2c714ea5-f0d1-42bd-bd02-48764c69ca09 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Distributional soft actor-critic with three refinements,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 03469d2b-161f-4484-a2dc-8fc8a03152a6 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Controlling overestimation bias with truncated mixture of continuous distributional quantile critics,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 885ddcd0-b9ed-4882-9a81-0af74ef93f18 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Improving Deep Reinforcement Learning by Reducing the Chain Effect of Value and Policy Churn
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 94520889-b3d5-47ca-9d65-34385a5a08c7 · outbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion Stop regressing: Training value functions via classification for scalable deep rl,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
No inbound Pith citation observations are available.