Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion

As of 7 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2506.20036.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20036 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:04:22.954743Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T08:48:14.710291Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 333d578b-a6ab-4fbd-ba2e-17357ea098cb · outbound

This paper cites Learning to Walk via Deep Reinforcement Learning.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning to Walk via Deep Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.748967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.748967Z digest=sha256:cccdb4028dc8d8eee5e9afa9ab138b5799a06e610828e41f945c852f792ca00f

Observation daa5474c-f09e-4b37-b8b3-29fe1abf5c44 · outbound

This paper cites Learning to Walk in the Real World with Minimal Human Effort.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning to Walk in the Real World with Minimal Human Effort

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.760351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.760351Z digest=sha256:a0f3b4bf18c927c1c718e177666856afd3a1ab543465d7eac71662af79ce73ee

Observation e9d06fc2-df31-45fa-be57-29e1d76e8880 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning to walk in minutes using massively parallel deep reinforcement learning,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.766963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.766963Z digest=sha256:217cd02e6c1e13111c521ead1bddcbbc1949c2acfaba9377f3a84ff45c024add

Observation f0bfe37c-3866-452e-8475-47e3d2a54b7e · outbound

This paper cites Learning fast adapta- tion with meta strategy optimization,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning fast adapta- tion with meta strategy optimization,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:25.169003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.785144Z digest=sha256:3a6976d4b8d957aebfe7aa7191504a4e03b554d84a277a74c0ccd534ee93cc27

Observation 0a0c4285-9394-447e-8d87-277170ab32fd · outbound

This paper cites Legged locomotion in challenging terrains using egocentric vision,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Legged locomotion in challenging terrains using egocentric vision,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.799138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.799138Z digest=sha256:9775e554884924865ff80f89aea6f4f40c1d78b9851265248a4bf76e2ae6621a

Observation fd0a9e0c-9437-4f0f-8f90-e1d92ea97ea1 · outbound

This paper cites Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal Transformers.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.808203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.808203Z digest=sha256:c18ad910ce6f8dee3e31865c6fbec9986d4b35960e7f06765124f93dae51225b

Observation b0f45bd1-3c9a-43e8-872a-7ac3102dc1b4 · outbound

This paper cites Policies modulating trajectory generators,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Policies modulating trajectory generators,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.857451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.818994Z digest=sha256:1bc795cb873c340ac83c9d63be50e45cde4c13cdfc48d0495823a88c88dd0978

Observation 8c895a94-052b-4928-9aad-6c3a901fe411 · outbound

This paper cites Learning quadrupedal locomotion over challenging terrain,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning quadrupedal locomotion over challenging terrain,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.834747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.834747Z digest=sha256:87f4e3a2045f6665b822b6397fc4753a57d71242fda5fd7fb18f1baff1474595

Observation 0397cb74-f44f-46c3-8611-9b99acfbc0fa · outbound

This paper cites Visual-locomotion: Learning to walk on complex terrains with vision,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Visual-locomotion: Learning to walk on complex terrains with vision,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.665324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.844742Z digest=sha256:874148ebbbfc3caaf43f17065f558cd4fb125ddc75fe537bf4f176c8ac1d848a

Observation 7b72015c-7d05-470d-a77f-4deb24715e94 · outbound

This paper cites Zero-Shot Terrain Generalization for Visual Locomotion Policies.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Zero-Shot Terrain Generalization for Visual Locomotion Policies

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:04:23.198739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.852843Z digest=sha256:76ce979d0e5ca802f1e7daff46097b4b18a8675fe52ccd6c904fe00e746cb701

Observation c4170442-e519-4122-bfae-8d29ba4a321f · outbound

This paper cites Learning agile locomotion skills with a mentor,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning agile locomotion skills with a mentor,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.411864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.860434Z digest=sha256:c4068f852d70d508f41819b3c59b93393804829c6902a59ccdc27d534ffbf397

Observation 41f936ba-2ed3-4416-b529-93719e526df9 · outbound

This paper cites Learning Agile Robotic Locomotion Skills by Imitating Animals.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning Agile Robotic Locomotion Skills by Imitating Animals

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.865462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.865462Z digest=sha256:4b792d015d9a31524a6fb5e144e8d0a259de2b94336329671c7b74e16f22e832

Observation 490190c5-0737-4616-9197-d5aa1f5a7118 · outbound

This paper cites Real-time trajectory adaptation for quadrupedal locomotion using deep reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Real-time trajectory adaptation for quadrupedal locomotion using deep reinforcement learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:24.154395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.871044Z digest=sha256:b960b3809145f9e4d744da6f01ba3a198b3911aa861e23b4cff94e6a34935663

Observation 73665f83-3395-4377-966d-2764f7468065 · outbound

This paper cites Guided constrained policy optimization for dynamic quadrupedal robot locomotion,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Guided constrained policy optimization for dynamic quadrupedal robot locomotion,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.876052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.876052Z digest=sha256:990aaf844f6fcfc775190e171e705362b338869eddc419dbb7b5dd9334cfebb1

Observation fac83b42-093f-4e70-9f79-15178228f663 · outbound

This paper cites Deepgait: Planning and control of quadrupedal gaits using deep reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Deepgait: Planning and control of quadrupedal gaits using deep reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.880917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.880917Z digest=sha256:62020d3d154cf07b65a23a8b6f91930668591c595be2074d425a382adf3716ba

Observation 659f3e14-aedc-4fec-a466-17373a2407d6 · outbound

This paper cites Allsteps: Curriculum-driven learning of stepping stone skills,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Allsteps: Curriculum-driven learning of stepping stone skills,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:23.744751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.887703Z digest=sha256:840f512cfa3eefa727e6d7c67335d4bb71ed238d12295af5a3ca12691bdcaf27

Observation 2dcdbbc2-1653-43c6-90c6-bc843fb53338 · outbound

This paper cites Learning gen- eralizable locomotion skills with hierarchical reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Learning gen- eralizable locomotion skills with hierarchical reinforcement learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:23.593917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.892820Z digest=sha256:d0885c1081a33ecd04c8c5208f88b3b287329292ebee1648f7d064ef75b1a34f

Observation a7aaf8dc-b9aa-4de5-86ba-380b572d6619 · outbound

This paper cites Scalable deep reinforcement learning for vision-based robotic manipulation,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Scalable deep reinforcement learning for vision-based robotic manipulation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:04:23.519591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:04:22.898822Z digest=sha256:f9736b916a7080ac48695710f8e16be1790318db24d821242a0e9798eafc84d3

Observation c88ac69d-0779-4e24-b38a-7f1f285a8fd8 · outbound

This paper cites Hierarchical plan- ning through goal-conditioned offline reinforcement learning,.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Hierarchical plan- ning through goal-conditioned offline reinforcement learning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.904750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.904750Z digest=sha256:a607a2be8f677f96db378f972b5423717a78414a8cdef4dcc7caf74611463444

Observation f58c29f1-7e1f-4e03-8604-376fb57eed20 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Proximal Policy Optimization Algorithms

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.917740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.917740Z digest=sha256:37639c8358d3ec2706fbc29dc64ecf2dfb0afb0f81598d9945e2804c3f3b56ed

Observation 11e69b46-db17-4390-b2c5-2987336495f1 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.923351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.923351Z digest=sha256:0c11d6724d35e368c34f5a534e8053018c2bf78af7b370c35bbb1c7ea92ab884

Observation ab0e3be2-9288-4487-aa34-9c158f627305 · outbound

This paper cites Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning.

Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:04:22.954743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:04:22.954743Z digest=sha256:2ba671d1766cf780486ff3a1c6acb55341f1783bcb0626257eccdc7d548ecf82

Pith citing papers

Observation 1f1692dc-f9fe-4483-9ca6-3963e6a63028 · inbound

PUMA: Perception-driven Unified Foothold Prior for Mobility Augmented Quadruped Parkour cites this paper.

PUMA: Perception-driven Unified Foothold Prior for Mobility Augmented Quadruped Parkour Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T08:48:14.710291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:48:14.710291Z digest=sha256:e899f6b1f2b75132e8301e8b391acb13a205307a4690d036aef52aa08db7df63