Pith. sign in

Paper Citation Record · LEDGER

Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.08299.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.08299 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:01:24.053493Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T12:52:17.884464Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 856d1602-f593-4176-85e7-ac6be5519e6b · inbound

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion cites this paper.

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:52:17.886099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T12:50:17.902979Z digest=sha256:022e845d10ecec4bb14b2347b4f3b6c288a461f3cb0fa04dc35c723090138f26

Observation 2bb77020-12c5-4c6b-8420-129b419420fe · inbound

LOVON: Legged Open-Vocabulary Object Navigator cites this paper.

LOVON: Legged Open-Vocabulary Object Navigator Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:01:24.053493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:01:24.053493Z digest=sha256:57eed4d0292d00b322e1417053995f9ebd99b3a4aed0d9861fd8df6a073ce043

Observation 2bdec20f-3c79-4083-aa19-c20f56dfd9e5 · inbound

Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots cites this paper.

Humanoid Occupancy: Enabling A Generalized Multimodal Occupancy Perception System on Humanoid Robots Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T13:49:08.538358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:49:08.538358Z digest=sha256:259bcbf673e0a4ff31ef876e787a3765b226797b27d0bdf04fbfe797bf8fe366

Observation 522cdbdd-760a-46ad-b5e2-7003e268af78 · inbound

DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction cites this paper.

DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T11:03:10.563965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:03:10.563965Z digest=sha256:b44b2e102afa10b6d20caca2e27ef82fcadd82e6466d1b7867ee0a0d471f7830

Observation 38f54183-f9cb-48f4-a584-dbf9ffd60c9a · inbound

Pretraining in Actor-Critic Reinforcement Learning for Locomotion cites this paper.

Pretraining in Actor-Critic Reinforcement Learning for Locomotion Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T10:03:33.681668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:03:33.681668Z digest=sha256:ec903980499de06f2af1178f140d07c0d8079abadeec198aeb8a1812b18705e9

Observation 43010c18-8556-48e8-b9d2-b435ff33790b · inbound

Learning Agile Striker Skills for Humanoid Soccer Robots from Noisy Sensory Input cites this paper.

Learning Agile Striker Skills for Humanoid Soccer Robots from Noisy Sensory Input Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:21:23.433854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T00:20:43.058346Z digest=sha256:6053537f39f4b3ac51c8b7717c9a55635f4e76a877b7d16421d165a55492b26e