Pith. sign in

Paper Citation Record · LEDGER

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning

As of 12 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2505.19054.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19054 v2

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-19T13:50:42.091642Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact6
  • verified fuzzy4
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 55344bec-6a96-43ff-a4e4-f44fc07d2a7f · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.290963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:f8dbd29ebb62bb132ee5c03c2bef08baa50c80b1a9b0f7a0f3394ab70b68eb79

Observation b6513092-1d96-4e77-bb30-43132cda5277 · outbound

This paper cites 1889-1897.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning 1889-1897

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:52:20.300952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:a4046748e670958ad51d1fbaaaae4b7d4817567ecae3a23a3448797e84429fda

Observation b35a1235-6277-469d-bf9a-50818136defc · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.295666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:936cf6a30d03babc32e759c136c7c0b4f3c6874aea3f3abce58b486069888441

Observation 60ae0425-41f3-4057-b218-76d458be3cd1 · outbound

This paper cites 1861-1870.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning 1861-1870

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:52:20.288641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:f87c33dc6c9fb8be6ffcc18365c7129919eae08ff9ea3bd49a637d1e9c5566bd

Observation 83578248-09de-43ff-bc08-e8314399a6aa · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.286043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:0965021437d198dfe135a0c12708610224e80edd1a6934d6dd8cc9dce44b7085

Observation 15b43c41-1e60-4c5b-a852-f1e56d80819a · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.324168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:93f10de42f8f689ebda382e0b5f5ebe40b2ef2e6f8a32c4d6f1916e8f1f433d2

Observation c0628d36-f101-45b4-b998-9291eaea52c4 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.326703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:3dd2856cb2f70fb794c92b90a0c59593e52a1524000e2b08b6f5434b1750dbe5

Observation 2bc58868-833b-4cc8-ac5b-d90b259b47c8 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.303800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:0691be7c5fb52a40259794f1246f53dc0ea6b2ea55da5362111410bdfa180154

Observation 0d5c7311-39db-48ee-bec9-a868f09e10e0 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.319197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:17876c6267291893fbe10acba75c509294ce6113a86e0e657d6f98d5ee53234b

Observation 5b6676e2-c9ff-42a7-93e7-644db5374122 · outbound

This paper cites RMA: Rapid Motor Adaptation for Legged Robots.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning RMA: Rapid Motor Adaptation for Legged Robots

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-19T13:52:19.860921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:51394f4b68dc87cd4341f0e9a83a38e7cb349c430879770389c21f31641a9703

Observation 308649c4-7701-45a1-9d9a-53abd625105b · outbound

This paper cites Blind Bipedal Stair Traversal via Sim-to-Real Reinforcement Learning.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Blind Bipedal Stair Traversal via Sim-to-Real Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:52:19.853610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:686bb5d8c73a22069dabda9ef655abd062c818f49a9cef8305dd58e9484f3634

Observation 81b7dc85-f269-4da6-9f2e-a48822c4a872 · outbound

This paper cites 46, 2008, pp 265-271.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning 46, 2008, pp 265-271

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:52:20.329112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:7f12e757e47a9ce8ba57bfca0c38ebd50df23440051ef775307d979fae3c8d81

Observation 456ab353-0c74-4167-9360-ba8c98673a50 · outbound

This paper cites 21.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning 21

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-19T13:52:20.331455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:2e0b506c2c9c64ca97acd1e79bb1ff2514637891374ae4aac61ee531844e8bfa

Observation 3c8b006f-04ff-4f1c-b838-7822b20a58ad · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.308931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:a3a73fcbf6da36a6cc639c3a4989118bfc1ae89125f2d3636ea9febdc28abf09

Observation 76f32a22-55cd-45b0-b2fc-1bc12023387c · outbound

This paper cites Proximal Policy Optimization Algorithms.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Proximal Policy Optimization Algorithms

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-19T13:52:19.865256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:4734c2ccf020a9db3773a236b64af3b71f26f531ac59021c5e2c41ef6592ae6e

Observation 35b99efa-b75c-4e7a-9405-e40a0cc0c287 · outbound

This paper cites Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-19T13:52:19.869636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:6a3fe2cfa031fb4980be40fecb4043c16f850574120f395f288033599bc3143e

Observation 03c6d283-1aa8-4f62-a89d-b4a258d62429 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-19T13:52:19.848012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:5d998c9542774ec82444fff735db5e859bc98a9f9096cac0c917612d70fcb9fe

Observation a3390950-101d-4830-87f2-7da1be86c979 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.321876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:3b546c1a54d84717f34c8ba1abf8f1e774790faced9b78e59279b0fb553d028d

Observation 7973f8e0-1565-4793-af2d-96127b4d7fa8 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.317020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:bf12666504e9a2472b162ec88219d5aa9db91fac8e646e8a0482134fd1a30c3d

Observation 7924e550-1b45-4498-96d9-be1a596ad8ec · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.312113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:f95d356b49a9cd592a395b4f1fe094ee50cc73ce62a24dff13439408eb1b886d

Observation 1d9d7b39-90de-4592-aa44-9cd04a01cbf5 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.314713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:d63355fad049b3e7b7d37825d90144bdcd14c84985a2adfd8959a0d6c793279b

Observation 76367036-1c3b-411d-b406-cb6b50f0cb1e · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.293490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:daee7095834017e3dfee5e228efa24ebcac89d0628bb9598b47bfca77887ba08

Observation 1014af90-6d43-45ed-9f0f-79021f000033 · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.298076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:2bd2e4a5ec3342671e3184a50b252e604292678e4ca6409bdd79549e9032c9df

Observation e3eff874-ee67-4213-b251-0276d2d2a56c · outbound

This paper cites an unresolved cited work.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-19T13:52:20.306068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:cd6995de4146068956fe0a29937ff3d6cc4e76d86a81a4e506cd15c1cde15e77

Observation f5578e8b-4e33-4e1a-8997-c08afc1f36da · outbound

This paper cites RSL-RL: A Learning Library for Robotics Research.

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning RSL-RL: A Learning Library for Robotics Research

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:52:19.873923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T13:50:42.091642Z digest=sha256:99ebea97ae2a6e036af354cc167ef8e56adabe159475c1851264d715e3d9aac7

Pith citing papers

No inbound Pith citation observations are available.