Pith. sign in

Paper Citation Record · LEDGER

RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2412.09858.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09858 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:26:17.660152Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:49:58.308235Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3364acc7-9aa7-4326-8492-fcdad8c6e926 · inbound

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning cites this paper.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:55:40.358784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:88fcf41940a0fef7c4b49b038ce0a3a73c54b828c6dc5285d43dc7a354bf0659

Observation 360cd47c-54bd-481e-a586-d2ea7a05add6 · inbound

Integrating Diffusion-based Multi-task Learning with Online Reinforcement Learning for Robust Quadruped Robot Control cites this paper.

Integrating Diffusion-based Multi-task Learning with Online Reinforcement Learning for Robust Quadruped Robot Control RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T19:26:17.660152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:26:17.660152Z digest=sha256:4b53fb99050f9be80ed605075af1a6d4d0ea629cb117c51b5e5151300b44b366

Observation 898b5141-124a-41f3-8456-6dccdf659324 · inbound

Arnold: a generalist muscle transformer policy cites this paper.

Arnold: a generalist muscle transformer policy RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T16:41:54.320097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:41:54.320097Z digest=sha256:ab774aba446b6cd9e2d4758cc3a7540af5b8c98acc9f1f3b01d6c4b6e0665e72

Observation 298d2bfe-c995-4e35-9711-a71c32863790 · inbound

$\pi^{*}_{0.6}$: a VLA That Learns From Experience cites this paper.

$\pi^{*}_{0.6}$: a VLA That Learns From Experience RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:34:59.374709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T10:34:59.134604Z digest=sha256:9ec3f2dbbbe2078af076ba15d7d46aee71b1cdf580e2b2616f40b97e9442b75c

Observation 1fe7b4aa-83b4-4e27-a0c9-959c76ea977d · inbound

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training cites this paper.

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:00.200266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:00.200266Z digest=sha256:ea6a5b67d174bc435eefe923a940189bf17ab14b573623cb83187f6b6a1f0b51

Observation a6bac5bf-d19d-453b-b1a2-575a2262b002 · inbound

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities cites this paper.

${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:21.728091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T11:42:34.409651Z digest=sha256:ddccbf355170b619454d3333725eb58a636279b59804d4cf2c89e0a133c5efb3

Observation 1a33a8ec-78ce-4ec8-acf0-89a077ad6181 · inbound

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies cites this paper.

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:36:10.916898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T19:31:38.069592Z digest=sha256:95613c42aed7331227a24709c869d495c1a1e396259823a54565633fbaf8251d

Observation 8f56a2a4-f713-460f-b1f0-517b52dff45c · inbound

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies cites this paper.

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:15:32.390805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T08:05:47.128354Z digest=sha256:3343320d14c9c2fce315d9885712e8cf93055de696086af73d1fcdc213124a6f

Observation 6008e264-f2bc-4cc8-b535-58f40c74c76b · inbound

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning cites this paper.

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:39.422800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T12:09:27.409746Z digest=sha256:e4f662ef5c31a1430e55010c9b34a777a8ec6c1a2fe0f847271cdc05330fe46d

Observation 17a9adad-e6d9-44dc-9c31-3d559f5d3484 · inbound

Flow-based Policy Adaptation without Policy Updates cites this paper.

Flow-based Policy Adaptation without Policy Updates RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:26:59.171127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T01:18:08.001674Z digest=sha256:6843f1b5af7e2ac853670ef9e39f8cfe1c7746109521fb892afd0471185cb849

Observation 8270259e-b914-4de0-851f-481fa9787007 · inbound

DexPIE: Stable Dexterous Policy Improvement from Real-World Experience cites this paper.

DexPIE: Stable Dexterous Policy Improvement from Real-World Experience RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:07:30.753147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T16:47:22.504175Z digest=sha256:d508082604533f35e19c521469cf02bcbe5075735a710de235500ee225d42df9

Observation cfbea7c4-f093-4a13-9c56-293551becfd1 · inbound

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning cites this paper.

AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:07:38.967081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:25:59.194721Z digest=sha256:7e3c5d6db0e913aebb6b62f0e73fa2bb718fde732f4d6799ae700a24302cad48

Observation dc49a3a3-079f-48c7-82cc-22ea3731d402 · inbound

Improving Robotic Generalist Policies via Flow Reversal Steering cites this paper.

Improving Robotic Generalist Policies via Flow Reversal Steering RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:35.751705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T06:20:19.209180Z digest=sha256:ab2670e314f4ae2a7d7dd5db651f1e73e9f2e2f6606bf5d0c3cd46b29d6e2846

Observation 62d40101-1ed8-4d44-9e87-b271bf9dba30 · inbound

JoyAI-Sim: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid cites this paper.

JoyAI-Sim: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:34:36.299082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T10:31:54.292897Z digest=sha256:574d818d20f3e1b90336921e00185cf4636f10321c1ef09dddeec623eb3dcd0d

Observation e2a28978-94d0-46ca-aa7e-c48aa2622140 · inbound

Scalable Multi-Task Data Generation via Reinforcement Learning for Language-Conditioned Bimanual Dexterous Manipulation cites this paper.

Scalable Multi-Task Data Generation via Reinforcement Learning for Language-Conditioned Bimanual Dexterous Manipulation RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:19:42.939427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T10:18:01.230648Z digest=sha256:7d8cdf287d9ffac73915f4a1bd1068d1b36e007568ac889829e9a99c8c783568

Observation 0fb83220-c451-4f5b-8c9c-c6527a00d5d1 · inbound

Scalable Multi-Task Data Generation via Reinforcement Learning for Language-Conditioned Bimanual Dexterous Manipulation cites this paper.

Scalable Multi-Task Data Generation via Reinforcement Learning for Language-Conditioned Bimanual Dexterous Manipulation RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:54:36.971120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T10:46:26.385071Z digest=sha256:75c62aaf200df3e5b42e9afbe8777b02dbdc6ffe4b2c67f48cce79428ba094e9

Observation e64a8b7e-6135-4813-bd8b-490bde4e5899 · inbound

InSight: Self-Guided Skill Acquisition via Steerable VLAs cites this paper.

InSight: Self-Guided Skill Acquisition via Steerable VLAs RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:49:58.309660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T00:10:51.721485Z digest=sha256:39be2848c93e5deeb4bf93634385d61bd76a6b53ff0e07bfd8dbf83bd7fd0b76

Observation 6f15f21a-c235-49f2-ad1e-e3fc6bf16095 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning

Reference 189

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:52.278110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:52.278110Z digest=sha256:00a8703af8765e9f28301764461c6dcb96ecf4403049eca67c6d9a87818b0332