Pith. sign in

Paper Citation Record · LEDGER

Residual Reinforcement Learning for Robot Control

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:1812.03201.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1812.03201 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:14:54.312313Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T11:39:46.881274Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7cbfcf38-225d-4b6b-86dd-a8881ceed2c1 · inbound

A Comparison of Action Spaces for Learning Manipulation Tasks cites this paper.

A Comparison of Action Spaces for Learning Manipulation Tasks Residual Reinforcement Learning for Robot Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T11:35:46.497355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:35:46.497355Z digest=sha256:c9b1e0e4cd03b35d508c843e65331a54b73af94b6e55c51a446ded7c15262a7b

Observation 7119433e-63c7-4298-b95b-6aefad78ae80 · inbound

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance cites this paper.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Residual Reinforcement Learning for Robot Control

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.312313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.312313Z digest=sha256:c4924f49a49ba4fd4761e8411ca95c4f0f7241fd5c7bcfe283db15129a669437

Observation 8ce3016e-b097-4050-92bf-5297e3ec3799 · inbound

Gray-Box Computed Torque Control for Differential-Drive Mobile Robot Tracking cites this paper.

Gray-Box Computed Torque Control for Differential-Drive Mobile Robot Tracking Residual Reinforcement Learning for Robot Control

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T13:31:27.531515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:31:27.531515Z digest=sha256:c491d7cc1476ad22cc99be857feb7dcc4b7d4f086ab2d7ca8a19b5327c66f89e

Observation f1f749c5-d387-44fe-83df-a196f3fa2d02 · inbound

HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies cites this paper.

HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies Residual Reinforcement Learning for Robot Control

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-15T11:49:58.839323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T11:48:13.368844Z digest=sha256:0718bbfa6e0a4054b45e091bd0753fcfac60dc2fa851cdb68f3394a1af1d6958

Observation 60af5cfe-a0dd-4186-a658-502e982feb52 · inbound

HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies cites this paper.

HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies Residual Reinforcement Learning for Robot Control

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-21T11:55:04.216099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T11:54:57.866685Z digest=sha256:a0d4a84af765c6eab33036e5cee51438fa201f2b8c082eecdc23e546a4bed8c7

Observation 234ac503-a19e-45c7-9930-6f19e4ce2861 · inbound

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation cites this paper.

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation Residual Reinforcement Learning for Robot Control

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-22T05:41:08.548801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T05:37:26.290308Z digest=sha256:af71eacf4c7352b395e84497074e124a9397436ee98ca1f4830b75ff2525c771

Observation 9541a2f8-49ec-46b3-ac82-f0ac0fe5695b · inbound

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation cites this paper.

CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation Residual Reinforcement Learning for Robot Control

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:54:57.936474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T16:54:39.269896Z digest=sha256:f4186d60a9e1176f0c21762c60afdaac2198e0a608beec99fab8159324d55e0e

Observation d9cba8f6-dd07-489d-bb94-55d0fe0ea713 · inbound

Application of Reinforcement Learning for Multigroup Energy Grid Optimization for Neutron Transport Criticality Problems cites this paper.

Application of Reinforcement Learning for Multigroup Energy Grid Optimization for Neutron Transport Criticality Problems Residual Reinforcement Learning for Robot Control

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-06-29T09:53:17.539724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T09:46:28.025449Z digest=sha256:9d2b488d7e823da67ca10086ab406b38797409240abf70174493b7bb516cbe5c

Observation bebe6664-399a-4ba9-b720-4dff9787566e · inbound

An Agency-Transferring Model-Free Policy Enhancement Technique cites this paper.

An Agency-Transferring Model-Free Policy Enhancement Technique Residual Reinforcement Learning for Robot Control

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-03T00:27:29.870338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T17:11:45.240357Z digest=sha256:c280839b828dd870937b558fa156ecd77f4bba73af0d0d21288b7b309681e424

Observation 3761bc26-95ef-42b8-b939-6853d3dcc301 · inbound

Improving Robotic Generalist Policies via Flow Reversal Steering cites this paper.

Improving Robotic Generalist Policies via Flow Reversal Steering Residual Reinforcement Learning for Robot Control

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-07-03T15:48:35.768547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T06:20:19.209180Z digest=sha256:97a6ae76d75af96490090728f02c7bd806504ede12408f954e9df4c04f6604eb

Observation c11115b6-9542-46fe-a824-fa97fed4c48e · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL Residual Reinforcement Learning for Robot Control

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-04T09:59:44.446333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:2c18fc3acb6fe70625400d6006d071679e74924c6dffcef732b4453bcb2228d7

Observation e8529928-3a4f-42e9-8364-524f5585d5f5 · inbound

Enforcing Human-like Kinematics in Dexterous Piano Playing via Adversarial Posture Regularization cites this paper.

Enforcing Human-like Kinematics in Dexterous Piano Playing via Adversarial Posture Regularization Residual Reinforcement Learning for Robot Control

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:39:46.882694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T07:50:18.384238Z digest=sha256:e9b8f6059f9addd40a57a36ec65e62f3a523e7a241b12007afedff875bf18b12

Observation d08ac51b-1aec-47fe-9596-488bcd4bbe77 · inbound

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies cites this paper.

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies Residual Reinforcement Learning for Robot Control

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:58:05.715130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-03T11:50:49.215046Z digest=sha256:332fe518298333a91f9b6cce685a8f11237fe16c5f43240e6e384a7648306221

Observation 1fa53bb1-1a65-4934-a5ca-9dbd428d84a3 · inbound

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies cites this paper.

Guided Action Flow: Q-Guided Inference for Flow-Matching Vision-Language-Action Policies Residual Reinforcement Learning for Robot Control

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T08:26:59.428524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:26:59.428524Z digest=sha256:f1472c47c1d93b5bfe95f57984c5b0eedd99e9cec480d5f1ba0652ac15e7d13f

Observation 20293ff1-8855-489a-8e40-1989dc659d38 · inbound

FlowDAgger: Human-in-the-Loop Adaptation of Generative Robot Policies in Latent Space cites this paper.

FlowDAgger: Human-in-the-Loop Adaptation of Generative Robot Policies in Latent Space Residual Reinforcement Learning for Robot Control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T06:05:18.192691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:05:18.192691Z digest=sha256:b8afb49f61098cced77165c18783cf71f7740ce225aa5b79c693ea835ca910d3

Observation 552d55ff-da37-46a2-9d0d-199328349152 · inbound

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC cites this paper.

RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC Residual Reinforcement Learning for Robot Control

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T22:35:37.354247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:35:37.354247Z digest=sha256:976fa5e885570df5199b663694e7e9f7d2e9fd3d031f96c68317b28ee5347833

Observation 2db3407d-4d38-4fbe-814c-718dca8eb7b6 · inbound

Conformal Constraint Tightening for Chance-Constrained Motion Planning with Unknown Dynamics cites this paper.

Conformal Constraint Tightening for Chance-Constrained Motion Planning with Unknown Dynamics Residual Reinforcement Learning for Robot Control

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T04:54:12.185761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:54:12.185761Z digest=sha256:003660cd63b26f31061339d2502ed330aa90765e5b46c6f6d2419e6117f818d0

Observation f5db7fea-11db-4820-981c-ddc72416d264 · inbound

Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? cites this paper.

Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Residual Reinforcement Learning for Robot Control

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-30T11:06:22.954758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T11:06:22.954758Z digest=sha256:331be319fbe87ba05f3eafe41cd870326188b89750b67d1bf808259e6f81a3a7

Observation 4ef21e19-4a49-49a0-bd4e-28aec6a260ac · inbound

Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? cites this paper.

Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Residual Reinforcement Learning for Robot Control

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T04:27:43.895668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:27:43.895668Z digest=sha256:3dcc591983c9606520568f92703ddc4de556e5a65d6f29c9379933dc6dede9d8