Pith. sign in

Paper Citation Record · LEDGER

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

As of 14 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2603.13707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.13707 v3

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T18:15:50.566587Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T09:27:43.453005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:40:06.781891Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e6b7662f-cde7-401e-9e3a-835639d176c8 · outbound

This paper cites Tai- loring solution accuracy for fast whole-body model predictive control of legged robots,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Tai- loring solution accuracy for fast whole-body model predictive control of legged robots,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.683331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.683331Z digest=sha256:e01a073b0b5e4bda66b6d748379e7e4b10b17e9aa495b48d9619f9eb93de2a35

Observation cb16dc8d-5e96-4cbc-976b-0b522170c050 · outbound

This paper cites Seec: Stable end- effector control with model-enhanced residual learning for humanoid loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Seec: Stable end- effector control with model-enhanced residual learning for humanoid loco-manipulation,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.732651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.732651Z digest=sha256:1b093cc55d7ab0a94f946dc6a4247462c4b9de03200b6cf85fd31beb7f8459f7

Observation 2295d57f-360b-4992-a9d3-82cd7fa09cb2 · outbound

This paper cites Omnih2o: Universal and dexterous human-to-humanoid whole-body teleoperation and learning.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Omnih2o: Universal and dexterous human-to-humanoid whole-body teleoperation and learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.790300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.790300Z digest=sha256:1174f897578550550f443d0358b41390b5e0e7ac3079a98a8a7a8b7aef91c2c4

Observation 84b3c53a-dbff-4d45-9410-8d34a65a5fb3 · outbound

This paper cites Humanplus: Hu- manoid shadowing and imitation from humans,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Humanplus: Hu- manoid shadowing and imitation from humans,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.842412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.842412Z digest=sha256:ce0bfffbe49455abd687d822064a5ee731bbd510500265051031ab2f41ec6eb6

Observation edd77b51-2099-48ca-a89d-46c35361b14d · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Diffusion policy: Visuomotor policy learning via action diffusion,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.914916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.914916Z digest=sha256:779d83dd7a4fbe5ea6e0f04928e411dec4df6e1a85e34f78c7a4167562c4ab11

Observation 7d07d19c-11c7-4a8e-a39e-c1fab143f97c · outbound

This paper cites Diffusion policy policy optimization,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Diffusion policy policy optimization,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:46.991325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:46.991325Z digest=sha256:93914db2fd78c26821d43bac89846216256b1cb4008fcdf2d4ec16a0ab88652b

Observation 07be161b-a45e-434d-99a7-ae6c5e9fdf43 · outbound

This paper cites Ppf: Pre-training and preservative fine-tuning of humanoid locomotion via model-assumption- based regularization,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Ppf: Pre-training and preservative fine-tuning of humanoid locomotion via model-assumption- based regularization,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.147433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.147433Z digest=sha256:1d0b426d6c442f127ef5445cf31bfdd1129fb9e0a1bd816a13c0754212fd64b2

Observation 47997f94-1242-468f-a641-6d34497aa582 · outbound

This paper cites Humanoid locomotion and manipulation: Current progress and challenges in control, planning, and learning,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Humanoid locomotion and manipulation: Current progress and challenges in control, planning, and learning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.274577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.274577Z digest=sha256:0eac3e3ac12f3a9ac285e5f02f1393d7216823baf087852a664a6570b6189071

Observation a9a935df-5963-4bfc-951e-ffd88bd011e3 · outbound

This paper cites Twist2: Scalable, portable, and holistic humanoid data collection system,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Twist2: Scalable, portable, and holistic humanoid data collection system,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.413631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.413631Z digest=sha256:e5bddac7f6f441905dd4dafd8c9c47b60e7dcb9b38b762f32157e6d4f057245d

Observation 7ca0f4b8-9443-4b8f-8845-4eded4a28ad2 · outbound

This paper cites Falcon: Learning force-adaptive humanoid loco- manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Falcon: Learning force-adaptive humanoid loco- manipulation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.556653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.556653Z digest=sha256:13d069cc3efaf13a08d5c19d5159482bcbe4cde235b95805c3658e1bf378f1fb

Observation 32751e62-7876-4883-b9af-9eee0fad104d · outbound

This paper cites Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.698722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.698722Z digest=sha256:ebd9fd883de41436ab1b562bda055cf6899d48b38afd87fdae6342dee3222630

Observation dca8669d-b1e6-4fca-9040-c7b2c8ed063b · outbound

This paper cites HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.782052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.782052Z digest=sha256:d52ed7ecd9e592e9fd5bd6dd9196c232f71f3bdc7dec933ae3db9ac1e93d729f

Observation 5f0f052b-fb0e-4015-b6f2-5ba8b3f8715f · outbound

This paper cites Wococo: Learning whole-body humanoid control with sequential contacts,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Wococo: Learning whole-body humanoid control with sequential contacts,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.894940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.894940Z digest=sha256:9e065011eb572b269acce9b736bcc3d1bca28f97cbfe45851099c3c258a7e87f

Observation 65a300ab-954f-470b-a9bf-cfa9e50c9245 · outbound

This paper cites Curiosity-driven learning of joint locomotion and manipulation tasks,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Curiosity-driven learning of joint locomotion and manipulation tasks,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:47.969528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:47.969528Z digest=sha256:7e5404e81f4c465605b8e914c586d2b0ec69371afb04e40079ae4ee2442097b1

Observation 9670c6cc-c8b7-4d52-840a-2308ebb17a2d · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Learning agile soccer skills for a bipedal robot with deep reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.052472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.052472Z digest=sha256:d3ba52e75b5c9434bd178b72b2667bca979c68571742d622bcd90cf9f8292bc0

Observation 4ede54d1-788d-4f23-ae2f-730e90754b08 · outbound

This paper cites Opening the sim-to-real door for humanoid pixel-to- action policy transfer,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Opening the sim-to-real door for humanoid pixel-to- action policy transfer,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.158430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.158430Z digest=sha256:0a5e2cbd69082f7dbe1b7e8b340034a1b150f4a4291b7666e2b0f6609f9a9ea5

Observation ad6f6515-edfb-4643-b4e8-bb7021c761e3 · outbound

This paper cites Sim-to-real learning for humanoid box loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Sim-to-real learning for humanoid box loco-manipulation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.240523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.240523Z digest=sha256:552b2d94ea7188c735c4c6b3d9835805b75c1ea9de19bab6b24b5ddbb76e7abd

Observation cda389a6-1076-4912-a836-ce51652aa1ef · outbound

This paper cites Opt2skill: Imitating dynamically-feasible whole-body trajectories for versatile humanoid loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Opt2skill: Imitating dynamically-feasible whole-body trajectories for versatile humanoid loco-manipulation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.351210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.351210Z digest=sha256:328eb78099cfc40bbb74414778413600d5faabab1f7d551fa42df69dca869717

Observation 5d206791-09fb-47d7-b75b-c5e671578c54 · outbound

This paper cites Hier- archical planning and control for box loco-manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Hier- archical planning and control for box loco-manipulation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.460920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.460920Z digest=sha256:3266679e55fbb37af3116fa6c9a7f12559c0c33dbc3cd63456b0b59d7c6cd0c8

Observation 5673059d-8830-4ad6-9cf8-3cadadb9e32b · outbound

This paper cites Learning fine-grained bimanual manipulation with low-cost hardware,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Learning fine-grained bimanual manipulation with low-cost hardware,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.602126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.602126Z digest=sha256:a6fcfc60675f00c321b90bca5cbd1dbd8b6e32492d32e726f132ecffab5df1e6

Observation f72cddf3-1bdd-4d01-a9ab-da6283afaff1 · outbound

This paper cites Visualmimic: Visual humanoid loco-manipulation via motion tracking and generation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Visualmimic: Visual humanoid loco-manipulation via motion tracking and generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.703596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.703596Z digest=sha256:a05110e9820d88a5c281d6339242713a6a68a705e249580d2be39674391db76f

Observation b0343bb9-65fe-4e55-b041-d126507fe94e · outbound

This paper cites Hdmi: Learning interactive humanoid whole-body control from human videos,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Hdmi: Learning interactive humanoid whole-body control from human videos,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.879849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.879849Z digest=sha256:1ad42a182de38e08a650f31fb853109e2a315f9c575fb9db07b003495d06b35d

Observation 305bb488-6dba-4fc5-b99a-b45adbe7e23e · outbound

This paper cites From imitation to refinement – residual rl for precise assembly,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning From imitation to refinement – residual rl for precise assembly,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:48.992696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:48.992696Z digest=sha256:2ea279a037681dbdec14792a536777c4ee14f22580bd999c843e36c1d3a1a920

Observation db184c3c-63c1-4d1c-a119-5f2ecc1f6247 · outbound

This paper cites Rfs: Reinforce- ment learning with residual flow steering for dexterous manipulation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Rfs: Reinforce- ment learning with residual flow steering for dexterous manipulation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.131081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.131081Z digest=sha256:5ba10a4c819d549c973fee44f973436d45da575ac6c6d25443440b7d2cc2f40e

Observation 2b63d6d6-8220-4e3e-b889-07204454c797 · outbound

This paper cites Residual off-policy rl for finetuning behavior cloning policies,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Residual off-policy rl for finetuning behavior cloning policies,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.220783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.220783Z digest=sha256:9a13cc73962d5d8619468247bfcad20bd31ceaa4394446acdfc0820ce7d5e6d5

Observation ba798607-ea78-49cd-8cab-0fc7deef9f1d · outbound

This paper cites Proximal Policy Optimization Algorithms.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.360402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.360402Z digest=sha256:67da7fa04d20b1b0ff7ad346697db1e8024d7d43e73db2bbd862e0771c72d45e

Observation a74d813b-ea64-4e86-8bcb-913bf209e99a · outbound

This paper cites Efficient online reinforcement learning for diffusion policy,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Efficient online reinforcement learning for diffusion policy,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.499122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.499122Z digest=sha256:7eecc6c74caf4d3207ab267ad8d749233069bb090091d66beddaa209c7cc140e

Observation 2d72b892-502b-4cf9-859d-f9f25451a1e1 · outbound

This paper cites π RL: Online rl fine-tuning for flow-based vision- language-action models,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning π RL: Online rl fine-tuning for flow-based vision- language-action models,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.582649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.582649Z digest=sha256:5541871401738db775d6b4f3e5851b3dad57326147930d411e504b81e87a7e91

Observation 71651857-4908-47c2-b5e4-35b53fd4ce40 · outbound

This paper cites Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Isaac Lab: A GPU-Accelerated Simulation Framework for Multi-Modal Robot Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.694818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.694818Z digest=sha256:4dbe1838fdc3c91f0229538e991e2a1630b835e8a4f4137a3679fb227d3b47ad

Observation c50a1bac-1116-4886-b20a-53a708757a20 · outbound

This paper cites BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning BeamDojo: Learning Agile Humanoid Locomotion on Sparse Footholds

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.743659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.743659Z digest=sha256:74f243fc84b4c4245ec7f65e4e0ab33dfa973b164e1a113c984cb7f0730e5f7e

Observation 6331e3f7-12eb-42d1-bd60-c692afb82ef9 · outbound

This paper cites Re- inforcement learning-based footstep control for humanoid robots on complex terrain,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Re- inforcement learning-based footstep control for humanoid robots on complex terrain,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.850902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.850902Z digest=sha256:04bf07465c04a08bdc32132a675778fb562698520ed09ebc3e8d699d563c5efe

Observation 9a56f533-4561-4441-948a-38df0a363957 · outbound

This paper cites Deepmimic: example-guided deep reinforcement learning of physics-based character skills,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Deepmimic: example-guided deep reinforcement learning of physics-based character skills,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:49.972201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:49.972201Z digest=sha256:037a932078074e6ab02bb20d29234e0e678f542d396acbd8e4339a122b44febc

Observation 9b0f6f8a-4800-4fe7-bcaf-7c045b10541a · outbound

This paper cites Denoising diffusion probabilistic models,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Denoising diffusion probabilistic models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.018839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.018839Z digest=sha256:f88c06ac4280accbf30fa2d02e6f5cb244600df100412fc61721c491d9894f06

Observation e83cbd15-2b66-4150-b438-78460c0ee39d · outbound

This paper cites High- dimensional continuous control using generalized advantage estimation,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning High- dimensional continuous control using generalized advantage estimation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.103256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.103256Z digest=sha256:5c59915a8c7012b7fcc8259ab5423187d3aab26de9bb2e36b2d0aef316a9fd26

Observation 8284605d-37e2-41f7-ad39-7cd5e70c5cb0 · outbound

This paper cites Improved Denoising Diffusion Proba- bilistic Models,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Improved Denoising Diffusion Proba- bilistic Models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.196798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.196798Z digest=sha256:e253d7f61531a9823e608b4c683d79fda76055ef482274f4864a54771461d643

Observation 00433e31-41fc-4383-aaae-c9853f160f01 · outbound

This paper cites Stageact: Stage-conditioned imitation for robust humanoid door opening,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning Stageact: Stage-conditioned imitation for robust humanoid door opening,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.306809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.306809Z digest=sha256:5996fad88ffd7393ca45243faa494dd2d5069b2744eb75ec85312676c0bec0b3

Observation 1d79d6f4-75fa-4a49-bd1e-2b3d38e2fc76 · outbound

This paper cites A behavior architecture for fast humanoid robot door traversals,.

REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning A behavior architecture for fast humanoid robot door traversals,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T18:15:50.566587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:15:50.566587Z digest=sha256:d99877411f949a47b4c9f9e8632468ce41fc4f988787876441e47e11d0a728a2

Pith citing papers

Observation f2339b25-7a1a-432e-bffe-48e6c843d5f8 · inbound

EgoEngine: From Egocentric Human Videos to High-Fidelity Dexterous Robot Demonstrations cites this paper.

EgoEngine: From Egocentric Human Videos to High-Fidelity Dexterous Robot Demonstrations REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-31T02:03:17.213017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T09:27:43.453005Z digest=sha256:3b7c9661a509f402cd559bc64f099313d1aadede424dba721308c4fc254775d5

Observation a50faa90-7ed9-4a1f-bec1-50924d3641f0 · inbound

Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots cites this paper.

Learning Asynchronous Upper-body Task-space Trajectory Tracking Policy for Humanoid Robots REFINE-DP: Diffusion Policy Fine-tuning for Humanoid Loco-manipulation via Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-31T02:03:17.213017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-25T21:08:43.127148Z digest=sha256:6d2200b834f17018ce2bf2b751238df68a4e632ca277443b3bccdf58d07131a9