Pith. sign in

Paper Citation Record · LEDGER

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos

As of 19 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2505.08367.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.08367 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:00:49.583586Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f8be5de1-f33e-431c-ac7e-497fb28cd909 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning to walk in minutes using massively parallel deep reinforcement learning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.385469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.385469Z digest=sha256:e54ed354c5859bb254731509c198d8fecb2deb16d45793fe979cb66b1913b494

Observation 003a80c2-2005-4243-a7d7-bd74aaacc32b · outbound

This paper cites Advanced skills through multiple adversarial motion priors in reinforcement learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Advanced skills through multiple adversarial motion priors in reinforcement learning,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.391273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.391273Z digest=sha256:55182718d658bd7af76a09e040092f3358f5db57940a97adeac2fd2b40599a18

Observation 5e4954a8-c1fe-493c-98a1-46e87dec2118 · outbound

This paper cites Walk these ways: Tuning robot control for generalization with multiplicity of behavior,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Walk these ways: Tuning robot control for generalization with multiplicity of behavior,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.396041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.396041Z digest=sha256:17cd2642341fef342dc0d1a00e8d28af4ff7f1702ea8ca5065c7232b02fed704

Observation 5076248c-1874-4214-ba74-ff551c7ac670 · outbound

This paper cites Adversarial motion priors make good substitutes for complex reward functions,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Adversarial motion priors make good substitutes for complex reward functions,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.400849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.400849Z digest=sha256:04cb744040fd6c985aaf1c4695b154ac524c65a3d196a794a47ff7af02faa816

Observation 932d7c1d-dcb2-451c-a8f9-d17482c80fa3 · outbound

This paper cites Learning agile skills via adversarial imitation of rough partial demonstrations,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning agile skills via adversarial imitation of rough partial demonstrations,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.405588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.405588Z digest=sha256:264153ffcef5e2c350198424b794361b051a50bfff54a12e296596f2fe2534b2

Observation ecf7c619-6db0-49e5-955b-2667f4665dca · outbound

This paper cites Roboclip: One demonstration is enough to learn robot policies,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Roboclip: One demonstration is enough to learn robot policies,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.410414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.410414Z digest=sha256:e3af50e3fabe5b002d85c5d8411f7a1b1b77577d59f1e45e32606d2e41138391

Observation 07c77394-d7b3-4da4-bb8e-24925943c9f6 · outbound

This paper cites SDS -- See it, Do it, Sorted: Quadruped Skill Synthesis from Single Video Demonstration.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos SDS -- See it, Do it, Sorted: Quadruped Skill Synthesis from Single Video Demonstration

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.415620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.415620Z digest=sha256:7e74d1b66cdcb7e3c14b8fc504910651047f86e1c4dda43e1117720c9a15adba

Observation 484b11ab-6ba6-4b8b-a544-307fdf227a3b · outbound

This paper cites Mocapact: A multi-task dataset for simulated humanoid control,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Mocapact: A multi-task dataset for simulated humanoid control,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.114786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.420628Z digest=sha256:d8ef7ea8593809c3de023510b99868cbb04f1c11c64088384296e9f646024e40

Observation b3bd7250-9d93-42ec-9382-8c115272a73b · outbound

This paper cites Sfv: Re- inforcement learning of physical skills from videos,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Sfv: Re- inforcement learning of physical skills from videos,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.098226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.425222Z digest=sha256:c9295804ecabec366de8acdc96c692078543448bffac07e4511e19269a730bdd

Observation 67addd62-dcf0-4bcc-9750-e6200c510896 · outbound

This paper cites Deep reinforcement learning-based safe interaction for industrial human-robot collaboration using intrinsic reward function,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning-based safe interaction for industrial human-robot collaboration using intrinsic reward function,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.080062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.429884Z digest=sha256:e5492bc48a42ce13fd70e44f7786154e39df118a79a85e08b6227aca84e041d1

Observation c941bb07-8860-41d9-ac7e-eba2f9e96484 · outbound

This paper cites Achieving Stable High-Speed Locomotion for Humanoid Robots with Deep Reinforcement Learning.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Achieving Stable High-Speed Locomotion for Humanoid Robots with Deep Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.434912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.434912Z digest=sha256:1a902f0d80a490791fe53c898a17605320e94db88a0093b9d820b6b5ff02cfe4

Observation 95bb94f0-1151-43fa-ac5e-497b86b37efb · outbound

This paper cites Mastering the game of go without human knowledge,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Mastering the game of go without human knowledge,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.440252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.440252Z digest=sha256:5920ff8a85acc6fa0a7989c6552e7778d7c9236fc3f55d1cf9733f4af66c8891

Observation 31060761-a577-4c0d-9ba6-26e231f2386b · outbound

This paper cites Drl- dclp: A deep reinforcement learning-based dimension-configurable local planner for robot navigation,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Drl- dclp: A deep reinforcement learning-based dimension-configurable local planner for robot navigation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.046842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.444732Z digest=sha256:ca115d8d0c335d11eae442a4ed7ddda0b0e6493514e2e36d8ebaed8169839534

Observation 4153d326-90bf-41f3-947b-3ade23c80fcb · outbound

This paper cites Deep reinforcement learning for general game playing,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for general game playing,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.028772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.450044Z digest=sha256:789a0346c57d59df4d486dfe37e6d0341daf1d70800f05aa5661bf946d762937

Observation 25a76d6d-0f0e-4160-90a8-b934d47d40f4 · outbound

This paper cites Novel automated interactive reinforcement learning framework with a constraint-based supervisor for procedural tasks,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Novel automated interactive reinforcement learning framework with a constraint-based supervisor for procedural tasks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.011469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.455909Z digest=sha256:d1d0a1d170f7193eaba5fb5263a046eec3ad6fe897dbe519cd0b50fcbef64d6b

Observation 352517f0-15ad-4000-b022-6b5c979e6642 · outbound

This paper cites Deep reinforcement learning for unsu- pervised video summarization with diversity-representativeness reward,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for unsu- pervised video summarization with diversity-representativeness reward,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.995498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.461630Z digest=sha256:df7e48cb181bcb484eecd8def62c68aa4e5ee9ac0999fa612196d036f15d8d0f

Observation 43bd12f5-44d7-46a2-b595-eed7858e2752 · outbound

This paper cites Imitation from observation: Learning to imitate behaviors from raw video via context translation,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Imitation from observation: Learning to imitate behaviors from raw video via context translation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.978334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.467566Z digest=sha256:eff807cd802f004b82c7ae00300b0bdaea58973cc7857f6859a60b89c1bc9957

Observation 4e3c512d-7408-471d-9524-e7278458e02c · outbound

This paper cites Reinforcement Learning with Videos: Combining Offline Observations with Interaction.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Reinforcement Learning with Videos: Combining Offline Observations with Interaction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.473148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.473148Z digest=sha256:c471b3164ea2c25cdb0958d65d8c750ff660fb9151f90955247eef86eb6d3462

Observation b6d92a21-8be0-405f-ab27-4605ccb67ed0 · outbound

This paper cites Learning agile robotic locomotion skills by imitating animals,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning agile robotic locomotion skills by imitating animals,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.961396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.478333Z digest=sha256:504602ac6a5f3f28c57e24640e15076aa49087eebf2de199e39c5f270ece1760

Observation 3017d10b-0e12-4fe8-afe6-014639dd870c · outbound

This paper cites Progprompt: Generating situated robot task plans using large language models,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Progprompt: Generating situated robot task plans using large language models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.487945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.487945Z digest=sha256:122467f2b97e7fecf3edf64c71f576779305462e2b74786bc76ece4d84aa5a85

Observation 6ed62dec-8c63-4f13-903d-56de5e93add8 · outbound

This paper cites Language to Rewards for Robotic Skill Synthesis.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Language to Rewards for Robotic Skill Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.493036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.493036Z digest=sha256:613304c4e63ad06cb29af65e2f0fec2b91097ffc395a17b473e7473cc4edd97e

Observation 65b89694-9417-4811-b140-548212581532 · outbound

This paper cites Eureka: Human-level reward design via coding large language models,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Eureka: Human-level reward design via coding large language models,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.932703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.498626Z digest=sha256:9cd3d88919b4c93403604e1521364e3529a9f05ef0f1b81bff961ea402d496a7

Observation d15b3b39-863e-4178-b19f-18cb9ab12934 · outbound

This paper cites Dreureka: Language model guided sim-to-real transfer,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Dreureka: Language model guided sim-to-real transfer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.915270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.503715Z digest=sha256:409b092ec914c813b827b773529bd1ee9b1ffe12720e4f2922c8631da1172300

Observation 7bc36ccd-dfc0-4d5e-8fca-c861741619e7 · outbound

This paper cites Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.508669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.508669Z digest=sha256:51268777acaa723cbd1418d89b1fb4f3925299c43cc22b0e72053bd157fa2372

Observation d7dd6408-6193-40aa-992b-1acf1df247d1 · outbound

This paper cites Slomo: A general system for legged robot motion imitation from casual videos,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Slomo: A general system for legged robot motion imitation from casual videos,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.897916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.514457Z digest=sha256:245563543272036b32e903f6fe08d6d08a5f49f248df852b25ddb5afef41d707

Observation c1bdfc47-eca9-4aed-8658-397f473c065a · outbound

This paper cites Continuous control with deep reinforcement learning.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Continuous control with deep reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.520309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.520309Z digest=sha256:c8fc26c0677389b34b3a4564065f861798bebb21a60fce45e4bd74ede4649c95

Observation 73ce0b37-5164-4b53-8015-a7ea39dc84c2 · outbound

This paper cites A survey on offline reinforcement learning: Taxonomy, review, and open problems,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos A survey on offline reinforcement learning: Taxonomy, review, and open problems,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.525521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.525521Z digest=sha256:3f3f5195eefff01042995904cfddcdc0b092879ae09fb5772ddf811aec30654a

Observation 3beca98d-57b5-48f5-a907-3a37b8e4195c · outbound

This paper cites Deep reinforcement learning for autonomous driving: A survey,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for autonomous driving: A survey,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.530208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.530208Z digest=sha256:76592734cbb399ca5d87d92ce2c5fd9b604bf8ea7150ad575b448a388f11d054

Observation 3cafdd46-5731-469e-b948-468bf8ed1934 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.536073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.536073Z digest=sha256:82dd72c4290650d967785287d269c484c57c71f3f6b4b538d5bd8bd41a3d6618

Observation 04157bdb-9c1f-4901-b673-dca6b0ad588c · outbound

This paper cites Off-policy deep reinforcement learning without exploration,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Off-policy deep reinforcement learning without exploration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.861601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.541876Z digest=sha256:3e7439ab8427bcd4234a4c42d49e53765a05413ad13654d95ea497deb046b01e

Observation e7753f54-1fc3-431a-9cd5-c61893628c5c · outbound

This paper cites Conservative q-learning for offline reinforcement learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Conservative q-learning for offline reinforcement learning,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.546597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.546597Z digest=sha256:a835e01f359cab1e29128109aa16ca405793191e11db586471fd2a128d0dad73

Observation 295d9e74-32b4-447a-80c4-d3c1f9b5ac67 · outbound

This paper cites Where do rewards come from,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Where do rewards come from,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.833836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.551132Z digest=sha256:fd0bb60671b791dafcc8419f717a1ed83af27988b3a5b6d8f1cd4ff285920882

Observation 8a7be170-d85b-480a-8fd9-549766640f3a · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.555874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.555874Z digest=sha256:181b125e61b6400c9c3e2add9f58163803cebba1b70b12cb04e108ae00d77fdb

Observation 5c0c7afb-5947-4699-a4d4-81eebc13707b · outbound

This paper cites Two-frame motion estimation based on polynomial ex- pansion,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Two-frame motion estimation based on polynomial ex- pansion,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.560695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.560695Z digest=sha256:8919357b4726f85e4ed505412a7a9d004cf3e596caf3b1d26903f522459a9a7b

Observation a87dc0db-53cf-4429-8db2-264127c04271 · outbound

This paper cites Off-policy deep reinforcement learning without exploration,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Off-policy deep reinforcement learning without exploration,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.565552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.565552Z digest=sha256:18694f05d706f4d15c30829dba00bf639871252920660252b53e7fd5e62defc5

Observation cc0bab48-ef4b-4e1b-b151-ef80d67bb7c9 · outbound

This paper cites Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.569935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.569935Z digest=sha256:b28c8ab85c1abe13012bcbbe8f07b7bc61a56aa344c053505f0194584cd242c8

Observation a9c4b02f-1ae7-47e3-a9dd-368cef854201 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Proximal Policy Optimization Algorithms

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.574470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.574470Z digest=sha256:fe5b780265b36a645d63ed1bcca460e71dc1f110fa0a8be898f5279bff0b97ea

Observation 1f284fcd-3f53-418c-94b7-56ab0636e3b0 · outbound

This paper cites Offline reinforcement learning with implicit q-learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline reinforcement learning with implicit q-learning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.579142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.579142Z digest=sha256:eecec5a826e06d4042017f8e0b801dd5e92ece1a6f494474f1548ee4b17afc8d

Observation bd2b8eea-ebbf-482e-b39a-0b46dbdd8291 · outbound

This paper cites Gpt-4v(ision) system card,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Gpt-4v(ision) system card,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.773116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:00:49.583586Z digest=sha256:0c8a01ce6031ab2fa3eca5ad23c65f35fbfe35ab3a5222193d16c7d0a018e27a

Observation 717a2144-7deb-455f-ade6-3ca707320a95 · outbound

This paper cites Available: https://doi.org/10.15607/RSS.2020.XVI.064.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Available: https://doi.org/10.15607/RSS.2020.XVI.064

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.483083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.483083Z digest=sha256:005cee38bb04c7d4a66df015c475b117b0965e06f6c9223b97a653752fd2c40c

Pith citing papers

No inbound Pith citation observations are available.