Pith. sign in

Paper Citation Record · LEDGER

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos

As of 18 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2505.08367.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.08367 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:00:49.583586Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f8be5de1-f33e-431c-ac7e-497fb28cd909 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning to walk in minutes using massively parallel deep reinforcement learning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.385469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.385469Z digest=sha256:eefcdd0ee7ac98da2f299002241170077fafbd4e2af8c4f4d9c111f39623751f

Observation 003a80c2-2005-4243-a7d7-bd74aaacc32b · outbound

This paper cites Advanced skills through multiple adversarial motion priors in reinforcement learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Advanced skills through multiple adversarial motion priors in reinforcement learning,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.391273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.391273Z digest=sha256:6101152a3dc22eb575dca0c86792565653541b4426d4490680dc297a3065f637

Observation 5e4954a8-c1fe-493c-98a1-46e87dec2118 · outbound

This paper cites Walk these ways: Tuning robot control for generalization with multiplicity of behavior,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Walk these ways: Tuning robot control for generalization with multiplicity of behavior,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.396041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.396041Z digest=sha256:b1dddc26aa50e9d2d6e28aedb4aaf5584a96041d05cf4bd5208d1e0f689c7e50

Observation 5076248c-1874-4214-ba74-ff551c7ac670 · outbound

This paper cites Adversarial motion priors make good substitutes for complex reward functions,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Adversarial motion priors make good substitutes for complex reward functions,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.400849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.400849Z digest=sha256:6acbe0cfe08ebb863854eb266082f1e925534ae1d6f82a59443ff8676d184994

Observation 932d7c1d-dcb2-451c-a8f9-d17482c80fa3 · outbound

This paper cites Learning agile skills via adversarial imitation of rough partial demonstrations,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning agile skills via adversarial imitation of rough partial demonstrations,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.405588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.405588Z digest=sha256:d6d4f7cb2cf39ff42dd89c81bc7c8ad5c30c2d7fde678bfa93e5dbffb8865454

Observation ecf7c619-6db0-49e5-955b-2667f4665dca · outbound

This paper cites Roboclip: One demonstration is enough to learn robot policies,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Roboclip: One demonstration is enough to learn robot policies,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.410414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.410414Z digest=sha256:9496e2b906322a3b69282d371e2e065afdf26899fc0b19f736b501d1e4c2fab3

Observation 07c77394-d7b3-4da4-bb8e-24925943c9f6 · outbound

This paper cites SDS -- See it, Do it, Sorted: Quadruped Skill Synthesis from Single Video Demonstration.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos SDS -- See it, Do it, Sorted: Quadruped Skill Synthesis from Single Video Demonstration

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.415620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.415620Z digest=sha256:3beec9c68f35e8f3029458b29da893d8f0e04de630efd791352197781fe4cfe2

Observation 484b11ab-6ba6-4b8b-a544-307fdf227a3b · outbound

This paper cites Mocapact: A multi-task dataset for simulated humanoid control,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Mocapact: A multi-task dataset for simulated humanoid control,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.114786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.420628Z digest=sha256:6d9472d20bf40e75ec8e06399221bff386a329794d5652e58a425c845d82d99b

Observation b3bd7250-9d93-42ec-9382-8c115272a73b · outbound

This paper cites Sfv: Re- inforcement learning of physical skills from videos,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Sfv: Re- inforcement learning of physical skills from videos,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.098226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.425222Z digest=sha256:1be57ce930b0bdc3c789cd6c6c9c2d4083b454494e5bb987950111dc1e584344

Observation 67addd62-dcf0-4bcc-9750-e6200c510896 · outbound

This paper cites Deep reinforcement learning-based safe interaction for industrial human-robot collaboration using intrinsic reward function,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning-based safe interaction for industrial human-robot collaboration using intrinsic reward function,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.080062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.429884Z digest=sha256:04fc931739e8cf5b2bd53bde3b2fcef4a7a7bdabef10967fefc255608687032e

Observation c941bb07-8860-41d9-ac7e-eba2f9e96484 · outbound

This paper cites Achieving Stable High-Speed Locomotion for Humanoid Robots with Deep Reinforcement Learning.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Achieving Stable High-Speed Locomotion for Humanoid Robots with Deep Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.434912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.434912Z digest=sha256:3c9505e729c02046654212e01a6e6567eac18d4a2bb648c912a8805403793b55

Observation 95bb94f0-1151-43fa-ac5e-497b86b37efb · outbound

This paper cites Mastering the game of go without human knowledge,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Mastering the game of go without human knowledge,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.440252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.440252Z digest=sha256:caaf591a23c0a222abd5fc7869111cddb47f8fa70a5a0837e695495d03d0a22e

Observation 31060761-a577-4c0d-9ba6-26e231f2386b · outbound

This paper cites Drl- dclp: A deep reinforcement learning-based dimension-configurable local planner for robot navigation,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Drl- dclp: A deep reinforcement learning-based dimension-configurable local planner for robot navigation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.046842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.444732Z digest=sha256:a35ce1085f485fcb0434c626ff5ee25162fa4495361d2ffc461c7109417157c4

Observation 4153d326-90bf-41f3-947b-3ade23c80fcb · outbound

This paper cites Deep reinforcement learning for general game playing,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for general game playing,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.028772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.450044Z digest=sha256:187933b87d000205a43e711cd41bac796e8d845d773e6bd11cb8d77e6604e271

Observation 25a76d6d-0f0e-4160-90a8-b934d47d40f4 · outbound

This paper cites Novel automated interactive reinforcement learning framework with a constraint-based supervisor for procedural tasks,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Novel automated interactive reinforcement learning framework with a constraint-based supervisor for procedural tasks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:50.011469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.455909Z digest=sha256:77e3a16cf3c9f4f57227fba440d68f2b24fcff5f4cf27de80686d26b153c1190

Observation 352517f0-15ad-4000-b022-6b5c979e6642 · outbound

This paper cites Deep reinforcement learning for unsu- pervised video summarization with diversity-representativeness reward,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for unsu- pervised video summarization with diversity-representativeness reward,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.995498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.461630Z digest=sha256:82681d5e67a09e93d144bfc2db2839c9ebe0babd7440f09ff79f9b5d56594c23

Observation 43bd12f5-44d7-46a2-b595-eed7858e2752 · outbound

This paper cites Imitation from observation: Learning to imitate behaviors from raw video via context translation,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Imitation from observation: Learning to imitate behaviors from raw video via context translation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.978334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.467566Z digest=sha256:83aa437314c23e936020f56ed925e6a40f56131f0140a6adce7ed3ce96bfe059

Observation 4e3c512d-7408-471d-9524-e7278458e02c · outbound

This paper cites Reinforcement Learning with Videos: Combining Offline Observations with Interaction.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Reinforcement Learning with Videos: Combining Offline Observations with Interaction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.473148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.473148Z digest=sha256:58c17c0fa7622b3a703dc1a028e5e0244742bc2b1f0b13b7ce141afc636c75c1

Observation b6d92a21-8be0-405f-ab27-4605ccb67ed0 · outbound

This paper cites Learning agile robotic locomotion skills by imitating animals,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning agile robotic locomotion skills by imitating animals,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.961396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.478333Z digest=sha256:f1e4269da45f344696d27b6dbc17559b6da90a568ae9f6c750be426a4e4e0299

Observation 3017d10b-0e12-4fe8-afe6-014639dd870c · outbound

This paper cites Progprompt: Generating situated robot task plans using large language models,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Progprompt: Generating situated robot task plans using large language models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.487945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.487945Z digest=sha256:378b151229899ef0ab921f562092d4578e92d3e160776d5927a91e9aa2f7312c

Observation 6ed62dec-8c63-4f13-903d-56de5e93add8 · outbound

This paper cites Language to Rewards for Robotic Skill Synthesis.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Language to Rewards for Robotic Skill Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.493036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.493036Z digest=sha256:f1ed61767a9c81efaede5d104bec31f60155b0716dfea58e31c3927e3e42f18c

Observation 65b89694-9417-4811-b140-548212581532 · outbound

This paper cites Eureka: Human-level reward design via coding large language models,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Eureka: Human-level reward design via coding large language models,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.932703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.498626Z digest=sha256:0ef299a407f68387cd604a2fcbfb37c7d827052cd557b4dc9252af4d1c27d8ca

Observation d15b3b39-863e-4178-b19f-18cb9ab12934 · outbound

This paper cites Dreureka: Language model guided sim-to-real transfer,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Dreureka: Language model guided sim-to-real transfer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.915270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.503715Z digest=sha256:4421ed3f5bda0f4d04e7ade41459bb059a8ecb12121c0b97aa7f276429b05087

Observation 7bc36ccd-dfc0-4d5e-8fca-c861741619e7 · outbound

This paper cites Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.508669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.508669Z digest=sha256:0f126284080a2811d9463499c6e78df8163dd643bc22b5f72edb280f449708c5

Observation d7dd6408-6193-40aa-992b-1acf1df247d1 · outbound

This paper cites Slomo: A general system for legged robot motion imitation from casual videos,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Slomo: A general system for legged robot motion imitation from casual videos,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.897916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.514457Z digest=sha256:b9216f7803a4c26d9143524e32dca2733f5cb180d765c5f588c8ed210769945e

Observation c1bdfc47-eca9-4aed-8658-397f473c065a · outbound

This paper cites Continuous control with deep reinforcement learning.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Continuous control with deep reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.520309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.520309Z digest=sha256:500b0a337054c8ebebdb87868463c52b6c67d98a1cdd1389daa55cf10a6ed2ab

Observation 73ce0b37-5164-4b53-8015-a7ea39dc84c2 · outbound

This paper cites A survey on offline reinforcement learning: Taxonomy, review, and open problems,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos A survey on offline reinforcement learning: Taxonomy, review, and open problems,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.525521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.525521Z digest=sha256:531b2bab7d501ce3aacafb9a6a2afeefcb7e4cf3d9304d6b7751601f363fb4d5

Observation 3beca98d-57b5-48f5-a907-3a37b8e4195c · outbound

This paper cites Deep reinforcement learning for autonomous driving: A survey,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for autonomous driving: A survey,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.530208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.530208Z digest=sha256:b1bd0860887274db0d9e94bf41c3320b5ed24672a182f10a1a32eb400d95caee

Observation 3cafdd46-5731-469e-b948-468bf8ed1934 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.536073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.536073Z digest=sha256:22ef5d8c63bd942ae0f90880b2371f6230202d9454606ab11334710578be231b

Observation 04157bdb-9c1f-4901-b673-dca6b0ad588c · outbound

This paper cites Off-policy deep reinforcement learning without exploration,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Off-policy deep reinforcement learning without exploration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.861601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.541876Z digest=sha256:7d9ba1113f87a62e3df6a39ba8ce9c29c5d426a2775bd32fb69e938c7cc81fa0

Observation e7753f54-1fc3-431a-9cd5-c61893628c5c · outbound

This paper cites Conservative q-learning for offline reinforcement learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Conservative q-learning for offline reinforcement learning,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.546597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.546597Z digest=sha256:d456859da7e8022a0c3da6c6baaa12baebd47c44efb9f3ec4a4b3386c5382f4d

Observation 295d9e74-32b4-447a-80c4-d3c1f9b5ac67 · outbound

This paper cites Where do rewards come from,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Where do rewards come from,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.833836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.551132Z digest=sha256:5f57f40e3c7ab1a2a66135055bf3a3316a4702032d629942fafc9ef1a3f8c77d

Observation 8a7be170-d85b-480a-8fd9-549766640f3a · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.555874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.555874Z digest=sha256:e90e53ce1dc3e2e62e2b2c4f34dc2e5edb65f6d3b446a56ec8217d131bc0717c

Observation 5c0c7afb-5947-4699-a4d4-81eebc13707b · outbound

This paper cites Two-frame motion estimation based on polynomial ex- pansion,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Two-frame motion estimation based on polynomial ex- pansion,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.560695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.560695Z digest=sha256:8b032859ce8a3a619475e2cc58e6b532964742632e1adcda1bf1568625c53c75

Observation a87dc0db-53cf-4429-8db2-264127c04271 · outbound

This paper cites Off-policy deep reinforcement learning without exploration,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Off-policy deep reinforcement learning without exploration,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.565552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.565552Z digest=sha256:3b25ef1dd9439d88b8881d4aef1bcc5586477afe3f7d83bf56a25522fb722e03

Observation cc0bab48-ef4b-4e1b-b151-ef80d67bb7c9 · outbound

This paper cites Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.569935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.569935Z digest=sha256:f5853297df2d3b56126bcd1f55767bc0729cb75bdaec16243a77577f812191e8

Observation a9c4b02f-1ae7-47e3-a9dd-368cef854201 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Proximal Policy Optimization Algorithms

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.574470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.574470Z digest=sha256:9399e8543f74337ad7b0ce516dcec85bb938ba09d98287f86b4deb03685762af

Observation 1f284fcd-3f53-418c-94b7-56ab0636e3b0 · outbound

This paper cites Offline reinforcement learning with implicit q-learning,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline reinforcement learning with implicit q-learning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.579142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.579142Z digest=sha256:8ec34f25eed2c3a7f2fafc76531f195561e45d60b72a11fe4e7d1072bfde4698

Observation bd2b8eea-ebbf-482e-b39a-0b46dbdd8291 · outbound

This paper cites Gpt-4v(ision) system card,.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Gpt-4v(ision) system card,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:00:49.773116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:00:49.583586Z digest=sha256:fdd52147ce679fa3a7bd543f7b4bc80cf7ec96d3e9ec6f298009938a2f28ea3a

Observation 717a2144-7deb-455f-ade6-3ca707320a95 · outbound

This paper cites Available: https://doi.org/10.15607/RSS.2020.XVI.064.

MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Available: https://doi.org/10.15607/RSS.2020.XVI.064

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-15T22:00:49.483083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:00:49.483083Z digest=sha256:266d653d55d67252f0b0a8c207c470bef587ee86e52ffa9e990187a90960d1e3

Pith citing papers

No inbound Pith citation observations are available.