Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:00:49.583586Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2505.08367.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:00:49.583586Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f8be5de1-f33e-431c-ac7e-497fb28cd909 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning to walk in minutes using massively parallel deep reinforcement learning,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 003a80c2-2005-4243-a7d7-bd74aaacc32b · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Advanced skills through multiple adversarial motion priors in reinforcement learning,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e4954a8-c1fe-493c-98a1-46e87dec2118 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Walk these ways: Tuning robot control for generalization with multiplicity of behavior,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5076248c-1874-4214-ba74-ff551c7ac670 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Adversarial motion priors make good substitutes for complex reward functions,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 932d7c1d-dcb2-451c-a8f9-d17482c80fa3 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning agile skills via adversarial imitation of rough partial demonstrations,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecf7c619-6db0-49e5-955b-2667f4665dca · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Roboclip: One demonstration is enough to learn robot policies,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07c77394-d7b3-4da4-bb8e-24925943c9f6 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos SDS -- See it, Do it, Sorted: Quadruped Skill Synthesis from Single Video Demonstration
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 484b11ab-6ba6-4b8b-a544-307fdf227a3b · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Mocapact: A multi-task dataset for simulated humanoid control,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b3bd7250-9d93-42ec-9382-8c115272a73b · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Sfv: Re- inforcement learning of physical skills from videos,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 67addd62-dcf0-4bcc-9750-e6200c510896 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning-based safe interaction for industrial human-robot collaboration using intrinsic reward function,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c941bb07-8860-41d9-ac7e-eba2f9e96484 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Achieving Stable High-Speed Locomotion for Humanoid Robots with Deep Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95bb94f0-1151-43fa-ac5e-497b86b37efb · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Mastering the game of go without human knowledge,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31060761-a577-4c0d-9ba6-26e231f2386b · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Drl- dclp: A deep reinforcement learning-based dimension-configurable local planner for robot navigation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4153d326-90bf-41f3-947b-3ade23c80fcb · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for general game playing,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 25a76d6d-0f0e-4160-90a8-b934d47d40f4 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Novel automated interactive reinforcement learning framework with a constraint-based supervisor for procedural tasks,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 352517f0-15ad-4000-b022-6b5c979e6642 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for unsu- pervised video summarization with diversity-representativeness reward,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 43bd12f5-44d7-46a2-b595-eed7858e2752 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Imitation from observation: Learning to imitate behaviors from raw video via context translation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4e3c512d-7408-471d-9524-e7278458e02c · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Reinforcement Learning with Videos: Combining Offline Observations with Interaction
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d92a21-8be0-405f-ab27-4605ccb67ed0 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Learning agile robotic locomotion skills by imitating animals,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3017d10b-0e12-4fe8-afe6-014639dd870c · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Progprompt: Generating situated robot task plans using large language models,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ed62dec-8c63-4f13-903d-56de5e93add8 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Language to Rewards for Robotic Skill Synthesis
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65b89694-9417-4811-b140-548212581532 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Eureka: Human-level reward design via coding large language models,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d15b3b39-863e-4178-b19f-18cb9ab12934 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Dreureka: Language model guided sim-to-real transfer,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7bc36ccd-dfc0-4d5e-8fca-c861741619e7 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7dd6408-6193-40aa-992b-1acf1df247d1 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Slomo: A general system for legged robot motion imitation from casual videos,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c1bdfc47-eca9-4aed-8658-397f473c065a · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Continuous control with deep reinforcement learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73ce0b37-5164-4b53-8015-a7ea39dc84c2 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos A survey on offline reinforcement learning: Taxonomy, review, and open problems,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3beca98d-57b5-48f5-a907-3a37b8e4195c · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Deep reinforcement learning for autonomous driving: A survey,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cafdd46-5731-469e-b948-468bf8ed1934 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04157bdb-9c1f-4901-b673-dca6b0ad588c · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Off-policy deep reinforcement learning without exploration,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e7753f54-1fc3-431a-9cd5-c61893628c5c · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Conservative q-learning for offline reinforcement learning,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 295d9e74-32b4-447a-80c4-d3c1f9b5ac67 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Where do rewards come from,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8a7be170-d85b-480a-8fd9-549766640f3a · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c0c7afb-5947-4699-a4d4-81eebc13707b · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Two-frame motion estimation based on polynomial ex- pansion,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a87dc0db-53cf-4429-8db2-264127c04271 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Off-policy deep reinforcement learning without exploration,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc0bab48-ef4b-4e1b-b151-ef80d67bb7c9 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline-to-online reinforcement learning via balanced replay and pessimistic q-ensemble,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9c4b02f-1ae7-47e3-a9dd-368cef854201 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Proximal Policy Optimization Algorithms
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f284fcd-3f53-418c-94b7-56ab0636e3b0 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Offline reinforcement learning with implicit q-learning,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd2b8eea-ebbf-482e-b39a-0b46dbdd8291 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Gpt-4v(ision) system card,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 717a2144-7deb-455f-ade6-3ca707320a95 · outbound
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos Available: https://doi.org/10.15607/RSS.2020.XVI.064
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.