Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T01:04:04.161124Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2502.03717.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T01:04:04.161124Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9862a8bb-a89c-4762-90ec-f217dcae7027 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Walk these ways: Tuning robot control for generalization with multiplicity of behavior,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27d7851e-e09c-4253-8ab9-8c60dbdf1b89 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning RMA: Rapid Motor Adaptation for Legged Robots
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 060ce7d0-de42-43e0-be15-115deb11a287 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Learning quadrupedal locomotion over challenging terrain,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2de40d92-583b-412d-805c-9cd47179d2d4 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Inverse reward design,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4837da17-58ba-4590-bce9-46fbd5cafc6c · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Reward Design with Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64b65b33-7600-4aa1-b7b9-cf8a7245ee5a · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Language to Rewards for Robotic Skill Synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d771b9c7-bfe8-4072-9f82-70ce3b44ccde · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning SayTap: Language to Quadrupedal Locomotion
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dc50f98-2973-42fb-ac68-accebfe8d797 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 539c8285-fa94-41a4-ab54-15348654d59b · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Interactive learning from policy- dependent human feedback,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b0ac61ae-d3c3-4e39-95b7-97d63f2418bf · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Deep reinforcement learning from human preferences,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5750ba4-345c-479c-b362-073007f18f0d · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1355008-a2f1-4702-bc83-86db8485d1f4 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Few-shot preference learning for human- in-the-loop rl,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f9dd4e5d-dff2-4841-aa77-11be2ec4c5b2 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Barkour: Benchmarking Animal-level Agility with Quadruped Robots
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae7638f6-b09f-4d7b-9d81-b13095502303 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Minimizing Energy Consumption Leads to the Emergence of Gaits in Legged Robots
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65fc7f5b-293d-440c-aa85-5d35205685b8 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Fast and efficient locomotion via learned gait transitions,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 24ec8c59-73b2-439e-b7e2-8cffc0b41ec1 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Code as policies: Language model programs for embodied control,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d9a400-f896-46ae-b6a0-0c9723b8bbeb · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Do as i can, not as i say: Grounding language in robotic affordances,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 425c8df5-9d11-49ef-a939-9b06610a7654 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Generative expressive robot behaviors using large language models,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3519c8d9-9308-441f-abba-0962489ea263 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Text2Interaction: Establishing Safe and Preferable Human-Robot Interaction
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7170756-a754-423d-95b7-0c7413b27345 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Learning to Learn Faster from Human Feedback with Language Model Predictive Control
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba02c72b-6728-44ef-8118-682170279fd8 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Language instructed reinforcement learning for human-ai coordination,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 19ba90b6-35f8-4d14-996f-9b0f03fd7e62 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfce0453-4214-4ed9-8f30-82ccf5d2a35a · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Learning human objectives from sequences of physical corrections,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d51c9db1-cb4e-458a-b2b7-0072325c16e5 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Interactively shaping agents via human reinforcement: The tamer framework,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 89db36bf-8add-4e6b-a911-0bbc2542b9d6 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Learning multimodal rewards from rankings,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2fa5e376-b293-42ee-a8ec-e5ca97fd3662 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Extrapolating beyond suboptimal demonstrations via inverse reinforcement learning from observations,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5a44fc6-fb89-4297-93ad-68c6bff49fa4 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Sadigh, A
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 940e35dc-0d03-4f4d-9fe6-32276b1a422b · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Fine-Tuning Language Models from Human Preferences
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d1b2b3c-9d17-4128-8b5e-f31e0f16c049 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Preference-conditioned language- guided abstraction,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9466e122-8789-4dcf-bc52-52526e0b088e · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning MAPLE: A Framework for Active Preference Learning Guided by Large Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d4789680-6288-45d6-b030-618b320d2ccc · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Batch active preference-based learning of reward functions,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 35ac569d-eb5e-48ca-be22-58e922b18a07 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning ICPL: Few-shot In-context Preference Learning via LLMs
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36d0384a-c41d-4b7b-8547-6004e5a37dce · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning B-Pref: Benchmarking Preference-Based Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d86b9fff-eb9e-41d3-9310-4f373a8b727b · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning A bayesian approach for policy learning from trajectory preference queries,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6df6945f-7cda-4f7e-a133-e155c9af6660 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Rank analysis of incomplete block designs: I. the method of paired comparisons,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c02d76a-77b5-477f-8028-75d4157eff5d · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Safe imitation learning via fast bayesian reward inference from preferences,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 27c8227b-bf54-4de5-8750-a1e287986fcf · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning SURF: Semi-supervised Reward Learning with Data Augmentation for Feedback-efficient Preference-based Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e3833e5-4445-4fa7-8fc6-4177bf36919c · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 163a41e4-fc20-47aa-9ab8-c13d796a35c5 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Proximal Policy Optimization Algorithms
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1d407fa-3b44-4c3b-a192-a1c49da205f4 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning GPT-4 Technical Report
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ac6859e-a717-431a-9250-38252ec7df62 · outbound
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning Chain-of-thought prompting elicits reasoning in large language models,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.