Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T10:03:34.319088Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2510.12363.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T10:03:34.319088Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
62 of 62 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f15addc7-28da-460f-b71d-971be214492b · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2affba33-b178-456a-94eb-b934c1f5b208 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning markov state abstractions for deep reinforcement learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9f7fe64-278c-4da5-8650-2a6a0b25df32 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Pedipulate: Enabling Manipulation Skills using a Quadruped Robot 's Leg , 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef68a4bb-1729-4e5d-85f4-cd39121c6e7c · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Scaling mlps: A tale of inductive bias
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80327dcc-2f64-48fe-b0db-a5206da5b4be · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b07a8e93-b328-4795-96e7-077d6dc56b0b · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Dario Bellicoso, Koen Krämer, Markus Stäuble, Dhionis Sako, Fabian Jenelten, Marko Bjelonic, and Marco Hutter
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5795325-1ce7-43c4-8afd-fea62184fc21 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3514963b-92f6-4a69-99f2-bd323b52994a · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion RT -2: Vision - Language - Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c592f4f9-47c4-4a93-a14a-e06d16c38e58 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Symmetric Reinforcement Learning Loss for Robust Learning on Diverse Tasks and Model Scales
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f716a29-44bd-4cbd-b1ff-2b30ce74574c · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning quadrupedal locomotion on deformable terrain
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e3d4d30-45c6-4df2-a7f1-e33638c2408f · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Transfer from Simulation to Real World through Learning Deep Inverse Dynamics Model , 2016
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7b80cc8-74d3-4886-baea-b64c506064ab · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Deep reinforcement learning in a handful of trials using probabilistic dynamics models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16ba2e58-1971-48ef-8948-a52e7b83341a · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Efficient model-based reinforcement learning through optimistic policy search and planning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edaf0117-a842-4c2e-86fd-223adf5a3d99 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion BERT : Pre-training of deep bidirectional transformers for language understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac754f07-fbaa-4e61-a25c-35c30c927275 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Roloma: Robust loco-manipulation for quadruped robots with arms
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b79247-7cdd-428f-a2c6-7c0e617fe6c7 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41482c13-611f-48bf-a375-14c43c80edf8 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Masked autoencoders are scalable vision learners
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5594a5f6-a13f-4886-a309-84a9f27e46e1 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion ANYmal Parkour : Learning Agile Navigation for Quadrupedal Robots , 2023
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6340f53c-ba44-446c-9a66-8a8d01786479 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Dario Bellicoso, Vassilios Tsounis, Jemin Hwangbo, Karen Bodie, Peter Fankhauser, Michael Bloesch, Remo Diethelm, Samuel Bachmann, Amir Melzer, and Mark Hoepflinger
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 669f2f07-67dd-4779-bbc0-846e3ae63956 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning agile and dynamic motor skills for legged robots
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c998ed-53b6-4904-8ad3-5c7c84b0ce88 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Bellman Eluder Dimension: New Rich Classes of RL Problems, and Sample-Efficient Algorithms
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08439231-3013-4d01-9fd5-9dd36269bf6b · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion The Role of Domain Randomization in Training Diffusion Policies for Whole-Body Humanoid Control
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74fc8085-0de7-424d-8198-0aedeeb842ca · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Adam: A Method for Stochastic Optimization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecfe473e-fcf0-4964-8b31-caf6109d252e · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Actor-critic algorithms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f01d23c-e5f1-4e4e-b97b-61dd58461855 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning quadrupedal locomotion over challenging terrain
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc4d0f2-b871-4947-b693-7132bc5c1050 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning to Walk from Three Minutes of Real - World Data with Semi -structured Dynamics Models , 2024
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a4d8fb1-e292-4dc3-925e-709f08923fa9 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion A Survey : Learning Embodied Intelligence from Physical Simulators and World Models , 2025
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af29eafa-30a8-482a-a7a8-4fa71a91ed1b · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fdc2f0b-89a6-46b4-bf0c-d451a8c7dd29 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Combining physics and deep learning to learn continuous-time dynamics models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3b3c0a4f-6c29-4485-a882-771da65cbd07 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51032244-5508-49df-a6f8-c77b79d71d82 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning robust perceptive locomotion for quadrupedal robots in the wild
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f1577d2-c062-4549-b5ca-1eb502819c71 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Orbit: A unified simulation framework for interactive robot learning environments
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d1ddd80-807e-42f9-8996-cee712795bd9 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Symmetry considerations for learning task symmetric robot policies
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b8908e4-05e2-4906-b8bb-6f0a35fb9460 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Murphy, Benjamin J
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e64f4e7d-64d3-46be-9ee8-a7e3687320a1 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Information-Directed Exploration for Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef271075-0b1c-45af-9c5d-b0198420bb88 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion AMP: Adversarial Motion Priors for Stylized Physics-Based Character Control
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 064f7195-f83e-4fd7-baea-b83735b3e5ca · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion ASE : large-scale reusable adversarial skill embeddings for physically simulated characters
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7746778e-0817-4026-9edc-9cf8c71397ca · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Whole-body end-effector pose tracking, 2024
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccd69885-dcd6-4e78-bd2f-9fc122b88fd0 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Language models are unsupervised multitask learners
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6062ef33-4f90-4afc-9db4-82b0d4a3a133 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Learning to walk in minutes using massively parallel deep reinforcement learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 808ac16b-76dd-43dc-9d59-e05db3c19195 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Parkour in the Wild : Learning a General and Extensible Agile Locomotion Policy Using Multi -expert Distillation and RL Fine -tuning, 2025
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df215aec-7d86-4320-884b-0e5eb467c7ef · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Proximal Policy Optimization Algorithms
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22d664b8-f072-4b50-acea-3046f2fa3812 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion RSL-RL: A Learning Library for Robotics Research
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d75942fd-dba1-42da-890a-104e6d089818 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Devon Hjelm, Philip Bachman, and Aaron C
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 635783d2-0bd9-4137-9b8e-e962ae853a9d · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Planning to explore via self-supervised world models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd9f9ead-8bea-47ce-bb5c-4210fca988c7 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Blind Bipedal Stair Traversal via Sim-to-Real Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c2add18-d72e-4490-8f31-36d43eb586a7 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion A Unified MPC Framework for Whole-Body Dynamic Locomotion and Manipulation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 556918e7-efe7-47be-b5b0-a76e796d2ce2 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Guided Reinforcement Learning for Robust Multi - Contact Loco - Manipulation , 2024
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b55a044d-5d02-447e-9d50-e8e07a1bb76e · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Perceptive Pedipulation with Local Obstacle Avoidance , 2024
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b6cdf5-1ab6-4da4-9f48-d6a3ca6eddea · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Gemini Robotics: Bringing AI into the Physical World
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 729cd458-865d-47fd-8b9e-db05840368d5 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Octo: An Open-Source Generalist Robot Policy
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c9ac8a9-0987-4798-ba2b-c9eea7696d2b · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67680e13-674a-45b2-8344-3c1051849caf · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Advanced Skills through Multiple Adversarial Motion Priors in Reinforcement Learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09a85d7d-c0f3-4016-8ec8-9ef83cfdd9a4 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Pretraining in Deep Reinforcement Learning : A Survey , 2022
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cadf785-2002-4825-9e71-e5fb6d831ca3 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Neural Robot Dynamics
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0720e17b-93fc-47e2-baf5-2e48e902188f · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Neural Volumetric Memory for Visual Locomotion Control
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38f54183-f9cb-48f4-a584-dbf9ffd60c9a · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9732485e-cbf4-4a40-8b4f-fd73d5922e81 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Intention- Conditioned Flow Occupancy Models , 2025
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671711d3-e693-4cf3-ab56-581b6eff25de · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion write newline
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6fdb2e7-a1c5-4abf-af22-05565b03b945 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion @esa (Ref
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2278c11-d25d-4377-bf2c-d4b6aae46917 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Unresolved cited work
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 316195b7-0025-4aaf-ad1a-4d3106a29923 · outbound
Pretraining in Actor-Critic Reinforcement Learning for Locomotion Unresolved cited work
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.