Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:30:30.203171Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2508.10423.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:30:30.203171Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 77978675-00ca-4485-8090-d5e4154448f3 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion HOVER: Versatile neural whole-body controller for humanoid robots,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a7977c1f-a96e-4aa1-8d43-42653518fb8a · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Distributional policy gradient with distributional value function,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fe4023d4-eb8b-4e26-a495-6bf5f45352f6 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Walk these ways: Tuning robot control for generalization with multiplicity of behavior,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 985e80ca-683d-4f4e-bf17-4fa640ee2d8b · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7c9e068d-c598-4426-a027-ebc4b4664807 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Biped dynamic walking using reinforcement learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c396c3dd-413b-41e3-8349-abb58f1d29e6 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning vision-based bipedal locomotion for challeng- ing terrain,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0ee0e265-40be-47bb-83e0-2ab770211a81 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c919e3cd-cf6a-4840-a824-a510f9f64f00 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimization-based control for dynamic legged robots,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b8305bd-9467-4bbb-a3b7-aca9f7e45c64 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Versatile multicontact planning and control for legged loco-manipulation,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34a470d6-66e6-426a-89b5-32f07bea1041 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db20e7da-e494-42b6-9fa5-c72d95d9758c · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Real-world humanoid locomotion with reinforcement learning,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 796fbf8d-98bb-459e-9f76-c6f243150cdc · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Not only rewards but also constraints: Applications on legged robot locomotion,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c7eb530-016d-42ac-9dea-1a50c79ae653 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Diffusion policy: Visuomotor policy learning via ac- tion diffusion,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deca8a10-e073-4628-953f-b38508e1f524 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning-based legged locomotion: State of the art and future perspectives,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4becff69-4421-4263-b090-a7df74190e7a · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 46252b04-f802-4a0a-88e7-4d329b2778cf · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning agile soccer skills for a bipedal robot with deep reinforcement learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cd324124-89f6-4c88-95c0-b2eacca536ee · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Visual whole-body control for legged loco-manipulation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bbbe3bc7-956b-48c5-b565-3979b034076e · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Teleoperation of humanoid robots: A survey,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b80c870b-c2c3-448b-bcea-1eaf90a43018 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real robotic sketching using behavior cloning and reinforcement learning,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a823bc1a-a4d1-4025-b640-27f1e7692334 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation edfb33f5-fe21-4ca7-b58f-03b78e567cab · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Towards human-level bimanual dexterous manipulation with reinforcement learning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3ea2df64-31e5-4e72-b1f1-1387ac8df8af · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Monotonic value function factorisation for deep multi- agent reinforcement learning,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7c4e3bfc-c2bd-4179-8d5b-6abc35ee01bf · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Data efficient deep reinforcement learning with action-ranked temporal difference learning,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 578cc9d9-97ad-47af-97ce-8f23987a7746 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e95550f3-e507-4b68-b817-530007f7ea10 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Expressive whole-body control for humanoid robots,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 33c3c267-bcbf-482c-ac3e-13ba11571aa4 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Mobile-television: Predictive motion priors for hu- manoid whole-body control,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c211d33c-216c-4a26-977a-e24b062e3fac · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Wococo: Learning whole-body humanoid control with sequential contacts,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 58e79e0c-8b28-4c1c-a432-294f41703d36 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning human-to-humanoid real-time whole-body teleoperation,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a15a4892-8a69-4a42-8efe-d4763eda588c · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Humanplus: Humanoid shadowing and imitation from humans,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2e493578-05f6-4ac4-85d4-c5190a5de706 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Okami: Teaching humanoid robots manipulation skills through single video imitation,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4883b107-7c53-4dc5-b71c-e1776702102c · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bd3bae68-5775-4fe0-afac-cf3684981887 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Perpetual humanoid control for real-time simulated avatars,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12984db4-c843-4ea9-8300-f16881223180 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Robust and versatile bipedal jumping control through reinforcement learning,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d69fa8f1-b5db-4d51-9298-7c8e8479ace1 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Feedback control for cassie with deep reinforcement learning,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 069b8b47-abbf-4977-ae2a-80882c267943 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real learning of all common bipedal gaits via periodic reward composition,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a0efdb97-5646-43d2-8fc7-8d0310bd2377 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Amp: Adversarial motion priors for stylized physics-based character control,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d23ade03-2480-4779-9685-53f1e942aa9c · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b35ec145-653e-4fa0-8472-c6763c42a85b · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5d7d7353-b146-4285-a9bf-f050193c4721 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Smarts: An open-source scalable multi-agent rl training school for autonomous driving,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 68817526-329f-484e-bd46-8fcd099f7438 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 14e78a57-7704-4df5-badf-01df49855e08 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d1d5d281-6f41-495d-b014-22f98124cffd · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0364da9-194a-4e12-940b-d733d99dcb95 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Proximal Policy Optimization Algorithms
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbe5f669-2f7d-4588-8e75-28bf798b0af0 · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Pomdps for robotic tasks with mixed observability
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29ca6b55-c894-4014-b506-112b19d6c1eb · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion The surprising effectiveness of ppo in cooperative multi-agent games,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9dd80522-a6bc-4869-8701-4b789a0946cb · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Stabilising experience replay for deep multi-agent rein- forcement learning,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 90a0c123-743d-4d53-b95a-8c19013752fc · outbound
MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Isaac gym: High performance gpu based physics simulation for robot learning,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.