Pith. sign in

Paper Citation Record · LEDGER

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion

As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2508.10423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10423 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:30:30.203171Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy37
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 77978675-00ca-4485-8090-d5e4154448f3 · outbound

This paper cites HOVER: Versatile neural whole-body controller for humanoid robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion HOVER: Versatile neural whole-body controller for humanoid robots,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.115929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:26.580920Z digest=sha256:9e1651e9492369ede95116d5a7d02cd90b95bff41fdcaf71dfcdb1134e50c23d

Observation a7977c1f-a96e-4aa1-8d43-42653518fb8a · outbound

This paper cites Distributional policy gradient with distributional value function,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Distributional policy gradient with distributional value function,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.098429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:26.649591Z digest=sha256:a2512ae8d3da593795c934179591c3941eb68f1da03d9a0872725776342727e4

Observation fe4023d4-eb8b-4e26-a495-6bf5f45352f6 · outbound

This paper cites Walk these ways: Tuning robot control for generalization with multiplicity of behavior,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Walk these ways: Tuning robot control for generalization with multiplicity of behavior,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.081654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:26.738893Z digest=sha256:caf0b5801424c5989673e2361dd23c0d55f57d5bc7146bc26d28006f35fd1c07

Observation 985e80ca-683d-4f4e-bf17-4fa640ee2d8b · outbound

This paper cites Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.063656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:26.795999Z digest=sha256:de0c40f80b19480ac0f5f81c2fa68ed069f3bacc91090e2f8f3e295b35569c32

Observation 7c9e068d-c598-4426-a027-ebc4b4664807 · outbound

This paper cites Biped dynamic walking using reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Biped dynamic walking using reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.045484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:26.847516Z digest=sha256:a4477129a56e1585864e2b0a22edb222f3da9f2bbf2c29364c2c2be186a3817c

Observation c396c3dd-413b-41e3-8349-abb58f1d29e6 · outbound

This paper cites Learning vision-based bipedal locomotion for challeng- ing terrain,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning vision-based bipedal locomotion for challeng- ing terrain,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.024478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:26.945183Z digest=sha256:42f5cfd757c69ebcb7deb6c0dfa38af173f75e436b131af71add3a4a7456f38d

Observation 0ee0e265-40be-47bb-83e0-2ab770211a81 · outbound

This paper cites Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.993595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:27.019704Z digest=sha256:b7893d67028dcdc9f097fde4fe4cb9ef516106e644da851830212d8a79909461

Observation c919e3cd-cf6a-4840-a824-a510f9f64f00 · outbound

This paper cites Optimization-based control for dynamic legged robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimization-based control for dynamic legged robots,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.972502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:27.111040Z digest=sha256:7c8a2f755d7b09cf7bea0f0ad06f42b1c0c7a8b5bc07220c53a565afdc3343a3

Observation 6b8305bd-9467-4bbb-a3b7-aca9f7e45c64 · outbound

This paper cites Versatile multicontact planning and control for legged loco-manipulation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Versatile multicontact planning and control for legged loco-manipulation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.279483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.279483Z digest=sha256:c94526575e3026c50b51efa479fc399451d8dbb34e6b77d2acf068da36e6dbb6

Observation 34a470d6-66e6-426a-89b5-32f07bea1041 · outbound

This paper cites Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.398322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.398322Z digest=sha256:3c1ac70f13483f5199d3e2ad6bcd23fee36b1fecbf6270c947ab8681e5cf413c

Observation db20e7da-e494-42b6-9fa5-c72d95d9758c · outbound

This paper cites Real-world humanoid locomotion with reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Real-world humanoid locomotion with reinforcement learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.567294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.567294Z digest=sha256:b56443ab4781cab7558857c665d82c7c6a33d47192ac591051858abbc2e72abd

Observation 796fbf8d-98bb-459e-9f76-c6f243150cdc · outbound

This paper cites Not only rewards but also constraints: Applications on legged robot locomotion,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Not only rewards but also constraints: Applications on legged robot locomotion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.661298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.661298Z digest=sha256:63c3b4fe1a7f1d6bb150951c48ca8f1ede56bd567655bd3400da8d6edce39881

Observation 4c7eb530-016d-42ac-9dea-1a50c79ae653 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.776297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.776297Z digest=sha256:25c02f0819ec7546df1cf977360aa17ef6e223b53161da04e824ea9112ebe8c3

Observation deca8a10-e073-4628-953f-b38508e1f524 · outbound

This paper cites Learning-based legged locomotion: State of the art and future perspectives,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning-based legged locomotion: State of the art and future perspectives,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.888448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:27.906470Z digest=sha256:33dfe130e9a4f9533b1703afaf09ba0a957e0a6517bc90340ce3b0d89e969dbf

Observation 4becff69-4421-4263-b090-a7df74190e7a · outbound

This paper cites Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.870338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.024263Z digest=sha256:5d5d733758c3f560cfe9d4dad76111f0392e29f8a9e1694081271a7cbccae84b

Observation 46252b04-f802-4a0a-88e7-4d329b2778cf · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning agile soccer skills for a bipedal robot with deep reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.853551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.188458Z digest=sha256:969ee2471699fa72ef48253fe2be90c2d4a028b754717533a9f96139a94db30e

Observation cd324124-89f6-4c88-95c0-b2eacca536ee · outbound

This paper cites Visual whole-body control for legged loco-manipulation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Visual whole-body control for legged loco-manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.837324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.307712Z digest=sha256:2c2617680a80f6f51e04c21263a2f69f345e78c96ba2bdc9ced303b39901e060

Observation bbbe3bc7-956b-48c5-b565-3979b034076e · outbound

This paper cites Teleoperation of humanoid robots: A survey,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Teleoperation of humanoid robots: A survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.822545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.432305Z digest=sha256:a206871201fdd07cc02a9c2e257ba1de97a6062242af05b5a47c22f21e23ea6d

Observation b80c870b-c2c3-448b-bcea-1eaf90a43018 · outbound

This paper cites Sim-to-real robotic sketching using behavior cloning and reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real robotic sketching using behavior cloning and reinforcement learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.805259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.565492Z digest=sha256:dd554ae2b4664355ee09520070af427432f2c824d88e42bebba4eb66f0177e7e

Observation a823bc1a-a4d1-4025-b640-27f1e7692334 · outbound

This paper cites A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.787160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.728125Z digest=sha256:ab6c0f81366e9f8779da2503138f87ccd051c06c01a53e204b7611f622af15c6

Observation edfb33f5-fe21-4ca7-b58f-03b78e567cab · outbound

This paper cites Towards human-level bimanual dexterous manipulation with reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Towards human-level bimanual dexterous manipulation with reinforcement learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.770371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:28.852754Z digest=sha256:c9fbb72f2712648c78a468909fc3b9c985d814ae70bbdaa1e598716db2292cb7

Observation 3ea2df64-31e5-4e72-b1f1-1387ac8df8af · outbound

This paper cites Monotonic value function factorisation for deep multi- agent reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Monotonic value function factorisation for deep multi- agent reinforcement learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.753663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.025150Z digest=sha256:3240cf49209c317e36241d181305207dbc726e9e67d15c21d18b08f975c74e70

Observation 7c4e3bfc-c2bd-4179-8d5b-6abc35ee01bf · outbound

This paper cites Data efficient deep reinforcement learning with action-ranked temporal difference learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Data efficient deep reinforcement learning with action-ranked temporal difference learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.736318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.167712Z digest=sha256:a00b1234188bea8cdd7f2cab4c22db8969cb9dc89e695db53f96b91d6190eb89

Observation 578cc9d9-97ad-47af-97ce-8f23987a7746 · outbound

This paper cites MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:30:30.324419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.314929Z digest=sha256:db2cafa0cde3a3705ea95c860d56a3ddc114022ff9357a8b7f39cc226fe6cd90

Observation e95550f3-e507-4b68-b817-530007f7ea10 · outbound

This paper cites Expressive whole-body control for humanoid robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Expressive whole-body control for humanoid robots,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.719256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.389280Z digest=sha256:c798cb067cb807645ab9d471f6198a6b069b30b02bcad590f825a3a8e35ceedb

Observation 33c3c267-bcbf-482c-ac3e-13ba11571aa4 · outbound

This paper cites Mobile-television: Predictive motion priors for hu- manoid whole-body control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Mobile-television: Predictive motion priors for hu- manoid whole-body control,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.697923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.460025Z digest=sha256:24cf9e05ea9395ba5e4491cd822cc34bff0585be551c377d901bab570381f384

Observation c211d33c-216c-4a26-977a-e24b062e3fac · outbound

This paper cites Wococo: Learning whole-body humanoid control with sequential contacts,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Wococo: Learning whole-body humanoid control with sequential contacts,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.677407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.550583Z digest=sha256:f8f19b9fb6c141a50d20c096373030e0ae42f11b212c578450b6ebda806cf0fa

Observation 58e79e0c-8b28-4c1c-a432-294f41703d36 · outbound

This paper cites Learning human-to-humanoid real-time whole-body teleoperation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning human-to-humanoid real-time whole-body teleoperation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.659984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.637244Z digest=sha256:d7a733670d2364e23248820659eedb0e517a62c088ebd8d1bd55eb42627b2b4c

Observation a15a4892-8a69-4a42-8efe-d4763eda588c · outbound

This paper cites Humanplus: Humanoid shadowing and imitation from humans,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Humanplus: Humanoid shadowing and imitation from humans,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.640690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.738800Z digest=sha256:50ceee16e864c4dacea566a7b73e3279ec1eff942ec9c0c6f3b53e3230e566e2

Observation 2e493578-05f6-4ac4-85d4-c5190a5de706 · outbound

This paper cites Okami: Teaching humanoid robots manipulation skills through single video imitation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Okami: Teaching humanoid robots manipulation skills through single video imitation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.624099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.880304Z digest=sha256:9e78f8bf3dfd0ad0ef20af62866cac28a824da93a8f2f9f3d13bb4e1fea3e39d

Observation 4883b107-7c53-4dc5-b71c-e1776702102c · outbound

This paper cites OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.606108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:29.953977Z digest=sha256:500687438cc1cb3554f85b19523462b2bb3e7abbda91d6ad543ace8c441b6357

Observation bd3bae68-5775-4fe0-afac-cf3684981887 · outbound

This paper cites Perpetual humanoid control for real-time simulated avatars,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Perpetual humanoid control for real-time simulated avatars,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.035721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.035721Z digest=sha256:703ed3136eda454a8ab64aa7f0eed161efff3fdcac7333cb27a583e48f3ee035

Observation 12984db4-c843-4ea9-8300-f16881223180 · outbound

This paper cites Robust and versatile bipedal jumping control through reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Robust and versatile bipedal jumping control through reinforcement learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.576370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.113842Z digest=sha256:88c139a4bb582b51b8ce99a75ab9af3fdba4cc2161484f6d4bcf22eeea32a7ad

Observation d69fa8f1-b5db-4d51-9298-7c8e8479ace1 · outbound

This paper cites Feedback control for cassie with deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Feedback control for cassie with deep reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.560029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.119322Z digest=sha256:aaab0a60a9d2c36d391cc101736904c24f6213e280bb1cabe0e2675136e67cf1

Observation 069b8b47-abbf-4977-ae2a-80882c267943 · outbound

This paper cites Sim-to-real learning of all common bipedal gaits via periodic reward composition,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real learning of all common bipedal gaits via periodic reward composition,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.543052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.125124Z digest=sha256:04f26202830cb415f76589eb8bf96435a90ab56d2e50ffedddc8d50bea3c3152

Observation a0efdb97-5646-43d2-8fc7-8d0310bd2377 · outbound

This paper cites Amp: Adversarial motion priors for stylized physics-based character control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Amp: Adversarial motion priors for stylized physics-based character control,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.523071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.131020Z digest=sha256:8d39fc7cdf1505edd3003c6c09d911a86392640874dd013b8f1e66bafdfdf6e9

Observation d23ade03-2480-4779-9685-53f1e942aa9c · outbound

This paper cites Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.501610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.136272Z digest=sha256:297eadb80a2600c176e6be5f902e5647a8a45e6b9ad45074decccd4bb78b86c6

Observation b35ec145-653e-4fa0-8472-c6763c42a85b · outbound

This paper cites Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.480618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.142441Z digest=sha256:cc2af252a2e739647f67e66d5666cee4cc8c8d0b97c815415e56c15264b8f184

Observation 5d7d7353-b146-4285-a9bf-f050193c4721 · outbound

This paper cites Smarts: An open-source scalable multi-agent rl training school for autonomous driving,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Smarts: An open-source scalable multi-agent rl training school for autonomous driving,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.460116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.149327Z digest=sha256:7e23707e155b6bf8bc4a120e7eb20c178177b7a40450c06961c8e1d3d12111d1

Observation 68817526-329f-484e-bd46-8fcd099f7438 · outbound

This paper cites Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.441939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.155078Z digest=sha256:9d1de78b4a2490c696ed5bfa5097061c63271debefd88221c940b3cc728b9efe

Observation 14e78a57-7704-4df5-badf-01df49855e08 · outbound

This paper cites Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:30:30.294226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.161628Z digest=sha256:03d86677fa8009506f23ede5f475bd4039ff1895dcadefc1ee0b4c67a8b8b455

Observation d1d5d281-6f41-495d-b014-22f98124cffd · outbound

This paper cites an unresolved cited work.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.169226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.169226Z digest=sha256:57f3e2cf26de961d7fcde3e0cfad0e9357db628ecbc80c727769697aa11599cf

Observation b0364da9-194a-4e12-940b-d733d99dcb95 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Proximal Policy Optimization Algorithms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.174196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.174196Z digest=sha256:a982c6fb3e5239a38512c5fc5ecaabf7d23a42d8980bcd587308edb5f78193e5

Observation cbe5f669-2f7d-4588-8e75-28bf798b0af0 · outbound

This paper cites Pomdps for robotic tasks with mixed observability.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Pomdps for robotic tasks with mixed observability

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.405800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.180731Z digest=sha256:4851ea1dfd27b59db019d88a418405822776da166d76cbd65c4871c47416d5c7

Observation 29ca6b55-c894-4014-b506-112b19d6c1eb · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion The surprising effectiveness of ppo in cooperative multi-agent games,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.381407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.187585Z digest=sha256:0b5bd5dbf8ce6fd0be2134eacc368631c672b9c4a7c29ee7df8c4c24d80952f9

Observation 9dd80522-a6bc-4869-8701-4b789a0946cb · outbound

This paper cites Stabilising experience replay for deep multi-agent rein- forcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Stabilising experience replay for deep multi-agent rein- forcement learning,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.362079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.193423Z digest=sha256:a0c5cc61ca8c84c93c9aecff6e637a2cb9e9e87b9b10b1db4b4f8e6c8e4acc65

Observation 90a0c123-743d-4d53-b95a-8c19013752fc · outbound

This paper cites Isaac gym: High performance gpu based physics simulation for robot learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Isaac gym: High performance gpu based physics simulation for robot learning,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.344647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T20:30:30.203171Z digest=sha256:61d28874717437cf7e0d9071358898e4b9ba91bafdbdb563c4f2a39fdf004d9f

Pith citing papers

No inbound Pith citation observations are available.