Pith. sign in

Paper Citation Record · LEDGER

MuJoCo Playground

As of 19 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 57 inbound Pith citation observations for arXiv:2502.08844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08844 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:34:11.289646Z

measured 135 of 135 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 57 of 57 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:59:53.054525Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved45
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

1
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation fa9376a8-b7cd-439f-85c6-3d5b511d2d18 · outbound

This paper cites Legged locomotion in challenging ter- rains using egocentric vision.

MuJoCo Playground Legged locomotion in challenging ter- rains using egocentric vision

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.929512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.851080Z digest=sha256:0b639581766d5430729447a4219fc9a7e9f8542c47eced6c1b5c8ab5683a88e5

Observation 9ad439b4-ac58-423d-91d5-1bc0d0002a10 · outbound

This paper cites ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation.

MuJoCo Playground ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.856636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.856636Z digest=sha256:7fdbd328d0ddbfc0d9c1b1f6d167b01b63d884f66693f735517149b7c7fcd419

Observation 87b35a59-608e-447d-9477-3e97bb1daa7f · outbound

This paper cites What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study.

MuJoCo Playground What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.862569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.862569Z digest=sha256:442bb95bf63b0aeb61fd64cc53018d541afbe7c3aa9e7dc436a4bc3896d11985

Observation ed7408b3-39d1-47b7-b262-95c71e0b1096 · outbound

This paper cites Learning dexterous in-hand manipula- tion.

MuJoCo Playground Learning dexterous in-hand manipula- tion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.905619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.867778Z digest=sha256:90389deafb7839998c4e31f166940fc4f8da573444d042a275ffb5c7b22f5f97

Observation 8e80af11-6a3d-4509-abea-b9b58a968bff · outbound

This paper cites JAX: composable transformations of Python+NumPy programs, 2018.

MuJoCo Playground JAX: composable transformations of Python+NumPy programs, 2018

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.872779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.872779Z digest=sha256:0c5751ce51fc938ff37a7a5e2e22e426e9f22c5771cba60fcf0da6940645768e

Observation f3a609f5-f2c5-49c1-95e3-c4f42d4b6e4d · outbound

This paper cites Barkour: Benchmarking Animal-level Agility with Quadruped Robots.

MuJoCo Playground Barkour: Benchmarking Animal-level Agility with Quadruped Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.878775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.878775Z digest=sha256:a3a532ee721be046d4937cd2a8c0bec9da93f36419bdaa50408b309a75b84aed

Observation 24bf15bb-17de-44c7-bc89-2c05396331e7 · outbound

This paper cites Closing the sim-to-real loop: Adapting simulation randomization with real world experience.

MuJoCo Playground Closing the sim-to-real loop: Adapting simulation randomization with real world experience

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.884603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.884603Z digest=sha256:3d99e7f0582231fa25efc361251c59c3e8979596982613afb792d81fcb28d203

Observation 17b15cd7-c543-46b4-94f7-d923ed44f1ee · outbound

This paper cites A system for general in-hand object re-orientation.

MuJoCo Playground A system for general in-hand object re-orientation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.889487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.889487Z digest=sha256:d4157ecc8ea06142f11fb23cf08dfb0cca04ce5154184a947feb1528580f87e9

Observation e533a724-fecc-4274-ac44-ed0cb95a89ff · outbound

This paper cites Extreme parkour with legged robots.

MuJoCo Playground Extreme parkour with legged robots

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.759864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.894221Z digest=sha256:f05e21d9f7dcefc8fcd25dd0c8a957819b9e39a84df817cfabebb46140fd03d3

Observation dd315f5e-2687-4a76-a2cc-8cefcdfb4714 · outbound

This paper cites CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects.

MuJoCo Playground CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.899764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.899764Z digest=sha256:2be3f48b416fdb34fb2509b9e551aef28a6cc583641a6ad040b5a3a07f7f1d49

Observation e2690e60-8b56-4e79-bb01-4e6fb3a9001f · outbound

This paper cites Onnx runtime.

MuJoCo Playground Onnx runtime

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.727976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.905027Z digest=sha256:488ed645a45932601b49a23b64cb345d9b0065939141e6a9b47d528a510acb9a

Observation fae33878-d52d-4882-96ec-920ca0513b62 · outbound

This paper cites Flayols, A.

MuJoCo Playground Flayols, A

Reference 12

Resolution
malformed identifier
no resolver link, observed 2026-08-07T23:34:10.910360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.910360Z digest=sha256:11609b9f25a6321bcecc330d279db4d73d109fc22f9bf3cface60de98732e58a

Observation 5799c609-edaf-4daf-b68e-38df5f53f7a2 · outbound

This paper cites Brax-a differentiable physics engine for large scale rigid body simulation, 2021.

MuJoCo Playground Brax-a differentiable physics engine for large scale rigid body simulation, 2021

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.688721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.915072Z digest=sha256:636e7c203f185fbc122c2bb1d78924077a31755f6c74dd189a10afa3e8a7865d

Observation d770e87e-6209-4ab7-9464-3816a08ddcdd · outbound

This paper cites Genesis: A universal and generative physics engine for robotics and beyond, December 2024.

MuJoCo Playground Genesis: A universal and generative physics engine for robotics and beyond, December 2024

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.657010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.919972Z digest=sha256:7f22045e56519946894a995156015ca69fa4633163336815bb91f7b383105adb

Observation 8d763b7c-4a47-48cb-943c-c8f118751cfc · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

MuJoCo Playground Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.628768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.925154Z digest=sha256:ab2b832fb76c19f4c81f40cb2059f76b322551962cf6873d7306297e3d9ff8dd

Observation e3378a3d-945a-4fb8-84fa-d0d77522e814 · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning.

MuJoCo Playground Learning agile soccer skills for a bipedal robot with deep reinforcement learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.929790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.929790Z digest=sha256:68cdd4ade7c0e8a2acfa04934b4083a657e6331e179dc87dc8457fab0a764256

Observation 691c9860-bb41-4d12-846a-d93bd0d631a3 · outbound

This paper cites Mastering Diverse Domains through World Models.

MuJoCo Playground Mastering Diverse Domains through World Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.934582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.934582Z digest=sha256:9db3661ed1bea9e832ae43a95c3409e44db640cf384001bc65b030085e673c1c

Observation 935dd99f-a9aa-47e8-93b2-5979c989e479 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:13.560878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.940144Z digest=sha256:33185665d0f26b4b78c0b63a3ee522eb10adfdc8b18ffe31e0ed02c3634b30ec

Observation 44b7610a-505e-466f-b7f6-4ecff143359b · outbound

This paper cites Dextreme: Transfer of agile in-hand manipulation from simulation to reality.

MuJoCo Playground Dextreme: Transfer of agile in-hand manipulation from simulation to reality

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.530876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.950291Z digest=sha256:3fbc2973e5268ed6144e60a67f2d15604bd9423527df5cabbc9afbc65fab281c

Observation d10f5a5c-74d9-4f77-a4a2-f69e9a424e1f · outbound

This paper cites Td-mpc2: Scalable, robust world models for continuous control, 2024.

MuJoCo Playground Td-mpc2: Scalable, robust world models for continuous control, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.955204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.955204Z digest=sha256:40b5b2cbdc8502e546f9ad45770e80c2e88ec72d030bf3b526445a6cda34688e

Observation bb4627f3-50fb-46a1-a884-8e8230f94f1c · outbound

This paper cites Analytical inverse kinematics for franka emika panda – a geometrical solver for 7- dof manipulators with unconventional design.

MuJoCo Playground Analytical inverse kinematics for franka emika panda – a geometrical solver for 7- dof manipulators with unconventional design

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.960123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.960123Z digest=sha256:3e1d42250eb171ef7e36be4d96a9a5eae829b9e3ebb70021ad8be618b9a8fbc2

Observation 0054a8c2-9ca3-43aa-9d17-62dfd2a3e038 · outbound

This paper cites Evolving control: Evolved high frequency control for continuous control tasks.

MuJoCo Playground Evolving control: Evolved high frequency control for continuous control tasks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.473859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.965210Z digest=sha256:ee6713d60122091d276fa7e08d86a35e9a816f4e97ea19ee4df098ec8eb63ce4

Observation f6ec3702-f13b-4b5f-884f-34de4b0b3146 · outbound

This paper cites DiffTaichi: Differentiable Programming for Physical Simulation.

MuJoCo Playground DiffTaichi: Differentiable Programming for Physical Simulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.970819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.970819Z digest=sha256:e39335a1954b364d1d8fdbaefc33896be1f0397279fc75d388662c3735a590cd

Observation 00660834-3599-431c-8219-4fee5f345484 · outbound

This paper cites How to train your robot with deep reinforcement learning: lessons we have learned.

MuJoCo Playground How to train your robot with deep reinforcement learning: lessons we have learned

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.440821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.976572Z digest=sha256:1967334e165af0b8c88f483e196a3ce0056fbb68da31890fd86d4389aac2a3e3

Observation e5fd55ac-48d6-4f5c-9e10-6e2ab527c8e2 · outbound

This paper cites Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion.

MuJoCo Playground Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion

Reference 25

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T23:34:11.854545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.981630Z digest=sha256:e07453b569f3e350b15152eca22570f3529c584131f29bd433cdb41d8d01bdf7

Observation 14d9023e-45df-48bc-9e6b-17eeaa399a5d · outbound

This paper cites Champion-level drone racing using deep rein- forcement learning.

MuJoCo Playground Champion-level drone racing using deep rein- forcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.986346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.986346Z digest=sha256:1857bfa754d08e61d2e142a3b9aa85fbd6044a7afd9955d72c7b3d77ce7d5d4e

Observation ceb68d23-e1bf-4d9c-af27-0f8a500cd4d2 · outbound

This paper cites Reinforce- ment learning in robotics: A survey.

MuJoCo Playground Reinforce- ment learning in robotics: A survey

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.385541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.993062Z digest=sha256:676dec4f85f4717f9bf2fa2883385b201fe2b456bd0e568078c2e43252dfa914

Observation d10424d4-203a-4b74-81e2-8ccf7f0bf569 · outbound

This paper cites Design and use paradigms for gazebo, an open-source multi-robot sim- ulator.

MuJoCo Playground Design and use paradigms for gazebo, an open-source multi-robot sim- ulator

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.345296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:10.999033Z digest=sha256:3fe71696200a9d160c6fcdc440746ed02c54bd4d9ec9ae5f7e7145ea36ab43ae

Observation 77b33af0-387c-40b6-9719-ece1c4f4f707 · outbound

This paper cites Reinforcement Learning with Augmented Data.

MuJoCo Playground Reinforcement Learning with Augmented Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.003868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.003868Z digest=sha256:e706ccde93709204fe7f8e1d8306c5a2c90f06c6d11a49c499b5b93e3d406b68

Observation 9570e471-2e99-4075-aa0a-070f3178745d · outbound

This paper cites Robust Recovery Controller for a Quadrupedal Robot using Deep Reinforcement Learning.

MuJoCo Playground Robust Recovery Controller for a Quadrupedal Robot using Deep Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.011005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.011005Z digest=sha256:e57429ff2a7c7464b1a0d57bc312b00ee7b0949ce3119781125dabd0798a6eb8

Observation 66826b4e-17cc-4718-baca-d9370cc49f35 · outbound

This paper cites rsl rl: Fast and simple implementation of rl algorithms, designed to run fully on gpu.

MuJoCo Playground rsl rl: Fast and simple implementation of rl algorithms, designed to run fully on gpu

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.316051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.016884Z digest=sha256:aa892b488bb411c543e0c14867254fea50ffca1a5e314ea31a9a5ea94c1e51a0

Observation ab58bfc9-c97b-4375-a64d-44d3782ff0b3 · outbound

This paper cites DROP: Dexterous Reorientation via Online Planning.

MuJoCo Playground DROP: Dexterous Reorientation via Online Planning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.021849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.021849Z digest=sha256:be85499cbfacb005d33632e80f625cece0a7933324f58f5210755d9b37823927

Observation e8ceb63d-94d0-4b3b-b425-ab2393557214 · outbound

This paper cites Rein- forcement learning for versatile, dynamic, and robust bipedal locomotion control.

MuJoCo Playground Rein- forcement learning for versatile, dynamic, and robust bipedal locomotion control

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.282664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.027334Z digest=sha256:e2fca2830381944440f2a0b6cf6307b47cde3b8afbf1e1a6b6ccfa6bd7da35e2

Observation 0644c9fd-46ba-454d-a4b8-0025b40929fd · outbound

This paper cites Gpu- accelerated robotic simulation for distributed reinforce- ment learning.

MuJoCo Playground Gpu- accelerated robotic simulation for distributed reinforce- ment learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.253941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.033633Z digest=sha256:e715ecf5c74b9184ecf8c5ddcf24bb85b9734196364c33c4207a1bda06615404

Observation f5cc8e5d-fe0a-4a6c-bafa-f7850f672064 · outbound

This paper cites Berkeley Humanoid: A Research Platform for Learning-based Control.

MuJoCo Playground Berkeley Humanoid: A Research Platform for Learning-based Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.038539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.038539Z digest=sha256:f3ef33fa8fe3d738608c76cde94e0d0ea0a185e533ce9bc56a36647b1c1e45a6

Observation b655bd5c-76a3-40d5-94ec-4926230a15fb · outbound

This paper cites Learning Humanoid Locomotion with Perceptive Internal Model.

MuJoCo Playground Learning Humanoid Locomotion with Perceptive Internal Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.043978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.043978Z digest=sha256:ce0452bc2b9cdd73adb20ce1c7661781ec9a360b2f20c83b60e3e249d99eb730

Observation 342d6944-cba7-4d91-b2d8-7df0d0d15643 · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

MuJoCo Playground Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.049309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.049309Z digest=sha256:9de6efd2b571be680bf513a0cf91d0118884db81b1f10d96a3f1af678af00335

Observation a4ba11e7-350f-463b-91fa-866a32e4ff96 · outbound

This paper cites Warp: A high-performance python frame- work for gpu simulation and graphics.

MuJoCo Playground Warp: A high-performance python frame- work for gpu simulation and graphics

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.218346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.054659Z digest=sha256:70c7c83b7dae8eba1f876fdf20a1fda244723b07d3d837e8e6356c1429feb064

Observation 8c0cf1ea-4b0e-409e-8339-b5dd660cd4b5 · outbound

This paper cites Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning.

MuJoCo Playground Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.060494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.060494Z digest=sha256:8e79c00eb9a76da903a6060d12c7d19430c7406df483dcdca4a5ca080b94c4e1

Observation 5f108b7e-6dcd-4bb8-9d1f-c4314e2c852e · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild.

MuJoCo Playground Learning robust perceptive locomotion for quadrupedal robots in the wild

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.065799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.065799Z digest=sha256:6060e39c202fd13c00fc94294758d06fc7ec35637976cc8daeadf8d5bd93db1c

Observation 906a61e5-0dc5-4ad2-9e2e-4b71a46d304e · outbound

This paper cites Orbit: A unified simulation framework for interactive robot learning envi- ronments.

MuJoCo Playground Orbit: A unified simulation framework for interactive robot learning envi- ronments

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.152950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.070919Z digest=sha256:9b787f986c4ec339d5a2e6e4143cb2bbe1f2761b3b93d718c5640650672a3a34

Observation 6573d4d5-0bc9-4911-a552-8f4f3d1d9797 · outbound

This paper cites Rusu, Joel Veness, Marc G.

MuJoCo Playground Rusu, Joel Veness, Marc G

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.075894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.075894Z digest=sha256:60a451d4d5099c6332d4fe0bc317d85b2995351fdbc239280c1c8f11b5a7b0bb

Observation 418792c6-3c60-4e88-b1e2-c50e0fbc594b · outbound

This paper cites MuJoCo XLA (MJX).

MuJoCo Playground MuJoCo XLA (MJX)

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.128171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.080924Z digest=sha256:e53f6401718c76899520a76b70db8dddf609e450344b9be3fc703a579bb63dd1

Observation 1062fdd3-caee-42ab-8ebe-7424b0b1ceaf · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:13.094626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.085861Z digest=sha256:1c4aa33f45d4f9af2b3ab1155a4633d2cd4ef8b77c1e6a30a45f532389586720

Observation c2137486-9195-4059-b5be-fd090b836d89 · outbound

This paper cites Dexpbt: Scaling up dexterous manipulation for hand-arm systems with pop- ulation based training.

MuJoCo Playground Dexpbt: Scaling up dexterous manipulation for hand-arm systems with pop- ulation based training

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.054620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.091829Z digest=sha256:4273d56e878f1eaf664af136a2e65888e74ed410c8b447df72500de0758aec46

Observation 2965fe43-a207-4124-84e8-5539bce9b0e8 · outbound

This paper cites Asymmetric actor critic for image-based robot learning.

MuJoCo Playground Asymmetric actor critic for image-based robot learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.024184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.096988Z digest=sha256:6d80746f8359c1cacd711333c261b115eb7d7d90f3eeab7e70bd90ba1e17d6d3

Observation 62910e05-e945-40be-9c07-91d93d404f63 · outbound

This paper cites Learning Humanoid Locomotion over Challenging Terrain.

MuJoCo Playground Learning Humanoid Locomotion over Challenging Terrain

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.104616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.104616Z digest=sha256:d4b37912487520f29a5feeab41c7167c871a9f56d88b571e68d70d293041eb3d

Observation 89b51cdc-e91d-4853-9396-da97800cd9d7 · outbound

This paper cites Searching for Activation Functions.

MuJoCo Playground Searching for Activation Functions

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.111293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.111293Z digest=sha256:dd94d1a6d17f53a5a4c3ed4d41e6bc0e1dfc781f65a7a4108cf97c7fe197bf1e

Observation 1cc1ce82-6809-4784-a0e3-627ae8ea2de0 · outbound

This paper cites High-throughput batch rendering for embodied ai.

MuJoCo Playground High-throughput batch rendering for embodied ai

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.995268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.123334Z digest=sha256:62436b6a36e586c4227c7d1c8860fa6e0aaa22e04f1eda9a5a65c8efdb8c3f92

Observation d967741b-536c-448f-b46a-e25b5a198219 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning.

MuJoCo Playground Learning to walk in minutes using massively parallel deep reinforcement learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.128925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.128925Z digest=sha256:ce3d62e794c9203dccb6d17a22157b2c15dda9a1e3125c147ca1d21c948fd16f

Observation d4ce7748-f9da-437f-985d-9a5e305e3d29 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MuJoCo Playground Proximal Policy Optimization Algorithms

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.134894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.134894Z digest=sha256:cf325c2d00ce3934685514c06845f6a6c74c3f7c43c580574242515248aac584

Observation 68850074-924c-4f0a-9018-c0e978992cdc · outbound

This paper cites Humanoidbench: Simulated humanoid benchmark for whole-body locomo- tion and manipulation.

MuJoCo Playground Humanoidbench: Simulated humanoid benchmark for whole-body locomo- tion and manipulation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.949895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.140354Z digest=sha256:28e334c30d3da1ed9cea6c43437184b5a65816c243f991851b72ee0032fe0c1b

Observation 0c85de97-6627-4729-8e90-c7121be3a591 · outbound

This paper cites An ex- tensible, data-oriented architecture for high-performance, many-world simulation.

MuJoCo Playground An ex- tensible, data-oriented architecture for high-performance, many-world simulation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.920513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.145147Z digest=sha256:2cd9b6f4ae5c41c1ea750829178d50d4422e59bacf031da6a8a08fd4b6897054

Observation b5aacce7-50bf-4058-8771-404c27fccc03 · outbound

This paper cites An ex- tensible, data-oriented architecture for high-performance, many-world simulation.

MuJoCo Playground An ex- tensible, data-oriented architecture for high-performance, many-world simulation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.884234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.150542Z digest=sha256:297b1817477b98fbfa98cab3892ed3504d6bc0246318e453687104b85bc7aa59

Observation 91b2ba84-6fb0-47be-aa3b-2c15cabe3519 · outbound

This paper cites Learning free gait tran- sition for quadruped robots via phase-guided controller.

MuJoCo Playground Learning free gait tran- sition for quadruped robots via phase-guided controller

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.845238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.156529Z digest=sha256:7f3b02fd61ccde29e7ea9def92bce6b10d15a32d17182db68b512fb605eefac8

Observation a93b3d01-da96-4cf4-b84d-c063c2b89d7e · outbound

This paper cites Leap hand: Low-cost, efficient, and anthropomorphic hand for robot learning.

MuJoCo Playground Leap hand: Low-cost, efficient, and anthropomorphic hand for robot learning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.817950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.161472Z digest=sha256:c840ecb89cd790da6ef81a97772c27bc7332f55f67720b60eca00bd86b283e8a

Observation 62ce00af-ffb7-45ac-99e0-340b3e07cad3 · outbound

This paper cites DextrAH-RGB: Visuomotor Policies to Grasp Anything with Dexterous Hands.

MuJoCo Playground DextrAH-RGB: Visuomotor Policies to Grasp Anything with Dexterous Hands

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.166853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.166853Z digest=sha256:cad52501e1ecff7ac72b59ce245440da8011ceb2608eb56a22c63ebc665dc4bd

Observation 43b14fb3-bd20-4552-b533-ea9c678aeb52 · outbound

This paper cites Legged robots that keep on learning: Fine-tuning locomotion policies in the real world.

MuJoCo Playground Legged robots that keep on learning: Fine-tuning locomotion policies in the real world

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.773652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.172371Z digest=sha256:6090df012be21a22449b082101e8a1f5b75c40edecf9581c43dcc6dfa16899ff

Observation 410bc3d4-9c5d-4ee7-87e6-f1a7aa67bd98 · outbound

This paper cites Sim-to-Real: Learning Agile Locomotion For Quadruped Robots.

MuJoCo Playground Sim-to-Real: Learning Agile Locomotion For Quadruped Robots

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.177602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.177602Z digest=sha256:08b0ad2204e4b854428ba3bb72ab6096548307c358da0e4e67311960fed3868f

Observation b2411a8b-6e36-4ac3-ba0a-703c1103808b · outbound

This paper cites ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI.

MuJoCo Playground ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.182563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.182563Z digest=sha256:94282ca0b12f674b5c4ef98a3a73cd266f4d87e08f7bf31acd226f8758449666

Observation cf10259e-ff76-491c-bc8b-5636cf06afc0 · outbound

This paper cites DeepMind Control Suite.

MuJoCo Playground DeepMind Control Suite

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.187882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.187882Z digest=sha256:4d207200a1f5b884cde5037bed557d4843e9d2b5a0fee69bb4b56a6fb0564d30

Observation e335c1fd-2f2d-408d-8b1d-807a1bd3436a · outbound

This paper cites Domain ran- domization for transferring deep neural networks from simulation to the real world.

MuJoCo Playground Domain ran- domization for transferring deep neural networks from simulation to the real world

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.735348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.194861Z digest=sha256:39ba6cb50f2f84347cf543faff74d559ba6a4930fd48545c862af9efc31c91ff

Observation fae85ff3-2bc8-4285-9a0b-80c8f8087eed · outbound

This paper cites Mujoco: A physics engine for model-based control.

MuJoCo Playground Mujoco: A physics engine for model-based control

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.203907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.203907Z digest=sha256:5f301273b21986c547b643c701345c06e9de031b6d8bbd3af4c7dac433f216ab

Observation add7841a-ae70-4bd3-9f32-7480d8d99c7c · outbound

This paper cites EfficientZero V2: Mastering Discrete and Continuous Control with Limited Data.

MuJoCo Playground EfficientZero V2: Mastering Discrete and Continuous Control with Limited Data

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.209457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.209457Z digest=sha256:1430458e4271183b725d1bfdeb15544d30c91d6ac271161225b02934654f6b42

Observation 29bbe30d-1116-4efb-a6e7-c071fa207b46 · outbound

This paper cites Bench- marking the performance and energy efficiency of ai accelerators for ai training.

MuJoCo Playground Bench- marking the performance and energy efficiency of ai accelerators for ai training

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.685171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.215677Z digest=sha256:2dfa7f0055cfece314f76c5b5414677ab95d98530937c5d46b728abdb1bf0988

Observation 5fcdcdeb-4b7c-4e43-acbc-ed90631002d3 · outbound

This paper cites Full-Order Sampling-Based MPC for Torque-Level Locomotion Control via Diffusion-Style Annealing.

MuJoCo Playground Full-Order Sampling-Based MPC for Torque-Level Locomotion Control via Diffusion-Style Annealing

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.222719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.222719Z digest=sha256:fa943acba41cc4c559f61fc6257e52ecde2fba40c60f003299a17c1c86769fef

Observation 716c539f-0f4d-4404-8eb0-acbfdbc9227e · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

MuJoCo Playground Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.228893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.228893Z digest=sha256:cf381317eac4843edeae18187b7bd55b3446c5e58c4c065cf5e6ccb6c7b61306

Observation 50f51a09-192b-4464-bc76-ed84f699daa4 · outbound

This paper cites MuJoCo Menagerie: A collection of high- quality simulation models for MuJoCo, 2022.

MuJoCo Playground MuJoCo Menagerie: A collection of high- quality simulation models for MuJoCo, 2022

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.234964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.234964Z digest=sha256:a3557f06562113b7c276a840b9eeb8911053397882c2754aa65f46af6305b249

Observation 97cd2f7b-df40-4362-833c-8e3cce22a425 · outbound

This paper cites Sim-to-real transfer in deep reinforcement learning for robotics: a survey.

MuJoCo Playground Sim-to-real transfer in deep reinforcement learning for robotics: a survey

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.633904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.240472Z digest=sha256:1351b6e3ad2ee005e229b51cce5a63474d703598697d1900bd6ac573b1d045db

Observation 85620289-00ba-4918-91ea-30ab4a2728b5 · outbound

This paper cites Robot Parkour Learning.

MuJoCo Playground Robot Parkour Learning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.246073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.246073Z digest=sha256:3765c070b269e9200df89608469f908014f2505896fa708501b14ecb732cb9c9

Observation 2c91a006-ea14-4b75-94e6-6735b5153476 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.601436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.252413Z digest=sha256:ad07632d5b0d3fa3a8ad46100c19d50b054ac30d17489a64490bf0d41d65182c

Observation 74ea8abe-f4e1-4889-a458-205b0991da61 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.567978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.257818Z digest=sha256:eb63b16baaafab58d1ef270270143248b5f2a0d66919b7d7b9054dbcef88bc8d

Observation a299a52c-802e-4844-bd09-bcc20087fc13 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.541681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.264364Z digest=sha256:fcbfd8ab55ac131bb9c78a423345750d8a50090f3690e6915ba3ca4b02ddcdda

Observation 11bcb761-4e63-49f6-8a25-45a9a596f140 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.513319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.271284Z digest=sha256:b3434cb9521f834ab153ed1b7622e028d37d7aaa25be1c1052017337aab2dc58

Observation c7228853-32c7-40ee-8b33-42669d190ced · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.485281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.277260Z digest=sha256:f5f47ac46b0f772a2429e01720895153abed9699534e6e791189f9e702c5908e

Observation 440a8768-6154-4064-80f7-d4f9fe137af2 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.455345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.283456Z digest=sha256:e3397e311115dc7bc1c7195d275cc56a0b7cc841bfba5c630168b75fa2f694fe

Observation 9eece729-f9fe-41d6-a135-d2d0c65524db · outbound

This paper cites injections.

MuJoCo Playground injections

Reference 78

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T23:34:12.425243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T23:34:11.289646Z digest=sha256:a42a358f60224997497caab8f1e272bdb6a24c913d1502b1fd40541b869e2a61

Observation bda7f170-6c68-4735-9cdd-5a0f194b7468 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.945235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.945235Z digest=sha256:30c3f90766d5f58518091ea50412a1b033b52d921823b459a139a3254ce54bbf

Pith citing papers

Observation 4c1ade24-0ab4-4d65-9264-0ef667cb2976 · inbound

Unreal Robotics Lab: A High-Fidelity Robotics Simulator with Advanced Physics and Rendering cites this paper.

Unreal Robotics Lab: A High-Fidelity Robotics Simulator with Advanced Physics and Rendering MuJoCo Playground

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:51:57.332577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T18:50:57.738602Z digest=sha256:f6baa74b08251c7b77d4e3608c189da3c832da49126783b63aa3b8fce43f213c

Observation aa772c30-5458-4448-8183-e2a2046b1e05 · inbound

MOSAIC: Skill-Centric Manipulation Planning with Physics Simulation cites this paper.

MOSAIC: Skill-Centric Manipulation Planning with Physics Simulation MuJoCo Playground

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:53.054525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:53.054525Z digest=sha256:1194fa2c4258f4c4ee5181b8027fb6a7070d2f0599f1122a5927e4dadd5665ed

Observation b164d273-89f8-49aa-92c1-55c0e14cf947 · inbound

Visual Imitation Enables Contextual Humanoid Control cites this paper.

Visual Imitation Enables Contextual Humanoid Control MuJoCo Playground

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T23:48:57.968507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:48:57.968507Z digest=sha256:7d3ec69afcd9bfb1c14bee4ac09d38a6cc17bdd2823e1677215dfe2d9ed48cc4

Observation 0184d00b-5bf5-4fab-ac08-c62caccdaf01 · inbound

Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control cites this paper.

Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control MuJoCo Playground

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:45:35.064770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:45:35.064770Z digest=sha256:f7522ec1b6b3195e991b2ddeb7d05d8a38bb7131ac9bddaf09129fabafe01919

Observation 61157ede-b9ad-44e5-b8be-61cee418ef4f · inbound

SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics cites this paper.

SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics MuJoCo Playground

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:31.382654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:31.382654Z digest=sha256:93dde404718f74e324c95157d2f151669fab837f83b4a910a5daa326248a6879

Observation 2591ff17-1047-4344-9c37-9a96b3937f3d · inbound

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control cites this paper.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control MuJoCo Playground

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.735651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.735651Z digest=sha256:98d1da8896130044a397db093340fb1f1ec5a65d664ce4996a6fa2a28e55467f

Observation 45f4eb52-46d9-4505-93a4-632df8cc7b34 · inbound

Booster Gym: An End-to-End Reinforcement Learning Framework for Humanoid Robot Locomotion cites this paper.

Booster Gym: An End-to-End Reinforcement Learning Framework for Humanoid Robot Locomotion MuJoCo Playground

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T19:45:00.972627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:45:00.972627Z digest=sha256:1dd49b7d4ff4046230ed5a8a1373630ad5f8e376059b9a947e4aa6e6d9029f19

Observation 0e5d95f8-c6c6-4f32-9034-9d550f6f1e88 · inbound

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training cites this paper.

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training MuJoCo Playground

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:06.275060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:06.275060Z digest=sha256:73bfd9c218383cc96e7b7169db93faa2019a61b936dc6dfe643d7235a7fef9d7

Observation 91c46414-c837-4d29-8744-b3125bb80fa6 · inbound

Flow Matching Policy Gradients cites this paper.

Flow Matching Policy Gradients MuJoCo Playground

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:09.516529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:09.516529Z digest=sha256:49ab8c0b334b54ef9eefb4e7e15d7b2385777676948fce29e7d58db7488c3993

Observation ae6cdf95-ba5c-4a97-9682-e5ab0e3b39c9 · inbound

Viser: Imperative, Web-based 3D Visualization in Python cites this paper.

Viser: Imperative, Web-based 3D Visualization in Python MuJoCo Playground

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T11:13:12.409367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:13:12.409367Z digest=sha256:6de53986b164895d540922ca31590d09b9e98dcfa7b3d71b4e125420bd10bd21

Observation 1915220b-0f3f-4a41-b915-29c425c5bd3e · inbound

Simultaneous Contact Sequence and Patch Planning for Dynamic Locomotion cites this paper.

Simultaneous Contact Sequence and Patch Planning for Dynamic Locomotion MuJoCo Playground

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:21:25.270760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:21:25.270760Z digest=sha256:cde0f6bec3a806cf720a66190524a31cf64008177d57f523cff804f751dce395

Observation 7717e797-ead4-4540-a0a2-a07d8add70f1 · inbound

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges cites this paper.

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges MuJoCo Playground

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-05T16:55:52.309377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:55:52.309377Z digest=sha256:4aeb0cfab42dd683cea1b93499ef497b7fcc2473b6aa167aed69af5e3040548d

Observation 6f27ce27-967c-4cec-812d-5bf4fb298b3c · inbound

RecoWorld: Building Simulated Environments for Agentic Recommender Systems cites this paper.

RecoWorld: Building Simulated Environments for Agentic Recommender Systems MuJoCo Playground

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T17:56:38.873029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:56:38.873029Z digest=sha256:04be5bfe3d5367ea4f46031e4dd16a27e0ed4d1be0d55d6bdf395bc59558343b

Observation 3ae413aa-1f6b-4178-af74-66361ca0ea93 · inbound

MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning cites this paper.

MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning MuJoCo Playground

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T22:58:38.900601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:58:38.900601Z digest=sha256:f266079f0bc69ba83051dcf5526bfcd8a3129ce862e9dc4ef04882c4d6b919f9

Observation e0fe0836-7a91-45e0-b36e-41a57adab712 · inbound

What Matters for Simulation to Online Reinforcement Learning on Real Robots cites this paper.

What Matters for Simulation to Online Reinforcement Learning on Real Robots MuJoCo Playground

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T21:36:15.812654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:36:15.812654Z digest=sha256:958168a32e8e6a5a29c33f964ad706ae494db587ce8f7b0b2da4f942ef16d458

Observation 4aaa48ef-34b2-4e8b-beba-ff37f25dc449 · inbound

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation cites this paper.

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation MuJoCo Playground

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:36:17.295232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T16:35:34.179788Z digest=sha256:3deff4bc79e9012bde8490d7b4d2e3727203295129be5bbacda2553b2f955ac1

Observation f5fd9519-96ff-4027-9d3f-7769c260314c · inbound

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation cites this paper.

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation MuJoCo Playground

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T18:52:42.535173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:52:42.535173Z digest=sha256:0c1c82937e103e607e75dc793047ac7731faf8685040d5d8f692879743041848

Observation febfeef0-bb3b-4482-8c81-822975be3825 · inbound

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control cites this paper.

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control MuJoCo Playground

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:20:00.675540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T12:17:59.392158Z digest=sha256:107d7cf7b159fa3769169205bdd9b2b9dea3fecd214b15d367d9d3353c3dfdf8

Observation 04a3f884-5279-4807-b3ed-50095a520332 · inbound

Research Novelty in Information Systems Journals After ChatGPT: Differences Across Institutional Language Contexts cites this paper.

Research Novelty in Information Systems Journals After ChatGPT: Differences Across Institutional Language Contexts MuJoCo Playground

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-15T11:52:28.028968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:52:28.028968Z digest=sha256:7d634f69ea1f0f1fdcaafc3208f50d46b4054ced9a4c50c356399681a7ef320c

Observation be7679be-271e-4c71-9b7d-e472d73b5518 · inbound

Learning Dexterous Grasping from Sparse Taxonomy Guidance cites this paper.

Learning Dexterous Grasping from Sparse Taxonomy Guidance MuJoCo Playground

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:00.984225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T17:05:06.300013Z digest=sha256:06e2d9e7632d667b95a3cfde87418e961997bfea95d3cd74d472796224a4c0c0

Observation 891402cb-515a-4e4d-8899-39c91b272998 · inbound

On Data Thinning for Model Validation in Small Area Estimation cites this paper.

On Data Thinning for Model Validation in Small Area Estimation MuJoCo Playground

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T11:14:37.619203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:14:37.619203Z digest=sha256:0910e0b6f3a0b3e59e9620f7cf0f33e1991788ceb6bc8cd4d082e6b3f3e1728f

Observation a2c1e83c-5c0f-4d91-887b-3a0aa009567f · inbound

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control cites this paper.

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control MuJoCo Playground

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:15:49.708901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T20:04:56.512544Z digest=sha256:cd33c345d708e6efb9a5807bb38f996e7ca4c8fc7a84e6dbf7e439fe238a9073

Observation 627ae881-2b5f-42d0-9613-b815e006808b · inbound

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control cites this paper.

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control MuJoCo Playground

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:12:41.315328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T17:08:31.770889Z digest=sha256:b481002c04740d9ca82a5a34348842f4dd0588fd184a0f17733651bf60861540

Observation d1c0e1be-db11-421f-8535-fa344fdcb0b6 · inbound

Simulation-Driven Evolutionary Motion Parameterization for Contact-Rich Granular Scooping with a Soft Conical Robotic Hand cites this paper.

Simulation-Driven Evolutionary Motion Parameterization for Contact-Rich Granular Scooping with a Soft Conical Robotic Hand MuJoCo Playground

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:15:50.383745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T20:04:06.942355Z digest=sha256:7fc97bc9c0c3dc84b14b96dd48e5bae902da14c928d524e48501fda249493f4a

Observation 7945ad17-eae3-4831-ba49-324b0c9fb612 · inbound

A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies cites this paper.

A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies MuJoCo Playground

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:30:22.763830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T12:29:38.306670Z digest=sha256:99168ccb65777b707c5c7d3d6d70337e5af9eca067ee74be9bb977eee8e2f734

Observation 0b7b07e6-1325-452d-bb3a-3d873f7f2bcf · inbound

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics cites this paper.

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics MuJoCo Playground

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:26:14.865036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T02:45:48.819070Z digest=sha256:2fac126f8b3a2812a63f113e090342eb426edc33cfe50f96d15b899237d4bcdf

Observation 5f5b781d-d0c8-4574-a622-6feae5ff7257 · inbound

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics cites this paper.

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics MuJoCo Playground

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:06:16.506077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T03:27:02.669953Z digest=sha256:4be58f03ac4c72e3b96603b6aef50ad2fbf7af7576e2a18bb98d2dc08e132e96

Observation a1e497f0-6063-4745-85d7-a093f89f590f · inbound

GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning cites this paper.

GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning MuJoCo Playground

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:51:17.041890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T16:12:45.211236Z digest=sha256:170d38cf9b532e4906730f204707284d95d12e28424c51959db81f18dab8489a

Observation 1c47de61-72df-42d8-b304-69c78636cb4f · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders MuJoCo Playground

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:33:04.209666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T05:28:50.354662Z digest=sha256:4b4edceef5b294a43287fd0ed7c46faa89e7b745608e884ef56ea6cb374fa7e4

Observation baaa4e01-c023-44a7-a32e-be4d220f4bc3 · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders MuJoCo Playground

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:39:49.242236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T07:36:12.214949Z digest=sha256:88cc02e300503432e216d639cb18f7535ff7bb2e15dac9bab69c479b14191201

Observation b3607232-2a62-42a6-b2cc-e971d0e1aaf7 · inbound

MuJoCoUni:Persistent Batched Runtime Primitives for MuJoCo cites this paper.

MuJoCoUni:Persistent Batched Runtime Primitives for MuJoCo MuJoCo Playground

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:47.814957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T01:13:54.860724Z digest=sha256:866a9c455120267c8a8ec3122560939df8059c1f84f8c0d50c66ab8e08af07aa

Observation eea67457-2825-414d-ae0e-a716ad6e0727 · inbound

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion cites this paper.

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion MuJoCo Playground

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:05:49.719887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T00:54:12.099045Z digest=sha256:69bbfe9685a72df289f93b65de32c0786bb405c9a8dbe07b5d1d2b00c672eae9

Observation a61681b2-7eac-4a12-8491-f38d87ce9c57 · inbound

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient cites this paper.

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient MuJoCo Playground

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.609847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T17:34:41.053725Z digest=sha256:b37490786a0e6364581724e9672e0af19595935a05b3e8dae218484ef8fcf86b

Observation 274fdd30-bc50-4669-be9d-66a72a1599c5 · inbound

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms cites this paper.

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms MuJoCo Playground

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:13:16.503863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T07:09:19.137932Z digest=sha256:89da5492892c4e97fa214608c7c6ebe54dde5f110dfbab775c74237eed4ba2d6

Observation a4a01c17-dfe4-4ad5-a76a-d480c3cba122 · inbound

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning cites this paper.

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning MuJoCo Playground

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T22:22:43.501943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T22:20:37.637669Z digest=sha256:6a92c41afbe921a9d7c0d8d6689ae5caa1e1fe33c7781d20b43e51466b0ca797

Observation 1c23d7ed-caf9-4752-adf0-0569e307b743 · inbound

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It) cites this paper.

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It) MuJoCo Playground

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:42:38.047426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T18:19:58.499360Z digest=sha256:0414650533c0ee8067d854eec60710edf482dcda9a17eb2e4eb821dbee7efea8

Observation 069d9282-7e8c-46a9-adad-cb43d4e823e2 · inbound

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment cites this paper.

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment MuJoCo Playground

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.277639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T06:20:18.312301Z digest=sha256:ffcbb4f9ee0aba398039a880bdc976e296427d8edfe996f54921c28d0cfee186

Observation 9b571c89-d364-428f-a374-292e2fd19428 · inbound

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation cites this paper.

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation MuJoCo Playground

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.979243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T21:55:15.993627Z digest=sha256:720b446611f1be61f32b8c9b5c45e85ea538abcba418aa555ca99f1be76eadfe

Observation 8316a452-9a63-4786-99de-4c2f667cf8b6 · inbound

Embedding Hybrid Systems into Continuous Latent Vector Fields cites this paper.

Embedding Hybrid Systems into Continuous Latent Vector Fields MuJoCo Playground

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:17:37.141265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T14:04:40.078357Z digest=sha256:f9c658c613e462b87218f34c26b9d1cd1732d39f2aebb0aa3b69713d31df019d

Observation c3193287-4333-42ec-a088-9d66bf7d96e0 · inbound

Blind Dexterous Grasping via Real2Sim2Real Tactile Policy Learning cites this paper.

Blind Dexterous Grasping via Real2Sim2Real Tactile Policy Learning MuJoCo Playground

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:48:02.597110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T09:50:02.635507Z digest=sha256:cd1372fc90ae43faa7e59b0445faabb3a0609ec4b83e47a7d2808c36b1cd3707

Observation 2dfda7d7-0a30-4c9e-af46-c8d1603b103a · inbound

Benchmarking Action Spaces in Reinforcement Learning for Vision-based Robotic Manipulation cites this paper.

Benchmarking Action Spaces in Reinforcement Learning for Vision-based Robotic Manipulation MuJoCo Playground

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:12.994967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T21:22:53.551550Z digest=sha256:539ad17637c9eeec13f9f20800ea2adc1acded459bdb8b9749f6b4775fa337be

Observation f4881233-2679-4885-ad83-f92832e45226 · inbound

Simulating Robotic Locomotion in Sand: Resistive Force Theory in an Open-Source Physics Engine cites this paper.

Simulating Robotic Locomotion in Sand: Resistive Force Theory in an Open-Source Physics Engine MuJoCo Playground

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:59:21.067835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T20:45:44.072738Z digest=sha256:a0f671f5dde2119d7e921dc5c0c102014542032f518b4cb49a23fa06aed9b0bd

Observation 73b9c8a3-f9b9-4561-8e6c-cbb6e6c8f9ec · inbound

CRAX: Fast Safe Reinforcement Learning Benchmarking cites this paper.

CRAX: Fast Safe Reinforcement Learning Benchmarking MuJoCo Playground

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:09:30.117010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T18:22:36.891186Z digest=sha256:900f441c092fedd213b45b86e92546e4a387a778a25d92c0ae2b0276ba19d627

Observation f8da729b-4043-4834-9878-f75dc173af1d · inbound

ReFPO: Reflow Regularization for Flow Matching Policy Gradients cites this paper.

ReFPO: Reflow Regularization for Flow Matching Policy Gradients MuJoCo Playground

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:19:37.785912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T14:38:54.754049Z digest=sha256:a13a8d32e948f3fe6feae0fb9b29311c7a782c757c89dc6f5776ae30dbfe0952

Observation 4c37d03e-ff90-4e0a-8964-077c453c63b9 · inbound

Long-Distance Real-World Navigation of the Legged-Wheeled Robot Go2-W Using Deep Reinforcement Learning cites this paper.

Long-Distance Real-World Navigation of the Legged-Wheeled Robot Go2-W Using Deep Reinforcement Learning MuJoCo Playground

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T13:59:30.584615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T13:57:34.489168Z digest=sha256:a0371f260872f7daa170362bf22c33b316132f06a0e06760e0f5d24dd831c014

Observation 040babbf-c22e-432f-98bb-4b4687ef0c68 · inbound

Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System cites this paper.

Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System MuJoCo Playground

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-26T11:39:24.547908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T11:39:18.595140Z digest=sha256:e3b32eeabd1850bac8d7e08498fd22ab0d7db08ef7abe75b10f9d8b5c05098be

Observation 021a8f26-846d-4e74-b55d-7fcbdfd40bdc · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering MuJoCo Playground

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.191604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:4ee2422e78c7eeff8b4f4505f05ef1da2737a6b5e12b30825ed936f3dd1d6276

Observation b38a5207-aeea-4be4-b093-599e7d27c379 · inbound

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications cites this paper.

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications MuJoCo Playground

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:06:55.353101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-02T11:59:04.103002Z digest=sha256:30716d039d446aaeec51a76308f84a6fe23767d1f94992ecfac0b5b324a09242

Observation 5e963b18-f5ca-43ae-9527-e34373a971c2 · inbound

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space cites this paper.

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space MuJoCo Playground

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T01:30:09.928609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:30:09.928609Z digest=sha256:64f057336aecf16e5fc5f136b277a661bf9a7bec1c0bdd864d1264be7ea3f2b8

Observation a257f697-0539-4b3c-9940-d9204394965c · inbound

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics cites this paper.

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics MuJoCo Playground

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T09:34:47.883931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-08T09:33:49.773169Z digest=sha256:25ee1a3c3026fa15a5cfb8b68a43584e3ccd2b1544a4fca7c732a72d9643fa79

Observation 1e155c6f-39bb-42ab-9841-927edd53210b · inbound

Rethinking the Suitability of Reinforcement Learning Algorithms Under Practical Transfer Constraints cites this paper.

Rethinking the Suitability of Reinforcement Learning Algorithms Under Practical Transfer Constraints MuJoCo Playground

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T18:24:35.875407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:24:35.875407Z digest=sha256:5398135b29a545c17694f30c12bf7598bd4a981d775808350c6427e91cd7e3e7

Observation fd659c5f-e75e-464b-92c9-abe51ca702e2 · inbound

AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation cites this paper.

AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation MuJoCo Playground

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T07:02:19.931906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:02:19.931906Z digest=sha256:7eac0f768c80b773bbdb50c1afff7a90735b38bfe69474a571040f02c78a4ec5

Observation d5868d46-7211-4538-97f8-d5f854ff5616 · inbound

$\pi\mathbf{R}^2$: Reactive Real-time Flow Policies cites this paper.

$\pi\mathbf{R}^2$: Reactive Real-time Flow Policies MuJoCo Playground

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T00:49:36.679003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:49:36.679003Z digest=sha256:0b255cbc0f66fbdcaec9ccab1f8461de453f922a2414aad4b7573a9358f50131

Observation 6fe240df-fd29-4458-bda9-e87a47ac0078 · inbound

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts cites this paper.

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts MuJoCo Playground

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T00:17:16.320668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:17:16.320668Z digest=sha256:cccf0e581791f75adf12b16df589342f1536f670cb71c8828568a58b7e323591

Observation 496ca157-cab4-4c9c-9a02-a3f9daebf011 · inbound

Foundations of Reinforcement Learning and Control:Connections and New Perspectives cites this paper.

Foundations of Reinforcement Learning and Control:Connections and New Perspectives MuJoCo Playground

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-04T07:32:30.427620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:32:30.427620Z digest=sha256:c6854f936df297f48a1491828daaec26356e2b26fa1403ef3abaac222c45f175

Observation 68d00cd2-65a6-4fa3-afa5-14e64dc04d4f · inbound

ATP: Anatomical Torque with Passivity-based Control Framework for Safe Upper-Limb Exoskeleton Assistance cites this paper.

ATP: Anatomical Torque with Passivity-based Control Framework for Safe Upper-Limb Exoskeleton Assistance MuJoCo Playground

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:42:26.860884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:42:26.860884Z digest=sha256:d2f0a2acf47241bbe7973518891274db2de55ef7d55b8e1fb39d73862c88a3a0

Observation 5b8b3ed1-8047-43e6-97c2-5d88df470f2f · inbound

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control cites this paper.

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control MuJoCo Playground

Reference 284

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:47.076898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:47.076898Z digest=sha256:7ed3cdf8e99778defb40df88acaae691db78e4d297d55d609350be5b5a933f9d