Pith. sign in

Paper Citation Record · LEDGER

MuJoCo Playground

As of 20 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 57 inbound Pith citation observations for arXiv:2502.08844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08844 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:34:11.289646Z

measured 135 of 135 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 57 of 57 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:59:53.054525Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved45
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

1
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation fa9376a8-b7cd-439f-85c6-3d5b511d2d18 · outbound

This paper cites Legged locomotion in challenging ter- rains using egocentric vision.

MuJoCo Playground Legged locomotion in challenging ter- rains using egocentric vision

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.929512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.851080Z digest=sha256:05c09b198b207e0c67b684e743ec812ed5246504692e477835617af7ddd768e8

Observation 9ad439b4-ac58-423d-91d5-1bc0d0002a10 · outbound

This paper cites ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation.

MuJoCo Playground ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.856636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.856636Z digest=sha256:098a03dfd382ceb5ce564675584eff32f6394ab4a8a3e2761c77bcb1406cb404

Observation 87b35a59-608e-447d-9477-3e97bb1daa7f · outbound

This paper cites What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study.

MuJoCo Playground What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.862569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.862569Z digest=sha256:2a1d7c0542ace368f14fdc94559fff553c12477819dad6d2b1cf709fb4f0c6de

Observation ed7408b3-39d1-47b7-b262-95c71e0b1096 · outbound

This paper cites Learning dexterous in-hand manipula- tion.

MuJoCo Playground Learning dexterous in-hand manipula- tion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.905619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.867778Z digest=sha256:2bc6d3c0b31beefa91c7de35191e4d69459d619c4a79a141b3a88ae19b991110

Observation 8e80af11-6a3d-4509-abea-b9b58a968bff · outbound

This paper cites JAX: composable transformations of Python+NumPy programs, 2018.

MuJoCo Playground JAX: composable transformations of Python+NumPy programs, 2018

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.872779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.872779Z digest=sha256:87b8e2c3ee2427503690d73f03bf19512bf7a64eec9027a276e4dff5226f5752

Observation f3a609f5-f2c5-49c1-95e3-c4f42d4b6e4d · outbound

This paper cites Barkour: Benchmarking Animal-level Agility with Quadruped Robots.

MuJoCo Playground Barkour: Benchmarking Animal-level Agility with Quadruped Robots

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.878775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.878775Z digest=sha256:c33300f62742da867e4c851d08f79544ae569aaa0741615dd0b080287f73223c

Observation 24bf15bb-17de-44c7-bc89-2c05396331e7 · outbound

This paper cites Closing the sim-to-real loop: Adapting simulation randomization with real world experience.

MuJoCo Playground Closing the sim-to-real loop: Adapting simulation randomization with real world experience

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.884603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.884603Z digest=sha256:a1960b009beb74586a9d9e16fb4555c3ed0418a3f3ae2b9dcd558307ae0122a3

Observation 17b15cd7-c543-46b4-94f7-d923ed44f1ee · outbound

This paper cites A system for general in-hand object re-orientation.

MuJoCo Playground A system for general in-hand object re-orientation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.889487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.889487Z digest=sha256:391797245be2f35203025f3616bbdb33b6b2f82bcd81dd7f143e75aabfe6ee15

Observation e533a724-fecc-4274-ac44-ed0cb95a89ff · outbound

This paper cites Extreme parkour with legged robots.

MuJoCo Playground Extreme parkour with legged robots

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.759864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.894221Z digest=sha256:363fb248dcee10c6b7ea1508e341f50b90e3f44701df67cfd1f3730e21473996

Observation dd315f5e-2687-4a76-a2cc-8cefcdfb4714 · outbound

This paper cites CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects.

MuJoCo Playground CORN: Contact-based Object Representation for Nonprehensile Manipulation of General Unseen Objects

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.899764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.899764Z digest=sha256:b51763ae8ab6801588bd00ec039c6fc69df74e50be59fad2cc22fc181492f787

Observation e2690e60-8b56-4e79-bb01-4e6fb3a9001f · outbound

This paper cites Onnx runtime.

MuJoCo Playground Onnx runtime

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.727976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.905027Z digest=sha256:e2de4b0eddd69f0e00d8aa1db99736579ec9aaf978d0e5bcfe478ba88291462f

Observation fae33878-d52d-4882-96ec-920ca0513b62 · outbound

This paper cites Flayols, A.

MuJoCo Playground Flayols, A

Reference 12

Resolution
malformed identifier
no resolver link, observed 2026-08-07T23:34:10.910360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.910360Z digest=sha256:82b30d1db152a0430ada51536a79e48fe432b0dab192cb55ac8a614360861030

Observation 5799c609-edaf-4daf-b68e-38df5f53f7a2 · outbound

This paper cites Brax-a differentiable physics engine for large scale rigid body simulation, 2021.

MuJoCo Playground Brax-a differentiable physics engine for large scale rigid body simulation, 2021

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.688721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.915072Z digest=sha256:327b8a4f3b2fdcb96cf95a7fdf20b04a801727beed63c9f960853b0610e7df55

Observation d770e87e-6209-4ab7-9464-3816a08ddcdd · outbound

This paper cites Genesis: A universal and generative physics engine for robotics and beyond, December 2024.

MuJoCo Playground Genesis: A universal and generative physics engine for robotics and beyond, December 2024

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.657010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.919972Z digest=sha256:bf25a2b821a468353f4248b4b1d5b8d2142e3817517a434af603b70e7b029e82

Observation 8d763b7c-4a47-48cb-943c-c8f118751cfc · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

MuJoCo Playground Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.628768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.925154Z digest=sha256:cd75577ba46b361ac9fda42293d7cf59de09be0ab51943d9c62362fd20256922

Observation e3378a3d-945a-4fb8-84fa-d0d77522e814 · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning.

MuJoCo Playground Learning agile soccer skills for a bipedal robot with deep reinforcement learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.929790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.929790Z digest=sha256:f4db052e22634415889daebef1a7370a12f5afc06afdcda98ded49da1be3bd70

Observation 691c9860-bb41-4d12-846a-d93bd0d631a3 · outbound

This paper cites Mastering Diverse Domains through World Models.

MuJoCo Playground Mastering Diverse Domains through World Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.934582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.934582Z digest=sha256:d1e3108aadc0c2cddfee615f5982f3fdb17f83cbbff4bbf40a52490b7db3249f

Observation 935dd99f-a9aa-47e8-93b2-5979c989e479 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:13.560878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.940144Z digest=sha256:bc110f48c411d036f66400e71612cd868a6c73d01ff59e1ce9b777d2b3a6a6a0

Observation 44b7610a-505e-466f-b7f6-4ecff143359b · outbound

This paper cites Dextreme: Transfer of agile in-hand manipulation from simulation to reality.

MuJoCo Playground Dextreme: Transfer of agile in-hand manipulation from simulation to reality

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.530876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.950291Z digest=sha256:796d92201ff65208c07e31197312659567a18684c50e9b57532bf4a6e31aad6f

Observation d10f5a5c-74d9-4f77-a4a2-f69e9a424e1f · outbound

This paper cites Td-mpc2: Scalable, robust world models for continuous control, 2024.

MuJoCo Playground Td-mpc2: Scalable, robust world models for continuous control, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.955204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.955204Z digest=sha256:22a523517ec6aefcb3a70ccc992f972a96605a1a573cbfb7c8db253c33f24fac

Observation bb4627f3-50fb-46a1-a884-8e8230f94f1c · outbound

This paper cites Analytical inverse kinematics for franka emika panda – a geometrical solver for 7- dof manipulators with unconventional design.

MuJoCo Playground Analytical inverse kinematics for franka emika panda – a geometrical solver for 7- dof manipulators with unconventional design

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.960123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.960123Z digest=sha256:99b095762b197504845394800cb0a2b1fde0dd3875de0864d6eb4494669d530b

Observation 0054a8c2-9ca3-43aa-9d17-62dfd2a3e038 · outbound

This paper cites Evolving control: Evolved high frequency control for continuous control tasks.

MuJoCo Playground Evolving control: Evolved high frequency control for continuous control tasks

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.473859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.965210Z digest=sha256:92f97e46c24669c1f6374241eef11c9a75d3388121efd53f473c368011d19979

Observation f6ec3702-f13b-4b5f-884f-34de4b0b3146 · outbound

This paper cites DiffTaichi: Differentiable Programming for Physical Simulation.

MuJoCo Playground DiffTaichi: Differentiable Programming for Physical Simulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.970819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.970819Z digest=sha256:966b69aeb132c5fbdd1d0d29d5fdddd19498797bcc57f2d42d898ea5913fd7e6

Observation 00660834-3599-431c-8219-4fee5f345484 · outbound

This paper cites How to train your robot with deep reinforcement learning: lessons we have learned.

MuJoCo Playground How to train your robot with deep reinforcement learning: lessons we have learned

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.440821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.976572Z digest=sha256:4aa8c283b3dcd5268eb71fb2a170f6af7e0b15c5ae782ab1c3e0c781fa9c4327

Observation e5fd55ac-48d6-4f5c-9e10-6e2ab527c8e2 · outbound

This paper cites Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion.

MuJoCo Playground Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion

Reference 25

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T23:34:11.854545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.981630Z digest=sha256:0467e32da3c6d7b643cc891edeb2baa959176e0812eb8c967b65c47d1606cfd2

Observation 14d9023e-45df-48bc-9e6b-17eeaa399a5d · outbound

This paper cites Champion-level drone racing using deep rein- forcement learning.

MuJoCo Playground Champion-level drone racing using deep rein- forcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.986346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.986346Z digest=sha256:dc56a4191de116401d50519934227ff2030f02d0a7a8eb8c72ff4705cd108d92

Observation ceb68d23-e1bf-4d9c-af27-0f8a500cd4d2 · outbound

This paper cites Reinforce- ment learning in robotics: A survey.

MuJoCo Playground Reinforce- ment learning in robotics: A survey

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.385541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.993062Z digest=sha256:9c5712cf8493369962fc8747aec8cba6ae149367423d563b91589ffbea5339fe

Observation d10424d4-203a-4b74-81e2-8ccf7f0bf569 · outbound

This paper cites Design and use paradigms for gazebo, an open-source multi-robot sim- ulator.

MuJoCo Playground Design and use paradigms for gazebo, an open-source multi-robot sim- ulator

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.345296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:10.999033Z digest=sha256:5cd4d8626bc306e0143e7d4c3a216ef4c50ad89a2029eb930277e56e617cbf19

Observation 77b33af0-387c-40b6-9719-ece1c4f4f707 · outbound

This paper cites Reinforcement Learning with Augmented Data.

MuJoCo Playground Reinforcement Learning with Augmented Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.003868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.003868Z digest=sha256:a1d972939520150bd94394b61503149dc6b1814abe036c17dbab7394be67698f

Observation 9570e471-2e99-4075-aa0a-070f3178745d · outbound

This paper cites Robust Recovery Controller for a Quadrupedal Robot using Deep Reinforcement Learning.

MuJoCo Playground Robust Recovery Controller for a Quadrupedal Robot using Deep Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.011005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.011005Z digest=sha256:db8040ddbd2766850dd32a845e4ad789453717fa789035cd81fad2f4734f08e9

Observation 66826b4e-17cc-4718-baca-d9370cc49f35 · outbound

This paper cites rsl rl: Fast and simple implementation of rl algorithms, designed to run fully on gpu.

MuJoCo Playground rsl rl: Fast and simple implementation of rl algorithms, designed to run fully on gpu

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.316051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.016884Z digest=sha256:2695df9f77ec9110b05d0688b50bb4a811a1bd3732fbbe886a757af24c733a38

Observation ab58bfc9-c97b-4375-a64d-44d3782ff0b3 · outbound

This paper cites DROP: Dexterous Reorientation via Online Planning.

MuJoCo Playground DROP: Dexterous Reorientation via Online Planning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.021849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.021849Z digest=sha256:21cf1e488563a6219367308d5c7356156027441b50acd38670463e840bf434c8

Observation e8ceb63d-94d0-4b3b-b425-ab2393557214 · outbound

This paper cites Rein- forcement learning for versatile, dynamic, and robust bipedal locomotion control.

MuJoCo Playground Rein- forcement learning for versatile, dynamic, and robust bipedal locomotion control

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.282664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.027334Z digest=sha256:01ce623e5b4f06b96c27b2c479472e0d185032716163a5d0527e75a51dd8d2d7

Observation 0644c9fd-46ba-454d-a4b8-0025b40929fd · outbound

This paper cites Gpu- accelerated robotic simulation for distributed reinforce- ment learning.

MuJoCo Playground Gpu- accelerated robotic simulation for distributed reinforce- ment learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.253941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.033633Z digest=sha256:285a5da6028712a85bcb8794a2c23c6f0b9d172156ac51992ebae8e013ff5dfa

Observation f5cc8e5d-fe0a-4a6c-bafa-f7850f672064 · outbound

This paper cites Berkeley Humanoid: A Research Platform for Learning-based Control.

MuJoCo Playground Berkeley Humanoid: A Research Platform for Learning-based Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.038539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.038539Z digest=sha256:bda85f868e7dc586fb97699e8a4995cb770454f541b75b0a6011d590f0811402

Observation b655bd5c-76a3-40d5-94ec-4926230a15fb · outbound

This paper cites Learning Humanoid Locomotion with Perceptive Internal Model.

MuJoCo Playground Learning Humanoid Locomotion with Perceptive Internal Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.043978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.043978Z digest=sha256:0a1e4a0162e25eabe64e796d7b45fbb7537b27a47c0ef58afa595eb264e74d82

Observation 342d6944-cba7-4d91-b2d8-7df0d0d15643 · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

MuJoCo Playground Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.049309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.049309Z digest=sha256:bb57ceb00d7f63c74ca5b7f3d16b30d36153ac4adfb6a331d032abe57e84c56f

Observation a4ba11e7-350f-463b-91fa-866a32e4ff96 · outbound

This paper cites Warp: A high-performance python frame- work for gpu simulation and graphics.

MuJoCo Playground Warp: A high-performance python frame- work for gpu simulation and graphics

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.218346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.054659Z digest=sha256:fc413c524b325a7beb68d6b812504b5e93e302281f9f674591e82b06f21468d5

Observation 8c0cf1ea-4b0e-409e-8339-b5dd660cd4b5 · outbound

This paper cites Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning.

MuJoCo Playground Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.060494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.060494Z digest=sha256:063bfc29552378bcca822983758e91db7b29d90aaa75177b20b087bfa89b0605

Observation 5f108b7e-6dcd-4bb8-9d1f-c4314e2c852e · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild.

MuJoCo Playground Learning robust perceptive locomotion for quadrupedal robots in the wild

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.065799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.065799Z digest=sha256:32142bf963061d330e4e8cb0c34f147e348089e6830bff644e8af91ce17d6afb

Observation 906a61e5-0dc5-4ad2-9e2e-4b71a46d304e · outbound

This paper cites Orbit: A unified simulation framework for interactive robot learning envi- ronments.

MuJoCo Playground Orbit: A unified simulation framework for interactive robot learning envi- ronments

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.152950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.070919Z digest=sha256:afe758f2b864c4c9141f3ff12e44a5177dc6f07a214fda0d9f64b4e4a592edab

Observation 6573d4d5-0bc9-4911-a552-8f4f3d1d9797 · outbound

This paper cites Rusu, Joel Veness, Marc G.

MuJoCo Playground Rusu, Joel Veness, Marc G

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.075894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.075894Z digest=sha256:fdfa5e004b5faa37e00559c525785964e0a1c739be5a078dc9d7a48918681b24

Observation 418792c6-3c60-4e88-b1e2-c50e0fbc594b · outbound

This paper cites MuJoCo XLA (MJX).

MuJoCo Playground MuJoCo XLA (MJX)

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.128171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.080924Z digest=sha256:42211167e6267f1e9f4809639a82fafd9e146ab808671047a7086b607eaceaf1

Observation 1062fdd3-caee-42ab-8ebe-7424b0b1ceaf · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:13.094626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.085861Z digest=sha256:ea00c97d9c1207a5f6670dce386fba3fbd75750cad4378f422e67c38c9fa7813

Observation c2137486-9195-4059-b5be-fd090b836d89 · outbound

This paper cites Dexpbt: Scaling up dexterous manipulation for hand-arm systems with pop- ulation based training.

MuJoCo Playground Dexpbt: Scaling up dexterous manipulation for hand-arm systems with pop- ulation based training

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.054620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.091829Z digest=sha256:faa1085c9f44d356786692a28377421760729f3989da9d90cd4344c849ab9dbf

Observation 2965fe43-a207-4124-84e8-5539bce9b0e8 · outbound

This paper cites Asymmetric actor critic for image-based robot learning.

MuJoCo Playground Asymmetric actor critic for image-based robot learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:13.024184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.096988Z digest=sha256:3f1fa3de01c89002e97cca1cc95be38d841ebc29e8b32bc546cd2f3e912985f6

Observation 62910e05-e945-40be-9c07-91d93d404f63 · outbound

This paper cites Learning Humanoid Locomotion over Challenging Terrain.

MuJoCo Playground Learning Humanoid Locomotion over Challenging Terrain

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.104616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.104616Z digest=sha256:9959966d2fa70fbbbe8ba6f4441e48518a56092aaade97dd596dbbb1c0bb7df2

Observation 89b51cdc-e91d-4853-9396-da97800cd9d7 · outbound

This paper cites Searching for Activation Functions.

MuJoCo Playground Searching for Activation Functions

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.111293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.111293Z digest=sha256:aee08022cc04085a3e2766613c96a81693eff60d95c6bc5db8d0649945cb6dc5

Observation 1cc1ce82-6809-4784-a0e3-627ae8ea2de0 · outbound

This paper cites High-throughput batch rendering for embodied ai.

MuJoCo Playground High-throughput batch rendering for embodied ai

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.995268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.123334Z digest=sha256:a87c9d59cf511aa86f79ef13bed4552dcc8b13b5686c60c62c2009f4aba4b566

Observation d967741b-536c-448f-b46a-e25b5a198219 · outbound

This paper cites Learning to walk in minutes using massively parallel deep reinforcement learning.

MuJoCo Playground Learning to walk in minutes using massively parallel deep reinforcement learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.128925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.128925Z digest=sha256:75f4029a11492c5bf505538872de759b17ccc21f69093d8eac2b0fbf0d5a87df

Observation d4ce7748-f9da-437f-985d-9a5e305e3d29 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MuJoCo Playground Proximal Policy Optimization Algorithms

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.134894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.134894Z digest=sha256:3d2d484997f2d3b618bf24094f9405c2901e67f259b8a058c6c616fb06fa9299

Observation 68850074-924c-4f0a-9018-c0e978992cdc · outbound

This paper cites Humanoidbench: Simulated humanoid benchmark for whole-body locomo- tion and manipulation.

MuJoCo Playground Humanoidbench: Simulated humanoid benchmark for whole-body locomo- tion and manipulation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.949895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.140354Z digest=sha256:850f9edab4746a2a1f7d7c7adc7ec7225bfb8ec7f7b30c7f8e7b828ec43357d7

Observation 0c85de97-6627-4729-8e90-c7121be3a591 · outbound

This paper cites An ex- tensible, data-oriented architecture for high-performance, many-world simulation.

MuJoCo Playground An ex- tensible, data-oriented architecture for high-performance, many-world simulation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.920513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.145147Z digest=sha256:e49c58010bd523edf36681c0c9ae2fcc84ae347b4a0e5d0b15a85c2143a4f94d

Observation b5aacce7-50bf-4058-8771-404c27fccc03 · outbound

This paper cites An ex- tensible, data-oriented architecture for high-performance, many-world simulation.

MuJoCo Playground An ex- tensible, data-oriented architecture for high-performance, many-world simulation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.884234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.150542Z digest=sha256:88425e173cfac26dc788061946d16c32b8e8cd74c48139a697b3e37a11b044ed

Observation 91b2ba84-6fb0-47be-aa3b-2c15cabe3519 · outbound

This paper cites Learning free gait tran- sition for quadruped robots via phase-guided controller.

MuJoCo Playground Learning free gait tran- sition for quadruped robots via phase-guided controller

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.845238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.156529Z digest=sha256:f8fde59821b431a76fc4c9288b20316a56c04adef1ce34c63b8292c578bfbe05

Observation a93b3d01-da96-4cf4-b84d-c063c2b89d7e · outbound

This paper cites Leap hand: Low-cost, efficient, and anthropomorphic hand for robot learning.

MuJoCo Playground Leap hand: Low-cost, efficient, and anthropomorphic hand for robot learning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.817950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.161472Z digest=sha256:063b08e39aed01354510824e878fbf63f82e85d609f867d172fe175e8b644ecc

Observation 62ce00af-ffb7-45ac-99e0-340b3e07cad3 · outbound

This paper cites DextrAH-RGB: Visuomotor Policies to Grasp Anything with Dexterous Hands.

MuJoCo Playground DextrAH-RGB: Visuomotor Policies to Grasp Anything with Dexterous Hands

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.166853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.166853Z digest=sha256:93b6c7edf0e72a85bd3849e7350dd757481edefc051ea66c87f4134c5439130f

Observation 43b14fb3-bd20-4552-b533-ea9c678aeb52 · outbound

This paper cites Legged robots that keep on learning: Fine-tuning locomotion policies in the real world.

MuJoCo Playground Legged robots that keep on learning: Fine-tuning locomotion policies in the real world

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.773652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.172371Z digest=sha256:193799a1463bba6713ca76805f86fcd13e0ffb7e3038fa456cc15547ba450748

Observation 410bc3d4-9c5d-4ee7-87e6-f1a7aa67bd98 · outbound

This paper cites Sim-to-Real: Learning Agile Locomotion For Quadruped Robots.

MuJoCo Playground Sim-to-Real: Learning Agile Locomotion For Quadruped Robots

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.177602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.177602Z digest=sha256:5730507713a8dd4bb3ffc4899ab4c3d8e6cff08ae935c00915f1a89f251c3006

Observation b2411a8b-6e36-4ac3-ba0a-703c1103808b · outbound

This paper cites ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI.

MuJoCo Playground ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.182563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.182563Z digest=sha256:2560cb275da51fa1986fe4be3acf26d664bf60273f72a580bd4ec43f7f638c5a

Observation cf10259e-ff76-491c-bc8b-5636cf06afc0 · outbound

This paper cites DeepMind Control Suite.

MuJoCo Playground DeepMind Control Suite

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.187882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.187882Z digest=sha256:76145fe84ad319ddec5d93a24c54327d321a9bb8b0acfb4e072570bde9d0eb91

Observation e335c1fd-2f2d-408d-8b1d-807a1bd3436a · outbound

This paper cites Domain ran- domization for transferring deep neural networks from simulation to the real world.

MuJoCo Playground Domain ran- domization for transferring deep neural networks from simulation to the real world

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.735348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.194861Z digest=sha256:9c898e4b93857029a35fe2f5a1786ee5ad87d898e9afa246af7da5e478359a48

Observation fae85ff3-2bc8-4285-9a0b-80c8f8087eed · outbound

This paper cites Mujoco: A physics engine for model-based control.

MuJoCo Playground Mujoco: A physics engine for model-based control

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.203907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.203907Z digest=sha256:d4ed76c7a98f3710e18d0ddcb84de640d0855f6dba79de17fcd1f0072f22ff82

Observation add7841a-ae70-4bd3-9f32-7480d8d99c7c · outbound

This paper cites EfficientZero V2: Mastering Discrete and Continuous Control with Limited Data.

MuJoCo Playground EfficientZero V2: Mastering Discrete and Continuous Control with Limited Data

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.209457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.209457Z digest=sha256:48fd58d496619d707562feb68a4218fd801e04e6da72d97e3e4e00939c3b4d3d

Observation 29bbe30d-1116-4efb-a6e7-c071fa207b46 · outbound

This paper cites Bench- marking the performance and energy efficiency of ai accelerators for ai training.

MuJoCo Playground Bench- marking the performance and energy efficiency of ai accelerators for ai training

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.685171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.215677Z digest=sha256:9daa97d88105477e2f07944dad1a4769ddd8fa4a9785b81e36773eddf24b881f

Observation 5fcdcdeb-4b7c-4e43-acbc-ed90631002d3 · outbound

This paper cites Full-Order Sampling-Based MPC for Torque-Level Locomotion Control via Diffusion-Style Annealing.

MuJoCo Playground Full-Order Sampling-Based MPC for Torque-Level Locomotion Control via Diffusion-Style Annealing

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.222719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.222719Z digest=sha256:ae9ebf2800ee72cf1d07e147cdab4c210846c77269eb0483c2f91139f8951a1c

Observation 716c539f-0f4d-4404-8eb0-acbfdbc9227e · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

MuJoCo Playground Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.228893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.228893Z digest=sha256:aa12aee6ee7e4f22adedb80043896669cf307190028fd51c74a743f81b49d2a1

Observation 50f51a09-192b-4464-bc76-ed84f699daa4 · outbound

This paper cites MuJoCo Menagerie: A collection of high- quality simulation models for MuJoCo, 2022.

MuJoCo Playground MuJoCo Menagerie: A collection of high- quality simulation models for MuJoCo, 2022

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.234964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.234964Z digest=sha256:0c2f08938b7b534352cfcda7dbad2d4ab239e384f6fe1c83ddb5adc807244f6d

Observation 97cd2f7b-df40-4362-833c-8e3cce22a425 · outbound

This paper cites Sim-to-real transfer in deep reinforcement learning for robotics: a survey.

MuJoCo Playground Sim-to-real transfer in deep reinforcement learning for robotics: a survey

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:34:12.633904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.240472Z digest=sha256:14708a53058681b092b1fb3cb7e1232df09b73180ed80a4af4349e99c3420de1

Observation 85620289-00ba-4918-91ea-30ab4a2728b5 · outbound

This paper cites Robot Parkour Learning.

MuJoCo Playground Robot Parkour Learning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:11.246073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:11.246073Z digest=sha256:ef71712e8a3b31fec2cc60e76642268d227bf2903d26a74bf87335e2c3d2f5af

Observation 2c91a006-ea14-4b75-94e6-6735b5153476 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.601436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.252413Z digest=sha256:30f069df8592a56ce79df72372ef25e66f9858652a6fa2db9acf84d9c950cbe5

Observation 74ea8abe-f4e1-4889-a458-205b0991da61 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.567978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.257818Z digest=sha256:f4140c62c5cbc43944f0a98e3e6c203376ad3099e51d8e82911956360fa5cee8

Observation a299a52c-802e-4844-bd09-bcc20087fc13 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.541681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.264364Z digest=sha256:8f2c47738aba7a28d78f68c079f0df8b398f2a1f7144f653afa501b04e390990

Observation 11bcb761-4e63-49f6-8a25-45a9a596f140 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.513319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.271284Z digest=sha256:a8566f8a6d49cb1558215c5bc8a97bfaaebe3f9fe4d71da145cfe71e0e9ee72e

Observation c7228853-32c7-40ee-8b33-42669d190ced · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.485281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.277260Z digest=sha256:29e1becc420f107035e26ee76766f6c3559a33d4c16b4688da5a29d6d77ebf3d

Observation 440a8768-6154-4064-80f7-d4f9fe137af2 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:34:12.455345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.283456Z digest=sha256:eab861c0979053886dc06cc25fedbd88d3d887d8a72511464bc181f704f041bf

Observation 9eece729-f9fe-41d6-a135-d2d0c65524db · outbound

This paper cites injections.

MuJoCo Playground injections

Reference 78

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T23:34:12.425243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T23:34:11.289646Z digest=sha256:0333c1d6c2b146f833c014309eca0103c14c7c608a9b9b1ddb2b2837cb113df3

Observation bda7f170-6c68-4735-9cdd-5a0f194b7468 · outbound

This paper cites an unresolved cited work.

MuJoCo Playground Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T23:34:10.945235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:34:10.945235Z digest=sha256:dca1cbc4851e4b4188de6c3704e556285509f94e5b58dbea903834f6da2ddd26

Pith citing papers

Observation 4c1ade24-0ab4-4d65-9264-0ef667cb2976 · inbound

Unreal Robotics Lab: A High-Fidelity Robotics Simulator with Advanced Physics and Rendering cites this paper.

Unreal Robotics Lab: A High-Fidelity Robotics Simulator with Advanced Physics and Rendering MuJoCo Playground

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:51:57.332577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T18:50:57.738602Z digest=sha256:963957ad4fd01fcf6caeec2252f894773fc160bf6ea854806c9d99cb8d9dab1d

Observation aa772c30-5458-4448-8183-e2a2046b1e05 · inbound

MOSAIC: Skill-Centric Manipulation Planning with Physics Simulation cites this paper.

MOSAIC: Skill-Centric Manipulation Planning with Physics Simulation MuJoCo Playground

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:53.054525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:53.054525Z digest=sha256:8aa2a51dfeea9e5caecc8a7222f0f4949aaa1a271671f7c1015be5edfdf25d26

Observation b164d273-89f8-49aa-92c1-55c0e14cf947 · inbound

Visual Imitation Enables Contextual Humanoid Control cites this paper.

Visual Imitation Enables Contextual Humanoid Control MuJoCo Playground

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T23:48:57.968507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:48:57.968507Z digest=sha256:73568d5bccc5c395e06c689df039993c2d70a9ab78c57000dc6d83083540d2cc

Observation 0184d00b-5bf5-4fab-ac08-c62caccdaf01 · inbound

Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control cites this paper.

Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control MuJoCo Playground

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:45:35.064770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:45:35.064770Z digest=sha256:ce5e9a575297e2a8ef002a1b52a907fc192e3da02b4ae27933820893e7e55f9f

Observation 61157ede-b9ad-44e5-b8be-61cee418ef4f · inbound

SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics cites this paper.

SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics MuJoCo Playground

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:31.382654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:31.382654Z digest=sha256:6970deba95800cc1d1fe047981834962adad06239d5b884eea2f1f720042bb46

Observation 2591ff17-1047-4344-9c37-9a96b3937f3d · inbound

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control cites this paper.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control MuJoCo Playground

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.735651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.735651Z digest=sha256:eca772396a9df10ff5c476d39ab4f10cb9a127e50541ab3097481e64e461e4ff

Observation 45f4eb52-46d9-4505-93a4-632df8cc7b34 · inbound

Booster Gym: An End-to-End Reinforcement Learning Framework for Humanoid Robot Locomotion cites this paper.

Booster Gym: An End-to-End Reinforcement Learning Framework for Humanoid Robot Locomotion MuJoCo Playground

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T19:45:00.972627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:45:00.972627Z digest=sha256:7bd262348b18847efbb1fcfb08454c189a8c0984e0c738d0855a62885c38a026

Observation 0e5d95f8-c6c6-4f32-9034-9d550f6f1e88 · inbound

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training cites this paper.

SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training MuJoCo Playground

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:06.275060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:06.275060Z digest=sha256:0277c5d51463a0467701e13b80a32a675d4a88e0e2737cbf7ea072ad86b67c15

Observation 91c46414-c837-4d29-8744-b3125bb80fa6 · inbound

Flow Matching Policy Gradients cites this paper.

Flow Matching Policy Gradients MuJoCo Playground

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:09.516529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:09.516529Z digest=sha256:fe30da27d07c1235387d71d1d8fc66afc8b620799fbcee7bddb2f7ef83d181c0

Observation ae6cdf95-ba5c-4a97-9682-e5ab0e3b39c9 · inbound

Viser: Imperative, Web-based 3D Visualization in Python cites this paper.

Viser: Imperative, Web-based 3D Visualization in Python MuJoCo Playground

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T11:13:12.409367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:13:12.409367Z digest=sha256:c171b4f7c6ab450edf16a24e4b74512be1d2f60758d4fcaf686f99fa99d540bb

Observation 1915220b-0f3f-4a41-b915-29c425c5bd3e · inbound

Simultaneous Contact Sequence and Patch Planning for Dynamic Locomotion cites this paper.

Simultaneous Contact Sequence and Patch Planning for Dynamic Locomotion MuJoCo Playground

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:21:25.270760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:21:25.270760Z digest=sha256:3f03dc01e50b1d4f21705d144a446ffda6f0835d473956d234f698f45669ddc9

Observation 7717e797-ead4-4540-a0a2-a07d8add70f1 · inbound

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges cites this paper.

Robotic Manipulation via Imitation Learning: Taxonomy, Evolution, Benchmark, and Challenges MuJoCo Playground

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-05T16:55:52.309377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:55:52.309377Z digest=sha256:4904f4d034c8d1d92cd2407bfdc24c853dd24e4f0807421c95c6a0719c25ba8e

Observation 6f27ce27-967c-4cec-812d-5bf4fb298b3c · inbound

RecoWorld: Building Simulated Environments for Agentic Recommender Systems cites this paper.

RecoWorld: Building Simulated Environments for Agentic Recommender Systems MuJoCo Playground

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T17:56:38.873029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:56:38.873029Z digest=sha256:313acc1cbe966638a89ff53510dd97f40cb8cefe13a80afe7f37422916670449

Observation 3ae413aa-1f6b-4178-af74-66361ca0ea93 · inbound

MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning cites this paper.

MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning MuJoCo Playground

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T22:58:38.900601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:58:38.900601Z digest=sha256:03f80aadcb2bb21fc53ad31352f62600cf845f4bd7b9d0e9cfa29eeaa288f521

Observation e0fe0836-7a91-45e0-b36e-41a57adab712 · inbound

What Matters for Simulation to Online Reinforcement Learning on Real Robots cites this paper.

What Matters for Simulation to Online Reinforcement Learning on Real Robots MuJoCo Playground

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T21:36:15.812654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:36:15.812654Z digest=sha256:2c28ef0452eea46f2d8a8ce559180a20cde82972bf6ea58401b051b6fe2c8391

Observation 4aaa48ef-34b2-4e8b-beba-ff37f25dc449 · inbound

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation cites this paper.

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation MuJoCo Playground

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:36:17.295232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T16:35:34.179788Z digest=sha256:2263f992530c3370bb3565645aa0abbd44ae3721cdb2a320e5bc6da8e5ba9395

Observation f5fd9519-96ff-4027-9d3f-7769c260314c · inbound

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation cites this paper.

PTLD: Sim-to-real Privileged Tactile Latent Distillation for Dexterous Manipulation MuJoCo Playground

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T18:52:42.535173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:52:42.535173Z digest=sha256:f911021236ccea341ef5f13f7813cf554ed377be3e502ccb2b8acff3ee8d1149

Observation febfeef0-bb3b-4482-8c81-822975be3825 · inbound

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control cites this paper.

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control MuJoCo Playground

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:20:00.675540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T12:17:59.392158Z digest=sha256:545c0b083e538ac2ba59d2130dc45f917b83a321678965eb6d402c992e5559b8

Observation 04a3f884-5279-4807-b3ed-50095a520332 · inbound

Research Novelty in Information Systems Journals After ChatGPT: Differences Across Institutional Language Contexts cites this paper.

Research Novelty in Information Systems Journals After ChatGPT: Differences Across Institutional Language Contexts MuJoCo Playground

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-15T11:52:28.028968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:52:28.028968Z digest=sha256:3b9a1be962729769bc54ebd0e5cf52878188af9894ada0b596e43e903a8f62e4

Observation be7679be-271e-4c71-9b7d-e472d73b5518 · inbound

Learning Dexterous Grasping from Sparse Taxonomy Guidance cites this paper.

Learning Dexterous Grasping from Sparse Taxonomy Guidance MuJoCo Playground

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:00.984225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T17:05:06.300013Z digest=sha256:2012e4f92a9ec2702fb0c5cbe32f32bf083244ad6a8234f338dd024cebbc19c9

Observation 891402cb-515a-4e4d-8899-39c91b272998 · inbound

On Data Thinning for Model Validation in Small Area Estimation cites this paper.

On Data Thinning for Model Validation in Small Area Estimation MuJoCo Playground

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T11:14:37.619203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:14:37.619203Z digest=sha256:09eabe32ea45e6aed5fcc25b1d1d308d8c2c8bf6cd18ea5b0d6f3245264c8f79

Observation a2c1e83c-5c0f-4d91-887b-3a0aa009567f · inbound

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control cites this paper.

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control MuJoCo Playground

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:15:49.708901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T20:04:56.512544Z digest=sha256:2d6b313aa96199093131ed94ef7f013f4af9b45392c7a5b5a10c6c0b7f6c6bc3

Observation 627ae881-2b5f-42d0-9613-b815e006808b · inbound

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control cites this paper.

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control MuJoCo Playground

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:12:41.315328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T17:08:31.770889Z digest=sha256:c4f2703c5d1b8278299764ce01c5d244dc7aabaa52d747c30e9840adb946f9d8

Observation d1c0e1be-db11-421f-8535-fa344fdcb0b6 · inbound

Simulation-Driven Evolutionary Motion Parameterization for Contact-Rich Granular Scooping with a Soft Conical Robotic Hand cites this paper.

Simulation-Driven Evolutionary Motion Parameterization for Contact-Rich Granular Scooping with a Soft Conical Robotic Hand MuJoCo Playground

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:15:50.383745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T20:04:06.942355Z digest=sha256:9cab55380abf211507713e4c10db71736b7e1829c0605aee5dbaa3e93601a4c2

Observation 7945ad17-eae3-4831-ba49-324b0c9fb612 · inbound

A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies cites this paper.

A Mechanistic Analysis of Sim-and-Real Co-Training in Generative Robot Policies MuJoCo Playground

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:30:22.763830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T12:29:38.306670Z digest=sha256:773483ad1f2d1ef1125423b4ddd8b44bcb10dc90e3218c324f57b63d1a1df035

Observation 0b7b07e6-1325-452d-bb3a-3d873f7f2bcf · inbound

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics cites this paper.

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics MuJoCo Playground

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:26:14.865036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T02:45:48.819070Z digest=sha256:b258feadb778f71bd0cac1a57e9370582f55bfb4d68640bba7f36d93b389b4ee

Observation 5f5b781d-d0c8-4574-a622-6feae5ff7257 · inbound

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics cites this paper.

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics MuJoCo Playground

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:06:16.506077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T03:27:02.669953Z digest=sha256:420953446ec7c7953af110bd6c8c2c0481e4f6e06fc94228c99353043ba15f92

Observation a1e497f0-6063-4745-85d7-a093f89f590f · inbound

GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning cites this paper.

GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning MuJoCo Playground

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:51:17.041890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T16:12:45.211236Z digest=sha256:de5367bae38ed529c7d8dc867c10cbed2b5e0e0651a689a7113df6cc7b91b336

Observation 1c47de61-72df-42d8-b304-69c78636cb4f · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders MuJoCo Playground

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:33:04.209666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T05:28:50.354662Z digest=sha256:59553d0ce443de2fa0d24455da41086b1a38c1a9cc671f161c1c49b9cb34b82d

Observation baaa4e01-c023-44a7-a32e-be4d220f4bc3 · inbound

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders cites this paper.

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders MuJoCo Playground

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:39:49.242236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T07:36:12.214949Z digest=sha256:1d4d42cc5bc1fc08eb252610f5f3a3d881788f609fc3011f485f49553206434f

Observation b3607232-2a62-42a6-b2cc-e971d0e1aaf7 · inbound

MuJoCoUni:Persistent Batched Runtime Primitives for MuJoCo cites this paper.

MuJoCoUni:Persistent Batched Runtime Primitives for MuJoCo MuJoCo Playground

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:47.814957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T01:13:54.860724Z digest=sha256:de598b294a73877d2c20836b759db242cfddcdb2d3377322de35164d4893b457

Observation eea67457-2825-414d-ae0e-a716ad6e0727 · inbound

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion cites this paper.

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion MuJoCo Playground

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:05:49.719887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T00:54:12.099045Z digest=sha256:dc80319ba90b09276bf3af767a891536f6fcbbf07c7bd4361302b285fb8b934e

Observation a61681b2-7eac-4a12-8491-f38d87ce9c57 · inbound

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient cites this paper.

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient MuJoCo Playground

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.609847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T17:34:41.053725Z digest=sha256:e1d07eb84f533b2db85eec23e28d14b906bc447807cce8aa27b943c0cad63801

Observation 274fdd30-bc50-4669-be9d-66a72a1599c5 · inbound

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms cites this paper.

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms MuJoCo Playground

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:13:16.503863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T07:09:19.137932Z digest=sha256:1aa3527f0f6b76f2b96c1c2cbe2f5c36103c3b10bdbceefd8f63bc97d18df196

Observation a4a01c17-dfe4-4ad5-a76a-d480c3cba122 · inbound

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning cites this paper.

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning MuJoCo Playground

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T22:22:43.501943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T22:20:37.637669Z digest=sha256:370d0cb74b9576d40648cf17b1babf2a9ea19105ca52fcff0501dd93ee5b393a

Observation 1c23d7ed-caf9-4752-adf0-0569e307b743 · inbound

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It) cites this paper.

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It) MuJoCo Playground

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:42:38.047426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T18:19:58.499360Z digest=sha256:abf829e9b8bb7170fccf68cf7e9031033b943d38261fbf0dd0f6113c5b32e927

Observation 069d9282-7e8c-46a9-adad-cb43d4e823e2 · inbound

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment cites this paper.

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment MuJoCo Playground

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.277639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T06:20:18.312301Z digest=sha256:639809b8cbd930ac9139424061c09b8d3f433d5e253e2eb24a506dceb51db33b

Observation 9b571c89-d364-428f-a374-292e2fd19428 · inbound

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation cites this paper.

QuadVerse: An Integrated Framework Aligning Visual-Physical Reality for Quadruped Simulation MuJoCo Playground

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.979243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T21:55:15.993627Z digest=sha256:d7027e3b9fcc5b7cb81f08b4002fb694a22e5d3dd631eb3368a959944f962c8a

Observation 8316a452-9a63-4786-99de-4c2f667cf8b6 · inbound

Embedding Hybrid Systems into Continuous Latent Vector Fields cites this paper.

Embedding Hybrid Systems into Continuous Latent Vector Fields MuJoCo Playground

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:17:37.141265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T14:04:40.078357Z digest=sha256:662cdfc21a22e76748de8c450b6f1edd82acf09ad856afa597aadb774ac26623

Observation c3193287-4333-42ec-a088-9d66bf7d96e0 · inbound

Blind Dexterous Grasping via Real2Sim2Real Tactile Policy Learning cites this paper.

Blind Dexterous Grasping via Real2Sim2Real Tactile Policy Learning MuJoCo Playground

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:48:02.597110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T09:50:02.635507Z digest=sha256:4d858b659cad9e8a89a0ac2c844bc76d6d0501f297a66e87ee3ad75ab0246b3f

Observation 2dfda7d7-0a30-4c9e-af46-c8d1603b103a · inbound

Benchmarking Action Spaces in Reinforcement Learning for Vision-based Robotic Manipulation cites this paper.

Benchmarking Action Spaces in Reinforcement Learning for Vision-based Robotic Manipulation MuJoCo Playground

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:12.994967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T21:22:53.551550Z digest=sha256:7fd5c162ca665508ffbe66016da26829ca3bf3244bb0027af10bb0ff33840c4a

Observation f4881233-2679-4885-ad83-f92832e45226 · inbound

Simulating Robotic Locomotion in Sand: Resistive Force Theory in an Open-Source Physics Engine cites this paper.

Simulating Robotic Locomotion in Sand: Resistive Force Theory in an Open-Source Physics Engine MuJoCo Playground

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:59:21.067835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T20:45:44.072738Z digest=sha256:2352b3ef57cdecc090860acb61fda51873b084c3e954c3d7f2b574cfda7a3168

Observation 73b9c8a3-f9b9-4561-8e6c-cbb6e6c8f9ec · inbound

CRAX: Fast Safe Reinforcement Learning Benchmarking cites this paper.

CRAX: Fast Safe Reinforcement Learning Benchmarking MuJoCo Playground

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:09:30.117010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T18:22:36.891186Z digest=sha256:f309ba377fa8aa597485f2f006666765ad72cfdf2b37f7f3ee6becb315853757

Observation f8da729b-4043-4834-9878-f75dc173af1d · inbound

ReFPO: Reflow Regularization for Flow Matching Policy Gradients cites this paper.

ReFPO: Reflow Regularization for Flow Matching Policy Gradients MuJoCo Playground

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:19:37.785912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T14:38:54.754049Z digest=sha256:dd8da7ef82315b22d76de18aaaefc5ee7e7fe32715bc4d721d23e34fe854e47b

Observation 4c37d03e-ff90-4e0a-8964-077c453c63b9 · inbound

Long-Distance Real-World Navigation of the Legged-Wheeled Robot Go2-W Using Deep Reinforcement Learning cites this paper.

Long-Distance Real-World Navigation of the Legged-Wheeled Robot Go2-W Using Deep Reinforcement Learning MuJoCo Playground

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T13:59:30.584615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T13:57:34.489168Z digest=sha256:46c6fe9be1ad9d28298e347df6423bd21f41534dd0b3adea72707a101d4f4f7e

Observation 040babbf-c22e-432f-98bb-4b4687ef0c68 · inbound

Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System cites this paper.

Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System MuJoCo Playground

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-26T11:39:24.547908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T11:39:18.595140Z digest=sha256:b66a11d4fe72b530d5b6e80c017c1a928a964b5acbf0aa53310ba56a6c0bbf09

Observation 021a8f26-846d-4e74-b55d-7fcbdfd40bdc · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering MuJoCo Playground

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.191604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:6818f47f6c562e13865565bee87e36e58580817bf318f5367096d8a432dc9d09

Observation b38a5207-aeea-4be4-b093-599e7d27c379 · inbound

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications cites this paper.

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications MuJoCo Playground

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:06:55.353101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-02T11:59:04.103002Z digest=sha256:b22005eee4a16df4f869546b5ec5e38ca807d7dea1149150e20a4ac20c8e0c95

Observation 5e963b18-f5ca-43ae-9527-e34373a971c2 · inbound

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space cites this paper.

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space MuJoCo Playground

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T01:30:09.928609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:30:09.928609Z digest=sha256:674f1cfd1a2b544e5ea5fe1db989eaecf0544d425d8841f70c0cb7c31faba576

Observation a257f697-0539-4b3c-9940-d9204394965c · inbound

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics cites this paper.

OrchardBench: A Physically-Grounded, GPU-Parallel Apple-Orchard Simulation Benchmark for Agricultural Robotics MuJoCo Playground

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T09:34:47.883931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-08T09:33:49.773169Z digest=sha256:f41420e34003976c1b392289b64e558107c4bbe439e3f84ee57b0aebb6b3c7f4

Observation 1e155c6f-39bb-42ab-9841-927edd53210b · inbound

Rethinking the Suitability of Reinforcement Learning Algorithms Under Practical Transfer Constraints cites this paper.

Rethinking the Suitability of Reinforcement Learning Algorithms Under Practical Transfer Constraints MuJoCo Playground

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T18:24:35.875407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T18:24:35.875407Z digest=sha256:7f124bb77e9e9e11ead193ca332488503da98731e225c80ad5cc57a4b3095247

Observation fd659c5f-e75e-464b-92c9-abe51ca702e2 · inbound

AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation cites this paper.

AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation MuJoCo Playground

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T07:02:19.931906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:02:19.931906Z digest=sha256:60819d19ba939b73767daf001cfd3cebde9b26934b196ffdbab0afa14b1fc0ff

Observation d5868d46-7211-4538-97f8-d5f854ff5616 · inbound

$\pi\mathbf{R}^2$: Reactive Real-time Flow Policies cites this paper.

$\pi\mathbf{R}^2$: Reactive Real-time Flow Policies MuJoCo Playground

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T00:49:36.679003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:49:36.679003Z digest=sha256:1c7fb23595fb2ef8161528c94a2c4d40fe47d88dbde57c0ed15d5a927b5a450e

Observation 6fe240df-fd29-4458-bda9-e87a47ac0078 · inbound

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts cites this paper.

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts MuJoCo Playground

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T00:17:16.320668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:17:16.320668Z digest=sha256:1bb64301a49581c35c78fcd9c2acc380d0b51af72444734f5f6668d9792b365f

Observation 496ca157-cab4-4c9c-9a02-a3f9daebf011 · inbound

Foundations of Reinforcement Learning and Control:Connections and New Perspectives cites this paper.

Foundations of Reinforcement Learning and Control:Connections and New Perspectives MuJoCo Playground

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-04T07:32:30.427620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:32:30.427620Z digest=sha256:565e2ffcffae51aa55c8d8a28118e3eeb33f810df7da30e46919835787b90dfe

Observation 68d00cd2-65a6-4fa3-afa5-14e64dc04d4f · inbound

ATP: Anatomical Torque with Passivity-based Control Framework for Safe Upper-Limb Exoskeleton Assistance cites this paper.

ATP: Anatomical Torque with Passivity-based Control Framework for Safe Upper-Limb Exoskeleton Assistance MuJoCo Playground

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:42:26.860884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:42:26.860884Z digest=sha256:5fa0f4fdf7ae2526c3c3273d935dbda6f28021e191155d5ac63b30a4388b3821

Observation 5b8b3ed1-8047-43e6-97c2-5d88df470f2f · inbound

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control cites this paper.

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control MuJoCo Playground

Reference 284

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:47.076898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:47.076898Z digest=sha256:593203bf723a454478703bcbf69f03a5ed28c5f872bcc490ef209dbb81c3bff5