Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:03:32.220111Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 3 inbound Pith citation observations for arXiv:2507.23172.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:03:32.220111Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T09:47:59.248127Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T11:46:55.739072Z
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e0259aa2-b285-4ab6-9435-193e1a95d66b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Deep reinforcement learning at the edge of the statistical precipice
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5fa4944-397e-44e1-8606-27211b6b0bf6 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Locomujoco: A comprehensive imitation learning benchmark for locomotion
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4bad7e69-22dd-4008-acfc-f4649fc76bef · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Transferring Dexterous Manipulation from GPU Simulation to a Remote Real-World TriFinger
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3827d9ad-3158-4b90-8100-960951aa70db · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Layer Normalization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b10ff1-ae20-4779-bf81-9719f4b0d411 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Reinforcement Learning through Asynchronous Advantage Actor-Critic on a GPU
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 148d73d9-d505-4656-9469-ff55e32f826b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAX
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 491e358b-41aa-4f9b-9660-e1946ef5fc3b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks DaXBench: Benchmarking Deformable Object Manipulation with Differentiable Physics
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 312e53a3-d87f-4fef-81a8-051d536222d3 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Extreme parkour with legged robots
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35b71f9c-0e24-4310-bd0f-c3675fd4cd26 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Leveraging procedural generation to benchmark reinforcement learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation faac1cd1-27bd-4bce-b510-9776a76ae743 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Sample-efficient reinforcement learning by breaking the replay ratio barrier
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfddbf17-7840-428c-8808-b0106f042cda · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3870f6e1-a552-42ea-b335-6bc2d11dee6b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Franka emika panda robot, 2017
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54764a4d-0c94-40c1-abf4-a62653d42ccc · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccd60bd3-ae62-47bd-a072-3bd43242188b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Deep whole-body control: learning a unified policy for manipulation and locomotion
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab8b6e32-c3dd-4343-aa03-25fb4bda2d62 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Simplifying Deep Temporal Difference Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8459246d-52e0-4e88-b6a0-342a40ad399c · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a321585-e291-49c4-9d6e-32adc78b4397 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d59c6f04-43fc-48a7-9b93-b6c32e809d66 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Multi-task deep reinforcement learning with popart
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a46b7bb0-bc39-44d3-8dbf-881ccef46813 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Distributed Prioritized Experience Replay
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9ffe28-7bb2-4d4b-b8a9-1b7cd0c3554b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Learning agile and dynamic motor skills for legged robots
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3258f4a-b1b9-43b5-b886-798a84a21ba2 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd30d924-2919-47fd-aa93-5fedc6e7b561 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ab0480f-88c2-482e-8a09-ab1a4f53cc69 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks A survey of zero-shot generalisation in deep reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4395375f-7a86-46d0-b96e-855ff9522a56 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Pgx: Hardware-accelerated parallel game simulators for reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0e12560f-4329-4fbf-8dd1-372ed53574d4 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks gymnax : A JAX -based reinforcement learning environment library, 2022
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a40fe96-9794-4111-bade-18bcb46fbbfa · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Learning quadrupedal locomotion over challenging terrain
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae0ec9ea-4485-46d2-9123-ad50587f3a1b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Parallel q -learning: Scaling off-policy reinforcement learning under massively parallel simulation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 364232bc-313b-4d15-8146-02d236d6c375 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Rllib: Abstractions for distributed reinforcement learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ab1ae18-ad23-4ec2-a8df-b474ab50a07c · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Gpu-accelerated robotic simulation for distributed reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d37a556b-2458-43da-919f-b8703e104bfc · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Eurekaverse: Environment Curriculum Generation via Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47eec33a-a586-4242-80be-6a08a54dd6b5 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks FAMO: Fast Adaptive Multitask Optimization
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 089e5355-0b7b-4230-b443-348793f2127b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dcc7e5e-a8a9-45e2-b9b0-eefce8af99a2 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Conflict-Averse Gradient Descent for Multi-task Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e5a73cd-437f-4547-aef0-4bc354ce73fd · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Perpetual humanoid control for real-time simulated avatars
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54044885-8e42-4ff7-82e6-3826c246efd0 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks rl-games: A high-performance framework for reinforcement learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce8db6ca-13f9-4303-8011-c66f60ef4308 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51f0bb4e-13d4-4c88-b0d5-0942b92d19ff · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Rapid locomotion via reinforcement learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d54f2052-a4f1-4f94-88de-cc7103c26e2b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Craftax: A Lightning-Fast Benchmark for Open-Ended Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb8220b2-4ab6-4d99-bfc5-e802104815c9 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Orbit: A unified simulation framework for interactive robot learning environments
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 549915d5-7361-4756-a42a-b11ed3a2739d · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Playing Atari with Deep Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 657a0540-9a52-41b7-9faf-4514d3bfd4bb · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Rusu, Joel Veness, Marc G
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 042c9cc2-100a-4ca5-a884-dab096552ee5 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Asynchronous methods for deep reinforcement learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6adb8350-b470-4884-8cee-2a8e9ef67971 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks POPGym: Benchmarking Partially Observable Reinforcement Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60f8519e-7bb0-4d02-a5dd-db22594ca89e · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Massively Parallel Methods for Deep Reinforcement Learning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21094dcb-5b7d-4cdc-9496-eda1c508bba8 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Learning Dexterous In-Hand Manipulation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94b606da-e505-4439-a907-8aea010ac98d · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks OGBench: Benchmarking Offline Goal-Conditioned RL
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8025aec0-f76e-4f0f-80f3-3c664fa6c409 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Sample factory: Egocentric 3d control from pixels at 100000 fps with asynchronous reinforcement learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56cf0269-8829-47f6-9178-353ac92748be · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Learning to Push by Grasping: Using multiple tasks for effective learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a31f0dba-12cf-4f62-b3ad-670077d431ca · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Learning to walk in minutes using massively parallel deep reinforcement learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40a0afb0-41b0-49d2-8666-7bfd43cfeec4 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 842ede67-ce26-4d2e-9dcf-cef77f684586 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Proximal Policy Optimization Algorithms
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee4ca13-facb-48a3-9191-b09fb5d5866d · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Solving Continuous Control via Q-learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d968c205-8934-42fd-b2a7-ad78bab7f571 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8291e303-13cf-4d50-8383-79c9691d366f · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b2a84a9-6694-411d-a3a9-4ee50ab02404 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Mastering the game of go with deep neural networks and tree search
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8edd2085-984b-4cd6-a003-1f3f81c939a4 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Sapg: Split and aggregate policy gradients
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b44df3ee-a281-468c-8b6e-fd5c22f9e1ed · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Mtrl - multi task rl algorithms
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b98ea7ac-7bc8-4b88-9745-60fcd076016a · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Multi-task reinforcement learning with context-based representations
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73c4d6a4-1bee-416e-a3ef-9a44468a2bb7 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks PaCo: Parameter-Compositional Multi-Task Reinforcement Learning
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c905634-aadf-4819-961e-8212124c9f4b · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Value-Decomposition Networks For Cooperative Multi-Agent Learning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 914f3f29-610e-43fc-bde4-5c3aa65780a4 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Policy gradient methods for reinforcement learning with function approximation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ea2fe3f-5da3-46b9-84fa-74f99b29251f · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68c89edc-604f-4cc5-9864-f4518319b735 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks DeepMind Control Suite
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcf13415-3bb3-4bda-997f-b352ffbfb748 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Mujoco: A physics engine for model-based control
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e5a82a9-c46a-48eb-9fff-257bfa6a15c9 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Go1 User Manual
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 79daa933-3cb6-46f4-adc2-430add7b2e7d · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Dueling Network Architectures for Deep Reinforcement Learning
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76425055-dca4-44c7-be27-c1107a9aa233 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Outracing champion gran turismo drivers with deep reinforcement learning
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e35af45e-1f0b-448d-b13c-8b2845a3e1bb · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Stabilizing Reinforcement Learning in Differentiable Multiphysics Simulation
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b81ebc2e-3859-43b7-8cf5-1912b4c15131 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Multi-Task Reinforcement Learning with Soft Modularization
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f03fb84-dc42-4b67-b1a5-abfff31214d3 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Gradient Surgery for Multi-Task Learning
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54e0e1a9-2db9-4aba-94b0-ea133a08d417 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1687d3a9-5815-4d30-ad87-3707271c1528 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Kahrs, Carlo Sferrazza, Yuval Tassa, and Pieter Abbeel
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 745be101-15cf-404c-8fa4-113e362d9d90 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks robosuite: A Modular Simulation Framework and Benchmark for Robot Learning
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a99782d8-5d50-41a7-b8e4-d33ff230ed85 · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Robot Parkour Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3610a66a-7b43-46b3-a45f-9d670e20b58c · outbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks write newline
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1747023-19c3-4ee6-a954-0939abf28fc4 · inbound
Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f022c6-d5d9-416e-9437-40c857ca52a9 · inbound
TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 320feb81-15bf-4380-b99e-8748e88b3295 · inbound
Representation Learning Enables Scalable Multitask Deep Reinforcement Learning Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.