Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T07:44:14.707183Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 100 inbound Pith citation observations for arXiv:1801.00690.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-13T07:44:14.707183Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:33:00.901351Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
13 of 13 outbound references displayed
External citation measurements
524
pith, observed 2026-08-05T02:28:24.338817Z
Observation 807f2362-7982-4495-b448-78c7a93bdd25 · outbound
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3854d122-4ec6-4075-b192-77239ad98e89 · outbound
DeepMind Control Suite doi: 10.1109/TSMC.1983.6313077
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c67eab7a-90c3-417c-9ca5-afe34451d96e · outbound
DeepMind Control Suite A Distributional Perspective on Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5a9edccc-de44-4a74-82a2-bed0bbf8d511 · outbound
DeepMind Control Suite Simulation tools for model-based robotics: Comparison of bullet, havok, mujoco, ode and physx
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b009f6a1-d086-47c7-9cb7-80f05702153e · outbound
DeepMind Control Suite Reproducibility of Benchmarked Deep Reinforcement Learning Tasks for Continuous Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1a4731eb-07d8-4429-831f-2985294a14b7 · outbound
DeepMind Control Suite Adam: A Method for Stochastic Optimization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d1dcd187-9fbb-43ec-850f-2555dd775096 · outbound
DeepMind Control Suite Continuous control with deep reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6cce8354-23d2-47f4-a76e-1a25442f3eda · outbound
DeepMind Control Suite Learning human behaviors from motion capture by adversarial imitation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 696e50c7-cce6-4c80-9ffe-16489702aa79 · outbound
DeepMind Control Suite Asynchronous Methods for Deep Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 446edd7a-c21b-4b18-855c-86db467ad88b · outbound
DeepMind Control Suite Data-efficient Deep Reinforcement Learning for Dexterous Manipulation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0815a571-6fdf-45c5-9891-89fcae920a79 · outbound
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 704f26a3-6e1c-4c1a-9898-d80878ae1136 · outbound
DeepMind Control Suite Synthesis and stabilization of complex be- haviors through online trajectory optimization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dfba87fe-73f8-49e7-98a2-7f469ab8dc87 · outbound
DeepMind Control Suite Mujoco: A physics engine for model- based control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1da3fe04-a7a1-4e8c-96dd-3207e1013801 · inbound
Continual Reinforcement Learning with Diversity Exploration and Adversarial Self-Correction DeepMind Control Suite
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aee1c70b-09af-465f-9667-305e16b3d165 · inbound
Benchmarking Model-Based Reinforcement Learning DeepMind Control Suite
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e8a13acc-cb8d-414c-87af-0aa75e403a70 · inbound
An Actor-Critic-Attention Mechanism for Deep Reinforcement Learning in Multi-view Environments DeepMind Control Suite
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 16d83a09-8b7e-4ce8-9dcd-8ab0ea9f3f65 · inbound
Arena: a toolkit for Multi-Agent Reinforcement Learning DeepMind Control Suite
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8cbcd9de-289e-4009-88b3-7c1f6ddde25f · inbound
DoorGym: A Scalable Door Opening Environment And Baseline Agent DeepMind Control Suite
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeab421e-c2d9-4e84-9da8-1404c2f0436f · inbound
Behaviour Suite for Reinforcement Learning DeepMind Control Suite
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fce1452b-b1f0-4404-8e40-74b21e56e207 · inbound
Continuous Control for High-Dimensional State Spaces: An Interactive Learning Approach DeepMind Control Suite
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbf5f46a-224b-4cc9-9304-429984598fb9 · inbound
Reference 1999
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd1a0bee-e734-4c36-bf31-d8d88a866bea · inbound
Constraint Learning for Control Tasks with Limited Duration Barrier Functions DeepMind Control Suite
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d999fd21-c414-4317-8ded-21298ed57e68 · inbound
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88e64274-e44b-4b18-809c-bffb6c53e312 · inbound
Evolutionary reinforcement learning of dynamical large deviations DeepMind Control Suite
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2107af0d-8c8b-4f7b-a27d-5b7f95de44fa · inbound
Dream to Control: Learning Behaviors by Latent Imagination DeepMind Control Suite
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6571296c-dd51-420b-9669-e171941ef493 · inbound
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bf152135-6bf3-4b22-ad38-7a6f5f8116c5 · inbound
Mastering Diverse Domains through World Models DeepMind Control Suite
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a86d6b53-1fe0-464b-8612-2f61151fa780 · inbound
Learning Interactive Real-World Simulators DeepMind Control Suite
Reference 149
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b8715765-dea1-4596-84ed-0c0a23154d20 · inbound
BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation DeepMind Control Suite
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 204c8b6d-ebef-4c13-a9ed-afd580d1965b · inbound
A Survey on Vision-Language-Action Models for Embodied AI DeepMind Control Suite
Reference 191
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c13aac80-e2f4-463e-b2eb-7818c171ad2e · inbound
Plasticity Loss in Deep Reinforcement Learning: A Survey DeepMind Control Suite
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3b40c114-ac46-4e80-85c3-e0ecafcdfb9c · inbound
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning DeepMind Control Suite
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1ef484d7-9a17-47bd-864c-62a0bbb620b3 · inbound
The Surprising Ineffectiveness of Pre-Trained Visual Representations for Model-Based Reinforcement Learning DeepMind Control Suite
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7191024-6f99-4ea4-838a-21fc4fe43c2c · inbound
A Pre-Trained Graph-Based Model for Adaptive Sequencing of Educational Documents DeepMind Control Suite
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 936a2caa-1bd3-4ba7-a870-c0edf194c34f · inbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement DeepMind Control Suite
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3457a125-7d69-42c0-8586-49cd62777647 · inbound
Proto Successor Measure: Representing the Behavior Space of an RL Agent DeepMind Control Suite
Reference 9879
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1974da99-5b2e-4708-bdbb-e0c3cf56c6ce · inbound
Quantization-Aware Imitation-Learning for Resource-Efficient Robotic Control DeepMind Control Suite
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da024ce-5f1e-4160-baa2-d27fe26bfac3 · inbound
LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations DeepMind Control Suite
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 498ee4d0-47ef-408b-9e8d-d0e330dd76de · inbound
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning DeepMind Control Suite
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 653571ef-1a73-4cbe-afbd-5a380cf3ba5d · inbound
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization DeepMind Control Suite
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a25c4a6-ea0c-43d8-aea4-07b9f4459cdd · inbound
Equivariant Action Sampling for Reinforcement Learning and Planning DeepMind Control Suite
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60759d72-b6ed-4eb5-805d-bc931de3135f · inbound
When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning? DeepMind Control Suite
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 477d9800-9cd4-4f15-bd20-12d6563daecb · inbound
Mimicking-Bench: A Benchmark for Generalizable Humanoid-Scene Interaction Learning via Human Mimicking DeepMind Control Suite
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69f9d5b3-9ffe-4d3f-b8eb-03d142fef62d · inbound
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight DeepMind Control Suite
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dddd86c4-1003-4c99-8cda-baba5dd184d7 · inbound
Episodic Novelty Through Temporal Distance DeepMind Control Suite
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03f9be27-826b-443c-bd3f-733efbb4037e · inbound
Towards General-Purpose Model-Free Reinforcement Learning DeepMind Control Suite
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a101682c-8ed7-4680-83a8-a66b30983d51 · inbound
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning DeepMind Control Suite
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0571d6fd-0f68-4fae-ba71-14eab8a96cac · inbound
Trajectory World Models for Heterogeneous Environments DeepMind Control Suite
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e2de058-a7e6-4ff0-8168-187637aace92 · inbound
Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation DeepMind Control Suite
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff320b2a-65dc-42d1-90f9-2fc5efa76819 · inbound
Domain-Invariant Per-Frame Feature Extraction for Cross-Domain Imitation Learning with Visual Observations DeepMind Control Suite
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03b7e3bf-b9ad-4a06-9d3d-d41cd7ab7fc9 · inbound
TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint DeepMind Control Suite
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ed334e7-e600-487b-823e-b0184f09bff2 · inbound
Efficient Reinforcement Learning Through Adaptively Pretrained Visual Encoder DeepMind Control Suite
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6b13d53-5601-4e40-b84a-d9868ee27780 · inbound
Skill Expansion and Composition in Parameter Space DeepMind Control Suite
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2ce34cb-fe22-4df6-a1f3-e20a89ae72cc · inbound
DrugImproverGPT: A Large Language Model for Drug Optimization with Fine-Tuning via Structured Policy Optimization DeepMind Control Suite
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55ee6be8-6836-40f9-a899-c5f22e3d4788 · inbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning DeepMind Control Suite
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f15bc57-ea9e-49ba-816d-260ba7d40a00 · inbound
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization DeepMind Control Suite
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 939260c3-2479-44fa-b1c5-9df8e07854eb · inbound
Pre-Trained Video Generative Models as World Simulators DeepMind Control Suite
Reference 1998
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fe7e0bf-b594-493b-bf54-57c16464c53e · inbound
Salience-Invariant Consistent Policy Learning for Generalization in Visual Reinforcement Learning DeepMind Control Suite
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53b22708-c14d-4cac-9678-4e0d9d138082 · inbound
Learning Humanoid Standing-up Control across Diverse Postures DeepMind Control Suite
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf10259e-ff76-491c-bc8b-5636cf06afc0 · inbound
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2b0a6ae-571d-47cf-9568-3c9d58e618cf · inbound
Self-Consistent Model-based Adaptation for Visual Reinforcement Learning DeepMind Control Suite
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf8a8f19-063d-426b-8e24-397832700133 · inbound
Causal Information Prioritization for Efficient Reinforcement Learning DeepMind Control Suite
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2838999e-e3fd-40b8-8f79-d6fdce618c82 · inbound
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning DeepMind Control Suite
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 676a07e8-6f2f-4f10-a3f3-d893aae50b66 · inbound
Solving New Tasks by Adapting Internet Video Knowledge DeepMind Control Suite
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b096d802-c938-43e0-bebc-8689bfd7145c · inbound
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities DeepMind Control Suite
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5eca6816-fbfc-4852-a254-0379176db644 · inbound
Q-function Decomposition with Intervention Semantics with Factored Action Spaces DeepMind Control Suite
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5e35a59-79ac-48c1-99df-a0c872ebd9c8 · inbound
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49d62347-1b97-4297-a897-17f6214a9bfe · inbound
Trajectory Entropy Reinforcement Learning for Predictable and Robust Control DeepMind Control Suite
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4924f5c-39ff-4c37-8aa1-a33671b683b7 · inbound
Approximated Behavioral Metric-based State Projection for Federated Reinforcement Learning DeepMind Control Suite
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 512581c5-0282-48d5-a51f-b5d11898a776 · inbound
Learning Diverse Natural Behaviors for Enhancing the Agility of Quadrupedal Robots DeepMind Control Suite
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90bb5f1d-81a0-4d9e-96e5-70bce867cb92 · inbound
ImagineBench: Evaluating Reinforcement Learning with Large Language Model Rollouts DeepMind Control Suite
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b3a41f8-d149-45ed-8003-3e5229198334 · inbound
Zero-Shot Visual Generalization in Robot Manipulation DeepMind Control Suite
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba82ac2c-b88e-4ebc-8a14-984b6a5dac50 · inbound
TD-GRPC: Temporal Difference Learning with Group Relative Policy Constraint for Humanoid Locomotion DeepMind Control Suite
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f29855f9-5e6b-47a6-a869-b01169c27d81 · inbound
Saliency-Aware Quantized Imitation Learning for Efficient Robotic Control DeepMind Control Suite
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea192361-4a4e-46a7-8aef-1e9982e8c62c · inbound
Maximum Total Correlation Reinforcement Learning DeepMind Control Suite
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 600fe287-358e-4acc-8640-28b23aa8ec7b · inbound
ProphetDWM: A Driving World Model for Rolling Out Future Actions and Videos DeepMind Control Suite
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91cf2336-d0c9-4b15-bfcd-b2a770fe4511 · inbound
Beyond Domain Randomization: Event-Inspired Perception for Visually Robust Adversarial Imitation from Videos DeepMind Control Suite
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4797f4a-225b-42be-bab4-c83fa7fde7b1 · inbound
Bigger, Regularized, Categorical: High-Capacity Value Functions are Efficient Multi-Task Learners DeepMind Control Suite
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b36b0e9-bb24-401d-8b6f-dc40103a9cc9 · inbound
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer DeepMind Control Suite
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17e588d6-466d-4205-938a-469714d91c97 · inbound
CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries DeepMind Control Suite
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bdfd756-83ac-44a5-aafa-613e25fea1e5 · inbound
Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments DeepMind Control Suite
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b43741-5e71-4deb-b037-c7410b7eb4ad · inbound
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn DeepMind Control Suite
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 206a9620-d6d5-408f-9f2f-1fe6f85fa7b9 · inbound
LongDWM: Cross-Granularity Distillation for Building a Long-Term Driving World Model DeepMind Control Suite
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74f7e137-fc3e-4938-8567-ff461be12493 · inbound
Self-Predictive Dynamics for Generalization of Vision-based Reinforcement Learning DeepMind Control Suite
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3669cb44-9fb7-4db1-9904-ddb582f30561 · inbound
Dream to Generalize: Zero-Shot Model-Based Reinforcement Learning for Unseen Visual Distractions DeepMind Control Suite
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d0962a6-04d4-4af0-ae2b-2dceeda5975d · inbound
Intention-Conditioned Flow Occupancy Models DeepMind Control Suite
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 60bf0a01-09cd-4194-869f-ddc7a0b9ec41 · inbound
An Open-Source Software Toolkit & Benchmark Suite for the Evaluation and Adaptation of Multimodal Action Models DeepMind Control Suite
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a9c1ad1-640c-4584-826a-93eef9aba65b · inbound
Multi-Task Reward Learning from Human Ratings DeepMind Control Suite
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e3cff0f-aa7f-4865-8ef2-41ca97ea3b63 · inbound
SkillBlender: Towards Versatile Humanoid Whole-Body Loco-Manipulation via Skill Blending DeepMind Control Suite
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca32db8c-638b-4247-8b9d-8dbecc0ca744 · inbound
Flow-Based Policy for Online Reinforcement Learning DeepMind Control Suite
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 364092c5-6388-4b31-afdc-14272e675950 · inbound
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DeepMind Control Suite
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ea88ba8-32ca-403b-9456-f0b60699cb2d · inbound
PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning DeepMind Control Suite
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66b480bf-5b9a-445a-93bd-0ddbd0d3441a · inbound
Unsupervised Skill Discovery through Skill Regions Differentiation DeepMind Control Suite
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2f8887-3798-40b0-8beb-df1ae25b0f29 · inbound
Zero-Shot Reinforcement Learning Under Partial Observability DeepMind Control Suite
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 884867cb-5336-4f9c-8768-509e25106d40 · inbound
Network Sparsity Unlocks the Scaling Potential of Deep Reinforcement Learning DeepMind Control Suite
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9264219-78f5-4415-86dc-f119aa5a0721 · inbound
Unsupervised Data Generation for Offline Reinforcement Learning: A Perspective from Model DeepMind Control Suite
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69566d21-3b66-42e4-b757-8d852cfba287 · inbound
rQdia: Regularizing Q-Value Distributions With Image Augmentation DeepMind Control Suite
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f76eac-f5de-486b-a84d-587b9d6015f2 · inbound
RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation DeepMind Control Suite
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f34593ff-a47b-41dd-b8a5-c7ec3e295d37 · inbound
Distributional Soft Actor-Critic with Diffusion Policy DeepMind Control Suite
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3b7639b-7c2d-4fb1-9739-cf884d6fc7b2 · inbound
A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control DeepMind Control Suite
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62811960-c9fe-4002-94d9-59eee7640de7 · inbound
Epistemically-guided forward-backward exploration DeepMind Control Suite
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 988eac78-cb55-455f-8fa4-138eb22a2724 · inbound
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making DeepMind Control Suite
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbeeeba4-80c0-4693-9d85-79b85eef1422 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities DeepMind Control Suite
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c115efa-d84f-46cc-9419-68c1c54c748a · inbound
Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning DeepMind Control Suite
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4537ca1-3755-46b8-9780-389e03224349 · inbound
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68c89edc-604f-4cc5-9864-f4518319b735 · inbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks DeepMind Control Suite
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a627790-67da-4c89-ad59-60d9bc95771b · inbound
Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies DeepMind Control Suite
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671c5a40-c46a-437c-bf47-dbadb5d87f7f · inbound
Robust Remote Reinforcement Learning over Unreliable Communication Channels using Homomorphic State Encoding DeepMind Control Suite
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 41275ee9-78e0-4b8f-900f-bd5c02bcc698 · inbound
Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges DeepMind Control Suite
Reference 106
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7c63e72-f786-46c6-b334-0d12c0924c5c · inbound
Arnold: a generalist muscle transformer policy DeepMind Control Suite
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f64fb178-d907-4734-a8d3-54012a4eafb9 · inbound
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning DeepMind Control Suite
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6b28899f-3b84-4e68-9dde-d8958becdc0e · inbound
D2 Actor Critic: Diffusion Actor Meets Distributional Critic DeepMind Control Suite
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8fc0e4de-9b14-4c3c-966c-2eb11a5a805c · inbound
PAC-Bayesian Reinforcement Learning Trains Generalizable Policies DeepMind Control Suite
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.