Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:47:20.742361Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.01823.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:47:20.742361Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 04e8cbc0-f772-4358-9ce7-06de442aa9d1 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Learning dexterous in-hand manipulation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9263eaca-20d5-46ab-8758-8c32302f40e5 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Multitask learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c0e3251-fa33-452d-9d3b-12d0c98feaee · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Deep reinforce- ment learning in a handful of trials using probabilistic dynamics models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1468e31-2710-4527-8c86-414adf765016 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Distilling policy distillation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c1a1e3d9-3583-458f-8612-acbcbdb931ac · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Model-agnostic meta-learning for fast adaptation of deep networks
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bcaecc37-eca4-461b-a67f-6cf8a4848649 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Model predictive control: Theory and practice—a survey
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21171bc0-8935-4a9a-b243-2dcc2542cc23 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents PWM: Policy Learning with Multi-Task World Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc07b9b3-bef6-4e53-bf11-8ff2c6c5b35c · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4157507-8807-432c-8c42-20a3d1227fad · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ae87c9e5-3cd4-492a-9409-56fc1ef5fa57 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Dream to Control: Learning Behaviors by Latent Imagination
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32e106f6-2853-4358-a9fc-d2896720d3ec · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Mastering Diverse Domains through World Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d40029d-22c3-4768-b5a5-d251ef5f00fa · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents TD-MPC2: Scalable, Robust World Models for Continuous Control
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d46e977-1dfb-4464-b064-fed8fd283b8e · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Temporal difference learning for model predictive control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27ba02f8-e9c5-41b7-944d-3c14db794a55 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Distilling the knowledge in a neural network
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5348b85e-9f2b-4581-aa07-f682510f00ac · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents When to trust your model: Model-based policy optimization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b8835014-03a8-4cfd-8a00-e47530bd3960 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86f32bda-427f-4723-aa9e-e3279ec207e5 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Model- ensemble trust-region policy optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb48e649-4b62-4f98-bddb-3164cd34d21a · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Knowledge transfer in model-based reinforcement learning agents for efficient multi-task learning, 2025
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 730faa61-02d7-4fd2-ab3a-9bb662814a4e · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Multimodal reinforcement learning: A survey and taxonomy
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a7ab240-a1a4-4a21-a017-898f96770a69 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents End-to-end training of deep visuomotor policies
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5a911dd-e79b-49be-a946-c364248cf997 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783c56b4-a279-4c40-83f3-f2bed757516c · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Mixed Precision Training
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b6c0716-2fa1-4c1f-ad24-0ddef9e18747 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Playing atari with deep reinforcement learning, 2013
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb37187b-bee5-4690-8fff-bb312802380e · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Human-level control through deep reinforcement learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36319905-9737-4917-b7e8-8be34e6a0463 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Curriculum learning for reinforcement learning domains: A framework and survey
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6eb62eab-a102-4875-a46f-4e3602fb9356 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Actor-Mimic: Deep Multitask and Transfer Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3269a788-29a1-41cb-8b7d-4c914b02014d · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Policy Distillation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 302f9915-3495-492a-aec4-3b3abd2e996b · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Deep q-learning with quantized neural networks
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2cf4044-61c7-444f-902d-c6a24dd9a4ac · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Decoupling Representation Learning from Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f97136bf-5aa8-4c47-b6d4-b021f26006f6 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Learning to predict by the methods of temporal differences
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d15d1547-8fb8-4f92-982a-973fe84e01fb · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Policy gradient methods for reinforcement learning with function approximation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3be07ffd-7cf5-420b-a070-3daf4ac3127c · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents dm_control: Software and Tasks for Continuous Control
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ace5a2a-ea29-4d25-a4d0-dc9b791c8d92 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Distral: Robust multitask reinforcement learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 468cf8ab-d718-4d55-b6ae-ae543a14a9ec · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ee31bc2-f3f4-4faf-9008-9b8082672b42 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents Conservative q-learning for offline reinforcement learn- ing
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 382e6af6-99ca-431b-83f3-ecad44d038a3 · outbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents A survey on multi-task learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.