Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:06:27.128981Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2507.06615.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:06:27.128981Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6956b067-8341-46e0-bcbd-64795b02225e · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Dynamic programming
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af101913-8a09-4c41-8e5b-c7de8da6b9f2 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance OpenAI Gym
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34431820-cad1-4701-beac-8212532aa7fc · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Multitask learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0c35eeb8-bef0-40c6-9f1b-7770423608ce · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Gradnorm: Gradient normalization for adaptive loss balancing in deep multitask networks
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d248c66f-4a2b-44e9-91d2-24c7eca148c3 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Multi-task reinforcement learning with task representation method
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9de1f464-3e53-4755-8767-6d82f0270712 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Soft Actor-Critic for Discrete Action Settings
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ead5de88-ef60-4ba9-b730-12769cb6e820 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Divide- and-conquer reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6ef8dc4f-3342-4c2c-86c4-449f807a9ccb · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Reinforcement learning with deep energy-based policies
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 16bb38e4-780d-4e03-91d3-dec7f354261c · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0ebc75cf-1afd-4094-845d-af7ef6f61713 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Not all tasks are equally difficult: Multi-task deep reinforcement learning with dynamic depth routing
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4eb6ac94-69b8-4c8a-abea-48aed0d715bf · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Benchmark environments for multitask learning in continuous domains
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a8939339-cf84-4fdc-9900-350b4174c8ef · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance End-to-end training of deep visuomotor policies
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2176921b-bc88-4841-b2a8-0ba171809721 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Lillicrap, Jonathan J
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87174cbe-8b2a-4034-bce1-3a1152b3dbc6 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Conflict-averse gradient descent for multi-task learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b49c235-0949-4445-82fb-b8b63b001962 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Q- functionals for value-based continuous control
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1db05804-1920-4860-8fee-fadb7139911b · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Rusu, Joel Veness, Marc G
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f258522-cbea-488b-970d-d9e46d6f86fa · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Overcoming exploration in reinforcement learning with demonstrations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 163539aa-9b26-4d44-94bf-750e50195f4e · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Markov Decision Processes: Discrete Stochastic Dynamic Programming
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2194e526-57fb-41da-a89e-54786c8fc73d · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance An Overview of Multi-Task Learning in Deep Neural Networks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bef9725-1cdc-4db8-b924-a907e05a1350 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Hierarchical and interpretable skill acqui- sition in multi-task reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f460c02b-9ff7-401b-b39c-8a451cc4b64f · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Multi-task reinforcement learning with context- based representations
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f79f070-122f-407e-b8fc-4b0d5f47fbe5 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance PaCo: Parameter- compositional multi-task reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b9f67cb-074c-445a-83ca-e899be6f3b62 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Czarnecki, John Quan, James Kirkpatrick, Raia Hadsell, Nicolas Heess, and Razvan Pascanu
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 158b1812-c2ad-4633-84a0-8b8e28132152 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance MuJoCo: A physics engine for model-based control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation efcaf77a-b41a-4b39-b64b-899e88b3409e · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance A survey of multi-task deep reinforcement learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f491275-5314-4ac8-ae9a-aca732c37924 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Disentan- gling transfer in continual reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f63fe28-5d59-4ced-b086-b51e99ab9290 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Multi-task reinforcement learning with soft modularization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e64095ba-b40e-49bf-82fa-fdc5b6a898cc · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Mastering complex control in MOBA games with deep reinforcement learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 42273352-30d9-4d66-b59b-9720cee13b2d · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Conservative data sharing for multi-task offline reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6980ee7d-36ad-44b6-8cc9-56000f31aaa3 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Gradient surgery for multi-task learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f910f7ae-4f18-40d3-a21a-d8be669fb90d · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Meta-World: A benchmark and evaluation for multi-task and meta reinforcement learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ac343e8c-760f-49a8-ba43-6f5424df160a · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance QMP: Q-switch Mixture of Policies for Multi-Task Behavior Sharing
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 471e7da7-456e-48e7-941e-bb1777e50947 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance CUP: Critic-guided policy reuse
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 25376279-8414-45bf-8121-e2fe020ec629 · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance t+K−1X t′=t γt′−t (Ri(st′, at′) +αiH(πi(·|st′))) # + γKEst+K∼Pi h ˆQ˜g i (st+K, i) i = Eat′ ∼πi,st′+1∼Pi
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ef72a0e4-770a-45bb-ab68-b30d2601b72d · outbound
Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance Under the initial episode length setting of 150 timesteps, 42 tasks achieve a success rate exceeding 90%
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.