Pith. sign in

Paper Citation Record · LEDGER

PWM: Policy Learning with Multi-Task World Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2407.02466.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.02466 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:47:20.450908Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:44.908311Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 21171bc0-8935-4a9a-b243-2dcc2542cc23 · inbound

TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents cites this paper.

TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents PWM: Policy Learning with Multi-Task World Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:47:20.450908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:47:20.450908Z digest=sha256:8ca691c22341676057b552f21ce809d0be0acb9c05ba09fd55070f44c3d4c6c3

Observation 693e96f3-7121-4972-a386-c386cad89637 · inbound

First Order Model-Based RL through Decoupled Backpropagation cites this paper.

First Order Model-Based RL through Decoupled Backpropagation PWM: Policy Learning with Multi-Task World Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T13:55:20.496902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:55:20.496902Z digest=sha256:820ce309eb0fadc1b2ff71e078fad63e11ea3381e1c147a1e1a38709dc4398c8

Observation c724f07f-d680-4eaa-8913-584157685fc4 · inbound

Coupled Local and Global World Models for Efficient First Order RL cites this paper.

Coupled Local and Global World Models for Efficient First Order RL PWM: Policy Learning with Multi-Task World Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T04:03:56.304547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:03:56.304547Z digest=sha256:ffdc3c0d91d377bc08713b52a8c783dfb5b240363cb56f67296907e78db5c697

Observation 55b6c57f-8fa7-472c-b95a-99a30f340472 · inbound

Toward Safe Autonomous Robotic Endovascular Interventions using World Models cites this paper.

Toward Safe Autonomous Robotic Endovascular Interventions using World Models PWM: Policy Learning with Multi-Task World Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:59:49.492803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T00:58:30.063604Z digest=sha256:944f5594463b821ca7535e61d0e389084794e48f65742c55662c63d482cc079d

Observation 95903312-9e3b-40a1-bc1c-9b35b96f71bb · inbound

Intend, Reflect, Refine: An Adaptive Multimodal Reflection Framework for Autonomous Driving cites this paper.

Intend, Reflect, Refine: An Adaptive Multimodal Reflection Framework for Autonomous Driving PWM: Policy Learning with Multi-Task World Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:44.909718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T09:06:54.669489Z digest=sha256:60fce65c944828a7487ee6dd357cae1220a679d50d1a35fb35861965004d3264

Observation c11de667-f0c0-4929-91b8-855ae9d681c3 · inbound

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models cites this paper.

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models PWM: Policy Learning with Multi-Task World Models

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-02T04:49:19.477323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:49:19.477323Z digest=sha256:798ad4f53b71da9eeae5a393902b9900d88be3d181e1ddaee9e04bacad59b2bf