Pith. sign in

Paper Citation Record · LEDGER

On the role of planning in model-based deep reinforcement learning

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2011.04021.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2011.04021 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:31:07.765568Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:29:51.992455Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b235f8dc-118b-4247-878d-df8dc60994be · inbound

A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control cites this paper.

A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control On the role of planning in model-based deep reinforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:31:07.765568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:31:07.765568Z digest=sha256:b9d45e968dccb9016286e798a0ef49d41f52ad83472c6b489365719887ed0cae

Observation a1dfd49d-3e45-492d-8fed-bd514783d995 · inbound

Bounding Distributional Shifts in World Modeling through Novelty Detection cites this paper.

Bounding Distributional Shifts in World Modeling through Novelty Detection On the role of planning in model-based deep reinforcement learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T22:58:40.394538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:58:40.394538Z digest=sha256:6147f7f4168892feb338a49b9710454bbc8f6aea60a04830b325aa58aecf9a8c

Observation 46e32da9-dfa3-4e35-8e45-545312fb4b04 · inbound

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning cites this paper.

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning On the role of planning in model-based deep reinforcement learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:36:09.927407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-09T15:44:36.262834Z digest=sha256:b9d6c4df08ddd19ba21f73ff0ff1b8f9453a4d5cc06d8adcd874e60d7778290d

Observation 06f6ef56-7f5c-4095-a07b-72675bfd4dce · inbound

Learning to Theorize the World from Observation cites this paper.

Learning to Theorize the World from Observation On the role of planning in model-based deep reinforcement learning

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:21:26.250603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-07T17:15:43.429602Z digest=sha256:e8a10a6c7ce82b08f303d9364de599ab49a731b333124fb4ffd1043bdad5854f

Observation 30002609-2fb4-40a8-b118-0acdbe5a7820 · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL On the role of planning in model-based deep reinforcement learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.993954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T05:08:19.504454Z digest=sha256:aeb4f52090ca7e66120a723e8bbdcc5792f39d3e7bfc05f46a1cb78052d62fdc

Observation b31e46be-e7d9-40bd-acca-38e514eb0a16 · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL On the role of planning in model-based deep reinforcement learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:34:34.862751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-30T09:26:31.944405Z digest=sha256:20253a3899c9b2714f0dce7ca2f624a1948b520df6b7d1a0617bb357882657f8