Pith. sign in

Paper Citation Record · LEDGER

A Definition of Continual Reinforcement Learning

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2307.11046.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.11046 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T04:39:32.143002Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

9
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5010f47a-b89c-4f23-bfb8-9b9c404d5484 · inbound

Optimal control of the future via prospective learning with control cites this paper.

Optimal control of the future via prospective learning with control A Definition of Continual Reinforcement Learning

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T23:05:25.140051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T23:04:54.956095Z digest=sha256:6cd598e26accb2d0c68dc13dc21ff9279dce2df6dd27f8e1f3e3f48fc92461d5

Observation f33ef4e3-caea-41c7-9df6-15c81b830f82 · inbound

LIFE -- an energy efficient advanced continual learning agentic AI framework for frontier systems cites this paper.

LIFE -- an energy efficient advanced continual learning agentic AI framework for frontier systems A Definition of Continual Reinforcement Learning

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T15:50:33.822273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:48:38.759167Z digest=sha256:3d182f0cbbfa24f969be845fc617488059f9fe4c7d8b44e72bb083fd34dfd724

Observation 0da4288e-9a34-4605-acf7-71c73085f172 · inbound

Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought cites this paper.

Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought A Definition of Continual Reinforcement Learning

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:31:00.574005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-11T01:15:41.980346Z digest=sha256:ec3cf19997818d1be45480e5347a33e6f74fc1c1ad15cf6d1b93a1d6a4e066f9

Observation 2e5e56d3-7564-4c53-b48f-e87d48ac7841 · inbound

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning cites this paper.

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning A Definition of Continual Reinforcement Learning

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:35:58.632333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-11T01:13:34.836246Z digest=sha256:14bcba61eeac184bc16a77a6967af2a90967bd78b4b54c0afc7022935060c902

Observation 0dee8535-da51-443e-8a4a-a8ee3434d3cc · inbound

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning cites this paper.

Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning A Definition of Continual Reinforcement Learning

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:29:13.110473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-20T23:26:35.072127Z digest=sha256:c09188e7ec313e3136a7ac3551f75c76a315b7a62ae773b462d762d2e69ff07a

Observation 653fccfe-6c8b-4d2a-a556-d2faa10e6dfc · inbound

Adaptive Multi-Horizon Reinforcement Learning cites this paper.

Adaptive Multi-Horizon Reinforcement Learning A Definition of Continual Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T09:46:28.773333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T09:46:28.773333Z digest=sha256:29d02bec71fa24675474c5598cf2ca4d8de307ef9c8ccfdf833e6879c7db0689

Observation 194bd6c0-e714-41b6-9b8e-7097742fe61e · inbound

Continual-RL for Generalization in Autonomous Racing on the RoboRacer Platform cites this paper.

Continual-RL for Generalization in Autonomous Racing on the RoboRacer Platform A Definition of Continual Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T18:18:00.250483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T18:18:00.250483Z digest=sha256:3d78444e82b1af22084832fae09c486488f3f9debacca9fa75262b81d1930fb9

Observation 7e69388d-a935-45c0-8b90-449b23d3c3b0 · inbound

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback cites this paper.

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback A Definition of Continual Reinforcement Learning

Reference 251

Resolution
unresolved
no resolver link, observed 2026-08-03T04:39:32.143002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T04:39:32.143002Z digest=sha256:829a30a1412ddd17fbdb289ebfca57589531272b0c4896338520504765bfaad8