Pith. sign in

Paper Citation Record · LEDGER

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes

As of 17 August 2026, this Paper Citation Record lists 6 of 6 outbound references and 1 inbound Pith citation observation for arXiv:2512.14617.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.14617 v2

Coverage vector

measured 6 of 6 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:15:27.229715Z

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:43:00.121084Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T15:43:00.548193Z

Reference resolution

6 of 6 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 68411146-cfa5-4866-8cda-884838c60776 · outbound

This paper cites simulation-lemma.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes simulation-lemma

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.838755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.838755Z digest=sha256:212bf38cec998d3b3ca10f9b9f199fa2dd411567b94ead8df8f5f185cbef124c

Observation b9c49a03-2702-4384-8c3a-d6b0f4079b78 · outbound

This paper cites an unresolved cited work.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:27.021872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:27.021872Z digest=sha256:d2ff54b9e75a1ec7783aef51f8d63b21386a22722397e2857b69b1ec9cb51e42

Observation fecd9376-00bc-4af6-8f2f-2ba42e1f7d61 · outbound

This paper cites error + V πt ¯M (b, q)−V ∗ ¯M (b, q) | {z } est.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes error + V πt ¯M (b, q)−V ∗ ¯M (b, q) | {z } est

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:27.229715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:27.229715Z digest=sha256:06890868a0898c50914d95a4a7792c75e230415d178c237b448ac42fb1c1afa5

Observation 706d583f-ad28-4a7d-ae7d-d56150517726 · outbound

This paper cites Near-optimal Reinforcement Learning in Factored MDPs.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes Near-optimal Reinforcement Learning in Factored MDPs

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.717241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.717241Z digest=sha256:292d0bd26974df335588e5af0ede50232e6c2293b4b0696c2a3693388ab7995a

Observation 4bfe4a1d-56f7-4b16-a973-fa577001f1b7 · outbound

This paper cites InProc, ICAPS, volume 34, 13659–13662.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes InProc, ICAPS, volume 34, 13659–13662

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.591251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.591251Z digest=sha256:0ddd916580dc24c5cdb505e1a529efbc94b75f9a61d361882106c3b546a572d9

Observation a0d59e86-7840-4126-bf1b-495cf87f8e6e · outbound

This paper cites InInternational Conference on Artificial Intelligence and Statistics (AISTATS), 4114–4146.

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes InInternational Conference on Artificial Intelligence and Statistics (AISTATS), 4114–4146

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T16:15:26.476410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:15:26.476410Z digest=sha256:2de62109a31a0b2d2c21f3da1b76dc3ffc8be77d886e4e771efed789b426c106

Pith citing papers

Observation a3287bdd-5ede-4ab7-bf89-b864090bc05b · inbound

Theoretical Foundations of $\max$@$k$ Reinforcement Learning cites this paper.

Theoretical Foundations of $\max$@$k$ Reinforcement Learning Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-15T15:43:00.554958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:43:00.121084Z digest=sha256:e46caad26b93f1e7c99244a7ed6443d08153ee9a6107888a964d86a00e1129a4