Pith. sign in

Paper Citation Record · LEDGER

Constructing Non-Markovian Decision Process via History Aggregator

As of 22 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2506.24026.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.24026 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:45:13.631540Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-19T01:06:07.789346Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T01:06:57.322870Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 57842640-8dd7-4525-87e7-f169d00bfbec · outbound

This paper cites Learning and Solving Regular Decision Processes.

Constructing Non-Markovian Decision Process via History Aggregator Learning and Solving Regular Decision Processes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:11.780663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:11.780663Z digest=sha256:1a9199f3a801cc0163edc4262c5bf3c33cb15fde35966de079af68fe73d321ec

Observation 7b9c6881-b6a1-4201-96b8-00a6ca652048 · outbound

This paper cites Regular decision processes: A model for non-markovian domains.

Constructing Non-Markovian Decision Process via History Aggregator Regular decision processes: A model for non-markovian domains

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:18.034528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:11.863655Z digest=sha256:e9f2cda61dea18431457730154271c6cc0c53acfe8999d9d11864eb73f7a30b8

Observation 2ab29340-30e6-4d0b-a13b-9c267846545d · outbound

This paper cites Ltl and beyond: Formal languages for reward function specification in reinforcement learning.

Constructing Non-Markovian Decision Process via History Aggregator Ltl and beyond: Formal languages for reward function specification in reinforcement learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:17.842948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:11.945712Z digest=sha256:c205a52ca764e0c88f6de420c1431f0b9c092be6ec5496a9eabea42b1640caf9

Observation a07341f1-0e74-4c57-a9db-2e180f9b549b · outbound

This paper cites Reinforcement learning in non-markovian environments.

Constructing Non-Markovian Decision Process via History Aggregator Reinforcement learning in non-markovian environments

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:17.703454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.034868Z digest=sha256:d8c09f03ed237fe5ce606cd0d97bf7f729001d07f6a7ea40815820912af01309

Observation edce5df5-7ec1-427d-9b57-9a77aa77d334 · outbound

This paper cites Dynamics of non-markovian open quantum systems.

Constructing Non-Markovian Decision Process via History Aggregator Dynamics of non-markovian open quantum systems

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:17.561257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.188498Z digest=sha256:1f3530aacf3ec302e9188672b040e63f0c7454f11fe199a3a95efa488d700a22

Observation c32cec99-76c3-43f8-8239-5a053ddb37f3 · outbound

This paper cites Rl-baselines3-zoo, 2024.

Constructing Non-Markovian Decision Process via History Aggregator Rl-baselines3-zoo, 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:17.382502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.287244Z digest=sha256:9ed0aa19d4c9cb7a992afcc506f66ef888df720b6c92d9e3047be9eae0d36566

Observation 2af57d71-dbd3-447a-b4a2-fff599cefa49 · outbound

This paper cites Stable-baselines3, 2024.

Constructing Non-Markovian Decision Process via History Aggregator Stable-baselines3, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:17.159616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.356057Z digest=sha256:a253277a7655c91ad8c1c2c4f40cf96138cf8fe2de882578a925832a93f51e9f

Observation f2317c3c-95ff-4179-9cf0-61f1cad90578 · outbound

This paper cites Inferring probabilistic reward machines from non-markovian reward signals for reinforcement learning.

Constructing Non-Markovian Decision Process via History Aggregator Inferring probabilistic reward machines from non-markovian reward signals for reinforcement learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:17.004282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.437165Z digest=sha256:419c1d69b07123619c2280211e399ad66cfa334e5a85d244e9b8435d841886de

Observation 7d324916-2cda-442e-8ac6-6e3b63818d84 · outbound

This paper cites Gymnasium, 2023.

Constructing Non-Markovian Decision Process via History Aggregator Gymnasium, 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:16.808415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.521547Z digest=sha256:fa863a8e0fb4840327132ae919e66c83fa8580d9357e083ed8802ec257e44b79

Observation 6e2dc753-940e-4593-a543-b6fcb5137829 · outbound

This paper cites Reinforcement learning with non-markovian rewards.

Constructing Non-Markovian Decision Process via History Aggregator Reinforcement learning with non-markovian rewards

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:16.569043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.600611Z digest=sha256:be68382e4bbba9e16e731f22aa69eb3eaf38144d3a90d432ffdaf5c9320ba039

Observation 0e10c117-b73c-4caa-8a8b-298cdef37d22 · outbound

This paper cites Non-markovian reinforcement learning using fractional dynamics.

Constructing Non-Markovian Decision Process via History Aggregator Non-markovian reinforcement learning using fractional dynamics

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:16.256984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.678946Z digest=sha256:523fd468e2eb02ec948768a600624eb7a36850b546e1ec9390249f0c5155766e

Observation a88bcad1-5b00-4f06-abbb-79e56b382f31 · outbound

This paper cites Feature reinforcement learning: Part i.

Constructing Non-Markovian Decision Process via History Aggregator Feature reinforcement learning: Part i

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:15.925816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.773516Z digest=sha256:6917b47ad5b05639725857fc97403d006f14c081a3b5bb810b5192c09be5caa3

Observation 10dcacb8-d487-4619-8191-02492f15add6 · outbound

This paper cites The sample-complexity of general reinforcement learning.

Constructing Non-Markovian Decision Process via History Aggregator The sample-complexity of general reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:15.641125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.864855Z digest=sha256:d84347e34ededb18cf15ff3d030308fd7eaf8e5d0c7e4657a25ea34b401b9d30

Observation e83c53f6-c243-47e2-9261-f29434b67da6 · outbound

This paper cites Basic category theory , volume 143.

Constructing Non-Markovian Decision Process via History Aggregator Basic category theory , volume 143

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:12.926925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:12.926925Z digest=sha256:ffb849814de5b49b3748fec64e1b5d7414f99345542c30833eb7dbb73de83598

Observation 9f8318c8-35ba-4a53-b448-af111fcf08d6 · outbound

This paper cites Selecting the state-representation in reinforcement learning.

Constructing Non-Markovian Decision Process via History Aggregator Selecting the state-representation in reinforcement learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:15.406162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:12.995606Z digest=sha256:ddde937f706e281461cf8f9b475c071675e1b9cd3d4119a736f3a525481f95f8

Observation 2d76b730-ceca-49d8-9f27-c9349100913b · outbound

This paper cites On q-learning convergence for non-markov decision processes.

Constructing Non-Markovian Decision Process via History Aggregator On q-learning convergence for non-markov decision processes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:15.216850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.063997Z digest=sha256:2231107db768ebacae8554880c702eb8b50bb4a4f7836d65f582b6dbea9851ab

Observation 4c8a0e3a-3efc-4bb2-94e7-97ec9a671b95 · outbound

This paper cites Competing with an infinite set of models in reinforcement learning.

Constructing Non-Markovian Decision Process via History Aggregator Competing with an infinite set of models in reinforcement learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:15.045137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.117734Z digest=sha256:0356913dbd414c65d1064c26d40d80334e600ad11ee691b5f70822178a02684b

Observation 22601f24-f5c4-4472-811a-9cedb1d65401 · outbound

This paper cites Learning non-markovian decision-making from state-only sequences.

Constructing Non-Markovian Decision Process via History Aggregator Learning non-markovian decision-making from state-only sequences

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:14.789269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.165978Z digest=sha256:2e2be7b163c4451e0f7407badf98fa5a114fe52d249fa099b0c39b2b463c9a7a

Observation 93124684-39be-43f0-a505-9e378793009e · outbound

This paper cites Learning Non-Markovian Reward Models in MDPs.

Constructing Non-Markovian Decision Process via History Aggregator Learning Non-Markovian Reward Models in MDPs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:13.232447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:13.232447Z digest=sha256:2f4f754bbf572c52162c254f3255fae05eaa9999d46507130839cc33e2ce0d04

Observation 77a9e218-83dd-4d5b-bdb8-84e5b42ce8b8 · outbound

This paper cites Markov Abstractions for PAC Reinforcement Learning in Non-Markov Decision Processes.

Constructing Non-Markovian Decision Process via History Aggregator Markov Abstractions for PAC Reinforcement Learning in Non-Markov Decision Processes

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:45:13.804821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.324290Z digest=sha256:9a930cb466a3115e234ca236b4e07817c2d98a36859789e132705b87a921e649

Observation 62617c73-04cc-45be-ad57-7150db73969e · outbound

This paper cites Stable-baselines3-contrib, 2024.

Constructing Non-Markovian Decision Process via History Aggregator Stable-baselines3-contrib, 2024

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:14.516273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.408775Z digest=sha256:289044914660bc85c138e13dfc0bde9279b59e5b9422cd26f4a300fcb3342e98

Observation ebbbcfef-5172-4254-981d-1cde572d1d98 · outbound

This paper cites A monte-carlo aixi approximation.

Constructing Non-Markovian Decision Process via History Aggregator A monte-carlo aixi approximation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:14.337391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.472802Z digest=sha256:ea27126f6dc8ca344596227d4cfcb4b51843740b6e11e31c9b6f77cfe0e047e7

Observation 869dd012-a7e2-4b45-9e83-b552b3df264b · outbound

This paper cites Reinforcement learning of non-markov decision processes.

Constructing Non-Markovian Decision Process via History Aggregator Reinforcement learning of non-markov decision processes

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:45:14.065279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T21:45:13.565879Z digest=sha256:4af71e509a8e795ccca0f5857c02792a8504538361d9510c45615a75ea3d951d

Observation a38ea7d9-f9b3-405a-87c3-9406a96e2d1c · outbound

This paper cites write newline.

Constructing Non-Markovian Decision Process via History Aggregator write newline

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:13.631540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:13.631540Z digest=sha256:f2a3575f08836dc5f8d46eca38d7d4a3db31fea6bd38cf70debc04f398264a93

Pith citing papers

Observation 96355ccc-13e6-45c0-b9b1-5503b98c3f7c · inbound

Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling cites this paper.

Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling Constructing Non-Markovian Decision Process via History Aggregator

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:06:57.326486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T01:06:07.789346Z digest=sha256:18ea0a2131f52850c9217ee23e59a25b91b5ad49517d1d0f801df88c402e5707