Pith. sign in

Paper Citation Record · LEDGER

A Graphical Approach to State Variable Selection in Off-policy Learning

As of 23 August 2026, this Paper Citation Record lists 4 of 4 outbound references and 0 inbound Pith citation observations for arXiv:2501.00854.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.00854 v1

Coverage vector

measured 4 of 4 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:49:44.705573Z

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

4 of 4 outbound references displayed

  • verified exact1
  • verified fuzzy1
  • unresolved1
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2c7dcd50-e8d8-4c44-ba19-baa5d04cc952 · outbound

This paper cites Optimal control of Markov processes with incomplete state infor- mation.

A Graphical Approach to State Variable Selection in Off-policy Learning Optimal control of Markov processes with incomplete state infor- mation

Reference 1

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T22:49:44.840397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T22:49:44.691672Z digest=sha256:4356a9bb6782e22f23b734acdcf248573cb94b67cabff2ecc4283114b6ffd29d

Observation 3bec2679-d1ee-4554-b690-470b643442a8 · outbound

This paper cites Constructing dynamic treat- ment regimes over indefinite time horizons.

A Graphical Approach to State Variable Selection in Off-policy Learning Constructing dynamic treat- ment regimes over indefinite time horizons

Reference 76

Resolution
verified exact
raw_fallback, observed 2026-08-10T22:49:44.811276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T22:49:44.696494Z digest=sha256:2c71da1193b4895493c2edbf7346145375816229173357ab626d45c203709a7c

Observation 61ec9b4b-b276-463d-a1f3-fd9d0bbc8b8b · outbound

This paper cites Off-policy evaluation in partially observed Markov decision processes under sequential ignorability.

A Graphical Approach to State Variable Selection in Off-policy Learning Off-policy evaluation in partially observed Markov decision processes under sequential ignorability

Reference 311

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:49:44.827159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T22:49:44.700504Z digest=sha256:91f5ebd2ce0148e5515e125d8d4b1e3f8ca0059f5ccdf910247a2bb033464759

Observation 1ed861f4-25aa-4860-974f-4099b90c232b · outbound

This paper cites A Review of Off-Policy Evaluation in Reinforcement Learning.

A Graphical Approach to State Variable Selection in Off-policy Learning A Review of Off-Policy Evaluation in Reinforcement Learning

Reference 688

Resolution
unresolved
no resolver link, observed 2026-08-10T22:49:44.705573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:49:44.705573Z digest=sha256:d3b6672049f68a2d8bd379c72d0aae884dbfbc35f4cfffc844ffa166280316ca

Pith citing papers

No inbound Pith citation observations are available.