Pith. sign in

Paper Citation Record · LEDGER

Evaluating Reinforcement Learning Algorithms in Observational Health Settings

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:1805.12298.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1805.12298 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:04:38.836400Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T08:49:42.747936Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c85aec80-3401-46e5-a3e0-edbbcdfea887 · inbound

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems cites this paper.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 214

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:21.583455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:4e9894fb6daf3454f2bcb0d3c0980095be0ff4a40c4d029213d042966feb1c58

Observation 22c0192b-9c1c-4cbf-a725-bdb8b0389629 · inbound

Treatment, evidence, imitation, and chat cites this paper.

Treatment, evidence, imitation, and chat Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:22:10.855492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T08:21:05.698812Z digest=sha256:046f964c18616be5e6ee9a514768009e70807b537d8dbc34ddf7ccd511035df7

Observation 22e4a638-523a-46ca-bf1f-f1a50b3792db · inbound

Pragmatic Policy Development via Interpretable Behavior Cloning cites this paper.

Pragmatic Policy Development via Interpretable Behavior Cloning Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T15:04:38.836400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:04:38.836400Z digest=sha256:b051b2265eb2bcb81a188750a10d0395a22fed03819636b381083032f27bf0da

Observation 443f21f4-e2ed-4a34-81a7-0c1ba257ae11 · inbound

Outcome-Aware Spectral Feature Learning for Instrumental Variable Regression cites this paper.

Outcome-Aware Spectral Feature Learning for Instrumental Variable Regression Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-03T19:26:22.334398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:26:22.334398Z digest=sha256:efa45227756f7d8123f22061167703b7f6dde79db869e4d38a3362f4d2218ed6

Observation 34d53b90-b855-42ad-bde2-7c6f50195ee7 · inbound

Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning cites this paper.

Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T20:12:54.889602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T20:12:19.948073Z digest=sha256:a1a9f49168bdc4bbebc507384cbd1f1e020cca163c3e6e5b7034bcf770edfc4e

Observation cdd0321d-a40a-404c-92da-52a0c69b6e62 · inbound

MedGym:A Unified Continuous-Time Benchmark for Dynamic Medical Treatment Reinforcement Learning cites this paper.

MedGym:A Unified Continuous-Time Benchmark for Dynamic Medical Treatment Reinforcement Learning Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T21:06:13.701308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T17:29:37.376343Z digest=sha256:66bf10d4ad6132840771a47a01e66a7dce5b97f6ee66a51fbc9f4632ed40a4e4

Observation b01709d8-0743-45a2-9610-4ebf2aa32f5f · inbound

Stationary Robust Mean-Field Games under Model Mismatches cites this paper.

Stationary Robust Mean-Field Games under Model Mismatches Evaluating Reinforcement Learning Algorithms in Observational Health Settings

Reference 236

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T08:49:42.749139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T10:50:40.841967Z digest=sha256:180d27e2853d825ba92059dfe25a2760758054cca830a2998be8cf0cdff714d7