Pith. sign in

Paper Citation Record · LEDGER

Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2008.07146.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2008.07146 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:41:32.001981Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 82f8ef6c-baae-4381-8028-2f91e5cc99fc · inbound

Off-Policy Evaluation and Counterfactual Methods in Dynamic Auction Environments cites this paper.

Off-Policy Evaluation and Counterfactual Methods in Dynamic Auction Environments Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:18:15.884203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:18:15.884203Z digest=sha256:340242de478deaf75f1ba658baa3ddd4b09a746397700a1760b763a9540dfe47

Observation d6e41052-fc33-4904-935b-a26fad884d49 · inbound

Uncertainty Quantification and Causal Considerations for Off-Policy Decision Making cites this paper.

Uncertainty Quantification and Causal Considerations for Off-Policy Decision Making Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-08T17:13:44.932906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:13:44.932906Z digest=sha256:1bc57fb6a6db72302bcdbbd39bec039f65dd1ce84887f79bbbb43bed7dedc7a3

Observation 66ee9906-2376-4556-b60e-fd813689d1cb · inbound

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards cites this paper.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T23:02:49.084096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:02:49.084096Z digest=sha256:889b8503e56dbf7e83da7a9327a895e4d0808c7c1360749578febc210a3b1a51

Observation 95017b36-248e-408d-b4cf-c8a980077298 · inbound

Quick-Draw Bandits: Quickly Optimizing in Nonstationary Environments with Extremely Many Arms cites this paper.

Quick-Draw Bandits: Quickly Optimizing in Nonstationary Environments with Extremely Many Arms Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:24:43.002375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:24:43.002375Z digest=sha256:a976f5aac0fcb3ff471ad30899b5e66d66942ec1b6539307dc1e97e54a84e6b1

Observation 3733ae4d-b351-4922-b317-83bcdf8a200b · inbound

Off-Policy Evaluation of Ranking Policies via Embedding-Space User Behavior Modeling cites this paper.

Off-Policy Evaluation of Ranking Policies via Embedding-Space User Behavior Modeling Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:04.807882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:04.807882Z digest=sha256:abcdfeef4f5101c88050cd2152086e0275a26be0aae847162807fcfc91592e12

Observation e95b97d0-1929-46a9-9c6e-686c6ceab220 · inbound

Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation cites this paper.

Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:31:59.119580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T02:28:26.498038Z digest=sha256:f2b121511c9d4179157630d54fd22e9c13d23c63b5bbb38df49a295d216357fe

Observation 48b08b01-d849-4060-aa5e-92c330ce0668 · inbound

GrowthHacker: Automated Off-Policy Evaluation Optimization Using Code-Modifying LLM Agents cites this paper.

GrowthHacker: Automated Off-Policy Evaluation Optimization Using Code-Modifying LLM Agents Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T00:30:41.466085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:30:41.466085Z digest=sha256:4ca9859f1a89e4c43ce2277da13cfdb80f489bc5bf313dbd9283f66a0baba6c2

Observation 544832bf-46ba-4781-ad8a-da5a445b9fb9 · inbound

Asymptotically Log-Optimal Bayes-Assisted Confidence Sequences for Bounded Means cites this paper.

Asymptotically Log-Optimal Bayes-Assisted Confidence Sequences for Bounded Means Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:50:55.754079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-11T02:49:10.973353Z digest=sha256:e0a808f8daa487a70bdad11237c12c39f80d7eaf9ca97706c5733ed08218209c

Observation bfec7dcb-3d04-451d-b163-00048f9b1039 · inbound

Asymptotically Log-Optimal Bayes-Assisted Confidence Sequences for Bounded Means cites this paper.

Asymptotically Log-Optimal Bayes-Assisted Confidence Sequences for Bounded Means Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:21:26.678833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-12T03:26:33.296297Z digest=sha256:fa75193de611b32420b8aea74694ac9c42bf4f9a2b523c99dc175aa45cf20c8b

Observation ffca71a5-7013-4648-b8ad-04ad46741217 · inbound

Decision-Calibrated Conformal Uncertainty for Pacing Decisions in Streaming Advertising cites this paper.

Decision-Calibrated Conformal Uncertainty for Pacing Decisions in Streaming Advertising Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:57:38.241561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T14:26:00.746441Z digest=sha256:c2314df49bfc3dbd9260ba0cbd8c5964b097e2e2ed4fefab0266396f440c8346

Observation 5aacbe62-781b-4a95-a8f0-6c3296601067 · inbound

Fed-CausalDiff: Decoupled Synchronization for Federated Do-Simulation and Policy Evaluation cites this paper.

Fed-CausalDiff: Decoupled Synchronization for Federated Do-Simulation and Policy Evaluation Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:59:43.095641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T10:40:07.441340Z digest=sha256:fa4f12f0bd699628efa7754e974466bc42e2182d8886420eca396f552156469d

Observation e958fe33-b9e4-4576-bc8c-6328282ae43d · inbound

Estimating Causal Effects from Data Generated by Stochastic Algorithms cites this paper.

Estimating Causal Effects from Data Generated by Stochastic Algorithms Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T00:15:47.450402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-09T00:15:30.364680Z digest=sha256:caaffd17d7c5e1d461cdef1db775d061d2b9ea604e3d29ba5be884d53e275ebe

Observation d8c4b61f-e6c3-4bef-84c5-efd845d52295 · inbound

Actions Have Consequences: Detecting Outcome Performativity using Intervention Testing cites this paper.

Actions Have Consequences: Detecting Outcome Performativity using Intervention Testing Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-30T17:25:36.028836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T17:25:36.028836Z digest=sha256:c21ece3f9b7ac7cfb01e50bfbed88f5ddc809ad7f8a7695d5392c6e3637d228b

Observation 4d894979-df97-42d2-a391-f4f803c02ada · inbound

From Prediction to Incrementality: Causal Optimization for Large-Scale Targeting and Recommendation cites this paper.

From Prediction to Incrementality: Causal Optimization for Large-Scale Targeting and Recommendation Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:17:02.933176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:17:02.933176Z digest=sha256:75c282752503250f67180ada9256192e6621615056e0634babdd4096f86115fc

Observation 6e8f8f9c-daf3-499f-bd49-0a4c6de68077 · inbound

When Offline Evaluation Misleads: A Diagnostic Protocol for Reward and Policy Selection in Delayed-Feedback Contextual Bandits cites this paper.

When Offline Evaluation Misleads: A Diagnostic Protocol for Reward and Policy Selection in Delayed-Feedback Contextual Bandits Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T00:41:32.001981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:41:32.001981Z digest=sha256:635b4aeafb81bed8d4abc7cff2ae37b64f2f5293cae2e9b1d306fbf3f2ea20b3