Pith. sign in

Paper Citation Record · LEDGER

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards

As of 22 August 2026, this Paper Citation Record lists 9 of 9 outbound references and 0 inbound Pith citation observations for arXiv:2502.08993.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08993 v1

Coverage vector

measured 9 of 9 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:02:49.119324Z

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

9 of 9 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66ee9906-2376-4556-b60e-fd813689d1cb · outbound

This paper cites Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Open Bandit Dataset and Pipeline: Towards Realistic and Reproducible Off-Policy Evaluation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T23:02:49.084096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:02:49.084096Z digest=sha256:889b8503e56dbf7e83da7a9327a895e4d0808c7c1360749578febc210a3b1a51

Observation 6dc9f898-80e6-499f-83f4-c9f50e714ff4 · outbound

This paper cites Top-k off-policy correction for a REINFORCE recommender system.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Top-k off-policy correction for a REINFORCE recommender system

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.252327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.089212Z digest=sha256:30205e8a5dd7a5fe25ce915c0e276c7015ba782da734bbc745f40fe31e3631ac

Observation fdf27983-60c4-4e32-9a82-65d56b80bce6 · outbound

This paper cites Offline evaluation to make de- cisions about playlistrecommendation algorithms.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Offline evaluation to make de- cisions about playlistrecommendation algorithms

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.241097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.093479Z digest=sha256:f13ad0b74a3900a0af9e55fb223c82752b1b53ef7dc3e9c34a78d9755cfe1fd8

Observation 7a60b6a9-2d18-4fe4-b881-35d2d4b9f42f · outbound

This paper cites Bias and debias in recommender system: A survey and future directions.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Bias and debias in recommender system: A survey and future directions

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.229660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.097893Z digest=sha256:5b9188bffbf4c256cb22280b927b6f0f19f1c9bef3eaaaf1bb1ece6680e4a457

Observation 4d4905c1-f66d-40c6-94a5-cc738bcaf807 · outbound

This paper cites Recommendations as treat- ments: Debiasing learning and evaluation.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Recommendations as treat- ments: Debiasing learning and evaluation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.217102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.102283Z digest=sha256:593fdd03966a7a4f0f8f57272a950951a624739ed3c045ecce252292b42f0a24

Observation 19500047-ebae-427a-a5d6-a84700bf25c9 · outbound

This paper cites Unbiased recommender learning from missing-not-at-random implicit feedback.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Unbiased recommender learning from missing-not-at-random implicit feedback

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.204431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.106845Z digest=sha256:ef7c5ae8e3f362c3d1f36ddb2a3c30f64a47c19940a1ccdc2cca7bcc687e3a47

Observation c537b2fb-e1d4-4cee-b80b-2c4698de5a3a · outbound

This paper cites Coun- terfactual risk minimization: Learning from logged bandit feedback.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Coun- terfactual risk minimization: Learning from logged bandit feedback

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.192132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.111487Z digest=sha256:2a84f000f6623e3e228cd18c03fdb5df17300c108580a4b26e5de5eac8b7287d

Observation 74c97132-27a3-4fb0-b884-3a69fd579ac0 · outbound

This paper cites Off-Policy Eval- uation for Large Action Spaces via Embeddings.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Off-Policy Eval- uation for Large Action Spaces via Embeddings

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.179905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.115308Z digest=sha256:4d510a6158b106322ce693faf528ef603e6393b6352f8f5cb233c6c3fe8f562c

Observation 240b9a10-bf2b-41ce-80c3-d7927b26fe21 · outbound

This paper cites Offline evaluation of ranking poli- cies with click models.

Off-Policy Evaluation for Recommendations with Missing-Not-At-Random Rewards Offline evaluation of ranking poli- cies with click models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:02:49.166329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T23:02:49.119324Z digest=sha256:bbac0d3c22d34087413679d86d227d4bc6458b7013e48b9c98e28a55f8b6de42

Pith citing papers

No inbound Pith citation observations are available.