Pith. sign in

Paper Citation Record · LEDGER

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift

As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2607.23432.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23432 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T22:19:16.414341Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1f27d64e-d5bb-4789-b875-b4f904d23209 · outbound

This paper cites Bridging the gap between regret minimization and best arm identification, with application to a/b tests.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Bridging the gap between regret minimization and best arm identification, with application to a/b tests

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.385728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.385728Z digest=sha256:14d15284ffb831a67f61bc5ba8c868d13f6541e43394ad7222dbc53e7d8c8ef3

Observation 2999f7de-0df2-41a0-b9e1-664076d746d4 · outbound

This paper cites Therefore, we find that whenN < K1/3(T−N) 2/3, the instance-independent lower bound is Ω q K·(T−N) 2 N.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Therefore, we find that whenN < K1/3(T−N) 2/3, the instance-independent lower bound is Ω q K·(T−N) 2 N

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.410080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.410080Z digest=sha256:55a240fca4d54be76cb52c3ae230c9d28ba85c7f5c68216cd8ac64fc6e23e055

Observation 7a4fe39a-f5b2-412d-9c03-631be5ad27c4 · outbound

This paper cites an unresolved cited work.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.414341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.414341Z digest=sha256:4f400cc5e76cb0f13a61d0b1620585a142bf26ddffc9f8196d4a226027bd1a58

Observation 2573b3dc-1503-4e0d-8f2f-8b3718c81650 · outbound

This paper cites doi: 10.1023/A: 1013689704352.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift doi: 10.1023/A: 1013689704352

Reference 2002

Resolution
malformed identifier
no resolver link, observed 2026-07-30T22:19:16.374738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.374738Z digest=sha256:8ba45aa61aa8550dca497f6654a516a690048852f90ae027b58b3d5e6c760ce2

Observation 0ae4a0ee-641e-4b1d-b6b5-18d94d290788 · outbound

This paper cites Synthetically Controlled Bandits.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Synthetically Controlled Bandits

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.389455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.389455Z digest=sha256:dbedff013de40d0e7794fa9523908cefb2a8baa7ab18f5cddde515fa2d28e864

Observation 40d14575-9f01-43c5-aebf-57827a67cc8c · outbound

This paper cites Best arm identification with minimal regret.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Best arm identification with minimal regret

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.406817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.406817Z digest=sha256:fad6d74dc17a069c1d5feae20257ab8ee2a41e5c774aa4b01ac5232d0d2c7577

Observation adf93f16-5f70-4a7e-a683-d333e664ed95 · outbound

This paper cites Optimizing Adaptive Experiments: A Unified Approach to Regret Minimization and Best-Arm Identification.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Optimizing Adaptive Experiments: A Unified Approach to Regret Minimization and Best-Arm Identification

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.396798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.396798Z digest=sha256:ce6298d0163f09e6b01f417399210dd3cd5f4b375f00b1d2346c9637211aa109

Observation 9f56c8a0-aa6e-4094-a7b3-e2c783277adc · outbound

This paper cites Learning the Pareto Front Using Bootstrapped Observation Samples.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Learning the Pareto Front Using Bootstrapped Observation Samples

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.393248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.393248Z digest=sha256:485b88e13d1ae242bec225e1c30adc9445a0a10e6831880ed5a556402a3e657e

Observation 9df464bf-26d5-43ef-a65e-f74c49757060 · outbound

This paper cites Beatriz Pessoa de Araujo and Adam Robbins.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Beatriz Pessoa de Araujo and Adam Robbins

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.382071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.382071Z digest=sha256:1df12282e9dcc984ae16d8648dab89fe853ee8e7965ec9d3fb4ade5136ce874f

Observation 79d53273-a44a-443d-815a-471d1456506c · outbound

This paper cites Best arm identification in multi-armed bandits.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Best arm identification in multi-armed bandits

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.370418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.370418Z digest=sha256:62ba82882f4b5ba4a170c290b1d80059694425de6c05f0c0da536c62a3d92e9d

Observation 6ddb04e7-c24d-4527-859c-472354854044 · outbound

This paper cites Regret distribution in stochastic bandits: Optimal trade-off between expectation and tail risk.arXiv preprint arXiv:2304.04341,.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Regret distribution in stochastic bandits: Optimal trade-off between expectation and tail risk.arXiv preprint arXiv:2304.04341,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.400186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.400186Z digest=sha256:038812da6f08ccc057f00ada6f2b77072f129b0185d8cdaa8bf7bb2e135e553b

Observation 209f9bef-8dcb-442f-b81e-6adc84ab2cd0 · outbound

This paper cites Accessed 2026- 01-31.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Accessed 2026- 01-31

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.403687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.403687Z digest=sha256:437f93f8b0d2bd68040d8b849d7a612c0bd52e37869e987d8b34b63dd39d2328

Observation 44eaf60e-d5f9-4d3e-b8c7-9d3f8de089fb · outbound

This paper cites Sébastien Bubeck, Rémi Munos, and Gilles Stoltz.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Sébastien Bubeck, Rémi Munos, and Gilles Stoltz

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.378401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.378401Z digest=sha256:c2764524829a8992cc7609dd02b6126b8a8599533bd0c225e05227658d26ed1b

Pith citing papers

No inbound Pith citation observations are available.