Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-30T22:19:16.414341Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2607.23432.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-30T22:19:16.414341Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1f27d64e-d5bb-4789-b875-b4f904d23209 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Bridging the gap between regret minimization and best arm identification, with application to a/b tests
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2999f7de-0df2-41a0-b9e1-664076d746d4 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Therefore, we find that whenN < K1/3(T−N) 2/3, the instance-independent lower bound is Ω q K·(T−N) 2 N
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a4fe39a-f5b2-412d-9c03-631be5ad27c4 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2573b3dc-1503-4e0d-8f2f-8b3718c81650 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift doi: 10.1023/A: 1013689704352
Reference 2002
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ae4a0ee-641e-4b1d-b6b5-18d94d290788 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Synthetically Controlled Bandits
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40d14575-9f01-43c5-aebf-57827a67cc8c · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Best arm identification with minimal regret
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adf93f16-5f70-4a7e-a683-d333e664ed95 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Optimizing Adaptive Experiments: A Unified Approach to Regret Minimization and Best-Arm Identification
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f56c8a0-aa6e-4094-a7b3-e2c783277adc · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Learning the Pareto Front Using Bootstrapped Observation Samples
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df464bf-26d5-43ef-a65e-f74c49757060 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Beatriz Pessoa de Araujo and Adam Robbins
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d53273-a44a-443d-815a-471d1456506c · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Best arm identification in multi-armed bandits
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ddb04e7-c24d-4527-859c-472354854044 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Regret distribution in stochastic bandits: Optimal trade-off between expectation and tail risk.arXiv preprint arXiv:2304.04341,
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 209f9bef-8dcb-442f-b81e-6adc84ab2cd0 · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Accessed 2026- 01-31
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44eaf60e-d5f9-4d3e-b8c7-9d3f8de089fb · outbound
Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Sébastien Bubeck, Rémi Munos, and Gilles Stoltz
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.