Pith. sign in

Paper Citation Record · LEDGER

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift

As of 13 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2607.23432.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23432 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T22:19:16.414341Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1f27d64e-d5bb-4789-b875-b4f904d23209 · outbound

This paper cites Bridging the gap between regret minimization and best arm identification, with application to a/b tests.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Bridging the gap between regret minimization and best arm identification, with application to a/b tests

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.385728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.385728Z digest=sha256:2631d9b493ee6d3a4eb6d59aad8f28dc2829c5ec1bea277cec442934fd3ca883

Observation 2999f7de-0df2-41a0-b9e1-664076d746d4 · outbound

This paper cites Therefore, we find that whenN < K1/3(T−N) 2/3, the instance-independent lower bound is Ω q K·(T−N) 2 N.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Therefore, we find that whenN < K1/3(T−N) 2/3, the instance-independent lower bound is Ω q K·(T−N) 2 N

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.410080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.410080Z digest=sha256:8897d7a0eae965ee9b2892112a7c6f936a7b82536f6a4e1a04d7265a246f0b04

Observation 7a4fe39a-f5b2-412d-9c03-631be5ad27c4 · outbound

This paper cites an unresolved cited work.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.414341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.414341Z digest=sha256:58fb75cec8e7125e5682a3b9063b0f63607396e0b1bbfb44bcbf81449c7a395d

Observation 2573b3dc-1503-4e0d-8f2f-8b3718c81650 · outbound

This paper cites doi: 10.1023/A: 1013689704352.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift doi: 10.1023/A: 1013689704352

Reference 2002

Resolution
malformed identifier
no resolver link, observed 2026-07-30T22:19:16.374738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.374738Z digest=sha256:b646e9ef132922ac3c8792c4bee8e85e7a0dd8b2c62d160d582f55467ca3a8e9

Observation 0ae4a0ee-641e-4b1d-b6b5-18d94d290788 · outbound

This paper cites Synthetically Controlled Bandits.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Synthetically Controlled Bandits

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.389455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.389455Z digest=sha256:9d5c4b1512f5eda1fbe481d4b8ddfce1229706ebf234ef8b81f3f53608af037a

Observation 40d14575-9f01-43c5-aebf-57827a67cc8c · outbound

This paper cites Best arm identification with minimal regret.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Best arm identification with minimal regret

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.406817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.406817Z digest=sha256:21094c4e6878a37d35c3cf4206f92bf98d6bc938322420f78d2f1c7711e8dc66

Observation adf93f16-5f70-4a7e-a683-d333e664ed95 · outbound

This paper cites Optimizing Adaptive Experiments: A Unified Approach to Regret Minimization and Best-Arm Identification.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Optimizing Adaptive Experiments: A Unified Approach to Regret Minimization and Best-Arm Identification

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.396798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.396798Z digest=sha256:5e905fc331b60902e49d89f88e6318c90899f6995b8bc5d9c837d0aee4d61acb

Observation 9f56c8a0-aa6e-4094-a7b3-e2c783277adc · outbound

This paper cites Learning the Pareto Front Using Bootstrapped Observation Samples.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Learning the Pareto Front Using Bootstrapped Observation Samples

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.393248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.393248Z digest=sha256:57ea990e3365ffbe18dd20728c73cd44eeaafc5e2f8752ec6e9e84a8a4eed1b9

Observation 9df464bf-26d5-43ef-a65e-f74c49757060 · outbound

This paper cites Beatriz Pessoa de Araujo and Adam Robbins.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Beatriz Pessoa de Araujo and Adam Robbins

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.382071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.382071Z digest=sha256:8839067226ebac7e51a59dc3a8efb71277b9cee17e330344d59c039094530a12

Observation 79d53273-a44a-443d-815a-471d1456506c · outbound

This paper cites Best arm identification in multi-armed bandits.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Best arm identification in multi-armed bandits

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.370418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.370418Z digest=sha256:1c01b5841423a889c70796075899872eecfc4083752832acd4f05e403d5a924a

Observation 6ddb04e7-c24d-4527-859c-472354854044 · outbound

This paper cites Regret distribution in stochastic bandits: Optimal trade-off between expectation and tail risk.arXiv preprint arXiv:2304.04341,.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Regret distribution in stochastic bandits: Optimal trade-off between expectation and tail risk.arXiv preprint arXiv:2304.04341,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.400186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.400186Z digest=sha256:1852f56dfdddf64a7ebefe808a45345e35e5148f2920598522bbfef7da21eb18

Observation 209f9bef-8dcb-442f-b81e-6adc84ab2cd0 · outbound

This paper cites Accessed 2026- 01-31.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Accessed 2026- 01-31

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.403687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.403687Z digest=sha256:18fd8bf00541c3b3ab0b2860879ab2057f7e8beef166c331483e7332e4cb82e2

Observation 44eaf60e-d5f9-4d3e-b8c7-9d3f8de089fb · outbound

This paper cites Sébastien Bubeck, Rémi Munos, and Gilles Stoltz.

Short-Term Pain for Long-Term Gain: Adaptive Experiment with Post-Commitment Reward Shift Sébastien Bubeck, Rémi Munos, and Gilles Stoltz

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-07-30T22:19:16.378401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T22:19:16.378401Z digest=sha256:0d71fc9e5d028116af32b0aa347051a81d096fc425555279cc26abc640d0bec9

Pith citing papers

No inbound Pith citation observations are available.