Pith. sign in

Paper Citation Record · LEDGER

Off-policy estimation with adaptively collected data: the power of online learning

As of 13 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 0 inbound Pith citation observations for arXiv:2411.12786.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.12786 v1

Coverage vector

measured 69 of 69 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:43:58.594829Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

69 of 69 outbound references displayed

  • verified exact1
  • verified fuzzy60
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bdd83350-4a3f-4855-a636-79ceb0361447 · outbound

This paper cites Effective evaluation using logged bandit feedback from multiple loggers.

Off-policy estimation with adaptively collected data: the power of online learning Effective evaluation using logged bandit feedback from multiple loggers

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.256076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.384377Z digest=sha256:e67873caf38335c6afa9227167e49caaafd8715008ebdd208fa73e09f5561fdb

Observation 13ea1470-4f55-4404-8e50-78c360375fc7 · outbound

This paper cites Thompson sampling for contextu al bandits with linear payoffs.

Off-policy estimation with adaptively collected data: the power of online learning Thompson sampling for contextu al bandits with linear payoffs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.388449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.388449Z digest=sha256:b512d580ba95822fd7ca4cc6dd71e9e872cc7aad792c61d733a03259c2c789e9

Observation ce097e39-82ee-4281-b98d-a34f8d4f4eef · outbound

This paper cites Finite-sample optimal e stimation and inference on average treatment effects under unconfoundedness.

Off-policy estimation with adaptively collected data: the power of online learning Finite-sample optimal e stimation and inference on average treatment effects under unconfoundedness

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.240714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.391895Z digest=sha256:480ca30fb35950109d7be13549348853d9992ccb719865c4c542c91743be77ff

Observation 8bf30fc0-a75e-412b-bc2b-87bb306bf313 · outbound

This paper cites Counter factual reasoning and learning systems: The example of computational advertising.

Off-policy estimation with adaptively collected data: the power of online learning Counter factual reasoning and learning systems: The example of computational advertising

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.230750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.395022Z digest=sha256:9c199cbee148ca678710c66931a8e08195b9c0ef3b60d67bdbc77e6c03e24349

Observation fd40c365-92ad-4ac5-bbc8-19596f448401 · outbound

This paper cites Double/debiased/neyman machine learning of treatment eff ects.

Off-policy estimation with adaptively collected data: the power of online learning Double/debiased/neyman machine learning of treatment eff ects

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.221129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.397978Z digest=sha256:6a11369cf3f143d6f11112b880a2ffe6ded4ed0be34c7c89ffe7dad3ab2f6a5b

Observation e32cfe94-d42a-41ad-9186-b6f943c93bf2 · outbound

This paper cites Double/debiased machine learning for tre atment and structural parameters, 2018.

Off-policy estimation with adaptively collected data: the power of online learning Double/debiased machine learning for tre atment and structural parameters, 2018

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.212056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.401013Z digest=sha256:d1d3e36814c6dab39d92dcc87142b82c15ce0965560cddc68562ba076bee71e1

Observation 2e61cfd6-12a8-4a77-a4c6-12f39503cd48 · outbound

This paper cites Semiparametric e fficient inference in adaptive experiments.

Off-policy estimation with adaptively collected data: the power of online learning Semiparametric e fficient inference in adaptive experiments

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.202829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.404140Z digest=sha256:7dba5eaf452a96fb26090e07892d5078d1a0434eddb870589610381b6483750a

Observation 491d4f36-c436-434a-a1b3-0c90b1b30727 · outbound

This paper cites Clip-ogd: An experimental design for adaptive neyman allocation in sequential experiments.

Off-policy estimation with adaptively collected data: the power of online learning Clip-ogd: An experimental design for adaptive neyman allocation in sequential experiments

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.193720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.406886Z digest=sha256:72a4fa57fc167b126344e24325ab3d0254b15859b7d530a1898a5e2f64cb10ca

Observation dfdd3bd8-5e89-4030-b644-7e841cfa0b74 · outbound

This paper cites Do ubly Robust Policy Evaluation and Optimization.

Off-policy estimation with adaptively collected data: the power of online learning Do ubly Robust Policy Evaluation and Optimization

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.184495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.409474Z digest=sha256:c3a49405db2a54accb08d05c217e171c784ca1c19c171ffce64758e5b8c9ba39

Observation dff7eb14-9d79-402b-ad90-bf38a1f04644 · outbound

This paper cites Doubly Robust Policy Evaluation and Learning.

Off-policy estimation with adaptively collected data: the power of online learning Doubly Robust Policy Evaluation and Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.412086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.412086Z digest=sha256:ff5660e72d45c1e7739966cb68995c1b82a21b07058f7d6098a06c62c775c680

Observation 7d9430af-0516-438e-bd0d-a37e3f968b60 · outbound

This paper cites Overlap in observational studies with high-dimensional covariates.

Off-policy estimation with adaptively collected data: the power of online learning Overlap in observational studies with high-dimensional covariates

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.175673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.415711Z digest=sha256:155b8a44f04a172a0a29b9b7406f7a25309ed39182b9df2050b68ddaf98ff75e

Observation 2b0c36a5-6efe-4731-b99e-726bed0c03d4 · outbound

This paper cites More robust doubly robust off- policy evaluation.

Off-policy estimation with adaptively collected data: the power of online learning More robust doubly robust off- policy evaluation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.166818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.419407Z digest=sha256:74dcb933597fa4a6a185b75a83320674567e1bacf8d00ac248fae51b3dfce5c8

Observation a62be84f-d5b7-47bf-ae67-fe25208a1530 · outbound

This paper cites Off-policy evalua- tion with deficient support using side information.

Off-policy estimation with adaptively collected data: the power of online learning Off-policy evalua- tion with deficient support using side information

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.157620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.422621Z digest=sha256:6371ac8677748fa66d4936ef3191a8865d4f9c89ddca4e875b2448af78206905

Observation 42f631f7-b5c5-4b08-ba6e-5e56dd33b700 · outbound

This paper cites On choosing and bounding pr obability metrics.

Off-policy estimation with adaptively collected data: the power of online learning On choosing and bounding pr obability metrics

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.148297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.425943Z digest=sha256:c986b6ff4769cd315ff7e4722b24e3d8830cf066ca76418ebf7a02ca4335d935

Observation 4efef9cc-4459-402c-82b4-b956e067a7e3 · outbound

This paper cites Some limit theorems for empirical pro cesses.

Off-policy estimation with adaptively collected data: the power of online learning Some limit theorems for empirical pro cesses

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.138932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.429094Z digest=sha256:742c88532b7a7fe160f8a935883c59869f65dc27f308bdeee9f945e6a9efa31a

Observation 41617ce4-8cec-4cd5-a09a-1744a5b61858 · outbound

This paper cites Confidence Intervals for Policy Evaluation in Adaptive Experiments.

Off-policy estimation with adaptively collected data: the power of online learning Confidence Intervals for Policy Evaluation in Adaptive Experiments

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.432460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.432460Z digest=sha256:a324119f53268e3765c3fbd7afd5278221fa242b8269d5b6741eb3c6f7489cd8

Observation 0766f0a5-fa30-4f30-a69c-f9b62ed2d9d7 · outbound

This paper cites Confidence inter- vals for policy evaluation in adaptive experiments.

Off-policy estimation with adaptively collected data: the power of online learning Confidence inter- vals for policy evaluation in adaptive experiments

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.129076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.436031Z digest=sha256:318e3015a57f4e04fa779e9f81610df1df17ec3aaf87e430e77f153bd045bd13

Observation 692f9d41-76df-4b6a-a04a-de2cafda0a08 · outbound

This paper cites Introduction to online convex optimization.

Off-policy estimation with adaptively collected data: the power of online learning Introduction to online convex optimization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.119387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.439220Z digest=sha256:ebc28438464e06739ef444f53a7ed19bd9c61d550de9fe7ac2b10be3ba4c2d84

Observation 51e107c3-03f6-491f-bbdb-60360cedec4b · outbound

This paper cites Weighted average importance sampling and def ensive mixture distributions.

Off-policy estimation with adaptively collected data: the power of online learning Weighted average importance sampling and def ensive mixture distributions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.109820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.442438Z digest=sha256:a835f9b37de24e0803aa352fdb376e213397b8b314e28c8e36dbdab43a9eccb5

Observation fddd75ce-9f02-4eaa-b5f8-1d1e2d9534fe · outbound

This paper cites Efficient estim ation of average treatment effects using the estimated propensity score.

Off-policy estimation with adaptively collected data: the power of online learning Efficient estim ation of average treatment effects using the estimated propensity score

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.100505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.445622Z digest=sha256:1e49e189091bdf38a11bfd568f018c236f14779819c4d66854be3d4e6d8d449b

Observation ff2cfef6-8ffe-4454-9e4f-cef40093e66d · outbound

This paper cites A generalization of sa mpling without replacement from a finite universe.

Off-policy estimation with adaptively collected data: the power of online learning A generalization of sa mpling without replacement from a finite universe

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.091187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.448907Z digest=sha256:7b8d86034f940bcca258ac74058ce1899412bb2435edce1815e3bb43d67177ce

Observation 765d4f6a-cc66-4f1b-98c3-af5a1bfbf90c · outbound

This paper cites Howard, Aaditya Ramdas, Jon McAuliffe, and Jasjeet Sekhon.

Off-policy estimation with adaptively collected data: the power of online learning Howard, Aaditya Ramdas, Jon McAuliffe, and Jasjeet Sekhon

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.080484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.452186Z digest=sha256:eb1ab1c3a784ccde0111725b70d99968dea86e0aacd7d973e8fd6c4080d2c1b5

Observation 42d8ad55-67c2-44d3-b7e5-37b9018e7402 · outbound

This paper cites Nonparametric estimation of average treatme nt effects under exogeneity: A review.

Off-policy estimation with adaptively collected data: the power of online learning Nonparametric estimation of average treatme nt effects under exogeneity: A review

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.071509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.455371Z digest=sha256:2ab86bda0d2a319d0c8cb3a4dc9f8013db6074fef3883f35a7106bda16105ddc

Observation ab7db0ae-314d-4100-bf24-ca9e988eb639 · outbound

This paper cites Causal inference in statistics, social, and biomedical sci ences.

Off-policy estimation with adaptively collected data: the power of online learning Causal inference in statistics, social, and biomedical sci ences

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.062808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.458429Z digest=sha256:2a020fd409b88e56712d773c9c55d9c6c8a4d2c597bba25fa7708310b533b395

Observation 54c95186-535e-4c35-b68f-a97bb9c66e74 · outbound

This paper cites Truncated importance sampling.

Off-policy estimation with adaptively collected data: the power of online learning Truncated importance sampling

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.053695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.461628Z digest=sha256:cd4c2aeb95444c36724dd531b38a919d32caea1cce6c44f6f2f58d587d020093

Observation 33ecec2d-451b-42fb-8475-e8cb28bee0c2 · outbound

This paper cites Policy learning "without" overlap: Pessimism and generalized empirical Bernstein's inequality.

Off-policy estimation with adaptively collected data: the power of online learning Policy learning "without" overlap: Pessimism and generalized empirical Bernstein's inequality

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.464904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.464904Z digest=sha256:c4d7a9b4858564c9683bc46dcfd98ef81ba1f2b194ee36cd3ed2e0a74944e4fa

Observation 73586c0e-5b7c-464f-9f2b-1fd543b59673 · outbound

This paper cites Optimal off- policy evaluation from multiple logging policies.

Off-policy estimation with adaptively collected data: the power of online learning Optimal off- policy evaluation from multiple logging policies

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.044609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.468517Z digest=sha256:4b45f38b12d80c90c6000439b5a9795c6aac241b9937b7178bd76b83cce050fc

Observation 1415e787-1b6c-4ec2-8a14-6cd007dc1bf8 · outbound

This paper cites Policy evaluation and optimization w ith continuous treatments.

Off-policy estimation with adaptively collected data: the power of online learning Policy evaluation and optimization w ith continuous treatments

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.034761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.471851Z digest=sha256:711b88fd4651513db4348febe47a9d50ffb0c7c3dccaa21e10b90092af8d47da

Observation 7fe2337c-4657-423b-b733-7b91b5f3a839 · outbound

This paper cites an unresolved cited work.

Off-policy estimation with adaptively collected data: the power of online learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:43:59.024823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.474877Z digest=sha256:9b20b17bdc6d29e23de52582129eefa9cbb84b2388d10908524681ee55e71f33

Observation 2bd7d3f3-db7f-4c00-af86-588d5af2e177 · outbound

This paper cites Off-po licy confidence sequences.

Off-policy estimation with adaptively collected data: the power of online learning Off-po licy confidence sequences

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.015260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.477767Z digest=sha256:9d125586357705e9ed1a5cf620ea7f70e53f43e35f4cc427cdc5fd7d7afd8b46

Observation ef8513de-cbe9-4a65-8952-3db835d22697 · outbound

This paper cites Efficient Adaptive Experimental Design for Average Treatment Effect Estimation.

Off-policy estimation with adaptively collected data: the power of online learning Efficient Adaptive Experimental Design for Average Treatment Effect Estimation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.481508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.481508Z digest=sha256:fdfcf972cf25d37ad56e6a493eb1f41a15875e24891e3d70f83e8541739cc82e

Observation 70e4033f-2abc-4256-af5e-2f3bbe37c276 · outbound

This paper cites Irregular identification, support conditions, and inverse weight estima- tion.

Off-policy estimation with adaptively collected data: the power of online learning Irregular identification, support conditions, and inverse weight estima- tion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:59.005109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.484430Z digest=sha256:9df1410d7a1cb5e7de86c0f579d12081edb988117ff7838571414c966486b984

Observation c9c49b9a-9ba4-470a-b4ba-53cfe7aa30da · outbound

This paper cites Asymptotically efficient ada ptive allocation rules.

Off-policy estimation with adaptively collected data: the power of online learning Asymptotically efficient ada ptive allocation rules

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.995176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.487162Z digest=sha256:3e971bfb38a48930762e29ac7405eb7f7c60f7fb871d539fc9586bad66f31125

Observation 8340bea5-6a62-4d7d-9302-6f9027d95a2e · outbound

This paper cites Bandit algorithms.

Off-policy estimation with adaptively collected data: the power of online learning Bandit algorithms

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.489678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.489678Z digest=sha256:8235829ddf5727422e45282b6496ed391de7f42a4255341dc15c91a0f66dcd73

Observation 71145d63-232b-40ae-af3c-f59f7086d9c7 · outbound

This paper cites Local metric learning for off-policy evaluation in contextua l bandits with continuous actions.

Off-policy estimation with adaptively collected data: the power of online learning Local metric learning for off-policy evaluation in contextua l bandits with continuous actions

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.979962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.492191Z digest=sha256:a575305fe1aa25048841414a6f7df1a0fee92fe6206aec0b9f39c9315b9be448

Observation 59b946f9-4f6c-4513-956b-4e5b40cb731d · outbound

This paper cites Distribution-free assessment of population overlap in observational studies.

Off-policy estimation with adaptively collected data: the power of online learning Distribution-free assessment of population overlap in observational studies

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.970619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.494876Z digest=sha256:41c75e3186424b070937f286df64f9ca82befd7d432649a2aac7e07873d156b0

Observation 74c1930e-a5a0-49b4-a907-d89f17998b77 · outbound

This paper cites Sharp high-probability sample complexities for policy evaluation with linear function approximation.

Off-policy estimation with adaptively collected data: the power of online learning Sharp high-probability sample complexities for policy evaluation with linear function approximation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.961990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.497374Z digest=sha256:72a79f15438b0451e2b3399c11cbc4cc40657ea596eb8216afd1400b9c906dde

Observation 514c7f5d-4fe5-41ad-b7f4-1a26385c3c6b · outbound

This paper cites Toward minimaxoff-policy value estimation.

Off-policy estimation with adaptively collected data: the power of online learning Toward minimaxoff-policy value estimation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.952463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.500400Z digest=sha256:4d064bea5165385a529346193abe70f67bfbd66075f0c8942428d323dcb55ff2

Observation 702998d8-e69f-46f9-a818-9ed76ae8d96e · outbound

This paper cites Statistical analysis with missing data , volume 793.

Off-policy estimation with adaptively collected data: the power of online learning Statistical analysis with missing data , volume 793

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.942657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.503586Z digest=sha256:2a4a98f3c36d18bc4ef8fce01dc753711bb4e1c801d8df35fc002c771e2db73b

Observation 4f599360-83c2-46fe-a12f-0d273eef2325 · outbound

This paper cites Statistical infer ence for the mean outcome under a possibly non-unique optimal treatment strategy.

Off-policy estimation with adaptively collected data: the power of online learning Statistical infer ence for the mean outcome under a possibly non-unique optimal treatment strategy

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.933079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.506578Z digest=sha256:85e0e364671e3276dbefd431da154bb9db5effae242fa38ba5b27f1bb0c75663

Observation 880a205a-4902-443e-ad93-6d7f137db380 · outbound

This paper cites Min imax off-policy evaluation for multi-armed bandits.

Off-policy estimation with adaptively collected data: the power of online learning Min imax off-policy evaluation for multi-armed bandits

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.923814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.509538Z digest=sha256:b6b5acdaeac21462d4cedc7f6c2eea02585cf34983439ef8d326764d9560dc21

Observation bceb8c21-f1ac-4b72-a0dd-d67635c8d435 · outbound

This paper cites Off-policy estimation of linear functionals: Non-asymptotic theory for semi-parametric efficiency.

Off-policy estimation with adaptively collected data: the power of online learning Off-policy estimation of linear functionals: Non-asymptotic theory for semi-parametric efficiency

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:43:58.640071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.512501Z digest=sha256:3ce97f86082b5568daa81b57d9d8941c52f29aa7369c05c4e93097eb683856b0

Observation 54415f53-83e2-4302-bbeb-06c16943141e · outbound

This paper cites Efficient counter factual learning from bandit feedback.

Off-policy estimation with adaptively collected data: the power of online learning Efficient counter factual learning from bandit feedback

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.914524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.516147Z digest=sha256:73f5a3def47c9eed402291904ae8ccd8b1790662703507c900eeb1e57d224fb0

Observation c620aecc-93d6-472a-8f1d-1ebf3afd61b4 · outbound

This paper cites Offline policy evaluation in large action spaces via outcome-oriented action group ing.

Off-policy estimation with adaptively collected data: the power of online learning Offline policy evaluation in large action spaces via outcome-oriented action group ing

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.905104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.519430Z digest=sha256:33a63396c113a0ff24a8893d1cfb62f0f33a0f63e84d6f7133b2a0bc307f4014

Observation 70e47d55-6214-43f6-80ad-4964406a5e51 · outbound

This paper cites Online non-parametric regression.

Off-policy estimation with adaptively collected data: the power of online learning Online non-parametric regression

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.896509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.522574Z digest=sha256:91bf01d7c35e2ce0355fbd7f1634971cad4d3c6bb88a1075b9842ca4c0230274

Observation 6a50acc1-fa07-4676-9d78-ace03cdeeb35 · outbound

This paper cites Seque ntial complexities and uniform mar- tingale laws of large numbers.

Off-policy estimation with adaptively collected data: the power of online learning Seque ntial complexities and uniform mar- tingale laws of large numbers

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.888627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.525570Z digest=sha256:cf3a96fc01c4cce46889db4ae11c09145f57d0f05133c886c8cdd3f2b25a1a36

Observation 0ad9acdd-50b3-4190-98a3-c0138b3c4caf · outbound

This paper cites Relax and ra ndomize: From value to algorithms.

Off-policy estimation with adaptively collected data: the power of online learning Relax and ra ndomize: From value to algorithms

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.879912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.528627Z digest=sha256:0f0d24e44954746825c4c825c2254df41a7411373679d1cbf134a49ea4354a9e

Observation 61554459-13d2-46f6-85d9-e83c8c8689db · outbound

This paper cites Comment: Performance of double-robust estimators when” inverse probability” weights ar e highly variable.

Off-policy estimation with adaptively collected data: the power of online learning Comment: Performance of double-robust estimators when” inverse probability” weights ar e highly variable

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.870744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.531540Z digest=sha256:2265d58d52979278652adf0e70b8a85189f8c60fbf2ca25eee9f3aada25d5b82

Observation 17d2bccb-68fd-429f-bd91-b582d4ae5e5d · outbound

This paper cites Semiparametric efficiency in multivariate regression models with missing data.

Off-policy estimation with adaptively collected data: the power of online learning Semiparametric efficiency in multivariate regression models with missing data

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.860460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.534555Z digest=sha256:1d7d4bc93d30e9bcaef72c898e3f5ffafa0922c33308dc604f50b1e81a3e89f4

Observation 2ef21956-0e7a-464d-b7e4-54e89463da98 · outbound

This paper cites Estimatio n of regression coefficients when some regressors are not always observed.

Off-policy estimation with adaptively collected data: the power of online learning Estimatio n of regression coefficients when some regressors are not always observed

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.851091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.537719Z digest=sha256:3bb1dc4134551b883f18148cf65c9e679fdec54cc9e1a15a1a8fdbcb09e96312

Observation d4c43c02-d4e0-4815-9509-efa4d20b4331 · outbound

This paper cites Analysis o f semiparametric regression models for repeated outcomes in the presence of missing data.

Off-policy estimation with adaptively collected data: the power of online learning Analysis o f semiparametric regression models for repeated outcomes in the presence of missing data

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.841575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.540801Z digest=sha256:22ae70178283ef5249e6df08b5f42f7e34606391f710078e9f8dfc0ada149e77

Observation df24683d-c045-4d2b-b50f-09943e68d932 · outbound

This paper cites A tutorial on thompson sampling.

Off-policy estimation with adaptively collected data: the power of online learning A tutorial on thompson sampling

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.831606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.543919Z digest=sha256:a9f0e13c6b88656e388374586e0f536dbefcec6ac594df94e0ed17a4b8bfd6e1

Observation 7dcc1f49-cef1-4b27-a663-9803d30b5ea5 · outbound

This paper cites Off-Policy Evaluation for Large Action Spaces via Embeddings.

Off-policy estimation with adaptively collected data: the power of online learning Off-Policy Evaluation for Large Action Spaces via Embeddings

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T17:43:58.546929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:43:58.546929Z digest=sha256:ab2da2650b18814da4b334ede42c2551de4cb3ebf57ab85a2a4017795e7f5bb8

Observation de50d675-da0e-4cb4-8ef2-25962971d72f · outbound

This paper cites Off-policy ev aluation for large action spaces via conjunct effect modeling.

Off-policy estimation with adaptively collected data: the power of online learning Off-policy ev aluation for large action spaces via conjunct effect modeling

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.822558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.550231Z digest=sha256:9a3caf50528e123e88f11d3f828acbc4c8c1d2066d936c1f1ddf06cb17c82178

Observation 2b63066d-06c7-41c8-b289-5469d9008d7f · outbound

This paper cites Lear ning from logged implicit exploration data.

Off-policy estimation with adaptively collected data: the power of online learning Lear ning from logged implicit exploration data

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.814022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.553443Z digest=sha256:455caa70e6e258c79a91c1b010424210a1cc0a351613b40bda6c1bb9bd2b3e63

Observation b6bdbb72-811a-4e54-9836-42f2ed0d9109 · outbound

This paper cites Doubly robust off-policy evaluation with shrinkage.

Off-policy estimation with adaptively collected data: the power of online learning Doubly robust off-policy evaluation with shrinkage

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.805656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.556638Z digest=sha256:420f608caa0d4e39a28bbe5bad2bc43ff0d75e98cd44a5e93fb8a3a43e8477a2

Observation 8b319e47-8d66-4cf5-ad49-7037c7300635 · outbound

This paper cites Cab: Continuous adaptive blend- ing for policy evaluation and learning.

Off-policy estimation with adaptively collected data: the power of online learning Cab: Continuous adaptive blend- ing for policy evaluation and learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.797671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.559742Z digest=sha256:e2c234282ff5d67e2a70957f3788331856c1c03304532958439d3ab1e52e0608

Observation ecde691c-e9cc-4d95-b1c5-d1544122943d · outbound

This paper cites The self-normalized estimator for counterfactual learning.

Off-policy estimation with adaptively collected data: the power of online learning The self-normalized estimator for counterfactual learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.789061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.562987Z digest=sha256:f3e5ef227eb7155dd775e21133760b821fa5a9907ccda8394e94d66ba7aa9986

Observation e38b07e5-19f9-49d3-a04b-42ea05405a31 · outbound

This paper cites Data-efficient off-policy policy eva luation for reinforcement learn- ing.

Off-policy estimation with adaptively collected data: the power of online learning Data-efficient off-policy policy eva luation for reinforcement learn- ing

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.780101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.566553Z digest=sha256:ff609220d67fadcee8e3af45cd68a1468b2fa8c18f5cd485e75b114460ebe543

Observation c73b29ef-8313-497e-9d84-de6f5950c774 · outbound

This paper cites On the likelihood that one unknown probability ex ceeds another in view of the evidence of two samples.

Off-policy estimation with adaptively collected data: the power of online learning On the likelihood that one unknown probability ex ceeds another in view of the evidence of two samples

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.771008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.569592Z digest=sha256:75f8dfd41dc29ef58e8523e9d4dd7b433bc3b11b869c28cad86908b0256813c3

Observation d4634878-2d7b-4c84-a5b6-a07f3bddd642 · outbound

This paper cites The construction and analysis of adaptive group sequential designs.

Off-policy estimation with adaptively collected data: the power of online learning The construction and analysis of adaptive group sequential designs

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.761483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.572740Z digest=sha256:e98a36c34fd42893352348340d137f0f734004d798b473481776eee2c51450e6

Observation 55ff0e85-6000-421d-8fef-6a9d3b9cd250 · outbound

This paper cites High-dimensional statistics: A non-asymptotic viewpoint , volume 48.

Off-policy estimation with adaptively collected data: the power of online learning High-dimensional statistics: A non-asymptotic viewpoint , volume 48

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.752267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.575839Z digest=sha256:53cb231e946e440a270fe10eb189f5ba282cbb82856f26307d16e3a0cb532d89

Observation 0525e20e-def8-4f05-90d4-6c53dfd5e888 · outbound

This paper cites Oracle-e fficient pessimism: Offline policy optimization in contextual bandits.

Off-policy estimation with adaptively collected data: the power of online learning Oracle-e fficient pessimism: Offline policy optimization in contextual bandits

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.743269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.578928Z digest=sha256:1d07e74490db37495809e731513f846f0571788cd080c22f444796969a0b0d4f

Observation 0ac61766-8332-43d3-b8f1-f7864532a8c4 · outbound

This paper cites Optimal and ad aptive off-policy evaluation in contextual bandits.

Off-policy estimation with adaptively collected data: the power of online learning Optimal and ad aptive off-policy evaluation in contextual bandits

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.733433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.581648Z digest=sha256:452afb25b520b0ae1535e3b21d4beaa8af5c38e648c5c0b8c1e3b7ad2dacc913

Observation 649e6e36-317d-4fea-b7b2-eea3bb886478 · outbound

This paper cites Anytime-valid off-policy inference for contextual bandits.

Off-policy estimation with adaptively collected data: the power of online learning Anytime-valid off-policy inference for contextual bandits

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.724253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.584182Z digest=sha256:78c25c337d42d7bfa199c5f46fd47a027a1b4f7d5fb8801a41ff8f7b6bdfd6d5

Observation 88d8437e-3d16-4f29-bd16-c7139087686d · outbound

This paper cites Asymptotic inference of causal effects with o bservational studies trimmed by the estimated propensity scores.

Off-policy estimation with adaptively collected data: the power of online learning Asymptotic inference of causal effects with o bservational studies trimmed by the estimated propensity scores

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.715417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.586736Z digest=sha256:7f27710e846c8eeaac4a9efc1a4911689829d12fdd531a737861903713e51544

Observation 101db0d1-b20d-4d34-b628-83ec133cee91 · outbound

This paper cites Off-policy evaluation via adaptive weighting with data from contextual bandits.

Off-policy estimation with adaptively collected data: the power of online learning Off-policy evaluation via adaptive weighting with data from contextual bandits

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.706252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.589198Z digest=sha256:c7d03c622b9e7fb08405357e72b049e487a993b2de2a945a425caf23ed9405af

Observation 7a83b08d-89d3-434a-b576-448588353f15 · outbound

This paper cites Policy learning with adaptively collected data.

Off-policy estimation with adaptively collected data: the power of online learning Policy learning with adaptively collected data

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.696569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.591909Z digest=sha256:6189007e6cf247f8777fb0c09f8f1546a7c21eea4bd96b086544da48cb3cdb74

Observation 755424cb-481b-4825-a410-50b2442d6e48 · outbound

This paper cites Inference for batched bandits.

Off-policy estimation with adaptively collected data: the power of online learning Inference for batched bandits

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:43:58.686245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T17:43:58.594829Z digest=sha256:a8a7ec104852d7328266d8e278c158f458354e3d2a0fb200427ffdd1a56e4e61

Pith citing papers

No inbound Pith citation observations are available.