REVIEW 2 cited by
Tackling Interference Induced by Data Training Loops in A/B Tests: A Weighted Training Approach
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In modern recommendation systems, the standard pipeline involves training machine learning models on historical data to predict user behaviors and improve recommendations continuously. However, these data training loops can introduce interference in A/B tests, where data generated by control and treatment algorithms, potentially with different distributions, are combined. To address these challenges, we introduce a novel approach called weighted training. This approach entails training a model to predict the probability of each data point appearing in either the treatment or control data and subsequently applying weighted losses during model training. We demonstrate that this approach achieves the least variance among all estimators that do not cause shifts in the training distributions. Through simulation studies, we demonstrate the lower bias and variance of our approach compared to other methods.
Forward citations
Cited by 2 Pith papers
-
Causal Estimation of Share-Induced Engagement with Flywheel Effects
A flow-balance identity yields a closed-form geometric-amplification estimator of the global treatment effect of sharing features under network flywheel interference, with consistency under homogeneity and valid A/A i...
-
Experimental Designs for Multi-Item Multi-Period Inventory Control
Switchback experiments underestimate the global treatment effect in shared-capacity inventory systems, item-level randomization overestimates it, and a pairwise item-time design has intermediate bias.
Discussion (0). Continue with ORCID to comment.