|国家预印本平台
首页|Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback Experiments

Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback Experiments

Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback Experiments

来源:Arxiv_logoArxiv
英文摘要

A/B testing has become the gold standard for policy evaluation in modern technological industries. Motivated by the widespread use of switchback experiments in A/B testing, this paper conducts a comprehensive comparative analysis of various switchback designs in Markovian environments. Unlike many existing works which derive the optimal design based on specific and relatively simple estimators, our analysis covers a range of state-of-the-art estimators developed in the reinforcement learning (RL) literature. It reveals that the effectiveness of different switchback designs depends crucially on (i) the size of the carryover effect and (ii) the auto-correlations among reward errors over time. Meanwhile, these findings are estimator-agnostic, i.e., they apply to most RL estimators. Based on these insights, we provide a workflow to offer guidelines for practitioners on designing switchback experiments in A/B testing.

Qianglin Wen、Chengchun Shi、Yang Ying、Niansheng Tang、Hongtu Zhu

计算技术、计算机技术

Qianglin Wen,Chengchun Shi,Yang Ying,Niansheng Tang,Hongtu Zhu.Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback Experiments[EB/OL].(2025-07-11)[2025-07-16].https://arxiv.org/abs/2403.17285.点此复制

评论