Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback Experiments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wen, Qianglin, Shi, Chengchun, Yang, Ying, Tang, Niansheng, Zhu, Hongtu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robust Sequential Experimental Design for A/B Testing
von: Wen, Qianglin, et al.
Veröffentlicht: (2026)
von: Wen, Qianglin, et al.
Veröffentlicht: (2026)
Designing Time Series Experiments in A/B Testing with Transformer Reinforcement Learning
von: Wu, Xiangkun, et al.
Veröffentlicht: (2026)
von: Wu, Xiangkun, et al.
Veröffentlicht: (2026)
Combining Experimental and Historical Data for Policy Evaluation
von: Li, Ting, et al.
Veröffentlicht: (2024)
von: Li, Ting, et al.
Veröffentlicht: (2024)
A Two-armed Bandit Framework for A/B Testing
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025)
Deep Distributional Learning with Non-crossing Quantile Network
von: Shen, Guohao, et al.
Veröffentlicht: (2025)
von: Shen, Guohao, et al.
Veröffentlicht: (2025)
Beyond Passive Aggregation: Active Auditing and Topology-Aware Defense in Decentralized Federated Learning
von: Pan, Sheng, et al.
Veröffentlicht: (2026)
von: Pan, Sheng, et al.
Veröffentlicht: (2026)
Statistical Inference in Reinforcement Learning: A Selective Survey
von: Shi, Chengchun
Veröffentlicht: (2025)
von: Shi, Chengchun
Veröffentlicht: (2025)
Robust Offline Reinforcement learning with Heavy-Tailed Rewards
von: Zhu, Jin, et al.
Veröffentlicht: (2023)
von: Zhu, Jin, et al.
Veröffentlicht: (2023)
Causal Deepsets for Off-policy Evaluation under Spatial or Spatio-temporal Interferences
von: Dai, Runpeng, et al.
Veröffentlicht: (2024)
von: Dai, Runpeng, et al.
Veröffentlicht: (2024)
Detecting LLM-Generated Text with Performance Guarantees
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
Clustered Switchback Designs for Experimentation Under Spatio-temporal Interference
von: Jia, Su, et al.
Veröffentlicht: (2023)
von: Jia, Su, et al.
Veröffentlicht: (2023)
Demystifying the Paradox of Importance Sampling with an Estimated History-Dependent Behavior Policy in Off-Policy Evaluation
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2025)
Data-Driven Switchback Experiments: Theoretical Tradeoffs and Empirical Bayes Designs
von: Xiong, Ruoxuan, et al.
Veröffentlicht: (2024)
von: Xiong, Ruoxuan, et al.
Veröffentlicht: (2024)
Balancing Interference and Correlation in Spatial Experimental Designs: A Causal Graph Cut Approach
von: Zhu, Jin, et al.
Veröffentlicht: (2025)
von: Zhu, Jin, et al.
Veröffentlicht: (2025)
From Authors to Reviewers: Leveraging Rankings to Improve Peer Review
von: Wang, Weichen, et al.
Veröffentlicht: (2025)
von: Wang, Weichen, et al.
Veröffentlicht: (2025)
Demystifying Group Relative Policy Optimization: Its Policy Gradient is a U-Statistic
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
Dual Active Learning for Reinforcement Learning from Human Feedback
von: Liu, Pangpang, et al.
Veröffentlicht: (2024)
von: Liu, Pangpang, et al.
Veröffentlicht: (2024)
Reinforcement Learning from Human Feedback: A Statistical Perspective
von: Liu, Pangpang, et al.
Veröffentlicht: (2026)
von: Liu, Pangpang, et al.
Veröffentlicht: (2026)
Counterfactually Safe Reinforcement Learning
von: Li, Jingyi, et al.
Veröffentlicht: (2026)
von: Li, Jingyi, et al.
Veröffentlicht: (2026)
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
von: Zhou, Hongyi, et al.
Veröffentlicht: (2026)
Perturbation is All You Need for Extrapolating Language Models
von: Cen, Zetai, et al.
Veröffentlicht: (2026)
von: Cen, Zetai, et al.
Veröffentlicht: (2026)
ARMA-Design: Optimal Treatment Allocation Strategies for A/B Testing in Partially Observable Time Series Experiments
von: Sun, Ke, et al.
Veröffentlicht: (2024)
von: Sun, Ke, et al.
Veröffentlicht: (2024)
Fast Gaussian Process Approximations for Autocorrelated Data
von: Chokhachian, Ahmadreza, et al.
Veröffentlicht: (2025)
von: Chokhachian, Ahmadreza, et al.
Veröffentlicht: (2025)
Testing Stationarity and Change Point Detection in Reinforcement Learning
von: Li, Mengbing, et al.
Veröffentlicht: (2022)
von: Li, Mengbing, et al.
Veröffentlicht: (2022)
Double Fairness Policy Learning: Integrating Action Fairness and Outcome Fairness in Decision-making
von: Bian, Zeyu, et al.
Veröffentlicht: (2026)
von: Bian, Zeyu, et al.
Veröffentlicht: (2026)
Generalized Fitted Q-Iteration with Clustered Data
von: Hu, Liyuan, et al.
Veröffentlicht: (2025)
von: Hu, Liyuan, et al.
Veröffentlicht: (2025)
Spatio-temporal Prediction of Fine-Grained Origin-Destination Matrices with Applications in Ridesharing
von: Yang, Run, et al.
Veröffentlicht: (2025)
von: Yang, Run, et al.
Veröffentlicht: (2025)
Deep Autocorrelation Modeling for Time-Series Forecasting: Progress and Prospects
von: Wang, Hao, et al.
Veröffentlicht: (2026)
von: Wang, Hao, et al.
Veröffentlicht: (2026)
Analyzing and Bridging the Gap between Maximizing Total Reward and Discounted Reward in Deep Reinforcement Learning
von: Yin, Shuyu, et al.
Veröffentlicht: (2024)
von: Yin, Shuyu, et al.
Veröffentlicht: (2024)
Robust Reinforcement Learning from Human Feedback for Large Language Models Fine-Tuning
von: Ye, Kai, et al.
Veröffentlicht: (2025)
von: Ye, Kai, et al.
Veröffentlicht: (2025)
Partially Functional Dynamic Backdoor Diffusion-based Causal Model
von: Liu, Xinwen, et al.
Veröffentlicht: (2025)
von: Liu, Xinwen, et al.
Veröffentlicht: (2025)
Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning
von: Gong, Shijin, et al.
Veröffentlicht: (2026)
von: Gong, Shijin, et al.
Veröffentlicht: (2026)
Off-policy Evaluation in Doubly Inhomogeneous Environments
von: Bian, Zeyu, et al.
Veröffentlicht: (2023)
von: Bian, Zeyu, et al.
Veröffentlicht: (2023)
Breach in the Shield: Unveiling the Vulnerabilities of Large Language Models
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)
Spatially Randomized Designs Can Enhance Policy Evaluation
von: Yang, Ying, et al.
Veröffentlicht: (2024)
von: Yang, Ying, et al.
Veröffentlicht: (2024)
Pessimistic Causal Reinforcement Learning with Mediators for Confounded Offline Data
von: Wang, Danyang, et al.
Veröffentlicht: (2024)
von: Wang, Danyang, et al.
Veröffentlicht: (2024)
Enhancing Missing Data Imputation through Combined Bipartite Graph and Complete Directed Graph
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2024)
Sampling-guided Heterogeneous Graph Neural Network with Temporal Smoothing for Scalable Longitudinal Data Imputation
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2024)
von: Zhang, Zhaoyang, et al.
Veröffentlicht: (2024)
High-Dimensional Dynamic Covariance Models with Random Forests
von: Yu, Shuguang, et al.
Veröffentlicht: (2025)
von: Yu, Shuguang, et al.
Veröffentlicht: (2025)
Autocorrelation Matters: Understanding the Role of Initialization Schemes for State Space Models
von: Liu, Fusheng, et al.
Veröffentlicht: (2024)
von: Liu, Fusheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Robust Sequential Experimental Design for A/B Testing
von: Wen, Qianglin, et al.
Veröffentlicht: (2026) -
Designing Time Series Experiments in A/B Testing with Transformer Reinforcement Learning
von: Wu, Xiangkun, et al.
Veröffentlicht: (2026) -
Combining Experimental and Historical Data for Policy Evaluation
von: Li, Ting, et al.
Veröffentlicht: (2024) -
A Two-armed Bandit Framework for A/B Testing
von: Wang, Jinjuan, et al.
Veröffentlicht: (2025) -
Deep Distributional Learning with Non-crossing Quantile Network
von: Shen, Guohao, et al.
Veröffentlicht: (2025)