Simulation-Based Benchmarking of Reinforcement Learning Agents for Personalized Retail Promotions
Fuente:
arXiv
Saved in:
| Main Authors: | Xia, Yu, Narayanamoorthy, Sriram, Zhou, Zhengyuan, Mabry, Joshua |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Hybrid Framework for Reinsurance Optimization: Integrating Generative Models and Reinforcement Learning
by: Dong, Stella C.
Published: (2025)
by: Dong, Stella C.
Published: (2025)
LLM Personas as a Substitute for Field Experiments in Method Benchmarking
by: Kang, Enoch Hyunwook
Published: (2025)
by: Kang, Enoch Hyunwook
Published: (2025)
Multi-Agent Reinforcement Learning for Dynamic Pricing in Supply Chains: Benchmarking Strategic Agent Behaviours under Realistically Simulated Market Conditions
by: Hazenberg, Thomas, et al.
Published: (2025)
by: Hazenberg, Thomas, et al.
Published: (2025)
Management Decisions in Manufacturing using Causal Machine Learning -- To Rework, or not to Rework?
by: Schwarz, Philipp, et al.
Published: (2024)
by: Schwarz, Philipp, et al.
Published: (2024)
Reinforcement Learning for Monetary Policy Under Macroeconomic Uncertainty: Analyzing Tabular and Function Approximation Methods
by: Wang, Tony, et al.
Published: (2025)
by: Wang, Tony, et al.
Published: (2025)
Learning Individual Behavior in Agent-Based Models with Graph Diffusion Networks
by: Cozzi, Francesco, et al.
Published: (2025)
by: Cozzi, Francesco, et al.
Published: (2025)
Adaptive Experimental Design for Policy Learning
by: Kato, Masahiro, et al.
Published: (2024)
by: Kato, Masahiro, et al.
Published: (2024)
Selective Reviews of Bandit Problems in AI via a Statistical View
by: Zhou, Pengjie, et al.
Published: (2024)
by: Zhou, Pengjie, et al.
Published: (2024)
Learning Causal Representations from General Environments: Identifiability and Intrinsic Ambiguity
by: Jin, Jikai, et al.
Published: (2023)
by: Jin, Jikai, et al.
Published: (2023)
Learning from Double Positive and Unlabeled Data for Potential-Customer Identification
by: Kato, Masahiro, et al.
Published: (2025)
by: Kato, Masahiro, et al.
Published: (2025)
A Double Machine Learning Approach to Combining Experimental and Observational Data
by: Parikh, Harsh, et al.
Published: (2023)
by: Parikh, Harsh, et al.
Published: (2023)
Generating density nowcasts for U.S. GDP growth with deep learning: Bayes by Backprop and Monte Carlo dropout
by: Németh, Kristóf, et al.
Published: (2024)
by: Németh, Kristóf, et al.
Published: (2024)
Unveiling the Potential of Robustness in Selecting Conditional Average Treatment Effect Estimators
by: Huang, Yiyan, et al.
Published: (2024)
by: Huang, Yiyan, et al.
Published: (2024)
DeXposure-FM: A Time-series, Graph Foundation Model for Credit Exposures and Stability on Decentralized Financial Networks
by: Shu, Aijie, et al.
Published: (2026)
by: Shu, Aijie, et al.
Published: (2026)
Foundation Priors
by: Misra, Sanjog
Published: (2025)
by: Misra, Sanjog
Published: (2025)
Causal EpiNets: Precision-corrected Bounds on Individual Treatment Effects using Epistemic Neural Networks
by: Patil, Gandharv, et al.
Published: (2026)
by: Patil, Gandharv, et al.
Published: (2026)
Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice
by: Wang, Yingshuo, et al.
Published: (2026)
by: Wang, Yingshuo, et al.
Published: (2026)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
by: Kallus, Nathan
Published: (2025)
by: Kallus, Nathan
Published: (2025)
GDP nowcasting with artificial neural networks: How much does long-term memory matter?
by: Németh, Kristóf, et al.
Published: (2023)
by: Németh, Kristóf, et al.
Published: (2023)
An Empirical Risk Minimization Approach for Offline Inverse RL and Dynamic Discrete Choice Model
by: Kang, Enoch H., et al.
Published: (2025)
by: Kang, Enoch H., et al.
Published: (2025)
How Well Do LLMs Predict Human Behavior? A Measure of their Pretrained Knowledge
by: Gao, Wayne, et al.
Published: (2026)
by: Gao, Wayne, et al.
Published: (2026)
Structural Estimation of Markov Decision Processes in High-Dimensional State Space with Finite-Time Guarantees
by: Zeng, Siliang, et al.
Published: (2022)
by: Zeng, Siliang, et al.
Published: (2022)
Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies
by: Petrungaro, Bruno, et al.
Published: (2026)
by: Petrungaro, Bruno, et al.
Published: (2026)
Robust Time Series Causal Discovery for Agent-Based Model Validation
by: Yu, Gene, et al.
Published: (2024)
by: Yu, Gene, et al.
Published: (2024)
Towards Generalizing Inferences from Trials to Target Populations
by: Huang, Melody Y, et al.
Published: (2024)
by: Huang, Melody Y, et al.
Published: (2024)
Generalized Neyman Allocation for Locally Minimax Optimal Best-Arm Identification
by: Kato, Masahiro
Published: (2024)
by: Kato, Masahiro
Published: (2024)
Statistical Tests for Replacing Human Decision Makers with Algorithms
by: Feng, Kai, et al.
Published: (2023)
by: Feng, Kai, et al.
Published: (2023)
Global Ease of Living Index: a machine learning framework for longitudinal analysis of major economies
by: Panat, Tanay, et al.
Published: (2025)
by: Panat, Tanay, et al.
Published: (2025)
Multi-Band Variable-Lag Granger Causality: A Unified Framework for Causal Time Series Inference across Frequencies
by: Sookkongwaree, Chakattrai, et al.
Published: (2025)
by: Sookkongwaree, Chakattrai, et al.
Published: (2025)
From What Ifs to Insights: Counterfactuals in Causal Inference vs. Explainable AI
by: Shmueli, Galit, et al.
Published: (2025)
by: Shmueli, Galit, et al.
Published: (2025)
$ρ$-GNF: A Copula-based Sensitivity Analysis to Unobserved Confounding Using Normalizing Flows
by: Balgi, Sourabh, et al.
Published: (2022)
by: Balgi, Sourabh, et al.
Published: (2022)
Unified Causality Analysis Based on the Degrees of Freedom
by: Telcs, András, et al.
Published: (2024)
by: Telcs, András, et al.
Published: (2024)
Non-linear Phillips Curve for India: Evidence from Explainable Machine Learning
by: Sengupta, Shovon, et al.
Published: (2025)
by: Sengupta, Shovon, et al.
Published: (2025)
Transformers Handle Endogeneity in In-Context Linear Regression
by: Liang, Haodong, et al.
Published: (2024)
by: Liang, Haodong, et al.
Published: (2024)
Causality Elicitation from Large Language Models
by: Kameyama, Takashi, et al.
Published: (2026)
by: Kameyama, Takashi, et al.
Published: (2026)
Differentially Private Two-Stage Gradient Descent for Instrumental Variable Regression
by: Liang, Haodong, et al.
Published: (2025)
by: Liang, Haodong, et al.
Published: (2025)
What Is the Alignment Tax?
by: Young, Robin
Published: (2026)
by: Young, Robin
Published: (2026)
Post Reinforcement Learning Inference
by: Syrgkanis, Vasilis, et al.
Published: (2023)
by: Syrgkanis, Vasilis, et al.
Published: (2023)
Bellman Calibration for $V$-Learning in Offline Reinforcement Learning
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Policy Learning with Competing Agents
by: Sahoo, Roshni, et al.
Published: (2022)
by: Sahoo, Roshni, et al.
Published: (2022)
Similar Items
-
A Hybrid Framework for Reinsurance Optimization: Integrating Generative Models and Reinforcement Learning
by: Dong, Stella C.
Published: (2025) -
LLM Personas as a Substitute for Field Experiments in Method Benchmarking
by: Kang, Enoch Hyunwook
Published: (2025) -
Multi-Agent Reinforcement Learning for Dynamic Pricing in Supply Chains: Benchmarking Strategic Agent Behaviours under Realistically Simulated Market Conditions
by: Hazenberg, Thomas, et al.
Published: (2025) -
Management Decisions in Manufacturing using Causal Machine Learning -- To Rework, or not to Rework?
by: Schwarz, Philipp, et al.
Published: (2024) -
Reinforcement Learning for Monetary Policy Under Macroeconomic Uncertainty: Analyzing Tabular and Function Approximation Methods
by: Wang, Tony, et al.
Published: (2025)