Thompson Sampling for Repeated Newsvendor
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Li, Qin, Hanzhang, Xu, Yunbei, Zhu, Ruihao, Zhang, Weizhou |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On Pareto Optimality for Parametric Choice Bandits
von: Zuo, Jierui, et al.
Veröffentlicht: (2025)
von: Zuo, Jierui, et al.
Veröffentlicht: (2025)
Closing the Gaps: Optimality of Sample Average Approximation for Data-Driven Newsvendor Problems
von: Lyu, Jiameng, et al.
Veröffentlicht: (2024)
von: Lyu, Jiameng, et al.
Veröffentlicht: (2024)
Upper Counterfactual Confidence Bounds: a New Optimism Principle for Contextual Bandits
von: Xu, Yunbei, et al.
Veröffentlicht: (2020)
von: Xu, Yunbei, et al.
Veröffentlicht: (2020)
Deep Generative Demand Learning for Newsvendor and Pricing
von: Gong, Shijin, et al.
Veröffentlicht: (2024)
von: Gong, Shijin, et al.
Veröffentlicht: (2024)
Survey of Data-driven Newsvendor: Unified Analysis and Spectrum of Achievable Regrets
von: Chen, Zhuoxin, et al.
Veröffentlicht: (2024)
von: Chen, Zhuoxin, et al.
Veröffentlicht: (2024)
Sample-Mean Anchored Thompson Sampling for Offline-to-Online Learning with Distribution Shift
von: Li, Bochao, et al.
Veröffentlicht: (2026)
von: Li, Bochao, et al.
Veröffentlicht: (2026)
The Data-Driven Censored Newsvendor Problem
von: Hssaine, Chamsi, et al.
Veröffentlicht: (2024)
von: Hssaine, Chamsi, et al.
Veröffentlicht: (2024)
Constrained Linear Thompson Sampling
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025)
von: Gangrade, Aditya, et al.
Veröffentlicht: (2025)
Pointwise Generalization in Deep Neural Networks
von: Li, Shaojie, et al.
Veröffentlicht: (2026)
von: Li, Shaojie, et al.
Veröffentlicht: (2026)
Statistical Properties of Robust Satisficing
von: Li, Zhiyi, et al.
Veröffentlicht: (2024)
von: Li, Zhiyi, et al.
Veröffentlicht: (2024)
Thompson Sampling For Combinatorial Bandits: Polynomial Regret and Mismatched Sampling Paradox
von: Zhang, Raymond, et al.
Veröffentlicht: (2024)
von: Zhang, Raymond, et al.
Veröffentlicht: (2024)
Bayesian Design Principles for Frequentist Sequential Learning
von: Xu, Yunbei, et al.
Veröffentlicht: (2023)
von: Xu, Yunbei, et al.
Veröffentlicht: (2023)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
von: Namkoong, Hongseok, et al.
Veröffentlicht: (2020)
von: Namkoong, Hongseok, et al.
Veröffentlicht: (2020)
Autoregressive Learning in Joint KL: Sharp Oracle Bounds and Lower Bounds
von: Xu, Yunbei, et al.
Veröffentlicht: (2026)
von: Xu, Yunbei, et al.
Veröffentlicht: (2026)
On the Power of Adaptivity for $\varepsilon$-Best Arm Identification in Linear Bandits
von: Maiti, Arnab, et al.
Veröffentlicht: (2026)
von: Maiti, Arnab, et al.
Veröffentlicht: (2026)
Regenerative Particle Thompson Sampling
von: Zhou, Zeyu, et al.
Veröffentlicht: (2022)
von: Zhou, Zeyu, et al.
Veröffentlicht: (2022)
Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks
von: Xu, Yinglun, et al.
Veröffentlicht: (2024)
von: Xu, Yinglun, et al.
Veröffentlicht: (2024)
Adaptive Data Augmentation for Thompson Sampling
von: Kim, Wonyoung
Veröffentlicht: (2025)
von: Kim, Wonyoung
Veröffentlicht: (2025)
A Broader View of Thompson Sampling
von: Qu, Yanlin, et al.
Veröffentlicht: (2025)
von: Qu, Yanlin, et al.
Veröffentlicht: (2025)
Stochastically Constrained Best Arm Identification with Thompson Sampling
von: Yang, Le, et al.
Veröffentlicht: (2025)
von: Yang, Le, et al.
Veröffentlicht: (2025)
Understanding the Transferability of Representations via Task-Relatedness
von: Mehra, Akshay, et al.
Veröffentlicht: (2023)
von: Mehra, Akshay, et al.
Veröffentlicht: (2023)
Test-time Assessment of a Model's Performance on Unseen Domains via Optimal Transport
von: Mehra, Akshay, et al.
Veröffentlicht: (2024)
von: Mehra, Akshay, et al.
Veröffentlicht: (2024)
Risk-Aware Linear Bandits: Theory and Applications in Smart Order Routing
von: Ji, Jingwei, et al.
Veröffentlicht: (2022)
von: Ji, Jingwei, et al.
Veröffentlicht: (2022)
Graph Neural Thompson Sampling
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
Epsilon-Greedy Thompson Sampling to Bayesian Optimization
von: Do, Bach, et al.
Veröffentlicht: (2024)
von: Do, Bach, et al.
Veröffentlicht: (2024)
Gaussian Process Thompson Sampling via Rootfinding
von: Adebiyi, Taiwo A., et al.
Veröffentlicht: (2024)
von: Adebiyi, Taiwo A., et al.
Veröffentlicht: (2024)
Learning When to Restart: Nonstationary Newsvendor from Uncensored to Censored Demand
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Is Thompson Sampling Susceptible to Algorithmic Collusion?
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
von: Xiong, Yi, et al.
Veröffentlicht: (2024)
BFTS: Thompson Sampling with Bayesian Additive Regression Trees
von: Deng, Ruizhe, et al.
Veröffentlicht: (2026)
von: Deng, Ruizhe, et al.
Veröffentlicht: (2026)
Fast, Precise Thompson Sampling for Bayesian Optimization
von: Sweet, David
Veröffentlicht: (2024)
von: Sweet, David
Veröffentlicht: (2024)
Thompson Sampling in Partially Observable Contextual Bandits
von: Park, Hongju, et al.
Veröffentlicht: (2024)
von: Park, Hongju, et al.
Veröffentlicht: (2024)
Online Learning of Decision Trees with Thompson Sampling
von: Chaouki, Ayman, et al.
Veröffentlicht: (2024)
von: Chaouki, Ayman, et al.
Veröffentlicht: (2024)
On Regret Bounds of Thompson Sampling for Bayesian Optimization
von: Takeno, Shion, et al.
Veröffentlicht: (2026)
von: Takeno, Shion, et al.
Veröffentlicht: (2026)
EVA-0: Test-Time Model Evolution with Only Two Forward Passes per Sample
von: Chen, Guohao, et al.
Veröffentlicht: (2026)
von: Chen, Guohao, et al.
Veröffentlicht: (2026)
A Conformal Approach to Feature-based Newsvendor under Model Misspecification
von: Cao, Junyu
Veröffentlicht: (2024)
von: Cao, Junyu
Veröffentlicht: (2024)
What is the Value of Censored Data? An Exact Analysis for the Data-driven Newsvendor
von: Kumar, Rachitesh, et al.
Veröffentlicht: (2026)
von: Kumar, Rachitesh, et al.
Veröffentlicht: (2026)
Contextual Online Pricing with (Biased) Offline Data
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
On the Peril of (Even a Little) Nonstationarity in Satisficing Regret Minimization
von: Zhang, Yixuan, et al.
Veröffentlicht: (2026)
von: Zhang, Yixuan, et al.
Veröffentlicht: (2026)
Feel-Good Thompson Sampling for Contextual Dueling Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
Online Resource Allocation with Non-Stationary Customers
von: Zhang, Xiaoyue, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoyue, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
On Pareto Optimality for Parametric Choice Bandits
von: Zuo, Jierui, et al.
Veröffentlicht: (2025) -
Closing the Gaps: Optimality of Sample Average Approximation for Data-Driven Newsvendor Problems
von: Lyu, Jiameng, et al.
Veröffentlicht: (2024) -
Upper Counterfactual Confidence Bounds: a New Optimism Principle for Contextual Bandits
von: Xu, Yunbei, et al.
Veröffentlicht: (2020) -
Deep Generative Demand Learning for Newsvendor and Pricing
von: Gong, Shijin, et al.
Veröffentlicht: (2024) -
Survey of Data-driven Newsvendor: Unified Analysis and Spectrum of Achievable Regrets
von: Chen, Zhuoxin, et al.
Veröffentlicht: (2024)