Closing the Loop: Coordinating Inventory and Recommendation via Deep Reinforcement Learning on Multiple Timescales
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Jinyang, Han, Jinhui, Peng, Yijie, Zhang, Ying |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Classical and Deep Reinforcement Learning Inventory Control Policies for Pharmaceutical Supply Chains with Perishability and Non-Stationarity
by: Stranieri, Francesco, et al.
Published: (2025)
by: Stranieri, Francesco, et al.
Published: (2025)
Solving Integrated Process Planning and Scheduling Problem via Graph Neural Network Based Deep Reinforcement Learning
by: Li, Hongpei, et al.
Published: (2024)
by: Li, Hongpei, et al.
Published: (2024)
Integrated Offline and Online Learning to Solve a Large Class of Scheduling Problems
by: Liu, Anbang, et al.
Published: (2025)
by: Liu, Anbang, et al.
Published: (2025)
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
by: Jing, Gangshan, et al.
Published: (2021)
by: Jing, Gangshan, et al.
Published: (2021)
Action Dependency Graphs for Globally Optimal Coordinated Reinforcement Learning
by: Ding, Jianglin, et al.
Published: (2025)
by: Ding, Jianglin, et al.
Published: (2025)
A Deep Q-Network Based on Radial Basis Functions for Multi-Echelon Inventory Management
by: Cheng, Liqiang, et al.
Published: (2024)
by: Cheng, Liqiang, et al.
Published: (2024)
Deep Reinforcement Learning for Traveling Purchaser Problems
by: Yuan, Haofeng, et al.
Published: (2024)
by: Yuan, Haofeng, et al.
Published: (2024)
OptiRepair: Closed-Loop Diagnosis and Repair of Supply Chain Optimization Models with LLM Agents
by: Ao, Ruicheng, et al.
Published: (2026)
by: Ao, Ruicheng, et al.
Published: (2026)
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning
by: Li, Jingqi, et al.
Published: (2022)
by: Li, Jingqi, et al.
Published: (2022)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
by: Xiao, Minheng, et al.
Published: (2024)
by: Xiao, Minheng, et al.
Published: (2024)
Agentic Transformers Provably Learn to Search via Reinforcement Learning
by: Yang, Tong, et al.
Published: (2026)
by: Yang, Tong, et al.
Published: (2026)
Deep Reinforcement Learning for Solving the Fleet Size and Mix Vehicle Routing Problem
by: Wan, Pengfu, et al.
Published: (2025)
by: Wan, Pengfu, et al.
Published: (2025)
Faster Reinforcement Learning by Freezing Slow States
by: Wang, Yijia, et al.
Published: (2023)
by: Wang, Yijia, et al.
Published: (2023)
Stochastic Optimization of Inventory at Large-scale Supply Chains
by: Jin, Zhaoyang Larry, et al.
Published: (2025)
by: Jin, Zhaoyang Larry, et al.
Published: (2025)
Optimizing Inventory Routing: A Decision-Focused Learning Approach using Neural Networks
by: Islam, MD Shafikul, et al.
Published: (2023)
by: Islam, MD Shafikul, et al.
Published: (2023)
Accelerating Cutting-Plane Algorithms via Reinforcement Learning Surrogates
by: Mana, Kyle, et al.
Published: (2023)
by: Mana, Kyle, et al.
Published: (2023)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
by: Cui, Mingxuan, et al.
Published: (2025)
by: Cui, Mingxuan, et al.
Published: (2025)
Reinforcement Learning under Latent Dynamics: Toward Statistical and Algorithmic Modularity
by: Amortila, Philip, et al.
Published: (2024)
by: Amortila, Philip, et al.
Published: (2024)
Stochastic Approximation Methods for Distortion Risk Measure Optimization
by: Jiang, Jinyang, et al.
Published: (2025)
by: Jiang, Jinyang, et al.
Published: (2025)
Hierarchical Deep Reinforcement Learning Framework for Multi-Year Asset Management Under Budget Constraints
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
Universal Approximation Theorem for Deep Q-Learning via FBSDE System
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
AI2STOW: End-to-End Deep Reinforcement Learning to Construct Master Stowage Plans under Demand Uncertainty
by: Van Twiller, Jaike, et al.
Published: (2025)
by: Van Twiller, Jaike, et al.
Published: (2025)
Deterministic Policy Gradient for Reinforcement Learning with Continuous Time and State
by: Cheng, Ziheng, et al.
Published: (2025)
by: Cheng, Ziheng, et al.
Published: (2025)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
by: Anand, Akhil S, et al.
Published: (2025)
by: Anand, Akhil S, et al.
Published: (2025)
Joint Problems in Learning Multiple Dynamical Systems
by: Niu, Mengjia, et al.
Published: (2023)
by: Niu, Mengjia, et al.
Published: (2023)
R2L: Reliable Reinforcement Learning: Guaranteed Return & Reliable Policies in Reinforcement Learning
by: Farhi, Nadir
Published: (2025)
by: Farhi, Nadir
Published: (2025)
Learning the Riccati solution operator for time-varying LQR via Deep Operator Networks
by: Chen, Jun, et al.
Published: (2026)
by: Chen, Jun, et al.
Published: (2026)
Contextual Bilevel Reinforcement Learning for Incentive Alignment
by: Thoma, Vinzenz, et al.
Published: (2024)
by: Thoma, Vinzenz, et al.
Published: (2024)
Structure-Informed Deep Reinforcement Learning for Inventory Management
by: Maggiar, Alvaro, et al.
Published: (2025)
by: Maggiar, Alvaro, et al.
Published: (2025)
An Improved Finite-time Analysis of Temporal Difference Learning with Deep Neural Networks
by: Ke, Zhifa, et al.
Published: (2024)
by: Ke, Zhifa, et al.
Published: (2024)
Benchmarking Reinforcement Learning via Stochastic Converse Optimality: Generating Systems with Known Optimal Policies
by: Ibrahim, Sinan, et al.
Published: (2026)
by: Ibrahim, Sinan, et al.
Published: (2026)
Score as Action: Fine-Tuning Diffusion Generative Models by Continuous-time Reinforcement Learning
by: Zhao, Hanyang, et al.
Published: (2025)
by: Zhao, Hanyang, et al.
Published: (2025)
Adaptive Primal-Dual Method for Safe Reinforcement Learning
by: Chen, Weiqin, et al.
Published: (2024)
by: Chen, Weiqin, et al.
Published: (2024)
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Reinforcement Learning for Multi-Truck Vehicle Routing Problems
by: Levin, Joshua, et al.
Published: (2022)
by: Levin, Joshua, et al.
Published: (2022)
Hyperparameter Optimization for Driving Strategies Based on Reinforcement Learning
by: Adde, Nihal Acharya, et al.
Published: (2024)
by: Adde, Nihal Acharya, et al.
Published: (2024)
Understanding Optimization in Deep Learning with Central Flows
by: Cohen, Jeremy M., et al.
Published: (2024)
by: Cohen, Jeremy M., et al.
Published: (2024)
Combining Reinforcement Learning and Optimal Transport for the Traveling Salesman Problem
by: Goh, Yong Liang, et al.
Published: (2022)
by: Goh, Yong Liang, et al.
Published: (2022)
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning
by: Ma, Chaolun, et al.
Published: (2022)
by: Ma, Chaolun, et al.
Published: (2022)
Deep Learning for Two-Stage Robust Integer Optimization
by: Dumouchelle, Justin, et al.
Published: (2023)
by: Dumouchelle, Justin, et al.
Published: (2023)
Similar Items
-
Classical and Deep Reinforcement Learning Inventory Control Policies for Pharmaceutical Supply Chains with Perishability and Non-Stationarity
by: Stranieri, Francesco, et al.
Published: (2025) -
Solving Integrated Process Planning and Scheduling Problem via Graph Neural Network Based Deep Reinforcement Learning
by: Li, Hongpei, et al.
Published: (2024) -
Integrated Offline and Online Learning to Solve a Large Class of Scheduling Problems
by: Liu, Anbang, et al.
Published: (2025) -
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
by: Jing, Gangshan, et al.
Published: (2021) -
Action Dependency Graphs for Globally Optimal Coordinated Reinforcement Learning
by: Ding, Jianglin, et al.
Published: (2025)