Projection-based Lyapunov method for fully heterogeneous weakly-coupled MDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xiangcheng, Hong, Yige, Wang, Weina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Achieving Exponential Asymptotic Optimality in Average-Reward Restless Bandits without Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2024)
by: Hong, Yige, et al.
Published: (2024)
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2023)
by: Hong, Yige, et al.
Published: (2023)
Unichain and Aperiodicity are Sufficient for Asymptotic Optimality of Average-Reward Restless Bandits
by: Hong, Yige, et al.
Published: (2024)
by: Hong, Yige, et al.
Published: (2024)
Semi-Supervised Clustering of Sparse Graphs: Crossing the Information-Theoretic Threshold
by: Sheng, Junda, et al.
Published: (2022)
by: Sheng, Junda, et al.
Published: (2022)
A note on weak compactness of occupation measures for an absorbing Markov decision process
by: Dufour, François, et al.
Published: (2024)
by: Dufour, François, et al.
Published: (2024)
Adaptive Conditional Forest Sampling for Spectral Risk Optimisation under Decision-Dependent Uncertainty
by: Kurbucz, Marcell T.
Published: (2026)
by: Kurbucz, Marcell T.
Published: (2026)
Reinforcement Learning with Markov Risk Measures and Multipattern Risk Approximation
by: Ruszczynski, Andrzej, et al.
Published: (2026)
by: Ruszczynski, Andrzej, et al.
Published: (2026)
Measuring diversity. A review and an empirical analysis
by: Parreño, Francisco, et al.
Published: (2024)
by: Parreño, Francisco, et al.
Published: (2024)
Strategy Complexity of Limsup and Liminf Threshold Objectives in Countable MDPs, with Applications to Optimal Expected Payoffs
by: Mayr, Richard, et al.
Published: (2022)
by: Mayr, Richard, et al.
Published: (2022)
Variance Reduced Policy Gradient Method for Multi-Objective Reinforcement Learning
by: Guidobene, Davide, et al.
Published: (2025)
by: Guidobene, Davide, et al.
Published: (2025)
Multi-Objective Optimization with Desirability and Morris-Mitchell Criterion
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
Dynamic resource matching in manufacturing using deep reinforcement learning
by: Panda, Saunak Kumar, et al.
Published: (2026)
by: Panda, Saunak Kumar, et al.
Published: (2026)
Projected Inventory Level Policies for Lost Sales Inventory Systems: Asymptotic Optimality in Two Regimes
by: van Jaarsveld, Willem, et al.
Published: (2021)
by: van Jaarsveld, Willem, et al.
Published: (2021)
Cost-Effective Strategies for Infectious Diseases: A Multi-Objective Framework with an Interactive Dashboard
by: Lee, Jongmin, et al.
Published: (2025)
by: Lee, Jongmin, et al.
Published: (2025)
Multi-Objective Optimization and Hyperparameter Tuning With Desirability Functions
by: Bartz-Beielstein, Thomas
Published: (2025)
by: Bartz-Beielstein, Thomas
Published: (2025)
Asymptotic Optimality of Projected Inventory Level Policies for Lost Sales Inventory Systems with Large Leadtime and Penalty Cost
by: Moradi, Poulad, et al.
Published: (2025)
by: Moradi, Poulad, et al.
Published: (2025)
$\texttt{skwdro}$: a library for Wasserstein distributionally robust machine learning
by: Vincent, Florian, et al.
Published: (2024)
by: Vincent, Florian, et al.
Published: (2024)
Monotone and Conservative Policy Iteration Beyond the Tabular Case
by: Eshwar, S. R., et al.
Published: (2025)
by: Eshwar, S. R., et al.
Published: (2025)
A new $1/(1-ρ)$-scaling bound for multiserver queues via a leave-one-out technique
by: Hong, Yige
Published: (2025)
by: Hong, Yige
Published: (2025)
Amazon Locker Capacity Management
by: Sethuraman, Samyukta, et al.
Published: (2023)
by: Sethuraman, Samyukta, et al.
Published: (2023)
Universal Neural Optimal Transport
by: Geuter, Jonathan, et al.
Published: (2022)
by: Geuter, Jonathan, et al.
Published: (2022)
Modeling and analysis methods for early detection of leakage points in gas transmission systems
by: Aliyev, Ilgar
Published: (2025)
by: Aliyev, Ilgar
Published: (2025)
Multi-period Newsvendor Model
by: Khokhlov, Valentyn
Published: (2026)
by: Khokhlov, Valentyn
Published: (2026)
A Restless Bandit Model for Dynamic Ride Matching with Reneging Travelers
by: Fu, Jing, et al.
Published: (2021)
by: Fu, Jing, et al.
Published: (2021)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
by: Khanal, Kshitiz
Published: (2025)
by: Khanal, Kshitiz
Published: (2025)
Interior-Point Vanishing Problem in Semidefinite Relaxations for Neural Network Verification
by: Ueda, Ryota, et al.
Published: (2025)
by: Ueda, Ryota, et al.
Published: (2025)
Explainable and Class-Revealing Signal Feature Extraction via Scattering Transform and Constrained Zeroth-Order Optimization
by: Saito, Naoki, et al.
Published: (2025)
by: Saito, Naoki, et al.
Published: (2025)
Duality of causal distributionally robust optimization
by: Jiang, Yifan
Published: (2024)
by: Jiang, Yifan
Published: (2024)
Average-Cost MDPs with Infinite State and Action Sets: New Sufficient Conditions for Optimality Inequalities and Equations
by: Feinberg, Eugene A., et al.
Published: (2024)
by: Feinberg, Eugene A., et al.
Published: (2024)
Axis-Aligned Relaxations for Mixed-Integer Nonlinear Programming
by: Zhu, Haisheng, et al.
Published: (2026)
by: Zhu, Haisheng, et al.
Published: (2026)
Testing weak optimality of a given solution in interval linear programming revisited: NP-hardness proof, algorithm and some polynomial cases
by: Rada, Miroslav, et al.
Published: (2017)
by: Rada, Miroslav, et al.
Published: (2017)
Stochastic convergence of parallel asynchronous adaptive first-order methods
by: Gratton, Serge, et al.
Published: (2026)
by: Gratton, Serge, et al.
Published: (2026)
Robust Capacity Expansion Modelling for Renewable Energy Systems
by: Kebrich, Sebastian, et al.
Published: (2025)
by: Kebrich, Sebastian, et al.
Published: (2025)
Black-Box Uniform Stability for Non-Euclidean Empirical Risk Minimization
by: Vary, Simon, et al.
Published: (2024)
by: Vary, Simon, et al.
Published: (2024)
Achieving $\tilde{\mathcal{O}}(1/N)$ Optimality Gap in Restless Bandits through Gaussian Approximation
by: Yan, Chen, et al.
Published: (2024)
by: Yan, Chen, et al.
Published: (2024)
A random-key GRASP for combinatorial optimization
by: Chaves, Antonio A., et al.
Published: (2024)
by: Chaves, Antonio A., et al.
Published: (2024)
Random-key genetic algorithms: Principles and applications
by: Londe, Mariana A., et al.
Published: (2025)
by: Londe, Mariana A., et al.
Published: (2025)
A globally convergent SQP-type method with least constraint violation for nonlinear semidefinite programming
by: Fu, Wenhao, et al.
Published: (2023)
by: Fu, Wenhao, et al.
Published: (2023)
Comparative Evaluation of SDP, SOCP, and QC Convex Relaxations for Large-Scale Market-Based AC Optimal Power Flow
by: Keskin, Ata
Published: (2026)
by: Keskin, Ata
Published: (2026)
Zeroth-Order Methods for Nonconvex Stochastic Problems with Decision-Dependent Distributions
by: Hikima, Yuya, et al.
Published: (2024)
by: Hikima, Yuya, et al.
Published: (2024)
Similar Items
-
Achieving Exponential Asymptotic Optimality in Average-Reward Restless Bandits without Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2024) -
Restless Bandits with Average Reward: Breaking the Uniform Global Attractor Assumption
by: Hong, Yige, et al.
Published: (2023) -
Unichain and Aperiodicity are Sufficient for Asymptotic Optimality of Average-Reward Restless Bandits
by: Hong, Yige, et al.
Published: (2024) -
Semi-Supervised Clustering of Sparse Graphs: Crossing the Information-Theoretic Threshold
by: Sheng, Junda, et al.
Published: (2022) -
A note on weak compactness of occupation measures for an absorbing Markov decision process
by: Dufour, François, et al.
Published: (2024)