Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zhenghao, Wang, Shengbo, Si, Nian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
von: Chen, Zijun, et al.
Veröffentlicht: (2025)
von: Chen, Zijun, et al.
Veröffentlicht: (2025)
A Finite Sample Complexity Bound for Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
Bellman Optimality of Average-Reward Robust Markov Decision Processes with a Constant Gain
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Learning Optimal Distributionally Robust Stochastic Control in Continuous State Spaces
von: Wang, Shengbo, et al.
Veröffentlicht: (2024)
von: Wang, Shengbo, et al.
Veröffentlicht: (2024)
Distributionally Robust Model-Based Offline Reinforcement Learning with Near-Optimal Sample Complexity
von: Shi, Laixi, et al.
Veröffentlicht: (2022)
von: Shi, Laixi, et al.
Veröffentlicht: (2022)
ORVIT: Near-Optimal Online Distributionally Robust Reinforcement Learning
von: Ghosh, Debamita, et al.
Veröffentlicht: (2025)
von: Ghosh, Debamita, et al.
Veröffentlicht: (2025)
Achieving $\varepsilon^{-2}$ Dependence for Average-Reward Q-Learning with a New Contraction Principle
von: Chen, Zijun, et al.
Veröffentlicht: (2026)
von: Chen, Zijun, et al.
Veröffentlicht: (2026)
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Near-Optimal Sample Complexity in Reward-Free Kernel-Based Reinforcement Learning
von: Kayal, Aya, et al.
Veröffentlicht: (2025)
von: Kayal, Aya, et al.
Veröffentlicht: (2025)
Nearly Optimal Sample Complexity for Learning with Label Proportions
von: Busa-Fekete, Robert, et al.
Veröffentlicht: (2025)
von: Busa-Fekete, Robert, et al.
Veröffentlicht: (2025)
Near Optimal Non-asymptotic Sample Complexity of 1-Identification
von: Li, Zitian, et al.
Veröffentlicht: (2025)
von: Li, Zitian, et al.
Veröffentlicht: (2025)
Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model
von: Deng, Zilong, et al.
Veröffentlicht: (2025)
von: Deng, Zilong, et al.
Veröffentlicht: (2025)
Representation-Aware Distributionally Robust Optimization: A Knowledge Transfer Framework
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
Private Realizable-to-Agnostic Transformation with Near-Optimal Sample Complexity
von: Li, Bo, et al.
Veröffentlicht: (2025)
von: Li, Bo, et al.
Veröffentlicht: (2025)
Knowledge-Guided Wasserstein Distributionally Robust Optimization
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
Near-Optimal Sample Complexity for Online Constrained MDPs
von: Liu, Chang, et al.
Veröffentlicht: (2026)
von: Liu, Chang, et al.
Veröffentlicht: (2026)
Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithm
von: Lu, Miao, et al.
Veröffentlicht: (2024)
von: Lu, Miao, et al.
Veröffentlicht: (2024)
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
von: He, Yiting, et al.
Veröffentlicht: (2025)
von: He, Yiting, et al.
Veröffentlicht: (2025)
Near-Minimax-Optimal Distributional Reinforcement Learning with a Generative Model
von: Rowland, Mark, et al.
Veröffentlicht: (2024)
von: Rowland, Mark, et al.
Veröffentlicht: (2024)
Near-Optimality of Contrastive Divergence Algorithms
von: Glaser, Pierre, et al.
Veröffentlicht: (2025)
von: Glaser, Pierre, et al.
Veröffentlicht: (2025)
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
Central Limit Theorem for Two-Time-Scale Approximate Distributionally Robust RL
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
Near-Optimal Sample Complexity Bounds for Constrained Average-Reward MDPs
von: Wei, Yukuan, et al.
Veröffentlicht: (2025)
von: Wei, Yukuan, et al.
Veröffentlicht: (2025)
Multivariate Distributional Reinforcement Learning Using Sliced Divergences
von: Debes, Baptiste, et al.
Veröffentlicht: (2026)
von: Debes, Baptiste, et al.
Veröffentlicht: (2026)
Statistical Guarantees for Distributionally Robust Optimization with Optimal Transport and OT-Regularized Divergences
von: Birrell, Jeremiah, et al.
Veröffentlicht: (2026)
von: Birrell, Jeremiah, et al.
Veröffentlicht: (2026)
Achieving the Asymptotically Optimal Sample Complexity of Offline Reinforcement Learning: A DRO-Based Approach
von: Wang, Yue, et al.
Veröffentlicht: (2023)
von: Wang, Yue, et al.
Veröffentlicht: (2023)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
von: Provodin, Danil, et al.
Veröffentlicht: (2024)
von: Provodin, Danil, et al.
Veröffentlicht: (2024)
Towards Minimax Optimality of Model-based Robust Reinforcement Learning
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
von: Clavier, Pierre, et al.
Veröffentlicht: (2023)
Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2026)
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2026)
Minimax-Optimal Multi-Agent Robust Reinforcement Learning
von: Jiao, Yuchen, et al.
Veröffentlicht: (2024)
von: Jiao, Yuchen, et al.
Veröffentlicht: (2024)
Projection by Convolution: Optimal Sample Complexity for Reinforcement Learning in Continuous-Space MDPs
von: Maran, Davide, et al.
Veröffentlicht: (2024)
von: Maran, Davide, et al.
Veröffentlicht: (2024)
The Optimal Sample Complexity of Multiclass and List Learning
von: Pabbaraju, Chirag
Veröffentlicht: (2026)
von: Pabbaraju, Chirag
Veröffentlicht: (2026)
Near-Optimal Algorithms for Group Distributionally Robust Optimization and Beyond
von: Soma, Tasuku, et al.
Veröffentlicht: (2022)
von: Soma, Tasuku, et al.
Veröffentlicht: (2022)
Optimal Transport Regularized Divergences: Application to Adversarial Robustness
von: Birrell, Jeremiah, et al.
Veröffentlicht: (2023)
von: Birrell, Jeremiah, et al.
Veröffentlicht: (2023)
Tackling Interference Induced by Data Training Loops in A/B Tests: A Weighted Training Approach
von: Si, Nian
Veröffentlicht: (2023)
von: Si, Nian
Veröffentlicht: (2023)
Sample-Efficient Distributionally Robust Multi-Agent Reinforcement Learning via Online Interaction
von: Farhat, Zain Ulabedeen, et al.
Veröffentlicht: (2025)
von: Farhat, Zain Ulabedeen, et al.
Veröffentlicht: (2025)
Q-Measure-Learning for Continuous State RL: Efficient Implementation and Convergence
von: Wang, Shengbo
Veröffentlicht: (2026)
von: Wang, Shengbo
Veröffentlicht: (2026)
Ähnliche Einträge
-
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
von: Chen, Zijun, et al.
Veröffentlicht: (2025) -
A Finite Sample Complexity Bound for Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)