Learning Optimal Distributionally Robust Stochastic Control in Continuous State Spaces
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Shengbo, Meng, Jason, Si, Nian, Blanchet, Jose, Zhou, Zhengyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
A Finite Sample Complexity Bound for Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning
von: Li, Zhenghao, et al.
Veröffentlicht: (2025)
von: Li, Zhenghao, et al.
Veröffentlicht: (2025)
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
Bellman Optimality of Average-Reward Robust Markov Decision Processes with a Constant Gain
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Sample Complexity of Distributionally Robust Average-Reward Reinforcement Learning
von: Chen, Zijun, et al.
Veröffentlicht: (2025)
von: Chen, Zijun, et al.
Veröffentlicht: (2025)
Single-Trajectory Distributionally Robust Reinforcement Learning
von: Liang, Zhipeng, et al.
Veröffentlicht: (2023)
von: Liang, Zhipeng, et al.
Veröffentlicht: (2023)
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Learning an Optimal Assortment Policy under Observational Data
von: Han, Yuxuan, et al.
Veröffentlicht: (2025)
von: Han, Yuxuan, et al.
Veröffentlicht: (2025)
Achieving $\varepsilon^{-2}$ Dependence for Average-Reward Q-Learning with a New Contraction Principle
von: Chen, Zijun, et al.
Veröffentlicht: (2026)
von: Chen, Zijun, et al.
Veröffentlicht: (2026)
Q-Measure-Learning for Continuous State RL: Efficient Implementation and Convergence
von: Wang, Shengbo
Veröffentlicht: (2026)
von: Wang, Shengbo
Veröffentlicht: (2026)
Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithm
von: Lu, Miao, et al.
Veröffentlicht: (2024)
von: Lu, Miao, et al.
Veröffentlicht: (2024)
Distributionally Robust Policy Learning under Concept Drifts
von: Wang, Jingyuan, et al.
Veröffentlicht: (2024)
von: Wang, Jingyuan, et al.
Veröffentlicht: (2024)
Unifying Distributionally Robust Optimization via Optimal Transport Theory
von: Blanchet, Jose, et al.
Veröffentlicht: (2023)
von: Blanchet, Jose, et al.
Veröffentlicht: (2023)
Robust Assortment Optimization from Observational Data
von: Lu, Miao, et al.
Veröffentlicht: (2026)
von: Lu, Miao, et al.
Veröffentlicht: (2026)
Representation-Aware Distributionally Robust Optimization: A Knowledge Transfer Framework
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
Knowledge-Guided Wasserstein Distributionally Robust Optimization
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
von: Wang, Zitao, et al.
Veröffentlicht: (2025)
Nonconvex Stochastic Optimization under Heavy-Tailed Noises: Optimal Convergence without Gradient Clipping
von: Liu, Zijian, et al.
Veröffentlicht: (2024)
von: Liu, Zijian, et al.
Veröffentlicht: (2024)
Wasserstein Distributionally Robust Regret Optimization
von: Fiechtner, Lukas-Benedikt, et al.
Veröffentlicht: (2025)
von: Fiechtner, Lukas-Benedikt, et al.
Veröffentlicht: (2025)
ScoreFusion: Fusing Score-based Generative Models via Kullback-Leibler Barycenters
von: Liu, Hao, et al.
Veröffentlicht: (2024)
von: Liu, Hao, et al.
Veröffentlicht: (2024)
Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback
von: Wang, Yikai, et al.
Veröffentlicht: (2026)
von: Wang, Yikai, et al.
Veröffentlicht: (2026)
Central Limit Theorem for Two-Time-Scale Approximate Distributionally Robust RL
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)
DRO: A Python Library for Distributionally Robust Optimization in Machine Learning
von: Liu, Jiashuo, et al.
Veröffentlicht: (2025)
von: Liu, Jiashuo, et al.
Veröffentlicht: (2025)
Duality and Policy Evaluation in Distributionally Robust Bayesian Diffusion Control
von: Blanchet, Jose, et al.
Veröffentlicht: (2025)
von: Blanchet, Jose, et al.
Veröffentlicht: (2025)
Revisiting the Last-Iterate Convergence of Stochastic Gradient Methods
von: Liu, Zijian, et al.
Veröffentlicht: (2023)
von: Liu, Zijian, et al.
Veröffentlicht: (2023)
Stochastic Optimal Control for Diffusion Bridges in Function Spaces
von: Park, Byoungwoo, et al.
Veröffentlicht: (2024)
von: Park, Byoungwoo, et al.
Veröffentlicht: (2024)
Deep Learning for Markov Chains: Lyapunov Functions, Poisson's Equation, and Stationary Distributions
von: Qu, Yanlin, et al.
Veröffentlicht: (2025)
von: Qu, Yanlin, et al.
Veröffentlicht: (2025)
Limit Theorems for Stochastic Gradient Descent with Infinite Variance
von: Blanchet, Jose, et al.
Veröffentlicht: (2024)
von: Blanchet, Jose, et al.
Veröffentlicht: (2024)
Optimal Complexity in Byzantine-Robust Distributed Stochastic Optimization with Data Heterogeneity
von: Shi, Qiankun, et al.
Veröffentlicht: (2025)
von: Shi, Qiankun, et al.
Veröffentlicht: (2025)
Tackling Interference Induced by Data Training Loops in A/B Tests: A Weighted Training Approach
von: Si, Nian
Veröffentlicht: (2023)
von: Si, Nian
Veröffentlicht: (2023)
Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
von: Huang, Jingkai, et al.
Veröffentlicht: (2026)
von: Huang, Jingkai, et al.
Veröffentlicht: (2026)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
Optimal Multi-Distribution Learning
von: Zhang, Zihan, et al.
Veröffentlicht: (2023)
von: Zhang, Zihan, et al.
Veröffentlicht: (2023)
Review on Determining the Number of Communities in Network Data
von: Du, Zhengyuan, et al.
Veröffentlicht: (2025)
von: Du, Zhengyuan, et al.
Veröffentlicht: (2025)
Optimal Arm Elimination Algorithms for Combinatorial Bandits
von: Wen, Yuxiao, et al.
Veröffentlicht: (2025)
von: Wen, Yuxiao, et al.
Veröffentlicht: (2025)
Optimal Rates for Robust Stochastic Convex Optimization
von: Gao, Changyu, et al.
Veröffentlicht: (2024)
von: Gao, Changyu, et al.
Veröffentlicht: (2024)
Scalable and Precise Patch Robustness Certification for Deep Learning Models with Top-k Predictions
von: Zhou, Qilin, et al.
Veröffentlicht: (2025)
von: Zhou, Qilin, et al.
Veröffentlicht: (2025)
Distributionally Robust Model-based Reinforcement Learning with Large State Spaces
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2023)
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
On the Foundation of Distributionally Robust Reinforcement Learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
A Finite Sample Complexity Bound for Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
Sample Complexity of Variance-reduced Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023) -
Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning
von: Li, Zhenghao, et al.
Veröffentlicht: (2025) -
Non-Rectangular Average-Reward Robust MDPs: Optimal Policies and Their Transient Values
von: Wang, Shengbo, et al.
Veröffentlicht: (2026)