Distributionally Robust Off-Dynamics Reinforcement Learning: Provable Efficiency with Linear Function Approximation
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Zhishuai, Xu, Pan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning
por: Liu, Zhishuai, et al.
Publicado: (2024)
por: Liu, Zhishuai, et al.
Publicado: (2024)
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
por: He, Yiting, et al.
Publicado: (2025)
por: He, Yiting, et al.
Publicado: (2025)
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
por: Gu, Jingwen, et al.
Publicado: (2025)
por: Gu, Jingwen, et al.
Publicado: (2025)
Minimax Optimal and Computationally Efficient Algorithms for Distributionally Robust Offline Reinforcement Learning
por: Liu, Zhishuai, et al.
Publicado: (2024)
por: Liu, Zhishuai, et al.
Publicado: (2024)
Linear Mixture Distributionally Robust Markov Decision Processes
por: Liu, Zhishuai, et al.
Publicado: (2025)
por: Liu, Zhishuai, et al.
Publicado: (2025)
Robust Offline Reinforcement Learning with Linearly Structured f-Divergence Regularization
por: Tang, Cheng, et al.
Publicado: (2024)
por: Tang, Cheng, et al.
Publicado: (2024)
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning
por: Wang, Ruhan, et al.
Publicado: (2024)
por: Wang, Ruhan, et al.
Publicado: (2024)
Reinforcement Learning from Partial Observation: Linear Function Approximation with Provable Sample Efficiency
por: Cai, Qi, et al.
Publicado: (2022)
por: Cai, Qi, et al.
Publicado: (2022)
Provable Risk-Sensitive Distributional Reinforcement Learning with General Function Approximation
por: Chen, Yu, et al.
Publicado: (2024)
por: Chen, Yu, et al.
Publicado: (2024)
How to Provably Improve Return Conditioned Supervised Learning?
por: Liu, Zhishuai, et al.
Publicado: (2025)
por: Liu, Zhishuai, et al.
Publicado: (2025)
Provably Efficient Reinforcement Learning with Multinomial Logit Function Approximation
por: Li, Long-Fei, et al.
Publicado: (2024)
por: Li, Long-Fei, et al.
Publicado: (2024)
Provably Efficient Infinite-Horizon Average-Reward Reinforcement Learning with Linear Function Approximation
por: Chae, Woojin, et al.
Publicado: (2024)
por: Chae, Woojin, et al.
Publicado: (2024)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
por: Cho, Taehyun, et al.
Publicado: (2024)
por: Cho, Taehyun, et al.
Publicado: (2024)
Convergence of Distributionally Robust Q-Learning with Linear Function Approximation
por: Mandal, Saptarshi, et al.
Publicado: (2025)
por: Mandal, Saptarshi, et al.
Publicado: (2025)
Nonstationary Reinforcement Learning with Linear Function Approximation
por: Zhou, Huozhi, et al.
Publicado: (2020)
por: Zhou, Huozhi, et al.
Publicado: (2020)
Replicable Reinforcement Learning with Linear Function Approximation
por: Eaton, Eric, et al.
Publicado: (2025)
por: Eaton, Eric, et al.
Publicado: (2025)
Distributionally Robust Online Markov Game with Linear Function Approximation
por: Zheng, Zewu, et al.
Publicado: (2025)
por: Zheng, Zewu, et al.
Publicado: (2025)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
por: Guo, Yihong, et al.
Publicado: (2025)
por: Guo, Yihong, et al.
Publicado: (2025)
Communication-Efficient Federated Group Distributionally Robust Optimization
por: Guo, Zhishuai, et al.
Publicado: (2024)
por: Guo, Zhishuai, et al.
Publicado: (2024)
Reinforcement Learning with Function Approximation: From Linear to Nonlinear
por: Long, Jihao, et al.
Publicado: (2023)
por: Long, Jihao, et al.
Publicado: (2023)
Provably and Practically Efficient Adversarial Imitation Learning with General Function Approximation
por: Xu, Tian, et al.
Publicado: (2024)
por: Xu, Tian, et al.
Publicado: (2024)
Hybrid Transfer Reinforcement Learning: Provable Sample Efficiency from Shifted-Dynamics Data
por: Qu, Chengrui, et al.
Publicado: (2024)
por: Qu, Chengrui, et al.
Publicado: (2024)
Localized Dynamics-Aware Domain Adaption for Off-Dynamics Offline Reinforcement Learning
por: Xia, Zhangjie, et al.
Publicado: (2026)
por: Xia, Zhangjie, et al.
Publicado: (2026)
Online Robust Reinforcement Learning with General Function Approximation
por: Ghosh, Debamita, et al.
Publicado: (2025)
por: Ghosh, Debamita, et al.
Publicado: (2025)
Analysis of Off-Policy Multi-Step TD-Learning with Linear Function Approximation
por: Lee, Donghwan
Publicado: (2024)
por: Lee, Donghwan
Publicado: (2024)
Analysis of Off-Policy $n$-Step TD-Learning with Linear Function Approximation
por: Lim, Han-Dong, et al.
Publicado: (2025)
por: Lim, Han-Dong, et al.
Publicado: (2025)
Strategically Robust Multi-Agent Reinforcement Learning with Linear Function Approximation
por: Gonzales, Jake, et al.
Publicado: (2026)
por: Gonzales, Jake, et al.
Publicado: (2026)
Accelerated Distributional Temporal Difference Learning with Linear Function Approximation
por: Jin, Kaicheng, et al.
Publicado: (2025)
por: Jin, Kaicheng, et al.
Publicado: (2025)
Provably Adaptive Linear Approximation for the Shapley Value and Beyond
por: Li, Weida, et al.
Publicado: (2026)
por: Li, Weida, et al.
Publicado: (2026)
Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning
por: Yang, Yu, et al.
Publicado: (2026)
por: Yang, Yu, et al.
Publicado: (2026)
The Path Not Taken: RLVR Provably Learns Off the Principals
por: Zhu, Hanqing, et al.
Publicado: (2025)
por: Zhu, Hanqing, et al.
Publicado: (2025)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
por: Guo, Yihong, et al.
Publicado: (2024)
por: Guo, Yihong, et al.
Publicado: (2024)
Corruption-Robust Offline Reinforcement Learning with General Function Approximation
por: Ye, Chenlu, et al.
Publicado: (2023)
por: Ye, Chenlu, et al.
Publicado: (2023)
Provably Efficient RL under Episode-Wise Safety in Constrained MDPs with Linear Function Approximation
por: Kitamura, Toshinori, et al.
Publicado: (2025)
por: Kitamura, Toshinori, et al.
Publicado: (2025)
Provably Sample-Efficient Robust Reinforcement Learning with Average Reward
por: Roch, Zachary, et al.
Publicado: (2025)
por: Roch, Zachary, et al.
Publicado: (2025)
On the Statistical Efficiency of Mean-Field Reinforcement Learning with General Function Approximation
por: Huang, Jiawei, et al.
Publicado: (2023)
por: Huang, Jiawei, et al.
Publicado: (2023)
Provably Robust Federated Reinforcement Learning
por: Fang, Minghong, et al.
Publicado: (2025)
por: Fang, Minghong, et al.
Publicado: (2025)
A Finite Sample Analysis of Distributional TD Learning with Linear Function Approximation
por: Peng, Yang, et al.
Publicado: (2025)
por: Peng, Yang, et al.
Publicado: (2025)
ODRL: A Benchmark for Off-Dynamics Reinforcement Learning
por: Lyu, Jiafei, et al.
Publicado: (2024)
por: Lyu, Jiafei, et al.
Publicado: (2024)
Gap-Dependent Bounds for Nearly Minimax Optimal Reinforcement Learning with Linear Function Approximation
por: Zhang, Haochen, et al.
Publicado: (2026)
por: Zhang, Haochen, et al.
Publicado: (2026)
Ejemplares similares
-
Upper and Lower Bounds for Distributionally Robust Off-Dynamics Reinforcement Learning
por: Liu, Zhishuai, et al.
Publicado: (2024) -
Sample Complexity of Distributionally Robust Off-Dynamics Reinforcement Learning with Online Interaction
por: He, Yiting, et al.
Publicado: (2025) -
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
por: Gu, Jingwen, et al.
Publicado: (2025) -
Minimax Optimal and Computationally Efficient Algorithms for Distributionally Robust Offline Reinforcement Learning
por: Liu, Zhishuai, et al.
Publicado: (2024) -
Linear Mixture Distributionally Robust Markov Decision Processes
por: Liu, Zhishuai, et al.
Publicado: (2025)