Provable Domain Adaptation for Offline Reinforcement Learning with Limited Samples
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Weiqin, Zhang, Xinjie, Mishra, Sandipan, Paternain, Santiago |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Filtering Learning Histories Enhances In-Context Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2025)
von: Chen, Weiqin, et al.
Veröffentlicht: (2025)
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
Adaptive Primal-Dual Method for Safe Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
Transfer Learning for a Class of Cascade Dynamical Systems
von: Rabiei, Shima, et al.
Veröffentlicht: (2024)
von: Rabiei, Shima, et al.
Veröffentlicht: (2024)
A General Control-Theoretic Approach for Reinforcement Learning: Theory and Algorithms
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
von: Chen, Weiqin, et al.
Veröffentlicht: (2024)
ConstrainedSQL: Training LLMs for Text2SQL via Constrained Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2025)
von: Chen, Weiqin, et al.
Veröffentlicht: (2025)
Safety Guarantees in Zero-Shot Reinforcement Learning for Cascade Dynamical Systems
von: Rabiei, Shima, et al.
Veröffentlicht: (2026)
von: Rabiei, Shima, et al.
Veröffentlicht: (2026)
Provable Zero-Shot Generalization in Offline Reinforcement Learning
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
Provably Efficient Offline-to-Online Value Adaptation with General Function Approximation
von: Li, Shangzhe, et al.
Veröffentlicht: (2026)
von: Li, Shangzhe, et al.
Veröffentlicht: (2026)
Tensor and Matrix Low-Rank Value-Function Approximation in Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2022)
von: Rozada, Sergio, et al.
Veröffentlicht: (2022)
From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning
von: Zu, Lipeng, et al.
Veröffentlicht: (2025)
von: Zu, Lipeng, et al.
Veröffentlicht: (2025)
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
von: Liu, Xu-Hui, et al.
Veröffentlicht: (2024)
von: Liu, Xu-Hui, et al.
Veröffentlicht: (2024)
Provable Offline Reinforcement Learning for Structured Cyclic MDPs
von: Lee, Kyungbok, et al.
Veröffentlicht: (2026)
von: Lee, Kyungbok, et al.
Veröffentlicht: (2026)
Provably Sample-Efficient Robust Reinforcement Learning with Average Reward
von: Roch, Zachary, et al.
Veröffentlicht: (2025)
von: Roch, Zachary, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning with Domain-Unlabeled Data
von: Nishimori, Soichiro, et al.
Veröffentlicht: (2024)
von: Nishimori, Soichiro, et al.
Veröffentlicht: (2024)
A Tensor Low-Rank Approximation for Value Functions in Multi-Task Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
Sample-Efficient Policy Constraint Offline Deep Reinforcement Learning based on Sample Filtering
von: Chen, Yuanhao, et al.
Veröffentlicht: (2025)
von: Chen, Yuanhao, et al.
Veröffentlicht: (2025)
Imagination-Limited Q-Learning for Offline Reinforcement Learning
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
von: Liu, Wenhui, et al.
Veröffentlicht: (2025)
Diverse Randomized Value Functions: A Provably Pessimistic Approach for Offline Reinforcement Learning
von: Yu, Xudong, et al.
Veröffentlicht: (2024)
von: Yu, Xudong, et al.
Veröffentlicht: (2024)
Diffusion-DICE: In-Sample Diffusion Guidance for Offline Reinforcement Learning
von: Mao, Liyuan, et al.
Veröffentlicht: (2024)
von: Mao, Liyuan, et al.
Veröffentlicht: (2024)
Sample-Efficient Tabular Self-Play for Offline Robust Reinforcement Learning
von: Li, Na, et al.
Veröffentlicht: (2025)
von: Li, Na, et al.
Veröffentlicht: (2025)
Autonomous helicopter aerial refueling: controller design and performance guarantees
von: Jayarathne, Damsara, et al.
Veröffentlicht: (2025)
von: Jayarathne, Damsara, et al.
Veröffentlicht: (2025)
An Offline Adaptation Framework for Constrained Multi-Objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024)
von: Lin, Qian, et al.
Veröffentlicht: (2024)
Sample Efficient Active Algorithms for Offline Reinforcement Learning
von: Roy, Soumyadeep, et al.
Veröffentlicht: (2026)
von: Roy, Soumyadeep, et al.
Veröffentlicht: (2026)
Reinforced Domain Selection for Continuous Domain Adaptation
von: Liu, Hanbing, et al.
Veröffentlicht: (2025)
von: Liu, Hanbing, et al.
Veröffentlicht: (2025)
Contrastive Representation for Data Filtering in Cross-Domain Offline Reinforcement Learning
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2024)
Towards Provable Emergence of In-Context Reinforcement Learning
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
Provable Meta-Learning with Low-Rank Adaptations
von: Block, Jacob L., et al.
Veröffentlicht: (2024)
von: Block, Jacob L., et al.
Veröffentlicht: (2024)
Policy-Driven World Model Adaptation for Robust Offline Model-based Reinforcement Learning
von: Chen, Jiayu, et al.
Veröffentlicht: (2025)
von: Chen, Jiayu, et al.
Veröffentlicht: (2025)
Information-Directed Offline-to-Online Reinforcement Learning
von: Chen, Keru
Veröffentlicht: (2026)
von: Chen, Keru
Veröffentlicht: (2026)
Offline Trajectory Optimization for Offline Reinforcement Learning
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
On Sample-Efficient Offline Reinforcement Learning: Data Diversity, Posterior Sampling, and Beyond
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen-Tang, Thanh, et al.
Veröffentlicht: (2024)
Dual-Robust Cross-Domain Offline Reinforcement Learning Against Dynamics Shifts
von: Qiao, Zhongjian, et al.
Veröffentlicht: (2025)
von: Qiao, Zhongjian, et al.
Veröffentlicht: (2025)
When is Offline Policy Selection Sample Efficient for Reinforcement Learning?
von: Liu, Vincent, et al.
Veröffentlicht: (2023)
von: Liu, Vincent, et al.
Veröffentlicht: (2023)
Hybrid Transfer Reinforcement Learning: Provable Sample Efficiency from Shifted-Dynamics Data
von: Qu, Chengrui, et al.
Veröffentlicht: (2024)
von: Qu, Chengrui, et al.
Veröffentlicht: (2024)
Federated Offline Reinforcement Learning
von: Zhou, Doudou, et al.
Veröffentlicht: (2022)
von: Zhou, Doudou, et al.
Veröffentlicht: (2022)
Offline Reinforcement Learning and Sequence Modeling for Downlink Link Adaptation
von: Peri, Samuele, et al.
Veröffentlicht: (2024)
von: Peri, Samuele, et al.
Veröffentlicht: (2024)
Provable Risk-Sensitive Distributional Reinforcement Learning with General Function Approximation
von: Chen, Yu, et al.
Veröffentlicht: (2024)
von: Chen, Yu, et al.
Veröffentlicht: (2024)
Cross-Domain Offline Policy Adaptation via Selective Transition Correction
von: Yan, Mengbei, et al.
Veröffentlicht: (2026)
von: Yan, Mengbei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Filtering Learning Histories Enhances In-Context Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2025) -
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons
von: Chen, Weiqin, et al.
Veröffentlicht: (2024) -
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023) -
Adaptive Primal-Dual Method for Safe Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2024) -
Transfer Learning for a Class of Cascade Dynamical Systems
von: Rabiei, Shima, et al.
Veröffentlicht: (2024)