A Non-Monolithic Policy Approach of Offline-to-Online Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, JaeYoon, Xuan, Junyu, Liang, Christy, Hussain, Farookh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Autonomous Non-monolithic Agent with Multi-mode Exploration based on Options Framework
von: Kim, JaeYoon, et al.
Veröffentlicht: (2023)
von: Kim, JaeYoon, et al.
Veröffentlicht: (2023)
Decoupling Exploration and Exploitation for Unsupervised Pre-training with Successor Features
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024)
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024)
Online Pre-Training for Offline-to-Online Reinforcement Learning
von: Shin, Yongjae, et al.
Veröffentlicht: (2025)
von: Shin, Yongjae, et al.
Veröffentlicht: (2025)
A Behavior-Aware Approach for Deep Reinforcement Learning in Non-stationary Environments without Known Change Points
von: Liu, Zihe, et al.
Veröffentlicht: (2024)
von: Liu, Zihe, et al.
Veröffentlicht: (2024)
Latent Safety-Constrained Policy Approach for Safe Offline Reinforcement Learning
von: Koirala, Prajwal, et al.
Veröffentlicht: (2024)
von: Koirala, Prajwal, et al.
Veröffentlicht: (2024)
Robust Offline Reinforcement Learning for Non-Markovian Decision Processes
von: Huang, Ruiquan, et al.
Veröffentlicht: (2024)
von: Huang, Ruiquan, et al.
Veröffentlicht: (2024)
Online Policy Learning from Offline Preferences
von: Zhang, Guoxi, et al.
Veröffentlicht: (2024)
von: Zhang, Guoxi, et al.
Veröffentlicht: (2024)
Information-Directed Offline-to-Online Reinforcement Learning
von: Chen, Keru
Veröffentlicht: (2026)
von: Chen, Keru
Veröffentlicht: (2026)
MOORL: A Framework for Integrating Offline-Online Reinforcement Learning
von: Chaudhary, Gaurav, et al.
Veröffentlicht: (2025)
von: Chaudhary, Gaurav, et al.
Veröffentlicht: (2025)
Bayesian Design Principles for Offline-to-Online Reinforcement Learning
von: Hu, Hao, et al.
Veröffentlicht: (2024)
von: Hu, Hao, et al.
Veröffentlicht: (2024)
B3C: A Minimalist Approach to Offline Multi-Agent Reinforcement Learning
von: Kim, Woojun, et al.
Veröffentlicht: (2025)
von: Kim, Woojun, et al.
Veröffentlicht: (2025)
Compositional Conservatism: A Transductive Approach in Offline Reinforcement Learning
von: Song, Yeda, et al.
Veröffentlicht: (2024)
von: Song, Yeda, et al.
Veröffentlicht: (2024)
Online Optimization for Offline Safe Reinforcement Learning
von: Chemingui, Yassine, et al.
Veröffentlicht: (2025)
von: Chemingui, Yassine, et al.
Veröffentlicht: (2025)
The Three Regimes of Offline-to-Online Reinforcement Learning
von: Li, Lu, et al.
Veröffentlicht: (2025)
von: Li, Lu, et al.
Veröffentlicht: (2025)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
von: Wang, Qi, et al.
Veröffentlicht: (2023)
von: Wang, Qi, et al.
Veröffentlicht: (2023)
Active Reinforcement Learning Strategies for Offline Policy Improvement
von: Dukkipati, Ambedkar, et al.
Veröffentlicht: (2024)
von: Dukkipati, Ambedkar, et al.
Veröffentlicht: (2024)
Hypercube Policy Regularization Framework for Offline Reinforcement Learning
von: Shen, Yi, et al.
Veröffentlicht: (2024)
von: Shen, Yi, et al.
Veröffentlicht: (2024)
Hierarchical Subspaces of Policies for Continual Offline Reinforcement Learning
von: Kobanda, Anthony, et al.
Veröffentlicht: (2024)
von: Kobanda, Anthony, et al.
Veröffentlicht: (2024)
Mildly Constrained Evaluation Policy for Offline Reinforcement Learning
von: Xu, Linjie, et al.
Veröffentlicht: (2023)
von: Xu, Linjie, et al.
Veröffentlicht: (2023)
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026)
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026)
Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning
von: Jing, Tan, et al.
Veröffentlicht: (2025)
von: Jing, Tan, et al.
Veröffentlicht: (2025)
Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning
von: Gireesh, Nandiraju, et al.
Veröffentlicht: (2026)
von: Gireesh, Nandiraju, et al.
Veröffentlicht: (2026)
Offline Reinforcement Learning with Generative Trajectory Policies
von: Feng, Xinsong, et al.
Veröffentlicht: (2025)
von: Feng, Xinsong, et al.
Veröffentlicht: (2025)
Robot Policy Transfer with Online Demonstrations: An Active Reinforcement Learning Approach
von: Hou, Muhan, et al.
Veröffentlicht: (2025)
von: Hou, Muhan, et al.
Veröffentlicht: (2025)
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
von: Liu, Xu-Hui, et al.
Veröffentlicht: (2024)
von: Liu, Xu-Hui, et al.
Veröffentlicht: (2024)
Planning to Go Out-of-Distribution in Offline-to-Online Reinforcement Learning
von: McInroe, Trevor, et al.
Veröffentlicht: (2023)
von: McInroe, Trevor, et al.
Veröffentlicht: (2023)
Active Advantage-Aligned Online Reinforcement Learning with Offline Data
von: Liu, Xuefeng, et al.
Veröffentlicht: (2025)
von: Liu, Xuefeng, et al.
Veröffentlicht: (2025)
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
von: Bozkurt, Alper Kamil, et al.
Veröffentlicht: (2026)
von: Bozkurt, Alper Kamil, et al.
Veröffentlicht: (2026)
Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
von: Lei, Kun, et al.
Veröffentlicht: (2023)
von: Lei, Kun, et al.
Veröffentlicht: (2023)
Offline-Online Reinforcement Learning for Linear Mixture MDPs
von: Zhang, Zhongjun, et al.
Veröffentlicht: (2026)
von: Zhang, Zhongjun, et al.
Veröffentlicht: (2026)
Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning
von: Song, Chihyeon, et al.
Veröffentlicht: (2025)
von: Song, Chihyeon, et al.
Veröffentlicht: (2025)
Discrete Flow Matching for Offline-to-Online Reinforcement Learning
von: Khan, Fairoz Nower, et al.
Veröffentlicht: (2026)
von: Khan, Fairoz Nower, et al.
Veröffentlicht: (2026)
Offline Trajectory Optimization for Offline Reinforcement Learning
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
Residual Learning and Context Encoding for Adaptive Offline-to-Online Reinforcement Learning
von: Nakhaei, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Nakhaei, Mohammadreza, et al.
Veröffentlicht: (2024)
Enhancing Online Reinforcement Learning with Meta-Learned Objective from Offline Data
von: Deng, Shilong, et al.
Veröffentlicht: (2025)
von: Deng, Shilong, et al.
Veröffentlicht: (2025)
Towards Robust Offline-to-Online Reinforcement Learning via Uncertainty and Smoothness
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2023)
von: Wen, Xiaoyu, et al.
Veröffentlicht: (2023)
Efficient and Uncertainty-Aware Diffusion Framework for Offline-to-Online Reinforcement Learning
von: Bui, Ha Manh, et al.
Veröffentlicht: (2026)
von: Bui, Ha Manh, et al.
Veröffentlicht: (2026)
Policy-regularized Offline Multi-objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024)
von: Lin, Qian, et al.
Veröffentlicht: (2024)
Policy-Based Trajectory Clustering in Offline Reinforcement Learning
von: Hu, Hao, et al.
Veröffentlicht: (2025)
von: Hu, Hao, et al.
Veröffentlicht: (2025)
Evaluation-Time Policy Switching for Offline Reinforcement Learning
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
von: Neggatu, Natinael Solomon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
An Autonomous Non-monolithic Agent with Multi-mode Exploration based on Options Framework
von: Kim, JaeYoon, et al.
Veröffentlicht: (2023) -
Decoupling Exploration and Exploitation for Unsupervised Pre-training with Successor Features
von: Kim, JaeYoon, et al.
Veröffentlicht: (2024) -
Online Pre-Training for Offline-to-Online Reinforcement Learning
von: Shin, Yongjae, et al.
Veröffentlicht: (2025) -
A Behavior-Aware Approach for Deep Reinforcement Learning in Non-stationary Environments without Known Change Points
von: Liu, Zihe, et al.
Veröffentlicht: (2024) -
Latent Safety-Constrained Policy Approach for Safe Offline Reinforcement Learning
von: Koirala, Prajwal, et al.
Veröffentlicht: (2024)