Hybrid Reinforcement Learning from Offline Observation Alone
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Yuda, Bagnell, J. Andrew, Singh, Aarti |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
To Distill or Decide? Understanding the Algorithmic Trade-off in Partially Observable Reinforcement Learning
by: Song, Yuda, et al.
Published: (2025)
by: Song, Yuda, et al.
Published: (2025)
Expanding the Capabilities of Reinforcement Learning via Text Feedback
by: Song, Yuda, et al.
Published: (2026)
by: Song, Yuda, et al.
Published: (2026)
The Importance of Online Data: Understanding Preference Fine-tuning via Coverage
by: Song, Yuda, et al.
Published: (2024)
by: Song, Yuda, et al.
Published: (2024)
Hybrid Inverse Reinforcement Learning
by: Ren, Juntao, et al.
Published: (2024)
by: Ren, Juntao, et al.
Published: (2024)
Inverse Reinforcement Learning without Reinforcement Learning
by: Swamy, Gokul, et al.
Published: (2023)
by: Swamy, Gokul, et al.
Published: (2023)
Rich-Observation Reinforcement Learning with Continuous Latent Dynamics
by: Song, Yuda, et al.
Published: (2024)
by: Song, Yuda, et al.
Published: (2024)
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
by: Swamy, Gokul, et al.
Published: (2025)
by: Swamy, Gokul, et al.
Published: (2025)
Offline Reinforcement Learning with Imputed Rewards
by: Romeo, Carlo, et al.
Published: (2024)
by: Romeo, Carlo, et al.
Published: (2024)
Mildly Constrained Evaluation Policy for Offline Reinforcement Learning
by: Xu, Linjie, et al.
Published: (2023)
by: Xu, Linjie, et al.
Published: (2023)
Enhancing Offline Reinforcement Learning with Curriculum Learning-Based Trajectory Valuation
by: Abolfazli, Amir, et al.
Published: (2025)
by: Abolfazli, Amir, et al.
Published: (2025)
Offline Reinforcement Learning for Rotation Profile Control in Tokamaks
by: Sonker, Rohit, et al.
Published: (2026)
by: Sonker, Rohit, et al.
Published: (2026)
Adaptive Replay Buffer for Offline-to-Online Reinforcement Learning
by: Song, Chihyeon, et al.
Published: (2025)
by: Song, Chihyeon, et al.
Published: (2025)
Offline Trajectory Optimization for Offline Reinforcement Learning
by: Zhao, Ziqi, et al.
Published: (2024)
by: Zhao, Ziqi, et al.
Published: (2024)
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Hybrid Adaptive Conformal Offline Reinforcement Learning for Fair Population Health Management
by: Basu, Sanjay, et al.
Published: (2025)
by: Basu, Sanjay, et al.
Published: (2025)
Equivariant Offline Reinforcement Learning
by: Tangri, Arsh, et al.
Published: (2024)
by: Tangri, Arsh, et al.
Published: (2024)
Federated Offline Reinforcement Learning
by: Zhou, Doudou, et al.
Published: (2022)
by: Zhou, Doudou, et al.
Published: (2022)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
by: Xu, Qiushui, et al.
Published: (2025)
by: Xu, Qiushui, et al.
Published: (2025)
Epistemic Robust Offline Reinforcement Learning
by: Chenreddy, Abhilash Reddy, et al.
Published: (2026)
by: Chenreddy, Abhilash Reddy, et al.
Published: (2026)
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
by: Sestini, Alessandro, et al.
Published: (2025)
by: Sestini, Alessandro, et al.
Published: (2025)
Offline Multitask Representation Learning for Reinforcement Learning
by: Ishfaq, Haque, et al.
Published: (2024)
by: Ishfaq, Haque, et al.
Published: (2024)
A Benchmark Environment for Offline Reinforcement Learning in Racing Games
by: Macaluso, Girolamo, et al.
Published: (2024)
by: Macaluso, Girolamo, et al.
Published: (2024)
When is Offline Policy Selection Sample Efficient for Reinforcement Learning?
by: Liu, Vincent, et al.
Published: (2023)
by: Liu, Vincent, et al.
Published: (2023)
Unveiling Markov Heads in Pretrained Language Models for Offline Reinforcement Learning
by: Zhao, Wenhao, et al.
Published: (2024)
by: Zhao, Wenhao, et al.
Published: (2024)
Cooperative Multi-agent RL with Communication Constraints
by: Xiong, Nuoya, et al.
Published: (2026)
by: Xiong, Nuoya, et al.
Published: (2026)
Projection Optimization: A General Framework for Multi-Objective and Multi-Group RLHF
by: Xiong, Nuoya, et al.
Published: (2025)
by: Xiong, Nuoya, et al.
Published: (2025)
Maximum Likelihood Reinforcement Learning
by: Tajwar, Fahim, et al.
Published: (2026)
by: Tajwar, Fahim, et al.
Published: (2026)
Robust Offline Policy Learning with Observational Data from Multiple Sources
by: Carranza, Aldo Gael, et al.
Published: (2024)
by: Carranza, Aldo Gael, et al.
Published: (2024)
Compositional Conservatism: A Transductive Approach in Offline Reinforcement Learning
by: Song, Yeda, et al.
Published: (2024)
by: Song, Yeda, et al.
Published: (2024)
Temporal Abstraction in Reinforcement Learning with Offline Data
by: Ayyagari, Ranga Shaarad, et al.
Published: (2024)
by: Ayyagari, Ranga Shaarad, et al.
Published: (2024)
Offline Reinforcement Learning with Domain-Unlabeled Data
by: Nishimori, Soichiro, et al.
Published: (2024)
by: Nishimori, Soichiro, et al.
Published: (2024)
Rethinking Optimal Transport in Offline Reinforcement Learning
by: Asadulaev, Arip, et al.
Published: (2024)
by: Asadulaev, Arip, et al.
Published: (2024)
Sparse Offline Reinforcement Learning with Corruption Robustness
by: Tran, Nam Phuong, et al.
Published: (2025)
by: Tran, Nam Phuong, et al.
Published: (2025)
Information-Directed Offline-to-Online Reinforcement Learning
by: Chen, Keru
Published: (2026)
by: Chen, Keru
Published: (2026)
When to Trust Your Simulator: Dynamics-Aware Hybrid Offline-and-Online Reinforcement Learning
by: Niu, Haoyi, et al.
Published: (2022)
by: Niu, Haoyi, et al.
Published: (2022)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
by: Wang, Qi, et al.
Published: (2023)
by: Wang, Qi, et al.
Published: (2023)
Learning Decisions Offline from Censored Observations with ε-insensitive Operational Costs
by: Chen, Minxia, et al.
Published: (2024)
by: Chen, Minxia, et al.
Published: (2024)
Preference Elicitation for Offline Reinforcement Learning
by: Pace, Alizée, et al.
Published: (2024)
by: Pace, Alizée, et al.
Published: (2024)
Simple Ingredients for Offline Reinforcement Learning
by: Cetin, Edoardo, et al.
Published: (2024)
by: Cetin, Edoardo, et al.
Published: (2024)
Similar Items
-
To Distill or Decide? Understanding the Algorithmic Trade-off in Partially Observable Reinforcement Learning
by: Song, Yuda, et al.
Published: (2025) -
Expanding the Capabilities of Reinforcement Learning via Text Feedback
by: Song, Yuda, et al.
Published: (2026) -
The Importance of Online Data: Understanding Preference Fine-tuning via Coverage
by: Song, Yuda, et al.
Published: (2024) -
Hybrid Inverse Reinforcement Learning
by: Ren, Juntao, et al.
Published: (2024) -
Inverse Reinforcement Learning without Reinforcement Learning
by: Swamy, Gokul, et al.
Published: (2023)