Leveraging Unlabeled Data Sharing through Kernel Function Approximation in Offline Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lai, Yen-Ru, Chang, Fu-Chieh, Wu, Pei-Yuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Offline Reinforcement Learning with Domain-Unlabeled Data
von: Nishimori, Soichiro, et al.
Veröffentlicht: (2024)
von: Nishimori, Soichiro, et al.
Veröffentlicht: (2024)
Corruption-Robust Offline Reinforcement Learning with General Function Approximation
von: Ye, Chenlu, et al.
Veröffentlicht: (2023)
von: Ye, Chenlu, et al.
Veröffentlicht: (2023)
Augmenting Offline RL with Unlabeled Data
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
A Theoretical Framework for OOD Robustness in Transformers using Gevrey Classes
von: Wang, Yu, et al.
Veröffentlicht: (2025)
von: Wang, Yu, et al.
Veröffentlicht: (2025)
On Leveraging Unlabeled Data for Concurrent Positive-Unlabeled Classification and Robust Generation
von: Yu, Bing, et al.
Veröffentlicht: (2020)
von: Yu, Bing, et al.
Veröffentlicht: (2020)
Quantum Hierarchical Reinforcement Learning via Variational Quantum Circuits
von: Lee, Yu-Ting, et al.
Veröffentlicht: (2026)
von: Lee, Yu-Ting, et al.
Veröffentlicht: (2026)
Urban-Focused Multi-Task Offline Reinforcement Learning with Contrastive Data Sharing
von: Zhao, Xinbo, et al.
Veröffentlicht: (2024)
von: Zhao, Xinbo, et al.
Veröffentlicht: (2024)
Unveiling the Latent Directions of Reflection in Large Language Models
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2025)
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2025)
RL-STaR: Theoretical Analysis of Reinforcement Learning Frameworks for Self-Taught Reasoner
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2024)
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2024)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
von: Golowich, Noah, et al.
Veröffentlicht: (2024)
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
von: Bai, Chenjia, et al.
Veröffentlicht: (2024)
von: Bai, Chenjia, et al.
Veröffentlicht: (2024)
Pretraining a Shared Q-Network for Data-Efficient Offline Reinforcement Learning
von: Park, Jongchan, et al.
Veröffentlicht: (2025)
von: Park, Jongchan, et al.
Veröffentlicht: (2025)
Accelerated Policy Gradient: On the Convergence Rates of the Nesterov Momentum for Reinforcement Learning
von: Chen, Yen-Ju, et al.
Veröffentlicht: (2023)
von: Chen, Yen-Ju, et al.
Veröffentlicht: (2023)
Unraveling Arithmetic in Large Language Models: The Role of Algebraic Structures
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2024)
von: Chang, Fu-Chieh, et al.
Veröffentlicht: (2024)
Temporal Abstraction in Reinforcement Learning with Offline Data
von: Ayyagari, Ranga Shaarad, et al.
Veröffentlicht: (2024)
von: Ayyagari, Ranga Shaarad, et al.
Veröffentlicht: (2024)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
Offline Trajectory Optimization for Offline Reinforcement Learning
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhao, Ziqi, et al.
Veröffentlicht: (2024)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)
von: Huang, Xingshuai, et al.
Veröffentlicht: (2024)
PLEIADES: Building Temporal Kernels with Orthogonal Polynomials
von: Pei, Yan Ru, et al.
Veröffentlicht: (2024)
von: Pei, Yan Ru, et al.
Veröffentlicht: (2024)
Positive-Unlabeled Reinforcement Learning Distillation for On-Premise Small Models
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2026)
von: Kou, Zhiqiang, et al.
Veröffentlicht: (2026)
Diffusion Policies with Offline and Inverse Reinforcement Learning for Promoting Physical Activity in Older Adults Using Wearable Sensors
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
In-Context Positive-Unlabeled Learning
von: Liu, Siyan, et al.
Veröffentlicht: (2026)
von: Liu, Siyan, et al.
Veröffentlicht: (2026)
On the Complexity of Offline Reinforcement Learning with $Q^\star$-Approximation and Partial Coverage
von: Liu, Haolin, et al.
Veröffentlicht: (2026)
von: Liu, Haolin, et al.
Veröffentlicht: (2026)
Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration
von: Wilcoxson, Max, et al.
Veröffentlicht: (2024)
von: Wilcoxson, Max, et al.
Veröffentlicht: (2024)
Robust Offline Active Learning on Graphs
von: Wu, Yuanchen, et al.
Veröffentlicht: (2024)
von: Wu, Yuanchen, et al.
Veröffentlicht: (2024)
Learning from Uncertain Similarity and Unlabeled Data
von: Wei, Meng, et al.
Veröffentlicht: (2025)
von: Wei, Meng, et al.
Veröffentlicht: (2025)
Generalizing Beyond Suboptimality: Offline Reinforcement Learning Learns Effective Scheduling through Random Data
von: van Remmerden, Jesse, et al.
Veröffentlicht: (2025)
von: van Remmerden, Jesse, et al.
Veröffentlicht: (2025)
Trajectory-Level Data Augmentation for Offline Reinforcement Learning
von: Schmähling, Tobias, et al.
Veröffentlicht: (2026)
von: Schmähling, Tobias, et al.
Veröffentlicht: (2026)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
Benchmarks for Reinforcement Learning with Biased Offline Data and Imperfect Simulators
von: Linial, Ori, et al.
Veröffentlicht: (2024)
von: Linial, Ori, et al.
Veröffentlicht: (2024)
Active Advantage-Aligned Online Reinforcement Learning with Offline Data
von: Liu, Xuefeng, et al.
Veröffentlicht: (2025)
von: Liu, Xuefeng, et al.
Veröffentlicht: (2025)
Offline Constrained Reinforcement Learning under Partial Data Coverage
von: Ko, Seokmin, et al.
Veröffentlicht: (2025)
von: Ko, Seokmin, et al.
Veröffentlicht: (2025)
Goal-conditioned Offline Reinforcement Learning through State Space Partitioning
von: Wang, Mianchu, et al.
Veröffentlicht: (2023)
von: Wang, Mianchu, et al.
Veröffentlicht: (2023)
Instrumental Variable Value Iteration for Causal Offline Reinforcement Learning
von: Liao, Luofeng, et al.
Veröffentlicht: (2021)
von: Liao, Luofeng, et al.
Veröffentlicht: (2021)
Data-Incremental Continual Offline Reinforcement Learning
von: Gai, Sibo, et al.
Veröffentlicht: (2024)
von: Gai, Sibo, et al.
Veröffentlicht: (2024)
Leveraging Temporally Extended Behavior Sharing for Multi-task Reinforcement Learning
von: Lee, Gawon, et al.
Veröffentlicht: (2025)
von: Lee, Gawon, et al.
Veröffentlicht: (2025)
Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning
von: Jing, Tan, et al.
Veröffentlicht: (2025)
von: Jing, Tan, et al.
Veröffentlicht: (2025)
Nonstationary Reinforcement Learning with Linear Function Approximation
von: Zhou, Huozhi, et al.
Veröffentlicht: (2020)
von: Zhou, Huozhi, et al.
Veröffentlicht: (2020)
Replicable Reinforcement Learning with Linear Function Approximation
von: Eaton, Eric, et al.
Veröffentlicht: (2025)
von: Eaton, Eric, et al.
Veröffentlicht: (2025)
Interactive Symbolic Regression through Offline Reinforcement Learning: A Co-Design Framework
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Offline Reinforcement Learning with Domain-Unlabeled Data
von: Nishimori, Soichiro, et al.
Veröffentlicht: (2024) -
Corruption-Robust Offline Reinforcement Learning with General Function Approximation
von: Ye, Chenlu, et al.
Veröffentlicht: (2023) -
Augmenting Offline RL with Unlabeled Data
von: Wang, Zhao, et al.
Veröffentlicht: (2024) -
A Theoretical Framework for OOD Robustness in Transformers using Gevrey Classes
von: Wang, Yu, et al.
Veröffentlicht: (2025) -
On Leveraging Unlabeled Data for Concurrent Positive-Unlabeled Classification and Robust Generation
von: Yu, Bing, et al.
Veröffentlicht: (2020)