ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Kun, Zhao, Yinuo, Xu, Zhiyuan, Che, Zhengping, Yin, Chengxiang, Liu, Chi Harold, Feng, Feiferi, Tang, Jian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HACTS: a Human-As-Copilot Teleoperation System for Robot Learning
by: Xu, Zhiyuan, et al.
Published: (2025)
by: Xu, Zhiyuan, et al.
Published: (2025)
Real-world Reinforcement Learning from Suboptimal Interventions
by: Zhao, Yinuo, et al.
Published: (2025)
by: Zhao, Yinuo, et al.
Published: (2025)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)
by: Zhao, Yinuo, et al.
Published: (2024)
Training-free Generation of Temporally Consistent Rewards from VLMs
by: Zhao, Yinuo, et al.
Published: (2025)
by: Zhao, Yinuo, et al.
Published: (2025)
Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning
by: Gireesh, Nandiraju, et al.
Published: (2026)
by: Gireesh, Nandiraju, et al.
Published: (2026)
Learning from Imperfect Demonstrations with Self-Supervision for Robotic Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
ArtVIP: Articulated Digital Assets of Visual Realism, Modular Interaction, and Physical Fidelity for Robot Learning
by: Jin, Zhao, et al.
Published: (2025)
by: Jin, Zhao, et al.
Published: (2025)
A Survey on Robotics with Foundation Models: toward Embodied AI
by: Xu, Zhiyuan, et al.
Published: (2024)
by: Xu, Zhiyuan, et al.
Published: (2024)
RoboAug: One Annotation to Hundreds of Scenes via Region-Contrastive Data Augmentation for Robotic Manipulation
by: Wang, Xinhua, et al.
Published: (2026)
by: Wang, Xinhua, et al.
Published: (2026)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
FreqPolicy: Efficient Flow-based Visuomotor Policy via Frequency Consistency
by: Su, Yifei, et al.
Published: (2025)
by: Su, Yifei, et al.
Published: (2025)
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
by: Zhao, Kai, et al.
Published: (2023)
by: Zhao, Kai, et al.
Published: (2023)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
by: Li, Mingxuan, et al.
Published: (2026)
by: Li, Mingxuan, et al.
Published: (2026)
Trajectory-Level Data Augmentation for Offline Reinforcement Learning
by: Schmähling, Tobias, et al.
Published: (2026)
by: Schmähling, Tobias, et al.
Published: (2026)
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations
by: Fan, Shichao, et al.
Published: (2025)
by: Fan, Shichao, et al.
Published: (2025)
BiCQL-ML: A Bi-Level Conservative Q-Learning Framework for Maximum Likelihood Inverse Reinforcement Learning
by: Park, Junsung
Published: (2025)
by: Park, Junsung
Published: (2025)
Q-value Regularized Decision ConvFormer for Offline Reinforcement Learning
by: Yan, Teng, et al.
Published: (2024)
by: Yan, Teng, et al.
Published: (2024)
Improving Offline Reinforcement Learning with Inaccurate Simulators
by: Hou, Yiwen, et al.
Published: (2024)
by: Hou, Yiwen, et al.
Published: (2024)
Residual Learning and Context Encoding for Adaptive Offline-to-Online Reinforcement Learning
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
Conservative Offline Robot Policy Learning via Posterior-Transition Reweighting
by: Zhang, Wanpeng, et al.
Published: (2026)
by: Zhang, Wanpeng, et al.
Published: (2026)
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
by: Alles, Marvin, et al.
Published: (2025)
by: Alles, Marvin, et al.
Published: (2025)
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking
by: Choi, Andrew, et al.
Published: (2026)
by: Choi, Andrew, et al.
Published: (2026)
Towards an Adaptive Social Game-Playing Robot: An Offline Reinforcement Learning-Based Framework
by: Chu, Soon Jynn, et al.
Published: (2025)
by: Chu, Soon Jynn, et al.
Published: (2025)
Investigating Adaptive Tuning of Assistive Exoskeletons Using Offline Reinforcement Learning: Challenges and Insights
by: Findik, Yasin, et al.
Published: (2025)
by: Findik, Yasin, et al.
Published: (2025)
Equivariant Offline Reinforcement Learning
by: Tangri, Arsh, et al.
Published: (2024)
by: Tangri, Arsh, et al.
Published: (2024)
Visual Robotic Manipulation with Depth-Aware Pretraining
by: Wang, Wanying, et al.
Published: (2024)
by: Wang, Wanying, et al.
Published: (2024)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
by: Liu, Tenglong, et al.
Published: (2024)
by: Liu, Tenglong, et al.
Published: (2024)
Uncertainty-Aware Rank-One MIMO Q Network Framework for Accelerated Offline Reinforcement Learning
by: Nguyen, Thanh, et al.
Published: (2026)
by: Nguyen, Thanh, et al.
Published: (2026)
CRAFT: Adapting VLA Models to Contact-rich Manipulation via Force-aware Curriculum Fine-tuning
by: Zhang, Yike, et al.
Published: (2026)
by: Zhang, Yike, et al.
Published: (2026)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Safe Offline Reinforcement Learning with Real-Time Budget Constraints
by: Lin, Qian, et al.
Published: (2023)
by: Lin, Qian, et al.
Published: (2023)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
by: Huang, Xingshuai, et al.
Published: (2024)
by: Huang, Xingshuai, et al.
Published: (2024)
Offline Imitation Learning Through Graph Search and Retrieval
by: Yin, Zhao-Heng, et al.
Published: (2024)
by: Yin, Zhao-Heng, et al.
Published: (2024)
Robust Offline Reinforcement Learning with Linearly Structured f-Divergence Regularization
by: Tang, Cheng, et al.
Published: (2024)
by: Tang, Cheng, et al.
Published: (2024)
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
by: Ma, Yunchang, et al.
Published: (2025)
by: Ma, Yunchang, et al.
Published: (2025)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
by: Corrado, Nicholas E., et al.
Published: (2023)
by: Corrado, Nicholas E., et al.
Published: (2023)
Diffusion Trajectory-guided Policy for Long-horizon Robot Manipulation
by: Fan, Shichao, et al.
Published: (2025)
by: Fan, Shichao, et al.
Published: (2025)
Extremum Flow Matching for Offline Goal Conditioned Reinforcement Learning
by: Rouxel, Quentin, et al.
Published: (2025)
by: Rouxel, Quentin, et al.
Published: (2025)
Learning Vision-based Robotic Manipulation Tasks Sequentially in Offline Reinforcement Learning Settings
by: Yadav, Sudhir Pratap, et al.
Published: (2023)
by: Yadav, Sudhir Pratap, et al.
Published: (2023)
Q-SpiRL: Quantum Spiking Reinforcement Learning for Adaptive Robot Navigation
by: Altrabulsi, Mohamed Khair, et al.
Published: (2026)
by: Altrabulsi, Mohamed Khair, et al.
Published: (2026)
Similar Items
-
HACTS: a Human-As-Copilot Teleoperation System for Robot Learning
by: Xu, Zhiyuan, et al.
Published: (2025) -
Real-world Reinforcement Learning from Suboptimal Interventions
by: Zhao, Yinuo, et al.
Published: (2025) -
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024) -
Training-free Generation of Temporally Consistent Rewards from VLMs
by: Zhao, Yinuo, et al.
Published: (2025) -
Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning
by: Gireesh, Nandiraju, et al.
Published: (2026)