RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking
Fuente:
arXiv
Guardado en:
| Autores principales: | Choi, Andrew, Xu, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
por: Zhao, Kai, et al.
Publicado: (2023)
por: Zhao, Kai, et al.
Publicado: (2023)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
por: Li, Mingxuan, et al.
Publicado: (2026)
por: Li, Mingxuan, et al.
Publicado: (2026)
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
por: Wu, Kun, et al.
Publicado: (2024)
por: Wu, Kun, et al.
Publicado: (2024)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
por: Su, Huikang, et al.
Publicado: (2025)
por: Su, Huikang, et al.
Publicado: (2025)
Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
por: Li, Huanyu, et al.
Publicado: (2026)
por: Li, Huanyu, et al.
Publicado: (2026)
Adversarial Fine-tuning in Offline-to-Online Reinforcement Learning for Robust Robot Control
por: Ayabe, Shingo, et al.
Publicado: (2025)
por: Ayabe, Shingo, et al.
Publicado: (2025)
BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models
por: Chen, Zhongxi, et al.
Publicado: (2026)
por: Chen, Zhongxi, et al.
Publicado: (2026)
FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning
por: Alles, Marvin, et al.
Publicado: (2025)
por: Alles, Marvin, et al.
Publicado: (2025)
Toward Learning POMDPs Beyond Full-Rank Actions and State Observability
por: Shaw, Seiji, et al.
Publicado: (2026)
por: Shaw, Seiji, et al.
Publicado: (2026)
Integrating Offline Pre-Training with Online Fine-Tuning: A Reinforcement Learning Approach for Robot Social Navigation
por: Su, Run, et al.
Publicado: (2025)
por: Su, Run, et al.
Publicado: (2025)
Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
por: Liu, Tenglong, et al.
Publicado: (2024)
por: Liu, Tenglong, et al.
Publicado: (2024)
Resilient UAV Trajectory Planning via Few-Shot Meta-Offline Reinforcement Learning
por: Eldeeb, Eslam, et al.
Publicado: (2025)
por: Eldeeb, Eslam, et al.
Publicado: (2025)
Scaling Sim-to-Real Reinforcement Learning for Robot VLAs with Generative 3D Worlds
por: Choi, Andrew, et al.
Publicado: (2026)
por: Choi, Andrew, et al.
Publicado: (2026)
Cross-Embodiment Offline Reinforcement Learning for Heterogeneous Robot Datasets
por: Abe, Haruki, et al.
Publicado: (2026)
por: Abe, Haruki, et al.
Publicado: (2026)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
por: Seo, Younggyo, et al.
Publicado: (2024)
por: Seo, Younggyo, et al.
Publicado: (2024)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
por: Guo, Yihong, et al.
Publicado: (2025)
por: Guo, Yihong, et al.
Publicado: (2025)
Goal-Conditioned Decision Transformer for Multi-Goal Offline Reinforcement Learning
por: Gajewski, Paweł, et al.
Publicado: (2024)
por: Gajewski, Paweł, et al.
Publicado: (2024)
Offline Reinforcement Learning with Wasserstein Regularization via Optimal Transport Maps
por: Omura, Motoki, et al.
Publicado: (2025)
por: Omura, Motoki, et al.
Publicado: (2025)
Robust Offline Reinforcement Learning with Linearly Structured f-Divergence Regularization
por: Tang, Cheng, et al.
Publicado: (2024)
por: Tang, Cheng, et al.
Publicado: (2024)
Diffusion Policies with Value-Conditional Optimization for Offline Reinforcement Learning
por: Ma, Yunchang, et al.
Publicado: (2025)
por: Ma, Yunchang, et al.
Publicado: (2025)
Physically-Grounded Goal Imagination: Physics-Informed Variational Autoencoder for Self-Supervised Reinforcement Learning
por: Nguyen, Lan Thi Ha, et al.
Publicado: (2025)
por: Nguyen, Lan Thi Ha, et al.
Publicado: (2025)
Rapidly Learning Soft Robot Control via Implicit Time-Stepping
por: Choi, Andrew, et al.
Publicado: (2025)
por: Choi, Andrew, et al.
Publicado: (2025)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
Offline Reinforcement Learning with Discrete Diffusion Skills
por: Qiao, RuiXi, et al.
Publicado: (2025)
por: Qiao, RuiXi, et al.
Publicado: (2025)
Improving Offline Reinforcement Learning with Inaccurate Simulators
por: Hou, Yiwen, et al.
Publicado: (2024)
por: Hou, Yiwen, et al.
Publicado: (2024)
A Clean Slate for Offline Reinforcement Learning
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
TROFI: Trajectory-Ranked Offline Inverse Reinforcement Learning
por: Sestini, Alessandro, et al.
Publicado: (2025)
por: Sestini, Alessandro, et al.
Publicado: (2025)
Localized Dynamics-Aware Domain Adaption for Off-Dynamics Offline Reinforcement Learning
por: Xia, Zhangjie, et al.
Publicado: (2026)
por: Xia, Zhangjie, et al.
Publicado: (2026)
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement
por: Ayalew, Tewodros, et al.
Publicado: (2024)
por: Ayalew, Tewodros, et al.
Publicado: (2024)
Self-adapting Robotic Agents through Online Continual Reinforcement Learning with World Model Feedback
por: Domberg, Fabian, et al.
Publicado: (2026)
por: Domberg, Fabian, et al.
Publicado: (2026)
Forecasting in Offline Reinforcement Learning for Non-stationary Environments
por: Ada, Suzan Ece, et al.
Publicado: (2025)
por: Ada, Suzan Ece, et al.
Publicado: (2025)
Variational OOD State Correction for Offline Reinforcement Learning
por: Jiang, Ke, et al.
Publicado: (2025)
por: Jiang, Ke, et al.
Publicado: (2025)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
por: Huang, Xingshuai, et al.
Publicado: (2024)
por: Huang, Xingshuai, et al.
Publicado: (2024)
Graph-Assisted Stitching for Offline Hierarchical Reinforcement Learning
por: Baek, Seungho, et al.
Publicado: (2025)
por: Baek, Seungho, et al.
Publicado: (2025)
Robotic Manipulation Datasets for Offline Compositional Reinforcement Learning
por: Hussing, Marcel, et al.
Publicado: (2023)
por: Hussing, Marcel, et al.
Publicado: (2023)
Vintix: Action Model via In-Context Reinforcement Learning
por: Polubarov, Andrey, et al.
Publicado: (2025)
por: Polubarov, Andrey, et al.
Publicado: (2025)
Latent Action Priors for Locomotion with Deep Reinforcement Learning
por: Hausdörfer, Oliver, et al.
Publicado: (2024)
por: Hausdörfer, Oliver, et al.
Publicado: (2024)
V-OCBF: Learning Safety Filters from Offline Data via Value-Guided Offline Control Barrier Functions
por: Tayal, Mumuksh, et al.
Publicado: (2025)
por: Tayal, Mumuksh, et al.
Publicado: (2025)
A Recipe for Stable Offline Multi-agent Reinforcement Learning
por: Lee, Dongsu, et al.
Publicado: (2026)
por: Lee, Dongsu, et al.
Publicado: (2026)
Offline Actor-Critic Reinforcement Learning Scales to Large Models
por: Springenberg, Jost Tobias, et al.
Publicado: (2024)
por: Springenberg, Jost Tobias, et al.
Publicado: (2024)
Ejemplares similares
-
ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
por: Zhao, Kai, et al.
Publicado: (2023) -
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
por: Li, Mingxuan, et al.
Publicado: (2026) -
ACL-QL: Adaptive Conservative Level in Q-Learning for Offline Reinforcement Learning
por: Wu, Kun, et al.
Publicado: (2024) -
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
por: Su, Huikang, et al.
Publicado: (2025) -
Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
por: Li, Huanyu, et al.
Publicado: (2026)