Leveraging Sub-Optimal Data for Human-in-the-Loop Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Muslimani, Calarina, Taylor, Matthew E. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Reward Learning through Ranking Mean Squared Error
por: Kharyal, Chaitanya, et al.
Publicado: (2026)
por: Kharyal, Chaitanya, et al.
Publicado: (2026)
A Systematic Approach to Design Real-World Human-in-the-Loop Deep Reinforcement Learning: Salient Features, Challenges and Trade-offs
por: Arabneydi, Jalal, et al.
Publicado: (2025)
por: Arabneydi, Jalal, et al.
Publicado: (2025)
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
por: Muslimani, Calarina, et al.
Publicado: (2025)
por: Muslimani, Calarina, et al.
Publicado: (2025)
Reinforcement Teaching
por: Muslimani, Calarina, et al.
Publicado: (2022)
por: Muslimani, Calarina, et al.
Publicado: (2022)
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
por: Walter, Tim, et al.
Publicado: (2025)
por: Walter, Tim, et al.
Publicado: (2025)
Redefining Data Pairing for Motion Retargeting Leveraging a Human Body Prior
por: Figuera, Xiyana, et al.
Publicado: (2024)
por: Figuera, Xiyana, et al.
Publicado: (2024)
Offline Reinforcement Learning with Wasserstein Regularization via Optimal Transport Maps
por: Omura, Motoki, et al.
Publicado: (2025)
por: Omura, Motoki, et al.
Publicado: (2025)
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving
por: Zhang, Zhihao, et al.
Publicado: (2025)
por: Zhang, Zhihao, et al.
Publicado: (2025)
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
por: Yunis, David, et al.
Publicado: (2023)
por: Yunis, David, et al.
Publicado: (2023)
ReinforceGen: Hybrid Skill Policies with Automated Data Generation and Reinforcement Learning
por: Zhou, Zihan, et al.
Publicado: (2025)
por: Zhou, Zihan, et al.
Publicado: (2025)
Gradual Transition from Bellman Optimality Operator to Bellman Operator in Online Reinforcement Learning
por: Omura, Motoki, et al.
Publicado: (2025)
por: Omura, Motoki, et al.
Publicado: (2025)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
por: Huang, Xingshuai, et al.
Publicado: (2024)
por: Huang, Xingshuai, et al.
Publicado: (2024)
A Clean Slate for Offline Reinforcement Learning
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
por: Jackson, Matthew Thomas, et al.
Publicado: (2025)
Handling Delay in Real-Time Reinforcement Learning
por: Anokhin, Ivan, et al.
Publicado: (2025)
por: Anokhin, Ivan, et al.
Publicado: (2025)
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning
por: Li, Keyu, et al.
Publicado: (2021)
por: Li, Keyu, et al.
Publicado: (2021)
Uncertainty-Based Smooth Policy Regularisation for Reinforcement Learning with Few Demonstrations
por: Zhu, Yujie, et al.
Publicado: (2025)
por: Zhu, Yujie, et al.
Publicado: (2025)
Belief Aided Navigation using Bayesian Reinforcement Learning for Avoiding Humans in Blind Spots
por: Kim, Jinyeob, et al.
Publicado: (2024)
por: Kim, Jinyeob, et al.
Publicado: (2024)
HAIM-DRL: Enhanced Human-in-the-loop Reinforcement Learning for Safe and Efficient Autonomous Driving
por: Huang, Zilin, et al.
Publicado: (2024)
por: Huang, Zilin, et al.
Publicado: (2024)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
Data-Efficient Learning from Human Interventions for Mobile Robots
por: Peng, Zhenghao, et al.
Publicado: (2025)
por: Peng, Zhenghao, et al.
Publicado: (2025)
Replication of Impedance Identification Experiments on a Reinforcement-Learning-Controlled Digital Twin of Human Elbows
por: Yu, Hao, et al.
Publicado: (2024)
por: Yu, Hao, et al.
Publicado: (2024)
Planning the path with Reinforcement Learning: Optimal Robot Motion Planning in RoboCup Small Size League Environments
por: Machado, Mateus G., et al.
Publicado: (2024)
por: Machado, Mateus G., et al.
Publicado: (2024)
Coarse-to-fine Q-Network with Action Sequence for Data-Efficient Reinforcement Learning
por: Seo, Younggyo, et al.
Publicado: (2024)
por: Seo, Younggyo, et al.
Publicado: (2024)
Data-Efficient Hierarchical Goal-Conditioned Reinforcement Learning via Normalizing Flows
por: Garg, Shaswat, et al.
Publicado: (2026)
por: Garg, Shaswat, et al.
Publicado: (2026)
Body Transformer: Leveraging Robot Embodiment for Policy Learning
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
por: Sferrazza, Carmelo, et al.
Publicado: (2024)
Interactive Double Deep Q-network: Integrating Human Interventions and Evaluative Predictions in Reinforcement Learning of Autonomous Driving
por: Sygkounas, Alkis, et al.
Publicado: (2025)
por: Sygkounas, Alkis, et al.
Publicado: (2025)
TADPO: Reinforcement Learning Goes Off-road
por: Wu, Zhouchonghao, et al.
Publicado: (2026)
por: Wu, Zhouchonghao, et al.
Publicado: (2026)
Scalable Multi-Agent Reinforcement Learning for Warehouse Logistics with Robotic and Human Co-Workers
por: Krnjaic, Aleksandar, et al.
Publicado: (2022)
por: Krnjaic, Aleksandar, et al.
Publicado: (2022)
AutoLoop: Fast Visual SLAM Fine-tuning through Agentic Curriculum Learning
por: Lahiany, Assaf, et al.
Publicado: (2025)
por: Lahiany, Assaf, et al.
Publicado: (2025)
RILe: Reinforced Imitation Learning
por: Albaba, Mert, et al.
Publicado: (2024)
por: Albaba, Mert, et al.
Publicado: (2024)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Rating-based Reinforcement Learning
por: White, Devin, et al.
Publicado: (2023)
por: White, Devin, et al.
Publicado: (2023)
Leveraging Procedural Generation for Learning Autonomous Peg-in-Hole Assembly in Space
por: Orsula, Andrej, et al.
Publicado: (2024)
por: Orsula, Andrej, et al.
Publicado: (2024)
Robot Policy Transfer with Online Demonstrations: An Active Reinforcement Learning Approach
por: Hou, Muhan, et al.
Publicado: (2025)
por: Hou, Muhan, et al.
Publicado: (2025)
Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
por: Poddar, Sriyash, et al.
Publicado: (2024)
por: Poddar, Sriyash, et al.
Publicado: (2024)
Privileged Sensing Scaffolds Reinforcement Learning
por: Hu, Edward S., et al.
Publicado: (2024)
por: Hu, Edward S., et al.
Publicado: (2024)
Automating the Refinement of Reinforcement Learning Specifications
por: Ambadkar, Tanmay, et al.
Publicado: (2025)
por: Ambadkar, Tanmay, et al.
Publicado: (2025)
Sampling-Based Safe Reinforcement Learning
por: Vignola, Luca, et al.
Publicado: (2026)
por: Vignola, Luca, et al.
Publicado: (2026)
Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data
por: Müller, Marcus G, et al.
Publicado: (2026)
por: Müller, Marcus G, et al.
Publicado: (2026)
Causal Flow Q-Learning for Robust Offline Reinforcement Learning
por: Li, Mingxuan, et al.
Publicado: (2026)
por: Li, Mingxuan, et al.
Publicado: (2026)
Ejemplares similares
-
Reward Learning through Ranking Mean Squared Error
por: Kharyal, Chaitanya, et al.
Publicado: (2026) -
A Systematic Approach to Design Real-World Human-in-the-Loop Deep Reinforcement Learning: Salient Features, Challenges and Trade-offs
por: Arabneydi, Jalal, et al.
Publicado: (2025) -
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
por: Muslimani, Calarina, et al.
Publicado: (2025) -
Reinforcement Teaching
por: Muslimani, Calarina, et al.
Publicado: (2022) -
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
por: Walter, Tim, et al.
Publicado: (2025)