OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation
Fuente:
arXiv
Guardado en:
| Autores principales: | Mo, Yunyang, Li, Jian, Wu, Qiwei, Kang, Yihang, Xu, Renjing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Gentle Manipulation Policy Learning via Demonstrations from VLM Planned Atomic Skills
por: Zhou, Jiayu, et al.
Publicado: (2025)
por: Zhou, Jiayu, et al.
Publicado: (2025)
Tabero: Learning Gentle Manipulation with Closed-Loop Force Feedback from Vision, Touch, and Language
por: Wu, Qiwei, et al.
Publicado: (2026)
por: Wu, Qiwei, et al.
Publicado: (2026)
RM-RL: Role-Model Reinforcement Learning for Precise Robot Manipulation
por: Chen, Xiangyu, et al.
Publicado: (2025)
por: Chen, Xiangyu, et al.
Publicado: (2025)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
por: Lei, Kun, et al.
Publicado: (2025)
por: Lei, Kun, et al.
Publicado: (2025)
Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
por: Luo, Jianlan, et al.
Publicado: (2024)
por: Luo, Jianlan, et al.
Publicado: (2024)
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation
por: Xu, Qinwen, et al.
Publicado: (2026)
por: Xu, Qinwen, et al.
Publicado: (2026)
Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation
por: Zhu, Yihang, et al.
Publicado: (2025)
por: Zhu, Yihang, et al.
Publicado: (2025)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
por: Lu, Guanxing, et al.
Publicado: (2025)
por: Lu, Guanxing, et al.
Publicado: (2025)
Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
por: Li, Huanyu, et al.
Publicado: (2026)
por: Li, Huanyu, et al.
Publicado: (2026)
Morphology-Consistent Humanoid Interaction through Robot-Centric Video Synthesis
por: Xu, Weisheng, et al.
Publicado: (2026)
por: Xu, Weisheng, et al.
Publicado: (2026)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
por: Jiang, Zhennan, et al.
Publicado: (2025)
por: Jiang, Zhennan, et al.
Publicado: (2025)
PDF-HR: Pose Distance Fields for Humanoid Robots
por: Gu, Yi, et al.
Publicado: (2026)
por: Gu, Yi, et al.
Publicado: (2026)
Beyond Viewpoint Generalization: What Multi-View Demonstrations Offer and How to Synthesize Them for Robot Manipulation?
por: Cai, Boyang, et al.
Publicado: (2026)
por: Cai, Boyang, et al.
Publicado: (2026)
GazeVLA: Learning Human Intention for Robotic Manipulation
por: Li, Chengyang, et al.
Publicado: (2026)
por: Li, Chengyang, et al.
Publicado: (2026)
ContactRL: Safe Reinforcement Learning based Motion Planning for Contact based Human Robot Collaboration
por: Mulkana, Sundas Rafat, et al.
Publicado: (2025)
por: Mulkana, Sundas Rafat, et al.
Publicado: (2025)
Learning Human-Robot Handshaking Preferences for Quadruped Robots
por: Chappuis, Alessandra, et al.
Publicado: (2024)
por: Chappuis, Alessandra, et al.
Publicado: (2024)
XPG-RL: Reinforcement Learning with Explainable Priority Guidance for Efficiency-Boosted Mechanical Search
por: Zhang, Yiting, et al.
Publicado: (2025)
por: Zhang, Yiting, et al.
Publicado: (2025)
RL-GSBridge: 3D Gaussian Splatting Based Real2Sim2Real Method for Robotic Manipulation Learning
por: Wu, Yuxuan, et al.
Publicado: (2024)
por: Wu, Yuxuan, et al.
Publicado: (2024)
Human-in-the-loop Online Rejection Sampling for Robotic Manipulation
por: Lu, Guanxing, et al.
Publicado: (2025)
por: Lu, Guanxing, et al.
Publicado: (2025)
MoSA: Motion-constrained Stress Adaptation for Mitigating Real-to-Sim Gap in Continuum Dynamics via Learning Residual Anisotropy
por: Wang, Jiaxu, et al.
Publicado: (2026)
por: Wang, Jiaxu, et al.
Publicado: (2026)
Accelerating Robotic Reinforcement Learning with Agent Guidance
por: Chen, Haojun, et al.
Publicado: (2026)
por: Chen, Haojun, et al.
Publicado: (2026)
Dexterous Grasping with Real-World Robotic Reinforcement Learning
por: Huang, Dongchi, et al.
Publicado: (2025)
por: Huang, Dongchi, et al.
Publicado: (2025)
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
por: Yan, Hongyu, et al.
Publicado: (2026)
por: Yan, Hongyu, et al.
Publicado: (2026)
A Contact-Safe Reinforcement Learning Framework for Contact-Rich Robot Manipulation
por: Zhu, Xiang, et al.
Publicado: (2022)
por: Zhu, Xiang, et al.
Publicado: (2022)
Learning Manipulation Skills through Robot Chain-of-Thought with Sparse Failure Guidance
por: Zhang, Kaifeng, et al.
Publicado: (2024)
por: Zhang, Kaifeng, et al.
Publicado: (2024)
Discrete Policy: Learning Disentangled Action Space for Multi-Task Robotic Manipulation
por: Wu, Kun, et al.
Publicado: (2024)
por: Wu, Kun, et al.
Publicado: (2024)
GR-RL: Going Dexterous and Precise for Long-Horizon Robotic Manipulation
por: Li, Yunfei, et al.
Publicado: (2025)
por: Li, Yunfei, et al.
Publicado: (2025)
Towards Human-Centric Autonomous Driving: A Fast-Slow Architecture Integrating Large Language Model Guidance with Reinforcement Learning
por: Xu, Chengkai, et al.
Publicado: (2025)
por: Xu, Chengkai, et al.
Publicado: (2025)
Waypoint-Based Reinforcement Learning for Robot Manipulation Tasks
por: Mehta, Shaunak A., et al.
Publicado: (2024)
por: Mehta, Shaunak A., et al.
Publicado: (2024)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
por: Bo, Zitong, et al.
Publicado: (2025)
por: Bo, Zitong, et al.
Publicado: (2025)
Bootstrap Dynamic-Aware 3D Visual Representation for Scalable Robot Learning
por: Liang, Qiwei, et al.
Publicado: (2025)
por: Liang, Qiwei, et al.
Publicado: (2025)
LiPS: Large-Scale Humanoid Robot Reinforcement Learning with Parallel-Series Structures
por: Zhang, Qiang, et al.
Publicado: (2025)
por: Zhang, Qiang, et al.
Publicado: (2025)
Multimodal Reinforcement Learning for Robots Collaborating with Humans
por: Shervedani, Afagh Mehri, et al.
Publicado: (2023)
por: Shervedani, Afagh Mehri, et al.
Publicado: (2023)
RoboCopilot: Human-in-the-loop Interactive Imitation Learning for Robot Manipulation
por: Wu, Philipp, et al.
Publicado: (2025)
por: Wu, Philipp, et al.
Publicado: (2025)
Distillation-PPO: A Novel Two-Stage Reinforcement Learning Framework for Humanoid Robot Perceptive Locomotion
por: Zhang, Qiang, et al.
Publicado: (2025)
por: Zhang, Qiang, et al.
Publicado: (2025)
Speech-to-Trajectory: Learning Human-Like Verbal Guidance for Robot Motion
por: Bamani, Eran Beeri, et al.
Publicado: (2025)
por: Bamani, Eran Beeri, et al.
Publicado: (2025)
Unity RL Playground: A Versatile Reinforcement Learning Framework for Mobile Robots
por: Ye, Linqi, et al.
Publicado: (2025)
por: Ye, Linqi, et al.
Publicado: (2025)
Empowering Embodied Manipulation: A Bimanual-Mobile Robot Manipulation Dataset for Household Tasks
por: Zhang, Tianle, et al.
Publicado: (2024)
por: Zhang, Tianle, et al.
Publicado: (2024)
Human-Robot Gym: Benchmarking Reinforcement Learning in Human-Robot Collaboration
por: Thumm, Jakob, et al.
Publicado: (2023)
por: Thumm, Jakob, et al.
Publicado: (2023)
Robot Air Hockey: A Manipulation Testbed for Robot Learning with Reinforcement Learning
por: Chuck, Caleb, et al.
Publicado: (2024)
por: Chuck, Caleb, et al.
Publicado: (2024)
Ejemplares similares
-
Gentle Manipulation Policy Learning via Demonstrations from VLM Planned Atomic Skills
por: Zhou, Jiayu, et al.
Publicado: (2025) -
Tabero: Learning Gentle Manipulation with Closed-Loop Force Feedback from Vision, Touch, and Language
por: Wu, Qiwei, et al.
Publicado: (2026) -
RM-RL: Role-Model Reinforcement Learning for Precise Robot Manipulation
por: Chen, Xiangyu, et al.
Publicado: (2025) -
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
por: Lei, Kun, et al.
Publicado: (2025) -
Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
por: Luo, Jianlan, et al.
Publicado: (2024)