E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Deng, Haoyuan, Xue, Yuanjiang, Du, Haoyang, Zhou, Boyang, Wu, Zhenyu, Wang, Ziwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
vEDGAR -- Can CARLA Do HiL?
von: Gehrke, Nils, et al.
Veröffentlicht: (2025)
von: Gehrke, Nils, et al.
Veröffentlicht: (2025)
UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
von: Liu, Haichao, et al.
Veröffentlicht: (2026)
SafeBimanual: Diffusion-based Trajectory Optimization for Safe Bimanual Manipulation
von: Deng, Haoyuan, et al.
Veröffentlicht: (2025)
von: Deng, Haoyuan, et al.
Veröffentlicht: (2025)
VLA-Reasoner: Empowering Vision-Language-Action Models with Reasoning via Online Monte Carlo Tree Search
von: Guo, Wenkai, et al.
Veröffentlicht: (2025)
von: Guo, Wenkai, et al.
Veröffentlicht: (2025)
HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?
von: Trinh, Tu, et al.
Veröffentlicht: (2026)
von: Trinh, Tu, et al.
Veröffentlicht: (2026)
Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
von: Li, Yaxuan, et al.
Veröffentlicht: (2026)
MAP-VLA: Memory-Augmented Prompting for Vision-Language-Action Model in Robotic Manipulation
von: Li, Runhao, et al.
Veröffentlicht: (2025)
von: Li, Runhao, et al.
Veröffentlicht: (2025)
PlannerRFT: Reinforcing Diffusion Planners through Closed-Loop and Sample-Efficient Fine-Tuning
von: Li, Hongchen, et al.
Veröffentlicht: (2026)
von: Li, Hongchen, et al.
Veröffentlicht: (2026)
DockAnywhere: Data-Efficient Visuomotor Policy Learning for Mobile Manipulation via Novel Demonstration Generation
von: Shan, Ziyu, et al.
Veröffentlicht: (2026)
von: Shan, Ziyu, et al.
Veröffentlicht: (2026)
DexHiL: A Human-in-the-Loop Framework for Vision-Language-Action Model Post-Training in Dexterous Manipulation
von: Han, Yifan, et al.
Veröffentlicht: (2026)
von: Han, Yifan, et al.
Veröffentlicht: (2026)
Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior
von: Wu, Yuanye, et al.
Veröffentlicht: (2026)
von: Wu, Yuanye, et al.
Veröffentlicht: (2026)
SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training
von: Wu, Mingdong, et al.
Veröffentlicht: (2025)
von: Wu, Mingdong, et al.
Veröffentlicht: (2025)
A Systematic Approach to Design Real-World Human-in-the-Loop Deep Reinforcement Learning: Salient Features, Challenges and Trade-offs
von: Arabneydi, Jalal, et al.
Veröffentlicht: (2025)
von: Arabneydi, Jalal, et al.
Veröffentlicht: (2025)
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation
von: Xue, Yuquan, et al.
Veröffentlicht: (2025)
von: Xue, Yuquan, et al.
Veröffentlicht: (2025)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
von: Bo, Zitong, et al.
Veröffentlicht: (2025)
von: Bo, Zitong, et al.
Veröffentlicht: (2025)
Closed-Loop Sim-to-Real Reinforcement Learning for Deformable Microfiber Shape Control
von: Amici, Alessandro, et al.
Veröffentlicht: (2026)
von: Amici, Alessandro, et al.
Veröffentlicht: (2026)
Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
von: Luo, Jianlan, et al.
Veröffentlicht: (2024)
von: Luo, Jianlan, et al.
Veröffentlicht: (2024)
Development and Preliminary Evaluation of a Machine Vision‐Guided Smart Sprayer Prototype Toward Precision Vegetable Weeding
von: Boyang Deng, et al.
Veröffentlicht: (2025)
von: Boyang Deng, et al.
Veröffentlicht: (2025)
RoHIL: Robust Human-in-the-Loop Robotic Reinforcement Learning Against Illumination Variations
von: Zhang, Shuoqin, et al.
Veröffentlicht: (2026)
von: Zhang, Shuoqin, et al.
Veröffentlicht: (2026)
"Hi AirStar, Guide Me to the Badminton Court."
von: Wang, Ziqin, et al.
Veröffentlicht: (2025)
von: Wang, Ziqin, et al.
Veröffentlicht: (2025)
Sample-Efficient Reinforcement Learning with Symmetry-Guided Demonstrations for Robotic Manipulation
von: Enayati, Amir M. Soufi, et al.
Veröffentlicht: (2023)
von: Enayati, Amir M. Soufi, et al.
Veröffentlicht: (2023)
D-Optimality-Guided Reinforcement Learning for Efficient Open-Loop Calibration of a 3-DOF Ankle Rehabilitation Robot
von: Hu, Qifan, et al.
Veröffentlicht: (2026)
von: Hu, Qifan, et al.
Veröffentlicht: (2026)
Hi-Dyna Graph: Hierarchical Dynamic Scene Graph for Robotic Autonomy in Human-Centric Environments
von: Hou, Jiawei, et al.
Veröffentlicht: (2025)
von: Hou, Jiawei, et al.
Veröffentlicht: (2025)
Selective Progress-Aware Querying for Human-in-the-Loop Reinforcement Learning
von: Muraleedharan, Anujith, et al.
Veröffentlicht: (2025)
von: Muraleedharan, Anujith, et al.
Veröffentlicht: (2025)
HiCRISP: An LLM-based Hierarchical Closed-Loop Robotic Intelligent Self-Correction Planner
von: Ming, Chenlin, et al.
Veröffentlicht: (2023)
von: Ming, Chenlin, et al.
Veröffentlicht: (2023)
Efficient Reinforcement Learning by Guiding Generalist World Models with Non-Curated Data
von: Zhao, Yi, et al.
Veröffentlicht: (2025)
von: Zhao, Yi, et al.
Veröffentlicht: (2025)
RL-100: Performant Robotic Manipulation with Real-World Reinforcement Learning
von: Lei, Kun, et al.
Veröffentlicht: (2025)
von: Lei, Kun, et al.
Veröffentlicht: (2025)
CTE-MLO: Continuous-time and Efficient Multi-LiDAR Odometry with Localizability-aware Point Cloud Sampling
von: Shen, Hongming, et al.
Veröffentlicht: (2024)
von: Shen, Hongming, et al.
Veröffentlicht: (2024)
Computationally and Sample Efficient Safe Reinforcement Learning Using Adaptive Conformal Prediction
von: Zhou, Hao, et al.
Veröffentlicht: (2025)
von: Zhou, Hao, et al.
Veröffentlicht: (2025)
Learning to Sample: Reinforcement Learning-Guided Sampling for Autonomous Vehicle Motion Planning
von: Moller, Korbinian, et al.
Veröffentlicht: (2025)
von: Moller, Korbinian, et al.
Veröffentlicht: (2025)
Dexterous Grasping with Real-World Robotic Reinforcement Learning
von: Huang, Dongchi, et al.
Veröffentlicht: (2025)
von: Huang, Dongchi, et al.
Veröffentlicht: (2025)
ManiGaussian++: General Robotic Bimanual Manipulation with Hierarchical Gaussian World Model
von: Yu, Tengbo, et al.
Veröffentlicht: (2025)
von: Yu, Tengbo, et al.
Veröffentlicht: (2025)
Predictive Preference Learning from Human Interventions
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
AnyBimanual: Transferring Unimanual Policy for General Bimanual Manipulation
von: Lu, Guanxing, et al.
Veröffentlicht: (2024)
von: Lu, Guanxing, et al.
Veröffentlicht: (2024)
DemoSpeedup: Accelerating Visuomotor Policies via Entropy-Guided Demonstration Acceleration
von: Guo, Lingxiao, et al.
Veröffentlicht: (2025)
von: Guo, Lingxiao, et al.
Veröffentlicht: (2025)
L1 Sample Flow for Efficient Visuomotor Learning
von: Song, Weixi, et al.
Veröffentlicht: (2025)
von: Song, Weixi, et al.
Veröffentlicht: (2025)
HiCrowd: Hierarchical Crowd Flow Alignment for Dense Human Environments
von: Zhu, Yufei, et al.
Veröffentlicht: (2026)
von: Zhu, Yufei, et al.
Veröffentlicht: (2026)
Sample-Efficient Reinforcement Learning with Temporal Logic Objectives: Leveraging the Task Specification to Guide Exploration
von: Kantaros, Yiannis, et al.
Veröffentlicht: (2024)
von: Kantaros, Yiannis, et al.
Veröffentlicht: (2024)
NORA-1.5: A Vision-Language-Action Model Trained using World Model- and Action-based Preference Rewards
von: Hung, Chia-Yu, et al.
Veröffentlicht: (2025)
von: Hung, Chia-Yu, et al.
Veröffentlicht: (2025)
Loop Closure from Two Views: Revisiting PGO for Scalable Trajectory Estimation through Monocular Priors
von: Lim, Tian Yi, et al.
Veröffentlicht: (2025)
von: Lim, Tian Yi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
vEDGAR -- Can CARLA Do HiL?
von: Gehrke, Nils, et al.
Veröffentlicht: (2025) -
UniManip: General-Purpose Zero-Shot Robotic Manipulation with Agentic Operational Graph
von: Liu, Haichao, et al.
Veröffentlicht: (2026) -
SafeBimanual: Diffusion-based Trajectory Optimization for Safe Bimanual Manipulation
von: Deng, Haoyuan, et al.
Veröffentlicht: (2025) -
VLA-Reasoner: Empowering Vision-Language-Action Models with Reasoning via Online Monte Carlo Tree Search
von: Guo, Wenkai, et al.
Veröffentlicht: (2025) -
HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?
von: Trinh, Tu, et al.
Veröffentlicht: (2026)