Learning a High-quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yihong, Kang, Dongyeop, Ha, Sehoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PrivilegedDreamer: Explicit Imagination of Privileged Information for Rapid Adaptation of Learned Policies
von: Byrd, Morgan, et al.
Veröffentlicht: (2025)
von: Byrd, Morgan, et al.
Veröffentlicht: (2025)
Deep Reinforcement Learning for Multi-Agent Coordination
von: Aina, Kehinde O., et al.
Veröffentlicht: (2025)
von: Aina, Kehinde O., et al.
Veröffentlicht: (2025)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
von: Tang, Nan, et al.
Veröffentlicht: (2025)
von: Tang, Nan, et al.
Veröffentlicht: (2025)
Decoupling Task and Behavior: A Two-Stage Reward Curriculum in Reinforcement Learning for Robotics
von: Freitag, Kilian, et al.
Veröffentlicht: (2026)
von: Freitag, Kilian, et al.
Veröffentlicht: (2026)
Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion
von: Rho, Seungeun, et al.
Veröffentlicht: (2025)
von: Rho, Seungeun, et al.
Veröffentlicht: (2025)
Learning to enhance multi-legged robot on rugged landscapes
von: He, Juntao, et al.
Veröffentlicht: (2024)
von: He, Juntao, et al.
Veröffentlicht: (2024)
Curriculum Reinforcement Learning for Complex Reward Functions
von: Freitag, Kilian, et al.
Veröffentlicht: (2024)
von: Freitag, Kilian, et al.
Veröffentlicht: (2024)
Curriculum Imitation Learning of Distributed Multi-Robot Policies
von: Roche, Jesús, et al.
Veröffentlicht: (2025)
von: Roche, Jesús, et al.
Veröffentlicht: (2025)
Policy Learning from Large Vision-Language Model Feedback without Reward Modeling
von: Luu, Tung M., et al.
Veröffentlicht: (2025)
von: Luu, Tung M., et al.
Veröffentlicht: (2025)
Force-Modulated Visual Policy for Robot-Assisted Dressing with Arm Motions
von: Hao, Alexis Yihong, et al.
Veröffentlicht: (2025)
von: Hao, Alexis Yihong, et al.
Veröffentlicht: (2025)
Robot Policy Learning with Temporal Optimal Transport Reward
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
von: Fu, Yuwei, et al.
Veröffentlicht: (2024)
ELEMENTAL: Interactive Learning from Demonstrations and Vision-Language Models for Reward Design in Robotics
von: Chen, Letian, et al.
Veröffentlicht: (2024)
von: Chen, Letian, et al.
Veröffentlicht: (2024)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
von: Zargarbashi, Fatemeh, et al.
Veröffentlicht: (2024)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
von: Patel, Bhrij, et al.
Veröffentlicht: (2023)
AdaptManip: Learning Adaptive Whole-Body Object Lifting and Delivery with Online Recurrent State Estimation
von: Byrd, Morgan, et al.
Veröffentlicht: (2026)
von: Byrd, Morgan, et al.
Veröffentlicht: (2026)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
von: Miranda, Victor R. F., et al.
Veröffentlicht: (2022)
von: Miranda, Victor R. F., et al.
Veröffentlicht: (2022)
Using Temperature Sampling to Effectively Train Robot Learning Policies on Imbalanced Datasets
von: Patil, Basavasagar, et al.
Veröffentlicht: (2025)
von: Patil, Basavasagar, et al.
Veröffentlicht: (2025)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
von: Huang, Changxin, et al.
Veröffentlicht: (2024)
Language Guided Skill Discovery
von: Rho, Seungeun, et al.
Veröffentlicht: (2024)
von: Rho, Seungeun, et al.
Veröffentlicht: (2024)
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation
von: Feng, Youhe, et al.
Veröffentlicht: (2026)
von: Feng, Youhe, et al.
Veröffentlicht: (2026)
Reward Machine Inference for Robotic Manipulation
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?
von: Ou, Buqing, et al.
Veröffentlicht: (2026)
von: Ou, Buqing, et al.
Veröffentlicht: (2026)
Accelerating Visual-Policy Learning through Parallel Differentiable Simulation
von: You, Haoxiang, et al.
Veröffentlicht: (2025)
von: You, Haoxiang, et al.
Veröffentlicht: (2025)
Dynamic Policy Learning for Legged Robot with Simplified Model Pretraining and Model Homotopy Transfer
von: Kang, Dongyun, et al.
Veröffentlicht: (2025)
von: Kang, Dongyun, et al.
Veröffentlicht: (2025)
Contrast Sets for Evaluating Language-Guided Robot Policies
von: Anwar, Abrar, et al.
Veröffentlicht: (2024)
von: Anwar, Abrar, et al.
Veröffentlicht: (2024)
Robot Fleet Learning via Policy Merging
von: Wang, Lirui, et al.
Veröffentlicht: (2023)
von: Wang, Lirui, et al.
Veröffentlicht: (2023)
STRIDE: Automating Reward Design, Deep Reinforcement Learning Training and Feedback Optimization in Humanoid Robotics Locomotion
von: Wu, Zhenwei, et al.
Veröffentlicht: (2025)
von: Wu, Zhenwei, et al.
Veröffentlicht: (2025)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
Trajectory First: A Curriculum for Discovering Diverse Policies
von: Braun, Cornelius V., et al.
Veröffentlicht: (2025)
von: Braun, Cornelius V., et al.
Veröffentlicht: (2025)
Learning Dolly-In Filming From Demonstration Using a Ground-Based Robot
von: Lorimer, Philip, et al.
Veröffentlicht: (2025)
von: Lorimer, Philip, et al.
Veröffentlicht: (2025)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
ClevrSkills: Compositional Language and Visual Reasoning in Robotics
von: Haresh, Sanjay, et al.
Veröffentlicht: (2024)
von: Haresh, Sanjay, et al.
Veröffentlicht: (2024)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
Riemannian Flow Matching Policy for Robot Motion Learning
von: Braun, Max, et al.
Veröffentlicht: (2024)
von: Braun, Max, et al.
Veröffentlicht: (2024)
RobotDesignGPT: Automated Robot Design Synthesis using Vision Language Models
von: Sontakke, Nitish, et al.
Veröffentlicht: (2026)
von: Sontakke, Nitish, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PrivilegedDreamer: Explicit Imagination of Privileged Information for Rapid Adaptation of Learned Policies
von: Byrd, Morgan, et al.
Veröffentlicht: (2025) -
Deep Reinforcement Learning for Multi-Agent Coordination
von: Aina, Kehinde O., et al.
Veröffentlicht: (2025) -
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
von: Tang, Nan, et al.
Veröffentlicht: (2025) -
Decoupling Task and Behavior: A Two-Stage Reward Curriculum in Reinforcement Learning for Robotics
von: Freitag, Kilian, et al.
Veröffentlicht: (2026) -
Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion
von: Rho, Seungeun, et al.
Veröffentlicht: (2025)