Gespeichert in:
| Hauptverfasser: | Yang, Yanting, Chen, Minghao, Qiu, Qibo, Wu, Jiahao, Wang, Wenxiao, Lin, Binbin, Guan, Ziyu, He, Xiaofei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2407.14872 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
Adapting Image-based RL Policies via Predicted Rewards
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
von: Wang, Weiyao, et al.
Veröffentlicht: (2024)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
von: Lee, Tony, et al.
Veröffentlicht: (2026)
von: Lee, Tony, et al.
Veröffentlicht: (2026)
Contrast, Imitate, Adapt: Learning Robotic Skills From Raw Human Videos
von: Qian, Zhifeng, et al.
Veröffentlicht: (2024)
von: Qian, Zhifeng, et al.
Veröffentlicht: (2024)
RoVer: Robot Reward Model as Test-Time Verifier for Vision-Language-Action Model
von: Dai, Mingtong, et al.
Veröffentlicht: (2025)
von: Dai, Mingtong, et al.
Veröffentlicht: (2025)
Generalizable Dense Reward for Long-Horizon Robotic Tasks
von: Yong, Silong, et al.
Veröffentlicht: (2026)
von: Yong, Silong, et al.
Veröffentlicht: (2026)
RoboScape-R: Unified Reward-Observation World Models for Generalizable Robotics Training via RL
von: Tang, Yinzhou, et al.
Veröffentlicht: (2025)
von: Tang, Yinzhou, et al.
Veröffentlicht: (2025)
Learning Reward for Robot Skills Using Large Language Models via Self-Alignment
von: Zeng, Yuwei, et al.
Veröffentlicht: (2024)
von: Zeng, Yuwei, et al.
Veröffentlicht: (2024)
Adapting Robot's Explanation for Failures Based on Observed Human Behavior in Human-Robot Collaboration
von: Naoum, Andreas, et al.
Veröffentlicht: (2025)
von: Naoum, Andreas, et al.
Veröffentlicht: (2025)
Rank2Reward: Learning Shaped Reward Functions from Passive Video
von: Yang, Daniel, et al.
Veröffentlicht: (2024)
von: Yang, Daniel, et al.
Veröffentlicht: (2024)
Video2Reward: Generating Reward Function from Videos for Legged Robot Behavior Learning
von: Zeng, Runhao, et al.
Veröffentlicht: (2024)
von: Zeng, Runhao, et al.
Veröffentlicht: (2024)
From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models
von: Gumbsch, Christian, et al.
Veröffentlicht: (2026)
von: Gumbsch, Christian, et al.
Veröffentlicht: (2026)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
Training People to Reward Robots
von: Sun, Endong, et al.
Veröffentlicht: (2025)
von: Sun, Endong, et al.
Veröffentlicht: (2025)
Diffusion Reward: Learning Rewards via Conditional Video Diffusion
von: Huang, Tao, et al.
Veröffentlicht: (2023)
von: Huang, Tao, et al.
Veröffentlicht: (2023)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
AdaptPNP: Integrating Prehensile and Non-Prehensile Skills for Adaptive Robotic Manipulation
von: Zhu, Jinxuan, et al.
Veröffentlicht: (2025)
von: Zhu, Jinxuan, et al.
Veröffentlicht: (2025)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
von: Wu, Hao, et al.
Veröffentlicht: (2026)
von: Wu, Hao, et al.
Veröffentlicht: (2026)
Constraining Streaming Flow Models for Adapting Learned Robot Trajectory Distributions
von: Long, Jieting, et al.
Veröffentlicht: (2026)
von: Long, Jieting, et al.
Veröffentlicht: (2026)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
von: Huang, Changxin, et al.
Veröffentlicht: (2025)
von: Huang, Changxin, et al.
Veröffentlicht: (2025)
Adapting Neural Robot Dynamics on the Fly for Predictive Control
von: Altawaitan, Abdullah, et al.
Veröffentlicht: (2026)
von: Altawaitan, Abdullah, et al.
Veröffentlicht: (2026)
Learning Generalizable Robot Policy with Human Demonstration Video as a Prompt
von: Zhu, Xiang, et al.
Veröffentlicht: (2025)
von: Zhu, Xiang, et al.
Veröffentlicht: (2025)
SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation
von: Chen, Qianzhong, et al.
Veröffentlicht: (2025)
von: Chen, Qianzhong, et al.
Veröffentlicht: (2025)
GRAPPA: Generalizing and Adapting Robot Policies via Online Agentic Guidance
von: Bucker, Arthur, et al.
Veröffentlicht: (2024)
von: Bucker, Arthur, et al.
Veröffentlicht: (2024)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
Reward Training Wheels: Adaptive Auxiliary Rewards for Robotics Reinforcement Learning
von: Wang, Linji, et al.
Veröffentlicht: (2025)
von: Wang, Linji, et al.
Veröffentlicht: (2025)
Video-Language Critic: Transferable Reward Functions for Language-Conditioned Robotics
von: Alakuijala, Minttu, et al.
Veröffentlicht: (2024)
von: Alakuijala, Minttu, et al.
Veröffentlicht: (2024)
Motion Planning Diffusion: Learning and Adapting Robot Motion Planning with Diffusion Models
von: Carvalho, J., et al.
Veröffentlicht: (2024)
von: Carvalho, J., et al.
Veröffentlicht: (2024)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
Keyframe-Guided Structured Rewards for Reinforcement Learning in Long-Horizon Laboratory Robotics
von: Qiu, Yibo, et al.
Veröffentlicht: (2026)
von: Qiu, Yibo, et al.
Veröffentlicht: (2026)
Automating Robot Failure Recovery Using Vision-Language Models With Optimized Prompts
von: Chen, Hongyi, et al.
Veröffentlicht: (2024)
von: Chen, Hongyi, et al.
Veröffentlicht: (2024)
Reward Machine Inference for Robotic Manipulation
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
von: Baert, Mattijs, et al.
Veröffentlicht: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
von: Tang, Nan, et al.
Veröffentlicht: (2025)
von: Tang, Nan, et al.
Veröffentlicht: (2025)
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning
von: Schroeder, Philip, et al.
Veröffentlicht: (2026)
von: Schroeder, Philip, et al.
Veröffentlicht: (2026)
Generalizable and Actionable Parts Pose Estimation with Symmetry Annotation-Free Learning Strategy
von: Chen, Wenxiao, et al.
Veröffentlicht: (2026)
von: Chen, Wenxiao, et al.
Veröffentlicht: (2026)
Adapt On-the-Go: Behavior Modulation for Single-Life Robot Deployment
von: Chen, Annie S., et al.
Veröffentlicht: (2023)
von: Chen, Annie S., et al.
Veröffentlicht: (2023)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
von: Zhao, Zhikai, et al.
Veröffentlicht: (2026)
von: Zhao, Zhikai, et al.
Veröffentlicht: (2026)
ELEMENTAL: Interactive Learning from Demonstrations and Vision-Language Models for Reward Design in Robotics
von: Chen, Letian, et al.
Veröffentlicht: (2024)
von: Chen, Letian, et al.
Veröffentlicht: (2024)
RoboFAC: A Comprehensive Framework for Robotic Failure Analysis and Correction
von: Ye, Zewei, et al.
Veröffentlicht: (2025)
von: Ye, Zewei, et al.
Veröffentlicht: (2025)
Solving New Tasks by Adapting Internet Video Knowledge
von: Luo, Calvin, et al.
Veröffentlicht: (2025)
von: Luo, Calvin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026) -
Adapting Image-based RL Policies via Predicted Rewards
von: Wang, Weiyao, et al.
Veröffentlicht: (2024) -
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
von: Lee, Tony, et al.
Veröffentlicht: (2026) -
Contrast, Imitate, Adapt: Learning Robotic Skills From Raw Human Videos
von: Qian, Zhifeng, et al.
Veröffentlicht: (2024) -
RoVer: Robot Reward Model as Test-Time Verifier for Vision-Language-Action Model
von: Dai, Mingtong, et al.
Veröffentlicht: (2025)