Video2Reward: Generating Reward Function from Videos for Legged Robot Behavior Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Runhao, Zhou, Dingjie, Liang, Qiwei, Liu, Junlin, Li, Hui, Huang, Changxin, Li, Jianqiang, Hu, Xiping, Sun, Fuchun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
by: Huang, Changxin, et al.
Published: (2024)
by: Huang, Changxin, et al.
Published: (2024)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
by: Huang, Changxin, et al.
Published: (2025)
by: Huang, Changxin, et al.
Published: (2025)
Whole-Body Coordination for Dynamic Object Grasping with Legged Manipulators
by: Liang, Qiwei, et al.
Published: (2025)
by: Liang, Qiwei, et al.
Published: (2025)
Rank2Reward: Learning Shaped Reward Functions from Passive Video
by: Yang, Daniel, et al.
Published: (2024)
by: Yang, Daniel, et al.
Published: (2024)
Learning Task-Invariant Properties via Dreamer: Enabling Efficient Policy Transfer for Quadruped Robots
by: Liang, Junyang, et al.
Published: (2026)
by: Liang, Junyang, et al.
Published: (2026)
Diffusion Reward: Learning Rewards via Conditional Video Diffusion
by: Huang, Tao, et al.
Published: (2023)
by: Huang, Tao, et al.
Published: (2023)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
by: Lee, Tony, et al.
Published: (2026)
by: Lee, Tony, et al.
Published: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
by: Liu, Yuyang, et al.
Published: (2025)
by: Liu, Yuyang, et al.
Published: (2025)
Training People to Reward Robots
by: Sun, Endong, et al.
Published: (2025)
by: Sun, Endong, et al.
Published: (2025)
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
by: Yang, Yanting, et al.
Published: (2024)
by: Yang, Yanting, et al.
Published: (2024)
Video-Language Critic: Transferable Reward Functions for Language-Conditioned Robotics
by: Alakuijala, Minttu, et al.
Published: (2024)
by: Alakuijala, Minttu, et al.
Published: (2024)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
by: Wu, Hao, et al.
Published: (2026)
by: Wu, Hao, et al.
Published: (2026)
Revisiting Cross-Architecture Distillation: Adaptive Dual-Teacher Transfer for Lightweight Video Models
by: Peng, Ying, et al.
Published: (2025)
by: Peng, Ying, et al.
Published: (2025)
MA-ROESL: Motion-aware Rapid Reward Optimization for Efficient Robot Skill Learning from Single Videos
by: Wang, Xianghui, et al.
Published: (2025)
by: Wang, Xianghui, et al.
Published: (2025)
Morphology-Consistent Humanoid Interaction through Robot-Centric Video Synthesis
by: Xu, Weisheng, et al.
Published: (2026)
by: Xu, Weisheng, et al.
Published: (2026)
A Learning Framework for Diverse Legged Robot Locomotion Using Barrier-Based Style Rewards
by: Kim, Gijeong, et al.
Published: (2024)
by: Kim, Gijeong, et al.
Published: (2024)
TeViR: Text-to-Video Reward with Diffusion Models for Efficient Reinforcement Learning
by: Chen, Yuhui, et al.
Published: (2025)
by: Chen, Yuhui, et al.
Published: (2025)
Reward Training Wheels: Adaptive Auxiliary Rewards for Robotics Reinforcement Learning
by: Wang, Linji, et al.
Published: (2025)
by: Wang, Linji, et al.
Published: (2025)
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning
by: Schroeder, Philip, et al.
Published: (2026)
by: Schroeder, Philip, et al.
Published: (2026)
Video-to-BT: Generating Reactive Behavior Trees from Human Demonstration Videos for Robotic Assembly
by: Zhao, Xiwei, et al.
Published: (2025)
by: Zhao, Xiwei, et al.
Published: (2025)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
by: Biza, Ondrej, et al.
Published: (2024)
by: Biza, Ondrej, et al.
Published: (2024)
Learning to Recover: Dynamic Reward Shaping with Wheel-Leg Coordination for Fallen Robots
by: Deng, Boyuan, et al.
Published: (2025)
by: Deng, Boyuan, et al.
Published: (2025)
Not Only Rewards But Also Constraints: Applications on Legged Robot Locomotion
by: Kim, Yunho, et al.
Published: (2023)
by: Kim, Yunho, et al.
Published: (2023)
Keyframe-Guided Structured Rewards for Reinforcement Learning in Long-Horizon Laboratory Robotics
by: Qiu, Yibo, et al.
Published: (2026)
by: Qiu, Yibo, et al.
Published: (2026)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
by: Miranda, Victor R. F., et al.
Published: (2022)
by: Miranda, Victor R. F., et al.
Published: (2022)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
by: Wang, Ruixiang, et al.
Published: (2026)
by: Wang, Ruixiang, et al.
Published: (2026)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
by: Wu, Yanru, et al.
Published: (2026)
by: Wu, Yanru, et al.
Published: (2026)
Decoupling Task and Behavior: A Two-Stage Reward Curriculum in Reinforcement Learning for Robotics
by: Freitag, Kilian, et al.
Published: (2026)
by: Freitag, Kilian, et al.
Published: (2026)
HEADER: Hierarchical Robot Exploration via Attention-Based Deep Reinforcement Learning with Expert-Guided Reward
by: Cao, Yuhong, et al.
Published: (2025)
by: Cao, Yuhong, et al.
Published: (2025)
Video Generators are Robot Policies
by: Liang, Junbang, et al.
Published: (2025)
by: Liang, Junbang, et al.
Published: (2025)
CLIP-Motion: Learning Reward Functions for Robotic Actions Using Consecutive Observations
by: Dang, Xuzhe, et al.
Published: (2023)
by: Dang, Xuzhe, et al.
Published: (2023)
Learning Reward for Robot Skills Using Large Language Models via Self-Alignment
by: Zeng, Yuwei, et al.
Published: (2024)
by: Zeng, Yuwei, et al.
Published: (2024)
SuPLE: Robot Learning with Lyapunov Rewards
by: Nguyen, Phu, et al.
Published: (2024)
by: Nguyen, Phu, et al.
Published: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
by: Tang, Nan, et al.
Published: (2025)
by: Tang, Nan, et al.
Published: (2025)
Design of Reward Function on Reinforcement Learning for Automated Driving
by: Goto, Takeru, et al.
Published: (2025)
by: Goto, Takeru, et al.
Published: (2025)
Task-Oriented Grasping Using Reinforcement Learning with a Contextual Reward Machine
by: Li, Hui, et al.
Published: (2025)
by: Li, Hui, et al.
Published: (2025)
Language-Model-Assisted Bi-Level Programming for Reward Learning from Internet Videos
by: Mahesheka, Harsh, et al.
Published: (2024)
by: Mahesheka, Harsh, et al.
Published: (2024)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
by: Bo, Zitong, et al.
Published: (2025)
by: Bo, Zitong, et al.
Published: (2025)
Experience-Learning Inspired Two-Step Reward Method for Efficient Legged Locomotion Learning Towards Natural and Robust Gaits
by: Li, Yinghui, et al.
Published: (2024)
by: Li, Yinghui, et al.
Published: (2024)
Similar Items
-
Efficient Language-instructed Skill Acquisition via Reward-Policy Co-Evolution
by: Huang, Changxin, et al.
Published: (2024) -
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
by: Huang, Changxin, et al.
Published: (2025) -
Whole-Body Coordination for Dynamic Object Grasping with Legged Manipulators
by: Liang, Qiwei, et al.
Published: (2025) -
Rank2Reward: Learning Shaped Reward Functions from Passive Video
by: Yang, Daniel, et al.
Published: (2024) -
Learning Task-Invariant Properties via Dreamer: Enabling Efficient Policy Transfer for Quadruped Robots
by: Liang, Junyang, et al.
Published: (2026)