On-Robot Reinforcement Learning with Goal-Contrastive Rewards
Fuente:
arXiv
Guardado en:
| Autores principales: | Biza, Ondrej, Weng, Thomas, Sun, Lingfeng, Schmeckpeper, Karl, Kelestemur, Tarik, Ma, Yecheng Jason, Platt, Robert, van de Meent, Jan-Willem, Wong, Lawson L. S. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning
por: Schroeder, Philip, et al.
Publicado: (2026)
por: Schroeder, Philip, et al.
Publicado: (2026)
When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning
por: Dodeja, Lakshita, et al.
Publicado: (2026)
por: Dodeja, Lakshita, et al.
Publicado: (2026)
Imagination Policy: Using Generative Point Cloud Models for Learning Manipulation Policies
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
Equivariant Offline Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2024)
por: Tangri, Arsh, et al.
Publicado: (2024)
You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector
por: Patil, Omkar, et al.
Publicado: (2026)
por: Patil, Omkar, et al.
Publicado: (2026)
Theia: Distilling Diverse Vision Foundation Models for Robot Learning
por: Shang, Jinghuan, et al.
Publicado: (2024)
por: Shang, Jinghuan, et al.
Publicado: (2024)
Equivariant Goal Conditioned Contrastive Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2025)
por: Tangri, Arsh, et al.
Publicado: (2025)
One-Shot Cross-Geometry Skill Transfer through Part Decomposition
por: Thompson, Skye, et al.
Publicado: (2026)
por: Thompson, Skye, et al.
Publicado: (2026)
ThinkGrasp: A Vision-Language System for Strategic Part Grasping in Clutter
por: Qian, Yaoyao, et al.
Publicado: (2024)
por: Qian, Yaoyao, et al.
Publicado: (2024)
Accelerating Residual Reinforcement Learning with Uncertainty Estimation
por: Dodeja, Lakshita, et al.
Publicado: (2025)
por: Dodeja, Lakshita, et al.
Publicado: (2025)
ROVER: Recursive Reasoning Over Videos with Vision-Language Models for Embodied Tasks
por: Schroeder, Philip, et al.
Publicado: (2025)
por: Schroeder, Philip, et al.
Publicado: (2025)
Equivariant Diffusion Policy
por: Wang, Dian, et al.
Publicado: (2024)
por: Wang, Dian, et al.
Publicado: (2024)
ACDC: Adaptive Curriculum Planning with Dynamic Contrastive Control for Goal-Conditioned Reinforcement Learning in Robotic Manipulation
por: Wang, Xuerui, et al.
Publicado: (2026)
por: Wang, Xuerui, et al.
Publicado: (2026)
Robot Body Schema Learning from Full-body Extero/Proprioception Sensors
por: Jiang, Shuo, et al.
Publicado: (2024)
por: Jiang, Shuo, et al.
Publicado: (2024)
CuriousBot: Interactive Mobile Exploration via Actionable 3D Relational Object Graph
por: Wang, Yixuan, et al.
Publicado: (2025)
por: Wang, Yixuan, et al.
Publicado: (2025)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
por: Vasan, Gautham, et al.
Publicado: (2024)
por: Vasan, Gautham, et al.
Publicado: (2024)
Real-is-Sim: Bridging the Sim-to-Real Gap with a Dynamic Digital Twin
por: Abou-Chakra, Jad, et al.
Publicado: (2025)
por: Abou-Chakra, Jad, et al.
Publicado: (2025)
Sceniris: A Fast Procedural Scene Generation Framework
por: Shang, Jinghuan, et al.
Publicado: (2025)
por: Shang, Jinghuan, et al.
Publicado: (2025)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
por: Shi, Junyao, et al.
Publicado: (2024)
por: Shi, Junyao, et al.
Publicado: (2024)
RoboReward: General-Purpose Vision-Language Reward Models for Robotics
por: Lee, Tony, et al.
Publicado: (2026)
por: Lee, Tony, et al.
Publicado: (2026)
Reinforced Embodied Planning with Verifiable Reward for Real-World Robotic Manipulation
por: Bo, Zitong, et al.
Publicado: (2025)
por: Bo, Zitong, et al.
Publicado: (2025)
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy
por: Wang, Yixuan, et al.
Publicado: (2024)
por: Wang, Yixuan, et al.
Publicado: (2024)
Robot Tactile Gesture Recognition Based on Full-body Modular E-skin
por: Jiang, Shuo, et al.
Publicado: (2025)
por: Jiang, Shuo, et al.
Publicado: (2025)
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
por: Nakamura, Issa, et al.
Publicado: (2026)
por: Nakamura, Issa, et al.
Publicado: (2026)
Reward Training Wheels: Adaptive Auxiliary Rewards for Robotics Reinforcement Learning
por: Wang, Linji, et al.
Publicado: (2025)
por: Wang, Linji, et al.
Publicado: (2025)
Eureka: Human-Level Reward Design via Coding Large Language Models
por: Ma, Yecheng Jason, et al.
Publicado: (2023)
por: Ma, Yecheng Jason, et al.
Publicado: (2023)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
por: Venugopal, Aravind, et al.
Publicado: (2026)
por: Venugopal, Aravind, et al.
Publicado: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
por: Ishihara, Yu, et al.
Publicado: (2025)
por: Ishihara, Yu, et al.
Publicado: (2025)
Reinforcement Learning Goal-Reaching Control with Guaranteed Lyapunov-Like Stabilizer for Mobile Robots
por: Shahna, Mehdi Heydari, et al.
Publicado: (2026)
por: Shahna, Mehdi Heydari, et al.
Publicado: (2026)
EquAct: An SE(3)-Equivariant Multi-Task Transformer for Open-Loop Robotic Manipulation
por: Zhu, Xupeng, et al.
Publicado: (2025)
por: Zhu, Xupeng, et al.
Publicado: (2025)
Keyframe-Guided Structured Rewards for Reinforcement Learning in Long-Horizon Laboratory Robotics
por: Qiu, Yibo, et al.
Publicado: (2026)
por: Qiu, Yibo, et al.
Publicado: (2026)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
por: Miranda, Victor R. F., et al.
Publicado: (2022)
por: Miranda, Victor R. F., et al.
Publicado: (2022)
Hierarchical Reinforcement Learning in Multi-Goal Spatial Navigation with Autonomous Mobile Robots
por: Johnson, Brendon, et al.
Publicado: (2025)
por: Johnson, Brendon, et al.
Publicado: (2025)
Robotic Skill Diversification via Active Mutation of Reward Functions in Reinforcement Learning During a Liquid Pouring Task
por: van Buuren, Jannick, et al.
Publicado: (2025)
por: van Buuren, Jannick, et al.
Publicado: (2025)
Training People to Reward Robots
por: Sun, Endong, et al.
Publicado: (2025)
por: Sun, Endong, et al.
Publicado: (2025)
Gen-Drive: Enhancing Diffusion Generative Driving Policies with Reward Modeling and Reinforcement Learning Fine-tuning
por: Huang, Zhiyu, et al.
Publicado: (2024)
por: Huang, Zhiyu, et al.
Publicado: (2024)
Active Embodiment Identification with Reinforcement Learning for Legged Robots
por: Bohlinger, Nico, et al.
Publicado: (2026)
por: Bohlinger, Nico, et al.
Publicado: (2026)
Goal-Oriented End-User Programming of Robots
por: Porfirio, David, et al.
Publicado: (2024)
por: Porfirio, David, et al.
Publicado: (2024)
LiPS: Large-Scale Humanoid Robot Reinforcement Learning with Parallel-Series Structures
por: Zhang, Qiang, et al.
Publicado: (2025)
por: Zhang, Qiang, et al.
Publicado: (2025)
Fourier Transporter: Bi-Equivariant Robotic Manipulation in 3D
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
Ejemplares similares
-
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning
por: Schroeder, Philip, et al.
Publicado: (2026) -
When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning
por: Dodeja, Lakshita, et al.
Publicado: (2026) -
Imagination Policy: Using Generative Point Cloud Models for Learning Manipulation Policies
por: Huang, Haojie, et al.
Publicado: (2024) -
Equivariant Offline Reinforcement Learning
por: Tangri, Arsh, et al.
Publicado: (2024) -
You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector
por: Patil, Omkar, et al.
Publicado: (2026)