VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
Fuente:
arXiv
Saved in:
| Main Authors: | Singh, Anukriti, Bhaskar, Amisha, Yu, Peihong, Chakraborty, Souradip, Dasyam, Ruthwik, Bedi, Amrit, Tokekar, Pratap |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sketch-to-Skill: Bootstrapping Robot Learning with Human Drawn Trajectory Sketches
by: Yu, Peihong, et al.
Published: (2025)
by: Yu, Peihong, et al.
Published: (2025)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
LAVA: Long-horizon Visual Action based Food Acquisition
by: Bhaskar, Amisha, et al.
Published: (2024)
by: Bhaskar, Amisha, et al.
Published: (2024)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
by: Barakat, Anas, et al.
Published: (2024)
by: Barakat, Anas, et al.
Published: (2024)
Older Adults' Preferences for Feedback Cadence from an Exercise Coach Robot
by: Kaushik, Roshni, et al.
Published: (2026)
by: Kaushik, Roshni, et al.
Published: (2026)
Beyond Joint Demonstrations: Personalized Expert Guidance for Efficient Multi-Agent Reinforcement Learning
by: Yu, Peihong, et al.
Published: (2024)
by: Yu, Peihong, et al.
Published: (2024)
The Impact of VR and 2D Interfaces on Human Feedback in Preference-Based Robot Learning
by: de Heuvel, Jorge, et al.
Published: (2025)
by: de Heuvel, Jorge, et al.
Published: (2025)
Adaptive Visual Imitation Learning for Robotic Assisted Feeding Across Varied Bowl Configurations and Food Types
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
From "Thumbs Up" to "10 out of 10": Reconsidering Scalar Feedback in Interactive Reinforcement Learning
by: Yu, Hang, et al.
Published: (2023)
by: Yu, Hang, et al.
Published: (2023)
Low-Burden LLM-Based Preference Learning: Personalizing Assistive Robots from Natural Language Feedback for Users with Paralysis
by: Shankar, Keshav, et al.
Published: (2026)
by: Shankar, Keshav, et al.
Published: (2026)
IMRL: Integrating Visual, Physical, Temporal, and Geometric Representations for Enhanced Food Acquisition
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
SlicerRoboTMS: An Open-Source 3D Slicer Extension for Robot-Assisted Transcranial Magnetic Stimulation
by: Bai, Wenzhi, et al.
Published: (2026)
by: Bai, Wenzhi, et al.
Published: (2026)
Open-TeleVision: Teleoperation with Immersive Active Visual Feedback
by: Cheng, Xuxin, et al.
Published: (2024)
by: Cheng, Xuxin, et al.
Published: (2024)
Pre-Trained Masked Image Model for Mobile Robot Navigation
by: Sharma, Vishnu Dutt, et al.
Published: (2023)
by: Sharma, Vishnu Dutt, et al.
Published: (2023)
Vision-Language System using Open-Source LLMs for Gestures in Medical Interpreter Robots
by: Ngo, Thanh-Tung, et al.
Published: (2026)
by: Ngo, Thanh-Tung, et al.
Published: (2026)
Use of Winsome Robots for Understanding Human Feedback (UWU)
by: Eggers, Jessica, et al.
Published: (2025)
by: Eggers, Jessica, et al.
Published: (2025)
Alternative Interfaces for Human-initiated Natural Language Communication and Robot-initiated Haptic Feedback: Towards Better Situational Awareness in Human-Robot Collaboration
by: Bennie, Callum, et al.
Published: (2024)
by: Bennie, Callum, et al.
Published: (2024)
CHARM: Considering Human Attributes for Reinforcement Modeling
by: Fang, Qidi, et al.
Published: (2025)
by: Fang, Qidi, et al.
Published: (2025)
PRISM: Performer RS-IMLE for Single-pass Multisensory Imitation Learning
by: Bhaskar, Amisha, et al.
Published: (2026)
by: Bhaskar, Amisha, et al.
Published: (2026)
Design and Integration of Thermal and Vibrotactile Feedback for Lifelike Touch in Social Robots
by: Borgstedt, Jacqueline, et al.
Published: (2025)
by: Borgstedt, Jacqueline, et al.
Published: (2025)
GuideTouch: An Obstacle Avoidance Device with Tactile Feedback for Visually Impaired
by: Kozlov, Timofei, et al.
Published: (2026)
by: Kozlov, Timofei, et al.
Published: (2026)
Beyond Visuals: Investigating Force Feedback in Extended Reality for Robot Data Collection
by: Li, Xueyin, et al.
Published: (2025)
by: Li, Xueyin, et al.
Published: (2025)
EnchantedClothes: Visual and Tactile Feedback with an Abdomen-Attached Robot through Clothes
by: Yamamoto, Takumi, et al.
Published: (2024)
by: Yamamoto, Takumi, et al.
Published: (2024)
SoftNash: Entropy-Regularized Nash Games for Non-Fighting Virtual Fixtures
by: Inui, Tai, et al.
Published: (2025)
by: Inui, Tai, et al.
Published: (2025)
From Agent Autonomy to Casual Collaboration: A Design Investigation on Help-Seeking Urban Robots
by: Yu, Xinyan, et al.
Published: (2024)
by: Yu, Xinyan, et al.
Published: (2024)
RoboBlockly Studio: Conversational Block Programming with Embodied Robot Feedback for Computational Thinking
by: Li, Leyi, et al.
Published: (2026)
by: Li, Leyi, et al.
Published: (2026)
Effect of Haptic Feedback on Avoidance Behavior and Visual Exploration in Dynamic VR Pedestrian Environment
by: Ishibashi, Kyosuke, et al.
Published: (2025)
by: Ishibashi, Kyosuke, et al.
Published: (2025)
Vibrotactile Feedback for a Remote Operated Robot with Noise Subtraction Based on Perceived Intensity
by: Yamawaki, Ryoma, et al.
Published: (2024)
by: Yamawaki, Ryoma, et al.
Published: (2024)
Preserving Sense of Agency: User Preferences for Robot Autonomy and User Control across Household Tasks
by: Yang, Claire, et al.
Published: (2025)
by: Yang, Claire, et al.
Published: (2025)
Do Looks Matter? Exploring Functional and Aesthetic Design Preferences for a Robotic Guide Dog
by: Cohav, Aviv L., et al.
Published: (2025)
by: Cohav, Aviv L., et al.
Published: (2025)
REACT: Two Datasets for Analyzing Both Human Reactions and Evaluative Feedback to Robots Over Time
by: Candon, Kate, et al.
Published: (2024)
by: Candon, Kate, et al.
Published: (2024)
LLM-Glasses: GenAI-driven Glasses with Haptic Feedback for Navigation of Visually Impaired People
by: Tokmurziyev, Issatay, et al.
Published: (2025)
by: Tokmurziyev, Issatay, et al.
Published: (2025)
AeroHaptix: A Wearable Vibrotactile Feedback System for Enhancing Collision Avoidance in UAV Teleoperation
by: Huang, Bingjian, et al.
Published: (2024)
by: Huang, Bingjian, et al.
Published: (2024)
Automated Assessment and Adaptive Multimodal Formative Feedback Improves Psychomotor Skills Training Outcomes in Quadrotor Teleoperation
by: Jensen, Emily, et al.
Published: (2024)
by: Jensen, Emily, et al.
Published: (2024)
Vision Beyond Boundaries: An Initial Design Space of Domain-specific Large Vision Models in Human-robot Interaction
by: Zhang, Yuchong, et al.
Published: (2024)
by: Zhang, Yuchong, et al.
Published: (2024)
RFM-HRI : A Multimodal Dataset of Medical Robot Failure, User Reaction and Recovery Preferences for Item Retrieval Tasks
by: Batra, Yashika, et al.
Published: (2026)
by: Batra, Yashika, et al.
Published: (2026)
Indicating Robot Vision Capabilities with Augmented Reality
by: Wang, Hong, et al.
Published: (2025)
by: Wang, Hong, et al.
Published: (2025)
Effects of Wrist-Worn Haptic Feedback on Force Accuracy and Task Speed during a Teleoperated Robotic Surgery Task
by: Vuong, Brian B., et al.
Published: (2025)
by: Vuong, Brian B., et al.
Published: (2025)
"It's like a pet...but my pet doesn't collect data about me": Multi-person Households' Privacy Design Preferences for Household Robots
by: Li, Jennica, et al.
Published: (2026)
by: Li, Jennica, et al.
Published: (2026)
Model of Spatial Human-Agent Interaction with Consideration for Others
by: Sakamoto, Takafumi, et al.
Published: (2026)
by: Sakamoto, Takafumi, et al.
Published: (2026)
Similar Items
-
Sketch-to-Skill: Bootstrapping Robot Learning with Human Drawn Trajectory Sketches
by: Yu, Peihong, et al.
Published: (2025) -
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023) -
LAVA: Long-horizon Visual Action based Food Acquisition
by: Bhaskar, Amisha, et al.
Published: (2024) -
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
by: Barakat, Anas, et al.
Published: (2024) -
Older Adults' Preferences for Feedback Cadence from an Exercise Coach Robot
by: Kaushik, Roshni, et al.
Published: (2026)