Video-Language Critic: Transferable Reward Functions for Language-Conditioned Robotics
Fuente:
arXiv
Saved in:
| Main Authors: | Alakuijala, Minttu, McLean, Reginald, Woungang, Isaac, Farsad, Nariman, Kaski, Samuel, Marttinen, Pekka, Yuan, Kai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Task Reinforcement Learning Enables Parameter Scaling
by: McLean, Reginald, et al.
Published: (2025)
by: McLean, Reginald, et al.
Published: (2025)
Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search
by: Dainese, Nicola, et al.
Published: (2024)
by: Dainese, Nicola, et al.
Published: (2024)
Memento No More: Coaching AI Agents to Master Multiple Tasks via Hints Internalization
by: Alakuijala, Minttu, et al.
Published: (2025)
by: Alakuijala, Minttu, et al.
Published: (2025)
Recursive Decomposition with Dependencies for Generic Divide-and-Conquer Reasoning
by: Hernández-Gutiérrez, Sergio, et al.
Published: (2025)
by: Hernández-Gutiérrez, Sergio, et al.
Published: (2025)
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning
by: Schroeder, Philip, et al.
Published: (2026)
by: Schroeder, Philip, et al.
Published: (2026)
Adapt2Reward: Adapting Video-Language Models to Generalizable Robotic Rewards via Failure Prompts
by: Yang, Yanting, et al.
Published: (2024)
by: Yang, Yanting, et al.
Published: (2024)
Video2Reward: Generating Reward Function from Videos for Legged Robot Behavior Learning
by: Zeng, Runhao, et al.
Published: (2024)
by: Zeng, Runhao, et al.
Published: (2024)
ViPlan: A Benchmark for Visual Planning with Symbolic Predicates and Vision-Language Models
by: Merler, Matteo, et al.
Published: (2025)
by: Merler, Matteo, et al.
Published: (2025)
Improving Generalization of Language-Conditioned Robot Manipulation
by: Cui, Chenglin, et al.
Published: (2025)
by: Cui, Chenglin, et al.
Published: (2025)
In-Context Symbolic Regression: Leveraging Large Language Models for Function Discovery
by: Merler, Matteo, et al.
Published: (2024)
by: Merler, Matteo, et al.
Published: (2024)
Diffusion Reward: Learning Rewards via Conditional Video Diffusion
by: Huang, Tao, et al.
Published: (2023)
by: Huang, Tao, et al.
Published: (2023)
Language-Conditioned Robotic Manipulation with Fast and Slow Thinking
by: Zhu, Minjie, et al.
Published: (2024)
by: Zhu, Minjie, et al.
Published: (2024)
Leveraging Large Language Models in Human-Robot Interaction: A Critical Analysis of Potential and Pitfalls
by: Atuhurra, Jesse
Published: (2024)
by: Atuhurra, Jesse
Published: (2024)
Generating and Evolving Reward Functions for Highway Driving with Large Language Models
by: Han, Xu, et al.
Published: (2024)
by: Han, Xu, et al.
Published: (2024)
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks
by: Zhang, Shiduo, et al.
Published: (2024)
by: Zhang, Shiduo, et al.
Published: (2024)
Online Learning of Human Constraints from Feedback in Shared Autonomy
by: Zhu, Shibei, et al.
Published: (2024)
by: Zhu, Shibei, et al.
Published: (2024)
This&That: Language-Gesture Controlled Video Generation for Robot Planning
by: Wang, Boyang, et al.
Published: (2024)
by: Wang, Boyang, et al.
Published: (2024)
Strategies for Robust Deep Learning Based Deformable Registration
by: Honkamaa, Joel, et al.
Published: (2025)
by: Honkamaa, Joel, et al.
Published: (2025)
New multimodal similarity measure for image registration via modeling local functional dependence with linear combination of learned basis functions
by: Honkamaa, Joel, et al.
Published: (2025)
by: Honkamaa, Joel, et al.
Published: (2025)
SITReg: Multi-resolution architecture for symmetric, inverse consistent, and topology preserving image registration
by: Honkamaa, Joel, et al.
Published: (2023)
by: Honkamaa, Joel, et al.
Published: (2023)
Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data Augmentation
by: Hui, Chenyu, et al.
Published: (2026)
by: Hui, Chenyu, et al.
Published: (2026)
Evaluation of Habitat Robotics using Large Language Models
by: Li, William, et al.
Published: (2025)
by: Li, William, et al.
Published: (2025)
Improving Medical Multi-modal Contrastive Learning with Expert Annotations
by: Kumar, Yogesh, et al.
Published: (2024)
by: Kumar, Yogesh, et al.
Published: (2024)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
by: Xie, Tianbao, et al.
Published: (2023)
by: Xie, Tianbao, et al.
Published: (2023)
Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing
by: Li, Zhiyuan, et al.
Published: (2026)
by: Li, Zhiyuan, et al.
Published: (2026)
IROSA: Interactive Robot Skill Adaptation using Natural Language
by: Knauer, Markus, et al.
Published: (2026)
by: Knauer, Markus, et al.
Published: (2026)
EveryDayVLA: A Vision-Language-Action Model for Affordable Robotic Manipulation
by: Chopra, Samarth, et al.
Published: (2025)
by: Chopra, Samarth, et al.
Published: (2025)
Meta-World+: An Improved, Standardized, RL Benchmark
by: McLean, Reginald, et al.
Published: (2025)
by: McLean, Reginald, et al.
Published: (2025)
Device-Conditioned Neural Architecture Search for Efficient Robotic Manipulation
by: Wu, Yiming, et al.
Published: (2026)
by: Wu, Yiming, et al.
Published: (2026)
Benchmarking Local Language Models for Social Robots using Edge Devices
by: Lamouille, Dorian, et al.
Published: (2026)
by: Lamouille, Dorian, et al.
Published: (2026)
Using the Pepper Robot to Support Sign Language Communication
by: Botta, Giulia, et al.
Published: (2025)
by: Botta, Giulia, et al.
Published: (2025)
Critical Insights about Robots for Mental Wellbeing
by: Laban, Guy, et al.
Published: (2025)
by: Laban, Guy, et al.
Published: (2025)
Robot Tasks with Fuzzy Time Requirements from Natural Language Instructions
by: Sucker, Sascha, et al.
Published: (2024)
by: Sucker, Sascha, et al.
Published: (2024)
Towards Multimodal Social Conversations with Robots: Using Vision-Language Models
by: Janssens, Ruben, et al.
Published: (2025)
by: Janssens, Ruben, et al.
Published: (2025)
Timing the Message: Language-Based Notifications for Time-Critical Assistive Settings
by: Hsu, Ya-Chuan, et al.
Published: (2025)
by: Hsu, Ya-Chuan, et al.
Published: (2025)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
by: Nikitin, Alexander, et al.
Published: (2024)
by: Nikitin, Alexander, et al.
Published: (2024)
LGR2: Language Guided Reward Relabeling for Accelerating Hierarchical Reinforcement Learning
by: Singh, Utsav, et al.
Published: (2024)
by: Singh, Utsav, et al.
Published: (2024)
Manipulate-Anything: Automating Real-World Robots using Vision-Language Models
by: Duan, Jiafei, et al.
Published: (2024)
by: Duan, Jiafei, et al.
Published: (2024)
Precise Robot Command Understanding Using Grammar-Constrained Large Language Models
by: Huo, Xinyun, et al.
Published: (2026)
by: Huo, Xinyun, et al.
Published: (2026)
Towards Natural Language Environment: Understanding Seamless Natural-Language-Based Human-Multi-Robot Interactions
by: Liu, Ziyi, et al.
Published: (2026)
by: Liu, Ziyi, et al.
Published: (2026)
Similar Items
-
Multi-Task Reinforcement Learning Enables Parameter Scaling
by: McLean, Reginald, et al.
Published: (2025) -
Generating Code World Models with Large Language Models Guided by Monte Carlo Tree Search
by: Dainese, Nicola, et al.
Published: (2024) -
Memento No More: Coaching AI Agents to Master Multiple Tasks via Hints Internalization
by: Alakuijala, Minttu, et al.
Published: (2025) -
Recursive Decomposition with Dependencies for Generic Divide-and-Conquer Reasoning
by: Hernández-Gutiérrez, Sergio, et al.
Published: (2025) -
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning
by: Schroeder, Philip, et al.
Published: (2026)