Task Success is not Enough: Investigating the Use of Video-Language Models as Behavior Critics for Catching Undesirable Agent Behaviors
Fuente:
arXiv
Saved in:
| Main Authors: | Guan, Lin, Zhou, Yifan, Liu, Denis, Zha, Yantian, Amor, Heni Ben, Kambhampati, Subbarao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning from Ambiguous Demonstrations with Self-Explanation Guided Reinforcement Learning
by: Zha, Yantian, et al.
Published: (2021)
by: Zha, Yantian, et al.
Published: (2021)
Enabling Stateful Behaviors for Diffusion-based Policy Learning
by: Liu, Xiao, et al.
Published: (2024)
by: Liu, Xiao, et al.
Published: (2024)
Learning and Blending Robot Hugging Behaviors in Time and Space
by: Drolet, Michael, et al.
Published: (2022)
by: Drolet, Michael, et al.
Published: (2022)
SiSCo: Signal Synthesis for Effective Human-Robot Communication Via Large Language Models
by: Sonawani, Shubham, et al.
Published: (2024)
by: Sonawani, Shubham, et al.
Published: (2024)
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Repairing Neural Networks for Safety in Robotic Systems using Predictive Models
by: Majd, Keyvan, et al.
Published: (2024)
by: Majd, Keyvan, et al.
Published: (2024)
Ubiquitous Robot Control Through Multimodal Motion Capture Using Smartwatch and Smartphone Data
by: Weigend, Fabian C, et al.
Published: (2024)
by: Weigend, Fabian C, et al.
Published: (2024)
RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields
by: Sagar, Som, et al.
Published: (2024)
by: Sagar, Som, et al.
Published: (2024)
SAS-Prompt: Large Language Models as Numerical Optimizers for Robot Self-Improvement
by: Amor, Heni Ben, et al.
Published: (2025)
by: Amor, Heni Ben, et al.
Published: (2025)
Can Large Language Models Reason and Plan?
by: Kambhampati, Subbarao
Published: (2024)
by: Kambhampati, Subbarao
Published: (2024)
TwinTrack: Bridging Vision and Contact Physics for Real-Time Tracking of Unknown Objects in Contact-Rich Scenes
by: Yang, Wen, et al.
Published: (2025)
by: Yang, Wen, et al.
Published: (2025)
Anytime, Anywhere: Human Arm Pose from Smartwatch Data for Ubiquitous Robot Control and Teleoperation
by: Weigend, Fabian C, et al.
Published: (2023)
by: Weigend, Fabian C, et al.
Published: (2023)
Beyond Task Success: Behavioral and Representational Diagnostics for WAM and VLA
by: Mai, Hung, et al.
Published: (2026)
by: Mai, Hung, et al.
Published: (2026)
AAM-SEALS: Developing Aerial-Aquatic Manipulators in SEa, Air, and Land Simulator
by: Yang, William, et al.
Published: (2024)
by: Yang, William, et al.
Published: (2024)
LLM-BT: Performing Robotic Adaptive Tasks based on Large Language Models and Behavior Trees
by: Zhou, Haotian, et al.
Published: (2024)
by: Zhou, Haotian, et al.
Published: (2024)
Can large language models reason and plan?
by: Subbarao Kambhampati
Published: (2024)
by: Subbarao Kambhampati
Published: (2024)
A Comparison of Imitation Learning Algorithms for Bimanual Manipulation
by: Drolet, Michael, et al.
Published: (2024)
by: Drolet, Michael, et al.
Published: (2024)
NatSGLD: A Dataset with Speech, Gesture, Logic, and Demonstration for Robot Learning in Natural Human-Robot Interaction
by: Shrestha, Snehesh, et al.
Published: (2025)
by: Shrestha, Snehesh, et al.
Published: (2025)
NatSGD: A Dataset with Speech, Gestures, and Demonstrations for Robot Learning in Natural Human-Robot Interaction
by: Shrestha, Snehesh, et al.
Published: (2024)
by: Shrestha, Snehesh, et al.
Published: (2024)
On the Self-Verification Limitations of Large Language Models on Reasoning and Planning Tasks
by: Stechly, Kaya, et al.
Published: (2024)
by: Stechly, Kaya, et al.
Published: (2024)
When Context Is Not Enough: Modeling Unexplained Variability in Car-Following Behavior
by: Zhang, Chengyuan, et al.
Published: (2025)
by: Zhang, Chengyuan, et al.
Published: (2025)
iRoCo: Intuitive Robot Control From Anywhere Using a Smartwatch
by: Weigend, Fabian C, et al.
Published: (2024)
by: Weigend, Fabian C, et al.
Published: (2024)
Algorithmic Language Models with Neurally Compiled Libraries
by: Saldyt, Lucas, et al.
Published: (2024)
by: Saldyt, Lucas, et al.
Published: (2024)
Behaviorally Heterogeneous Multi-Agent Exploration Using Distributed Task Allocation
by: Mandal, Nirabhra, et al.
Published: (2025)
by: Mandal, Nirabhra, et al.
Published: (2025)
LLM-based Realistic Safety-Critical Driving Video Generation
by: Fu, Yongjie, et al.
Published: (2025)
by: Fu, Yongjie, et al.
Published: (2025)
Identifying and Extracting Pedestrian Behavior in Critical Traffic Situations
by: Schachner, Martin, et al.
Published: (2024)
by: Schachner, Martin, et al.
Published: (2024)
Static Is Not Enough: A Comparative Study of VR and SpaceMouse in Static and Dynamic Teleoperation Tasks
by: Zhou, Yijun, et al.
Published: (2026)
by: Zhou, Yijun, et al.
Published: (2026)
Multimodal Behavior Tree Generation: A Small Vision-Language Model for Robot Task Planning
by: Battistini, Cristiano, et al.
Published: (2026)
by: Battistini, Cristiano, et al.
Published: (2026)
BaTCAVe: Trustworthy Explanations for Robot Behaviors
by: Sagar, Som, et al.
Published: (2024)
by: Sagar, Som, et al.
Published: (2024)
Safe Multi-Agent Reinforcement Learning for Behavior-Based Cooperative Navigation
by: Dawood, Murad, et al.
Published: (2023)
by: Dawood, Murad, et al.
Published: (2023)
Task Matters: Investigating Human Questioning Behavior in Different Household Service for Learning by Asking Robots
by: Hu, Yuanda, et al.
Published: (2025)
by: Hu, Yuanda, et al.
Published: (2025)
Catch It! Learning to Catch in Flight with Mobile Dexterous Hands
by: Zhang, Yuanhang, et al.
Published: (2024)
by: Zhang, Yuanhang, et al.
Published: (2024)
Versatile Behavior Diffusion for Generalized Traffic Agent Simulation
by: Huang, Zhiyu, et al.
Published: (2024)
by: Huang, Zhiyu, et al.
Published: (2024)
Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation
by: Bhambri, Siddhant, et al.
Published: (2025)
by: Bhambri, Siddhant, et al.
Published: (2025)
BehaviorGPT: Smart Agent Simulation for Autonomous Driving with Next-Patch Prediction
by: Zhou, Zikang, et al.
Published: (2024)
by: Zhou, Zikang, et al.
Published: (2024)
BTGenBot: Behavior Tree Generation for Robotic Tasks with Lightweight LLMs
by: Izzo, Riccardo Andrea, et al.
Published: (2024)
by: Izzo, Riccardo Andrea, et al.
Published: (2024)
Three Dimensional Hydrodynamic Flow-Based Collision Avoidance for UAV Formations Facing Emergent Dynamic Obstacles
by: Sato, Suguru, et al.
Published: (2026)
by: Sato, Suguru, et al.
Published: (2026)
Mitigating Undesired Conditions in Flexible Production with Product-Process-Resource Asset Knowledge Graphs
by: Novak, Petr, et al.
Published: (2025)
by: Novak, Petr, et al.
Published: (2025)
Coordinated Diffusion: Generating Multi-Agent Behavior Without Multi-Agent Demonstrations
by: Peters, Lasse, et al.
Published: (2026)
by: Peters, Lasse, et al.
Published: (2026)
Reasoning Multi-Agent Behavioral Topology for Interactive Autonomous Driving
by: Liu, Haochen, et al.
Published: (2024)
by: Liu, Haochen, et al.
Published: (2024)
Similar Items
-
Learning from Ambiguous Demonstrations with Self-Explanation Guided Reinforcement Learning
by: Zha, Yantian, et al.
Published: (2021) -
Enabling Stateful Behaviors for Diffusion-based Policy Learning
by: Liu, Xiao, et al.
Published: (2024) -
Learning and Blending Robot Hugging Behaviors in Time and Space
by: Drolet, Michael, et al.
Published: (2022) -
SiSCo: Signal Synthesis for Effective Human-Robot Communication Via Large Language Models
by: Sonawani, Shubham, et al.
Published: (2024) -
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
by: Verma, Mudit, et al.
Published: (2024)