Agreeing to Interact in Human-Robot Interaction using Large Language Models and Vision Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Sasabuchi, Kazuhiro, Wake, Naoki, Kanehira, Atsushi, Takamatsu, Jun, Ikeuchi, Katsushi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLM-driven Behavior Tree for Context-aware Task Planning
by: Wake, Naoki, et al.
Published: (2025)
by: Wake, Naoki, et al.
Published: (2025)
GPT-4V(ision) for Robotics: Multimodal Task Planning from Human Demonstration
by: Wake, Naoki, et al.
Published: (2023)
by: Wake, Naoki, et al.
Published: (2023)
Plan-and-Act using Large Language Models for Interactive Agreement
by: Sasabuchi, Kazuhiro, et al.
Published: (2025)
by: Sasabuchi, Kazuhiro, et al.
Published: (2025)
Open-Vocabulary Action Localization with Iterative Visual Prompting
by: Wake, Naoki, et al.
Published: (2024)
by: Wake, Naoki, et al.
Published: (2024)
A Taxonomy of Self-Handover
by: Wake, Naoki, et al.
Published: (2025)
by: Wake, Naoki, et al.
Published: (2025)
IK Seed Generator for Dual-Arm Human-like Physicality Robot with Mobile Base
by: Takamatsu, Jun, et al.
Published: (2025)
by: Takamatsu, Jun, et al.
Published: (2025)
RL-Driven Data Generation for Robust Vision-Based Dexterous Grasping
by: Kanehira, Atsushi, et al.
Published: (2025)
by: Kanehira, Atsushi, et al.
Published: (2025)
Large Language Models as Zero-Shot Human Models for Human-Robot Interaction
by: Zhang, Bowen, et al.
Published: (2023)
by: Zhang, Bowen, et al.
Published: (2023)
Modality-Driven Design for Multi-Step Dexterous Manipulation: Insights from Neuroscience
by: Wake, Naoki, et al.
Published: (2024)
by: Wake, Naoki, et al.
Published: (2024)
Designing Library of Skill-Agents for Hardware-Level Reusability
by: Takamatsu, Jun, et al.
Published: (2024)
by: Takamatsu, Jun, et al.
Published: (2024)
TalkWithMachines: Enhancing Human-Robot Interaction for Interpretable Industrial Robotics Through Large/Vision Language Models
by: Abbas, Ammar N., et al.
Published: (2024)
by: Abbas, Ammar N., et al.
Published: (2024)
APriCoT: Action Primitives based on Contact-state Transition for In-Hand Tool Manipulation
by: Saito, Daichi, et al.
Published: (2024)
by: Saito, Daichi, et al.
Published: (2024)
Empathic Grounding: Explorations using Multimodal Interaction and Large Language Models with Conversational Agents
by: Arjmand, Mehdi, et al.
Published: (2024)
by: Arjmand, Mehdi, et al.
Published: (2024)
IROSA: Interactive Robot Skill Adaptation using Natural Language
by: Knauer, Markus, et al.
Published: (2026)
by: Knauer, Markus, et al.
Published: (2026)
Understanding Large-Language Model (LLM)-powered Human-Robot Interaction
by: Kim, Callie Y., et al.
Published: (2024)
by: Kim, Callie Y., et al.
Published: (2024)
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
Gaze-supported Large Language Model Framework for Bi-directional Human-Robot Interaction
by: Rüppel, Jens V., et al.
Published: (2025)
by: Rüppel, Jens V., et al.
Published: (2025)
Towards Multimodal Social Conversations with Robots: Using Vision-Language Models
by: Janssens, Ruben, et al.
Published: (2025)
by: Janssens, Ruben, et al.
Published: (2025)
Child Speech Recognition in Human-Robot Interaction: Problem Solved?
by: Janssens, Ruben, et al.
Published: (2024)
by: Janssens, Ruben, et al.
Published: (2024)
How Can Large Language Models Enable Better Socially Assistive Human-Robot Interaction: A Brief Survey
by: Shi, Zhonghao, et al.
Published: (2024)
by: Shi, Zhonghao, et al.
Published: (2024)
Interactive Task Planning with Language Models
by: Li, Boyi, et al.
Published: (2023)
by: Li, Boyi, et al.
Published: (2023)
A Framework for Adapting Human-Robot Interaction to Diverse User Groups
by: Rosin, Theresa Pekarek, et al.
Published: (2024)
by: Rosin, Theresa Pekarek, et al.
Published: (2024)
Spoken Language Interaction with Robots: Research Issues and Recommendations, Report from the NSF Future Directions Workshop
by: Marge, Matthew, et al.
Published: (2020)
by: Marge, Matthew, et al.
Published: (2020)
How Do We Research Human-Robot Interaction in the Age of Large Language Models? A Systematic Review
by: Wang, Yufeng, et al.
Published: (2026)
by: Wang, Yufeng, et al.
Published: (2026)
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
by: Verma, Mudit, et al.
Published: (2024)
by: Verma, Mudit, et al.
Published: (2024)
Are Large Language Models Aligned with People's Social Intuitions for Human-Robot Interactions?
by: Wachowiak, Lennart, et al.
Published: (2024)
by: Wachowiak, Lennart, et al.
Published: (2024)
Learning Multimodal Latent Dynamics for Human-Robot Interaction
by: Prasad, Vignesh, et al.
Published: (2023)
by: Prasad, Vignesh, et al.
Published: (2023)
Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation
by: Dai, Yinpei, et al.
Published: (2023)
by: Dai, Yinpei, et al.
Published: (2023)
Towards Natural Language Environment: Understanding Seamless Natural-Language-Based Human-Multi-Robot Interactions
by: Liu, Ziyi, et al.
Published: (2026)
by: Liu, Ziyi, et al.
Published: (2026)
Vision Beyond Boundaries: An Initial Design Space of Domain-specific Large Vision Models in Human-robot Interaction
by: Zhang, Yuchong, et al.
Published: (2024)
by: Zhang, Yuchong, et al.
Published: (2024)
USER-VLM 360: Personalized Vision Language Models with User-aware Tuning for Social Human-Robot Interactions
by: Rahimi, Hamed, et al.
Published: (2025)
by: Rahimi, Hamed, et al.
Published: (2025)
Learning Multimodal Confidence for Intention Recognition in Human-Robot Interaction
by: Zhao, Xiyuan, et al.
Published: (2024)
by: Zhao, Xiyuan, et al.
Published: (2024)
Human-Robot Kinaesthetic Interaction Based on Free Energy Principle
by: Sawada, Hiroki, et al.
Published: (2023)
by: Sawada, Hiroki, et al.
Published: (2023)
Integrating Perceptions: A Human-Centered Physical Safety Model for Human-Robot Interaction
by: Pandey, Pranav, et al.
Published: (2025)
by: Pandey, Pranav, et al.
Published: (2025)
Reframing Human-Robot Interaction Through Extended Reality: Unlocking Safer, Smarter, and More Empathic Interactions with Virtual Robots and Foundation Models
by: Zhang, Yuchong, et al.
Published: (2025)
by: Zhang, Yuchong, et al.
Published: (2025)
Designing for Fairness in Human-Robot Interactions
by: Claure, Houston
Published: (2024)
by: Claure, Houston
Published: (2024)
Does the Appearance of Autonomous Conversational Robots Affect User Spoken Behaviors in Real-World Conference Interactions?
by: Pang, Zi Haur, et al.
Published: (2025)
by: Pang, Zi Haur, et al.
Published: (2025)
CARMA: Context-Aware Situational Grounding of Human-Robot Group Interactions by Combining Vision-Language Models with Object and Action Recognition
by: Deigmoeller, Joerg, et al.
Published: (2025)
by: Deigmoeller, Joerg, et al.
Published: (2025)
Enhancing Human-Robot Collaborative Assembly in Manufacturing Systems Using Large Language Models
by: Lim, Jonghan, et al.
Published: (2024)
by: Lim, Jonghan, et al.
Published: (2024)
Quadrupped-Legged Robot Movement Plan Generation using Large Language Model
by: Muhtadin, et al.
Published: (2025)
by: Muhtadin, et al.
Published: (2025)
Similar Items
-
VLM-driven Behavior Tree for Context-aware Task Planning
by: Wake, Naoki, et al.
Published: (2025) -
GPT-4V(ision) for Robotics: Multimodal Task Planning from Human Demonstration
by: Wake, Naoki, et al.
Published: (2023) -
Plan-and-Act using Large Language Models for Interactive Agreement
by: Sasabuchi, Kazuhiro, et al.
Published: (2025) -
Open-Vocabulary Action Localization with Iterative Visual Prompting
by: Wake, Naoki, et al.
Published: (2024) -
A Taxonomy of Self-Handover
by: Wake, Naoki, et al.
Published: (2025)