Real-Time Human-Robot Interaction Intent Detection Using RGB-based Pose and Emotion Cues with Cross-Camera Model Generalization
Fuente:
arXiv
Saved in:
| Main Authors: | Mohsen, Farida, Safa, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MINT-RVAE: Multi-Cues Intention Prediction of Human-Robot Interaction using Human Pose and Emotion Information from RGB-only Camera Data
by: Mohsen, Farida, et al.
Published: (2025)
by: Mohsen, Farida, et al.
Published: (2025)
Deep Fusion of Ultra-Low-Resolution Thermal Camera and Gyroscope Data for Lighting-Robust and Compute-Efficient Rotational Odometry
by: Mohsen, Farida, et al.
Published: (2025)
by: Mohsen, Farida, et al.
Published: (2025)
Learning to Control Dynamical Agents via Spiking Neural Networks and Metropolis-Hastings Sampling
by: Safa, Ali, et al.
Published: (2025)
by: Safa, Ali, et al.
Published: (2025)
On the Importance of Neural Membrane Potential Leakage for LIDAR-based Robot Obstacle Avoidance using Spiking Neural Networks
by: Ali, Zainab, et al.
Published: (2025)
by: Ali, Zainab, et al.
Published: (2025)
Task-Oriented Edge-Assisted Cross-System Design for Real-Time Human-Robot Interaction in Industrial Metaverse
by: Chen, Kan, et al.
Published: (2025)
by: Chen, Kan, et al.
Published: (2025)
Real-Time Imitation of Human Head Motions, Blinks and Emotions by Nao Robot: A Closed-Loop Approach
by: Rayati, Keyhan, et al.
Published: (2025)
by: Rayati, Keyhan, et al.
Published: (2025)
Navigating the Human Maze: Real-Time Robot Pathfinding with Generative Imitation Learning
by: Moder, Martin, et al.
Published: (2024)
by: Moder, Martin, et al.
Published: (2024)
Vision-Language Models on the Edge for Real-Time Robotic Perception
by: Ahmad, Sarat, et al.
Published: (2026)
by: Ahmad, Sarat, et al.
Published: (2026)
DTRT: Enhancing Human Intent Estimation and Role Allocation for Physical Human-Robot Collaboration
by: Liu, Haotian, et al.
Published: (2025)
by: Liu, Haotian, et al.
Published: (2025)
Attention-based Estimation and Prediction of Human Intent to augment Haptic Glove aided Control of Robotic Hand
by: Ahmed, Muneeb, et al.
Published: (2021)
by: Ahmed, Muneeb, et al.
Published: (2021)
Crossing the Human-Robot Embodiment Gap with Sim-to-Real RL using One Human Demonstration
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
by: Lum, Tyler Ga Wei, et al.
Published: (2025)
BIMCaP: BIM-based AI-supported LiDAR-Camera Pose Refinement
by: Torres, Miguel Arturo Vega, et al.
Published: (2024)
by: Torres, Miguel Arturo Vega, et al.
Published: (2024)
Experimental Evaluation of ROS-Causal in Real-World Human-Robot Spatial Interaction Scenarios
by: Castri, Luca, et al.
Published: (2024)
by: Castri, Luca, et al.
Published: (2024)
Predicting the Intention to Interact with a Service Robot:the Role of Gaze Cues
by: Arreghini, Simone, et al.
Published: (2024)
by: Arreghini, Simone, et al.
Published: (2024)
ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
by: Chen, Jason, et al.
Published: (2025)
by: Chen, Jason, et al.
Published: (2025)
Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents
by: Nandkumar, Chandran, et al.
Published: (2024)
by: Nandkumar, Chandran, et al.
Published: (2024)
Airy: Reading Robot Intent through Height and Sky
by: Chen, Baoyang, et al.
Published: (2025)
by: Chen, Baoyang, et al.
Published: (2025)
RMS: Redundancy-Minimizing Point Cloud Sampling for Real-Time Pose Estimation
by: Petracek, Pavel, et al.
Published: (2023)
by: Petracek, Pavel, et al.
Published: (2023)
Theory of Mind for Explainable Human-Robot Interaction
by: Bauer, Marie S., et al.
Published: (2025)
by: Bauer, Marie S., et al.
Published: (2025)
PoseDiff: A Unified Diffusion Model Bridging Robot Pose Estimation and Video-to-Action Control
by: Zhang, Haozhuo, et al.
Published: (2025)
by: Zhang, Haozhuo, et al.
Published: (2025)
A Graph-to-Text Approach to Knowledge-Grounded Response Generation in Human-Robot Interaction
by: Walker, Nicholas Thomas, et al.
Published: (2023)
by: Walker, Nicholas Thomas, et al.
Published: (2023)
Using Natural Language for Human-Robot Collaboration in the Real World
by: Lindes, Peter, et al.
Published: (2025)
by: Lindes, Peter, et al.
Published: (2025)
Object Pose Estimation by Camera Arm Control Based on the Next Viewpoint Estimation
by: Mizuno, Tomoki, et al.
Published: (2025)
by: Mizuno, Tomoki, et al.
Published: (2025)
TextOp: Real-time Interactive Text-Driven Humanoid Robot Motion Generation and Control
by: Xie, Weiji, et al.
Published: (2026)
by: Xie, Weiji, et al.
Published: (2026)
An Approach to Combining Video and Speech with Large Language Models in Human-Robot Interaction
by: Shen, Guanting, et al.
Published: (2026)
by: Shen, Guanting, et al.
Published: (2026)
NVP-HRI: Zero Shot Natural Voice and Posture-based Human-Robot Interaction via Large Language Model
by: Lai, Yuzhi, et al.
Published: (2025)
by: Lai, Yuzhi, et al.
Published: (2025)
Gaze-Aware Task Progression Detection Framework for Human-Robot Interaction Using RGB Cameras
by: Cheng, Linlin, et al.
Published: (2026)
by: Cheng, Linlin, et al.
Published: (2026)
Real-Time Obstacle Avoidance for a Mobile Robot Using CNN-Based Sensor Fusion
by: Zain, Lamiaa H.
Published: (2025)
by: Zain, Lamiaa H.
Published: (2025)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
by: Kedia, Kushal, et al.
Published: (2023)
by: Kedia, Kushal, et al.
Published: (2023)
Autonomous Human-Robot Interaction via Operator Imitation
by: Christen, Sammy, et al.
Published: (2025)
by: Christen, Sammy, et al.
Published: (2025)
ROS-Causal: A ROS-based Causal Analysis Framework for Human-Robot Interaction Applications
by: Castri, Luca, et al.
Published: (2024)
by: Castri, Luca, et al.
Published: (2024)
A Unified Framework for Real-Time Failure Handling in Robotics Using Vision-Language Models, Reactive Planner and Behavior Trees
by: Ahmad, Faseeh, et al.
Published: (2025)
by: Ahmad, Faseeh, et al.
Published: (2025)
V-CAS: A Realtime Vehicle Anti Collision System Using Vision Transformer on Multi-Camera Streams
by: Ashraf, Muhammad Waqas, et al.
Published: (2024)
by: Ashraf, Muhammad Waqas, et al.
Published: (2024)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
Ablation Study of Multimodal Perception, Language Grounding, and Control for Human-Robot Interaction in an Object Detection and Grasping Task
by: Tian, Zi, et al.
Published: (2026)
by: Tian, Zi, et al.
Published: (2026)
Human-Humanoid Robots Cross-Embodiment Behavior-Skill Transfer Using Decomposed Adversarial Learning from Demonstration
by: Liu, Junjia, et al.
Published: (2024)
by: Liu, Junjia, et al.
Published: (2024)
Robots Can Multitask Too: Integrating a Memory Architecture and LLMs for Enhanced Cross-Task Robot Action Generation
by: Ali, Hassan, et al.
Published: (2024)
by: Ali, Hassan, et al.
Published: (2024)
Embodied Human Simulation for Quantitative Design and Analysis of Interactive Robotics
by: Zuo, Chenhui, et al.
Published: (2026)
by: Zuo, Chenhui, et al.
Published: (2026)
ROSGPT_Vision: Commanding Robots Using Only Language Models' Prompts
by: Benjdira, Bilel, et al.
Published: (2023)
by: Benjdira, Bilel, et al.
Published: (2023)
Beyond Text: Utilizing Vocal Cues to Improve Decision Making in LLMs for Robot Navigation Tasks
by: Sun, Xingpeng, et al.
Published: (2024)
by: Sun, Xingpeng, et al.
Published: (2024)
Similar Items
-
MINT-RVAE: Multi-Cues Intention Prediction of Human-Robot Interaction using Human Pose and Emotion Information from RGB-only Camera Data
by: Mohsen, Farida, et al.
Published: (2025) -
Deep Fusion of Ultra-Low-Resolution Thermal Camera and Gyroscope Data for Lighting-Robust and Compute-Efficient Rotational Odometry
by: Mohsen, Farida, et al.
Published: (2025) -
Learning to Control Dynamical Agents via Spiking Neural Networks and Metropolis-Hastings Sampling
by: Safa, Ali, et al.
Published: (2025) -
On the Importance of Neural Membrane Potential Leakage for LIDAR-based Robot Obstacle Avoidance using Spiking Neural Networks
by: Ali, Zainab, et al.
Published: (2025) -
Task-Oriented Edge-Assisted Cross-System Design for Real-Time Human-Robot Interaction in Industrial Metaverse
by: Chen, Kan, et al.
Published: (2025)