A Backbone for Long-Horizon Robot Task Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xiaoshuai, Chen, Wei, Lee, Dongmyoung, Ge, Yukun, Rojas, Nicolas, Kormushev, Petar |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unified Understanding of Environment, Task, and Human for Human-Robot Interaction in Real-World Environments
by: Yano, Yuga, et al.
Published: (2024)
by: Yano, Yuga, et al.
Published: (2024)
Lightweight Structured Multimodal Reasoning for Clinical Scene Understanding in Robotics
by: Jha, Saurav, et al.
Published: (2025)
by: Jha, Saurav, et al.
Published: (2025)
Robot Interaction Behavior Generation based on Social Motion Forecasting for Human-Robot Interaction
by: Mascaro, Esteve Valls, et al.
Published: (2024)
by: Mascaro, Esteve Valls, et al.
Published: (2024)
Synthetic data enables faster annotation and robust segmentation for multi-object grasping in clutter
by: Lee, Dongmyoung, et al.
Published: (2024)
by: Lee, Dongmyoung, et al.
Published: (2024)
iLearnRobot: An Interactive Learning-Based Multi-Modal Robot with Continuous Improvement
by: Wang, Kohou, et al.
Published: (2025)
by: Wang, Kohou, et al.
Published: (2025)
Queryable 3D Scene Representation: A Multi-Modal Framework for Semantic Reasoning and Robotic Task Planning
by: Li, Xun, et al.
Published: (2025)
by: Li, Xun, et al.
Published: (2025)
VLM-driven Behavior Tree for Context-aware Task Planning
by: Wake, Naoki, et al.
Published: (2025)
by: Wake, Naoki, et al.
Published: (2025)
Multi-face emotion detection for effective Human-Robot Interaction
by: Yahyaoui, Mohamed Ala, et al.
Published: (2025)
by: Yahyaoui, Mohamed Ala, et al.
Published: (2025)
DEXOP: A Device for Robotic Transfer of Dexterous Human Manipulation
by: Fang, Hao-Shu, et al.
Published: (2025)
by: Fang, Hao-Shu, et al.
Published: (2025)
Gaze Detection and Analysis for Initiating Joint Activity in Industrial Human-Robot Collaboration
by: Prajod, Pooja, et al.
Published: (2023)
by: Prajod, Pooja, et al.
Published: (2023)
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
by: Yakolli, Nivedan, et al.
Published: (2025)
by: Yakolli, Nivedan, et al.
Published: (2025)
Creativity and Visual Communication from Machine to Musician: Sharing a Score through a Robotic Camera
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
ASI-Seg: Audio-Driven Surgical Instrument Segmentation with Surgeon Intention Understanding
by: Chen, Zhen, et al.
Published: (2024)
by: Chen, Zhen, et al.
Published: (2024)
GuideNav: User-Informed Development of a Vision-Only Robotic Navigation Assistant For Blind Travelers
by: Hwang, Hochul, et al.
Published: (2025)
by: Hwang, Hochul, et al.
Published: (2025)
Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
by: Han, Kyungtae, et al.
Published: (2025)
by: Han, Kyungtae, et al.
Published: (2025)
SurgBox: Agent-Driven Operating Room Sandbox with Surgery Copilot
by: Wu, Jinlin, et al.
Published: (2024)
by: Wu, Jinlin, et al.
Published: (2024)
ShelfHelp: Empowering Humans to Perform Vision-Independent Manipulation Tasks with a Socially Assistive Robotic Cane
by: Agrawal, Shivendra, et al.
Published: (2024)
by: Agrawal, Shivendra, et al.
Published: (2024)
Acoustic Field Video for Multimodal Scene Understanding
by: Kim, Daehwa, et al.
Published: (2026)
by: Kim, Daehwa, et al.
Published: (2026)
A Multimodal Depth-Aware Method For Embodied Reference Understanding
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
AnyUser: Translating Sketched User Intent into Domestic Robots
by: Yang, Songyuan, et al.
Published: (2026)
by: Yang, Songyuan, et al.
Published: (2026)
ReSemAct: Advancing Fine-Grained Robotic Manipulation via Semantic Structuring and Affordance Refinement
by: Su, Chenyu, et al.
Published: (2025)
by: Su, Chenyu, et al.
Published: (2025)
Next-Best-Trajectory Planning of Robot Manipulators for Effective Observation and Exploration
by: Renz, Heiko, et al.
Published: (2025)
by: Renz, Heiko, et al.
Published: (2025)
Generating Robot Constitutions & Benchmarks for Semantic Safety
by: Sermanet, Pierre, et al.
Published: (2025)
by: Sermanet, Pierre, et al.
Published: (2025)
GarmentLab: A Unified Simulation and Benchmark for Garment Manipulation
by: Lu, Haoran, et al.
Published: (2024)
by: Lu, Haoran, et al.
Published: (2024)
Toward Human-Robot Teaming: Learning Handover Behaviors from 3D Scenes
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Real-Time Multimodal Signal Processing for HRI in RoboCup: Understanding a Human Referee
by: Ansalone, Filippo, et al.
Published: (2024)
by: Ansalone, Filippo, et al.
Published: (2024)
Social-LLaVA: Enhancing Robot Navigation through Human-Language Reasoning in Social Spaces
by: Payandeh, Amirreza, et al.
Published: (2024)
by: Payandeh, Amirreza, et al.
Published: (2024)
User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals
by: Miyoshi, Ryo, et al.
Published: (2025)
by: Miyoshi, Ryo, et al.
Published: (2025)
CD-TWINSAFE: A ROS-enabled Digital Twin for Scene Understanding and Safety Emerging V2I Technology
by: Khaled, Amro, et al.
Published: (2026)
by: Khaled, Amro, et al.
Published: (2026)
Inclusive STEAM Education: A Framework for Teaching Cod-2 ing and Robotics to Students with Visually Impairment Using 3 Advanced Computer Vision
by: Hamash, Mahmoud, et al.
Published: (2025)
by: Hamash, Mahmoud, et al.
Published: (2025)
CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence
by: Zeng, Tianle, et al.
Published: (2026)
by: Zeng, Tianle, et al.
Published: (2026)
Logic-Free Building Automation: Learning the Control of Room Facilities with Wall Switches and Ceiling Camera
by: Ochiai, Hideya, et al.
Published: (2024)
by: Ochiai, Hideya, et al.
Published: (2024)
Skeleton-Based Transformer for Classification of Errors and Better Feedback in Low Back Pain Physical Rehabilitation Exercises
by: Marusic, Aleksa, et al.
Published: (2025)
by: Marusic, Aleksa, et al.
Published: (2025)
ICPR 2024 Competition on Rider Intention Prediction
by: Gangisetty, Shankar, et al.
Published: (2025)
by: Gangisetty, Shankar, et al.
Published: (2025)
IndEgo: A Dataset of Industrial Scenarios and Collaborative Work for Egocentric Assistants
by: Chavan, Vivek, et al.
Published: (2025)
by: Chavan, Vivek, et al.
Published: (2025)
Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition
by: Luo, Shengcheng, et al.
Published: (2024)
by: Luo, Shengcheng, et al.
Published: (2024)
Monocular 3D Object Position Estimation with VLMs for Human-Robot Interaction
by: Wahl, Ari, et al.
Published: (2026)
by: Wahl, Ari, et al.
Published: (2026)
HAPI: A Model for Learning Robot Facial Expressions from Human Preferences
by: Yang, Dongsheng, et al.
Published: (2025)
by: Yang, Dongsheng, et al.
Published: (2025)
AppCopilot: Toward General, Accurate, Long-Horizon, and Efficient Mobile Agent
by: Fan, Jingru, et al.
Published: (2025)
by: Fan, Jingru, et al.
Published: (2025)
Intelligent Control of Robotic X-ray Devices using a Language-promptable Digital Twin
by: Killeen, Benjamin D., et al.
Published: (2024)
by: Killeen, Benjamin D., et al.
Published: (2024)
Similar Items
-
Unified Understanding of Environment, Task, and Human for Human-Robot Interaction in Real-World Environments
by: Yano, Yuga, et al.
Published: (2024) -
Lightweight Structured Multimodal Reasoning for Clinical Scene Understanding in Robotics
by: Jha, Saurav, et al.
Published: (2025) -
Robot Interaction Behavior Generation based on Social Motion Forecasting for Human-Robot Interaction
by: Mascaro, Esteve Valls, et al.
Published: (2024) -
Synthetic data enables faster annotation and robust segmentation for multi-object grasping in clutter
by: Lee, Dongmyoung, et al.
Published: (2024) -
iLearnRobot: An Interactive Learning-Based Multi-Modal Robot with Continuous Improvement
by: Wang, Kohou, et al.
Published: (2025)