Intent at a Glance: Gaze-Guided Robotic Manipulation via Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Tay, Tracey Yee Hsin, Yan, Xu, Ouyang, Jonathan, Wu, Daniel, Jiang, William, Kao, Jonathan, Cui, Yuchen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sticky-Glance: Robust Intent Recognition for Human Robot Collaboration via Single-Glance
by: Lai, Yuzhi, et al.
Published: (2026)
by: Lai, Yuzhi, et al.
Published: (2026)
Gaze-Guided Task Decomposition for Imitation Learning in Robotic Manipulation
by: Takizawa, Ryo, et al.
Published: (2025)
by: Takizawa, Ryo, et al.
Published: (2025)
GazeVLA: Learning Human Intention for Robotic Manipulation
by: Li, Chengyang, et al.
Published: (2026)
by: Li, Chengyang, et al.
Published: (2026)
Navi2Gaze: Leveraging Foundation Models for Navigation and Target Gazing
by: Zhu, Jun, et al.
Published: (2024)
by: Zhu, Jun, et al.
Published: (2024)
Gaze-Guided Robotic Vascular Ultrasound Leveraging Human Intention Estimation
by: Bi, Yuan, et al.
Published: (2025)
by: Bi, Yuan, et al.
Published: (2025)
Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation
by: Zuo, Kuangji, et al.
Published: (2026)
by: Zuo, Kuangji, et al.
Published: (2026)
IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation
by: Lian, Shijie, et al.
Published: (2026)
by: Lian, Shijie, et al.
Published: (2026)
Multi-Robot Data-Free Continual Communicative Learning (CCL) from Black-Box Visual Place Recognition Models
by: Tsukahara, Kenta, et al.
Published: (2025)
by: Tsukahara, Kenta, et al.
Published: (2025)
Transferring Foundation Models for Generalizable Robotic Manipulation
by: Yang, Jiange, et al.
Published: (2023)
by: Yang, Jiange, et al.
Published: (2023)
Mirror Eyes: Explainable Human-Robot Interaction at a Glance
by: Krüger, Matti, et al.
Published: (2025)
by: Krüger, Matti, et al.
Published: (2025)
DINOBot: Robot Manipulation via Retrieval and Alignment with Vision Foundation Models
by: Di Palo, Norman, et al.
Published: (2024)
by: Di Palo, Norman, et al.
Published: (2024)
Embodied Robot Manipulation in the Era of Foundation Models: Planning and Learning Perspectives
by: Bai, Shuanghao, et al.
Published: (2025)
by: Bai, Shuanghao, et al.
Published: (2025)
Constraint-Aware Intent Estimation for Dynamic Human-Robot Object Co-Manipulation
by: Shao, Yifei Simon, et al.
Published: (2024)
by: Shao, Yifei Simon, et al.
Published: (2024)
Enhancing Reusability of Learned Skills for Robot Manipulation via Gaze Information and Motion Bottlenecks
by: Takizawa, Ryo, et al.
Published: (2025)
by: Takizawa, Ryo, et al.
Published: (2025)
IntentReact: Guiding Reactive Object-Centric Navigation via Topological Intent
by: Jiao, Yanmei, et al.
Published: (2026)
by: Jiao, Yanmei, et al.
Published: (2026)
Automatic Behavior Tree Expansion with LLMs for Robotic Manipulation
by: Styrud, Jonathan, et al.
Published: (2024)
by: Styrud, Jonathan, et al.
Published: (2024)
RaycastGrasp: Eye-Gaze Interaction with Wearable Devices for Robotic Manipulation
by: Lin, Zitiantao, et al.
Published: (2025)
by: Lin, Zitiantao, et al.
Published: (2025)
Humanizing Robot Gaze Shifts: A Framework for Natural Gaze Shifts in Humanoid Robots
by: Wei, Jingchao, et al.
Published: (2026)
by: Wei, Jingchao, et al.
Published: (2026)
Human Gaze and Head Rotation during Navigation, Exploration and Object Manipulation in Shared Environments with Robots
by: Schreiter, Tim, et al.
Published: (2024)
by: Schreiter, Tim, et al.
Published: (2024)
KineSoft: Learning Proprioceptive Manipulation Policies with Soft Robot Hands
by: Yoo, Uksang, et al.
Published: (2025)
by: Yoo, Uksang, et al.
Published: (2025)
Long-horizon Locomotion and Manipulation on a Quadrupedal Robot with Large Language Models
by: Ouyang, Yutao, et al.
Published: (2024)
by: Ouyang, Yutao, et al.
Published: (2024)
Learning Instruction-Guided Manipulation Affordance via Large Models for Embodied Robotic Tasks
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
Autonomous Surface Selection For Manipulator-Based UV Disinfection In Hospitals Using Foundation Models
by: Oh, Xueyan, et al.
Published: (2025)
by: Oh, Xueyan, et al.
Published: (2025)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
by: Jiang, Feng, et al.
Published: (2026)
by: Jiang, Feng, et al.
Published: (2026)
InCoM: Intent-Driven Perception and Structured Coordination for Mobile Manipulation
by: Liu, Jiahao, et al.
Published: (2026)
by: Liu, Jiahao, et al.
Published: (2026)
DYMO-Hair: Generalizable Volumetric Dynamics Modeling for Robot Hair Manipulation
by: Zhao, Chengyang, et al.
Published: (2025)
by: Zhao, Chengyang, et al.
Published: (2025)
TGM-VLA: Task-Guided Mixup for Sampling-Efficient and Robust Robotic Manipulation
by: Pu, Fanqi, et al.
Published: (2026)
by: Pu, Fanqi, et al.
Published: (2026)
ManiFoundation Model for General-Purpose Robotic Manipulation of Contact Synthesis with Arbitrary Objects and Robots
by: Xu, Zhixuan, et al.
Published: (2024)
by: Xu, Zhixuan, et al.
Published: (2024)
What Foundation Models can Bring for Robot Learning in Manipulation : A Survey
by: Li, Dingzhe, et al.
Published: (2024)
by: Li, Dingzhe, et al.
Published: (2024)
Risk-Guided Diffusion: Toward Deploying Robot Foundation Models in Space, Where Failure Is Not An Option
by: Thakker, Rohan, et al.
Published: (2025)
by: Thakker, Rohan, et al.
Published: (2025)
GazeGrasp: DNN-Driven Robotic Grasping with Wearable Eye-Gaze Interface
by: Tokmurziyev, Issatay, et al.
Published: (2025)
by: Tokmurziyev, Issatay, et al.
Published: (2025)
Gaze-Guided 3D Hand Motion Prediction for Detecting Intent in Egocentric Grasping Tasks
by: He, Yufei, et al.
Published: (2025)
by: He, Yufei, et al.
Published: (2025)
RoPotter: Toward Robotic Pottery and Deformable Object Manipulation with Structural Priors
by: Yoo, Uksang, et al.
Published: (2024)
by: Yoo, Uksang, et al.
Published: (2024)
Distilling and Retrieving Generalizable Knowledge for Robot Manipulation via Language Corrections
by: Zha, Lihan, et al.
Published: (2023)
by: Zha, Lihan, et al.
Published: (2023)
Controlling Intent Expressiveness in Robot Motion with Diffusion Models
by: Shi, Wenli, et al.
Published: (2025)
by: Shi, Wenli, et al.
Published: (2025)
Temporal and Semantic Evaluation Metrics for Foundation Models in Post-Hoc Analysis of Robotic Sub-tasks
by: Salfity, Jonathan, et al.
Published: (2024)
by: Salfity, Jonathan, et al.
Published: (2024)
Dynamics-Guided Diffusion Model for Sensor-less Robot Manipulator Design
by: Xu, Xiaomeng, et al.
Published: (2024)
by: Xu, Xiaomeng, et al.
Published: (2024)
Learning Long-Horizon Robot Manipulation Skills via Privileged Action
by: Mao, Xiaofeng, et al.
Published: (2025)
by: Mao, Xiaofeng, et al.
Published: (2025)
Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO
by: Chen, Tianyang, et al.
Published: (2026)
by: Chen, Tianyang, et al.
Published: (2026)
FAM-HRI: Foundation-Model Assisted Multi-Modal Human-Robot Interaction Combining Gaze and Speech
by: Lai, Yuzhi, et al.
Published: (2025)
by: Lai, Yuzhi, et al.
Published: (2025)
Similar Items
-
Sticky-Glance: Robust Intent Recognition for Human Robot Collaboration via Single-Glance
by: Lai, Yuzhi, et al.
Published: (2026) -
Gaze-Guided Task Decomposition for Imitation Learning in Robotic Manipulation
by: Takizawa, Ryo, et al.
Published: (2025) -
GazeVLA: Learning Human Intention for Robotic Manipulation
by: Li, Chengyang, et al.
Published: (2026) -
Navi2Gaze: Leveraging Foundation Models for Navigation and Target Gazing
by: Zhu, Jun, et al.
Published: (2024) -
Gaze-Guided Robotic Vascular Ultrasound Leveraging Human Intention Estimation
by: Bi, Yuan, et al.
Published: (2025)