Gaze-Guided 3D Hand Motion Prediction for Detecting Intent in Egocentric Grasping Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | He, Yufei, Zhang, Xucong, Stienen, Arno H. A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
por: Chen, Mingfei, et al.
Publicado: (2025)
por: Chen, Mingfei, et al.
Publicado: (2025)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
3D Hand Pose Estimation in Everyday Egocentric Images
por: Prakash, Aditya, et al.
Publicado: (2023)
por: Prakash, Aditya, et al.
Publicado: (2023)
Bimanual Grasp Synthesis for Dexterous Robot Hands
por: Shao, Yanming, et al.
Publicado: (2024)
por: Shao, Yanming, et al.
Publicado: (2024)
From Scene to Object: Text-Guided Dual-Gaze Prediction
por: Ke, Zehong, et al.
Publicado: (2026)
por: Ke, Zehong, et al.
Publicado: (2026)
RealDex: Towards Human-like Grasping for Robotic Dexterous Hand
por: Liu, Yumeng, et al.
Publicado: (2024)
por: Liu, Yumeng, et al.
Publicado: (2024)
Combining Shape Completion and Grasp Prediction for Fast and Versatile Grasping with a Multi-Fingered Hand
por: Humt, Matthias, et al.
Publicado: (2023)
por: Humt, Matthias, et al.
Publicado: (2023)
Efficient Heatmap-Guided 6-Dof Grasp Detection in Cluttered Scenes
por: Chen, Siang, et al.
Publicado: (2024)
por: Chen, Siang, et al.
Publicado: (2024)
TAVIS: A Benchmark for Egocentric Active Vision and Anticipatory Gaze in Imitation Learning
por: Spigler, Giacomo
Publicado: (2026)
por: Spigler, Giacomo
Publicado: (2026)
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views
por: Ma, Junyi, et al.
Publicado: (2025)
por: Ma, Junyi, et al.
Publicado: (2025)
DNAct: Diffusion Guided Multi-Task 3D Policy Learning
por: Yan, Ge, et al.
Publicado: (2024)
por: Yan, Ge, et al.
Publicado: (2024)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
por: Lee, Andrew, et al.
Publicado: (2025)
por: Lee, Andrew, et al.
Publicado: (2025)
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
por: Wang, Wenze, et al.
Publicado: (2026)
por: Wang, Wenze, et al.
Publicado: (2026)
GraspClutter6D: A Large-scale Real-world Dataset for Robust Perception and Grasping in Cluttered Scenes
por: Back, Seunghyeok, et al.
Publicado: (2025)
por: Back, Seunghyeok, et al.
Publicado: (2025)
Ego-Grounding for Personalized Question-Answering in Egocentric Videos
por: Xiao, Junbin, et al.
Publicado: (2026)
por: Xiao, Junbin, et al.
Publicado: (2026)
AI Guide Dog: Egocentric Path Prediction on Smartphone
por: Jadhav, Aishwarya, et al.
Publicado: (2025)
por: Jadhav, Aishwarya, et al.
Publicado: (2025)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
por: Zhang, Erhang, et al.
Publicado: (2025)
por: Zhang, Erhang, et al.
Publicado: (2025)
An Efficient LiDAR-Camera Fusion Network for Multi-Class 3D Dynamic Object Detection and Trajectory Prediction
por: He, Yushen, et al.
Publicado: (2025)
por: He, Yushen, et al.
Publicado: (2025)
DexGrasp Anything: Towards Universal Robotic Dexterous Grasping with Physics Awareness
por: Zhong, Yiming, et al.
Publicado: (2025)
por: Zhong, Yiming, et al.
Publicado: (2025)
DRAMA-X: A Fine-grained Intent Prediction and Risk Reasoning Benchmark For Driving
por: Godbole, Mihir, et al.
Publicado: (2025)
por: Godbole, Mihir, et al.
Publicado: (2025)
VISTA: A Vision and Intent-Aware Social Attention Framework for Multi-Agent Trajectory Prediction
por: Martins, Stephane Da Silva, et al.
Publicado: (2025)
por: Martins, Stephane Da Silva, et al.
Publicado: (2025)
PhyGile: Physics-Prefix Guided Motion Generation for Agile General Humanoid Motion Tracking
por: Bao, Jiacheng, et al.
Publicado: (2026)
por: Bao, Jiacheng, et al.
Publicado: (2026)
In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
por: Cai, Xiongyi, et al.
Publicado: (2025)
por: Cai, Xiongyi, et al.
Publicado: (2025)
RDD4D: 4D Attention-Guided Road Damage Detection And Classification
por: Alkalbani, Asma, et al.
Publicado: (2025)
por: Alkalbani, Asma, et al.
Publicado: (2025)
Robot Instance Segmentation with Few Annotations for Grasping
por: Kimhi, Moshe, et al.
Publicado: (2024)
por: Kimhi, Moshe, et al.
Publicado: (2024)
g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks
por: Wang, Zihan, et al.
Publicado: (2024)
por: Wang, Zihan, et al.
Publicado: (2024)
Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
por: Lall, Vishakha, et al.
Publicado: (2025)
por: Lall, Vishakha, et al.
Publicado: (2025)
PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes
por: Abdelreheem, Ahmed, et al.
Publicado: (2025)
por: Abdelreheem, Ahmed, et al.
Publicado: (2025)
Whole-Body Conditioned Egocentric Video Prediction
por: Bai, Yutong, et al.
Publicado: (2025)
por: Bai, Yutong, et al.
Publicado: (2025)
HandDGP: Camera-Space Hand Mesh Prediction with Differentiable Global Positioning
por: Valassakis, Eugene, et al.
Publicado: (2024)
por: Valassakis, Eugene, et al.
Publicado: (2024)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
por: Fu, Hongming, et al.
Publicado: (2026)
por: Fu, Hongming, et al.
Publicado: (2026)
RAG-3DSG: Enhancing 3D Scene Graphs with Re-Shot Guided Retrieval-Augmented Generation
por: Chang, Yue, et al.
Publicado: (2026)
por: Chang, Yue, et al.
Publicado: (2026)
OpenEgo: A Large-Scale Multimodal Egocentric Dataset for Dexterous Manipulation
por: Jawaid, Ahad, et al.
Publicado: (2025)
por: Jawaid, Ahad, et al.
Publicado: (2025)
SPGrasp: Spatiotemporal Prompt-driven Grasp Synthesis in Dynamic Scenes
por: Mei, Yunpeng, et al.
Publicado: (2025)
por: Mei, Yunpeng, et al.
Publicado: (2025)
EmbodiedVSR: Dynamic Scene Graph-Guided Chain-of-Thought Reasoning for Visual Spatial Tasks
por: Zhang, Yi, et al.
Publicado: (2025)
por: Zhang, Yi, et al.
Publicado: (2025)
LiDAR-based 3D Change Detection at City Scale
por: Albagami, Hezam, et al.
Publicado: (2025)
por: Albagami, Hezam, et al.
Publicado: (2025)
SSL-Interactions: Pretext Tasks for Interactive Trajectory Prediction
por: Bhattacharyya, Prarthana, et al.
Publicado: (2024)
por: Bhattacharyya, Prarthana, et al.
Publicado: (2024)
Flow-guided Motion Prediction with Semantics and Dynamic Occupancy Grid Maps
por: Asghar, Rabbia, et al.
Publicado: (2024)
por: Asghar, Rabbia, et al.
Publicado: (2024)
FunGrasp: Functional Grasping for Diverse Dexterous Hands
por: Huang, Linyi, et al.
Publicado: (2024)
por: Huang, Linyi, et al.
Publicado: (2024)
Robotic Grasping of Harvested Tomato Trusses Using Vision and Online Learning
por: Bent, Luuk van den, et al.
Publicado: (2023)
por: Bent, Luuk van den, et al.
Publicado: (2023)
Ejemplares similares
-
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
por: Chen, Mingfei, et al.
Publicado: (2025) -
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
por: Banerjee, Prithviraj, et al.
Publicado: (2024) -
3D Hand Pose Estimation in Everyday Egocentric Images
por: Prakash, Aditya, et al.
Publicado: (2023) -
Bimanual Grasp Synthesis for Dexterous Robot Hands
por: Shao, Yanming, et al.
Publicado: (2024) -
From Scene to Object: Text-Guided Dual-Gaze Prediction
por: Ke, Zehong, et al.
Publicado: (2026)