EgoPAT3Dv2: Predicting 3D Action Target from 2D Egocentric Vision for Human-Robot Interaction
Fuente:
arXiv
Guardado en:
| Autores principales: | Fang, Irving, Chen, Yuzhong, Wang, Yifan, Zhang, Jianghan, Zhang, Qiushi, Xu, Jiali, He, Xibo, Gao, Weibo, Su, Hao, Li, Yiming, Feng, Chen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
por: Cho, Daesol, et al.
Publicado: (2026)
por: Cho, Daesol, et al.
Publicado: (2026)
EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy
por: Li, Jinzhao, et al.
Publicado: (2026)
por: Li, Jinzhao, et al.
Publicado: (2026)
EgoPrompt: Prompt Learning for Egocentric Action Recognition
por: Lyu, Huaihai, et al.
Publicado: (2025)
por: Lyu, Huaihai, et al.
Publicado: (2025)
EgoHandICL: Egocentric 3D Hand Reconstruction with In-Context Learning
por: Xie, Binzhu, et al.
Publicado: (2026)
por: Xie, Binzhu, et al.
Publicado: (2026)
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
por: Yang, Yuhang, et al.
Publicado: (2024)
por: Yang, Yuhang, et al.
Publicado: (2024)
PARSE-Ego4D: Personal Action Recommendation Suggestions for Egocentric Videos
por: Abreu, Steven, et al.
Publicado: (2024)
por: Abreu, Steven, et al.
Publicado: (2024)
EgoLifter: Open-world 3D Segmentation for Egocentric Perception
por: Gu, Qiao, et al.
Publicado: (2024)
por: Gu, Qiao, et al.
Publicado: (2024)
EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates
por: Peng, Weikun, et al.
Publicado: (2026)
por: Peng, Weikun, et al.
Publicado: (2026)
EventEgo3D: 3D Human Motion Capture from Egocentric Event Streams
por: Millerdurai, Christen, et al.
Publicado: (2024)
por: Millerdurai, Christen, et al.
Publicado: (2024)
EgoReAct: Egocentric Video-Driven 3D Human Reaction Generation
por: Zhang, Libo, et al.
Publicado: (2025)
por: Zhang, Libo, et al.
Publicado: (2025)
EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision
por: Zhao, Yiming, et al.
Publicado: (2024)
por: Zhao, Yiming, et al.
Publicado: (2024)
EgoDTM: Towards 3D-Aware Egocentric Video-Language Pretraining
por: Xu, Boshen, et al.
Publicado: (2025)
por: Xu, Boshen, et al.
Publicado: (2025)
AG-EgoPose: Leveraging Action-Guided Motion and Kinematic Joint Encoding for Egocentric 3D Pose Estimation
por: Azam, Md Mushfiqur, et al.
Publicado: (2026)
por: Azam, Md Mushfiqur, et al.
Publicado: (2026)
EgoGaussian: Dynamic Scene Understanding from Egocentric Video with 3D Gaussian Splatting
por: Zhang, Daiwei, et al.
Publicado: (2024)
por: Zhang, Daiwei, et al.
Publicado: (2024)
EgoSplat: Open-Vocabulary Egocentric Scene Understanding with Language Embedded 3D Gaussian Splatting
por: Li, Di, et al.
Publicado: (2025)
por: Li, Di, et al.
Publicado: (2025)
EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos
por: Yang, Ruihan, et al.
Publicado: (2025)
por: Yang, Ruihan, et al.
Publicado: (2025)
EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation
por: Yang, Chenhongyi, et al.
Publicado: (2024)
por: Yang, Chenhongyi, et al.
Publicado: (2024)
EgoPush: Learning End-to-End Egocentric Multi-Object Rearrangement for Mobile Robots
por: An, Boyuan, et al.
Publicado: (2026)
por: An, Boyuan, et al.
Publicado: (2026)
EventEgoHands: Event-based Egocentric 3D Hand Mesh Reconstruction
por: Hara, Ryosei, et al.
Publicado: (2025)
por: Hara, Ryosei, et al.
Publicado: (2025)
VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
por: Wang, Beichen, et al.
Publicado: (2024)
por: Wang, Beichen, et al.
Publicado: (2024)
EgoM2P: Egocentric Multimodal Multitask Pretraining
por: Li, Gen, et al.
Publicado: (2025)
por: Li, Gen, et al.
Publicado: (2025)
PAT3D: Physics-Augmented Text-to-3D Scene Generation
por: Lin, Guying, et al.
Publicado: (2025)
por: Lin, Guying, et al.
Publicado: (2025)
EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration
por: Shi, Modi, et al.
Publicado: (2026)
por: Shi, Modi, et al.
Publicado: (2026)
EgoDemoGen: Egocentric Demonstration Generation for Viewpoint Generalization in Robotic Manipulation
por: Xu, Yuan, et al.
Publicado: (2025)
por: Xu, Yuan, et al.
Publicado: (2025)
EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses
por: Pallotta, Enrico, et al.
Publicado: (2025)
por: Pallotta, Enrico, et al.
Publicado: (2025)
Real-time 3D Semantic Scene Perception for Egocentric Robots with Binocular Vision
por: Nguyen, K., et al.
Publicado: (2024)
por: Nguyen, K., et al.
Publicado: (2024)
Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation
por: Hu, Mu, et al.
Publicado: (2024)
por: Hu, Mu, et al.
Publicado: (2024)
EMAG: Ego-motion Aware and Generalizable 2D Hand Forecasting from Egocentric Videos
por: Hatano, Masashi, et al.
Publicado: (2024)
por: Hatano, Masashi, et al.
Publicado: (2024)
PCIE_Interaction Solution for Ego4D Social Interaction Challenge
por: Lertniphonphan, Kanokphan, et al.
Publicado: (2025)
por: Lertniphonphan, Kanokphan, et al.
Publicado: (2025)
DvD: Unleashing a Generative Paradigm for Document Dewarping via Coordinates-based Diffusion Model
por: Zhang, Weiguang, et al.
Publicado: (2025)
por: Zhang, Weiguang, et al.
Publicado: (2025)
TeleEgo: Benchmarking Egocentric AI Assistants in the Wild
por: Yan, Jiaqi, et al.
Publicado: (2025)
por: Yan, Jiaqi, et al.
Publicado: (2025)
EgoMotion: Hierarchical Reasoning and Diffusion for Egocentric Vision-Language Motion Generation
por: Hou, Ruibing, et al.
Publicado: (2026)
por: Hou, Ruibing, et al.
Publicado: (2026)
EgoGen: An Egocentric Synthetic Data Generator
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
EgoInteract: Synthetic Egocentric Videos Generation for Interaction Understanding and Anticipation
por: Leonardi, Rosario, et al.
Publicado: (2026)
por: Leonardi, Rosario, et al.
Publicado: (2026)
CaRe-Ego: Contact-aware Relationship Modeling for Egocentric Interactive Hand-object Segmentation
por: Su, Yuejiao, et al.
Publicado: (2024)
por: Su, Yuejiao, et al.
Publicado: (2024)
EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos
por: Ma, Junyi, et al.
Publicado: (2025)
por: Ma, Junyi, et al.
Publicado: (2025)
Interact with me: Joint Egocentric Forecasting of Intent to Interact, Attitude and Social Actions
por: Bian, Tongfei, et al.
Publicado: (2024)
por: Bian, Tongfei, et al.
Publicado: (2024)
EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
por: Fu, Zhiheng, et al.
Publicado: (2026)
por: Fu, Zhiheng, et al.
Publicado: (2026)
EgoLife: Towards Egocentric Life Assistant
por: Yang, Jingkang, et al.
Publicado: (2025)
por: Yang, Jingkang, et al.
Publicado: (2025)
EgoSim: Egocentric World Simulator for Embodied Interaction Generation
por: Hao, Jinkun, et al.
Publicado: (2026)
por: Hao, Jinkun, et al.
Publicado: (2026)
Ejemplares similares
-
EgoAVFlow: Robot Policy Learning with Active Vision from Human Egocentric Videos via 3D Flow
por: Cho, Daesol, et al.
Publicado: (2026) -
EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy
por: Li, Jinzhao, et al.
Publicado: (2026) -
EgoPrompt: Prompt Learning for Egocentric Action Recognition
por: Lyu, Huaihai, et al.
Publicado: (2025) -
EgoHandICL: Egocentric 3D Hand Reconstruction with In-Context Learning
por: Xie, Binzhu, et al.
Publicado: (2026) -
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
por: Yang, Yuhang, et al.
Publicado: (2024)