WHOLE: World-Grounded Hand-Object Lifted from Egocentric Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Yufei, Li, Jiaman, Rong, Ryan, Liu, C. Karen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lifting Motion to the 3D World via 2D Diffusion
von: Li, Jiaman, et al.
Veröffentlicht: (2024)
von: Li, Jiaman, et al.
Veröffentlicht: (2024)
AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
von: Li, Hongjie, et al.
Veröffentlicht: (2026)
von: Li, Hongjie, et al.
Veröffentlicht: (2026)
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026)
von: Li, Dayou, et al.
Veröffentlicht: (2026)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
von: Fu, Hongming, et al.
Veröffentlicht: (2026)
von: Fu, Hongming, et al.
Veröffentlicht: (2026)
Discovering Novel Actions from Open World Egocentric Videos with Object-Grounded Visual Commonsense Reasoning
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2023)
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2023)
Human-Object Interaction from Human-Level Instructions
von: Wu, Zhen, et al.
Veröffentlicht: (2024)
von: Wu, Zhen, et al.
Veröffentlicht: (2024)
Object-Shot Enhanced Grounding Network for Egocentric Video
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
Modeling Fine-Grained Hand-Object Dynamics for Egocentric Video Representation Learning
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
Get a Grip: Reconstructing Hand-Object Stable Grasps in Egocentric Videos
von: Zhu, Zhifan, et al.
Veröffentlicht: (2023)
von: Zhu, Zhifan, et al.
Veröffentlicht: (2023)
Beyond Language: Grounding Referring Expressions with Hand Pointing in Egocentric Vision
von: Li, Ling, et al.
Veröffentlicht: (2026)
von: Li, Ling, et al.
Veröffentlicht: (2026)
HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos
von: Zhang, Jinglei, et al.
Veröffentlicht: (2025)
von: Zhang, Jinglei, et al.
Veröffentlicht: (2025)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
von: Zhou, Bohan, et al.
Veröffentlicht: (2025)
von: Zhou, Bohan, et al.
Veröffentlicht: (2025)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
Controllable Human-Object Interaction Synthesis
von: Li, Jiaman, et al.
Veröffentlicht: (2023)
von: Li, Jiaman, et al.
Veröffentlicht: (2023)
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
Recognizing Hand Use and Hand Role at Home After Stroke from Egocentric Video
von: Tsai, Meng-Fen, et al.
Veröffentlicht: (2022)
von: Tsai, Meng-Fen, et al.
Veröffentlicht: (2022)
Put Myself in Your Shoes: Lifting the Egocentric Perspective from Exocentric Videos
von: Luo, Mi, et al.
Veröffentlicht: (2024)
von: Luo, Mi, et al.
Veröffentlicht: (2024)
Detecting Precise Hand Touch Moments in Egocentric Video
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
von: Bansal, Siddhant, et al.
Veröffentlicht: (2024)
von: Bansal, Siddhant, et al.
Veröffentlicht: (2024)
Are Synthetic Data Useful for Egocentric Hand-Object Interaction Detection?
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023)
von: Leonardi, Rosario, et al.
Veröffentlicht: (2023)
POV: Prompt-Oriented View-Agnostic Learning for Egocentric Hand-Object Interaction in the Multi-View World
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2024)
von: Kundu, Sanjoy, et al.
Veröffentlicht: (2024)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2025)
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2025)
Leveraging Synthetic Data for Enhancing Egocentric Hand-Object Interaction Detection
von: Leonardi, Rosario, et al.
Veröffentlicht: (2026)
von: Leonardi, Rosario, et al.
Veröffentlicht: (2026)
G-HOP: Generative Hand-Object Prior for Interaction Reconstruction and Grasp Synthesis
von: Ye, Yufei, et al.
Veröffentlicht: (2024)
von: Ye, Yufei, et al.
Veröffentlicht: (2024)
Grounded Question-Answering in Long Egocentric Videos
von: Di, Shangzhe, et al.
Veröffentlicht: (2023)
von: Di, Shangzhe, et al.
Veröffentlicht: (2023)
Hand2World: Autoregressive Egocentric Interaction Generation via Free-Space Hand Gestures
von: Wang, Yuxi, et al.
Veröffentlicht: (2026)
von: Wang, Yuxi, et al.
Veröffentlicht: (2026)
StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
von: Zeng, Huajian, et al.
Veröffentlicht: (2026)
LookOut: Real-World Humanoid Egocentric Navigation
von: Pan, Boxiao, et al.
Veröffentlicht: (2025)
von: Pan, Boxiao, et al.
Veröffentlicht: (2025)
Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints
von: Zhang, Chenyangguang, et al.
Veröffentlicht: (2026)
von: Zhang, Chenyangguang, et al.
Veröffentlicht: (2026)
Anticipating Next Active Objects for Egocentric Videos
von: Thakur, Sanket, et al.
Veröffentlicht: (2023)
von: Thakur, Sanket, et al.
Veröffentlicht: (2023)
Predicting 4D Hand Trajectory from Monocular Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2025)
von: Ye, Yufei, et al.
Veröffentlicht: (2025)
Fine-grained Spatiotemporal Grounding on Egocentric Videos
von: Liang, Shuo, et al.
Veröffentlicht: (2025)
von: Liang, Shuo, et al.
Veröffentlicht: (2025)
Grounded Multi-Hop VideoQA in Long-Form Egocentric Videos
von: Chen, Qirui, et al.
Veröffentlicht: (2024)
von: Chen, Qirui, et al.
Veröffentlicht: (2024)
Interaction-aware Representation Modeling with Co-occurrence Consistency for Egocentric Hand-Object Parsing
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
WorldWander: Bridging Egocentric and Exocentric Worlds in Video Generation
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
von: Song, Quanjian, et al.
Veröffentlicht: (2025)
Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
A Real-Time System for Egocentric Hand-Object Interaction Detection in Industrial Domains
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
von: Finocchiaro, Antonio, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Lifting Motion to the 3D World via 2D Diffusion
von: Li, Jiaman, et al.
Veröffentlicht: (2024) -
AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
von: Li, Hongjie, et al.
Veröffentlicht: (2026) -
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
von: Li, Dayou, et al.
Veröffentlicht: (2026) -
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025) -
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
von: Fu, Hongming, et al.
Veröffentlicht: (2026)