Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Aytekin, Ayce Idil, Rhodin, Helge, Dabral, Rishabh, Theobalt, Christian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions
by: Aytekin, Ayce Idil, et al.
Published: (2026)
by: Aytekin, Ayce Idil, et al.
Published: (2026)
Physics-based Human Pose Estimation from a Single Moving RGB Camera
by: Aytekin, Ayce Idil, et al.
Published: (2025)
by: Aytekin, Ayce Idil, et al.
Published: (2025)
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
by: Gopal, Suhas, et al.
Published: (2025)
by: Gopal, Suhas, et al.
Published: (2025)
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025)
by: Zhang, Wanyue, et al.
Published: (2025)
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
by: Ghosh, Anindita, et al.
Published: (2023)
by: Ghosh, Anindita, et al.
Published: (2023)
MIBURI: Towards Expressive Interactive Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2026)
by: Mughal, M. Hamza, et al.
Published: (2026)
ROAM: Robust and Object-Aware Motion Generation Using Neural Pose Descriptors
by: Zhang, Wanyue, et al.
Published: (2023)
by: Zhang, Wanyue, et al.
Published: (2023)
MetaCap: Meta-learning Priors from Multi-View Imagery for Sparse-view Human Performance Capture and Rendering
by: Sun, Guoxing, et al.
Published: (2024)
by: Sun, Guoxing, et al.
Published: (2024)
CoMoGen: COntrollable MOtion Dynamics and Interactions with Mask-Guided Video GENeration
by: Meric, Adil, et al.
Published: (2026)
by: Meric, Adil, et al.
Published: (2026)
E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation
by: Deshmukh, Mayur, et al.
Published: (2026)
by: Deshmukh, Mayur, et al.
Published: (2026)
PractiLight: Practical Light Control Using Foundational Diffusion Models
by: Erel, Yotam, et al.
Published: (2025)
by: Erel, Yotam, et al.
Published: (2025)
BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects
by: Zhang, Wanyue, et al.
Published: (2024)
by: Zhang, Wanyue, et al.
Published: (2024)
Real-time Free-view Human Rendering from Sparse-view RGB Videos using Double Unprojected Textures
by: Sun, Guoxing, et al.
Published: (2024)
by: Sun, Guoxing, et al.
Published: (2024)
SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
by: Ghosh, Anindita, et al.
Published: (2026)
by: Ghosh, Anindita, et al.
Published: (2026)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
by: Liu, Yumeng, et al.
Published: (2024)
by: Liu, Yumeng, et al.
Published: (2024)
Audio-Driven Universal Gaussian Head Avatars
by: Teotia, Kartik, et al.
Published: (2025)
by: Teotia, Kartik, et al.
Published: (2025)
Attention (as Discrete-Time Markov) Chains
by: Erel, Yotam, et al.
Published: (2025)
by: Erel, Yotam, et al.
Published: (2025)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
by: Mughal, Muhammad Hamza, et al.
Published: (2024)
by: Mughal, Muhammad Hamza, et al.
Published: (2024)
Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis
by: Mughal, M. Hamza, et al.
Published: (2024)
by: Mughal, M. Hamza, et al.
Published: (2024)
FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos
by: Delitzas, Alexandros, et al.
Published: (2026)
by: Delitzas, Alexandros, et al.
Published: (2026)
Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal Input
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
FRAME: Floor-aligned Representation for Avatar Motion from Egocentric Video
by: Camiletto, Andrea Boscolo, et al.
Published: (2025)
by: Camiletto, Andrea Boscolo, et al.
Published: (2025)
PocoLoco: A Point Cloud Diffusion Model of Human Shape in Loose Clothing
by: Seth, Siddharth, et al.
Published: (2024)
by: Seth, Siddharth, et al.
Published: (2024)
EmbodMocap: In-the-Wild 4D Human-Scene Reconstruction for Embodied Agents
by: Wang, Wenjia, et al.
Published: (2026)
by: Wang, Wenjia, et al.
Published: (2026)
Object and Contact Point Tracking in Demonstrations Using 3D Gaussian Splatting
by: Büttner, Michael, et al.
Published: (2024)
by: Büttner, Michael, et al.
Published: (2024)
Relightable Holoported Characters: Capturing and Relighting Dynamic Human Performance from Sparse Views
by: Singh, Kunwar Maheep, et al.
Published: (2025)
by: Singh, Kunwar Maheep, et al.
Published: (2025)
LatentKeypointGAN: Controlling Images via Latent Keypoints
by: He, Xingzhe, et al.
Published: (2021)
by: He, Xingzhe, et al.
Published: (2021)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
by: Zhu, Zhifan, et al.
Published: (2025)
by: Zhu, Zhifan, et al.
Published: (2025)
Gaussian Shadow Casting for Neural Characters
by: Bolanos, Luis, et al.
Published: (2024)
by: Bolanos, Luis, et al.
Published: (2024)
HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions
by: Xu, Hao, et al.
Published: (2024)
by: Xu, Hao, et al.
Published: (2024)
CHOIR: Contact-aware 4D Hand-Object Interaction Reconstruction
by: Xu, Hao, et al.
Published: (2026)
by: Xu, Hao, et al.
Published: (2026)
Interaction-Aware 4D Gaussian Splatting for Dynamic Hand-Object Interaction Reconstruction
by: Tian, Hao, et al.
Published: (2025)
by: Tian, Hao, et al.
Published: (2025)
G-HOP: Generative Hand-Object Prior for Interaction Reconstruction and Grasp Synthesis
by: Ye, Yufei, et al.
Published: (2024)
by: Ye, Yufei, et al.
Published: (2024)
Dynamic Reconstruction of Hand-Object Interaction with Distributed Force-aware Contact Representation
by: Yu, Zhenjun, et al.
Published: (2024)
by: Yu, Zhenjun, et al.
Published: (2024)
HandNeRF: Learning to Reconstruct Hand-Object Interaction Scene from a Single RGB Image
by: Choi, Hongsuk, et al.
Published: (2023)
by: Choi, Hongsuk, et al.
Published: (2023)
ReDepth Anything: Test-Time Depth Refinement via Self-Supervised Re-lighting
by: Bhattarai, Ananta R., et al.
Published: (2025)
by: Bhattarai, Ananta R., et al.
Published: (2025)
TexHOI: Reconstructing Textures of 3D Unknown Objects in Monocular Hand-Object Interaction Scenes
by: Aggarwal, Alakh, et al.
Published: (2025)
by: Aggarwal, Alakh, et al.
Published: (2025)
DuetGen: Music Driven Two-Person Dance Generation via Hierarchical Masked Modeling
by: Ghosh, Anindita, et al.
Published: (2025)
by: Ghosh, Anindita, et al.
Published: (2025)
In My Perspective, In My Hands: Accurate Egocentric 2D Hand Pose and Action Recognition
by: Mucha, Wiktor, et al.
Published: (2024)
by: Mucha, Wiktor, et al.
Published: (2024)
Making Video Models Adhere to User Intent with Minor Adjustments
by: Ajisafe, Daniel, et al.
Published: (2026)
by: Ajisafe, Daniel, et al.
Published: (2026)
Similar Items
-
Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions
by: Aytekin, Ayce Idil, et al.
Published: (2026) -
Physics-based Human Pose Estimation from a Single Moving RGB Camera
by: Aytekin, Ayce Idil, et al.
Published: (2025) -
Betsu-Betsu: Multi-View Separable 3D Reconstruction of Two Interacting Objects
by: Gopal, Suhas, et al.
Published: (2025) -
VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification
by: Zhang, Wanyue, et al.
Published: (2025) -
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
by: Ghosh, Anindita, et al.
Published: (2023)