MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xie, Yuechen, Jiang, Haobo, Yang, Jian, Zhang, Yigong, Xie, Jin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mask6D: Masked Pose Priors For 6D Object Pose Estimation
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
MV3DIS: Multi-View Mask Matching via 3D Guides for Zero-Shot 3D Instance Segmentation
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
MonoSE(3)-Diffusion: A Monocular SE(3) Diffusion Framework for Robust Camera-to-Robot Pose Estimation
von: Zhu, Kangjian, et al.
Veröffentlicht: (2025)
von: Zhu, Kangjian, et al.
Veröffentlicht: (2025)
Dataset Ownership Verification for Pre-trained Masked Models
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
von: Yang, Hui, et al.
Veröffentlicht: (2025)
von: Yang, Hui, et al.
Veröffentlicht: (2025)
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
von: Yang, Hui, et al.
Veröffentlicht: (2026)
von: Yang, Hui, et al.
Veröffentlicht: (2026)
FunHOI: Annotation-Free 3D Hand-Object Interaction Generation via Functional Text Guidanc
von: Tian, Yongqi, et al.
Veröffentlicht: (2025)
von: Tian, Yongqi, et al.
Veröffentlicht: (2025)
ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction Videos
von: Chen, Yuantao, et al.
Veröffentlicht: (2026)
von: Chen, Yuantao, et al.
Veröffentlicht: (2026)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
von: Xue, Zihui, et al.
Veröffentlicht: (2024)
von: Xue, Zihui, et al.
Veröffentlicht: (2024)
TexHOI: Reconstructing Textures of 3D Unknown Objects in Monocular Hand-Object Interaction Scenes
von: Aggarwal, Alakh, et al.
Veröffentlicht: (2025)
von: Aggarwal, Alakh, et al.
Veröffentlicht: (2025)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
von: Bansal, Siddhant, et al.
Veröffentlicht: (2024)
von: Bansal, Siddhant, et al.
Veröffentlicht: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
von: Peng, Xiaogang, et al.
Veröffentlicht: (2023)
von: Peng, Xiaogang, et al.
Veröffentlicht: (2023)
PartHOI: Part-based Hand-Object Interaction Transfer via Generalized Cylinders
von: Wang, Qiaochu, et al.
Veröffentlicht: (2025)
von: Wang, Qiaochu, et al.
Veröffentlicht: (2025)
Pan-cancer Histopathology WSI Pre-training with Position-aware Masked Autoencoder
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
Rethinking UMM Visual Generation: Masked Modeling for Efficient Image-Only Pre-training
von: Sun, Peng, et al.
Veröffentlicht: (2026)
von: Sun, Peng, et al.
Veröffentlicht: (2026)
GigaSLAM: Large-Scale Monocular SLAM with Hierarchical Gaussian Splats
von: Deng, Kai, et al.
Veröffentlicht: (2025)
von: Deng, Kai, et al.
Veröffentlicht: (2025)
Universal Image Restoration Pre-training via Masked Degradation Classification
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
von: Hu, JiaKui, et al.
Veröffentlicht: (2025)
F-HOI: Toward Fine-grained Semantic-Aligned 3D Human-Object Interactions
von: Yang, Jie, et al.
Veröffentlicht: (2024)
von: Yang, Jie, et al.
Veröffentlicht: (2024)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
von: Wang, Zikai, et al.
Veröffentlicht: (2026)
von: Wang, Zikai, et al.
Veröffentlicht: (2026)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2025)
DynaHOI: Benchmarking Hand-Object Interaction for Dynamic Target
von: Hu, BoCheng, et al.
Veröffentlicht: (2026)
von: Hu, BoCheng, et al.
Veröffentlicht: (2026)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
FUSER: Feed-Forward MUltiview 3D Registration Transformer and SE(3)$^N$ Diffusion Refinement
von: Jiang, Haobo, et al.
Veröffentlicht: (2025)
von: Jiang, Haobo, et al.
Veröffentlicht: (2025)
GenHOI: Towards Object-Consistent Hand-Object Interaction with Temporally Balanced and Spatially Selective Object Injection
von: Huang, Xuan, et al.
Veröffentlicht: (2026)
von: Huang, Xuan, et al.
Veröffentlicht: (2026)
Masked Pre-training Enables Universal Zero-shot Denoiser
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxiao, et al.
Veröffentlicht: (2024)
Mask Consistency Regularization in Object Removal
von: Yuan, Hua, et al.
Veröffentlicht: (2025)
von: Yuan, Hua, et al.
Veröffentlicht: (2025)
InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling
von: Javed, Muhammad Gohar, et al.
Veröffentlicht: (2024)
von: Javed, Muhammad Gohar, et al.
Veröffentlicht: (2024)
MaskHand: Generative Masked Modeling for Robust Hand Mesh Reconstruction in the Wild
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2024)
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2024)
Masks and Manuscripts: Advancing Medical Pre-training with End-to-End Masking and Narrative Structuring
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2024)
von: Gowda, Shreyank N, et al.
Veröffentlicht: (2024)
Enhancing SAR Object Detection with Self-Supervised Pre-training on Masked Auto-Encoders
von: Pu, Xinyang, et al.
Veröffentlicht: (2025)
von: Pu, Xinyang, et al.
Veröffentlicht: (2025)
ShadowMaskFormer: Mask Augmented Patch Embeddings for Shadow Removal
von: Li, Zhuohao, et al.
Veröffentlicht: (2024)
von: Li, Zhuohao, et al.
Veröffentlicht: (2024)
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
von: Li, Wenyu, et al.
Veröffentlicht: (2025)
von: Li, Wenyu, et al.
Veröffentlicht: (2025)
Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estimation
von: Yang, Yuchen, et al.
Veröffentlicht: (2023)
von: Yang, Yuchen, et al.
Veröffentlicht: (2023)
PA-HOI: A Physics-Aware Human and Object Interaction Dataset
von: Wang, Ruiyan, et al.
Veröffentlicht: (2025)
von: Wang, Ruiyan, et al.
Veröffentlicht: (2025)
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer
von: Shen, Zhelun, et al.
Veröffentlicht: (2025)
von: Shen, Zhelun, et al.
Veröffentlicht: (2025)
P3P: Pseudo-3D Pre-training for Scaling 3D Voxel-based Masked Autoencoders
von: Chen, Xuechao, et al.
Veröffentlicht: (2024)
von: Chen, Xuechao, et al.
Veröffentlicht: (2024)
Emerging Property of Masked Token for Effective Pre-training
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
von: Choi, Hyesong, et al.
Veröffentlicht: (2024)
Efficient Vision-Language Pre-training by Cluster Masking
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
von: Wei, Zihao, et al.
Veröffentlicht: (2024)
DualSplat: Robust 3D Gaussian Splatting via Pseudo-Mask Bootstrapping from Reconstruction Failures
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Mask6D: Masked Pose Priors For 6D Object Pose Estimation
von: Xie, Yuechen, et al.
Veröffentlicht: (2025) -
MV3DIS: Multi-View Mask Matching via 3D Guides for Zero-Shot 3D Instance Segmentation
von: Zhao, Yibo, et al.
Veröffentlicht: (2026) -
MonoSE(3)-Diffusion: A Monocular SE(3) Diffusion Framework for Robust Camera-to-Robot Pose Estimation
von: Zhu, Kangjian, et al.
Veröffentlicht: (2025) -
Dataset Ownership Verification for Pre-trained Masked Models
von: Xie, Yuechen, et al.
Veröffentlicht: (2025) -
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
von: Yang, Hui, et al.
Veröffentlicht: (2025)