OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zhenhao, Shi, Ye, Yang, Lingxiao, Ni, Suting, Ye, Qi, Wang, Jingya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gaze-guided Hand-Object Interaction Synthesis: Dataset and Method
by: Tian, Jie, et al.
Published: (2024)
by: Tian, Jie, et al.
Published: (2024)
UniHM: Unified Dexterous Hand Manipulation with Vision Language Model
by: Zhang, Zhenhao, et al.
Published: (2026)
by: Zhang, Zhenhao, et al.
Published: (2026)
HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment
by: Zhang, Juze, et al.
Published: (2024)
by: Zhang, Juze, et al.
Published: (2024)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
by: Liu, Yumeng, et al.
Published: (2024)
by: Liu, Yumeng, et al.
Published: (2024)
HOID-R1: Reinforcement Learning for Open-World Human-Object Interaction Detection Reasoning with Multimodal Large Language Model
by: Zhang, Zhenhao, et al.
Published: (2025)
by: Zhang, Zhenhao, et al.
Published: (2025)
ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction Videos
by: Chen, Yuantao, et al.
Published: (2026)
by: Chen, Yuantao, et al.
Published: (2026)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
by: Xue, Zihui, et al.
Published: (2024)
by: Xue, Zihui, et al.
Published: (2024)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
by: Bansal, Siddhant, et al.
Published: (2024)
by: Bansal, Siddhant, et al.
Published: (2024)
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
by: Lei, Ting, et al.
Published: (2024)
by: Lei, Ting, et al.
Published: (2024)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
by: Nguyen, Minh Anh, et al.
Published: (2026)
by: Nguyen, Minh Anh, et al.
Published: (2026)
DynaHOI: Benchmarking Hand-Object Interaction for Dynamic Target
by: Hu, BoCheng, et al.
Published: (2026)
by: Hu, BoCheng, et al.
Published: (2026)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
by: Wang, Zikai, et al.
Published: (2026)
by: Wang, Zikai, et al.
Published: (2026)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
by: Xie, Yuechen, et al.
Published: (2025)
by: Xie, Yuechen, et al.
Published: (2025)
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
by: Lei, Ting, et al.
Published: (2025)
by: Lei, Ting, et al.
Published: (2025)
I'M HOI: Inertia-aware Monocular Capture of 3D Human-Object Interactions
by: Zhao, Chengfeng, et al.
Published: (2023)
by: Zhao, Chengfeng, et al.
Published: (2023)
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
by: Chen, Junwen, et al.
Published: (2025)
by: Chen, Junwen, et al.
Published: (2025)
PartHOI: Part-based Hand-Object Interaction Transfer via Generalized Cylinders
by: Wang, Qiaochu, et al.
Published: (2025)
by: Wang, Qiaochu, et al.
Published: (2025)
GenHOI: Towards Object-Consistent Hand-Object Interaction with Temporally Balanced and Spatially Selective Object Injection
by: Huang, Xuan, et al.
Published: (2026)
by: Huang, Xuan, et al.
Published: (2026)
TexHOI: Reconstructing Textures of 3D Unknown Objects in Monocular Hand-Object Interaction Scenes
by: Aggarwal, Alakh, et al.
Published: (2025)
by: Aggarwal, Alakh, et al.
Published: (2025)
FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation
by: Zeng, Huajian, et al.
Published: (2026)
by: Zeng, Huajian, et al.
Published: (2026)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
by: Cai, Songjin, et al.
Published: (2026)
by: Cai, Songjin, et al.
Published: (2026)
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
by: Yang, Hui, et al.
Published: (2026)
by: Yang, Hui, et al.
Published: (2026)
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
by: Li, Dayou, et al.
Published: (2026)
by: Li, Dayou, et al.
Published: (2026)
Monocular Human-Object Reconstruction in the Wild
by: Huo, Chaofan, et al.
Published: (2024)
by: Huo, Chaofan, et al.
Published: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
by: Cha, Junuk, et al.
Published: (2024)
by: Cha, Junuk, et al.
Published: (2024)
G-HOP: Generative Hand-Object Prior for Interaction Reconstruction and Grasp Synthesis
by: Ye, Yufei, et al.
Published: (2024)
by: Ye, Yufei, et al.
Published: (2024)
StructBiHOI: Structured Articulation Modeling for Long--Horizon Bimanual Hand--Object Interaction Generation
by: Wang, Zhi, et al.
Published: (2026)
by: Wang, Zhi, et al.
Published: (2026)
SGC-Net: Stratified Granular Comparison Network for Open-Vocabulary HOI Detection
by: Lin, Xin, et al.
Published: (2025)
by: Lin, Xin, et al.
Published: (2025)
THOR: Text to Human-Object Interaction Diffusion via Relation Intervention
by: Wu, Qianyang, et al.
Published: (2024)
by: Wu, Qianyang, et al.
Published: (2024)
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model
by: Yu, Chunlin, et al.
Published: (2024)
by: Yu, Chunlin, et al.
Published: (2024)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
by: Yang, Jie, et al.
Published: (2024)
by: Yang, Jie, et al.
Published: (2024)
On Large Multimodal Models as Open-World Image Classifiers
by: Conti, Alessandro, et al.
Published: (2025)
by: Conti, Alessandro, et al.
Published: (2025)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
by: Zhang, Jiajun, et al.
Published: (2023)
by: Zhang, Jiajun, et al.
Published: (2023)
Towards Open-Vocabulary Industrial Defect Understanding with a Large-Scale Multimodal Dataset
by: Ni, TsaiChing, et al.
Published: (2025)
by: Ni, TsaiChing, et al.
Published: (2025)
Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy
by: Deng, Zekai, et al.
Published: (2025)
by: Deng, Zekai, et al.
Published: (2025)
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer
by: Shen, Zhelun, et al.
Published: (2025)
by: Shen, Zhelun, et al.
Published: (2025)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
by: Zhou, Bohan, et al.
Published: (2025)
by: Zhou, Bohan, et al.
Published: (2025)
ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
by: Zeng, Ling-An, et al.
Published: (2025)
by: Zeng, Ling-An, et al.
Published: (2025)
HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models
by: Peng, Xiaogang, et al.
Published: (2023)
by: Peng, Xiaogang, et al.
Published: (2023)
MineWorld: a Real-Time and Open-Source Interactive World Model on Minecraft
by: Guo, Junliang, et al.
Published: (2025)
by: Guo, Junliang, et al.
Published: (2025)
Similar Items
-
Gaze-guided Hand-Object Interaction Synthesis: Dataset and Method
by: Tian, Jie, et al.
Published: (2024) -
UniHM: Unified Dexterous Hand Manipulation with Vision Language Model
by: Zhang, Zhenhao, et al.
Published: (2026) -
HOI-M3:Capture Multiple Humans and Objects Interaction within Contextual Environment
by: Zhang, Juze, et al.
Published: (2024) -
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
by: Liu, Yumeng, et al.
Published: (2024) -
HOID-R1: Reinforcement Learning for Open-World Human-Object Interaction Detection Reasoning with Multimodal Large Language Model
by: Zhang, Zhenhao, et al.
Published: (2025)