Interacted Object Grounding in Spatio-Temporal Human-Object Interactions
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xiaoyang, Wen, Boran, Liu, Xinpeng, Zhou, Zizheng, Fan, Hongwei, Lu, Cewu, Ma, Lizhuang, Chen, Yulong, Li, Yong-Lu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
by: Wen, Boran, et al.
Published: (2025)
by: Wen, Boran, et al.
Published: (2025)
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
by: Wen, Boran, et al.
Published: (2025)
by: Wen, Boran, et al.
Published: (2025)
Bridging the Gap between Human Motion and Action Semantics via Kinematic Phrases
by: Liu, Xinpeng, et al.
Published: (2023)
by: Liu, Xinpeng, et al.
Published: (2023)
Revisit Human-Scene Interaction via Space Occupancy
by: Liu, Xinpeng, et al.
Published: (2023)
by: Liu, Xinpeng, et al.
Published: (2023)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
by: Mao, Aihua, et al.
Published: (2026)
by: Mao, Aihua, et al.
Published: (2026)
iTACO: Interactable Digital Twins of Articulated Objects from Casually Captured RGBD Videos
by: Peng, Weikun, et al.
Published: (2025)
by: Peng, Weikun, et al.
Published: (2025)
OmniSTVG: Toward Spatio-Temporal Omni-Object Video Grounding
by: Yao, Jiali, et al.
Published: (2025)
by: Yao, Jiali, et al.
Published: (2025)
Physically Ground Commonsense Knowledge for Articulated Object Manipulation with Analytic Concepts
by: Wei, Jiude, et al.
Published: (2025)
by: Wei, Jiude, et al.
Published: (2025)
HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
by: Huang, Ziyao, et al.
Published: (2025)
by: Huang, Ziyao, et al.
Published: (2025)
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
by: Xing, Hao, et al.
Published: (2024)
by: Xing, Hao, et al.
Published: (2024)
Dynamic Reconstruction of Hand-Object Interaction with Distributed Force-aware Contact Representation
by: Yu, Zhenjun, et al.
Published: (2024)
by: Yu, Zhenjun, et al.
Published: (2024)
Spatial-Temporal Human-Object Interaction Detection
by: Sun, Xu, et al.
Published: (2025)
by: Sun, Xu, et al.
Published: (2025)
VideoAfford: Grounding 3D Affordance from Human-Object-Interaction Videos via Multimodal Large Language Model
by: Wang, Hanqing, et al.
Published: (2026)
by: Wang, Hanqing, et al.
Published: (2026)
Homogeneous Dynamics Space for Heterogeneous Humans
by: Liu, Xinpeng, et al.
Published: (2024)
by: Liu, Xinpeng, et al.
Published: (2024)
InterMamba: Efficient Human-Human Interaction Generation with Adaptive Spatio-Temporal Mamba
by: Wu, Zizhao, et al.
Published: (2025)
by: Wu, Zizhao, et al.
Published: (2025)
Controllable Human-Object Interaction Synthesis
by: Li, Jiaman, et al.
Published: (2023)
by: Li, Jiaman, et al.
Published: (2023)
SynHLMA:Synthesizing Hand Language Manipulation for Articulated Object with Discrete Human Object Interaction Representation
by: zhi, Wang, et al.
Published: (2025)
by: zhi, Wang, et al.
Published: (2025)
HUMOTO: A 4D Dataset of Mocap Human Object Interactions
by: Lu, Jiaxin, et al.
Published: (2025)
by: Lu, Jiaxin, et al.
Published: (2025)
ImDy: Human Inverse Dynamics from Imitated Observations
by: Liu, Xinpeng, et al.
Published: (2024)
by: Liu, Xinpeng, et al.
Published: (2024)
GenHOI: Towards Object-Consistent Hand-Object Interaction with Temporally Balanced and Spatially Selective Object Injection
by: Huang, Xuan, et al.
Published: (2026)
by: Huang, Xuan, et al.
Published: (2026)
Egocentric World Model for Photorealistic Hand-Object Interaction Synthesis
by: Li, Dayou, et al.
Published: (2026)
by: Li, Dayou, et al.
Published: (2026)
HOComp: Interaction-Aware Human-Object Composition
by: Liang, Dong, et al.
Published: (2025)
by: Liang, Dong, et al.
Published: (2025)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
by: Zhang, Jiajun, et al.
Published: (2023)
by: Zhang, Jiajun, et al.
Published: (2023)
Contextualized Representation Learning for Effective Human-Object Interaction Detection
by: Li, Zhehao, et al.
Published: (2025)
by: Li, Zhehao, et al.
Published: (2025)
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
by: Zhou, Donghao, et al.
Published: (2026)
by: Zhou, Donghao, et al.
Published: (2026)
DISCO: Embodied Navigation and Interaction via Differentiable Scene Semantics and Dual-level Control
by: Xu, Xinyu, et al.
Published: (2024)
by: Xu, Xinyu, et al.
Published: (2024)
Interact-Custom: Customized Human Object Interaction Image Generation
by: Xu, Zhu, et al.
Published: (2025)
by: Xu, Zhu, et al.
Published: (2025)
Guiding Human-Object Interactions with Rich Geometry and Relations
by: Xue, Mengqing, et al.
Published: (2025)
by: Xue, Mengqing, et al.
Published: (2025)
Improving Skeleton-based Action Recognition with Interactive Object Information
by: Wen, Hao, et al.
Published: (2025)
by: Wen, Hao, et al.
Published: (2025)
InteractMove: Text-Controlled Human-Object Interaction Generation in 3D Scenes with Movable Objects
by: Cai, Xinhao, et al.
Published: (2025)
by: Cai, Xinhao, et al.
Published: (2025)
SuperMat: Physically Consistent PBR Material Estimation at Interactive Rates
by: Hong, Yijia, et al.
Published: (2024)
by: Hong, Yijia, et al.
Published: (2024)
Learning Human-Object Interaction as Groups
by: Hong, Jiajun, et al.
Published: (2025)
by: Hong, Jiajun, et al.
Published: (2025)
Towards Unconstrained Human-Object Interaction
by: Tonini, Francesco, et al.
Published: (2026)
by: Tonini, Francesco, et al.
Published: (2026)
Human-Object Interaction from Human-Level Instructions
by: Wu, Zhen, et al.
Published: (2024)
by: Wu, Zhen, et al.
Published: (2024)
An Image-like Diffusion Method for Human-Object Interaction Detection
by: Hui, Xiaofei, et al.
Published: (2025)
by: Hui, Xiaofei, et al.
Published: (2025)
InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
by: Zhang, Jinlu, et al.
Published: (2025)
by: Zhang, Jinlu, et al.
Published: (2025)
VISTA: Video Interaction Spatio-Temporal Analysis Benchmark
by: Aparcedo, Alejandro, et al.
Published: (2026)
by: Aparcedo, Alejandro, et al.
Published: (2026)
TIMotion: Temporal and Interactive Framework for Efficient Human-Human Motion Generation
by: Wang, Yabiao, et al.
Published: (2024)
by: Wang, Yabiao, et al.
Published: (2024)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
by: Zhou, Bohan, et al.
Published: (2025)
by: Zhou, Bohan, et al.
Published: (2025)
An Angular-Temporal Interaction Network for Light Field Object Tracking in Low-Light Scenes
by: Wang, Mianzhao, et al.
Published: (2025)
by: Wang, Mianzhao, et al.
Published: (2025)
Similar Items
-
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
by: Wen, Boran, et al.
Published: (2025) -
Efficient and Scalable Monocular Human-Object Interaction Motion Reconstruction
by: Wen, Boran, et al.
Published: (2025) -
Bridging the Gap between Human Motion and Action Semantics via Kinematic Phrases
by: Liu, Xinpeng, et al.
Published: (2023) -
Revisit Human-Scene Interaction via Space Occupancy
by: Liu, Xinpeng, et al.
Published: (2023) -
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
by: Mao, Aihua, et al.
Published: (2026)