CrossHOI-Bench: A Unified Benchmark for HOI Evaluation across Vision-Language Models and HOI-Specific Methods
Fuente:
arXiv
Saved in:
| Main Authors: | Lei, Qinqian, Wang, Bo, Tan, Robby T. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EZ-HOI: VLM Adaptation via Guided Prompt Learning for Zero-Shot HOI Detection
by: Lei, Qinqian, et al.
Published: (2024)
by: Lei, Qinqian, et al.
Published: (2024)
HOLa: Zero-Shot HOI Detection with Low-Rank Decomposed VLM Feature Adaptation
by: Lei, Qinqian, et al.
Published: (2025)
by: Lei, Qinqian, et al.
Published: (2025)
SHOE: Semantic HOI Open-Vocabulary Evaluation Metric
by: Noack, Maja, et al.
Published: (2026)
by: Noack, Maja, et al.
Published: (2026)
Funnel-HOI: Top-Down Perception for Zero-Shot HOI Detection
by: Sarma, Sandipan, et al.
Published: (2025)
by: Sarma, Sandipan, et al.
Published: (2025)
IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction
by: Zhang, Haoshen, et al.
Published: (2026)
by: Zhang, Haoshen, et al.
Published: (2026)
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
by: Zhao, Yubo, et al.
Published: (2026)
by: Zhao, Yubo, et al.
Published: (2026)
Controllable Hand Grasp Generation for HOI and Efficient Evaluation Methods
by: Ishant, et al.
Published: (2025)
by: Ishant, et al.
Published: (2025)
OneHOI: Unifying Human-Object Interaction Generation and Editing
by: Hoe, Jiun Tian, et al.
Published: (2026)
by: Hoe, Jiun Tian, et al.
Published: (2026)
DynaHOI: Benchmarking Hand-Object Interaction for Dynamic Target
by: Hu, BoCheng, et al.
Published: (2026)
by: Hu, BoCheng, et al.
Published: (2026)
Enhancing HOI Detection with Contextual Cues from Large Vision-Language Models
by: Zhan, Yu-Wei, et al.
Published: (2023)
by: Zhan, Yu-Wei, et al.
Published: (2023)
HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision
by: Bansal, Siddhant, et al.
Published: (2024)
by: Bansal, Siddhant, et al.
Published: (2024)
CL-HOI: Cross-Level Human-Object Interaction Distillation from Vision Large Language Models
by: Gao, Jianjun, et al.
Published: (2024)
by: Gao, Jianjun, et al.
Published: (2024)
HOI-aware Adaptive Network for Weakly-supervised Action Segmentation
by: Zhang, Runzhong, et al.
Published: (2026)
by: Zhang, Runzhong, et al.
Published: (2026)
Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
by: Lei, Ting, et al.
Published: (2024)
by: Lei, Ting, et al.
Published: (2024)
VLM-HOI: Vision Language Models for Interpretable Human-Object Interaction Analysis
by: Kang, Donggoo, et al.
Published: (2024)
by: Kang, Donggoo, et al.
Published: (2024)
Exploring Interactive Semantic Alignment for Efficient HOI Detection with Vision-language Model
by: Dong, Jihao, et al.
Published: (2024)
by: Dong, Jihao, et al.
Published: (2024)
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection
by: Lei, Ting, et al.
Published: (2024)
by: Lei, Ting, et al.
Published: (2024)
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
by: Lei, Ting, et al.
Published: (2025)
by: Lei, Ting, et al.
Published: (2025)
End-to-End HOI Reconstruction Transformer with Graph-based Encoding
by: Wang, Zhenrong, et al.
Published: (2025)
by: Wang, Zhenrong, et al.
Published: (2025)
PA-HOI: A Physics-Aware Human and Object Interaction Dataset
by: Wang, Ruiyan, et al.
Published: (2025)
by: Wang, Ruiyan, et al.
Published: (2025)
Ins-HOI: Instance Aware Human-Object Interactions Recovery
by: Zhang, Jiajun, et al.
Published: (2023)
by: Zhang, Jiajun, et al.
Published: (2023)
UAHOI: Uncertainty-aware Robust Interaction Learning for HOI Detection
by: Chen, Mu, et al.
Published: (2024)
by: Chen, Mu, et al.
Published: (2024)
ViHOI: Human-Object Interaction Synthesis with Visual Priors
by: Cai, Songjin, et al.
Published: (2026)
by: Cai, Songjin, et al.
Published: (2026)
DQEN: Dual Query Enhancement Network for DETR-based HOI Detection
by: Li, Zhehao, et al.
Published: (2025)
by: Li, Zhehao, et al.
Published: (2025)
UniHOI: Unified Human-Object Interaction Understanding via Unified Token Space
by: Yang, Panqi, et al.
Published: (2025)
by: Yang, Panqi, et al.
Published: (2025)
Uni-HOI:A Unified framework for Learning the Joint distribution of Text and Human-Object Interaction
by: Zhang, Mengfei, et al.
Published: (2026)
by: Zhang, Mengfei, et al.
Published: (2026)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
by: Zhang, Zhenhao, et al.
Published: (2025)
by: Zhang, Zhenhao, et al.
Published: (2025)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
HOI-Dyn: Learning Interaction Dynamics for Human-Object Motion Diffusion
by: Wu, Lin, et al.
Published: (2025)
by: Wu, Lin, et al.
Published: (2025)
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
by: Yang, Hui, et al.
Published: (2026)
by: Yang, Hui, et al.
Published: (2026)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
by: Xue, Zihui, et al.
Published: (2024)
by: Xue, Zihui, et al.
Published: (2024)
Dex2HOI: Dexterous Bimanual Two-Object Interaction Generation
by: Pratikaki, Chrysa, et al.
Published: (2026)
by: Pratikaki, Chrysa, et al.
Published: (2026)
ContextHOI: Spatial Context Learning for Human-Object Interaction Detection
by: Jia, Mingda, et al.
Published: (2024)
by: Jia, Mingda, et al.
Published: (2024)
HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance
by: Li, Lei, et al.
Published: (2025)
by: Li, Lei, et al.
Published: (2025)
Zero-shot HOI Detection with MLLM-based Detector-agnostic Interaction Recognition
by: Xuan, Shiyu, et al.
Published: (2026)
by: Xuan, Shiyu, et al.
Published: (2026)
PersonaHOI: Effortlessly Improving Personalized Face with Human-Object Interaction Generation
by: Hu, Xinting, et al.
Published: (2025)
by: Hu, Xinting, et al.
Published: (2025)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
by: Liu, Yumeng, et al.
Published: (2024)
by: Liu, Yumeng, et al.
Published: (2024)
Unseen No More: Unlocking the Potential of CLIP for Generative Zero-shot HOI Detection
by: Guo, Yixin, et al.
Published: (2024)
by: Guo, Yixin, et al.
Published: (2024)
What if Agents Could Imagine? Reinforcing Open-Vocabulary HOI Comprehension through Generation
by: Yuan, Zhenlong, et al.
Published: (2026)
by: Yuan, Zhenlong, et al.
Published: (2026)
CycleHOI: Improving Human-Object Interaction Detection with Cycle Consistency of Detection and Generation
by: Wang, Yisen, et al.
Published: (2024)
by: Wang, Yisen, et al.
Published: (2024)
Similar Items
-
EZ-HOI: VLM Adaptation via Guided Prompt Learning for Zero-Shot HOI Detection
by: Lei, Qinqian, et al.
Published: (2024) -
HOLa: Zero-Shot HOI Detection with Low-Rank Decomposed VLM Feature Adaptation
by: Lei, Qinqian, et al.
Published: (2025) -
SHOE: Semantic HOI Open-Vocabulary Evaluation Metric
by: Noack, Maja, et al.
Published: (2026) -
Funnel-HOI: Top-Down Perception for Zero-Shot HOI Detection
by: Sarma, Sandipan, et al.
Published: (2025) -
IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction
by: Zhang, Haoshen, et al.
Published: (2026)