Towards Zero-shot Human-Object Interaction Detection via Vision-Language Integration
Fuente:
arXiv
Salvato in:
| Autori principali: | Xue, Weiying, Liu, Qi, Xiong, Qiwei, Wang, Yuxiao, Wei, Zhenao, Xing, Xiaofen, Xu, Xiangmin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Review of Human-Object Interaction Detection
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
OpenVidVRD: Open-Vocabulary Video Visual Relation Detection via Prompt-Driven Semantic Space Alignment
di: Liu, Qi, et al.
Pubblicazione: (2025)
di: Liu, Qi, et al.
Pubblicazione: (2025)
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
FreeA: Human-object Interaction Detection using Free Annotation Labels
di: Liu, Qi, et al.
Pubblicazione: (2024)
di: Liu, Qi, et al.
Pubblicazione: (2024)
Prompt Guidance and Human Proximal Perception for HOT Prediction with Regional Joint Loss
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
PointCore: Efficient Unsupervised Point Cloud Anomaly Detector Using Local-Global Features
di: Zhao, Baozhu, et al.
Pubblicazione: (2024)
di: Zhao, Baozhu, et al.
Pubblicazione: (2024)
What-Meets-Where: Unified Learning of Action and Contact Localization in Images
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
di: Hu, Yupeng, et al.
Pubblicazione: (2025)
di: Hu, Yupeng, et al.
Pubblicazione: (2025)
Disentangled Pre-training for Human-Object Interaction Detection
di: Li, Zhuolong, et al.
Pubblicazione: (2024)
di: Li, Zhuolong, et al.
Pubblicazione: (2024)
Zero-shot Generalizable Incremental Learning for Vision-Language Object Detection
di: Deng, Jieren, et al.
Pubblicazione: (2024)
di: Deng, Jieren, et al.
Pubblicazione: (2024)
InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
di: Zhang, Jinlu, et al.
Pubblicazione: (2025)
di: Zhang, Jinlu, et al.
Pubblicazione: (2025)
Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors
di: Chen, Jiahe, et al.
Pubblicazione: (2026)
di: Chen, Jiahe, et al.
Pubblicazione: (2026)
AnchorHOI: Zero-shot Generation of 4D Human-Object Interaction via Anchor-based Prior Distillation
di: Dai, Sisi, et al.
Pubblicazione: (2025)
di: Dai, Sisi, et al.
Pubblicazione: (2025)
Boosting Single-domain Generalized Object Detection via Vision-Language Knowledge Interaction
di: Xu, Xiaoran, et al.
Pubblicazione: (2025)
di: Xu, Xiaoran, et al.
Pubblicazione: (2025)
AvatarGO: Zero-shot 4D Human-Object Interaction Generation and Animation
di: Cao, Yukang, et al.
Pubblicazione: (2024)
di: Cao, Yukang, et al.
Pubblicazione: (2024)
Compact Model Training by Low-Rank Projection with Energy Transfer
di: Guo, Kailing, et al.
Pubblicazione: (2022)
di: Guo, Kailing, et al.
Pubblicazione: (2022)
MimicParts: Part-aware Style Injection for Speech-Driven 3D Motion Generation
di: Liu, Lianlian, et al.
Pubblicazione: (2025)
di: Liu, Lianlian, et al.
Pubblicazione: (2025)
Exploring Hyperspectral Anomaly Detection with Human Vision: A Small Target Aware Detector
di: Ma, Jitao, et al.
Pubblicazione: (2024)
di: Ma, Jitao, et al.
Pubblicazione: (2024)
VideoCoT: A Video Chain-of-Thought Dataset with Active Annotation Tool
di: Wang, Yan, et al.
Pubblicazione: (2024)
di: Wang, Yan, et al.
Pubblicazione: (2024)
Explanatory Instructions: Towards Unified Vision Tasks Understanding and Zero-shot Generalization
di: Shen, Yang, et al.
Pubblicazione: (2024)
di: Shen, Yang, et al.
Pubblicazione: (2024)
Locality-Aware Zero-Shot Human-Object Interaction Detection
di: Kim, Sanghyun, et al.
Pubblicazione: (2025)
di: Kim, Sanghyun, et al.
Pubblicazione: (2025)
Zero-shot Action Localization via the Confidence of Large Vision-Language Models
di: Aklilu, Josiah, et al.
Pubblicazione: (2024)
di: Aklilu, Josiah, et al.
Pubblicazione: (2024)
RATE-Nav: Region-Aware Termination Enhancement for Zero-shot Object Navigation with Vision-Language Models
di: Li, Junjie, et al.
Pubblicazione: (2025)
di: Li, Junjie, et al.
Pubblicazione: (2025)
InterDreamer: Zero-Shot Text to 3D Dynamic Human-Object Interaction
di: Xu, Sirui, et al.
Pubblicazione: (2024)
di: Xu, Sirui, et al.
Pubblicazione: (2024)
Task-Specific Zero-shot Quantization-Aware Training for Object Detection
di: Li, Changhao, et al.
Pubblicazione: (2025)
di: Li, Changhao, et al.
Pubblicazione: (2025)
Exploring Fine-grained Retail Product Discrimination with Zero-shot Object Classification Using Vision-Language Models
di: Tur, Anil Osman, et al.
Pubblicazione: (2024)
di: Tur, Anil Osman, et al.
Pubblicazione: (2024)
ZeroComp: Zero-shot Object Compositing from Image Intrinsics via Diffusion
di: Zhang, Zitian, et al.
Pubblicazione: (2024)
di: Zhang, Zitian, et al.
Pubblicazione: (2024)
Chain of Visual Perception: Harnessing Multimodal Large Language Models for Zero-shot Camouflaged Object Detection
di: Tang, Lv, et al.
Pubblicazione: (2023)
di: Tang, Lv, et al.
Pubblicazione: (2023)
BackdoorIDS: Zero-shot Backdoor Detection for Pretrained Vision Encoder
di: Huang, Siquan, et al.
Pubblicazione: (2026)
di: Huang, Siquan, et al.
Pubblicazione: (2026)
AnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly Detection
di: Zhou, Qihang, et al.
Pubblicazione: (2023)
di: Zhou, Qihang, et al.
Pubblicazione: (2023)
Zero-shot Object Counting with Good Exemplars
di: Zhu, Huilin, et al.
Pubblicazione: (2024)
di: Zhu, Huilin, et al.
Pubblicazione: (2024)
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
di: Wang, Zhenzhi, et al.
Pubblicazione: (2023)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2023)
Mining Instance-Centric Vision-Language Contexts for Human-Object Interaction Detection
di: Seo, Soo Won, et al.
Pubblicazione: (2026)
di: Seo, Soo Won, et al.
Pubblicazione: (2026)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2025)
di: Hoe, Jiun Tian, et al.
Pubblicazione: (2025)
AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models
di: Wu, Yongjian, et al.
Pubblicazione: (2024)
di: Wu, Yongjian, et al.
Pubblicazione: (2024)
Towards Unconstrained Human-Object Interaction
di: Tonini, Francesco, et al.
Pubblicazione: (2026)
di: Tonini, Francesco, et al.
Pubblicazione: (2026)
ViTGaze: Gaze Following with Interaction Features in Vision Transformers
di: Song, Yuehao, et al.
Pubblicazione: (2024)
di: Song, Yuehao, et al.
Pubblicazione: (2024)
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
di: Chen, Junwen, et al.
Pubblicazione: (2025)
di: Chen, Junwen, et al.
Pubblicazione: (2025)
Efficient Human-Object-Interaction (EHOI) Detection via Interaction Label Coding and Conditional Decision
di: Yang, Tsung-Shan, et al.
Pubblicazione: (2024)
di: Yang, Tsung-Shan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Review of Human-Object Interaction Detection
di: Wang, Yuxiao, et al.
Pubblicazione: (2024) -
OpenVidVRD: Open-Vocabulary Video Visual Relation Detection via Prompt-Driven Semantic Space Alignment
di: Liu, Qi, et al.
Pubblicazione: (2025) -
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
di: Wang, Yuxiao, et al.
Pubblicazione: (2024) -
FreeA: Human-object Interaction Detection using Free Annotation Labels
di: Liu, Qi, et al.
Pubblicazione: (2024) -
Prompt Guidance and Human Proximal Perception for HOT Prediction with Regional Joint Loss
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)