Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lei, Ting, Yin, Shaofeng, Liu, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
von: Lei, Ting, et al.
Veröffentlicht: (2025)
von: Lei, Ting, et al.
Veröffentlicht: (2025)
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection
von: Lei, Ting, et al.
Veröffentlicht: (2024)
von: Lei, Ting, et al.
Veröffentlicht: (2024)
SGC-Net: Stratified Granular Comparison Network for Open-Vocabulary HOI Detection
von: Lin, Xin, et al.
Veröffentlicht: (2025)
von: Lin, Xin, et al.
Veröffentlicht: (2025)
SHOE: Semantic HOI Open-Vocabulary Evaluation Metric
von: Noack, Maja, et al.
Veröffentlicht: (2026)
von: Noack, Maja, et al.
Veröffentlicht: (2026)
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)
What if Agents Could Imagine? Reinforcing Open-Vocabulary HOI Comprehension through Generation
von: Yuan, Zhenlong, et al.
Veröffentlicht: (2026)
von: Yuan, Zhenlong, et al.
Veröffentlicht: (2026)
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
von: Chen, Junwen, et al.
Veröffentlicht: (2025)
von: Chen, Junwen, et al.
Veröffentlicht: (2025)
Exploring Interactive Semantic Alignment for Efficient HOI Detection with Vision-language Model
von: Dong, Jihao, et al.
Veröffentlicht: (2024)
von: Dong, Jihao, et al.
Veröffentlicht: (2024)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenhao, et al.
Veröffentlicht: (2025)
CrossHOI-Bench: A Unified Benchmark for HOI Evaluation across Vision-Language Models and HOI-Specific Methods
von: Lei, Qinqian, et al.
Veröffentlicht: (2025)
von: Lei, Qinqian, et al.
Veröffentlicht: (2025)
EZ-HOI: VLM Adaptation via Guided Prompt Learning for Zero-Shot HOI Detection
von: Lei, Qinqian, et al.
Veröffentlicht: (2024)
von: Lei, Qinqian, et al.
Veröffentlicht: (2024)
On the Potential of Open-Vocabulary Models for Object Detection in Unusual Street Scenes
von: Ilyas, Sadia, et al.
Veröffentlicht: (2024)
von: Ilyas, Sadia, et al.
Veröffentlicht: (2024)
Enhancing HOI Detection with Contextual Cues from Large Vision-Language Models
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2023)
von: Zhan, Yu-Wei, et al.
Veröffentlicht: (2023)
Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation
von: Shao, Tong, et al.
Veröffentlicht: (2024)
von: Shao, Tong, et al.
Veröffentlicht: (2024)
Unseen No More: Unlocking the Potential of CLIP for Generative Zero-shot HOI Detection
von: Guo, Yixin, et al.
Veröffentlicht: (2024)
von: Guo, Yixin, et al.
Veröffentlicht: (2024)
FrozenSeg: Harmonizing Frozen Foundation Models for Open-Vocabulary Segmentation
von: Chen, Xi, et al.
Veröffentlicht: (2024)
von: Chen, Xi, et al.
Veröffentlicht: (2024)
DAG: Unleash the Potential of Diffusion Model for Open-Vocabulary 3D Affordance Grounding
von: Wang, Hanqing, et al.
Veröffentlicht: (2025)
von: Wang, Hanqing, et al.
Veröffentlicht: (2025)
Exploiting Unlabeled Data with Multiple Expert Teachers for Open Vocabulary Aerial Object Detection and Its Orientation Adaptation
von: Li, Yan, et al.
Veröffentlicht: (2024)
von: Li, Yan, et al.
Veröffentlicht: (2024)
Funnel-HOI: Top-Down Perception for Zero-Shot HOI Detection
von: Sarma, Sandipan, et al.
Veröffentlicht: (2025)
von: Sarma, Sandipan, et al.
Veröffentlicht: (2025)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
von: Lee, Sanghoon, et al.
Veröffentlicht: (2026)
von: Lee, Sanghoon, et al.
Veröffentlicht: (2026)
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking
von: Qian, Zekun, et al.
Veröffentlicht: (2024)
von: Qian, Zekun, et al.
Veröffentlicht: (2024)
Open-Vocabulary Video Anomaly Detection
von: Wu, Peng, et al.
Veröffentlicht: (2023)
von: Wu, Peng, et al.
Veröffentlicht: (2023)
Boosting Open-Vocabulary Object Detection by Handling Background Samples
von: Zeng, Ruizhe, et al.
Veröffentlicht: (2024)
von: Zeng, Ruizhe, et al.
Veröffentlicht: (2024)
EasyHOI: Unleashing the Power of Large Models for Reconstructing Hand-Object Interactions in the Wild
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
von: Liu, Yumeng, et al.
Veröffentlicht: (2024)
RSVG-ZeroOV: Exploring a Training-Free Framework for Zero-Shot Open-Vocabulary Visual Grounding in Remote Sensing Images
von: Li, Ke, et al.
Veröffentlicht: (2025)
von: Li, Ke, et al.
Veröffentlicht: (2025)
UAHOI: Uncertainty-aware Robust Interaction Learning for HOI Detection
von: Chen, Mu, et al.
Veröffentlicht: (2024)
von: Chen, Mu, et al.
Veröffentlicht: (2024)
Scaling Open-Vocabulary Action Detection
von: Sia, Zhen Hao, et al.
Veröffentlicht: (2025)
von: Sia, Zhen Hao, et al.
Veröffentlicht: (2025)
Scaling Open-Vocabulary Object Detection
von: Minderer, Matthias, et al.
Veröffentlicht: (2023)
von: Minderer, Matthias, et al.
Veröffentlicht: (2023)
Unleashing the Multi-View Fusion Potential: Noise Correction in VLM for Open-Vocabulary 3D Scene Understanding
von: Yin, Xingyilang, et al.
Veröffentlicht: (2025)
von: Yin, Xingyilang, et al.
Veröffentlicht: (2025)
HOLa: Zero-Shot HOI Detection with Low-Rank Decomposed VLM Feature Adaptation
von: Lei, Qinqian, et al.
Veröffentlicht: (2025)
von: Lei, Qinqian, et al.
Veröffentlicht: (2025)
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction
von: Du, Penghui, et al.
Veröffentlicht: (2024)
von: Du, Penghui, et al.
Veröffentlicht: (2024)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
von: Hu, Yupeng, et al.
Veröffentlicht: (2025)
von: Hu, Yupeng, et al.
Veröffentlicht: (2025)
A Training-Free Guess What Vision Language Model from Snippets to Open-Vocabulary Object Detection
von: Zhu, Guiying, et al.
Veröffentlicht: (2026)
von: Zhu, Guiying, et al.
Veröffentlicht: (2026)
Hierarchically-Structured Open-Vocabulary Indoor Scene Synthesis with Pre-trained Large Language Model
von: Sun, Weilin, et al.
Veröffentlicht: (2025)
von: Sun, Weilin, et al.
Veröffentlicht: (2025)
Anomize: Better Open Vocabulary Video Anomaly Detection
von: Li, Fei, et al.
Veröffentlicht: (2025)
von: Li, Fei, et al.
Veröffentlicht: (2025)
Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation
von: Shi, Yuheng, et al.
Veröffentlicht: (2024)
von: Shi, Yuheng, et al.
Veröffentlicht: (2024)
Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation
von: Lee, Junha, et al.
Veröffentlicht: (2025)
von: Lee, Junha, et al.
Veröffentlicht: (2025)
DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection
von: Jiang, Donghong, et al.
Veröffentlicht: (2026)
von: Jiang, Donghong, et al.
Veröffentlicht: (2026)
Learning to Detect and Segment for Open Vocabulary Object Detection
von: Wang, Tao, et al.
Veröffentlicht: (2022)
von: Wang, Tao, et al.
Veröffentlicht: (2022)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
von: Yang, Shuai, et al.
Veröffentlicht: (2026)
von: Yang, Shuai, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
von: Lei, Ting, et al.
Veröffentlicht: (2025) -
Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection
von: Lei, Ting, et al.
Veröffentlicht: (2024) -
SGC-Net: Stratified Granular Comparison Network for Open-Vocabulary HOI Detection
von: Lin, Xin, et al.
Veröffentlicht: (2025) -
SHOE: Semantic HOI Open-Vocabulary Evaluation Metric
von: Noack, Maja, et al.
Veröffentlicht: (2026) -
ScriptHOI: Learning Scripted State Transitions for Open-Vocabulary Human-Object Interaction Detection
von: Nguyen, Minh Anh, et al.
Veröffentlicht: (2026)