OpenVidVRD: Open-Vocabulary Video Visual Relation Detection via Prompt-Driven Semantic Space Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Qi, Xue, Weiying, Wang, Yuxiao, Wei, Zhenao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Zero-shot Human-Object Interaction Detection via Vision-Language Integration
di: Xue, Weiying, et al.
Pubblicazione: (2024)
di: Xue, Weiying, et al.
Pubblicazione: (2024)
A Review of Human-Object Interaction Detection
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
Prompt Guidance and Human Proximal Perception for HOT Prediction with Regional Joint Loss
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
di: Wang, Yuxiao, et al.
Pubblicazione: (2024)
Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation
di: Truong, Hoang M., et al.
Pubblicazione: (2026)
di: Truong, Hoang M., et al.
Pubblicazione: (2026)
FreeA: Human-object Interaction Detection using Free Annotation Labels
di: Liu, Qi, et al.
Pubblicazione: (2024)
di: Liu, Qi, et al.
Pubblicazione: (2024)
What-Meets-Where: Unified Learning of Action and Contact Localization in Images
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
di: Wang, Yuxiao, et al.
Pubblicazione: (2025)
Exploiting VLM Localizability and Semantics for Open Vocabulary Action Detection
di: Bao, Wentao, et al.
Pubblicazione: (2024)
di: Bao, Wentao, et al.
Pubblicazione: (2024)
Open-Vocabulary Video Anomaly Detection
di: Wu, Peng, et al.
Pubblicazione: (2023)
di: Wu, Peng, et al.
Pubblicazione: (2023)
SDVPT: Semantic-Driven Visual Prompt Tuning for Open-World Object Counting
di: Zhao, Yiming, et al.
Pubblicazione: (2025)
di: Zhao, Yiming, et al.
Pubblicazione: (2025)
Stepping Out of Similar Semantic Space for Open-Vocabulary Segmentation
di: Liu, Yong, et al.
Pubblicazione: (2025)
di: Liu, Yong, et al.
Pubblicazione: (2025)
Open-Vocabulary Object Detection via Neighboring Region Attention Alignment
di: Qiang, Sunyuan, et al.
Pubblicazione: (2024)
di: Qiang, Sunyuan, et al.
Pubblicazione: (2024)
Lost in Translation? Vocabulary Alignment for Source-Free Adaptation in Open-Vocabulary Semantic Segmentation
di: Mazzucco, Silvio, et al.
Pubblicazione: (2025)
di: Mazzucco, Silvio, et al.
Pubblicazione: (2025)
Unified Embedding Alignment for Open-Vocabulary Video Instance Segmentation
di: Fang, Hao, et al.
Pubblicazione: (2024)
di: Fang, Hao, et al.
Pubblicazione: (2024)
Decompose and Transfer: CoT-Prompting Enhanced Alignment for Open-Vocabulary Temporal Action Detection
di: Zhu, Sa, et al.
Pubblicazione: (2026)
di: Zhu, Sa, et al.
Pubblicazione: (2026)
Anomize: Better Open Vocabulary Video Anomaly Detection
di: Li, Fei, et al.
Pubblicazione: (2025)
di: Li, Fei, et al.
Pubblicazione: (2025)
Prototype-Aware Multimodal Alignment for Open-Vocabulary Visual Grounding
di: Xie, Jiangnan, et al.
Pubblicazione: (2025)
di: Xie, Jiangnan, et al.
Pubblicazione: (2025)
A Hierarchical Semantic Distillation Framework for Open-Vocabulary Object Detection
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
UniVid: The Open-Source Unified Video Model
di: Luo, Jiabin, et al.
Pubblicazione: (2025)
di: Luo, Jiabin, et al.
Pubblicazione: (2025)
Open-Vocabulary Object Detection with Meta Prompt Representation and Instance Contrastive Optimization
di: Wang, Zhao, et al.
Pubblicazione: (2024)
di: Wang, Zhao, et al.
Pubblicazione: (2024)
In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation
di: Kang, Dahyun, et al.
Pubblicazione: (2024)
di: Kang, Dahyun, et al.
Pubblicazione: (2024)
Open-Vocabulary HOI Detection with Interaction-aware Prompt and Concept Calibration
di: Lei, Ting, et al.
Pubblicazione: (2025)
di: Lei, Ting, et al.
Pubblicazione: (2025)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
di: Kong, Lingdong, et al.
Pubblicazione: (2024)
Denoise and Align: Diffusion-Driven Foreground Knowledge Prompting for Open-Vocabulary Temporal Action Detection
di: Zhu, Sa, et al.
Pubblicazione: (2026)
di: Zhu, Sa, et al.
Pubblicazione: (2026)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
di: Reichard, Klara, et al.
Pubblicazione: (2025)
di: Reichard, Klara, et al.
Pubblicazione: (2025)
Open-Vocabulary Segmentation with Semantic-Assisted Calibration
di: Liu, Yong, et al.
Pubblicazione: (2023)
di: Liu, Yong, et al.
Pubblicazione: (2023)
FGAseg: Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentation
di: Li, Bingyu, et al.
Pubblicazione: (2025)
di: Li, Bingyu, et al.
Pubblicazione: (2025)
DPSeg: Dual-Prompt Cost Volume Learning for Open-Vocabulary Semantic Segmentation
di: Zhao, Ziyu, et al.
Pubblicazione: (2025)
di: Zhao, Ziyu, et al.
Pubblicazione: (2025)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
di: Zhang, Yupeng, et al.
Pubblicazione: (2025)
di: Zhang, Yupeng, et al.
Pubblicazione: (2025)
Open-Vocabulary Animal Keypoint Detection with Semantic-feature Matching
di: Zhang, Hao, et al.
Pubblicazione: (2023)
di: Zhang, Hao, et al.
Pubblicazione: (2023)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
di: Yang, Shuai, et al.
Pubblicazione: (2026)
di: Yang, Shuai, et al.
Pubblicazione: (2026)
Efficient Redundancy Reduction for Open-Vocabulary Semantic Segmentation
di: Chen, Lin, et al.
Pubblicazione: (2025)
di: Chen, Lin, et al.
Pubblicazione: (2025)
Open-Vocabulary Action Localization with Iterative Visual Prompting
di: Wake, Naoki, et al.
Pubblicazione: (2024)
di: Wake, Naoki, et al.
Pubblicazione: (2024)
Learning Background Prompts to Discover Implicit Knowledge for Open Vocabulary Object Detection
di: Li, Jiaming, et al.
Pubblicazione: (2024)
di: Li, Jiaming, et al.
Pubblicazione: (2024)
Enhancing Open-Vocabulary Object Detection through Multi-Level Fine-Grained Visual-Language Alignment
di: Zhang, Tianyi, et al.
Pubblicazione: (2026)
di: Zhang, Tianyi, et al.
Pubblicazione: (2026)
OpenDPR: Open-Vocabulary Change Detection via Vision-Centric Diffusion-Guided Prototype Retrieval for Remote Sensing Imagery
di: Guo, Qi, et al.
Pubblicazione: (2026)
di: Guo, Qi, et al.
Pubblicazione: (2026)
Part-Aware Open-Vocabulary 3D Affordance Grounding via Prototypical Semantic and Geometric Alignment
di: Gou, Dongqiang, et al.
Pubblicazione: (2026)
di: Gou, Dongqiang, et al.
Pubblicazione: (2026)
VrdONE: One-stage Video Visual Relation Detection
di: Jiang, Xinjie, et al.
Pubblicazione: (2024)
di: Jiang, Xinjie, et al.
Pubblicazione: (2024)
OV-SCAN: Semantically Consistent Alignment for Novel Object Discovery in Open-Vocabulary 3D Object Detection
di: Chow, Adrian, et al.
Pubblicazione: (2025)
di: Chow, Adrian, et al.
Pubblicazione: (2025)
Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection
di: Kim, Youbin, et al.
Pubblicazione: (2026)
di: Kim, Youbin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Towards Zero-shot Human-Object Interaction Detection via Vision-Language Integration
di: Xue, Weiying, et al.
Pubblicazione: (2024) -
A Review of Human-Object Interaction Detection
di: Wang, Yuxiao, et al.
Pubblicazione: (2024) -
Prompt Guidance and Human Proximal Perception for HOT Prediction with Regional Joint Loss
di: Wang, Yuxiao, et al.
Pubblicazione: (2025) -
Precision-Enhanced Human-Object Contact Detection via Depth-Aware Perspective Interaction and Object Texture Restoration
di: Wang, Yuxiao, et al.
Pubblicazione: (2024) -
Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation
di: Truong, Hoang M., et al.
Pubblicazione: (2026)