LV-OSD: Language-Vision-Complementary Open-Set Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yupeng, Han, Ruize, Feng, Wei, Wang, Song, Wan, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2025)
by: Zhang, Yupeng, et al.
Published: (2025)
NoOVD: Novel Category Discovery and Embedding for Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection
by: Zhang, Yupeng, et al.
Published: (2026)
by: Zhang, Yupeng, et al.
Published: (2026)
OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking
by: Liang, Haiji, et al.
Published: (2024)
by: Liang, Haiji, et al.
Published: (2024)
OCTrack: Benchmarking the Open-Corpus Multi-Object Tracking
by: Qian, Zekun, et al.
Published: (2024)
by: Qian, Zekun, et al.
Published: (2024)
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking
by: Qian, Zekun, et al.
Published: (2024)
by: Qian, Zekun, et al.
Published: (2024)
CLIPVehicle: A Unified Framework for Vision-based Vehicle Search
by: Wang, Likai, et al.
Published: (2025)
by: Wang, Likai, et al.
Published: (2025)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
by: Hu, Yupeng, et al.
Published: (2025)
by: Hu, Yupeng, et al.
Published: (2025)
COVTrack++: Learning Open-Vocabulary Multi-Object Tracking from Continuous Videos via a Synergistic Paradigm
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
From Indoor To Outdoor: Unsupervised Domain Adaptive Gait Recognition
by: Wang, Likai, et al.
Published: (2022)
by: Wang, Likai, et al.
Published: (2022)
BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning
by: Qian, Zekun, et al.
Published: (2026)
by: Qian, Zekun, et al.
Published: (2026)
MarvelOVD: Marrying Object Recognition and Vision-Language Models for Robust Open-Vocabulary Object Detection
by: Wang, Kuo, et al.
Published: (2024)
by: Wang, Kuo, et al.
Published: (2024)
From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera Calibration
by: Qian, Zekun, et al.
Published: (2022)
by: Qian, Zekun, et al.
Published: (2022)
Online Reasoning Video Object Segmentation
by: Liu, Jinyuan, et al.
Published: (2026)
by: Liu, Jinyuan, et al.
Published: (2026)
Grounding DINO 1.5: Advance the "Edge" of Open-Set Object Detection
by: Ren, Tianhe, et al.
Published: (2024)
by: Ren, Tianhe, et al.
Published: (2024)
Towards Generalized Few-Shot Open-Set Object Detection
by: Su, Binyi, et al.
Published: (2022)
by: Su, Binyi, et al.
Published: (2022)
YOLO-UniOW: Efficient Universal Open-World Object Detection
by: Liu, Lihao, et al.
Published: (2024)
by: Liu, Lihao, et al.
Published: (2024)
Synthetic-To-Real Video Person Re-ID
by: Zhang, Xiangqun, et al.
Published: (2024)
by: Zhang, Xiangqun, et al.
Published: (2024)
DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding
by: Ren, Tianhe, et al.
Published: (2024)
by: Ren, Tianhe, et al.
Published: (2024)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
by: Liu, Shilong, et al.
Published: (2023)
by: Liu, Shilong, et al.
Published: (2023)
Incremental Object Detection with CLIP
by: Huang, Ziyue, et al.
Published: (2023)
by: Huang, Ziyue, et al.
Published: (2023)
Collaborative Feature-Logits Contrastive Learning for Open-Set Semi-Supervised Object Detection
by: Zhong, Xinhao, et al.
Published: (2024)
by: Zhong, Xinhao, et al.
Published: (2024)
OS-W2S: An Automatic Labeling Engine for Language-Guided Open-Set Aerial Object Detection
by: Wei, Guoting, et al.
Published: (2025)
by: Wei, Guoting, et al.
Published: (2025)
Unveiling the Power of Self-supervision for Multi-view Multi-human Association and Tracking
by: Feng, Wei, et al.
Published: (2024)
by: Feng, Wei, et al.
Published: (2024)
Open-Set Object Detection By Aligning Known Class Representations
by: Sarkar, Hiran, et al.
Published: (2024)
by: Sarkar, Hiran, et al.
Published: (2024)
DetCLIPv3: Towards Versatile Generative Open-vocabulary Object Detection
by: Yao, Lewei, et al.
Published: (2024)
by: Yao, Lewei, et al.
Published: (2024)
UADet: A Remarkably Simple Yet Effective Uncertainty-Aware Open-Set Object Detection Framework
by: Cheng, Silin, et al.
Published: (2024)
by: Cheng, Silin, et al.
Published: (2024)
FM-OSD: Foundation Model-Enabled One-Shot Detection of Anatomical Landmarks
by: Miao, Juzheng, et al.
Published: (2024)
by: Miao, Juzheng, et al.
Published: (2024)
Open-Set Recognition in the Age of Vision-Language Models
by: Miller, Dimity, et al.
Published: (2024)
by: Miller, Dimity, et al.
Published: (2024)
More Pictures Say More: Visual Intersection Network for Open Set Object Detection
by: Dong, Bingcheng, et al.
Published: (2024)
by: Dong, Bingcheng, et al.
Published: (2024)
A Training-Free Guess What Vision Language Model from Snippets to Open-Vocabulary Object Detection
by: Zhu, Guiying, et al.
Published: (2026)
by: Zhu, Guiying, et al.
Published: (2026)
UniArt: Unified 3D Representation for Generating 3D Articulated Objects with Open-Set Articulation
by: Jin, Bu, et al.
Published: (2025)
by: Jin, Bu, et al.
Published: (2025)
A Unified Perspective on Adversarial Membership Manipulation in Vision Models
by: Gao, Ruize, et al.
Published: (2026)
by: Gao, Ruize, et al.
Published: (2026)
OpenPath: Open-Set Active Learning for Pathology Image Classification via Pre-trained Vision-Language Models
by: Zhong, Lanfeng, et al.
Published: (2025)
by: Zhong, Lanfeng, et al.
Published: (2025)
Beyond Known Objects: A Novel Framework for Open-Set Object Detection using Negative-Aware Norm
by: Zhang, Yuchen, et al.
Published: (2026)
by: Zhang, Yuchen, et al.
Published: (2026)
Zero-shot Generalizable Incremental Learning for Vision-Language Object Detection
by: Deng, Jieren, et al.
Published: (2024)
by: Deng, Jieren, et al.
Published: (2024)
OSR-ViT: A Simple and Modular Framework for Open-Set Object Detection and Discovery
by: Inkawhich, Matthew, et al.
Published: (2024)
by: Inkawhich, Matthew, et al.
Published: (2024)
Complementary Frequency-Varying Awareness Network for Open-Set Fine-Grained Image Recognition
by: Dong, Qiulei, et al.
Published: (2023)
by: Dong, Qiulei, et al.
Published: (2023)
$\mathbf{C}^2$Former: Calibrated and Complementary Transformer for RGB-Infrared Object Detection
by: Yuan, Maoxun, et al.
Published: (2023)
by: Yuan, Maoxun, et al.
Published: (2023)
Similar Items
-
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2025) -
NoOVD: Novel Category Discovery and Embedding for Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2026) -
VFM$^{4}$SDG: Unveiling the Power of VFMs for Single-Domain Generalized Object Detection
by: Zhang, Yupeng, et al.
Published: (2026) -
COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection
by: Zhang, Yupeng, et al.
Published: (2026) -
OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking
by: Liang, Haiji, et al.
Published: (2024)