Unlocking Textual and Visual Wisdom: Open-Vocabulary 3D Object Detection Enhanced by Comprehensive Guidance from Text and Image
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiao, Pengkun, Zhao, Na, Chen, Jingjing, Jiang, Yu-Gang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Domain Expansion and Boundary Growth for Open-Set Single-Source Domain Generalization
von: Jiao, Pengkun, et al.
Veröffentlicht: (2024)
von: Jiao, Pengkun, et al.
Veröffentlicht: (2024)
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning
von: Jiao, Pengkun, et al.
Veröffentlicht: (2024)
von: Jiao, Pengkun, et al.
Veröffentlicht: (2024)
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024)
von: Yao, Jin, et al.
Veröffentlicht: (2024)
RT-OVAD: Real-Time Open-Vocabulary Aerial Object Detection via Image-Text Collaboration
von: Wei, Guoting, et al.
Veröffentlicht: (2024)
von: Wei, Guoting, et al.
Veröffentlicht: (2024)
Visual Textualization for Image Prompted Object Detection
von: Wu, Yongjian, et al.
Veröffentlicht: (2025)
von: Wu, Yongjian, et al.
Veröffentlicht: (2025)
Comprehensive Multi-Modal Prototypes are Simple and Effective Classifiers for Vast-Vocabulary Object Detection
von: Chen, Yitong, et al.
Veröffentlicht: (2024)
von: Chen, Yitong, et al.
Veröffentlicht: (2024)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
von: Zhao, Youjun, et al.
Veröffentlicht: (2025)
von: Zhao, Youjun, et al.
Veröffentlicht: (2025)
Enhancing Open-Vocabulary Object Detection through Multi-Level Fine-Grained Visual-Language Alignment
von: Zhang, Tianyi, et al.
Veröffentlicht: (2026)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2026)
V3Det Challenge 2024 on Vast Vocabulary and Open Vocabulary Object Detection: Methods and Results
von: Wang, Jiaqi, et al.
Veröffentlicht: (2024)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2024)
EAGLE: Towards Efficient Arbitrary Referring Visual Prompts Comprehension for Multimodal Large Language Models
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
Open-Vocabulary 3D Semantic Segmentation with Text-to-Image Diffusion Models
von: Zhu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaoyu, et al.
Veröffentlicht: (2024)
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
RTGen: Generating Region-Text Pairs for Open-Vocabulary Object Detection
von: Chen, Fangyi, et al.
Veröffentlicht: (2024)
von: Chen, Fangyi, et al.
Veröffentlicht: (2024)
Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models
von: Tang, Ziyao, et al.
Veröffentlicht: (2026)
von: Tang, Ziyao, et al.
Veröffentlicht: (2026)
Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting
von: Ruis, Frank, et al.
Veröffentlicht: (2025)
von: Ruis, Frank, et al.
Veröffentlicht: (2025)
Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
von: Jiao, Pengkun, et al.
Veröffentlicht: (2025)
von: Jiao, Pengkun, et al.
Veröffentlicht: (2025)
Uncertainty Meets Diversity: A Comprehensive Active Learning Framework for Indoor 3D Object Detection
von: Wang, Jiangyi, et al.
Veröffentlicht: (2025)
von: Wang, Jiangyi, et al.
Veröffentlicht: (2025)
Taming Self-Training for Open-Vocabulary Object Detection
von: Zhao, Shiyu, et al.
Veröffentlicht: (2023)
von: Zhao, Shiyu, et al.
Veröffentlicht: (2023)
Scaling Open-Vocabulary Object Detection
von: Minderer, Matthias, et al.
Veröffentlicht: (2023)
von: Minderer, Matthias, et al.
Veröffentlicht: (2023)
OpenNav: Efficient Open Vocabulary 3D Object Detection for Smart Wheelchair Navigation
von: Rahman, Muhammad Rameez ur, et al.
Veröffentlicht: (2024)
von: Rahman, Muhammad Rameez ur, et al.
Veröffentlicht: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023)
von: Nguyen, Phuc D. A., et al.
Veröffentlicht: (2023)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
von: Huang, Tianyu, et al.
Veröffentlicht: (2023)
OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations
von: Hsu, Peng-Hao, et al.
Veröffentlicht: (2025)
von: Hsu, Peng-Hao, et al.
Veröffentlicht: (2025)
Parameter-Efficient Semantic Augmentation for Enhancing Open-Vocabulary Object Detection
von: Cao, Weihao, et al.
Veröffentlicht: (2026)
von: Cao, Weihao, et al.
Veröffentlicht: (2026)
Exploring Open-Vocabulary Object Recognition in Images using CLIP
von: Chen, Wei Yu, et al.
Veröffentlicht: (2026)
von: Chen, Wei Yu, et al.
Veröffentlicht: (2026)
Evaluating the Performance of Open-Vocabulary Object Detection in Low-quality Image
von: Wu, Po-Chih
Veröffentlicht: (2025)
von: Wu, Po-Chih
Veröffentlicht: (2025)
Collaborative Vision-Text Representation Optimizing for Open-Vocabulary Segmentation
von: Jiao, Siyu, et al.
Veröffentlicht: (2024)
von: Jiao, Siyu, et al.
Veröffentlicht: (2024)
SceneAssistant: A Visual Feedback Agent for Open-Vocabulary 3D Scene Generation
von: Luo, Jun, et al.
Veröffentlicht: (2026)
von: Luo, Jun, et al.
Veröffentlicht: (2026)
Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
von: Yuan, Zhihao, et al.
Veröffentlicht: (2023)
von: Yuan, Zhihao, et al.
Veröffentlicht: (2023)
Auto-Vocabulary 3D Object Detection
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025)
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025)
Learning to Detect and Segment for Open Vocabulary Object Detection
von: Wang, Tao, et al.
Veröffentlicht: (2022)
von: Wang, Tao, et al.
Veröffentlicht: (2022)
Retrieval-Augmented Open-Vocabulary Object Detection
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024)
Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection
von: Kim, Youbin, et al.
Veröffentlicht: (2026)
von: Kim, Youbin, et al.
Veröffentlicht: (2026)
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
von: Huang, Rui, et al.
Veröffentlicht: (2024)
von: Huang, Rui, et al.
Veröffentlicht: (2024)
ImOV3D: Learning Open-Vocabulary Point Clouds 3D Object Detection from Only 2D Images
von: Yang, Timing, et al.
Veröffentlicht: (2024)
von: Yang, Timing, et al.
Veröffentlicht: (2024)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
von: Zhang, Yupeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yupeng, et al.
Veröffentlicht: (2025)
State and Scene Enhanced Prototypes for Weakly Supervised Open-Vocabulary Object Detection
von: Zhou, Jiaying, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaying, et al.
Veröffentlicht: (2025)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
von: Yang, Shuai, et al.
Veröffentlicht: (2026)
von: Yang, Shuai, et al.
Veröffentlicht: (2026)
Open-Vocabulary Camouflaged Object Segmentation
von: Pang, Youwei, et al.
Veröffentlicht: (2023)
von: Pang, Youwei, et al.
Veröffentlicht: (2023)
OV-SCAN: Semantically Consistent Alignment for Novel Object Discovery in Open-Vocabulary 3D Object Detection
von: Chow, Adrian, et al.
Veröffentlicht: (2025)
von: Chow, Adrian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Domain Expansion and Boundary Growth for Open-Set Single-Source Domain Generalization
von: Jiao, Pengkun, et al.
Veröffentlicht: (2024) -
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning
von: Jiao, Pengkun, et al.
Veröffentlicht: (2024) -
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024) -
RT-OVAD: Real-Time Open-Vocabulary Aerial Object Detection via Image-Text Collaboration
von: Wei, Guoting, et al.
Veröffentlicht: (2024) -
Visual Textualization for Image Prompted Object Detection
von: Wu, Yongjian, et al.
Veröffentlicht: (2025)