Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
Fuente:
arXiv
Saved in:
| Main Authors: | Xiang, Xinhao, Peng, Kuan-Chuan, Lohit, Suhas, Jones, Michael J., Zhang, Jiawei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Auto-Vocabulary 3D Object Detection
by: Zhang, Haomeng, et al.
Published: (2025)
by: Zhang, Haomeng, et al.
Published: (2025)
Multimodal 3D Object Detection on Unseen Domains
by: Hegde, Deepti, et al.
Published: (2024)
by: Hegde, Deepti, et al.
Published: (2024)
Equivariant Spatio-Temporal Self-Supervision for LiDAR Object Detection
by: Hegde, Deepti, et al.
Published: (2024)
by: Hegde, Deepti, et al.
Published: (2024)
WISE: Weighted Iterative Society-of-Experts for Robust Multimodal Multi-Agent Debate
by: Cherian, Anoop, et al.
Published: (2025)
by: Cherian, Anoop, et al.
Published: (2025)
Open Vocabulary Monocular 3D Object Detection
by: Yao, Jin, et al.
Published: (2024)
by: Yao, Jin, et al.
Published: (2024)
Improving Open-World Object Localization by Discovering Background
by: Singh, Ashish, et al.
Published: (2025)
by: Singh, Ashish, et al.
Published: (2025)
Multimodal Diffusion Bridge with Attention-Based SAR Fusion for Satellite Image Cloud Removal
by: Hu, Yuyang, et al.
Published: (2025)
by: Hu, Yuyang, et al.
Published: (2025)
TTSA3R: Training-Free Temporal-Spatial Adaptive Persistent State for Streaming 3D Reconstruction
by: Zheng, Zhijie, et al.
Published: (2026)
by: Zheng, Zhijie, et al.
Published: (2026)
OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations
by: Hsu, Peng-Hao, et al.
Published: (2025)
by: Hsu, Peng-Hao, et al.
Published: (2025)
Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking
by: Ishaq, Ayesha, et al.
Published: (2024)
by: Ishaq, Ayesha, et al.
Published: (2024)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
by: Tai, Hanchen, et al.
Published: (2024)
by: Tai, Hanchen, et al.
Published: (2024)
Towards Zero-shot 3D Anomaly Localization
by: Wang, Yizhou, et al.
Published: (2024)
by: Wang, Yizhou, et al.
Published: (2024)
Retrieval-Augmented Open-Vocabulary Object Detection
by: Kim, Jooyeon, et al.
Published: (2024)
by: Kim, Jooyeon, et al.
Published: (2024)
OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection
by: Xia, Zhongyu, et al.
Published: (2024)
by: Xia, Zhongyu, et al.
Published: (2024)
Scaling Open-Vocabulary Object Detection
by: Minderer, Matthias, et al.
Published: (2023)
by: Minderer, Matthias, et al.
Published: (2023)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025)
by: Zhao, Youjun, et al.
Published: (2025)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
by: Zhang, Yupeng, et al.
Published: (2025)
by: Zhang, Yupeng, et al.
Published: (2025)
OpenNav: Efficient Open Vocabulary 3D Object Detection for Smart Wheelchair Navigation
by: Rahman, Muhammad Rameez ur, et al.
Published: (2024)
by: Rahman, Muhammad Rameez ur, et al.
Published: (2024)
Beyond Open Vocabulary: Multimodal Prompting for Object Detection in Remote Sensing Images
by: Yang, Shuai, et al.
Published: (2026)
by: Yang, Shuai, et al.
Published: (2026)
Evaluating Large Vision-and-Language Models on Children's Mathematical Olympiads
by: Cherian, Anoop, et al.
Published: (2024)
by: Cherian, Anoop, et al.
Published: (2024)
Joint Training of Image Generator and Detector for Road Defect Detection
by: Peng, Kuan-Chuan
Published: (2025)
by: Peng, Kuan-Chuan
Published: (2025)
Learning to Detect and Segment for Open Vocabulary Object Detection
by: Wang, Tao, et al.
Published: (2022)
by: Wang, Tao, et al.
Published: (2022)
Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection
by: Kim, Youbin, et al.
Published: (2026)
by: Kim, Youbin, et al.
Published: (2026)
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
by: Huang, Rui, et al.
Published: (2024)
by: Huang, Rui, et al.
Published: (2024)
OV-Uni3DETR: Towards Unified Open-Vocabulary 3D Object Detection via Cycle-Modality Propagation
by: Wang, Zhenyu, et al.
Published: (2024)
by: Wang, Zhenyu, et al.
Published: (2024)
Taming Self-Training for Open-Vocabulary Object Detection
by: Zhao, Shiyu, et al.
Published: (2023)
by: Zhao, Shiyu, et al.
Published: (2023)
Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding
by: Li, Ruihuang, et al.
Published: (2024)
by: Li, Ruihuang, et al.
Published: (2024)
Open-Vocabulary Video Anomaly Detection
by: Wu, Peng, et al.
Published: (2023)
by: Wu, Peng, et al.
Published: (2023)
Open-Vocabulary Object Detection via Neighboring Region Attention Alignment
by: Qiang, Sunyuan, et al.
Published: (2024)
by: Qiang, Sunyuan, et al.
Published: (2024)
LOVON: Legged Open-Vocabulary Object Navigator
by: Peng, Daojie, et al.
Published: (2025)
by: Peng, Daojie, et al.
Published: (2025)
Boosting Open-Vocabulary Object Detection by Handling Background Samples
by: Zeng, Ruizhe, et al.
Published: (2024)
by: Zeng, Ruizhe, et al.
Published: (2024)
OV-SCAN: Semantically Consistent Alignment for Novel Object Discovery in Open-Vocabulary 3D Object Detection
by: Chow, Adrian, et al.
Published: (2025)
by: Chow, Adrian, et al.
Published: (2025)
V3Det Challenge 2024 on Vast Vocabulary and Open Vocabulary Object Detection: Methods and Results
by: Wang, Jiaqi, et al.
Published: (2024)
by: Wang, Jiaqi, et al.
Published: (2024)
MQADet: A Plug-and-Play Paradigm for Enhancing Open-Vocabulary Object Detection via Multimodal Question Answering
by: Li, Caixiong, et al.
Published: (2025)
by: Li, Caixiong, et al.
Published: (2025)
Toward Open Vocabulary Aerial Object Detection with CLIP-Activated Student-Teacher Learning
by: Li, Yan, et al.
Published: (2023)
by: Li, Yan, et al.
Published: (2023)
WeDetect: Fast Open-Vocabulary Object Detection as Retrieval
by: Fu, Shenghao, et al.
Published: (2025)
by: Fu, Shenghao, et al.
Published: (2025)
Streamlined Open-Vocabulary Human-Object Interaction Detection
by: Sun, Chang, et al.
Published: (2026)
by: Sun, Chang, et al.
Published: (2026)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
by: Lee, Sanghoon, et al.
Published: (2026)
by: Lee, Sanghoon, et al.
Published: (2026)
Sparse Multiview Open-Vocabulary 3D Detection
by: Moliner, Olivier, et al.
Published: (2025)
by: Moliner, Olivier, et al.
Published: (2025)
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Similar Items
-
Auto-Vocabulary 3D Object Detection
by: Zhang, Haomeng, et al.
Published: (2025) -
Multimodal 3D Object Detection on Unseen Domains
by: Hegde, Deepti, et al.
Published: (2024) -
Equivariant Spatio-Temporal Self-Supervision for LiDAR Object Detection
by: Hegde, Deepti, et al.
Published: (2024) -
WISE: Weighted Iterative Society-of-Experts for Robust Multimodal Multi-Agent Debate
by: Cherian, Anoop, et al.
Published: (2025) -
Open Vocabulary Monocular 3D Object Detection
by: Yao, Jin, et al.
Published: (2024)