GLRD: Global-Local Collaborative Reason and Debate with PSL for 3D Open-Vocabulary Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Xingyu, Liu, Si, Gao, Chen, Bai, Yan, Mu, Beipeng, Wang, Xiaofei, Xia, Huaxia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Global-Local Collaborative Inference with LLM for Lidar-Based Open-Vocabulary Detection
von: Peng, Xingyu, et al.
Veröffentlicht: (2024)
von: Peng, Xingyu, et al.
Veröffentlicht: (2024)
Eliminating Cross-modal Conflicts in BEV Space for LiDAR-Camera 3D Object Detection
von: Fu, Jiahui, et al.
Veröffentlicht: (2024)
von: Fu, Jiahui, et al.
Veröffentlicht: (2024)
RATopo: Improving Lane Topology Reasoning via Redundancy Assignment
von: Li, Han, et al.
Veröffentlicht: (2025)
von: Li, Han, et al.
Veröffentlicht: (2025)
Unified Modeling of Lane and Lane Topology for Driving Scene Reasoning
von: Li, Han, et al.
Veröffentlicht: (2026)
von: Li, Han, et al.
Veröffentlicht: (2026)
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
von: Huang, Rui, et al.
Veröffentlicht: (2024)
von: Huang, Rui, et al.
Veröffentlicht: (2024)
COS3D: Collaborative Open-Vocabulary 3D Segmentation
von: Zhu, Runsong, et al.
Veröffentlicht: (2025)
von: Zhu, Runsong, et al.
Veröffentlicht: (2025)
Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
von: Xiang, Xinhao, et al.
Veröffentlicht: (2025)
RT-OVAD: Real-Time Open-Vocabulary Aerial Object Detection via Image-Text Collaboration
von: Wei, Guoting, et al.
Veröffentlicht: (2024)
von: Wei, Guoting, et al.
Veröffentlicht: (2024)
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024)
von: Yao, Jin, et al.
Veröffentlicht: (2024)
Sparse Multiview Open-Vocabulary 3D Detection
von: Moliner, Olivier, et al.
Veröffentlicht: (2025)
von: Moliner, Olivier, et al.
Veröffentlicht: (2025)
Open-Vocabulary Video Anomaly Detection
von: Wu, Peng, et al.
Veröffentlicht: (2023)
von: Wu, Peng, et al.
Veröffentlicht: (2023)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
von: Sun, Haowen, et al.
Veröffentlicht: (2026)
von: Sun, Haowen, et al.
Veröffentlicht: (2026)
OpenDAS: Open-Vocabulary Domain Adaptation for 2D and 3D Segmentation
von: Yilmaz, Gonca, et al.
Veröffentlicht: (2024)
von: Yilmaz, Gonca, et al.
Veröffentlicht: (2024)
MonoSR: Open-Vocabulary Spatial Reasoning from Monocular Images
von: Wang, Qirui, et al.
Veröffentlicht: (2025)
von: Wang, Qirui, et al.
Veröffentlicht: (2025)
PGOV3D: Open-Vocabulary 3D Semantic Segmentation with Partial-to-Global Curriculum
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
Unsupervised Open-Vocabulary Object Localization in Videos
von: Fan, Ke, et al.
Veröffentlicht: (2023)
von: Fan, Ke, et al.
Veröffentlicht: (2023)
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
von: Cao, Yang, et al.
Veröffentlicht: (2024)
von: Cao, Yang, et al.
Veröffentlicht: (2024)
OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations
von: Hsu, Peng-Hao, et al.
Veröffentlicht: (2025)
von: Hsu, Peng-Hao, et al.
Veröffentlicht: (2025)
Auto-Vocabulary 3D Object Detection
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025)
von: Zhang, Haomeng, et al.
Veröffentlicht: (2025)
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction
von: Du, Penghui, et al.
Veröffentlicht: (2024)
von: Du, Penghui, et al.
Veröffentlicht: (2024)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation
von: Zhou, Zhishan, et al.
Veröffentlicht: (2025)
von: Zhou, Zhishan, et al.
Veröffentlicht: (2025)
MemOVCD: Training-Free Open-Vocabulary Change Detection via Cross-Temporal Memory Reasoning and Global-Local Adaptive Rectification
von: Kuang, Zuzheng, et al.
Veröffentlicht: (2026)
von: Kuang, Zuzheng, et al.
Veröffentlicht: (2026)
Open-Vocabulary Octree-Graph for 3D Scene Understanding
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
von: Wang, Zhigang, et al.
Veröffentlicht: (2024)
OpenNav: Efficient Open Vocabulary 3D Object Detection for Smart Wheelchair Navigation
von: Rahman, Muhammad Rameez ur, et al.
Veröffentlicht: (2024)
von: Rahman, Muhammad Rameez ur, et al.
Veröffentlicht: (2024)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
von: Zhao, Youjun, et al.
Veröffentlicht: (2025)
von: Zhao, Youjun, et al.
Veröffentlicht: (2025)
LoGoSeg: Integrating Local and Global Features for Open-Vocabulary Semantic Segmentation
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
V3Det Challenge 2024 on Vast Vocabulary and Open Vocabulary Object Detection: Methods and Results
von: Wang, Jiaqi, et al.
Veröffentlicht: (2024)
von: Wang, Jiaqi, et al.
Veröffentlicht: (2024)
ReasonGrounder: LVLM-Guided Hierarchical Feature Splatting for Open-Vocabulary 3D Visual Grounding and Reasoning
von: Liu, Zhenyang, et al.
Veröffentlicht: (2025)
von: Liu, Zhenyang, et al.
Veröffentlicht: (2025)
Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection
von: Kim, Youbin, et al.
Veröffentlicht: (2026)
von: Kim, Youbin, et al.
Veröffentlicht: (2026)
OmniOVCD: Streamlining Open-Vocabulary Change Detection with SAM 3
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
OpenHuman4D: Open-Vocabulary 4D Human Parsing
von: Suzuki, Keito, et al.
Veröffentlicht: (2025)
von: Suzuki, Keito, et al.
Veröffentlicht: (2025)
Learning Background Prompts to Discover Implicit Knowledge for Open Vocabulary Object Detection
von: Li, Jiaming, et al.
Veröffentlicht: (2024)
von: Li, Jiaming, et al.
Veröffentlicht: (2024)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
von: Hu, Yupeng, et al.
Veröffentlicht: (2025)
von: Hu, Yupeng, et al.
Veröffentlicht: (2025)
Search3D: Hierarchical Open-Vocabulary 3D Segmentation
von: Takmaz, Ayca, et al.
Veröffentlicht: (2024)
von: Takmaz, Ayca, et al.
Veröffentlicht: (2024)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
von: Wang, Ziyi, et al.
Veröffentlicht: (2024)
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
von: Shao, Yawen, et al.
Veröffentlicht: (2024)
von: Shao, Yawen, et al.
Veröffentlicht: (2024)
Scaling Open-Vocabulary Action Detection
von: Sia, Zhen Hao, et al.
Veröffentlicht: (2025)
von: Sia, Zhen Hao, et al.
Veröffentlicht: (2025)
Scaling Open-Vocabulary Object Detection
von: Minderer, Matthias, et al.
Veröffentlicht: (2023)
von: Minderer, Matthias, et al.
Veröffentlicht: (2023)
Collaborative Vision-Text Representation Optimizing for Open-Vocabulary Segmentation
von: Jiao, Siyu, et al.
Veröffentlicht: (2024)
von: Jiao, Siyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Global-Local Collaborative Inference with LLM for Lidar-Based Open-Vocabulary Detection
von: Peng, Xingyu, et al.
Veröffentlicht: (2024) -
Eliminating Cross-modal Conflicts in BEV Space for LiDAR-Camera 3D Object Detection
von: Fu, Jiahui, et al.
Veröffentlicht: (2024) -
RATopo: Improving Lane Topology Reasoning via Redundancy Assignment
von: Li, Han, et al.
Veröffentlicht: (2025) -
Unified Modeling of Lane and Lane Topology for Driving Scene Reasoning
von: Li, Han, et al.
Veröffentlicht: (2026) -
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
von: Huang, Rui, et al.
Veröffentlicht: (2024)