BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion
Fuente:
arXiv
Saved in:
| Main Authors: | Lan, Yuqing, Zhu, Chenyang, Gao, Zhirui, Zhang, Jiazhao, Cao, Yihan, Yi, Renjiao, Wang, Yijie, Xu, Kai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BoxFusion: Reconstruction‐Free Open‐Vocabulary 3D Object Detection via Real‐Time Multi‐View Box Fusion
by: Yuqing Lan, et al.
Published: (2025)
by: Yuqing Lan, et al.
Published: (2025)
RemixFusion: Residual-based Mixed Representation for Large-scale Online RGB-D Reconstruction
by: Lan, Yuqing, et al.
Published: (2025)
by: Lan, Yuqing, et al.
Published: (2025)
Generic Objects as Pose Probes for Few-shot View Synthesis
by: Gao, Zhirui, et al.
Published: (2024)
by: Gao, Zhirui, et al.
Published: (2024)
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Curve-Aware Gaussian Splatting for 3D Parametric Curve Reconstruction
by: Gao, Zhirui, et al.
Published: (2025)
by: Gao, Zhirui, et al.
Published: (2025)
Self-supervised Learning of Hybrid Part-aware 3D Representations of 2D Gaussians and Superquadrics
by: Gao, Zhirui, et al.
Published: (2024)
by: Gao, Zhirui, et al.
Published: (2024)
Learning Accurate Template Matching with Differentiable Coarse-to-Fine Correspondence Refinement
by: Gao, Zhirui, et al.
Published: (2023)
by: Gao, Zhirui, et al.
Published: (2023)
OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging
by: Tang, Yijie, et al.
Published: (2025)
by: Tang, Yijie, et al.
Published: (2025)
Relighting Scenes with Object Insertions in Neural Radiance Fields
by: Zhu, Xuening, et al.
Published: (2024)
by: Zhu, Xuening, et al.
Published: (2024)
RaLiBEV: Radar and LiDAR BEV Fusion Learning for Anchor Box Free Object Detection Systems
by: Yang, Yanlong, et al.
Published: (2022)
by: Yang, Yanlong, et al.
Published: (2022)
Real-time Transformer-based Open-Vocabulary Detection with Efficient Fusion Head
by: Zhao, Tiancheng, et al.
Published: (2024)
by: Zhao, Tiancheng, et al.
Published: (2024)
YOLO-World: Real-Time Open-Vocabulary Object Detection
by: Cheng, Tianheng, et al.
Published: (2024)
by: Cheng, Tianheng, et al.
Published: (2024)
Changes in Real Time: Online Scene Change Detection with Multi-View Fusion
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2025)
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2025)
Retrieving Objects from 3D Scenes with Box-Guided Open-Vocabulary Instance Segmentation
by: Nguyen, Khanh, et al.
Published: (2025)
by: Nguyen, Khanh, et al.
Published: (2025)
Sampling Bag of Views for Open-Vocabulary Object Detection
by: Choi, Hojun, et al.
Published: (2024)
by: Choi, Hojun, et al.
Published: (2024)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
by: Sun, Haowen, et al.
Published: (2026)
by: Sun, Haowen, et al.
Published: (2026)
Confidence Aware SSD Ensemble with Weighted Boxes Fusion for Weapon Detection
by: Jadhav, Atharva, et al.
Published: (2025)
by: Jadhav, Atharva, et al.
Published: (2025)
OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
OpenGS-Fusion: Open-Vocabulary Dense Mapping with Hybrid 3D Gaussian Splatting for Refined Object-Level Understanding
by: Yang, Dianyi, et al.
Published: (2025)
by: Yang, Dianyi, et al.
Published: (2025)
Cross-View Open-Vocabulary Object Detection in Aerial Imagery
by: Kini, Jyoti, et al.
Published: (2025)
by: Kini, Jyoti, et al.
Published: (2025)
MaskClustering: View Consensus based Mask Graph Clustering for Open-Vocabulary 3D Instance Segmentation
by: Yan, Mi, et al.
Published: (2024)
by: Yan, Mi, et al.
Published: (2024)
F2M-Reg: Unsupervised RGB-D Point Cloud Registration with Frame-to-Model Optimization
by: Yu, Zhinan, et al.
Published: (2024)
by: Yu, Zhinan, et al.
Published: (2024)
Enhancing Pseudo-Boxes via Data-Level LiDAR-Camera Fusion for Unsupervised 3D Object Detection
by: Ji, Mingqian, et al.
Published: (2025)
by: Ji, Mingqian, et al.
Published: (2025)
REXO: Indoor Multi-View Radar Object Detection via 3D Bounding Box Diffusion
by: Yataka, Ryoma, et al.
Published: (2025)
by: Yataka, Ryoma, et al.
Published: (2025)
Unleashing the Multi-View Fusion Potential: Noise Correction in VLM for Open-Vocabulary 3D Scene Understanding
by: Yin, Xingyilang, et al.
Published: (2025)
by: Yin, Xingyilang, et al.
Published: (2025)
Neural Observation Field Guided Hybrid Optimization of Camera Placement
by: Cao, Yihan, et al.
Published: (2024)
by: Cao, Yihan, et al.
Published: (2024)
FACTOR: Counterfactual Training-Free Test-Time Adaptation for Open-Vocabulary Object Detection
by: Zhao, Kaixiang, et al.
Published: (2026)
by: Zhao, Kaixiang, et al.
Published: (2026)
Multi-View Reconstruction with Global Context for 3D Anomaly Detection
by: Sun, Yihan, et al.
Published: (2025)
by: Sun, Yihan, et al.
Published: (2025)
Open-Vocabulary Object Detection via Language Hierarchy
by: Huang, Jiaxing, et al.
Published: (2024)
by: Huang, Jiaxing, et al.
Published: (2024)
BAM: Box Abstraction Monitors for Real-time OoD Detection in Object Detection
by: Wu, Changshun, et al.
Published: (2024)
by: Wu, Changshun, et al.
Published: (2024)
AdaptOVCD: Training-Free Open-Vocabulary Remote Sensing Change Detection via Adaptive Information Fusion
by: Dou, Mingyu, et al.
Published: (2026)
by: Dou, Mingyu, et al.
Published: (2026)
VasTSD: Learning 3D Vascular Tree-state Space Diffusion Model for Angiography Synthesis
by: Wang, Zhifeng, et al.
Published: (2025)
by: Wang, Zhifeng, et al.
Published: (2025)
CycleDiff: Cycle Diffusion Models for Unpaired Image-to-image Translation
by: Zou, Shilong, et al.
Published: (2025)
by: Zou, Shilong, et al.
Published: (2025)
AdaptiveFusion: Adaptive Multi-Modal Multi-View Fusion for 3D Human Body Reconstruction
by: Chen, Anjun, et al.
Published: (2024)
by: Chen, Anjun, et al.
Published: (2024)
Multi-Grid Redundant Bounding Box Annotation for Accurate Object Detection
by: Tesema, Solomon Negussie, et al.
Published: (2022)
by: Tesema, Solomon Negussie, et al.
Published: (2022)
DART: An Automated End-to-End Object Detection Pipeline with Data Diversification, Open-Vocabulary Bounding Box Annotation, Pseudo-Label Review, and Model Training
by: Xin, Chen, et al.
Published: (2024)
by: Xin, Chen, et al.
Published: (2024)
MUFASA: Multi-View Fusion and Adaptation Network with Spatial Awareness for Radar Object Detection
by: Peng, Xiangyuan, et al.
Published: (2024)
by: Peng, Xiangyuan, et al.
Published: (2024)
MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection
by: Wang, Hanshi, et al.
Published: (2025)
by: Wang, Hanshi, et al.
Published: (2025)
Training-free Boost for Open-Vocabulary Object Detection with Confidence Aggregation
by: Zheng, Yanhao, et al.
Published: (2024)
by: Zheng, Yanhao, et al.
Published: (2024)
ProFuse: Efficient Cross-View Context Fusion for Open-Vocabulary 3D Gaussian Splatting
by: Chiou, Yen-Jen, et al.
Published: (2026)
by: Chiou, Yen-Jen, et al.
Published: (2026)
Similar Items
-
BoxFusion: Reconstruction‐Free Open‐Vocabulary 3D Object Detection via Real‐Time Multi‐View Box Fusion
by: Yuqing Lan, et al.
Published: (2025) -
RemixFusion: Residual-based Mixed Representation for Large-scale Online RGB-D Reconstruction
by: Lan, Yuqing, et al.
Published: (2025) -
Generic Objects as Pose Probes for Few-shot View Synthesis
by: Gao, Zhirui, et al.
Published: (2024) -
Collaborative Novel Object Discovery and Box-Guided Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Cao, Yang, et al.
Published: (2024) -
Curve-Aware Gaussian Splatting for 3D Parametric Curve Reconstruction
by: Gao, Zhirui, et al.
Published: (2025)