VoxelNextFusion: A Simple, Unified and Effective Voxel Fusion Framework for Multi-Modal 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Song, Ziying, Zhang, Guoxin, Xie, Jun, Liu, Lin, Jia, Caiyan, Xu, Shaoqing, Wang, Zhepeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboFusion: Towards Robust Multi-Modal 3D Object Detection via SAM
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
by: Jia, Feiyang, et al.
Published: (2025)
by: Jia, Feiyang, et al.
Published: (2025)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
FGU3R: Fine-Grained Fusion via Unified 3D Representation for Multimodal 3D Object Detection
by: Zhang, Guoxin, et al.
Published: (2025)
by: Zhang, Guoxin, et al.
Published: (2025)
SparseDet: A Simple and Effective Framework for Fully Sparse LiDAR-based 3D Object Detection
by: Liu, Lin, et al.
Published: (2024)
by: Liu, Lin, et al.
Published: (2024)
GraphAlign: Enhancing Accurate Feature Alignment by Graph matching for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2023)
by: Song, Ziying, et al.
Published: (2023)
PVAFN: Point-Voxel Attention Fusion Network with Multi-Pooling Enhancing for 3D Object Detection
by: Li, Yidi, et al.
Published: (2024)
by: Li, Yidi, et al.
Published: (2024)
ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
Robustness-Aware 3D Object Detection in Autonomous Driving: A Review and Outlook
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
Pillar-Voxel Fusion Network for 3D Object Detection in Airborne Hyperspectral Point Clouds
by: Jiang, Yanze, et al.
Published: (2025)
by: Jiang, Yanze, et al.
Published: (2025)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
by: Sun, Haowen, et al.
Published: (2026)
by: Sun, Haowen, et al.
Published: (2026)
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
by: Lu, Yuxuan, et al.
Published: (2024)
by: Lu, Yuxuan, et al.
Published: (2024)
VoxelKeypointFusion: Generalizable Multi-View Multi-Person Pose Estimation
by: Bermuth, Daniel, et al.
Published: (2024)
by: Bermuth, Daniel, et al.
Published: (2024)
SparseVoxFormer: Sparse Voxel-based Transformer for Multi-modal 3D Object Detection
by: Son, Hyeongseok, et al.
Published: (2025)
by: Son, Hyeongseok, et al.
Published: (2025)
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
by: Leng, Zhaoqi, et al.
Published: (2024)
by: Leng, Zhaoqi, et al.
Published: (2024)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
by: Xu, Shaoqing, et al.
Published: (2024)
by: Xu, Shaoqing, et al.
Published: (2024)
UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene Representation
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
VGA: Vision and Graph Fused Attention Network for Rumor Detection
by: Bai, Lin, et al.
Published: (2024)
by: Bai, Lin, et al.
Published: (2024)
GaussianPretrain: A Simple Unified 3D Gaussian Representation for Visual Pre-training in Autonomous Driving
by: Xu, Shaoqing, et al.
Published: (2024)
by: Xu, Shaoqing, et al.
Published: (2024)
SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels
by: Yu, Qiucheng, et al.
Published: (2025)
by: Yu, Qiucheng, et al.
Published: (2025)
4D Neural Voxel Splatting: Dynamic Scene Rendering with Voxelized Guassian Splatting
by: Wu, Chun-Tin, et al.
Published: (2025)
by: Wu, Chun-Tin, et al.
Published: (2025)
GVSynergy-Det: Synergistic Gaussian-Voxel Representations for Multi-View 3D Object Detection
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
LESV: Language Embedded Sparse Voxel Fusion for Open-Vocabulary 3D Scene Understanding
by: Wang, Fusang, et al.
Published: (2026)
by: Wang, Fusang, et al.
Published: (2026)
Poxel: Voxel Reconstruction for 3D Printing
by: Cao, Ruixiang, et al.
Published: (2025)
by: Cao, Ruixiang, et al.
Published: (2025)
Vox-Fusion++: Voxel-based Neural Implicit Dense Tracking and Mapping with Multi-maps
by: Zhai, Hongjia, et al.
Published: (2024)
by: Zhai, Hongjia, et al.
Published: (2024)
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
by: Cao, Haozhi, et al.
Published: (2024)
by: Cao, Haozhi, et al.
Published: (2024)
Progressive Multi-Modal Fusion for Robust 3D Object Detection
by: Mohan, Rohit, et al.
Published: (2024)
by: Mohan, Rohit, et al.
Published: (2024)
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
by: Huang, Yuhao, et al.
Published: (2023)
by: Huang, Yuhao, et al.
Published: (2023)
MsSVT++: Mixed-scale Sparse Voxel Transformer with Center Voting for 3D Object Detection
by: Li, Jianan, et al.
Published: (2024)
by: Li, Jianan, et al.
Published: (2024)
Self-Supervised Scene Flow Estimation with Point-Voxel Fusion and Surface Representation
by: Xiang, Xuezhi, et al.
Published: (2024)
by: Xiang, Xuezhi, et al.
Published: (2024)
PV-SSD: A Multi-Modal Point Cloud Feature Fusion Method for Projection Features and Variable Receptive Field Voxel Features
by: Shao, Yongxin, et al.
Published: (2023)
by: Shao, Yongxin, et al.
Published: (2023)
MR3D-Net: Dynamic Multi-Resolution 3D Sparse Voxel Grid Fusion for LiDAR-Based Collective Perception
by: Teufel, Sven, et al.
Published: (2024)
by: Teufel, Sven, et al.
Published: (2024)
VoxelRF: Voxelized Radiance Field for Fast Wireless Channel Modeling
by: Zeng, Zihang, et al.
Published: (2025)
by: Zeng, Zihang, et al.
Published: (2025)
C$^3$P-VoxelMap: Compact, Cumulative and Coalescible Probabilistic Voxel Mapping
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
Uncertainty-Encoded Multi-Modal Fusion for Robust Object Detection in Autonomous Driving
by: Lou, Yang, et al.
Published: (2023)
by: Lou, Yang, et al.
Published: (2023)
RCM-Fusion: Radar-Camera Multi-Level Fusion for 3D Object Detection
by: Kim, Jisong, et al.
Published: (2023)
by: Kim, Jisong, et al.
Published: (2023)
Fusion is Not Enough: Single Modal Attacks on Fusion Models for 3D Object Detection
by: Cheng, Zhiyuan, et al.
Published: (2023)
by: Cheng, Zhiyuan, et al.
Published: (2023)
OpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understanding
by: Huang, Sheng-Yu, et al.
Published: (2026)
by: Huang, Sheng-Yu, et al.
Published: (2026)
NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding
by: Liu, Shiyu, et al.
Published: (2025)
by: Liu, Shiyu, et al.
Published: (2025)
Voxel grave
by: D4N1L4
Published: (2020)
by: D4N1L4
Published: (2020)
Similar Items
-
RoboFusion: Towards Robust Multi-Modal 3D Object Detection via SAM
by: Song, Ziying, et al.
Published: (2024) -
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
by: Jia, Feiyang, et al.
Published: (2025) -
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024) -
FGU3R: Fine-Grained Fusion via Unified 3D Representation for Multimodal 3D Object Detection
by: Zhang, Guoxin, et al.
Published: (2025) -
SparseDet: A Simple and Effective Framework for Fully Sparse LiDAR-based 3D Object Detection
by: Liu, Lin, et al.
Published: (2024)