MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Zitian, Huang, Zehao, Gao, Yulu, Wang, Naiyan, Liu, Si |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Enhancing 3D Lane Detection and Topology Reasoning with 2D Lane Priors
por: Li, Han, et al.
Publicado: (2024)
por: Li, Han, et al.
Publicado: (2024)
SparseFusion: Efficient Sparse Multi-Modal Fusion Framework for Long-Range 3D Perception
por: Li, Yiheng, et al.
Publicado: (2024)
por: Li, Yiheng, et al.
Publicado: (2024)
Fully Sparse Fusion for 3D Object Detection
por: Li, Yingyan, et al.
Publicado: (2023)
por: Li, Yingyan, et al.
Publicado: (2023)
Geometry-Guided 3D Visual Token Pruning for Video-Language Models
por: Li, Han, et al.
Publicado: (2026)
por: Li, Han, et al.
Publicado: (2026)
Anchor3DLane++: 3D Lane Detection via Sample-Adaptive Sparse 3D Anchor Regression
por: Huang, Shaofei, et al.
Publicado: (2024)
por: Huang, Shaofei, et al.
Publicado: (2024)
Modality-Specific Hierarchical Enhancement for RGB-D Camouflaged Object Detection
por: Niu, Yuzhen, et al.
Publicado: (2026)
por: Niu, Yuzhen, et al.
Publicado: (2026)
Instruction-Oriented Preference Alignment for Enhancing Multi-Modal Comprehension Capability of MLLMs
por: Wang, Zitian, et al.
Publicado: (2025)
por: Wang, Zitian, et al.
Publicado: (2025)
Exploiting Modality-Specific Features For Multi-Modal Manipulation Detection And Grounding
por: Wang, Jiazhen, et al.
Publicado: (2023)
por: Wang, Jiazhen, et al.
Publicado: (2023)
Eliminating Cross-modal Conflicts in BEV Space for LiDAR-Camera 3D Object Detection
por: Fu, Jiahui, et al.
Publicado: (2024)
por: Fu, Jiahui, et al.
Publicado: (2024)
Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection
por: Wang, Hao, et al.
Publicado: (2026)
por: Wang, Hao, et al.
Publicado: (2026)
Rethinking Multi-Modal Object Detection from the Perspective of Mono-Modality Feature Learning
por: Zhao, Tianyi, et al.
Publicado: (2025)
por: Zhao, Tianyi, et al.
Publicado: (2025)
Towards Flexible 3D Perception: Object-Centric Occupancy Completion Augments 3D Object Detection
por: Zheng, Chaoda, et al.
Publicado: (2024)
por: Zheng, Chaoda, et al.
Publicado: (2024)
Learnable Graph Matching: A Practical Paradigm for Data Association
por: He, Jiawei, et al.
Publicado: (2023)
por: He, Jiawei, et al.
Publicado: (2023)
Modality Prompts for Arbitrary Modality Salient Object Detection
por: Huang, Nianchang, et al.
Publicado: (2024)
por: Huang, Nianchang, et al.
Publicado: (2024)
ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop
por: Li, Shuangzhi, et al.
Publicado: (2026)
por: Li, Shuangzhi, et al.
Publicado: (2026)
Progressive Multi-Modal Fusion for Robust 3D Object Detection
por: Mohan, Rohit, et al.
Publicado: (2024)
por: Mohan, Rohit, et al.
Publicado: (2024)
RoboFusion: Towards Robust Multi-Modal 3D Object Detection via SAM
por: Song, Ziying, et al.
Publicado: (2024)
por: Song, Ziying, et al.
Publicado: (2024)
CCF: Complementary Collaborative Fusion for Domain Generalized Multi-Modal 3D Object Detection
por: Wu, Yuchen, et al.
Publicado: (2026)
por: Wu, Yuchen, et al.
Publicado: (2026)
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
por: Jia, Feiyang, et al.
Publicado: (2025)
por: Jia, Feiyang, et al.
Publicado: (2025)
Exploring Modality-Aware Fusion and Decoupled Temporal Propagation for Multi-Modal Object Tracking
por: Wang, Shilei, et al.
Publicado: (2026)
por: Wang, Shilei, et al.
Publicado: (2026)
Contrast-Guided Cross-Modal Distillation for Thermal Object Detection
por: Kim, SiWoo, et al.
Publicado: (2025)
por: Kim, SiWoo, et al.
Publicado: (2025)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
por: Song, Ziying, et al.
Publicado: (2024)
por: Song, Ziying, et al.
Publicado: (2024)
SP3D: Boosting Sparsely-Supervised 3D Object Detection via Accurate Cross-Modal Semantic Prompts
por: Zhao, Shijia, et al.
Publicado: (2025)
por: Zhao, Shijia, et al.
Publicado: (2025)
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
por: Ding, Rui, et al.
Publicado: (2026)
por: Ding, Rui, et al.
Publicado: (2026)
EVT: Efficient View Transformation for Multi-Modal 3D Object Detection
por: Lee, Yongjin, et al.
Publicado: (2024)
por: Lee, Yongjin, et al.
Publicado: (2024)
CMF-IoU: Multi-Stage Cross-Modal Fusion 3D Object Detection with IoU Joint Prediction
por: Ning, Zhiwei, et al.
Publicado: (2025)
por: Ning, Zhiwei, et al.
Publicado: (2025)
AW-MoE: All-Weather Mixture of Experts for Robust Multi-Modal 3D Object Detection
por: Lin, Hongwei, et al.
Publicado: (2026)
por: Lin, Hongwei, et al.
Publicado: (2026)
STMI: Segmentation-Guided Token Modulation with Cross-Modal Hypergraph Interaction for Multi-Modal Object Re-Identification
por: Xu, Xingguo, et al.
Publicado: (2026)
por: Xu, Xingguo, et al.
Publicado: (2026)
Salient Object Detection From Arbitrary Modalities
por: Huang, Nianchang, et al.
Publicado: (2024)
por: Huang, Nianchang, et al.
Publicado: (2024)
You Only Need Two Detectors to Achieve Multi-Modal 3D Multi-Object Tracking
por: Wang, Xiyang, et al.
Publicado: (2023)
por: Wang, Xiyang, et al.
Publicado: (2023)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
por: Di Lorenzo, Gaia, et al.
Publicado: (2025)
por: Di Lorenzo, Gaia, et al.
Publicado: (2025)
M^3Detection: Multi-Frame Multi-Level Feature Fusion for Multi-Modal 3D Object Detection with Camera and 4D Imaging Radar
por: Li, Xiaozhi, et al.
Publicado: (2025)
por: Li, Xiaozhi, et al.
Publicado: (2025)
Leveraging Multi-Modal Saliency and Fusion for Gaze Target Detection
por: Mathew, Athul M., et al.
Publicado: (2025)
por: Mathew, Athul M., et al.
Publicado: (2025)
VoxelNextFusion: A Simple, Unified and Effective Voxel Fusion Framework for Multi-Modal 3D Object Detection
por: Song, Ziying, et al.
Publicado: (2024)
por: Song, Ziying, et al.
Publicado: (2024)
OV-Uni3DETR: Towards Unified Open-Vocabulary 3D Object Detection via Cycle-Modality Propagation
por: Wang, Zhenyu, et al.
Publicado: (2024)
por: Wang, Zhenyu, et al.
Publicado: (2024)
Multi-Modal Assistance for Unsupervised Domain Adaptation on Point Cloud 3D Object Detection
por: Zhao, Shenao, et al.
Publicado: (2025)
por: Zhao, Shenao, et al.
Publicado: (2025)
PoIFusion: Multi-Modal 3D Object Detection via Fusion at Points of Interest
por: Deng, Jiajun, et al.
Publicado: (2024)
por: Deng, Jiajun, et al.
Publicado: (2024)
RegTrack: Simplicity Beneath Complexity in Robust Multi-Modal 3D Multi-Object Tracking
por: Gu, Lipeng, et al.
Publicado: (2024)
por: Gu, Lipeng, et al.
Publicado: (2024)
Learning Multi-Modal Prototypes for Cross-Domain Few-Shot Object Detection
por: Wang, Wanqi, et al.
Publicado: (2026)
por: Wang, Wanqi, et al.
Publicado: (2026)
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
por: Cha, Juhan, et al.
Publicado: (2024)
por: Cha, Juhan, et al.
Publicado: (2024)
Ejemplares similares
-
Enhancing 3D Lane Detection and Topology Reasoning with 2D Lane Priors
por: Li, Han, et al.
Publicado: (2024) -
SparseFusion: Efficient Sparse Multi-Modal Fusion Framework for Long-Range 3D Perception
por: Li, Yiheng, et al.
Publicado: (2024) -
Fully Sparse Fusion for 3D Object Detection
por: Li, Yingyan, et al.
Publicado: (2023) -
Geometry-Guided 3D Visual Token Pruning for Video-Language Models
por: Li, Han, et al.
Publicado: (2026) -
Anchor3DLane++: 3D Lane Detection via Sample-Adaptive Sparse 3D Anchor Regression
por: Huang, Shaofei, et al.
Publicado: (2024)