SSLFusion: Scale & Space Aligned Latent Fusion Model for Multimodal 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Bonan, Xie, Jin, Nie, Jing, Cao, Jiale |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
by: Ding, Bonan, et al.
Published: (2024)
by: Ding, Bonan, et al.
Published: (2024)
S-LAM3D: Segmentation-Guided Monocular 3D Object Detection via Feature Space Fusion
by: Sas, Diana-Alexandra, et al.
Published: (2025)
by: Sas, Diana-Alexandra, et al.
Published: (2025)
HiddenObject: Modality-Agnostic Fusion for Multimodal Hidden Object Detection
by: Song, Harris, et al.
Published: (2025)
by: Song, Harris, et al.
Published: (2025)
QCaption: Video Captioning and Q&A through Fusion of Large Multimodal Models
by: Wang, Jiale, et al.
Published: (2026)
by: Wang, Jiale, et al.
Published: (2026)
A Multimodal Hybrid Late-Cascade Fusion Network for Enhanced 3D Object Detection
by: Sgaravatti, Carlo, et al.
Published: (2025)
by: Sgaravatti, Carlo, et al.
Published: (2025)
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
by: Li, Wenxuan, et al.
Published: (2026)
by: Li, Wenxuan, et al.
Published: (2026)
RayMamba: Ray-Aligned Serialization for Long-Range 3D Object Detection
by: Lu, Cheng, et al.
Published: (2026)
by: Lu, Cheng, et al.
Published: (2026)
On the Adversarial Robustness of Camera-based 3D Object Detection
by: Xie, Shaoyuan, et al.
Published: (2023)
by: Xie, Shaoyuan, et al.
Published: (2023)
Aerial Monocular 3D Object Detection
by: Hu, Yue, et al.
Published: (2022)
by: Hu, Yue, et al.
Published: (2022)
DASSF: Dynamic-Attention Scale-Sequence Fusion for Aerial Object Detection
by: Li, Haodong, et al.
Published: (2024)
by: Li, Haodong, et al.
Published: (2024)
3DTTNet: Multimodal Fusion-Based 3D Traversable Terrain Modeling for Off-Road Environments
by: Chen, Zitong, et al.
Published: (2024)
by: Chen, Zitong, et al.
Published: (2024)
COMO: Cross-Mamba Interaction and Offset-Guided Fusion for Multimodal Object Detection
by: Liu, Chang, et al.
Published: (2024)
by: Liu, Chang, et al.
Published: (2024)
LEAF-Mamba: Local Emphatic and Adaptive Fusion State Space Model for RGB-D Salient Object Detection
by: Wu, Lanhu, et al.
Published: (2025)
by: Wu, Lanhu, et al.
Published: (2025)
State Space Model Meets Transformer: A New Paradigm for 3D Object Detection
by: Wang, Chuxin, et al.
Published: (2025)
by: Wang, Chuxin, et al.
Published: (2025)
Align Your Query: Representation Alignment for Multimodality Medical Object Detection
by: Seo, Ara, et al.
Published: (2025)
by: Seo, Ara, et al.
Published: (2025)
CFMD: Dynamic Cross-layer Feature Fusion for Salient Object Detection
by: Lian, Jin, et al.
Published: (2025)
by: Lian, Jin, et al.
Published: (2025)
SimpleBEV: Improved LiDAR-Camera Fusion Architecture for 3D Object Detection
by: Zhao, Yun, et al.
Published: (2024)
by: Zhao, Yun, et al.
Published: (2024)
BiCo-Fusion: Bidirectional Complementary LiDAR-Camera Fusion for Semantic- and Spatial-Aware 3D Object Detection
by: Song, Yang, et al.
Published: (2024)
by: Song, Yang, et al.
Published: (2024)
Post Fusion Bird's Eye View Feature Stabilization for Robust Multimodal 3D Detection
by: Dong, Trung Tien, et al.
Published: (2026)
by: Dong, Trung Tien, et al.
Published: (2026)
Safety-Aligned 3D Object Detection: Single-Vehicle, Cooperative, and End-to-End Perspectives
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2026)
by: Liao, Brian Hsuan-Cheng, et al.
Published: (2026)
Latent Distillation for Continual Object Detection at the Edge
by: Pasti, Francesco, et al.
Published: (2024)
by: Pasti, Francesco, et al.
Published: (2024)
Contextual Object Detection with Multimodal Large Language Models
by: Zang, Yuhang, et al.
Published: (2023)
by: Zang, Yuhang, et al.
Published: (2023)
COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection
by: Peng, Peiran, et al.
Published: (2025)
by: Peng, Peiran, et al.
Published: (2025)
Timely Fusion of Surround Radar/Lidar for Object Detection in Autonomous Driving Systems
by: Xie, Wenjing, et al.
Published: (2023)
by: Xie, Wenjing, et al.
Published: (2023)
Fusion-Mamba for Cross-modality Object Detection
by: Dong, Wenhao, et al.
Published: (2024)
by: Dong, Wenhao, et al.
Published: (2024)
S3MOT: Monocular 3D Object Tracking with Selective State Space Model
by: Yan, Zhuohao, et al.
Published: (2025)
by: Yan, Zhuohao, et al.
Published: (2025)
EMIFF: Enhanced Multi-scale Image Feature Fusion for Vehicle-Infrastructure Cooperative 3D Object Detection
by: Wang, Zhe, et al.
Published: (2024)
by: Wang, Zhe, et al.
Published: (2024)
DEPFusion: Dual-Domain Enhancement and Priority-Guided Mamba Fusion for UAV Multispectral Object Detection
by: Li, Shucong, et al.
Published: (2025)
by: Li, Shucong, et al.
Published: (2025)
Object Agnostic 3D Lifting in Space and Time
by: Fusco, Christopher, et al.
Published: (2024)
by: Fusco, Christopher, et al.
Published: (2024)
Fore-Mamba3D: Mamba-based Foreground-Enhanced Encoding for 3D Object Detection
by: Ning, Zhiwei, et al.
Published: (2026)
by: Ning, Zhiwei, et al.
Published: (2026)
FASTer: Focal Token Acquiring-and-Scaling Transformer for Long-term 3D Object Detection
by: Dang, Chenxu, et al.
Published: (2025)
by: Dang, Chenxu, et al.
Published: (2025)
EarthCrafter: Scalable 3D Earth Generation via Dual-Sparse Latent Diffusion
by: Liu, Shang, et al.
Published: (2025)
by: Liu, Shang, et al.
Published: (2025)
GateFuseNet: An Adaptive 3D Multimodal Neuroimaging Fusion Network for Parkinson's Disease Diagnosis
by: Jin, Rui, et al.
Published: (2025)
by: Jin, Rui, et al.
Published: (2025)
Deformba: Vision State Space Model with Adaptive State Fusion
by: Ke, Hongyu, et al.
Published: (2026)
by: Ke, Hongyu, et al.
Published: (2026)
Depth-aware Fusion Method based on Image and 4D Radar Spectrum for 3D Object Detection
by: Sun, Yue, et al.
Published: (2025)
by: Sun, Yue, et al.
Published: (2025)
ObjectNLQ @ Ego4D Episodic Memory Challenge 2024
by: Feng, Yisen, et al.
Published: (2024)
by: Feng, Yisen, et al.
Published: (2024)
SIM-Net: A Multimodal Fusion Network Using Inferred 3D Object Shape Point Clouds from RGB Images for 2D Classification
by: Sklab, Youcef, et al.
Published: (2025)
by: Sklab, Youcef, et al.
Published: (2025)
Cross-Modal Purification and Fusion for Small-Object RGB-D Transmission-Line Defect Detection
by: Cui, Jiaming, et al.
Published: (2026)
by: Cui, Jiaming, et al.
Published: (2026)
MultiCorrupt: A Multi-Modal Robustness Dataset and Benchmark of LiDAR-Camera Fusion for 3D Object Detection
by: Beemelmanns, Till, et al.
Published: (2024)
by: Beemelmanns, Till, et al.
Published: (2024)
OT-ALD: Aligning Latent Distributions with Optimal Transport for Accelerated Image-to-Image Translation
by: Wang, Zhanpeng, et al.
Published: (2025)
by: Wang, Zhanpeng, et al.
Published: (2025)
Similar Items
-
VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
by: Ding, Bonan, et al.
Published: (2024) -
S-LAM3D: Segmentation-Guided Monocular 3D Object Detection via Feature Space Fusion
by: Sas, Diana-Alexandra, et al.
Published: (2025) -
HiddenObject: Modality-Agnostic Fusion for Multimodal Hidden Object Detection
by: Song, Harris, et al.
Published: (2025) -
QCaption: Video Captioning and Q&A through Fusion of Large Multimodal Models
by: Wang, Jiale, et al.
Published: (2026) -
A Multimodal Hybrid Late-Cascade Fusion Network for Enhanced 3D Object Detection
by: Sgaravatti, Carlo, et al.
Published: (2025)