SparseVoxFormer: Sparse Voxel-based Transformer for Multi-modal 3D Object Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | Son, Hyeongseok, He, Jia, Park, Seung-In, Min, Ying, Zhang, Yunhao, Yoo, ByungIn |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
por: Li, Wenxi, et al.
Publicado: (2025)
por: Li, Wenxi, et al.
Publicado: (2025)
MsSVT++: Mixed-scale Sparse Voxel Transformer with Center Voting for 3D Object Detection
por: Li, Jianan, et al.
Publicado: (2024)
por: Li, Jianan, et al.
Publicado: (2024)
HIMap: HybrId Representation Learning for End-to-end Vectorized HD Map Construction
por: Zhou, Yi, et al.
Publicado: (2024)
por: Zhou, Yi, et al.
Publicado: (2024)
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
por: Leng, Zhaoqi, et al.
Publicado: (2024)
por: Leng, Zhaoqi, et al.
Publicado: (2024)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
por: Sun, Haowen, et al.
Publicado: (2026)
por: Sun, Haowen, et al.
Publicado: (2026)
SparseDet: A Simple and Effective Framework for Fully Sparse LiDAR-based 3D Object Detection
por: Liu, Lin, et al.
Publicado: (2024)
por: Liu, Lin, et al.
Publicado: (2024)
PlaceFormer: Transformer-based Visual Place Recognition using Multi-Scale Patch Selection and Fusion
por: Kannan, Shyam Sundar, et al.
Publicado: (2024)
por: Kannan, Shyam Sundar, et al.
Publicado: (2024)
HeightFormer: Learning Height Prediction in Voxel Features for Roadside Vision Centric 3D Object Detection via Transformer
por: Zhang, Zhang, et al.
Publicado: (2025)
por: Zhang, Zhang, et al.
Publicado: (2025)
MapDistill: Boosting Efficient Camera-based HD Map Construction via Camera-LiDAR Fusion Model Distillation
por: Hao, Xiaoshuai, et al.
Publicado: (2024)
por: Hao, Xiaoshuai, et al.
Publicado: (2024)
No Dense Tensors Needed: Fully Sparse Object Detection on Event-Camera Voxel Grids
por: Sadoun, Mohamad Yazan, et al.
Publicado: (2026)
por: Sadoun, Mohamad Yazan, et al.
Publicado: (2026)
SVRecon: Sparse Voxel Rasterization for Surface Reconstruction
por: Oh, Seunghun, et al.
Publicado: (2025)
por: Oh, Seunghun, et al.
Publicado: (2025)
Structure-Adaptive Sparse Diffusion in Voxel Space for 3D Medical Image Enhancement
por: Jiang, Hongxu, et al.
Publicado: (2026)
por: Jiang, Hongxu, et al.
Publicado: (2026)
ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention
por: He, Chenhang, et al.
Publicado: (2024)
por: He, Chenhang, et al.
Publicado: (2024)
Vox-Fusion++: Voxel-based Neural Implicit Dense Tracking and Mapping with Multi-maps
por: Zhai, Hongjia, et al.
Publicado: (2024)
por: Zhai, Hongjia, et al.
Publicado: (2024)
VoxelNextFusion: A Simple, Unified and Effective Voxel Fusion Framework for Multi-Modal 3D Object Detection
por: Song, Ziying, et al.
Publicado: (2024)
por: Song, Ziying, et al.
Publicado: (2024)
Scene Adaptive Sparse Transformer for Event-based Object Detection
por: Peng, Yansong, et al.
Publicado: (2024)
por: Peng, Yansong, et al.
Publicado: (2024)
SPADE: Sparse Pillar-based 3D Object Detection Accelerator for Autonomous Driving
por: Lee, Minjae, et al.
Publicado: (2023)
por: Lee, Minjae, et al.
Publicado: (2023)
Selectively Dilated Convolution for Accuracy-Preserving Sparse Pillar-based Embedded 3D Object Detection
por: Park, Seongmin, et al.
Publicado: (2024)
por: Park, Seongmin, et al.
Publicado: (2024)
Fully Sparse Fusion for 3D Object Detection
por: Li, Yingyan, et al.
Publicado: (2023)
por: Li, Yingyan, et al.
Publicado: (2023)
SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels
por: Yu, Qiucheng, et al.
Publicado: (2025)
por: Yu, Qiucheng, et al.
Publicado: (2025)
TSC-PCAC: Voxel Transformer and Sparse Convolution Based Point Cloud Attribute Compression for 3D Broadcasting
por: Guo, Zixi, et al.
Publicado: (2024)
por: Guo, Zixi, et al.
Publicado: (2024)
CVT-xRF: Contrastive In-Voxel Transformer for 3D Consistent Radiance Fields from Sparse Inputs
por: Zhong, Yingji, et al.
Publicado: (2024)
por: Zhong, Yingji, et al.
Publicado: (2024)
SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
por: Park, Sangha, et al.
Publicado: (2025)
por: Park, Sangha, et al.
Publicado: (2025)
Differentiable Voxel-based X-ray Rendering Improves Sparse-View 3D CBCT Reconstruction
por: Momeni, Mohammadhossein, et al.
Publicado: (2024)
por: Momeni, Mohammadhossein, et al.
Publicado: (2024)
SFi-Former: Sparse Flow Induced Attention for Graph Transformer
por: Li, Zhonghao, et al.
Publicado: (2025)
por: Li, Zhonghao, et al.
Publicado: (2025)
Few-Shot Object Detection with Sparse Context Transformers
por: Mei, Jie, et al.
Publicado: (2024)
por: Mei, Jie, et al.
Publicado: (2024)
SigFormer: Sparse Signal-Guided Transformer for Multi-Modal Human Action Segmentation
por: Liu, Qi, et al.
Publicado: (2023)
por: Liu, Qi, et al.
Publicado: (2023)
VoxAct-B: Voxel-Based Acting and Stabilizing Policy for Bimanual Manipulation
por: Liu, I-Chun Arthur, et al.
Publicado: (2024)
por: Liu, I-Chun Arthur, et al.
Publicado: (2024)
SFMNet: Sparse Focal Modulation for 3D Object Detection
por: Shrout, Oren, et al.
Publicado: (2025)
por: Shrout, Oren, et al.
Publicado: (2025)
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
por: Lu, Yuxuan, et al.
Publicado: (2024)
por: Lu, Yuxuan, et al.
Publicado: (2024)
ShapeShifter: 3D Variations Using Multiscale and Sparse Point-Voxel Diffusion
por: Maruani, Nissim, et al.
Publicado: (2025)
por: Maruani, Nissim, et al.
Publicado: (2025)
XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies
por: Ren, Xuanchi, et al.
Publicado: (2023)
por: Ren, Xuanchi, et al.
Publicado: (2023)
TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding
por: Guo, Wenxuan, et al.
Publicado: (2025)
por: Guo, Wenxuan, et al.
Publicado: (2025)
SparseLIF: High-Performance Sparse LiDAR-Camera Fusion for 3D Object Detection
por: Zhang, Hongcheng, et al.
Publicado: (2024)
por: Zhang, Hongcheng, et al.
Publicado: (2024)
HopFormer: Sparse Graph Transformers with Explicit Receptive Field Control
por: Yun, Sanggeon, et al.
Publicado: (2026)
por: Yun, Sanggeon, et al.
Publicado: (2026)
Scaffold Diffusion: Sparse Multi-Category Voxel Structure Generation with Discrete Diffusion
por: Jung, Justin
Publicado: (2025)
por: Jung, Justin
Publicado: (2025)
Advancing Structured Priors for Sparse-Voxel Surface Reconstruction
por: Chi, Ting-Hsun, et al.
Publicado: (2026)
por: Chi, Ting-Hsun, et al.
Publicado: (2026)
FSHNet: Fully Sparse Hybrid Network for 3D Object Detection
por: Liu, Shuai, et al.
Publicado: (2025)
por: Liu, Shuai, et al.
Publicado: (2025)
LESV: Language Embedded Sparse Voxel Fusion for Open-Vocabulary 3D Scene Understanding
por: Wang, Fusang, et al.
Publicado: (2026)
por: Wang, Fusang, et al.
Publicado: (2026)
Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection
por: Zhang, Guowen, et al.
Publicado: (2024)
por: Zhang, Guowen, et al.
Publicado: (2024)
Ejemplares similares
-
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
por: Li, Wenxi, et al.
Publicado: (2025) -
MsSVT++: Mixed-scale Sparse Voxel Transformer with Center Voting for 3D Object Detection
por: Li, Jianan, et al.
Publicado: (2024) -
HIMap: HybrId Representation Learning for End-to-end Vectorized HD Map Construction
por: Zhou, Yi, et al.
Publicado: (2024) -
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
por: Leng, Zhaoqi, et al.
Publicado: (2024) -
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
por: Sun, Haowen, et al.
Publicado: (2026)