PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Leng, Zhaoqi, Sun, Pei, He, Tong, Anguelov, Dragomir, Tan, Mingxing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
S4-Driver: Scalable Self-Supervised Driving Multimodal Large Language Modelwith Spatio-Temporal Visual Representation
by: Xie, Yichen, et al.
Published: (2025)
by: Xie, Yichen, et al.
Published: (2025)
Enhanced Motion Forecasting with Plug-and-Play Multimodal Large Language Models
by: Luo, Katie, et al.
Published: (2025)
by: Luo, Katie, et al.
Published: (2025)
Scene Reconstruction as Mapping Priors for 3D Detection
by: Fu, Yang, et al.
Published: (2026)
by: Fu, Yang, et al.
Published: (2026)
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
by: Lu, Yuxuan, et al.
Published: (2024)
by: Lu, Yuxuan, et al.
Published: (2024)
LET-3D-AP: Longitudinal Error Tolerant 3D Average Precision for Camera-Only 3D Detection
by: Hung, Wei-Chih, et al.
Published: (2022)
by: Hung, Wei-Chih, et al.
Published: (2022)
HeightFormer: Learning Height Prediction in Voxel Features for Roadside Vision Centric 3D Object Detection via Transformer
by: Zhang, Zhang, et al.
Published: (2025)
by: Zhang, Zhang, et al.
Published: (2025)
WOMD-LiDAR: Raw Sensor Dataset Benchmark for Motion Forecasting
by: Chen, Kan, et al.
Published: (2023)
by: Chen, Kan, et al.
Published: (2023)
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
by: Huang, Yuhao, et al.
Published: (2023)
by: Huang, Yuhao, et al.
Published: (2023)
SparseVoxFormer: Sparse Voxel-based Transformer for Multi-modal 3D Object Detection
by: Son, Hyeongseok, et al.
Published: (2025)
by: Son, Hyeongseok, et al.
Published: (2025)
Pillar-Voxel Fusion Network for 3D Object Detection in Airborne Hyperspectral Point Clouds
by: Jiang, Yanze, et al.
Published: (2025)
by: Jiang, Yanze, et al.
Published: (2025)
Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection
by: Zhang, Guowen, et al.
Published: (2024)
by: Zhang, Guowen, et al.
Published: (2024)
PVAFN: Point-Voxel Attention Fusion Network with Multi-Pooling Enhancing for 3D Object Detection
by: Li, Yidi, et al.
Published: (2024)
by: Li, Yidi, et al.
Published: (2024)
MsSVT++: Mixed-scale Sparse Voxel Transformer with Center Voting for 3D Object Detection
by: Li, Jianan, et al.
Published: (2024)
by: Li, Jianan, et al.
Published: (2024)
GraVoS: Voxel Selection for 3D Point-Cloud Detection
by: Shrout, Oren, et al.
Published: (2022)
by: Shrout, Oren, et al.
Published: (2022)
STELLAR: Scaling 3D Perception Large Models for Autonomous Driving
by: Li, Yingwei, et al.
Published: (2026)
by: Li, Yingwei, et al.
Published: (2026)
VoxelNextFusion: A Simple, Unified and Effective Voxel Fusion Framework for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
Efficient and Scalable Point Cloud Generation with Sparse Point-Voxel Diffusion Models
by: Romanelis, Ioannis, et al.
Published: (2024)
by: Romanelis, Ioannis, et al.
Published: (2024)
SceneCrafter: Controllable Multi-View Driving Scene Editing
by: Zhu, Zehao, et al.
Published: (2025)
by: Zhu, Zehao, et al.
Published: (2025)
BEVSpread: Spread Voxel Pooling for Bird's-Eye-View Representation in Vision-based Roadside 3D Object Detection
by: Wang, Wenjie, et al.
Published: (2024)
by: Wang, Wenjie, et al.
Published: (2024)
GVSynergy-Det: Synergistic Gaussian-Voxel Representations for Multi-View 3D Object Detection
by: Zhang, Yi, et al.
Published: (2025)
by: Zhang, Yi, et al.
Published: (2025)
3DTMDet: A Dual-Path Synergy Network of Transformer and SSM for 3D Object Detection in Point Clouds
by: Qiu, Bingwen, et al.
Published: (2026)
by: Qiu, Bingwen, et al.
Published: (2026)
PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
by: Huang, Kuan-Chih, et al.
Published: (2023)
by: Huang, Kuan-Chih, et al.
Published: (2023)
Voxel Densification for Serialized 3D Object Detection: Mitigating Sparsity via Pre-serialization Expansion
by: Liu, Qifeng, et al.
Published: (2025)
by: Liu, Qifeng, et al.
Published: (2025)
Weakly Supervised Point Clouds Transformer for 3D Object Detection
by: Tang, Zuojin, et al.
Published: (2023)
by: Tang, Zuojin, et al.
Published: (2023)
MonoCD: Monocular 3D Object Detection with Complementary Depths
by: Yan, Longfei, et al.
Published: (2024)
by: Yan, Longfei, et al.
Published: (2024)
PointDC:Unsupervised Semantic Segmentation of 3D Point Clouds via Cross-modal Distillation and Super-Voxel Clustering
by: Chen, Zisheng, et al.
Published: (2023)
by: Chen, Zisheng, et al.
Published: (2023)
OpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understanding
by: Huang, Sheng-Yu, et al.
Published: (2026)
by: Huang, Sheng-Yu, et al.
Published: (2026)
SHTOcc: Effective 3D Occupancy Prediction with Sparse Head and Tail Voxels
by: Yu, Qiucheng, et al.
Published: (2025)
by: Yu, Qiucheng, et al.
Published: (2025)
Study of Dropout in PointPillars with 3D Object Detection
by: Sun, Xiaoxiang, et al.
Published: (2024)
by: Sun, Xiaoxiang, et al.
Published: (2024)
OBMO: One Bounding Box Multiple Objects for Monocular 3D Object Detection
by: Huang, Chenxi, et al.
Published: (2022)
by: Huang, Chenxi, et al.
Published: (2022)
MS23D: A 3D Object Detection Method Using Multi-Scale Semantic Feature Points to Construct 3D Feature Layer
by: Shao, Yongxin, et al.
Published: (2023)
by: Shao, Yongxin, et al.
Published: (2023)
AVS-Net: Point Sampling with Adaptive Voxel Size for 3D Scene Understanding
by: Yang, Hongcheng, et al.
Published: (2024)
by: Yang, Hongcheng, et al.
Published: (2024)
PointVoxelFormer -- Reviving point cloud networks for 3D medical imaging
by: Heinrich, Mattias Paul
Published: (2024)
by: Heinrich, Mattias Paul
Published: (2024)
UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene Representation
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
Assembler: Scalable 3D Part Assembly via Anchor Point Diffusion
by: Zhao, Wang, et al.
Published: (2025)
by: Zhao, Wang, et al.
Published: (2025)
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
by: Hou, Jinghua, et al.
Published: (2024)
by: Hou, Jinghua, et al.
Published: (2024)
SceneDiffuser++: City-Scale Traffic Simulation via a Generative World Model
by: Tan, Shuhan, et al.
Published: (2025)
by: Tan, Shuhan, et al.
Published: (2025)
Drive&Gen: Co-Evaluating End-to-End Driving and Video Generation Models
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
Hierarchical Point Attention for Indoor 3D Object Detection
by: Shu, Manli, et al.
Published: (2023)
by: Shu, Manli, et al.
Published: (2023)
ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention
by: He, Chenhang, et al.
Published: (2024)
by: He, Chenhang, et al.
Published: (2024)
Similar Items
-
S4-Driver: Scalable Self-Supervised Driving Multimodal Large Language Modelwith Spatio-Temporal Visual Representation
by: Xie, Yichen, et al.
Published: (2025) -
Enhanced Motion Forecasting with Plug-and-Play Multimodal Large Language Models
by: Luo, Katie, et al.
Published: (2025) -
Scene Reconstruction as Mapping Priors for 3D Detection
by: Fu, Yang, et al.
Published: (2026) -
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
by: Lu, Yuxuan, et al.
Published: (2024) -
LET-3D-AP: Longitudinal Error Tolerant 3D Average Precision for Camera-Only 3D Detection
by: Hung, Wei-Chih, et al.
Published: (2022)