DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Jiasen, Li, Zhenglin, Sun, Ke, Liu, Xianyuan, Zhou, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Multi-View BEV Object Detection with Mixed Pinhole and Fisheye Cameras
by: Liu, Xiangzhong, et al.
Published: (2026)
by: Liu, Xiangzhong, et al.
Published: (2026)
DuoSpaceNet: Leveraging Both Bird's-Eye-View and Perspective View Representations for 3D Object Detection
by: Huang, Zhe, et al.
Published: (2024)
by: Huang, Zhe, et al.
Published: (2024)
SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection
by: Papais, Sandro, et al.
Published: (2026)
by: Papais, Sandro, et al.
Published: (2026)
ForeSight: Multi-View Streaming Joint Object Detection and Trajectory Forecasting
by: Papais, Sandro, et al.
Published: (2025)
by: Papais, Sandro, et al.
Published: (2025)
UEVAVD: A Dataset for Developing UAV's Eye View Active Object Detection
by: Jiang, Xinhua, et al.
Published: (2024)
by: Jiang, Xinhua, et al.
Published: (2024)
Active 6D Pose Estimation for Textureless Objects using Multi-View RGB Frames
by: Yang, Jun, et al.
Published: (2025)
by: Yang, Jun, et al.
Published: (2025)
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025)
by: Yu, Jiangyong, et al.
Published: (2025)
Multi-View Video Diffusion Policy: A 3D Spatio-Temporal-Aware Video Action Model
by: Li, Peiyan, et al.
Published: (2026)
by: Li, Peiyan, et al.
Published: (2026)
FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
by: Su, Haisheng, et al.
Published: (2025)
by: Su, Haisheng, et al.
Published: (2025)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
by: Banerjee, Prithviraj, et al.
Published: (2024)
by: Banerjee, Prithviraj, et al.
Published: (2024)
DKPMV: Dense Keypoints Fusion from Multi-View RGB Frames for 6D Pose Estimation of Textureless Objects
by: Chen, Jiahong, et al.
Published: (2025)
by: Chen, Jiahong, et al.
Published: (2025)
DreamGrasp: Zero-Shot 3D Multi-Object Reconstruction from Partial-View Images for Robotic Manipulation
by: Kim, Young Hun, et al.
Published: (2025)
by: Kim, Young Hun, et al.
Published: (2025)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
by: Jeong, Seunghoon, et al.
Published: (2026)
by: Jeong, Seunghoon, et al.
Published: (2026)
PAct: Part-Decomposed Single-View Articulated Object Generation
by: Liu, Qingming, et al.
Published: (2026)
by: Liu, Qingming, et al.
Published: (2026)
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025)
by: Yu, Jiangyong, et al.
Published: (2025)
FastOcc: Accelerating 3D Occupancy Prediction by Fusing the 2D Bird's-Eye View and Perspective View
by: Hou, Jiawei, et al.
Published: (2024)
by: Hou, Jiawei, et al.
Published: (2024)
3D Extended Object Tracking based on Extruded B-Spline Side View Profiles
by: Han, Longfei, et al.
Published: (2025)
by: Han, Longfei, et al.
Published: (2025)
StreamMOS: Streaming Moving Object Segmentation with Multi-View Perception and Dual-Span Memory
by: Li, Zhiheng, et al.
Published: (2024)
by: Li, Zhiheng, et al.
Published: (2024)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
Crowd-Sourced NeRF: Collecting Data from Production Vehicles for 3D Street View Reconstruction
by: Qin, Tong, et al.
Published: (2024)
by: Qin, Tong, et al.
Published: (2024)
MV-SSM: Multi-View State Space Modeling for 3D Human Pose Estimation
by: Chharia, Aviral, et al.
Published: (2025)
by: Chharia, Aviral, et al.
Published: (2025)
Boundary Exploration of Next Best View Policy in 3D Robotic Scanning
by: Li, Leihui, et al.
Published: (2024)
by: Li, Leihui, et al.
Published: (2024)
A Multi-Level Similarity Approach for Single-View Object Grasping: Matching, Planning, and Fine-Tuning
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
DRoPE: Directional Rotary Position Embedding for Efficient Agent Interaction Modeling
by: Zhao, Jianbo, et al.
Published: (2025)
by: Zhao, Jianbo, et al.
Published: (2025)
Active Implicit Object Reconstruction using Uncertainty-guided Next-Best-View Optimization
by: Yan, Dongyu, et al.
Published: (2023)
by: Yan, Dongyu, et al.
Published: (2023)
DM-OSVP++: One-Shot View Planning Using 3D Diffusion Models for Active RGB-Based Object Reconstruction
by: Pan, Sicong, et al.
Published: (2025)
by: Pan, Sicong, et al.
Published: (2025)
POp-GS: Next Best View in 3D-Gaussian Splatting with P-Optimality
by: Wilson, Joey, et al.
Published: (2025)
by: Wilson, Joey, et al.
Published: (2025)
Safe-Construct: Redefining Construction Safety Violation Recognition as 3D Multi-View Engagement Task
by: Chharia, Aviral, et al.
Published: (2025)
by: Chharia, Aviral, et al.
Published: (2025)
Robotic Arm Platform for Multi-View Image Acquisition and 3D Reconstruction in Minimally Invasive Surgery
by: Saikia, Alexander, et al.
Published: (2024)
by: Saikia, Alexander, et al.
Published: (2024)
Mobile Robotic Multi-View Photometric Stereo
by: Kumar, Suryansh
Published: (2025)
by: Kumar, Suryansh
Published: (2025)
BFA: Best-Feature-Aware Fusion for Multi-View Fine-grained Manipulation
by: Lan, Zihan, et al.
Published: (2025)
by: Lan, Zihan, et al.
Published: (2025)
UniScale: Unified Scale-Aware 3D Reconstruction for Multi-View Understanding via Prior Injection for Robotic Perception
by: Mahdavian, Mohammad, et al.
Published: (2026)
by: Mahdavian, Mohammad, et al.
Published: (2026)
Visual SLAM with 3D Gaussian Primitives and Depth Priors Enabling Novel View Synthesis
by: Qu, Zhongche, et al.
Published: (2024)
by: Qu, Zhongche, et al.
Published: (2024)
NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation
by: Liu, Youzhi, et al.
Published: (2024)
by: Liu, Youzhi, et al.
Published: (2024)
Perspective-Invariant 3D Object Detection
by: Liang, Ao, et al.
Published: (2025)
by: Liang, Ao, et al.
Published: (2025)
WLTCL: Wide Field-of-View 3-D LiDAR Truck Compartment Automatic Localization System
by: Sun, Guodong, et al.
Published: (2025)
by: Sun, Guodong, et al.
Published: (2025)
Toward General Object-level Mapping from Sparse Views with 3D Diffusion Priors
by: Liao, Ziwei, et al.
Published: (2024)
by: Liao, Ziwei, et al.
Published: (2024)
MrGS: Multi-modal Radiance Fields with 3D Gaussian Splatting for RGB-Thermal Novel View Synthesis
by: Kweon, Minseong, et al.
Published: (2025)
by: Kweon, Minseong, et al.
Published: (2025)
Similar Items
-
Benchmarking Multi-View BEV Object Detection with Mixed Pinhole and Fisheye Cameras
by: Liu, Xiangzhong, et al.
Published: (2026) -
DuoSpaceNet: Leveraging Both Bird's-Eye-View and Perspective View Representations for 3D Object Detection
by: Huang, Zhe, et al.
Published: (2024) -
SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection
by: Papais, Sandro, et al.
Published: (2026) -
ForeSight: Multi-View Streaming Joint Object Detection and Trajectory Forecasting
by: Papais, Sandro, et al.
Published: (2025) -
UEVAVD: A Dataset for Developing UAV's Eye View Active Object Detection
by: Jiang, Xinhua, et al.
Published: (2024)