FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Su, Haisheng, Zhang, Junjie, Song, Feixiang, Zhou, Sanping, Wu, Wei, Zheng, Nanning, Yan, Junchi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
by: Huang, Yuhao, et al.
Published: (2023)
by: Huang, Yuhao, et al.
Published: (2023)
DriveMamba: Task-Centric Scalable State Space Model for Efficient End-to-End Autonomous Driving
by: Su, Haisheng, et al.
Published: (2026)
by: Su, Haisheng, et al.
Published: (2026)
RoboSense: Large-scale Dataset and Benchmark for Egocentric Robot Perception and Navigation in Crowded and Unstructured Environments
by: Su, Haisheng, et al.
Published: (2024)
by: Su, Haisheng, et al.
Published: (2024)
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025)
by: Yu, Jiangyong, et al.
Published: (2025)
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025)
by: Yu, Jiangyong, et al.
Published: (2025)
DVPE: Divided View Position Embedding for Multi-View 3D Object Detection
by: Wang, Jiasen, et al.
Published: (2024)
by: Wang, Jiasen, et al.
Published: (2024)
PMT: Progressive Mean Teacher via Exploring Temporal Consistency for Semi-Supervised Medical Image Segmentation
by: Gao, Ning, et al.
Published: (2024)
by: Gao, Ning, et al.
Published: (2024)
UniMamba: Unified Spatial-Channel Representation Learning with Group-Efficient Mamba for LiDAR-based 3D Object Detection
by: Jin, Xin, et al.
Published: (2025)
by: Jin, Xin, et al.
Published: (2025)
Robust Noisy Label Learning via Two-Stream Sample Distillation
by: Bai, Sihan, et al.
Published: (2024)
by: Bai, Sihan, et al.
Published: (2024)
Selective Transfer Learning of Cross-Modality Distillation for Monocular 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
FreqGRL: Suppressing Low-Frequency Bias and Mining High-Frequency Knowledge for Cross-Domain Few-Shot Learning
by: Hui, Siqi, et al.
Published: (2025)
by: Hui, Siqi, et al.
Published: (2025)
Towards Generalizable Multi-Object Tracking
by: Qin, Zheng, et al.
Published: (2024)
by: Qin, Zheng, et al.
Published: (2024)
Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection
by: Chen, Shitao, et al.
Published: (2024)
by: Chen, Shitao, et al.
Published: (2024)
FreqMoE: Dynamic Frequency Enhancement for Neural PDE Solvers
by: Chen, Tianyu, et al.
Published: (2025)
by: Chen, Tianyu, et al.
Published: (2025)
RayD3D: Distilling Depth Knowledge Along the Ray for Robust Multi-View 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
DAMap: Distance-aware MapNet for High Quality HD Map Construction
by: Dong, Jinpeng, et al.
Published: (2025)
by: Dong, Jinpeng, et al.
Published: (2025)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
by: Shen, Yanqing, et al.
Published: (2025)
by: Shen, Yanqing, et al.
Published: (2025)
Learning to Infer Unseen Single-/Multi-Attribute-Object Compositions with Graph Networks
by: Chen, Hui, et al.
Published: (2020)
by: Chen, Hui, et al.
Published: (2020)
OPEN: Object-wise Position Embedding for Multi-view 3D Object Detection
by: Hou, Jinghua, et al.
Published: (2024)
by: Hou, Jinghua, et al.
Published: (2024)
DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection
by: Guo, Junjie, et al.
Published: (2024)
by: Guo, Junjie, et al.
Published: (2024)
PR-DETR: Injecting Position and Relation Prior for Dense Video Captioning
by: Li, Yizhe, et al.
Published: (2025)
by: Li, Yizhe, et al.
Published: (2025)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
ARS-DETR: Aspect Ratio-Sensitive Detection Transformer for Aerial Oriented Object Detection
by: Zeng, Ying, et al.
Published: (2023)
by: Zeng, Ying, et al.
Published: (2023)
FreqX: Analyze the Attribution Methods in Another Domain
by: Liu, Zechen, et al.
Published: (2024)
by: Liu, Zechen, et al.
Published: (2024)
EVT: Efficient View Transformation for Multi-Modal 3D Object Detection
by: Lee, Yongjin, et al.
Published: (2024)
by: Lee, Yongjin, et al.
Published: (2024)
FreqTrack: Frequency Learning based Vision Transformer for RGB-Event Object Tracking
by: You, Jinlin, et al.
Published: (2026)
by: You, Jinlin, et al.
Published: (2026)
Molecule Design by Latent Prompt Transformer
by: Kong, Deqian, et al.
Published: (2024)
by: Kong, Deqian, et al.
Published: (2024)
Exploring Surround-View Fisheye Camera 3D Object Detection
by: Li, Changcai, et al.
Published: (2025)
by: Li, Changcai, et al.
Published: (2025)
MIC-BEV: Multi-Infrastructure Camera Bird's-Eye-View Transformer with Relation-Aware Fusion for 3D Object Detection
by: Zhang, Yun, et al.
Published: (2025)
by: Zhang, Yun, et al.
Published: (2025)
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
by: Dong, Yitong, et al.
Published: (2024)
by: Dong, Yitong, et al.
Published: (2024)
Exploring Hardware Friendly Bottleneck Architecture in CNN for Embedded Computing Systems
by: Lei, Xing, et al.
Published: (2024)
by: Lei, Xing, et al.
Published: (2024)
MDHA: Multi-Scale Deformable Transformer with Hybrid Anchors for Multi-View 3D Object Detection
by: Adeline, Michelle, et al.
Published: (2024)
by: Adeline, Michelle, et al.
Published: (2024)
FreqSem
by: Beller, Stephen
Published: (2025)
by: Beller, Stephen
Published: (2025)
From Detection to Association: Learning Discriminative Object Embeddings for Multi-Object Tracking
by: Shao, Yuqing, et al.
Published: (2025)
by: Shao, Yuqing, et al.
Published: (2025)
Rethinking Vision Transformer Depth via Structural Reparameterization
by: Zhou, Chengwei, et al.
Published: (2025)
by: Zhou, Chengwei, et al.
Published: (2025)
Rethinking Remote Sensing Change Detection With A Mask View
by: Ma, Xiaowen, et al.
Published: (2024)
by: Ma, Xiaowen, et al.
Published: (2024)
Rethinking Large Language Models For Irregular Time Series Classification In Critical Care
by: Zheng, Feixiang, et al.
Published: (2026)
by: Zheng, Feixiang, et al.
Published: (2026)
Breaking through the learning plateaus of in-context learning in Transformer
by: Fu, Jingwen, et al.
Published: (2023)
by: Fu, Jingwen, et al.
Published: (2023)
MonoCD: Monocular 3D Object Detection with Complementary Depths
by: Yan, Longfei, et al.
Published: (2024)
by: Yan, Longfei, et al.
Published: (2024)
FreqBlender: Enhancing DeepFake Detection by Blending Frequency Knowledge
by: Li, Hanzhe, et al.
Published: (2024)
by: Li, Hanzhe, et al.
Published: (2024)
Similar Items
-
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
by: Huang, Yuhao, et al.
Published: (2023) -
DriveMamba: Task-Centric Scalable State Space Model for Efficient End-to-End Autonomous Driving
by: Su, Haisheng, et al.
Published: (2026) -
RoboSense: Large-scale Dataset and Benchmark for Egocentric Robot Perception and Navigation in Crowded and Unstructured Environments
by: Su, Haisheng, et al.
Published: (2024) -
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025) -
FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection
by: Yu, Jiangyong, et al.
Published: (2025)