WidthFormer: Toward Efficient Transformer-based BEV View Transformation
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Chenhongyi, Lin, Tianwei, Huang, Lichao, Crowley, Elliot J. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Plug and Play Active Learning for Object Detection
by: Yang, Chenhongyi, et al.
Published: (2022)
by: Yang, Chenhongyi, et al.
Published: (2022)
EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation
by: Yang, Chenhongyi, et al.
Published: (2024)
by: Yang, Chenhongyi, et al.
Published: (2024)
There is no SAMantics! Exploring SAM as a Backbone for Visual Understanding Tasks
by: Espinosa, Miguel, et al.
Published: (2024)
by: Espinosa, Miguel, et al.
Published: (2024)
No time to train! Training-Free Reference-Based Instance Segmentation
by: Espinosa, Miguel, et al.
Published: (2025)
by: Espinosa, Miguel, et al.
Published: (2025)
DualBEV: Unifying Dual View Transformation with Probabilistic Correspondences
by: Li, Peidong, et al.
Published: (2024)
by: Li, Peidong, et al.
Published: (2024)
CountFormer: Multi-View Crowd Counting Transformer
by: Mo, Hong, et al.
Published: (2024)
by: Mo, Hong, et al.
Published: (2024)
DenseBEV: Transforming BEV Grid Cells into 3D Objects
by: Dähling, Marius, et al.
Published: (2025)
by: Dähling, Marius, et al.
Published: (2025)
PlainMamba: Improving Non-Hierarchical Mamba in Visual Recognition
by: Yang, Chenhongyi, et al.
Published: (2024)
by: Yang, Chenhongyi, et al.
Published: (2024)
Focus on BEV: Self-calibrated Cycle View Transformation for Monocular Birds-Eye-View Segmentation
by: Zhao, Jiawei, et al.
Published: (2024)
by: Zhao, Jiawei, et al.
Published: (2024)
BEV-CV: Birds-Eye-View Transform for Cross-View Geo-Localisation
by: Shore, Tavis, et al.
Published: (2023)
by: Shore, Tavis, et al.
Published: (2023)
COP-GEN: Latent Diffusion Transformer for Copernicus Earth Observation Data
by: Espinosa, Miguel, et al.
Published: (2026)
by: Espinosa, Miguel, et al.
Published: (2026)
Video2BEV: Transforming Drone Videos to BEVs for Video-based Geo-localization
by: Ju, Hao, et al.
Published: (2024)
by: Ju, Hao, et al.
Published: (2024)
ViewFormer: Exploring Spatiotemporal Modeling for Multi-View 3D Occupancy Perception via View-Guided Transformers
by: Li, Jinke, et al.
Published: (2024)
by: Li, Jinke, et al.
Published: (2024)
MIC-BEV: Multi-Infrastructure Camera Bird's-Eye-View Transformer with Relation-Aware Fusion for 3D Object Detection
by: Zhang, Yun, et al.
Published: (2025)
by: Zhang, Yun, et al.
Published: (2025)
Inter2Former: Dynamic Hybrid Attention for Efficient High-Precision Interactive
by: Huang, You, et al.
Published: (2025)
by: Huang, You, et al.
Published: (2025)
CycleBEV: Regularizing View Transformation Networks via View Cycle Consistency for Bird's-Eye-View Semantic Segmentation
by: Hong, Jeongbin, et al.
Published: (2026)
by: Hong, Jeongbin, et al.
Published: (2026)
LowFormer: Hardware Efficient Design for Convolutional Transformer Backbones
by: Nottebaum, Moritz, et al.
Published: (2024)
by: Nottebaum, Moritz, et al.
Published: (2024)
ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention
by: He, Chenhang, et al.
Published: (2024)
by: He, Chenhang, et al.
Published: (2024)
MixFormerV2: Efficient Fully Transformer Tracking
by: Cui, Yutao, et al.
Published: (2023)
by: Cui, Yutao, et al.
Published: (2023)
SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
GraphBEV: Towards Robust BEV Feature Alignment for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2024)
by: Song, Ziying, et al.
Published: (2024)
PosFormer: Recognizing Complex Handwritten Mathematical Expression with Position Forest Transformer
by: Guan, Tongkun, et al.
Published: (2024)
by: Guan, Tongkun, et al.
Published: (2024)
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
by: Zhao, Xiao, et al.
Published: (2024)
by: Zhao, Xiao, et al.
Published: (2024)
FakeFormer: Efficient Vulnerability-Driven Transformers for Generalisable Deepfake Detection
by: Nguyen, Dat, et al.
Published: (2024)
by: Nguyen, Dat, et al.
Published: (2024)
BEVPredFormer: Spatio-temporal Attention for BEV Instance Prediction in Autonomous Driving
by: Antunes-García, Miguel, et al.
Published: (2026)
by: Antunes-García, Miguel, et al.
Published: (2026)
TransMed: Large Language Models Enhance Vision Transformer for Biomedical Image Classification
by: Zheng, Kaipeng, et al.
Published: (2023)
by: Zheng, Kaipeng, et al.
Published: (2023)
To View Transform or Not to View Transform: NeRF-based Pre-training Perspective
by: Jeong, Hyeonjun, et al.
Published: (2026)
by: Jeong, Hyeonjun, et al.
Published: (2026)
BRepFormer: Transformer-Based B-rep Geometric Feature Recognition
by: Dai, Yongkang, et al.
Published: (2025)
by: Dai, Yongkang, et al.
Published: (2025)
RoadBEV: Road Surface Reconstruction in Bird's Eye View
by: Zhao, Tong, et al.
Published: (2024)
by: Zhao, Tong, et al.
Published: (2024)
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning
by: Zhong, Hanwen, et al.
Published: (2025)
by: Zhong, Hanwen, et al.
Published: (2025)
PC-BEV: An Efficient Polar-Cartesian BEV Fusion Framework for LiDAR Semantic Segmentation
by: Qiu, Shoumeng, et al.
Published: (2024)
by: Qiu, Shoumeng, et al.
Published: (2024)
DA-BEV: Unsupervised Domain Adaptation for Bird's Eye View Perception
by: Jiang, Kai, et al.
Published: (2024)
by: Jiang, Kai, et al.
Published: (2024)
Proto-Former: Unified Facial Landmark Detection by Prototype Transformer
by: Hu, Shengkai, et al.
Published: (2025)
by: Hu, Shengkai, et al.
Published: (2025)
einspace: Searching for Neural Architectures from Fundamental Operations
by: Ericsson, Linus, et al.
Published: (2024)
by: Ericsson, Linus, et al.
Published: (2024)
SeaFormer++: Squeeze-enhanced Axial Transformer for Mobile Visual Recognition
by: Wan, Qiang, et al.
Published: (2023)
by: Wan, Qiang, et al.
Published: (2023)
ImplantFormer: Vision Transformer based Implant Position Regression Using Dental CBCT Data
by: Yang, Xinquan, et al.
Published: (2022)
by: Yang, Xinquan, et al.
Published: (2022)
SegFormer3D: an Efficient Transformer for 3D Medical Image Segmentation
by: Perera, Shehan, et al.
Published: (2024)
by: Perera, Shehan, et al.
Published: (2024)
ROA-BEV: 2D Region-Oriented Attention for BEV-based 3D Object Detection
by: Chen, Jiwei, et al.
Published: (2024)
by: Chen, Jiwei, et al.
Published: (2024)
AnchorFormer: Differentiable Anchor Attention for Efficient Vision Transformer
by: Shan, Jiquan, et al.
Published: (2025)
by: Shan, Jiquan, et al.
Published: (2025)
Transforming Image Super-Resolution: A ConvFormer-based Efficient Approach
by: Wu, Gang, et al.
Published: (2024)
by: Wu, Gang, et al.
Published: (2024)
Similar Items
-
Plug and Play Active Learning for Object Detection
by: Yang, Chenhongyi, et al.
Published: (2022) -
EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation
by: Yang, Chenhongyi, et al.
Published: (2024) -
There is no SAMantics! Exploring SAM as a Backbone for Visual Understanding Tasks
by: Espinosa, Miguel, et al.
Published: (2024) -
No time to train! Training-Free Reference-Based Instance Segmentation
by: Espinosa, Miguel, et al.
Published: (2025) -
DualBEV: Unifying Dual View Transformation with Probabilistic Correspondences
by: Li, Peidong, et al.
Published: (2024)