Window-to-Window BEV Representation Learning for Limited FoV Cross-View Geo-localization
Fuente:
arXiv
Saved in:
| Main Authors: | Cheng, Lei, Wang, Teng, Meng, Lingquan, Sun, Changyin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition
by: Wang, Teng, et al.
Published: (2024)
by: Wang, Teng, et al.
Published: (2024)
Cross-modal Fundus Image Registration under Large FoV Disparity
by: Li, Hongyang, et al.
Published: (2025)
by: Li, Hongyang, et al.
Published: (2025)
FoV-Net: Rotation-Invariant CAD B-rep Learning via Field-of-View Ray Casting
by: Ballegeer, Matteo, et al.
Published: (2026)
by: Ballegeer, Matteo, et al.
Published: (2026)
OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback
by: Luo, Kai, et al.
Published: (2025)
by: Luo, Kai, et al.
Published: (2025)
Sequence-Preserving Dual-FoV Defense for Traffic Sign and Light Recognition in Autonomous Vehicles
by: Joshi, Abhishek, et al.
Published: (2025)
by: Joshi, Abhishek, et al.
Published: (2025)
GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection
by: Zhang, Jinqing, et al.
Published: (2024)
by: Zhang, Jinqing, et al.
Published: (2024)
SG-BEV: Satellite-Guided BEV Fusion for Cross-View Semantic Segmentation
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
P2U-SLAM: A Monocular Wide-FoV SLAM System Based on Point Uncertainty and Pose Uncertainty
by: Zhang, Yufan, et al.
Published: (2024)
by: Zhang, Yufan, et al.
Published: (2024)
BEV-CV: Birds-Eye-View Transform for Cross-View Geo-Localisation
by: Shore, Tavis, et al.
Published: (2023)
by: Shore, Tavis, et al.
Published: (2023)
A Modern Look at Simplicity Bias in Image Classification Tasks
by: Chang, Xiaoguang, et al.
Published: (2025)
by: Chang, Xiaoguang, et al.
Published: (2025)
Learning Ego-Centric BEV Representations from a Perspective-Privileged View: Cross-View Supervision for Online HD Map Construction
by: Lengerer, Daniel, et al.
Published: (2026)
by: Lengerer, Daniel, et al.
Published: (2026)
VA-AR: Learning Velocity-Aware Action Representations with Mixture of Window Attention
by: Wei, Jiangning, et al.
Published: (2025)
by: Wei, Jiangning, et al.
Published: (2025)
Multi-Scale Representations by Varying Window Attention for Semantic Segmentation
by: Yan, Haotian, et al.
Published: (2024)
by: Yan, Haotian, et al.
Published: (2024)
MamBEV: Enabling State Space Models to Learn Birds-Eye-View Representations
by: Ke, Hongyu, et al.
Published: (2025)
by: Ke, Hongyu, et al.
Published: (2025)
Cross-level Attention with Overlapped Windows for Camouflaged Object Detection
by: Li, Jiepan, et al.
Published: (2023)
by: Li, Jiepan, et al.
Published: (2023)
RoadBEV: Road Surface Reconstruction in Bird's Eye View
by: Zhao, Tong, et al.
Published: (2024)
by: Zhao, Tong, et al.
Published: (2024)
FoVA-Depth: Field-of-View Agnostic Depth Estimation for Cross-Dataset Generalization
by: Lichy, Daniel, et al.
Published: (2024)
by: Lichy, Daniel, et al.
Published: (2024)
InstanceBEV: Unifying Instance and BEV Representation for 3D Panoptic Segmentation
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
Shifted Window Fourier Transform And Retention For Image Captioning
by: Hu, Jia Cheng, et al.
Published: (2024)
by: Hu, Jia Cheng, et al.
Published: (2024)
Hierarchical and Decoupled BEV Perception Learning Framework for Autonomous Driving
by: Dai, Yuqi, et al.
Published: (2024)
by: Dai, Yuqi, et al.
Published: (2024)
R3D-SWIN:Use Shifted Window Attention for Single-View 3D Reconstruction
by: Li, Chenhuan, et al.
Published: (2023)
by: Li, Chenhuan, et al.
Published: (2023)
Transcending the Limit of Local Window: Advanced Super-Resolution Transformer with Adaptive Token Dictionary
by: Zhang, Leheng, et al.
Published: (2024)
by: Zhang, Leheng, et al.
Published: (2024)
Cross-view image geo-localization with Panorama-BEV Co-Retrieval Network
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
Referring Camouflaged Object Detection With Multi-Context Overlapped Windows Cross-Attention
by: Wen, Yu, et al.
Published: (2025)
by: Wen, Yu, et al.
Published: (2025)
Refine-and-Contrast: Adaptive Instance-Aware BEV Representations for Multi-UAV Collaborative Object Detection
by: Li, Zhongyao, et al.
Published: (2025)
by: Li, Zhongyao, et al.
Published: (2025)
WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization
by: Tao, Wei, et al.
Published: (2026)
by: Tao, Wei, et al.
Published: (2026)
MCMS: Multi-Category Information and Multi-Scale Stripe Attention for Blind Motion Deblurring
by: Qiao, Nianzu, et al.
Published: (2024)
by: Qiao, Nianzu, et al.
Published: (2024)
Single Image Super-Resolution Based on Global-Local Information Synergy
by: Qiao, Nianzu, et al.
Published: (2024)
by: Qiao, Nianzu, et al.
Published: (2024)
FishBEV: Distortion-Resilient Bird's Eye View Segmentation with Surround-View Fisheye Cameras
by: Li, Hang, et al.
Published: (2025)
by: Li, Hang, et al.
Published: (2025)
SWARD: Stochastic Window-Attention-Based Relational Distillation for Cross-Architectural Semantic Segmentation
by: Makineni, Aditya, et al.
Published: (2026)
by: Makineni, Aditya, et al.
Published: (2026)
Video2BEV: Transforming Drone Videos to BEVs for Video-based Geo-localization
by: Ju, Hao, et al.
Published: (2024)
by: Ju, Hao, et al.
Published: (2024)
OnlineBEV: Recurrent Temporal Fusion in Bird's Eye View Representations for Multi-Camera 3D Perception
by: Koh, Junho, et al.
Published: (2025)
by: Koh, Junho, et al.
Published: (2025)
MaskBEV: Towards A Unified Framework for BEV Detection and Map Segmentation
by: Zhao, Xiao, et al.
Published: (2024)
by: Zhao, Xiao, et al.
Published: (2024)
DualBEV: Unifying Dual View Transformation with Probabilistic Correspondences
by: Li, Peidong, et al.
Published: (2024)
by: Li, Peidong, et al.
Published: (2024)
KD360-VoxelBEV: LiDAR and 360-degree Camera Cross Modality Knowledge Distillation for Bird's-Eye-View Segmentation
by: E, Wenke, et al.
Published: (2025)
by: E, Wenke, et al.
Published: (2025)
RFR-WWANet: Weighted Window Attention-Based Recovery Feature Resolution Network for Unsupervised Image Registration
by: Ma, Mingrui, et al.
Published: (2023)
by: Ma, Mingrui, et al.
Published: (2023)
TinyBEV: Cross Modal Knowledge Distillation for Efficient Multi Task Bird's Eye View Perception and Planning
by: Khan, Reeshad, et al.
Published: (2025)
by: Khan, Reeshad, et al.
Published: (2025)
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
by: Li, Bohan, et al.
Published: (2023)
by: Li, Bohan, et al.
Published: (2023)
Where am I? Cross-View Geo-localization with Natural Language Descriptions
by: Ye, Junyan, et al.
Published: (2024)
by: Ye, Junyan, et al.
Published: (2024)
Sliding-Window Merging for Compacting Patch-Redundant Layers in LLMs
by: Ding, Xuan, et al.
Published: (2025)
by: Ding, Xuan, et al.
Published: (2025)
Similar Items
-
LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition
by: Wang, Teng, et al.
Published: (2024) -
Cross-modal Fundus Image Registration under Large FoV Disparity
by: Li, Hongyang, et al.
Published: (2025) -
FoV-Net: Rotation-Invariant CAD B-rep Learning via Field-of-View Ray Casting
by: Ballegeer, Matteo, et al.
Published: (2026) -
OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback
by: Luo, Kai, et al.
Published: (2025) -
Sequence-Preserving Dual-FoV Defense for Traffic Sign and Light Recognition in Autonomous Vehicles
by: Joshi, Abhishek, et al.
Published: (2025)