Learnable Query Aggregation with KV Routing for Cross-view Geo-localisation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Hualin, Liu, Bingxi, Du, Jixiang, Qin, Yu, Chen, Ziyi, Zhang, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VICI: VLM-Instructed Cross-view Image-localisation
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025)
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
Image Aesthetics Assessment via Learnable Queries
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023)
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023)
Multi-modal Learnable Queries for Image Aesthetics Assessment
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2024)
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2024)
Dual-R-DETR: Resolving Query Competition with Pairwise Routing in Transformer Decoders
von: Zhang, Ye, et al.
Veröffentlicht: (2025)
von: Zhang, Ye, et al.
Veröffentlicht: (2025)
Micro-gesture Online Recognition using Learnable Query Points
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries
von: Zhu, Yijie, et al.
Veröffentlicht: (2025)
von: Zhu, Yijie, et al.
Veröffentlicht: (2025)
Geo$^\textbf{2}$: Geometry-Guided Cross-view Geo-Localization and Image Synthesis
von: Zhang, Yancheng, et al.
Veröffentlicht: (2026)
von: Zhang, Yancheng, et al.
Veröffentlicht: (2026)
EmbodiedPlace: Learning Mixture-of-Features with Embodied Constraints for Visual Place Recognition
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
Hierarchical Visual Relocalization with Nearest View Synthesis from Feature Gaussian Splatting
von: Tao, Huaqi, et al.
Veröffentlicht: (2026)
von: Tao, Huaqi, et al.
Veröffentlicht: (2026)
MT-PCR: Hybrid Mamba-Transformer Network with Spatial Serialization for Point Cloud Registration
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
MobileGeo: Exploring Hierarchical Knowledge Distillation for Resource-Efficient Cross-view Drone Geo-Localization
von: Sun, Jian, et al.
Veröffentlicht: (2025)
von: Sun, Jian, et al.
Veröffentlicht: (2025)
TrackingMiM: Efficient Mamba-in-Mamba Serialization for Real-time UAV Object Tracking
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
von: Liu, Bingxi, et al.
Veröffentlicht: (2025)
Multi-view Aggregation Network for Dichotomous Image Segmentation
von: Yu, Qian, et al.
Veröffentlicht: (2024)
von: Yu, Qian, et al.
Veröffentlicht: (2024)
Boosting Multi-view Stereo with Late Cost Aggregation
von: Wu, Jiang, et al.
Veröffentlicht: (2024)
von: Wu, Jiang, et al.
Veröffentlicht: (2024)
Drone-assisted Road Gaussian Splatting with Cross-view Uncertainty
von: Zhang, Saining, et al.
Veröffentlicht: (2024)
von: Zhang, Saining, et al.
Veröffentlicht: (2024)
Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models
von: Ji, Yicheng, et al.
Veröffentlicht: (2026)
von: Ji, Yicheng, et al.
Veröffentlicht: (2026)
A Unified Hierarchical Framework for Fine-grained Cross-view Geo-localization over Large-scale Scenarios
von: Song, Zhuo, et al.
Veröffentlicht: (2025)
von: Song, Zhuo, et al.
Veröffentlicht: (2025)
Anchor-free Cross-view Object Geo-localization with Gaussian Position Encoding and Cross-view Association
von: Ling, Xingtao, et al.
Veröffentlicht: (2025)
von: Ling, Xingtao, et al.
Veröffentlicht: (2025)
Cross-view Localization and Synthesis -- Datasets, Challenges and Opportunities
von: Xu, Ningli, et al.
Veröffentlicht: (2025)
von: Xu, Ningli, et al.
Veröffentlicht: (2025)
Learning Cross-view Visual Geo-localization without Ground Truth
von: Li, Haoyuan, et al.
Veröffentlicht: (2024)
von: Li, Haoyuan, et al.
Veröffentlicht: (2024)
PostCam: Camera-Controllable Novel-View Video Generation with Query-Shared Cross-Attention
von: Chen, Yipeng, et al.
Veröffentlicht: (2025)
von: Chen, Yipeng, et al.
Veröffentlicht: (2025)
BoQ: A Place is Worth a Bag of Learnable Queries
von: Ali-Bey, Amar, et al.
Veröffentlicht: (2024)
von: Ali-Bey, Amar, et al.
Veröffentlicht: (2024)
Enhancing DETRs Variants through Improved Content Query and Similar Query Aggregation
von: Zhang, Yingying, et al.
Veröffentlicht: (2024)
von: Zhang, Yingying, et al.
Veröffentlicht: (2024)
JRN-Geo: A Joint Perception Network based on RGB and Normal images for Cross-view Geo-localization
von: Zhou, Hongyu, et al.
Veröffentlicht: (2025)
von: Zhou, Hongyu, et al.
Veröffentlicht: (2025)
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
von: Shehzadi, Tahira, et al.
Veröffentlicht: (2024)
TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification
von: Tao, Huaqi, et al.
Veröffentlicht: (2025)
von: Tao, Huaqi, et al.
Veröffentlicht: (2025)
PanoFree: Tuning-Free Holistic Multi-view Image Generation with Cross-view Self-Guidance
von: Liu, Aoming, et al.
Veröffentlicht: (2024)
von: Liu, Aoming, et al.
Veröffentlicht: (2024)
ConGeo: Robust Cross-view Geo-localization across Ground View Variations
von: Mi, Li, et al.
Veröffentlicht: (2024)
von: Mi, Li, et al.
Veröffentlicht: (2024)
Improving Cross-view Object Geo-localization: A Dual Attention Approach with Cross-view Interaction and Multi-Scale Spatial Features
von: Zhu, Xingtao Ling Yingying
Veröffentlicht: (2025)
von: Zhu, Xingtao Ling Yingying
Veröffentlicht: (2025)
LVLMs as inspectors: an agentic framework for category-level structural defect annotation
von: Jiang, Sheng, et al.
Veröffentlicht: (2025)
von: Jiang, Sheng, et al.
Veröffentlicht: (2025)
BREEN: Bridge Data-Efficient Encoder-Free Multimodal Learning with Learnable Queries
von: Li, Tianle, et al.
Veröffentlicht: (2025)
von: Li, Tianle, et al.
Veröffentlicht: (2025)
Monocular Depth Estimation via Neural Network with Learnable Algebraic Group and Ring Structures
von: Wang, Qianlei, et al.
Veröffentlicht: (2026)
von: Wang, Qianlei, et al.
Veröffentlicht: (2026)
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
von: Hu, Yuxi, et al.
Veröffentlicht: (2025)
von: Hu, Yuxi, et al.
Veröffentlicht: (2025)
Composable Visual Tokenizers with Generator-Free Diagnostics of Learnability
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
Learning Intra-view and Cross-view Geometric Knowledge for Stereo Matching
von: Gong, Rui, et al.
Veröffentlicht: (2024)
von: Gong, Rui, et al.
Veröffentlicht: (2024)
Cross-view image geo-localization with Panorama-BEV Co-Retrieval Network
von: Ye, Junyan, et al.
Veröffentlicht: (2024)
von: Ye, Junyan, et al.
Veröffentlicht: (2024)
LQ-Adapter: ViT-Adapter with Learnable Queries for Gallbladder Cancer Detection from Ultrasound Image
von: Madan, Chetan, et al.
Veröffentlicht: (2024)
von: Madan, Chetan, et al.
Veröffentlicht: (2024)
Adjacent-view Transformers for Supervised Surround-view Depth Estimation
von: Guo, Xianda, et al.
Veröffentlicht: (2023)
von: Guo, Xianda, et al.
Veröffentlicht: (2023)
Learnable Shape Prototypes with Occlusion-Geometry-Guided Injection for Amodal Instance Segmentation
von: Zhang, Fufan, et al.
Veröffentlicht: (2026)
von: Zhang, Fufan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VICI: VLM-Instructed Cross-view Image-localisation
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025) -
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
von: Liu, Bingxi, et al.
Veröffentlicht: (2025) -
Image Aesthetics Assessment via Learnable Queries
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2023) -
Multi-modal Learnable Queries for Image Aesthetics Assessment
von: Xiong, Zhiwei, et al.
Veröffentlicht: (2024) -
Dual-R-DETR: Resolving Query Competition with Pairwise Routing in Transformer Decoders
von: Zhang, Ye, et al.
Veröffentlicht: (2025)