Faster or Stronger: Towards Flexible Visual Place Recognition via Weighted Aggregation and Token Pruning
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Zichao, Goo, June Moh, Zheng, Junwei, Fan, Weijia, Zhang, Jiaming, Stiefelhagen, Rainer, Boehm, Jan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-shot detection of buildings in mobile LiDAR using Language Vision Model
by: Goo, June Moh, et al.
Published: (2024)
by: Goo, June Moh, et al.
Published: (2024)
A Comprehensive Survey on Deep Learning-Based LiDAR Super-Resolution for Autonomous Driving
by: Goo, June Moh, et al.
Published: (2026)
by: Goo, June Moh, et al.
Published: (2026)
Real-Time LiDAR Super-Resolution via Frequency-Aware Multi-Scale Fusion
by: Goo, June Moh, et al.
Published: (2025)
by: Goo, June Moh, et al.
Published: (2025)
Exploring Single Domain Generalization of LiDAR-based Semantic Segmentation under Imperfect Labels
by: Kong, Weitong, et al.
Published: (2025)
by: Kong, Weitong, et al.
Published: (2025)
Zero-shot Building Age Classification from Facade Image Using GPT-4
by: Zeng, Zichao, et al.
Published: (2024)
by: Zeng, Zichao, et al.
Published: (2024)
Comb, Prune, Distill: Towards Unified Pruning for Vision Model Compression
by: Schmitt, Jonas, et al.
Published: (2024)
by: Schmitt, Jonas, et al.
Published: (2024)
Checkerboard Target Measurement in Unordered Point Clouds with Coloured ICP
by: Goo, June Moh, et al.
Published: (2025)
by: Goo, June Moh, et al.
Published: (2025)
Data-Driven Socio-Economic Deprivation Prediction via Dimensionality Reduction: The Power of Diffusion Maps
by: Goo, June Moh
Published: (2023)
by: Goo, June Moh
Published: (2023)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
by: Zheng, Junwei, et al.
Published: (2023)
by: Zheng, Junwei, et al.
Published: (2023)
Hybrid-Segmentor: A Hybrid Approach to Automated Fine-Grained Crack Segmentation in Civil Infrastructure
by: Goo, June Moh, et al.
Published: (2024)
by: Goo, June Moh, et al.
Published: (2024)
More than the Sum: Panorama-Language Models for Adverse Omni-Scenes
by: Fan, Weijia, et al.
Published: (2026)
by: Fan, Weijia, et al.
Published: (2026)
Deformable Mamba for Wide Field of View Segmentation
by: Hu, Jie, et al.
Published: (2024)
by: Hu, Jie, et al.
Published: (2024)
Scene-agnostic Pose Regression for Visual Localization
by: Zheng, Junwei, et al.
Published: (2025)
by: Zheng, Junwei, et al.
Published: (2025)
V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024)
by: Jongwiriyanurak, Natchapon, et al.
Published: (2024)
RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization
by: Zheng, Junwei, et al.
Published: (2026)
by: Zheng, Junwei, et al.
Published: (2026)
Structured Pruning for Efficient Visual Place Recognition
by: Grainge, Oliver, et al.
Published: (2024)
by: Grainge, Oliver, et al.
Published: (2024)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
by: Wei, Jiale, et al.
Published: (2024)
by: Wei, Jiale, et al.
Published: (2024)
SGR3 Model: Scene Graph Retrieval-Reasoning Model in 3D
by: Wang, Zirui, et al.
Published: (2026)
by: Wang, Zirui, et al.
Published: (2026)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
by: Wei, Yiping, et al.
Published: (2023)
by: Wei, Yiping, et al.
Published: (2023)
Optimal Transport Aggregation for Visual Place Recognition
by: Izquierdo, Sergio, et al.
Published: (2023)
by: Izquierdo, Sergio, et al.
Published: (2023)
PointOBB-v2: Towards Simpler, Faster, and Stronger Single Point Supervised Oriented Object Detection
by: Ren, Botao, et al.
Published: (2024)
by: Ren, Botao, et al.
Published: (2024)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
by: Fan, Linjuan, et al.
Published: (2025)
by: Fan, Linjuan, et al.
Published: (2025)
Skeleton-Based Human Action Recognition with Noisy Labels
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
Snap, Segment, Deploy: A Visual Data and Detection Pipeline for Wearable Industrial Assistants
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
by: Chen, Yifei, et al.
Published: (2023)
by: Chen, Yifei, et al.
Published: (2023)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
by: Chen, Yufan, et al.
Published: (2024)
by: Chen, Yufan, et al.
Published: (2024)
HybriDLA: Hybrid Generation for Document Layout Analysis
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
Graph-based Document Structure Analysis
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
@Bench: Benchmarking Vision-Language Models for Human-centered Assistive Technology
by: Jiang, Xin, et al.
Published: (2024)
by: Jiang, Xin, et al.
Published: (2024)
PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models
by: Dhouib, Mohamed, et al.
Published: (2025)
by: Dhouib, Mohamed, et al.
Published: (2025)
RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning
by: Vogel, Alexander, et al.
Published: (2025)
by: Vogel, Alexander, et al.
Published: (2025)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
by: Tao, Mingzhe, et al.
Published: (2026)
by: Tao, Mingzhe, et al.
Published: (2026)
VST++: Efficient and Stronger Visual Saliency Transformer
by: Liu, Nian, et al.
Published: (2023)
by: Liu, Nian, et al.
Published: (2023)
Referring Atomic Video Action Recognition
by: Peng, Kunyu, et al.
Published: (2024)
by: Peng, Kunyu, et al.
Published: (2024)
Similarity-Aware Token Pruning: Your VLM but Faster
by: Jeddi, Ahmadreza, et al.
Published: (2025)
by: Jeddi, Ahmadreza, et al.
Published: (2025)
SuperPlace: The Renaissance of Classical Feature Aggregation for Visual Place Recognition in the Era of Foundation Models
by: Liu, Bingxi, et al.
Published: (2025)
by: Liu, Bingxi, et al.
Published: (2025)
Open Panoramic Segmentation
by: Zheng, Junwei, et al.
Published: (2024)
by: Zheng, Junwei, et al.
Published: (2024)
DC-VLAQ: Query-Residual Aggregation for Robust Visual Place Recognition
by: Zhu, Hanyu, et al.
Published: (2026)
by: Zhu, Hanyu, et al.
Published: (2026)
Query-Based Adaptive Aggregation for Multi-Dataset Joint Training Toward Universal Visual Place Recognition
by: Xiao, Jiuhong, et al.
Published: (2025)
by: Xiao, Jiuhong, et al.
Published: (2025)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
by: Shen, Yanqing, et al.
Published: (2025)
by: Shen, Yanqing, et al.
Published: (2025)
Similar Items
-
Zero-shot detection of buildings in mobile LiDAR using Language Vision Model
by: Goo, June Moh, et al.
Published: (2024) -
A Comprehensive Survey on Deep Learning-Based LiDAR Super-Resolution for Autonomous Driving
by: Goo, June Moh, et al.
Published: (2026) -
Real-Time LiDAR Super-Resolution via Frequency-Aware Multi-Scale Fusion
by: Goo, June Moh, et al.
Published: (2025) -
Exploring Single Domain Generalization of LiDAR-based Semantic Segmentation under Imperfect Labels
by: Kong, Weitong, et al.
Published: (2025) -
Zero-shot Building Age Classification from Facade Image Using GPT-4
by: Zeng, Zichao, et al.
Published: (2024)