Towards Seamless Adaptation of Pre-trained Models for Visual Place Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Feng, Zhang, Lijun, Lan, Xiangyuan, Dong, Shuting, Wang, Yaowei, Yuan, Chun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Homography Estimation for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024)
von: Lu, Feng, et al.
Veröffentlicht: (2024)
SelaVPR++: Towards Seamless Adaptation of Foundation Models for Efficient Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2025)
von: Lu, Feng, et al.
Veröffentlicht: (2025)
CricaVPR: Cross-image Correlation-aware Representation Learning for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024)
von: Lu, Feng, et al.
Veröffentlicht: (2024)
Towards Implicit Aggregation: Robust Image Representation for Place Recognition in the Transformer Era
von: Lu, Feng, et al.
Veröffentlicht: (2025)
von: Lu, Feng, et al.
Veröffentlicht: (2025)
Pair-VPR: Place-Aware Pre-training and Contrastive Pair Classification for Visual Place Recognition with Vision Transformers
von: Hausler, Stephen, et al.
Veröffentlicht: (2024)
von: Hausler, Stephen, et al.
Veröffentlicht: (2024)
Efficient Adversarial Training via Criticality-Aware Fine-Tuning
von: Li, Wenyun, et al.
Veröffentlicht: (2026)
von: Li, Wenyun, et al.
Veröffentlicht: (2026)
Contribution-based Low-Rank Adaptation with Pre-training Model for Real Image Restoration
von: Park, Donwon, et al.
Veröffentlicht: (2024)
von: Park, Donwon, et al.
Veröffentlicht: (2024)
Efficient Adaptation of Pre-trained Vision Transformer via Householder Transformation
von: Dong, Wei, et al.
Veröffentlicht: (2024)
von: Dong, Wei, et al.
Veröffentlicht: (2024)
DM-Adapter: Domain-Aware Mixture-of-Adapters for Text-Based Person Retrieval
von: Liu, Yating, et al.
Veröffentlicht: (2025)
von: Liu, Yating, et al.
Veröffentlicht: (2025)
Spatio-Temporal Side Tuning Pre-trained Foundation Models for Video-based Pedestrian Attribute Recognition
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment
von: Xing, Yifei, et al.
Veröffentlicht: (2024)
von: Xing, Yifei, et al.
Veröffentlicht: (2024)
Efficient Adaptation of Pre-trained Vision Transformer underpinned by Approximately Orthogonal Fine-Tuning Strategy
von: Yang, Yiting, et al.
Veröffentlicht: (2025)
von: Yang, Yiting, et al.
Veröffentlicht: (2025)
EffoVPR: Effective Foundation Model Utilization for Visual Place Recognition
von: Tzachor, Issar, et al.
Veröffentlicht: (2024)
von: Tzachor, Issar, et al.
Veröffentlicht: (2024)
On Pre-training of Multimodal Language Models Customized for Chart Understanding
von: Fan, Wan-Cyuan, et al.
Veröffentlicht: (2024)
von: Fan, Wan-Cyuan, et al.
Veröffentlicht: (2024)
Generative Pre-trained Autoregressive Diffusion Transformer
von: Zhang, Yuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yuan, et al.
Veröffentlicht: (2025)
Faster or Stronger: Towards Flexible Visual Place Recognition via Weighted Aggregation and Token Pruning
von: Zeng, Zichao, et al.
Veröffentlicht: (2026)
von: Zeng, Zichao, et al.
Veröffentlicht: (2026)
KappaPlace: Learning Hyperspherical Uncertainty for Visual Place Recognition via Prototype-Anchored Supervision
von: Yanko, Maya, et al.
Veröffentlicht: (2026)
von: Yanko, Maya, et al.
Veröffentlicht: (2026)
EPRBench: A High-Quality Benchmark Dataset for Event Stream Based Visual Place Recognition
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
Towards Visual Grounding: A Survey
von: Xiao, Linhui, et al.
Veröffentlicht: (2024)
von: Xiao, Linhui, et al.
Veröffentlicht: (2024)
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
Long-Tailed Recognition on Binary Networks by Calibrating A Pre-trained Model
von: Kim, Jihun, et al.
Veröffentlicht: (2024)
von: Kim, Jihun, et al.
Veröffentlicht: (2024)
Transferable Adversarial Face Attack with Text Controlled Attribute
von: Li, Wenyun, et al.
Veröffentlicht: (2024)
von: Li, Wenyun, et al.
Veröffentlicht: (2024)
In-Depth and In-Breadth: Pre-training Multimodal Language Models Customized for Comprehensive Chart Understanding
von: Fan, Wan-Cyuan, et al.
Veröffentlicht: (2025)
von: Fan, Wan-Cyuan, et al.
Veröffentlicht: (2025)
UniPre3D: Unified Pre-training of 3D Point Cloud Models with Cross-Modal Gaussian Splatting
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
Towards Fine-Grained Recognition with Large Visual Language Models: Benchmark and Optimization Strategies
von: Pang, Cong, et al.
Veröffentlicht: (2025)
von: Pang, Cong, et al.
Veröffentlicht: (2025)
Bayesian Exploration of Pre-trained Models for Low-shot Image Classification
von: Miao, Yibo, et al.
Veröffentlicht: (2024)
von: Miao, Yibo, et al.
Veröffentlicht: (2024)
EDTformer: An Efficient Decoder Transformer for Visual Place Recognition
von: Jin, Tong, et al.
Veröffentlicht: (2024)
von: Jin, Tong, et al.
Veröffentlicht: (2024)
RGB-Event HyperGraph Prompt for Kilometer Marker Recognition based on Pre-trained Foundation Models
von: Xian, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Xian, Xiaoyu, et al.
Veröffentlicht: (2026)
Do Pre-trained Vision-Language Models Encode Object States?
von: Newman, Kaleb, et al.
Veröffentlicht: (2024)
von: Newman, Kaleb, et al.
Veröffentlicht: (2024)
VITAL: Vision-Encoder-centered Pre-training for LMMs in Visual Quality Assessment
von: Jia, Ziheng, et al.
Veröffentlicht: (2025)
von: Jia, Ziheng, et al.
Veröffentlicht: (2025)
VELoRA: A Low-Rank Adaptation Approach for Efficient RGB-Event based Recognition
von: Chen, Lan, et al.
Veröffentlicht: (2024)
von: Chen, Lan, et al.
Veröffentlicht: (2024)
Adjusting Logit in Gaussian Form for Long-Tailed Visual Recognition
von: Li, Mengke, et al.
Veröffentlicht: (2023)
von: Li, Mengke, et al.
Veröffentlicht: (2023)
Going Places: Place Recognition in Artificial and Natural Systems
von: Milford, Michael, et al.
Veröffentlicht: (2025)
von: Milford, Michael, et al.
Veröffentlicht: (2025)
Traj-LLM: A New Exploration for Empowering Trajectory Prediction with Pre-trained Large Language Models
von: Lan, Zhengxing, et al.
Veröffentlicht: (2024)
von: Lan, Zhengxing, et al.
Veröffentlicht: (2024)
Practical Continual Forgetting for Pre-trained Vision Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
PersGuard: Preventing Malicious Personalization via Backdoor Attacks on Pre-trained Text-to-Image Diffusion Models
von: Liu, Xinwei, et al.
Veröffentlicht: (2025)
von: Liu, Xinwei, et al.
Veröffentlicht: (2025)
Towards A Comprehensive Visual Saliency Explanation Framework for AI-based Face Recognition Systems
von: Lu, Yuhang, et al.
Veröffentlicht: (2024)
von: Lu, Yuhang, et al.
Veröffentlicht: (2024)
Pre-trained Vision-Language Models Learn Discoverable Visual Concepts
von: Zang, Yuan, et al.
Veröffentlicht: (2024)
von: Zang, Yuan, et al.
Veröffentlicht: (2024)
Enhancing Adversarial Transferability in Visual-Language Pre-training Models via Local Shuffle and Sample-based Attack
von: Liu, Xin, et al.
Veröffentlicht: (2025)
von: Liu, Xin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Deep Homography Estimation for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024) -
SelaVPR++: Towards Seamless Adaptation of Foundation Models for Efficient Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2025) -
CricaVPR: Cross-image Correlation-aware Representation Learning for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024) -
Towards Implicit Aggregation: Robust Image Representation for Place Recognition in the Transformer Era
von: Lu, Feng, et al.
Veröffentlicht: (2025) -
Pair-VPR: Place-Aware Pre-training and Contrastive Pair Classification for Visual Place Recognition with Vision Transformers
von: Hausler, Stephen, et al.
Veröffentlicht: (2024)