Map-Free Visual Relocalization Enhanced by Instance Knowledge and Depth Knowledge
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Mingyu, Chen, Runze, Luo, Haiyong, Zhao, Fang, Wang, Juan, Ma, Xuepeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Structure-Centric Robust Monocular Depth Estimation via Knowledge Distillation
by: Chen, Runze, et al.
Published: (2024)
by: Chen, Runze, et al.
Published: (2024)
Visual Instance-aware Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
KARL: Knowledge-Aware Reasoning and Reinforcement Learning for Knowledge-Intensive Visual Grounding
by: Ma, Xinyu, et al.
Published: (2025)
by: Ma, Xinyu, et al.
Published: (2025)
No Pose Estimation? No Problem: Pose-Agnostic and Instance-Aware Test-Time Adaptation for Monocular Depth Estimation
by: Sung, Mingyu, et al.
Published: (2025)
by: Sung, Mingyu, et al.
Published: (2025)
Teeth-SEG: An Efficient Instance Segmentation Framework for Orthodontic Treatment based on Anthropic Prior Knowledge
by: Zou, Bo, et al.
Published: (2024)
by: Zou, Bo, et al.
Published: (2024)
Fine-Grained Knowledge Structuring and Retrieval for Visual Question Answering
by: Zhang, Zhengxuan, et al.
Published: (2025)
by: Zhang, Zhengxuan, et al.
Published: (2025)
Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs Reasoning
by: Liu, Junming, et al.
Published: (2025)
by: Liu, Junming, et al.
Published: (2025)
MIRAGE: Knowledge Graph-Guided Cross-Cohort MRI Synthesis for Alzheimer's Disease Prediction
by: Wu, Guanchen, et al.
Published: (2026)
by: Wu, Guanchen, et al.
Published: (2026)
Instance-Guided Radar Depth Estimation for 3D Object Detection
by: Lo, Chen-Chou, et al.
Published: (2026)
by: Lo, Chen-Chou, et al.
Published: (2026)
Medical-Knowledge Driven Multiple Instance Learning for Classifying Severe Abdominal Anomalies on Prenatal Ultrasound
by: Liang, Huanwen, et al.
Published: (2025)
by: Liang, Huanwen, et al.
Published: (2025)
MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization
by: Xiao, Zhendong, et al.
Published: (2025)
by: Xiao, Zhendong, et al.
Published: (2025)
Enhancing Visual Token Representations for Video Large Language Models via Training-Free Spatial-Temporal Pooling and Gridding
by: Luo, Bingjun, et al.
Published: (2026)
by: Luo, Bingjun, et al.
Published: (2026)
Towards Effective Data-Free Knowledge Distillation via Diverse Diffusion Augmentation
by: Li, Muquan, et al.
Published: (2024)
by: Li, Muquan, et al.
Published: (2024)
Source-Free Cross-Modal Knowledge Transfer by Unleashing the Potential of Task-Irrelevant Data
by: Zhu, Jinjing, et al.
Published: (2024)
by: Zhu, Jinjing, et al.
Published: (2024)
A Knowledge Noise Mitigation Framework for Knowledge-based Visual Question Answering
by: Liu, Zhiyue, et al.
Published: (2025)
by: Liu, Zhiyue, et al.
Published: (2025)
Depth Map Denoising Network and Lightweight Fusion Network for Enhanced 3D Face Recognition
by: Xu, Ruizhuo, et al.
Published: (2024)
by: Xu, Ruizhuo, et al.
Published: (2024)
Robust MLLM Unlearning via Visual Knowledge Distillation
by: Wang, Yuhang, et al.
Published: (2025)
by: Wang, Yuhang, et al.
Published: (2025)
Hierarchical Knowledge Graphs for Story Understanding in Visual Narratives
by: Chen, Yi-Chun
Published: (2025)
by: Chen, Yi-Chun
Published: (2025)
VIKSER: Visual Knowledge-Driven Self-Reinforcing Reasoning Framework
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
StaR-KVQA: Structured Reasoning Traces for Implicit-Knowledge Visual Question Answering
by: Wen, Zhihao, et al.
Published: (2025)
by: Wen, Zhihao, et al.
Published: (2025)
Knowledge-Intensive Video Generation
by: Wang, Chenxu, et al.
Published: (2026)
by: Wang, Chenxu, et al.
Published: (2026)
SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling
by: Wang, Eileen, et al.
Published: (2024)
by: Wang, Eileen, et al.
Published: (2024)
IVLMap: Instance-Aware Visual Language Grounding for Consumer Robot Navigation
by: Huang, Jiacui, et al.
Published: (2024)
by: Huang, Jiacui, et al.
Published: (2024)
CSS: Overcoming Pose and Scene Challenges in Crowd-Sourced 3D Gaussian Splatting
by: Chen, Runze, et al.
Published: (2024)
by: Chen, Runze, et al.
Published: (2024)
FreeBind: Free Lunch in Unified Multimodal Space via Knowledge Fusion
by: Wang, Zehan, et al.
Published: (2024)
by: Wang, Zehan, et al.
Published: (2024)
Grounded Knowledge-Enhanced Medical Vision-Language Pre-training for Chest X-Ray
by: Deng, Qiao, et al.
Published: (2024)
by: Deng, Qiao, et al.
Published: (2024)
A Knowledge-guided Adversarial Defense for Resisting Malicious Visual Manipulation
by: Zhou, Dawei, et al.
Published: (2025)
by: Zhou, Dawei, et al.
Published: (2025)
Camera-Invariant Meta-Learning Network for Single-Camera-Training Person Re-identification
by: Pei, Jiangbo, et al.
Published: (2024)
by: Pei, Jiangbo, et al.
Published: (2024)
Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning
by: Wu, Ruiqi, et al.
Published: (2024)
by: Wu, Ruiqi, et al.
Published: (2024)
A Knowledge-driven Adaptive Collaboration of LLMs for Enhancing Medical Decision-making
by: Wu, Xiao, et al.
Published: (2025)
by: Wu, Xiao, et al.
Published: (2025)
See in Depth: Training-Free Surgical Scene Segmentation with Monocular Depth Priors
by: Yang, Kunyi, et al.
Published: (2025)
by: Yang, Kunyi, et al.
Published: (2025)
Knowledge-based Visual Question Answer with Multimodal Processing, Retrieval and Filtering
by: Hong, Yuyang, et al.
Published: (2025)
by: Hong, Yuyang, et al.
Published: (2025)
World to Code: Multi-modal Data Generation via Self-Instructed Compositional Captioning and Filtering
by: Wang, Jiacong, et al.
Published: (2024)
by: Wang, Jiacong, et al.
Published: (2024)
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
by: Mei, Guofeng, et al.
Published: (2024)
by: Mei, Guofeng, et al.
Published: (2024)
VGR: Visual Grounded Reasoning
by: Wang, Jiacong, et al.
Published: (2025)
by: Wang, Jiacong, et al.
Published: (2025)
IDMR: Towards Instance-Driven Precise Visual Correspondence in Multimodal Retrieval
by: Liu, Bangwei, et al.
Published: (2025)
by: Liu, Bangwei, et al.
Published: (2025)
Small Scale Data-Free Knowledge Distillation
by: Liu, He, et al.
Published: (2024)
by: Liu, He, et al.
Published: (2024)
KEPT: Knowledge-Enhanced Prediction of Trajectories from Consecutive Driving Frames with Vision-Language Models
by: Wang, Yujin, et al.
Published: (2025)
by: Wang, Yujin, et al.
Published: (2025)
HMID-Net: An Exploration of Masked Image Modeling and Knowledge Distillation in Hyperbolic Space
by: Wang, Changli, et al.
Published: (2025)
by: Wang, Changli, et al.
Published: (2025)
Knowledge Adaptation Network for Few-Shot Class-Incremental Learning
by: Wang, Ye, et al.
Published: (2024)
by: Wang, Ye, et al.
Published: (2024)
Similar Items
-
Structure-Centric Robust Monocular Depth Estimation via Knowledge Distillation
by: Chen, Runze, et al.
Published: (2024) -
Visual Instance-aware Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025) -
KARL: Knowledge-Aware Reasoning and Reinforcement Learning for Knowledge-Intensive Visual Grounding
by: Ma, Xinyu, et al.
Published: (2025) -
No Pose Estimation? No Problem: Pose-Agnostic and Instance-Aware Test-Time Adaptation for Monocular Depth Estimation
by: Sung, Mingyu, et al.
Published: (2025) -
Teeth-SEG: An Efficient Instance Segmentation Framework for Orthodontic Treatment based on Anthropic Prior Knowledge
by: Zou, Bo, et al.
Published: (2024)