REGNav: Room Expert Guided Image-Goal Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Pengna, Wu, Kangyi, Fu, Jingwen, Zhou, Sanping |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching
by: Liu, Yuhan, et al.
Published: (2025)
by: Liu, Yuhan, et al.
Published: (2025)
Camera-aware Label Refinement for Unsupervised Person Re-identification
by: Li, Pengna, et al.
Published: (2024)
by: Li, Pengna, et al.
Published: (2024)
RSRNav: Reasoning Spatial Relationship for Image-Goal Navigation
by: Qin, Zheng, et al.
Published: (2025)
by: Qin, Zheng, et al.
Published: (2025)
Dual-Anchoring: Addressing State Drift in Vision-Language Navigation
by: Wu, Kangyi, et al.
Published: (2026)
by: Wu, Kangyi, et al.
Published: (2026)
SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation
by: Li, Pengna, et al.
Published: (2026)
by: Li, Pengna, et al.
Published: (2026)
DAMap: Distance-aware MapNet for High Quality HD Map Construction
by: Dong, Jinpeng, et al.
Published: (2025)
by: Dong, Jinpeng, et al.
Published: (2025)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
by: Shen, Yanqing, et al.
Published: (2025)
by: Shen, Yanqing, et al.
Published: (2025)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
by: Lyu, Kailin, et al.
Published: (2026)
by: Lyu, Kailin, et al.
Published: (2026)
Instance-aware Exploration-Verification-Exploitation for Instance ImageGoal Navigation
by: Lei, Xiaohan, et al.
Published: (2024)
by: Lei, Xiaohan, et al.
Published: (2024)
Instruction-as-State: Environment-Guided and State-Conditioned Semantic Understanding for Embodied Navigation
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
PMT: Progressive Mean Teacher via Exploring Temporal Consistency for Semi-Supervised Medical Image Segmentation
by: Gao, Ning, et al.
Published: (2024)
by: Gao, Ning, et al.
Published: (2024)
PR-DETR: Injecting Position and Relation Prior for Dense Video Captioning
by: Li, Yizhe, et al.
Published: (2025)
by: Li, Yizhe, et al.
Published: (2025)
PIG-Nav: Key Insights for Pretrained Image Goal Navigation Models
by: Wan, Jiansong, et al.
Published: (2025)
by: Wan, Jiansong, et al.
Published: (2025)
Learning Goal-Oriented Vision-and-Language Navigation with Self-Improving Demonstrations at Scale
by: Li, Songze, et al.
Published: (2025)
by: Li, Songze, et al.
Published: (2025)
Diversifying Query: Region-Guided Transformer for Temporal Sentence Grounding
by: Sun, Xiaolong, et al.
Published: (2024)
by: Sun, Xiaolong, et al.
Published: (2024)
UniGoal: Towards Universal Zero-shot Goal-oriented Navigation
by: Yin, Hang, et al.
Published: (2025)
by: Yin, Hang, et al.
Published: (2025)
Towards Generalizable Multi-Object Tracking
by: Qin, Zheng, et al.
Published: (2024)
by: Qin, Zheng, et al.
Published: (2024)
LayoutRAG: Retrieval-Augmented Model for Content-agnostic Conditional Layout Generation
by: Wu, Yuxuan, et al.
Published: (2025)
by: Wu, Yuxuan, et al.
Published: (2025)
Action Hints: Semantic Typicality and Context Uniqueness for Generalizable Skeleton-based Video Anomaly Detection
by: Tang, Canhui, et al.
Published: (2025)
by: Tang, Canhui, et al.
Published: (2025)
Advancing Pre-trained Teacher: Towards Robust Feature Discrepancy for Anomaly Detection
by: Tang, Canhui, et al.
Published: (2024)
by: Tang, Canhui, et al.
Published: (2024)
DiffusionAgent: Navigating Expert Models for Agentic Image Generation
by: Qin, Jie, et al.
Published: (2024)
by: Qin, Jie, et al.
Published: (2024)
Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation
by: Bao, Muyi, et al.
Published: (2026)
by: Bao, Muyi, et al.
Published: (2026)
LiteVLoc: Map-Lite Visual Localization for Image Goal Navigation
by: Jiao, Jianhao, et al.
Published: (2024)
by: Jiao, Jianhao, et al.
Published: (2024)
SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation
by: Lu, Yujie, et al.
Published: (2026)
by: Lu, Yujie, et al.
Published: (2026)
The Essence of Balance for Self-Improving Agents in Vision-and-Language Navigation
by: Liu, Zhen, et al.
Published: (2026)
by: Liu, Zhen, et al.
Published: (2026)
From Mapping to Composing: A Two-Stage Framework for Zero-shot Composed Image Retrieval
by: Wang, Yabing, et al.
Published: (2025)
by: Wang, Yabing, et al.
Published: (2025)
DR.Experts: Differential Refinement of Distortion-Aware Experts for Blind Image Quality Assessment
by: Fu, Bohan, et al.
Published: (2026)
by: Fu, Bohan, et al.
Published: (2026)
Hierarchical Scoring with 3D Gaussian Splatting for Instance Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2025)
by: Deng, Yijie, et al.
Published: (2025)
AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2026)
by: Deng, Yijie, et al.
Published: (2026)
UniLayDiff: A Unified Diffusion Transformer for Content-Aware Layout Generation
by: Liu, Zeyang, et al.
Published: (2025)
by: Liu, Zeyang, et al.
Published: (2025)
Labeled-to-Unlabeled Distribution Alignment for Partially-Supervised Multi-Organ Medical Image Segmentation
by: Jiang, Xixi, et al.
Published: (2024)
by: Jiang, Xixi, et al.
Published: (2024)
uLayout: Unified Room Layout Estimation for Perspective and Panoramic Images
by: Lee, Jonathan, et al.
Published: (2025)
by: Lee, Jonathan, et al.
Published: (2025)
FreqGRL: Suppressing Low-Frequency Bias and Mining High-Frequency Knowledge for Cross-Domain Few-Shot Learning
by: Hui, Siqi, et al.
Published: (2025)
by: Hui, Siqi, et al.
Published: (2025)
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
by: Huang, Yuhao, et al.
Published: (2023)
by: Huang, Yuhao, et al.
Published: (2023)
Boosting Semi-Supervised Temporal Action Localization by Learning from Non-Target Classes
by: Xia, Kun, et al.
Published: (2024)
by: Xia, Kun, et al.
Published: (2024)
Enhancing Mamba Decoder with Bidirectional Interaction in Multi-Task Dense Prediction
by: Cao, Mang, et al.
Published: (2025)
by: Cao, Mang, et al.
Published: (2025)
Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
by: Li, Badi, et al.
Published: (2025)
by: Li, Badi, et al.
Published: (2025)
Statistical Characteristic-Guided Denoising for Rapid High-Resolution Transmission Electron Microscopy Imaging
by: Li, Hesong, et al.
Published: (2026)
by: Li, Hesong, et al.
Published: (2026)
WorldMesh: Generating Navigable Multi-Room 3D Scenes via Mesh-Conditioned Image Diffusion
by: Schneider, Manuel-Andreas, et al.
Published: (2026)
by: Schneider, Manuel-Andreas, et al.
Published: (2026)
FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
by: Su, Haisheng, et al.
Published: (2025)
by: Su, Haisheng, et al.
Published: (2025)
Similar Items
-
Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching
by: Liu, Yuhan, et al.
Published: (2025) -
Camera-aware Label Refinement for Unsupervised Person Re-identification
by: Li, Pengna, et al.
Published: (2024) -
RSRNav: Reasoning Spatial Relationship for Image-Goal Navigation
by: Qin, Zheng, et al.
Published: (2025) -
Dual-Anchoring: Addressing State Drift in Vision-Language Navigation
by: Wu, Kangyi, et al.
Published: (2026) -
SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation
by: Li, Pengna, et al.
Published: (2026)