DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ji, Zihe, Lin, Huangxuan, Gao, Yue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DyGeoVLN: Infusing Dynamic Geometry Foundation Model into Vision-Language Navigation
von: Liu, Xiangchen, et al.
Veröffentlicht: (2026)
von: Liu, Xiangchen, et al.
Veröffentlicht: (2026)
AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models
von: Heo, Hyeongjun, et al.
Veröffentlicht: (2026)
von: Heo, Hyeongjun, et al.
Veröffentlicht: (2026)
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2024)
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2024)
LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision Language Navigation in Continuous Environments
von: Ding, Hongyu, et al.
Veröffentlicht: (2025)
von: Ding, Hongyu, et al.
Veröffentlicht: (2025)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
von: Lyu, Kailin, et al.
Veröffentlicht: (2026)
von: Lyu, Kailin, et al.
Veröffentlicht: (2026)
VLN-Game: Vision-Language Equilibrium Search for Zero-Shot Semantic Navigation
von: Yu, Bangguo, et al.
Veröffentlicht: (2024)
von: Yu, Bangguo, et al.
Veröffentlicht: (2024)
CATNAV: Cached Vision-Language Traversability for Efficient Zero-Shot Robot Navigation
von: Potnis, Aditya, et al.
Veröffentlicht: (2026)
von: Potnis, Aditya, et al.
Veröffentlicht: (2026)
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)
T2Nav Algebraic Topology Aware Temporal Graph Memory and Loop Detection for ZeroShot Visual Navigation
von: D., Quang-Anh N., et al.
Veröffentlicht: (2026)
von: D., Quang-Anh N., et al.
Veröffentlicht: (2026)
OnFly: Onboard Zero-Shot Aerial Vision-Language Navigation toward Safety and Efficiency
von: Zheng, Guiyong, et al.
Veröffentlicht: (2026)
von: Zheng, Guiyong, et al.
Veröffentlicht: (2026)
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
von: Chen, Kehan, et al.
Veröffentlicht: (2024)
von: Chen, Kehan, et al.
Veröffentlicht: (2024)
OpenFMNav: Towards Open-Set Zero-Shot Object Navigation via Vision-Language Foundation Models
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Kuang, Yuxuan, et al.
Veröffentlicht: (2024)
VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
von: Xu, Runsen, et al.
Veröffentlicht: (2024)
von: Xu, Runsen, et al.
Veröffentlicht: (2024)
Exploring Bottlenecks in VLM-LLM Navigation: How 3D Scene Understanding Capability Impacts Zero-Shot VLN
von: Xia, Ziyi, et al.
Veröffentlicht: (2026)
von: Xia, Ziyi, et al.
Veröffentlicht: (2026)
TriHelper: Zero-Shot Object Navigation with Dynamic Assistance
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
History-Augmented Vision-Language Models for Frontier-Based Zero-Shot Object Navigation
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
USS-Nav: Unified Spatio-Semantic Scene Graph for Lightweight UAV Zero-Shot Object Navigation
von: Gai, Weiqi, et al.
Veröffentlicht: (2026)
von: Gai, Weiqi, et al.
Veröffentlicht: (2026)
NORM-Nav: Zero-Shot Mobile Robot Navigation with Natural Language Behavioral Constraints
von: Huo, Dongjie, et al.
Veröffentlicht: (2026)
von: Huo, Dongjie, et al.
Veröffentlicht: (2026)
Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
von: Li, Zerui, et al.
Veröffentlicht: (2025)
von: Li, Zerui, et al.
Veröffentlicht: (2025)
PanoNav: Mapless Zero-Shot Object Navigation with Panoramic Scene Parsing and Dynamic Memory
von: Jin, Qunchao, et al.
Veröffentlicht: (2025)
von: Jin, Qunchao, et al.
Veröffentlicht: (2025)
GoalVLM: VLM-driven Object Goal Navigation for Multi-Agent System
von: James, MoniJesu, et al.
Veröffentlicht: (2026)
von: James, MoniJesu, et al.
Veröffentlicht: (2026)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
VLM-Empowered Multi-Mode System for Efficient and Safe Planetary Navigation
von: Cheng, Sinuo, et al.
Veröffentlicht: (2025)
von: Cheng, Sinuo, et al.
Veröffentlicht: (2025)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
von: Wang, Yunheng, et al.
Veröffentlicht: (2025)
von: Wang, Yunheng, et al.
Veröffentlicht: (2025)
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024)
von: Cheng, An-Chieh, et al.
Veröffentlicht: (2024)
Learning Motion Skills with Adaptive Assistive Curriculum Force in Humanoid Robots
von: Cao, Zhanxiang, et al.
Veröffentlicht: (2025)
von: Cao, Zhanxiang, et al.
Veröffentlicht: (2025)
Zero-shot Object Navigation with Vision-Language Models Reasoning
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
SwarmVLM: VLM-Guided Impedance Control for Autonomous Navigation of Heterogeneous Robots in Dynamic Warehousing
von: Zafar, Malaika, et al.
Veröffentlicht: (2025)
von: Zafar, Malaika, et al.
Veröffentlicht: (2025)
Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
NavDreamer: Video Models as Zero-Shot 3D Navigators
von: Huang, Xijie, et al.
Veröffentlicht: (2026)
von: Huang, Xijie, et al.
Veröffentlicht: (2026)
CL-CoTNav: Closed-Loop Hierarchical Chain-of-Thought for Zero-Shot Object-Goal Navigation with Vision-Language Models
von: Cai, Yuxin, et al.
Veröffentlicht: (2025)
von: Cai, Yuxin, et al.
Veröffentlicht: (2025)
VLM-RRT: Vision Language Model Guided RRT Search for Autonomous UAV Navigation
von: Ye, Jianlin, et al.
Veröffentlicht: (2025)
von: Ye, Jianlin, et al.
Veröffentlicht: (2025)
VLM-Social-Nav: Socially Aware Robot Navigation through Scoring using Vision-Language Models
von: Song, Daeun, et al.
Veröffentlicht: (2024)
von: Song, Daeun, et al.
Veröffentlicht: (2024)
SFCo-Nav: Efficient Zero-Shot Visual Language Navigation via Collaboration of Slow LLM and Fast Attributed Graph Alignment
von: Xiong, Chaoran, et al.
Veröffentlicht: (2026)
von: Xiong, Chaoran, et al.
Veröffentlicht: (2026)
VLM-DEWM: Dynamic External World Model for Verifiable and Resilient Vision-Language Planning in Manufacturing
von: Tang, Guoqin, et al.
Veröffentlicht: (2026)
von: Tang, Guoqin, et al.
Veröffentlicht: (2026)
TagaVLM: Topology-Aware Global Action Reasoning for Vision-Language Navigation
von: Liu, Jiaxing, et al.
Veröffentlicht: (2026)
von: Liu, Jiaxing, et al.
Veröffentlicht: (2026)
Multi-Floor Zero-Shot Object Navigation Policy
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2024)
DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments
von: Ma, Ji, et al.
Veröffentlicht: (2024)
von: Ma, Ji, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DyGeoVLN: Infusing Dynamic Geometry Foundation Model into Vision-Language Navigation
von: Liu, Xiangchen, et al.
Veröffentlicht: (2026) -
AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models
von: Heo, Hyeongjun, et al.
Veröffentlicht: (2026) -
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025) -
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation
von: Zhang, Jiazhao, et al.
Veröffentlicht: (2024) -
LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision Language Navigation in Continuous Environments
von: Ding, Hongyu, et al.
Veröffentlicht: (2025)