SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Xiangyu, Li, Zerui, Lyu, Wenqi, Xia, Jiatong, Dayoub, Feras, Qiao, Yanyuan, Wu, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation
by: Lyu, Wenqi, et al.
Published: (2025)
by: Lyu, Wenqi, et al.
Published: (2025)
Embodied Domain Adaptation for Object Detection
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
by: Qiao, Yanyuan, et al.
Published: (2024)
by: Qiao, Yanyuan, et al.
Published: (2024)
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
by: Li, Zerui, et al.
Published: (2025)
by: Li, Zerui, et al.
Published: (2025)
Improving Online Source-free Domain Adaptation for Object Detection by Unsupervised Data Acquisition
by: Shi, Xiangyu, et al.
Published: (2023)
by: Shi, Xiangyu, et al.
Published: (2023)
Enhancing Embodied Object Detection through Language-Image Pre-training and Implicit Object Memory
by: Chapman, Nicolas Harvey, et al.
Published: (2024)
by: Chapman, Nicolas Harvey, et al.
Published: (2024)
QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries
by: Chapman, Nicolas Harvey, et al.
Published: (2025)
by: Chapman, Nicolas Harvey, et al.
Published: (2025)
SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
Effective Tuning Strategies for Generalist Robot Manipulation Policies
by: Zhang, Wenbo, et al.
Published: (2024)
by: Zhang, Wenbo, et al.
Published: (2024)
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
by: Wang, Wenze, et al.
Published: (2026)
by: Wang, Wenze, et al.
Published: (2026)
Hybrid Navigation Acceptability and Safety
by: Clement, Benoit, et al.
Published: (2024)
by: Clement, Benoit, et al.
Published: (2024)
COSMO: Combination of Selective Memorization for Low-cost Vision-and-Language Navigation
by: Zhang, Siqi, et al.
Published: (2025)
by: Zhang, Siqi, et al.
Published: (2025)
Skill-Nav: Enhanced Navigation with Versatile Quadrupedal Locomotion via Waypoint Interface
by: Wang, Dewei, et al.
Published: (2025)
by: Wang, Dewei, et al.
Published: (2025)
FlexVLN: Flexible Adaptation for Diverse Vision-and-Language Navigation Tasks
by: Zhang, Siqi, et al.
Published: (2025)
by: Zhang, Siqi, et al.
Published: (2025)
Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for Robotics
by: Abou-Chakra, Jad, et al.
Published: (2024)
by: Abou-Chakra, Jad, et al.
Published: (2024)
LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision Language Navigation in Continuous Environments
by: Ding, Hongyu, et al.
Published: (2025)
by: Ding, Hongyu, et al.
Published: (2025)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
by: Hosseinzadeh, Mehdi, et al.
Published: (2026)
by: Hosseinzadeh, Mehdi, et al.
Published: (2026)
Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams
by: Holden, Lachlan, et al.
Published: (2026)
by: Holden, Lachlan, et al.
Published: (2026)
VLN-Game: Vision-Language Equilibrium Search for Zero-Shot Semantic Navigation
by: Yu, Bangguo, et al.
Published: (2024)
by: Yu, Bangguo, et al.
Published: (2024)
CATNAV: Cached Vision-Language Traversability for Efficient Zero-Shot Robot Navigation
by: Potnis, Aditya, et al.
Published: (2026)
by: Potnis, Aditya, et al.
Published: (2026)
AARK: An Open Toolkit for Autonomous Racing Research
by: Bockman, James, et al.
Published: (2024)
by: Bockman, James, et al.
Published: (2024)
To Ask or Not to Ask? Detecting Absence of Information in Vision and Language Navigation
by: Abraham, Savitha Sam, et al.
Published: (2024)
by: Abraham, Savitha Sam, et al.
Published: (2024)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
by: Lyu, Kailin, et al.
Published: (2026)
by: Lyu, Kailin, et al.
Published: (2026)
NORM-Nav: Zero-Shot Mobile Robot Navigation with Natural Language Behavioral Constraints
by: Huo, Dongjie, et al.
Published: (2026)
by: Huo, Dongjie, et al.
Published: (2026)
Predictive and adaptive maps for long-term visual navigation in changing environments
by: Halodova, Lucie, et al.
Published: (2026)
by: Halodova, Lucie, et al.
Published: (2026)
WayEx: Waypoint Exploration using a Single Demonstration
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals
by: Podgorski, Stefan, et al.
Published: (2025)
by: Podgorski, Stefan, et al.
Published: (2025)
OnFly: Onboard Zero-Shot Aerial Vision-Language Navigation toward Safety and Efficiency
by: Zheng, Guiyong, et al.
Published: (2026)
by: Zheng, Guiyong, et al.
Published: (2026)
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
by: Chen, Kehan, et al.
Published: (2024)
by: Chen, Kehan, et al.
Published: (2024)
Boosting Zero-Shot VLN via Abstract Obstacle Map-Based Waypoint Prediction with TopoGraph-and-VisitInfo-Aware Prompting
by: Li, Boqi, et al.
Published: (2025)
by: Li, Boqi, et al.
Published: (2025)
Beyond Waypoints: Dual-Heatmap Grounding for Cross-Embodiment Semantic Navigation
by: Yun, Kaijie, et al.
Published: (2026)
by: Yun, Kaijie, et al.
Published: (2026)
History-Augmented Vision-Language Models for Frontier-Based Zero-Shot Object Navigation
by: Habibpour, Mobin, et al.
Published: (2025)
by: Habibpour, Mobin, et al.
Published: (2025)
Neural Implicit Action Fields: From Discrete Waypoints to Continuous Functions for Vision-Language-Action Models
by: Liu, Haoyun, et al.
Published: (2026)
by: Liu, Haoyun, et al.
Published: (2026)
Lang2Morph: Language-Driven Morphological Design of Robotic Hands
by: Qiao, Yanyuan, et al.
Published: (2025)
by: Qiao, Yanyuan, et al.
Published: (2025)
One Agent to Guide Them All: Empowering MLLMs for Vision-and-Language Navigation via Explicit World Representation
by: Li, Zerui, et al.
Published: (2026)
by: Li, Zerui, et al.
Published: (2026)
FocusNav: Spatial Selective Attention with Waypoint Guidance for Humanoid Local Navigation
by: Zhang, Yang, et al.
Published: (2026)
by: Zhang, Yang, et al.
Published: (2026)
DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory
by: Ji, Zihe, et al.
Published: (2025)
by: Ji, Zihe, et al.
Published: (2025)
RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
by: Garg, Sourav, et al.
Published: (2024)
by: Garg, Sourav, et al.
Published: (2024)
Similar Items
-
Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025) -
BadNAVer: Exploring Jailbreak Attacks On Vision-and-Language Navigation
by: Lyu, Wenqi, et al.
Published: (2025) -
Embodied Domain Adaptation for Object Detection
by: Shi, Xiangyu, et al.
Published: (2025) -
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
by: Qiao, Yanyuan, et al.
Published: (2024) -
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
by: Li, Zerui, et al.
Published: (2025)