Fast-SmartWay: Panoramic-Free End-to-End Zero-Shot Vision-and-Language Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shi, Xiangyu, Li, Zerui, Qiao, Yanyuan, Wu, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
von: Qiao, Yanyuan, et al.
Veröffentlicht: (2024)
von: Qiao, Yanyuan, et al.
Veröffentlicht: (2024)
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
von: Li, Zerui, et al.
Veröffentlicht: (2025)
von: Li, Zerui, et al.
Veröffentlicht: (2025)
SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)
COSMO: Combination of Selective Memorization for Low-cost Vision-and-Language Navigation
von: Zhang, Siqi, et al.
Veröffentlicht: (2025)
von: Zhang, Siqi, et al.
Veröffentlicht: (2025)
UAV-VLN: End-to-End Vision Language guided Navigation for UAVs
von: Saxena, Pranav, et al.
Veröffentlicht: (2025)
von: Saxena, Pranav, et al.
Veröffentlicht: (2025)
PanoNav: Mapless Zero-Shot Object Navigation with Panoramic Scene Parsing and Dynamic Memory
von: Jin, Qunchao, et al.
Veröffentlicht: (2025)
von: Jin, Qunchao, et al.
Veröffentlicht: (2025)
FlexVLN: Flexible Adaptation for Diverse Vision-and-Language Navigation Tasks
von: Zhang, Siqi, et al.
Veröffentlicht: (2025)
von: Zhang, Siqi, et al.
Veröffentlicht: (2025)
Embodied Domain Adaptation for Object Detection
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
MonoDream: Monocular Vision-Language Navigation with Panoramic Dreaming
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
von: Goetting, Dylan, et al.
Veröffentlicht: (2024)
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
von: Chen, Kehan, et al.
Veröffentlicht: (2024)
von: Chen, Kehan, et al.
Veröffentlicht: (2024)
Senna: Bridging Large Vision-Language Models and End-to-End Autonomous Driving
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
von: Jiang, Bo, et al.
Veröffentlicht: (2024)
AerialVLA: A Vision-Language-Action Model for UAV Navigation via Minimalist End-to-End Control
von: Xu, Peng, et al.
Veröffentlicht: (2026)
von: Xu, Peng, et al.
Veröffentlicht: (2026)
Temporal Sampling Frequency Matters: A Capacity-Aware Study of End-to-End Driving Trajectory Prediction
von: Liu, Yumao, et al.
Veröffentlicht: (2026)
von: Liu, Yumao, et al.
Veröffentlicht: (2026)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
von: Lyu, Kailin, et al.
Veröffentlicht: (2026)
von: Lyu, Kailin, et al.
Veröffentlicht: (2026)
Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation
von: Zheng, Wanrong, et al.
Veröffentlicht: (2026)
von: Zheng, Wanrong, et al.
Veröffentlicht: (2026)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
Exploring the Causality of End-to-End Autonomous Driving
von: Li, Jiankun, et al.
Veröffentlicht: (2024)
von: Li, Jiankun, et al.
Veröffentlicht: (2024)
VLNVerse: A Benchmark for Vision-Language Navigation with Versatile, Embodied, Realistic Simulation and Evaluation
von: Lin, Sihao, et al.
Veröffentlicht: (2025)
von: Lin, Sihao, et al.
Veröffentlicht: (2025)
Action Images: End-to-End Policy Learning via Multiview Video Generation
von: Zhen, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhen, Haoyu, et al.
Veröffentlicht: (2026)
Bench2FreeAD: A Benchmark for Vision-based End-to-end Navigation in Unstructured Robotic Environments
von: Peng, Yuhang, et al.
Veröffentlicht: (2025)
von: Peng, Yuhang, et al.
Veröffentlicht: (2025)
One-Shot Real-to-Sim via End-to-End Differentiable Simulation and Rendering
von: Zhu, Yifan, et al.
Veröffentlicht: (2024)
von: Zhu, Yifan, et al.
Veröffentlicht: (2024)
AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
von: Huang, Wenhui, et al.
Veröffentlicht: (2026)
SEAL: Vision-Language Model-Based Safe End-to-End Cooperative Autonomous Driving with Adaptive Long-Tail Modeling
von: You, Junwei, et al.
Veröffentlicht: (2025)
von: You, Junwei, et al.
Veröffentlicht: (2025)
PanoGen++: Domain-Adapted Text-Guided Panoramic Environment Generation for Vision-and-Language Navigation
von: Wang, Sen, et al.
Veröffentlicht: (2025)
von: Wang, Sen, et al.
Veröffentlicht: (2025)
Does Peer Observation Help? Vision-Sharing Collaboration for Vision-Language Navigation
von: Jin, Qunchao, et al.
Veröffentlicht: (2026)
von: Jin, Qunchao, et al.
Veröffentlicht: (2026)
ZTRS: Zero-Imitation End-to-end Autonomous Driving with Trajectory Scoring
von: Li, Zhenxin, et al.
Veröffentlicht: (2025)
von: Li, Zhenxin, et al.
Veröffentlicht: (2025)
NavigateDiff: Visual Predictors are Zero-Shot Navigation Assistants
von: Qin, Yiran, et al.
Veröffentlicht: (2025)
von: Qin, Yiran, et al.
Veröffentlicht: (2025)
Poutine: Vision-Language-Trajectory Pre-Training and Reinforcement Learning Post-Training Enable Robust End-to-End Autonomous Driving
von: Rowe, Luke, et al.
Veröffentlicht: (2025)
von: Rowe, Luke, et al.
Veröffentlicht: (2025)
RAD: Training an End-to-End Driving Policy via Large-Scale 3DGS-based Reinforcement Learning
von: Gao, Hao, et al.
Veröffentlicht: (2025)
von: Gao, Hao, et al.
Veröffentlicht: (2025)
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
von: Li, Yongkang, et al.
Veröffentlicht: (2025)
Zero-Shot 3D Visual Grounding from Vision-Language Models
von: Li, Rong, et al.
Veröffentlicht: (2025)
von: Li, Rong, et al.
Veröffentlicht: (2025)
Towards Realistic UAV Vision-Language Navigation: Platform, Benchmark, and Methodology
von: Wang, Xiangyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiangyu, et al.
Veröffentlicht: (2024)
DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving
von: Zhang, Lingjun, et al.
Veröffentlicht: (2026)
von: Zhang, Lingjun, et al.
Veröffentlicht: (2026)
From Representational Complementarity to Dual Systems: Synergizing VLM and Vision-Only Backbones for End-to-End Driving
von: Ang, Sining, et al.
Veröffentlicht: (2026)
von: Ang, Sining, et al.
Veröffentlicht: (2026)
2nd Place Solution for CVPR2024 E2E Challenge: End-to-End Autonomous Driving Using Vision Language Model
von: Guo, Zilong, et al.
Veröffentlicht: (2025)
von: Guo, Zilong, et al.
Veröffentlicht: (2025)
HAD: Combining Hierarchical Diffusion with Metric-Decoupled RL for End-to-End Driving
von: Yao, Wenhao, et al.
Veröffentlicht: (2026)
von: Yao, Wenhao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025) -
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
von: Qiao, Yanyuan, et al.
Veröffentlicht: (2024) -
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
von: Li, Zerui, et al.
Veröffentlicht: (2025) -
SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026) -
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
von: Zhang, Jiwen, et al.
Veröffentlicht: (2026)