GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Jiahao, Wang, Zihan, Li, Xiangyang, Zhu, Xing, Shen, Yujun, Xu, Yinghao, Jiang, Shuqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lookahead Exploration with Neural Radiance Representation for Continuous Vision-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
LookasideVLN: Direction-Aware Aerial Vision-and-Language Navigation
von: Ning, Yuwei, et al.
Veröffentlicht: (2026)
von: Ning, Yuwei, et al.
Veröffentlicht: (2026)
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026)
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026)
MiniVLN: Efficient Vision-and-Language Navigation by Progressive Knowledge Distillation
von: Zhu, Junyou, et al.
Veröffentlicht: (2024)
von: Zhu, Junyou, et al.
Veröffentlicht: (2024)
Efficient-VLN: A Training-Efficient Vision-Language Navigation Model
von: Zheng, Duo, et al.
Veröffentlicht: (2025)
von: Zheng, Duo, et al.
Veröffentlicht: (2025)
FantasyVLN: Unified Multimodal Chain-of-Thought Reasoning for Vision-Language Navigation
von: Zuo, Jing, et al.
Veröffentlicht: (2026)
von: Zuo, Jing, et al.
Veröffentlicht: (2026)
ActiveVLN: Towards Active Exploration via Multi-Turn RL in Vision-and-Language Navigation
von: Zhang, Zekai, et al.
Veröffentlicht: (2025)
von: Zhang, Zekai, et al.
Veröffentlicht: (2025)
JanusVLN: Decoupling Semantics and Spatiality with Dual Implicit Memory for Vision-Language Navigation
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
von: Zeng, Shuang, et al.
Veröffentlicht: (2025)
VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions
von: Su, Hung-Ting, et al.
Veröffentlicht: (2026)
von: Su, Hung-Ting, et al.
Veröffentlicht: (2026)
GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
AgriVLN: Vision-and-Language Navigation for Agricultural Robots
von: Zhao, Xiaobei, et al.
Veröffentlicht: (2025)
von: Zhao, Xiaobei, et al.
Veröffentlicht: (2025)
StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling
von: Wei, Meng, et al.
Veröffentlicht: (2025)
von: Wei, Meng, et al.
Veröffentlicht: (2025)
SpatialFly: Geometry-Guided Representation Alignment for UAV Vision-and-Language Navigation in Urban Environments
von: Jiang, Wen, et al.
Veröffentlicht: (2026)
von: Jiang, Wen, et al.
Veröffentlicht: (2026)
VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning
von: Qi, Zhangyang, et al.
Veröffentlicht: (2025)
von: Qi, Zhangyang, et al.
Veröffentlicht: (2025)
FlexVLN: Flexible Adaptation for Diverse Vision-and-Language Navigation Tasks
von: Zhang, Siqi, et al.
Veröffentlicht: (2025)
von: Zhang, Siqi, et al.
Veröffentlicht: (2025)
UAV-VLN: End-to-End Vision Language guided Navigation for UAVs
von: Saxena, Pranav, et al.
Veröffentlicht: (2025)
von: Saxena, Pranav, et al.
Veröffentlicht: (2025)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
von: Zhao, Baining, et al.
Veröffentlicht: (2026)
von: Zhao, Baining, et al.
Veröffentlicht: (2026)
UnitedVLN: Generalizable Gaussian Splatting for Continuous Vision-Language Navigation
von: Dai, Guangzhao, et al.
Veröffentlicht: (2024)
von: Dai, Guangzhao, et al.
Veröffentlicht: (2024)
VLN-Video: Utilizing Driving Videos for Outdoor Vision-and-Language Navigation
von: Li, Jialu, et al.
Veröffentlicht: (2024)
von: Li, Jialu, et al.
Veröffentlicht: (2024)
Spatial-VLN: Zero-Shot Vision-and-Language Navigation With Explicit Spatial Perception and Exploration
von: Yue, Lu, et al.
Veröffentlicht: (2026)
von: Yue, Lu, et al.
Veröffentlicht: (2026)
Navigation Instruction Generation with BEV Perception and Large Language Models
von: Fan, Sheng, et al.
Veröffentlicht: (2024)
von: Fan, Sheng, et al.
Veröffentlicht: (2024)
SE-VLN: A Self-Evolving Vision-Language Navigation Framework Based on Multimodal Large Language Models
von: Dong, Xiangyu, et al.
Veröffentlicht: (2025)
von: Dong, Xiangyu, et al.
Veröffentlicht: (2025)
Implicit Geometry Representations for Vision-and-Language Navigation from Web Videos
von: Han, Mingfei, et al.
Veröffentlicht: (2026)
von: Han, Mingfei, et al.
Veröffentlicht: (2026)
GMF-Drive: Gated Mamba Fusion with Spatial-Aware BEV Representation for End-to-End Autonomous Driving
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
von: Lyu, Kailin, et al.
Veröffentlicht: (2026)
von: Lyu, Kailin, et al.
Veröffentlicht: (2026)
VLN-MME: Diagnosing MLLMs as Language-guided Visual Navigation agents
von: Zhao, Xunyi, et al.
Veröffentlicht: (2025)
von: Zhao, Xunyi, et al.
Veröffentlicht: (2025)
Volumetric Environment Representation for Vision-Language Navigation
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
InstanceBEV: Unifying Instance and BEV Representation for 3D Panoptic Segmentation
von: Li, Feng, et al.
Veröffentlicht: (2025)
von: Li, Feng, et al.
Veröffentlicht: (2025)
TiGDistill-BEV: Multi-view BEV 3D Object Detection via Target Inner-Geometry Learning Distillation
von: Xu, Shaoqing, et al.
Veröffentlicht: (2024)
von: Xu, Shaoqing, et al.
Veröffentlicht: (2024)
Multi-Scale Gaussian-Language Map for Zero-shot Embodied Navigation and Reasoning
von: Zhang, Sixian, et al.
Veröffentlicht: (2026)
von: Zhang, Sixian, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Gaussian Map for Vision-Language Navigation
von: Gao, Jianzhe, et al.
Veröffentlicht: (2026)
von: Gao, Jianzhe, et al.
Veröffentlicht: (2026)
FLARE: Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Views
von: Zhang, Shangzhan, et al.
Veröffentlicht: (2025)
von: Zhang, Shangzhan, et al.
Veröffentlicht: (2025)
Interspatial Attention for Efficient 4D Human Video Generation
von: Shao, Ruizhi, et al.
Veröffentlicht: (2025)
von: Shao, Ruizhi, et al.
Veröffentlicht: (2025)
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
von: Li, Bohan, et al.
Veröffentlicht: (2023)
von: Li, Bohan, et al.
Veröffentlicht: (2023)
LSSInst: Improving Geometric Modeling in LSS-Based BEV Perception with Instance Representation
von: Ma, Weijie, et al.
Veröffentlicht: (2024)
von: Ma, Weijie, et al.
Veröffentlicht: (2024)
AdaVLN: Towards Visual Language Navigation in Continuous Indoor Environments with Moving Humans
von: Loh, Dillon, et al.
Veröffentlicht: (2024)
von: Loh, Dillon, et al.
Veröffentlicht: (2024)
ViSA-Enhanced Aerial VLN: A Visual-Spatial Reasoning Enhanced Framework for Aerial Vision-Language Navigation
von: Tong, Haoyu, et al.
Veröffentlicht: (2026)
von: Tong, Haoyu, et al.
Veröffentlicht: (2026)
Reconstruction Matters: Learning Geometry-Aligned BEV Representation through 3D Gaussian Splatting
von: Lu, Yiren, et al.
Veröffentlicht: (2026)
von: Lu, Yiren, et al.
Veröffentlicht: (2026)
GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation
von: Xu, Yinghao, et al.
Veröffentlicht: (2024)
von: Xu, Yinghao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Lookahead Exploration with Neural Radiance Representation for Continuous Vision-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024) -
Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation
von: Wang, Zihan, et al.
Veröffentlicht: (2024) -
LookasideVLN: Direction-Aware Aerial Vision-and-Language Navigation
von: Ning, Yuwei, et al.
Veröffentlicht: (2026) -
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
von: Guo, Wenxuan, et al.
Veröffentlicht: (2026) -
MiniVLN: Efficient Vision-and-Language Navigation by Progressive Knowledge Distillation
von: Zhu, Junyou, et al.
Veröffentlicht: (2024)