VLN-Game: Vision-Language Equilibrium Search for Zero-Shot Semantic Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Bangguo, Liu, Yuzhen, Han, Lei, Kasaei, Hamidreza, Li, Tingguang, Cao, Ming |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PANav: Toward Privacy-Aware Robot Navigation via Vision-Language Models
by: Yu, Bangguo, et al.
Published: (2024)
by: Yu, Bangguo, et al.
Published: (2024)
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
by: Yu, Bangguo, et al.
Published: (2023)
by: Yu, Bangguo, et al.
Published: (2023)
MAG-Nav: Language-Driven Object Navigation Leveraging Memory-Reserved Active Grounding
by: Zhang, Weifan, et al.
Published: (2025)
by: Zhang, Weifan, et al.
Published: (2025)
Single-Shot 6DoF Pose and 3D Size Estimation for Robotic Strawberry Harvesting
by: Li, Lun, et al.
Published: (2024)
by: Li, Lun, et al.
Published: (2024)
VITAL: Interactive Few-Shot Imitation Learning via Visual Human-in-the-Loop Corrections
by: Kasaei, Hamidreza, et al.
Published: (2024)
by: Kasaei, Hamidreza, et al.
Published: (2024)
Towards Open-World Grasping with Large Vision-Language Models
by: Tziafas, Georgios, et al.
Published: (2024)
by: Tziafas, Georgios, et al.
Published: (2024)
Harnessing the Synergy between Pushing, Grasping, and Throwing to Enhance Object Manipulation in Cluttered Scenarios
by: Kasaei, Hamidreza, et al.
Published: (2024)
by: Kasaei, Hamidreza, et al.
Published: (2024)
Lifelong Ensemble Learning based on Multiple Representations for Few-Shot Object Recognition
by: Kasaei, Hamidreza, et al.
Published: (2022)
by: Kasaei, Hamidreza, et al.
Published: (2022)
AgentVLN: Towards Agentic Vision-and-Language Navigation
by: Xin, Zihao, et al.
Published: (2026)
by: Xin, Zihao, et al.
Published: (2026)
OpenVLN: Open-world Aerial Vision-Language Navigation
by: Lin, Peican, et al.
Published: (2025)
by: Lin, Peican, et al.
Published: (2025)
Lifelong Robot Library Learning: Bootstrapping Composable and Generalizable Skills for Embodied Control with Language Models
by: Tziafas, Georgios, et al.
Published: (2024)
by: Tziafas, Georgios, et al.
Published: (2024)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
by: Lyu, Kailin, et al.
Published: (2026)
by: Lyu, Kailin, et al.
Published: (2026)
Enhanced View Planning for Robotic Harvesting: Tackling Occlusions with Imitation Learning
by: Li, Lun, et al.
Published: (2025)
by: Li, Lun, et al.
Published: (2025)
T-araVLN: Translator for Agricultural Robotic Agents on Vision-and-Language Navigation
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
DecoVLN: Decoupling Observation, Reasoning, and Correction for Vision-and-Language Navigation
by: Xin, Zihao, et al.
Published: (2026)
by: Xin, Zihao, et al.
Published: (2026)
DV-VLN: Dual Verification for Reliable LLM-Based Vision-and-Language Navigation
by: Li, Zijun, et al.
Published: (2026)
by: Li, Zijun, et al.
Published: (2026)
LM-MCVT: A Lightweight Multi-modal Multi-view Convolutional-Vision Transformer Approach for 3D Object Recognition
by: Xiong, Songsong, et al.
Published: (2025)
by: Xiong, Songsong, et al.
Published: (2025)
VLM-TDP: VLM-guided Trajectory-conditioned Diffusion Policy for Robust Long-Horizon Manipulation
by: Huang, Kefeng, et al.
Published: (2025)
by: Huang, Kefeng, et al.
Published: (2025)
VLN-Zero: Rapid Exploration and Cache-Enabled Neurosymbolic Vision-Language Planning for Zero-Shot Transfer in Robot Navigation
by: Bhatt, Neel P., et al.
Published: (2025)
by: Bhatt, Neel P., et al.
Published: (2025)
AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
by: Guo, Wenxuan, et al.
Published: (2026)
by: Guo, Wenxuan, et al.
Published: (2026)
MDE-AgriVLN: Agricultural Vision-and-Language Navigation with Monocular Depth Estimation
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
AgriVLN: Vision-and-Language Navigation for Agricultural Robots
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
Learning Dual-Arm Push and Grasp Synergy in Dense Clutter
by: Wang, Yongliang, et al.
Published: (2024)
by: Wang, Yongliang, et al.
Published: (2024)
HMT-Grasp: A Hybrid Mamba-Transformer Approach for Robot Grasping in Cluttered Environments
by: Xiong, Songsong, et al.
Published: (2024)
by: Xiong, Songsong, et al.
Published: (2024)
Fast Trajectory Planner with a Reinforcement Learning-based Controller for Robotic Manipulators
by: Wang, Yongliang, et al.
Published: (2025)
by: Wang, Yongliang, et al.
Published: (2025)
LiveVLN: Breaking the Stop-and-Go Loop in Vision-Language Navigation
by: Wang, Xiangchen, et al.
Published: (2026)
by: Wang, Xiangchen, et al.
Published: (2026)
SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation
by: Huang, Jingzhi, et al.
Published: (2026)
by: Huang, Jingzhi, et al.
Published: (2026)
DyGeoVLN: Infusing Dynamic Geometry Foundation Model into Vision-Language Navigation
by: Liu, Xiangchen, et al.
Published: (2026)
by: Liu, Xiangchen, et al.
Published: (2026)
JanusVLN: Decoupling Semantics and Spatiality with Dual Implicit Memory for Vision-Language Navigation
by: Zeng, Shuang, et al.
Published: (2025)
by: Zeng, Shuang, et al.
Published: (2025)
SUM-AgriVLN: Spatial Understanding Memory for Agricultural Vision-and-Language Navigation
by: Zhao, Xiaobei, et al.
Published: (2025)
by: Zhao, Xiaobei, et al.
Published: (2025)
FloorPlan-VLN: A New Paradigm for Floor Plan Guided Vision-Language Navigation
by: Chen, Kehan, et al.
Published: (2026)
by: Chen, Kehan, et al.
Published: (2026)
Learning Dual-Arm Coordination for Grasping Large Flat Objects
by: Wang, Yongliang, et al.
Published: (2025)
by: Wang, Yongliang, et al.
Published: (2025)
FlexVLN: Flexible Adaptation for Diverse Vision-and-Language Navigation Tasks
by: Zhang, Siqi, et al.
Published: (2025)
by: Zhang, Siqi, et al.
Published: (2025)
IMAC-AgriVLN: Can Agricultural Vision-and-Language Navigation Agents be Aware of Instruction Mistakes?
by: Zhao, Xiaobei, et al.
Published: (2026)
by: Zhao, Xiaobei, et al.
Published: (2026)
SkyVLN: Vision-and-Language Navigation and NMPC Control for UAVs in Urban Environments
by: Li, Tianshun, et al.
Published: (2025)
by: Li, Tianshun, et al.
Published: (2025)
LogisticsVLN: Vision-Language Navigation For Low-Altitude Terminal Delivery Based on Agentic UAVs
by: Zhang, Xinyuan, et al.
Published: (2025)
by: Zhang, Xinyuan, et al.
Published: (2025)
Exploring Bottlenecks in VLM-LLM Navigation: How 3D Scene Understanding Capability Impacts Zero-Shot VLN
by: Xia, Ziyi, et al.
Published: (2026)
by: Xia, Ziyi, et al.
Published: (2026)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
by: Zhao, Baining, et al.
Published: (2026)
by: Zhao, Baining, et al.
Published: (2026)
FSR-VLN: Fast and Slow Reasoning for Vision-Language Navigation with Hierarchical Multi-modal Scene Graph
by: Zhou, Xiaolin, et al.
Published: (2025)
by: Zhou, Xiaolin, et al.
Published: (2025)
UAV-VLN: End-to-End Vision Language guided Navigation for UAVs
by: Saxena, Pranav, et al.
Published: (2025)
by: Saxena, Pranav, et al.
Published: (2025)
Similar Items
-
PANav: Toward Privacy-Aware Robot Navigation via Vision-Language Models
by: Yu, Bangguo, et al.
Published: (2024) -
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
by: Yu, Bangguo, et al.
Published: (2023) -
MAG-Nav: Language-Driven Object Navigation Leveraging Memory-Reserved Active Grounding
by: Zhang, Weifan, et al.
Published: (2025) -
Single-Shot 6DoF Pose and 3D Size Estimation for Robotic Strawberry Harvesting
by: Li, Lun, et al.
Published: (2024) -
VITAL: Interactive Few-Shot Imitation Learning via Visual Human-in-the-Loop Corrections
by: Kasaei, Hamidreza, et al.
Published: (2024)