Meanings and Measurements: Multi-Agent Probabilistic Grounding for Vision-Language Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Padhan, Swagat, Jain, Lakshya, Shah, Bhavya Minesh, Patil, Omkar, Nguyen, Thao, Gopalan, Nakul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Composing Diffusion Policies for Few-shot Learning of Movement Trajectories
by: Patil, Omkar, et al.
Published: (2024)
by: Patil, Omkar, et al.
Published: (2024)
Continual Robot Skill and Task Learning via Dialogue
by: Gu, Weiwei, et al.
Published: (2024)
by: Gu, Weiwei, et al.
Published: (2024)
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
by: Li, Zerui, et al.
Published: (2025)
by: Li, Zerui, et al.
Published: (2025)
StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models
by: Pangaonkar, Kartikay Milind, et al.
Published: (2026)
by: Pangaonkar, Kartikay Milind, et al.
Published: (2026)
Do Visual Imaginations Improve Vision-and-Language Navigation Agents?
by: Perincherry, Akhil, et al.
Published: (2025)
by: Perincherry, Akhil, et al.
Published: (2025)
Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments
by: Hong, Haodong, et al.
Published: (2024)
by: Hong, Haodong, et al.
Published: (2024)
Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation
by: Chen, Jiaqi, et al.
Published: (2024)
by: Chen, Jiaqi, et al.
Published: (2024)
Factorizing Diffusion Policies for Observation Modality Prioritization
by: Patil, Omkar, et al.
Published: (2025)
by: Patil, Omkar, et al.
Published: (2025)
Learning Sequential Kinematic Models from Demonstrations for Multi-Jointed Articulated Objects
by: Gupta, Anmol, et al.
Published: (2025)
by: Gupta, Anmol, et al.
Published: (2025)
OpenFMNav: Towards Open-Set Zero-Shot Object Navigation via Vision-Language Foundation Models
by: Kuang, Yuxuan, et al.
Published: (2024)
by: Kuang, Yuxuan, et al.
Published: (2024)
What Limits Vision-and-Language Navigation ?
by: Wang, Yunheng, et al.
Published: (2026)
by: Wang, Yunheng, et al.
Published: (2026)
PokeNet: Learning Kinematic Models of Articulated Objects from Human Observations
by: Gupta, Anmol, et al.
Published: (2026)
by: Gupta, Anmol, et al.
Published: (2026)
ETPNav: Evolving Topological Planning for Vision-Language Navigation in Continuous Environments
by: An, Dong, et al.
Published: (2023)
by: An, Dong, et al.
Published: (2023)
Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
by: Taioli, Francesco, et al.
Published: (2024)
by: Taioli, Francesco, et al.
Published: (2024)
Vision-and-Language Navigation Generative Pretrained Transformer
by: Hanlin, Wen
Published: (2024)
by: Hanlin, Wen
Published: (2024)
Empathic Grounding: Explorations using Multimodal Interaction and Large Language Models with Conversational Agents
by: Arjmand, Mehdi, et al.
Published: (2024)
by: Arjmand, Mehdi, et al.
Published: (2024)
VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions
by: Su, Hung-Ting, et al.
Published: (2026)
by: Su, Hung-Ting, et al.
Published: (2026)
Automotive innovation landscaping using LLM
by: Gorain, Raju, et al.
Published: (2024)
by: Gorain, Raju, et al.
Published: (2024)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
by: Goetting, Dylan, et al.
Published: (2024)
by: Goetting, Dylan, et al.
Published: (2024)
ProGAL-VLA: Grounded Alignment through Prospective Reasoning in Vision-Language-Action Models
by: Darabi, Nastaran, et al.
Published: (2026)
by: Darabi, Nastaran, et al.
Published: (2026)
Stable Language Guidance for Vision-Language-Action Models
by: Zhan, Zhihao, et al.
Published: (2026)
by: Zhan, Zhihao, et al.
Published: (2026)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
by: Wang, Zhaowei, et al.
Published: (2024)
by: Wang, Zhaowei, et al.
Published: (2024)
Grounding Language Plans in Demonstrations Through Counterfactual Perturbations
by: Wang, Yanwei, et al.
Published: (2024)
by: Wang, Yanwei, et al.
Published: (2024)
DRAGON: A Dialogue-Based Robot for Assistive Navigation with Visual Language Grounding
by: Liu, Shuijing, et al.
Published: (2023)
by: Liu, Shuijing, et al.
Published: (2023)
CAMON: Cooperative Agents for Multi-Object Navigation with LLM-based Conversations
by: Wu, Pengying, et al.
Published: (2024)
by: Wu, Pengying, et al.
Published: (2024)
EdgeVLA: Efficient Vision-Language-Action Models
by: Budzianowski, Paweł, et al.
Published: (2025)
by: Budzianowski, Paweł, et al.
Published: (2025)
Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration
by: Wang, Jun, et al.
Published: (2026)
by: Wang, Jun, et al.
Published: (2026)
Can LLMs Translate Human Instructions into a Reinforcement Learning Agent's Internal Emergent Symbolic Representation?
by: Ma, Ziqi, et al.
Published: (2025)
by: Ma, Ziqi, et al.
Published: (2025)
NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models
by: Zhou, Gengze, et al.
Published: (2024)
by: Zhou, Gengze, et al.
Published: (2024)
CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model
by: Yu, Zhuoyuan, et al.
Published: (2025)
by: Yu, Zhuoyuan, et al.
Published: (2025)
Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation
by: Zhang, Pingrui, et al.
Published: (2025)
by: Zhang, Pingrui, et al.
Published: (2025)
Multi-Agent Consensus Seeking via Large Language Models
by: Chen, Huaben, et al.
Published: (2023)
by: Chen, Huaben, et al.
Published: (2023)
TRAVEL: Training-Free Retrieval and Alignment for Vision-and-Language Navigation
by: Rajabi, Navid, et al.
Published: (2025)
by: Rajabi, Navid, et al.
Published: (2025)
Rethinking the Embodied Gap in Vision-and-Language Navigation: A Holistic Study of Physical and Visual Disparities
by: Wang, Liuyi, et al.
Published: (2025)
by: Wang, Liuyi, et al.
Published: (2025)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
by: Wang, Yunheng, et al.
Published: (2025)
by: Wang, Yunheng, et al.
Published: (2025)
NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning
by: Lin, Bingqian, et al.
Published: (2024)
by: Lin, Bingqian, et al.
Published: (2024)
Multi-Turn Multi-Agent Dialogue for Collaborative Reconstruction Improves VLM Performance on Spatial Reasoning, But Only Barely
by: Kranti, Chalamalasetti, et al.
Published: (2026)
by: Kranti, Chalamalasetti, et al.
Published: (2026)
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
by: Werby, Abdelrhman, et al.
Published: (2024)
by: Werby, Abdelrhman, et al.
Published: (2024)
Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation
by: Wei, Ziming, et al.
Published: (2025)
by: Wei, Ziming, et al.
Published: (2025)
AgentThink: A Unified Framework for Tool-Augmented Chain-of-Thought Reasoning in Vision-Language Models for Autonomous Driving
by: Qian, Kangan, et al.
Published: (2025)
by: Qian, Kangan, et al.
Published: (2025)
Similar Items
-
Composing Diffusion Policies for Few-shot Learning of Movement Trajectories
by: Patil, Omkar, et al.
Published: (2024) -
Continual Robot Skill and Task Learning via Dialogue
by: Gu, Weiwei, et al.
Published: (2024) -
Ground-level Viewpoint Vision-and-Language Navigation in Continuous Environments
by: Li, Zerui, et al.
Published: (2025) -
StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models
by: Pangaonkar, Kartikay Milind, et al.
Published: (2026) -
Do Visual Imaginations Improve Vision-and-Language Navigation Agents?
by: Perincherry, Akhil, et al.
Published: (2025)