OpenFMNav: Towards Open-Set Zero-Shot Object Navigation via Vision-Language Foundation Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Kuang, Yuxuan, Lin, Hai, Jiang, Meng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
di: Wang, Zhaowei, et al.
Pubblicazione: (2024)
di: Wang, Zhaowei, et al.
Pubblicazione: (2024)
Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation
di: Chen, Jiaqi, et al.
Pubblicazione: (2024)
di: Chen, Jiaqi, et al.
Pubblicazione: (2024)
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024)
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
REFLEX: Metacognitive Reasoning for Reflective Zero-Shot Robotic Planning with Large Language Models
di: Lin, Wenjie, et al.
Pubblicazione: (2025)
di: Lin, Wenjie, et al.
Pubblicazione: (2025)
DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments
di: Ma, Ji, et al.
Pubblicazione: (2024)
di: Ma, Ji, et al.
Pubblicazione: (2024)
Can an Embodied Agent Find Your "Cat-shaped Mug"? LLM-Guided Exploration for Zero-Shot Object Navigation
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2023)
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2023)
Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation
di: Wei, Ziming, et al.
Pubblicazione: (2025)
di: Wei, Ziming, et al.
Pubblicazione: (2025)
HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System
di: Lyu, Kailin, et al.
Pubblicazione: (2026)
di: Lyu, Kailin, et al.
Pubblicazione: (2026)
osmAG-LLM: Zero-Shot Open-Vocabulary Object Navigation via Semantic Maps and Large Language Models Reasoning
di: Xie, Fujing, et al.
Pubblicazione: (2025)
di: Xie, Fujing, et al.
Pubblicazione: (2025)
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
di: Mitra, Chancharik, et al.
Pubblicazione: (2025)
di: Mitra, Chancharik, et al.
Pubblicazione: (2025)
LOC-ZSON: Language-driven Object-Centric Zero-Shot Object Retrieval and Navigation
di: Guan, Tianrui, et al.
Pubblicazione: (2024)
di: Guan, Tianrui, et al.
Pubblicazione: (2024)
Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation
di: Dai, Yinpei, et al.
Pubblicazione: (2023)
di: Dai, Yinpei, et al.
Pubblicazione: (2023)
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models
di: Wang, Tianyu, et al.
Pubblicazione: (2024)
di: Wang, Tianyu, et al.
Pubblicazione: (2024)
Stable Language Guidance for Vision-Language-Action Models
di: Zhan, Zhihao, et al.
Pubblicazione: (2026)
di: Zhan, Zhihao, et al.
Pubblicazione: (2026)
Language Models as Zero-Shot Trajectory Generators
di: Kwon, Teyun, et al.
Pubblicazione: (2023)
di: Kwon, Teyun, et al.
Pubblicazione: (2023)
Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments
di: Hong, Haodong, et al.
Pubblicazione: (2024)
di: Hong, Haodong, et al.
Pubblicazione: (2024)
Open-Source Image Editing Models Are Zero-Shot Vision Learners
di: Liu, Wei, et al.
Pubblicazione: (2026)
di: Liu, Wei, et al.
Pubblicazione: (2026)
Large Language Models as Zero-Shot Human Models for Human-Robot Interaction
di: Zhang, Bowen, et al.
Pubblicazione: (2023)
di: Zhang, Bowen, et al.
Pubblicazione: (2023)
OpenVLN: Open-world Aerial Vision-Language Navigation
di: Lin, Peican, et al.
Pubblicazione: (2025)
di: Lin, Peican, et al.
Pubblicazione: (2025)
FetchBot: Learning Generalizable Object Fetching in Cluttered Scenes via Zero-Shot Sim2Real
di: Liu, Weiheng, et al.
Pubblicazione: (2025)
di: Liu, Weiheng, et al.
Pubblicazione: (2025)
ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models
di: Zhang, Ying, et al.
Pubblicazione: (2025)
di: Zhang, Ying, et al.
Pubblicazione: (2025)
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA
di: Li, Junlong, et al.
Pubblicazione: (2022)
di: Li, Junlong, et al.
Pubblicazione: (2022)
Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigation
di: He, Yu, et al.
Pubblicazione: (2025)
di: He, Yu, et al.
Pubblicazione: (2025)
PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
di: Holk, Simon, et al.
Pubblicazione: (2024)
di: Holk, Simon, et al.
Pubblicazione: (2024)
Constraint-Aware Zero-Shot Vision-Language Navigation in Continuous Environments
di: Chen, Kehan, et al.
Pubblicazione: (2024)
di: Chen, Kehan, et al.
Pubblicazione: (2024)
Open-Loop Planning, Closed-Loop Verification: Speculative Verification for VLA
di: Wang, Zihua, et al.
Pubblicazione: (2026)
di: Wang, Zihua, et al.
Pubblicazione: (2026)
History-Augmented Vision-Language Models for Frontier-Based Zero-Shot Object Navigation
di: Habibpour, Mobin, et al.
Pubblicazione: (2025)
di: Habibpour, Mobin, et al.
Pubblicazione: (2025)
End-to-End Navigation with Vision Language Models: Transforming Spatial Reasoning into Question-Answering
di: Goetting, Dylan, et al.
Pubblicazione: (2024)
di: Goetting, Dylan, et al.
Pubblicazione: (2024)
Towards Multimodal Social Conversations with Robots: Using Vision-Language Models
di: Janssens, Ruben, et al.
Pubblicazione: (2025)
di: Janssens, Ruben, et al.
Pubblicazione: (2025)
Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
di: Taioli, Francesco, et al.
Pubblicazione: (2024)
What Limits Vision-and-Language Navigation ?
di: Wang, Yunheng, et al.
Pubblicazione: (2026)
di: Wang, Yunheng, et al.
Pubblicazione: (2026)
NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning
di: Lin, Bingqian, et al.
Pubblicazione: (2024)
di: Lin, Bingqian, et al.
Pubblicazione: (2024)
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
di: Werby, Abdelrhman, et al.
Pubblicazione: (2024)
di: Werby, Abdelrhman, et al.
Pubblicazione: (2024)
Building Explicit World Model for Zero-Shot Open-World Object Manipulation
di: Li, Xiaotong, et al.
Pubblicazione: (2026)
di: Li, Xiaotong, et al.
Pubblicazione: (2026)
SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models
di: Debnath, Arnab, et al.
Pubblicazione: (2025)
di: Debnath, Arnab, et al.
Pubblicazione: (2025)
Zero-shot Object Navigation with Vision-Language Models Reasoning
di: Wen, Congcong, et al.
Pubblicazione: (2024)
di: Wen, Congcong, et al.
Pubblicazione: (2024)
Open-Set 3D Semantic Instance Maps for Vision Language Navigation -- O3D-SIM
di: Nanwani, Laksh, et al.
Pubblicazione: (2024)
di: Nanwani, Laksh, et al.
Pubblicazione: (2024)
Towards Open-World Grasping with Large Vision-Language Models
di: Tziafas, Georgios, et al.
Pubblicazione: (2024)
di: Tziafas, Georgios, et al.
Pubblicazione: (2024)
EdgeVLA: Efficient Vision-Language-Action Models
di: Budzianowski, Paweł, et al.
Pubblicazione: (2025)
di: Budzianowski, Paweł, et al.
Pubblicazione: (2025)
Documenti analoghi
-
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
di: Wang, Zhaowei, et al.
Pubblicazione: (2024) -
Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation
di: Chen, Jiaqi, et al.
Pubblicazione: (2024) -
Open-Nav: Exploring Zero-Shot Vision-and-Language Navigation in Continuous Environment with Open-Source LLMs
di: Qiao, Yanyuan, et al.
Pubblicazione: (2024) -
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
di: Wang, Yunheng, et al.
Pubblicazione: (2025) -
REFLEX: Metacognitive Reasoning for Reflective Zero-Shot Robotic Planning with Large Language Models
di: Lin, Wenjie, et al.
Pubblicazione: (2025)