History-Augmented Vision-Language Models for Frontier-Based Zero-Shot Object Navigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Habibpour, Mobin, Afghah, Fatemeh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
di: Habibpour, Mobin, et al.
Pubblicazione: (2025)
di: Habibpour, Mobin, et al.
Pubblicazione: (2025)
Zero-shot Object Navigation with Vision-Language Models Reasoning
di: Wen, Congcong, et al.
Pubblicazione: (2024)
di: Wen, Congcong, et al.
Pubblicazione: (2024)
SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models
di: Debnath, Arnab, et al.
Pubblicazione: (2025)
di: Debnath, Arnab, et al.
Pubblicazione: (2025)
FIRE-VLM: A Vision-Language-Driven Reinforcement Learning Framework for UAV Wildfire Tracking in a Physics-Grounded Fire Digital Twin
di: Webb, Chris, et al.
Pubblicazione: (2026)
di: Webb, Chris, et al.
Pubblicazione: (2026)
Meta Reinforcement Learning Approach for Adaptive Resource Optimization in O-RAN
di: Lotfi, Fatemeh, et al.
Pubblicazione: (2024)
di: Lotfi, Fatemeh, et al.
Pubblicazione: (2024)
Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigation
di: He, Yu, et al.
Pubblicazione: (2025)
di: He, Yu, et al.
Pubblicazione: (2025)
LGR: LLM-Guided Ranking of Frontiers for Object Goal Navigation
di: Uno, Mitsuaki, et al.
Pubblicazione: (2025)
di: Uno, Mitsuaki, et al.
Pubblicazione: (2025)
Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots
di: Shi, Junyao, et al.
Pubblicazione: (2025)
di: Shi, Junyao, et al.
Pubblicazione: (2025)
SpatialNav: Leveraging Spatial Scene Graphs for Zero-Shot Vision-and-Language Navigation
di: Zhang, Jiwen, et al.
Pubblicazione: (2026)
di: Zhang, Jiwen, et al.
Pubblicazione: (2026)
R2F: Repurposing Ray Frontiers for LLM-free Object Navigation
di: Argenziano, Francesco, et al.
Pubblicazione: (2026)
di: Argenziano, Francesco, et al.
Pubblicazione: (2026)
DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
di: Wang, Yunheng, et al.
Pubblicazione: (2025)
Leveraging Unknown Objects to Construct Labeled-Unlabeled Meta-Relationships for Zero-Shot Object Navigation
di: Zheng, Yanwei, et al.
Pubblicazione: (2024)
di: Zheng, Yanwei, et al.
Pubblicazione: (2024)
Zero-Shot Generalization of Vision-Based RL Without Data Augmentation
di: Batra, Sumeet, et al.
Pubblicazione: (2024)
di: Batra, Sumeet, et al.
Pubblicazione: (2024)
3DGSNav: Enhancing Vision-Language Model Reasoning for Object Navigation via Active 3D Gaussian Splatting
di: Zheng, Wancai, et al.
Pubblicazione: (2026)
di: Zheng, Wancai, et al.
Pubblicazione: (2026)
REST: Receding Horizon Explorative Steiner Tree for Zero-Shot Object-Goal Navigation
di: Xiao, Shuqi, et al.
Pubblicazione: (2026)
di: Xiao, Shuqi, et al.
Pubblicazione: (2026)
Reliable Semantic Understanding for Real World Zero-shot Object Goal Navigation
di: Unlu, Halil Utku, et al.
Pubblicazione: (2024)
di: Unlu, Halil Utku, et al.
Pubblicazione: (2024)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
di: Singh, Harsh, et al.
Pubblicazione: (2024)
di: Singh, Harsh, et al.
Pubblicazione: (2024)
Vision-Language Navigation with Continual Learning
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
Can an Embodied Agent Find Your "Cat-shaped Mug"? LLM-Guided Exploration for Zero-Shot Object Navigation
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2023)
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2023)
Fly0: Decoupling Semantic Grounding from Geometric Planning for Zero-Shot Aerial Navigation
di: Xu, Zhenxing, et al.
Pubblicazione: (2026)
di: Xu, Zhenxing, et al.
Pubblicazione: (2026)
Enhancing Agricultural Environment Perception via Active Vision and Zero-Shot Learning
di: La Greca, Michele Carlo, et al.
Pubblicazione: (2024)
di: La Greca, Michele Carlo, et al.
Pubblicazione: (2024)
SimToolReal: An Object-Centric Policy for Zero-Shot Dexterous Tool Manipulation
di: Kedia, Kushal, et al.
Pubblicazione: (2026)
di: Kedia, Kushal, et al.
Pubblicazione: (2026)
Zero-Shot Metric Depth Estimation via Monocular Visual-Inertial Rescaling for Autonomous Aerial Navigation
di: Yang, Steven, et al.
Pubblicazione: (2025)
di: Yang, Steven, et al.
Pubblicazione: (2025)
Interactive Navigation in Environments with Traversable Obstacles Using Large Language and Vision-Language Models
di: Zhang, Zhen, et al.
Pubblicazione: (2023)
di: Zhang, Zhen, et al.
Pubblicazione: (2023)
Active Test-time Vision-Language Navigation
di: Ko, Heeju, et al.
Pubblicazione: (2025)
di: Ko, Heeju, et al.
Pubblicazione: (2025)
ShapeGrasp: Zero-Shot Task-Oriented Grasping with Large Language Models through Geometric Decomposition
di: Li, Samuel, et al.
Pubblicazione: (2024)
di: Li, Samuel, et al.
Pubblicazione: (2024)
VLM-RRT: Vision Language Model Guided RRT Search for Autonomous UAV Navigation
di: Ye, Jianlin, et al.
Pubblicazione: (2025)
di: Ye, Jianlin, et al.
Pubblicazione: (2025)
HTNav: A Hybrid Navigation Framework with Tiered Structure for Urban Aerial Vision-and-Language Navigation
di: Fan, Chengjie, et al.
Pubblicazione: (2026)
di: Fan, Chengjie, et al.
Pubblicazione: (2026)
AINav: Large Language Model-Based Adaptive Interactive Navigation
di: Zhou, Kangjie, et al.
Pubblicazione: (2025)
di: Zhou, Kangjie, et al.
Pubblicazione: (2025)
Exploring the Reliability of Foundation Model-Based Frontier Selection in Zero-Shot Object Goal Navigation
di: Yuan, Shuaihang, et al.
Pubblicazione: (2024)
di: Yuan, Shuaihang, et al.
Pubblicazione: (2024)
Cog-GA: A Large Language Models-based Generative Agent for Vision-Language Navigation in Continuous Environments
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
VLN-Zero: Rapid Exploration and Cache-Enabled Neurosymbolic Vision-Language Planning for Zero-Shot Transfer in Robot Navigation
di: Bhatt, Neel P., et al.
Pubblicazione: (2025)
di: Bhatt, Neel P., et al.
Pubblicazione: (2025)
Nav-EE: Navigation-Guided Early Exiting for Efficient Vision-Language Models in Autonomous Driving
di: Hu, Haibo, et al.
Pubblicazione: (2025)
di: Hu, Haibo, et al.
Pubblicazione: (2025)
One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation
di: Busch, Finn Lukas, et al.
Pubblicazione: (2024)
di: Busch, Finn Lukas, et al.
Pubblicazione: (2024)
NVP-HRI: Zero Shot Natural Voice and Posture-based Human-Robot Interaction via Large Language Model
di: Lai, Yuzhi, et al.
Pubblicazione: (2025)
di: Lai, Yuzhi, et al.
Pubblicazione: (2025)
Endowing Embodied Agents with Spatial Reasoning Capabilities for Vision-and-Language Navigation
di: Bai, Qianqian, et al.
Pubblicazione: (2025)
di: Bai, Qianqian, et al.
Pubblicazione: (2025)
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding
di: Zhang, Yuhang, et al.
Pubblicazione: (2025)
di: Zhang, Yuhang, et al.
Pubblicazione: (2025)
SeqWalker: Sequential-Horizon Vision-and-Language Navigation with Hierarchical Planning
di: Han, Zebin, et al.
Pubblicazione: (2026)
di: Han, Zebin, et al.
Pubblicazione: (2026)
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
di: Yu, Bangguo, et al.
Pubblicazione: (2023)
di: Yu, Bangguo, et al.
Pubblicazione: (2023)
OpenObject-NAV: Open-Vocabulary Object-Oriented Navigation Based on Dynamic Carrier-Relationship Scene Graph
di: Tang, Yujie, et al.
Pubblicazione: (2024)
di: Tang, Yujie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
di: Habibpour, Mobin, et al.
Pubblicazione: (2025) -
Zero-shot Object Navigation with Vision-Language Models Reasoning
di: Wen, Congcong, et al.
Pubblicazione: (2024) -
SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models
di: Debnath, Arnab, et al.
Pubblicazione: (2025) -
FIRE-VLM: A Vision-Language-Driven Reinforcement Learning Framework for UAV Wildfire Tracking in a Physics-Grounded Fire Digital Twin
di: Webb, Chris, et al.
Pubblicazione: (2026) -
Meta Reinforcement Learning Approach for Adaptive Resource Optimization in O-RAN
di: Lotfi, Fatemeh, et al.
Pubblicazione: (2024)