OpenNav: Open-World Navigation with Multimodal Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Yuan, Mingfeng, Wang, Letian, Waslander, Steven L. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STaR: Scalable Task-Conditioned Retrieval for Long-Horizon Multimodal Robot Memory
by: Yuan, Mingfeng, et al.
Published: (2026)
by: Yuan, Mingfeng, et al.
Published: (2026)
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
by: Ortega-Peimbert, Jesús, et al.
Published: (2025)
by: Ortega-Peimbert, Jesús, et al.
Published: (2025)
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
by: Shao, Hao, et al.
Published: (2026)
by: Shao, Hao, et al.
Published: (2026)
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
by: Zhou, Yang, et al.
Published: (2026)
by: Zhou, Yang, et al.
Published: (2026)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
by: Zhao, Zhikai, et al.
Published: (2026)
by: Zhao, Zhikai, et al.
Published: (2026)
SayNav: Grounding Large Language Models for Dynamic Planning to Navigation in New Environments
by: Rajvanshi, Abhinav, et al.
Published: (2023)
by: Rajvanshi, Abhinav, et al.
Published: (2023)
SayCoNav: Utilizing Large Language Models for Adaptive Collaboration in Decentralized Multi-Robot Navigation
by: Rajvanshi, Abhinav, et al.
Published: (2025)
by: Rajvanshi, Abhinav, et al.
Published: (2025)
SmartPretrain: Model-Agnostic and Dataset-Agnostic Representation Learning for Motion Prediction
by: Zhou, Yang, et al.
Published: (2024)
by: Zhou, Yang, et al.
Published: (2024)
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
by: Yu, Bangguo, et al.
Published: (2023)
by: Yu, Bangguo, et al.
Published: (2023)
Nav-EE: Navigation-Guided Early Exiting for Efficient Vision-Language Models in Autonomous Driving
by: Hu, Haibo, et al.
Published: (2025)
by: Hu, Haibo, et al.
Published: (2025)
SmartRefine: A Scenario-Adaptive Refinement Framework for Efficient Motion Prediction
by: Zhou, Yang, et al.
Published: (2024)
by: Zhou, Yang, et al.
Published: (2024)
LLM-State: Open World State Representation for Long-horizon Task Planning with Large Language Model
by: Chen, Siwei, et al.
Published: (2023)
by: Chen, Siwei, et al.
Published: (2023)
EfficientNav: Towards On-Device Object-Goal Navigation with Navigation Map Caching and Retrieval
by: Yang, Zebin, et al.
Published: (2025)
by: Yang, Zebin, et al.
Published: (2025)
DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning
by: Zhou, Yang, et al.
Published: (2026)
by: Zhou, Yang, et al.
Published: (2026)
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding
by: Zhang, Yuhang, et al.
Published: (2025)
by: Zhang, Yuhang, et al.
Published: (2025)
NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models
by: Zhou, Gengze, et al.
Published: (2024)
by: Zhou, Gengze, et al.
Published: (2024)
Humanoid World Models: Open World Foundation Models for Humanoid Robotics
by: Ali, Muhammad Qasim, et al.
Published: (2025)
by: Ali, Muhammad Qasim, et al.
Published: (2025)
DKPROMPT: Domain Knowledge Prompting Vision-Language Models for Open-World Planning
by: Zhang, Xiaohan, et al.
Published: (2024)
by: Zhang, Xiaohan, et al.
Published: (2024)
CityNavAgent: Aerial Vision-and-Language Navigation with Hierarchical Semantic Planning and Global Memory
by: Zhang, Weichen, et al.
Published: (2025)
by: Zhang, Weichen, et al.
Published: (2025)
OpenObject-NAV: Open-Vocabulary Object-Oriented Navigation Based on Dynamic Carrier-Relationship Scene Graph
by: Tang, Yujie, et al.
Published: (2024)
by: Tang, Yujie, et al.
Published: (2024)
PM-Nav: Priori-Map Guided Embodied Navigation in Functional Buildings
by: Gao, Jiang, et al.
Published: (2026)
by: Gao, Jiang, et al.
Published: (2026)
MemoNav: Working Memory Model for Visual Navigation
by: Li, Hongxin, et al.
Published: (2024)
by: Li, Hongxin, et al.
Published: (2024)
HyPerNav: Hybrid Perception for Object-Oriented Navigation in Unknown Environment
by: Yin, Zecheng, et al.
Published: (2025)
by: Yin, Zecheng, et al.
Published: (2025)
PSG-Nav: Probabilistic Scene Graph Navigation via Multiverse Decision Making
by: Chen, Rufeng, et al.
Published: (2026)
by: Chen, Rufeng, et al.
Published: (2026)
Toward General Object-level Mapping from Sparse Views with 3D Diffusion Priors
by: Liao, Ziwei, et al.
Published: (2024)
by: Liao, Ziwei, et al.
Published: (2024)
How Secure Are Large Language Models (LLMs) for Navigation in Urban Environments?
by: Wen, Congcong, et al.
Published: (2024)
by: Wen, Congcong, et al.
Published: (2024)
PlaceNav: Topological Navigation through Place Recognition
by: Suomela, Lauri, et al.
Published: (2023)
by: Suomela, Lauri, et al.
Published: (2023)
SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models
by: Debnath, Arnab, et al.
Published: (2025)
by: Debnath, Arnab, et al.
Published: (2025)
GSON: A Group-based Social Navigation Framework with Large Multimodal Model
by: Luo, Shangyi, et al.
Published: (2024)
by: Luo, Shangyi, et al.
Published: (2024)
MASt3R-Nav: WayPixel Navigation in Relative 3D Maps
by: Garg, Vansh, et al.
Published: (2026)
by: Garg, Vansh, et al.
Published: (2026)
ProCompNav: Proactive Instance Navigation with Comparative Judgment for Ambiguous User Queries
by: Kwon, Junhyuk, et al.
Published: (2026)
by: Kwon, Junhyuk, et al.
Published: (2026)
CarDreamer: Open-Source Learning Platform for World Model based Autonomous Driving
by: Gao, Dechen, et al.
Published: (2024)
by: Gao, Dechen, et al.
Published: (2024)
Grounded World Model for Semantically Generalizable Planning
by: Li, Quanyi, et al.
Published: (2026)
by: Li, Quanyi, et al.
Published: (2026)
OctoNav: Towards Generalist Embodied Navigation
by: Gao, Chen, et al.
Published: (2025)
by: Gao, Chen, et al.
Published: (2025)
Open-World Drone Active Tracking with Goal-Centered Rewards
by: Sun, Haowei, et al.
Published: (2024)
by: Sun, Haowei, et al.
Published: (2024)
How VLAs (Really) Work In Open-World Environments
by: Rasouli, Amir, et al.
Published: (2026)
by: Rasouli, Amir, et al.
Published: (2026)
Creating and Repairing Robot Programs in Open-World Domains
by: Schlesinger, Claire, et al.
Published: (2024)
by: Schlesinger, Claire, et al.
Published: (2024)
Contextual Safety Reasoning and Grounding for Open-World Robots
by: Ravichandran, Zachary, et al.
Published: (2026)
by: Ravichandran, Zachary, et al.
Published: (2026)
Robot Task Planning and Situation Handling in Open Worlds
by: Ding, Yan, et al.
Published: (2022)
by: Ding, Yan, et al.
Published: (2022)
LagMemo: Language 3D Gaussian Splatting Memory for Multi-modal Open-vocabulary Multi-goal Visual Navigation
by: Zhou, Haotian, et al.
Published: (2025)
by: Zhou, Haotian, et al.
Published: (2025)
Similar Items
-
STaR: Scalable Task-Conditioned Retrieval for Long-Horizon Multimodal Robot Memory
by: Yuan, Mingfeng, et al.
Published: (2026) -
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
by: Ortega-Peimbert, Jesús, et al.
Published: (2025) -
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
by: Shao, Hao, et al.
Published: (2026) -
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
by: Zhou, Yang, et al.
Published: (2026) -
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
by: Zhao, Zhikai, et al.
Published: (2026)