Reliable Semantic Understanding for Real World Zero-shot Object Goal Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Unlu, Halil Utku, Yuan, Shuaihang, Wen, Congcong, Huang, Hao, Tzes, Anthony, Fang, Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring the Reliability of Foundation Model-Based Frontier Selection in Zero-Shot Object Goal Navigation
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2024)
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2024)
GAMap: Zero-Shot Object Goal Navigation with Multi-Scale Geometric-Affordance Guidance
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2024)
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2024)
Zero-shot Object Navigation with Vision-Language Models Reasoning
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
One-shot Adaptation of Humanoid Whole-body Motion with Walking Priors
von: Huang, Hao, et al.
Veröffentlicht: (2025)
von: Huang, Hao, et al.
Veröffentlicht: (2025)
How Secure Are Large Language Models (LLMs) for Navigation in Urban Environments?
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
MapBERT: Bitwise Masked Modeling for Real-Time Semantic Mapping Generation
von: Deng, Yijie, et al.
Veröffentlicht: (2025)
von: Deng, Yijie, et al.
Veröffentlicht: (2025)
Humanoid Agent via Embodied Chain-of-Action Reasoning with Multimodal Foundation Models for Zero-Shot Loco-Manipulation
von: Wen, Congcong, et al.
Veröffentlicht: (2025)
von: Wen, Congcong, et al.
Veröffentlicht: (2025)
Efficient and Distributed Large-Scale 3D Map Registration using Tomographic Features
von: Unlu, Halil Utku, et al.
Veröffentlicht: (2024)
von: Unlu, Halil Utku, et al.
Veröffentlicht: (2024)
Wavelet Policy: Lifting Scheme for Policy Learning in Long-Horizon Tasks
von: Huang, Hao, et al.
Veröffentlicht: (2025)
von: Huang, Hao, et al.
Veröffentlicht: (2025)
AudioScene: Integrating Object-Event Audio into 3D Scenes
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2025)
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2025)
Hierarchical Scoring with 3D Gaussian Splatting for Instance Image-Goal Navigation
von: Deng, Yijie, et al.
Veröffentlicht: (2025)
von: Deng, Yijie, et al.
Veröffentlicht: (2025)
Socially-Aware Robot Navigation Enhanced by Bidirectional Natural Language Conversations Using Large Language Models
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
von: Wen, Congcong, et al.
Veröffentlicht: (2024)
H2-COMPACT: Human-Humanoid Co-Manipulation via Adaptive Contact Trajectory Policies
von: Bethala, Geeta Chandra Raju, et al.
Veröffentlicht: (2025)
von: Bethala, Geeta Chandra Raju, et al.
Veröffentlicht: (2025)
AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation
von: Deng, Yijie, et al.
Veröffentlicht: (2026)
von: Deng, Yijie, et al.
Veröffentlicht: (2026)
Semantic Environment Atlas for Object-Goal Navigation
von: Kim, Nuri, et al.
Veröffentlicht: (2024)
von: Kim, Nuri, et al.
Veröffentlicht: (2024)
Think, Remember, Navigate: Zero-Shot Object-Goal Navigation with VLM-Powered Reasoning
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2025)
SEEK: Semantic Reasoning for Object Goal Navigation in Real World Inspection Tasks
von: Ginting, Muhammad Fadhil, et al.
Veröffentlicht: (2024)
von: Ginting, Muhammad Fadhil, et al.
Veröffentlicht: (2024)
One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation
von: Busch, Finn Lukas, et al.
Veröffentlicht: (2024)
von: Busch, Finn Lukas, et al.
Veröffentlicht: (2024)
SemNav: A Model-Based Planner for Zero-Shot Object Goal Navigation Using Vision-Foundation Models
von: Debnath, Arnab, et al.
Veröffentlicht: (2025)
von: Debnath, Arnab, et al.
Veröffentlicht: (2025)
LGR: LLM-Guided Ranking of Frontiers for Object Goal Navigation
von: Uno, Mitsuaki, et al.
Veröffentlicht: (2025)
von: Uno, Mitsuaki, et al.
Veröffentlicht: (2025)
EfficientNav: Towards On-Device Object-Goal Navigation with Navigation Map Caching and Retrieval
von: Yang, Zebin, et al.
Veröffentlicht: (2025)
von: Yang, Zebin, et al.
Veröffentlicht: (2025)
REST: Receding Horizon Explorative Steiner Tree for Zero-Shot Object-Goal Navigation
von: Xiao, Shuqi, et al.
Veröffentlicht: (2026)
von: Xiao, Shuqi, et al.
Veröffentlicht: (2026)
Learning World Models for Unconstrained Goal Navigation
von: Duan, Yuanlin, et al.
Veröffentlicht: (2024)
von: Duan, Yuanlin, et al.
Veröffentlicht: (2024)
Goal-VLA: Image-Generative VLMs as Object-Centric World Models Empowering Zero-shot Robot Manipulation
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
von: Chen, Haonan, et al.
Veröffentlicht: (2025)
UniGoal: Towards Universal Zero-shot Goal-oriented Navigation
von: Yin, Hang, et al.
Veröffentlicht: (2025)
von: Yin, Hang, et al.
Veröffentlicht: (2025)
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
Show and Grasp: Few-shot Semantic Segmentation for Robot Grasping through Zero-shot Foundation Models
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
A Chain-of-Thought Subspace Meta-Learning for Few-shot Image Captioning with Large Vision and Language Models
von: Huang, Hao, et al.
Veröffentlicht: (2025)
von: Huang, Hao, et al.
Veröffentlicht: (2025)
UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios
von: Gonzalez, Antonio Galiza Cerdeira, et al.
Veröffentlicht: (2024)
von: Gonzalez, Antonio Galiza Cerdeira, et al.
Veröffentlicht: (2024)
SPINE: Online Semantic Planning for Missions with Incomplete Natural Language Specifications in Unstructured Environments
von: Ravichandran, Zachary, et al.
Veröffentlicht: (2024)
von: Ravichandran, Zachary, et al.
Veröffentlicht: (2024)
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
von: Yokoyama, Naoki, et al.
Veröffentlicht: (2024)
von: Yokoyama, Naoki, et al.
Veröffentlicht: (2024)
MCNav: Memory-Aware Dynamic Cognitive Map for Zero-shot Goal-oriented Navigation
von: Li, Jingyu, et al.
Veröffentlicht: (2026)
von: Li, Jingyu, et al.
Veröffentlicht: (2026)
Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigation
von: He, Yu, et al.
Veröffentlicht: (2025)
von: He, Yu, et al.
Veröffentlicht: (2025)
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
MV-UMI: A Scalable Multi-View Interface for Cross-Embodiment Learning
von: Rayyan, Omar, et al.
Veröffentlicht: (2025)
von: Rayyan, Omar, et al.
Veröffentlicht: (2025)
Open Scene Graphs for Open World Object-Goal Navigation
von: Loo, Joel, et al.
Veröffentlicht: (2024)
von: Loo, Joel, et al.
Veröffentlicht: (2024)
Open Scene Graphs for Open-World Object-Goal Navigation
von: Loo, Joel, et al.
Veröffentlicht: (2025)
von: Loo, Joel, et al.
Veröffentlicht: (2025)
Zero-shot Interactive Perception
von: Sripada, Venkatesh, et al.
Veröffentlicht: (2026)
von: Sripada, Venkatesh, et al.
Veröffentlicht: (2026)
IPPON: Common Sense Guided Informative Path Planning for Object Goal Navigation
von: Qu, Kaixian, et al.
Veröffentlicht: (2024)
von: Qu, Kaixian, et al.
Veröffentlicht: (2024)
Object-Oriented Semantic Mapping for Reliable UAVs Navigation
von: Canh, Thanh Nguyen, et al.
Veröffentlicht: (2024)
von: Canh, Thanh Nguyen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Exploring the Reliability of Foundation Model-Based Frontier Selection in Zero-Shot Object Goal Navigation
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2024) -
GAMap: Zero-Shot Object Goal Navigation with Multi-Scale Geometric-Affordance Guidance
von: Yuan, Shuaihang, et al.
Veröffentlicht: (2024) -
Zero-shot Object Navigation with Vision-Language Models Reasoning
von: Wen, Congcong, et al.
Veröffentlicht: (2024) -
One-shot Adaptation of Humanoid Whole-body Motion with Walking Priors
von: Huang, Hao, et al.
Veröffentlicht: (2025) -
How Secure Are Large Language Models (LLMs) for Navigation in Urban Environments?
von: Wen, Congcong, et al.
Veröffentlicht: (2024)