Zero-shot Object-Centric Instruction Following: Integrating Foundation Models with Traditional Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Raychaudhuri, Sonia, Ta, Duy, Ashton, Katrina, Chang, Angel X., Wang, Jiuguang, Bucher, Bernadette |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ASHiTA: Automatic Scene-grounded HIerarchical Task Analysis
by: Chang, Yun, et al.
Published: (2025)
by: Chang, Yun, et al.
Published: (2025)
MOPA: Modular Object Navigation with PointGoal Agents
by: Raychaudhuri, Sonia, et al.
Published: (2023)
by: Raychaudhuri, Sonia, et al.
Published: (2023)
Semantic Mapping in Indoor Embodied AI -- A Survey on Advances, Challenges, and Future Directions
by: Raychaudhuri, Sonia, et al.
Published: (2025)
by: Raychaudhuri, Sonia, et al.
Published: (2025)
MLFM: Multi-Layered Feature Maps for Richer Language Understanding in Zero-Shot Semantic Navigation
by: Raychaudhuri, Sonia, et al.
Published: (2025)
by: Raychaudhuri, Sonia, et al.
Published: (2025)
VoroNav: Voronoi-based Zero-shot Object Navigation with Large Language Model
by: Wu, Pengying, et al.
Published: (2024)
by: Wu, Pengying, et al.
Published: (2024)
Continuously Improving Mobile Manipulation with Autonomous Real-World RL
by: Mendonca, Russell, et al.
Published: (2024)
by: Mendonca, Russell, et al.
Published: (2024)
LOC-ZSON: Language-driven Object-Centric Zero-Shot Object Retrieval and Navigation
by: Guan, Tianrui, et al.
Published: (2024)
by: Guan, Tianrui, et al.
Published: (2024)
Object-Centric Instruction Augmentation for Robotic Manipulation
by: Wen, Junjie, et al.
Published: (2024)
by: Wen, Junjie, et al.
Published: (2024)
Balancing Performance and Efficiency in Zero-shot Robotic Navigation
by: Kuzmenko, Dmytro, et al.
Published: (2024)
by: Kuzmenko, Dmytro, et al.
Published: (2024)
SG-Nav: Online 3D Scene Graph Prompting for LLM-based Zero-shot Object Navigation
by: Yin, Hang, et al.
Published: (2024)
by: Yin, Hang, et al.
Published: (2024)
FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation
by: Shao, Dian, et al.
Published: (2026)
by: Shao, Dian, et al.
Published: (2026)
VISOR: VIsual Spatial Object Reasoning for Language-driven Object Navigation
by: Taioli, Francesco, et al.
Published: (2026)
by: Taioli, Francesco, et al.
Published: (2026)
InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment
by: Long, Yuxing, et al.
Published: (2024)
by: Long, Yuxing, et al.
Published: (2024)
UniGoal: Towards Universal Zero-shot Goal-oriented Navigation
by: Yin, Hang, et al.
Published: (2025)
by: Yin, Hang, et al.
Published: (2025)
Zero-shot Reconstruction of In-Scene Object Manipulation from Video
by: Lin, Dixuan, et al.
Published: (2025)
by: Lin, Dixuan, et al.
Published: (2025)
ViTa-Zero: Zero-shot Visuotactile Object 6D Pose Estimation
by: Li, Hongyu, et al.
Published: (2025)
by: Li, Hongyu, et al.
Published: (2025)
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
by: Zhong, Linqing, et al.
Published: (2024)
by: Zhong, Linqing, et al.
Published: (2024)
DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments
by: Ma, Ji, et al.
Published: (2024)
by: Ma, Ji, et al.
Published: (2024)
Bridging the Indoor-Outdoor Gap: Vision-Centric Instruction-Guided Embodied Navigation for the Last Meters
by: Zhao, Yuxiang, et al.
Published: (2026)
by: Zhao, Yuxiang, et al.
Published: (2026)
Multimodal LLM Guided Exploration and Active Mapping using Fisher Information
by: Jiang, Wen, et al.
Published: (2024)
by: Jiang, Wen, et al.
Published: (2024)
The impact of Compositionality in Zero-shot Multi-label action recognition for Object-based tasks
by: Calabrese, Carmela, et al.
Published: (2024)
by: Calabrese, Carmela, et al.
Published: (2024)
Rethinking Progression of Memory State in Robotic Manipulation: An Object-Centric Perspective
by: Chung, Nhat, et al.
Published: (2025)
by: Chung, Nhat, et al.
Published: (2025)
CuriousBot: Interactive Mobile Exploration via Actionable 3D Relational Object Graph
by: Wang, Yixuan, et al.
Published: (2025)
by: Wang, Yixuan, et al.
Published: (2025)
Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigation
by: He, Yu, et al.
Published: (2025)
by: He, Yu, et al.
Published: (2025)
PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and Planning
by: Villar-Corrales, Angel, et al.
Published: (2025)
by: Villar-Corrales, Angel, et al.
Published: (2025)
Imagine2Real: Towards Zero-shot Humanoid-Object Interaction via Video Generative Priors
by: Chen, Jiahe, et al.
Published: (2026)
by: Chen, Jiahe, et al.
Published: (2026)
AERR-Nav: Adaptive Exploration-Recovery-Reminiscing Strategy for Zero-Shot Object Navigation
by: Huang, Jingzhi, et al.
Published: (2026)
by: Huang, Jingzhi, et al.
Published: (2026)
IntentionNav: A Benchmark for Intent-Driven Object Navigation from Implicit Human Instruction
by: Qian, Lin, et al.
Published: (2026)
by: Qian, Lin, et al.
Published: (2026)
NovaFlow: Zero-Shot Manipulation via Actionable Flow from Generated Videos
by: Li, Hongyu, et al.
Published: (2025)
by: Li, Hongyu, et al.
Published: (2025)
TDANet: Target-Directed Attention Network For Object-Goal Visual Navigation With Zero-Shot Ability
by: Lian, Shiwei, et al.
Published: (2024)
by: Lian, Shiwei, et al.
Published: (2024)
PanoNav: Mapless Zero-Shot Object Navigation with Panoramic Scene Parsing and Dynamic Memory
by: Jin, Qunchao, et al.
Published: (2025)
by: Jin, Qunchao, et al.
Published: (2025)
ReMemNav: A Rethinking and Memory-Augmented Framework for Zero-Shot Object Navigation
by: Wu, Feng, et al.
Published: (2026)
by: Wu, Feng, et al.
Published: (2026)
WMNav: Integrating Vision-Language Models into World Models for Object Goal Navigation
by: Nie, Dujun, et al.
Published: (2025)
by: Nie, Dujun, et al.
Published: (2025)
ZISVFM: Zero-Shot Object Instance Segmentation in Indoor Robotic Environments with Vision Foundation Models
by: Zhang, Ying, et al.
Published: (2025)
by: Zhang, Ying, et al.
Published: (2025)
Following the Human Thread in Social Navigation
by: Scofano, Luca, et al.
Published: (2024)
by: Scofano, Luca, et al.
Published: (2024)
ConsistNav: Closing the Action Consistency Gap in Zero-Shot Object Navigation with Semantic Executive Control
by: Wang, Haosen, et al.
Published: (2026)
by: Wang, Haosen, et al.
Published: (2026)
Leveraging Unknown Objects to Construct Labeled-Unlabeled Meta-Relationships for Zero-Shot Object Navigation
by: Zheng, Yanwei, et al.
Published: (2024)
by: Zheng, Yanwei, et al.
Published: (2024)
Zero-BEV: Zero-shot Projection of Any First-Person Modality to BEV Maps
by: Monaci, Gianluca, et al.
Published: (2024)
by: Monaci, Gianluca, et al.
Published: (2024)
NavigateDiff: Visual Predictors are Zero-Shot Navigation Assistants
by: Qin, Yiran, et al.
Published: (2025)
by: Qin, Yiran, et al.
Published: (2025)
Disentangled Object-Centric Image Representation for Robotic Manipulation
by: Emukpere, David, et al.
Published: (2025)
by: Emukpere, David, et al.
Published: (2025)
Similar Items
-
ASHiTA: Automatic Scene-grounded HIerarchical Task Analysis
by: Chang, Yun, et al.
Published: (2025) -
MOPA: Modular Object Navigation with PointGoal Agents
by: Raychaudhuri, Sonia, et al.
Published: (2023) -
Semantic Mapping in Indoor Embodied AI -- A Survey on Advances, Challenges, and Future Directions
by: Raychaudhuri, Sonia, et al.
Published: (2025) -
MLFM: Multi-Layered Feature Maps for Richer Language Understanding in Zero-Shot Semantic Navigation
by: Raychaudhuri, Sonia, et al.
Published: (2025) -
VoroNav: Voronoi-based Zero-shot Object Navigation with Large Language Model
by: Wu, Pengying, et al.
Published: (2024)