Where to Fetch: Extracting Visual Scene Representation from Large Pre-Trained Models for Robotic Goal Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yu, Li, Dayou, Zhao, Chenkun, Wang, Ruifeng, Song, Ran, Zhang, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MPGNet: Learning Move-Push-Grasping Synergy for Target-Oriented Grasping in Occluded Scenes
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
Learning Instruction-Guided Manipulation Affordance via Large Models for Embodied Robotic Tasks
by: Li, Dayou, et al.
Published: (2024)
by: Li, Dayou, et al.
Published: (2024)
FetchBench: A Simulation Benchmark for Robot Fetching
by: Han, Beining, et al.
Published: (2024)
by: Han, Beining, et al.
Published: (2024)
FetchBot: Learning Generalizable Object Fetching in Cluttered Scenes via Zero-Shot Sim2Real
by: Liu, Weiheng, et al.
Published: (2025)
by: Liu, Weiheng, et al.
Published: (2025)
Enhancing Exploratory Capability of Visual Navigation Using Uncertainty of Implicit Scene Representation
by: Wang, Yichen, et al.
Published: (2024)
by: Wang, Yichen, et al.
Published: (2024)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
by: Shi, Junyao, et al.
Published: (2024)
by: Shi, Junyao, et al.
Published: (2024)
VANP: Learning Where to See for Navigation with Self-Supervised Vision-Action Pre-Training
by: Nazeri, Mohammad, et al.
Published: (2024)
by: Nazeri, Mohammad, et al.
Published: (2024)
SpatialAnt: Autonomous Zero-Shot Robot Navigation via Active Scene Reconstruction and Visual Anticipation
by: Zhang, Jiwen, et al.
Published: (2026)
by: Zhang, Jiwen, et al.
Published: (2026)
Pre-Trained Masked Image Model for Mobile Robot Navigation
by: Sharma, Vishnu Dutt, et al.
Published: (2023)
by: Sharma, Vishnu Dutt, et al.
Published: (2023)
Open Scene Graphs for Open World Object-Goal Navigation
by: Loo, Joel, et al.
Published: (2024)
by: Loo, Joel, et al.
Published: (2024)
Open Scene Graphs for Open-World Object-Goal Navigation
by: Loo, Joel, et al.
Published: (2025)
by: Loo, Joel, et al.
Published: (2025)
Meta-reasoning Using Attention Maps and Its Applications in Cloud Robotics
by: Lendinez, Adrian, et al.
Published: (2025)
by: Lendinez, Adrian, et al.
Published: (2025)
Safe Planner: Empowering Safety Awareness in Large Pre-Trained Models for Robot Task Planning
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
by: Li, Xinting, et al.
Published: (2023)
by: Li, Xinting, et al.
Published: (2023)
FSUNav: A Cerebrum-Cerebellum Architecture for Fast, Safe, and Universal Zero-Shot Goal-Oriented Navigation
by: Tan, Mingao, et al.
Published: (2026)
by: Tan, Mingao, et al.
Published: (2026)
Self-Predictive Representation for Autonomous UAV Object-Goal Navigation
by: Ayala, Angel, et al.
Published: (2026)
by: Ayala, Angel, et al.
Published: (2026)
SignScene: Visual Sign Grounding for Mapless Navigation
by: Zimmerman, Nicky, et al.
Published: (2026)
by: Zimmerman, Nicky, et al.
Published: (2026)
Advancing Object Goal Navigation Through LLM-enhanced Object Affinities Transfer
by: Lin, Mengying, et al.
Published: (2024)
by: Lin, Mengying, et al.
Published: (2024)
MSGField: A Unified Scene Representation Integrating Motion, Semantics, and Geometry for Robotic Manipulation
by: Sheng, Yu, et al.
Published: (2024)
by: Sheng, Yu, et al.
Published: (2024)
Robotic Scene Cloning:Advancing Zero-Shot Robotic Scene Adaptation in Manipulation via Visual Prompt Editing
by: Huang, Binyuan, et al.
Published: (2026)
by: Huang, Binyuan, et al.
Published: (2026)
GA-TEB: Goal-Adaptive Framework for Efficient Navigation Based on Goal Lines
by: Zhang, Qianyi, et al.
Published: (2024)
by: Zhang, Qianyi, et al.
Published: (2024)
MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding
by: Liang, Jing, et al.
Published: (2025)
by: Liang, Jing, et al.
Published: (2025)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
by: Wang, Zhaowei, et al.
Published: (2024)
by: Wang, Zhaowei, et al.
Published: (2024)
VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation
by: Zhao, Wentao, et al.
Published: (2024)
by: Zhao, Wentao, et al.
Published: (2024)
Immersive Explainability: Visualizing Robot Navigation Decisions through XAI Semantic Scene Projections in Virtual Reality
by: de Heuvel, Jorge, et al.
Published: (2025)
by: de Heuvel, Jorge, et al.
Published: (2025)
DRL-TH: Jointly Utilizing Temporal Graph Attention and Hierarchical Fusion for UGV Navigation in Crowded Environments
by: Li, Ruitong, et al.
Published: (2025)
by: Li, Ruitong, et al.
Published: (2025)
Efficient Image-Goal Navigation with Representative Latent World Model
by: Zhang, Zhiwei, et al.
Published: (2025)
by: Zhang, Zhiwei, et al.
Published: (2025)
MCNav: Memory-Aware Dynamic Cognitive Map for Zero-shot Goal-oriented Navigation
by: Li, Jingyu, et al.
Published: (2026)
by: Li, Jingyu, et al.
Published: (2026)
CANINE: Coaching Visually Impaired Users for Interactive Navigation with a Robot Guide Dog
by: Yu, Cunjun, et al.
Published: (2026)
by: Yu, Cunjun, et al.
Published: (2026)
THUD++: Large-Scale Dynamic Indoor Scene Dataset and Benchmark for Mobile Robots
by: Li, Zeshun, et al.
Published: (2024)
by: Li, Zeshun, et al.
Published: (2024)
CeRLP: A Cross-embodiment Robot Local Planning Framework for Visual Navigation
by: Xi, Haoyu, et al.
Published: (2026)
by: Xi, Haoyu, et al.
Published: (2026)
SplatSearch: Instance Image Goal Navigation for Mobile Robots using 3D Gaussian Splatting and Diffusion Models
by: Narasimhan, Siddarth, et al.
Published: (2025)
by: Narasimhan, Siddarth, et al.
Published: (2025)
Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
by: Li, Badi, et al.
Published: (2025)
by: Li, Badi, et al.
Published: (2025)
What Is The Best 3D Scene Representation for Robotics? From Geometric to Foundation Models
by: Deng, Tianchen, et al.
Published: (2025)
by: Deng, Tianchen, et al.
Published: (2025)
MagRobot:An Open Simulator for Magnetically Navigated Robots
by: Wang, Heng, et al.
Published: (2026)
by: Wang, Heng, et al.
Published: (2026)
MORN: Metacognitive Object-Goal Regulation for Resource-Rational Long-Horizon Navigation
by: Lin, Xi, et al.
Published: (2026)
by: Lin, Xi, et al.
Published: (2026)
Co-NavGPT: Multi-Robot Cooperative Visual Semantic Navigation Using Vision Language Models
by: Yu, Bangguo, et al.
Published: (2023)
by: Yu, Bangguo, et al.
Published: (2023)
EfficientNav: Towards On-Device Object-Goal Navigation with Navigation Map Caching and Retrieval
by: Yang, Zebin, et al.
Published: (2025)
by: Yang, Zebin, et al.
Published: (2025)
Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets
by: Jiang, Guangqi, et al.
Published: (2024)
by: Jiang, Guangqi, et al.
Published: (2024)
Resolving Positional Ambiguity in Dialogues by Vision-Language Models for Robot Navigation
by: Chen, Kuan-Lin, et al.
Published: (2024)
by: Chen, Kuan-Lin, et al.
Published: (2024)
Similar Items
-
MPGNet: Learning Move-Push-Grasping Synergy for Target-Oriented Grasping in Occluded Scenes
by: Li, Dayou, et al.
Published: (2024) -
Learning Instruction-Guided Manipulation Affordance via Large Models for Embodied Robotic Tasks
by: Li, Dayou, et al.
Published: (2024) -
FetchBench: A Simulation Benchmark for Robot Fetching
by: Han, Beining, et al.
Published: (2024) -
FetchBot: Learning Generalizable Object Fetching in Cluttered Scenes via Zero-Shot Sim2Real
by: Liu, Weiheng, et al.
Published: (2025) -
Enhancing Exploratory Capability of Visual Navigation Using Uncertainty of Implicit Scene Representation
by: Wang, Yichen, et al.
Published: (2024)