RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
Fuente:
arXiv
Saved in:
| Main Authors: | Garg, Sourav, Rana, Krishan, Hosseinzadeh, Mehdi, Mares, Lachlan, Sünderhauf, Niko, Dayoub, Feras, Reid, Ian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals
by: Podgorski, Stefan, et al.
Published: (2025)
by: Podgorski, Stefan, et al.
Published: (2025)
Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for Robotics
by: Abou-Chakra, Jad, et al.
Published: (2024)
by: Abou-Chakra, Jad, et al.
Published: (2024)
ObjectReact: Learning Object-Relative Control for Visual Navigation
by: Garg, Sourav, et al.
Published: (2025)
by: Garg, Sourav, et al.
Published: (2025)
To Ask or Not to Ask? Detecting Absence of Information in Vision and Language Navigation
by: Abraham, Savitha Sam, et al.
Published: (2024)
by: Abraham, Savitha Sam, et al.
Published: (2024)
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
by: Wang, Wenze, et al.
Published: (2026)
by: Wang, Wenze, et al.
Published: (2026)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
by: Hosseinzadeh, Mehdi, et al.
Published: (2026)
by: Hosseinzadeh, Mehdi, et al.
Published: (2026)
SceneEdited: A City-Scale Benchmark for 3D HD Map Updating via Image-Guided Change Detection
by: Lin, Chun-Jung, et al.
Published: (2025)
by: Lin, Chun-Jung, et al.
Published: (2025)
Robust Scene Change Detection Using Visual Foundation Models and Cross-Attention Mechanisms
by: Lin, Chun-Jung, et al.
Published: (2024)
by: Lin, Chun-Jung, et al.
Published: (2024)
BEVPose: Unveiling Scene Semantics through Pose-Guided Multi-Modal BEV Alignment
by: Hosseinzadeh, Mehdi, et al.
Published: (2024)
by: Hosseinzadeh, Mehdi, et al.
Published: (2024)
Learn 2 Rage: Experiencing The Emotional Roller Coaster That Is Reinforcement Learning
by: Mares, Lachlan, et al.
Published: (2024)
by: Mares, Lachlan, et al.
Published: (2024)
LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation
by: Miao, Bo, et al.
Published: (2026)
by: Miao, Bo, et al.
Published: (2026)
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer
by: Deng, Jiajun, et al.
Published: (2025)
by: Deng, Jiajun, et al.
Published: (2025)
PoIFusion: Multi-Modal 3D Object Detection via Fusion at Points of Interest
by: Deng, Jiajun, et al.
Published: (2024)
by: Deng, Jiajun, et al.
Published: (2024)
Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams
by: Holden, Lachlan, et al.
Published: (2026)
by: Holden, Lachlan, et al.
Published: (2026)
Segment Beyond View: Handling Partially Missing Modality for Audio-Visual Semantic Segmentation
by: Wu, Renjie, et al.
Published: (2023)
by: Wu, Renjie, et al.
Published: (2023)
Temporal Attention for Cross-View Sequential Image Localization
by: Yuan, Dong, et al.
Published: (2024)
by: Yuan, Dong, et al.
Published: (2024)
Wasserstein Distance-based Expansion of Low-Density Latent Regions for Unknown Class Detection
by: Mallick, Prakash, et al.
Published: (2024)
by: Mallick, Prakash, et al.
Published: (2024)
LHManip: A Dataset for Long-Horizon Language-Grounded Manipulation Tasks in Cluttered Tabletop Environments
by: Ceola, Federico, et al.
Published: (2023)
by: Ceola, Federico, et al.
Published: (2023)
Open-Set Recognition in the Age of Vision-Language Models
by: Miller, Dimity, et al.
Published: (2024)
by: Miller, Dimity, et al.
Published: (2024)
Learning from 10 Demos: Generalisable and Sample-Efficient Policy Learning with Oriented Affordance Frames
by: Rana, Krishan, et al.
Published: (2024)
by: Rana, Krishan, et al.
Published: (2024)
Detecting Precise Hand Touch Moments in Egocentric Video
by: Nguyen, Huy Anh, et al.
Published: (2026)
by: Nguyen, Huy Anh, et al.
Published: (2026)
G3Splat: Geometrically Consistent Generalizable Gaussian Splatting
by: Hosseinzadeh, Mehdi, et al.
Published: (2025)
by: Hosseinzadeh, Mehdi, et al.
Published: (2025)
QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries
by: Chapman, Nicolas Harvey, et al.
Published: (2025)
by: Chapman, Nicolas Harvey, et al.
Published: (2025)
Embodied Domain Adaptation for Object Detection
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
Multi-Modal 3D Scene Graph Updater for Shared and Dynamic Environments
by: Olivastri, Emilio, et al.
Published: (2024)
by: Olivastri, Emilio, et al.
Published: (2024)
AIMC-Spec: A Benchmark Dataset for Automatic Intrapulse Modulation Classification under Variable Noise Conditions
by: Cocks, Sebastian L., et al.
Published: (2026)
by: Cocks, Sebastian L., et al.
Published: (2026)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
Improving Online Source-free Domain Adaptation for Object Detection by Unsupervised Data Acquisition
by: Shi, Xiangyu, et al.
Published: (2023)
by: Shi, Xiangyu, et al.
Published: (2023)
QueSTMaps: Queryable Semantic Topological Maps for 3D Scene Understanding
by: Mehan, Yash, et al.
Published: (2024)
by: Mehan, Yash, et al.
Published: (2024)
Category Level 6D Object Pose Estimation from a Single RGB Image using Diffusion
by: Bethell, Adam, et al.
Published: (2024)
by: Bethell, Adam, et al.
Published: (2024)
AARK: An Open Toolkit for Autonomous Racing Research
by: Bockman, James, et al.
Published: (2024)
by: Bockman, James, et al.
Published: (2024)
Changes in Real Time: Online Scene Change Detection with Multi-View Fusion
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2025)
by: Galappaththige, Chamuditha Jayanga, et al.
Published: (2025)
"Which LLM should I use?": Evaluating LLMs for tasks performed by Undergraduate Computer Science Students
by: Agarwal, Vibhor, et al.
Published: (2024)
by: Agarwal, Vibhor, et al.
Published: (2024)
GeoWorld: Geometric World Models
by: Zhang, Zeyu, et al.
Published: (2026)
by: Zhang, Zeyu, et al.
Published: (2026)
ItTakesTwo: Leveraging Peer Representations for Semi-supervised LiDAR Semantic Segmentation
by: Liu, Yuyuan, et al.
Published: (2024)
by: Liu, Yuyuan, et al.
Published: (2024)
Team Samsung-RAL: Technical Report for 2024 RoboDrive Challenge-Robust Map Segmentation Track
by: Hao, Xiaoshuai, et al.
Published: (2024)
by: Hao, Xiaoshuai, et al.
Published: (2024)
SegMASt3R: Geometry Grounded Segment Matching
by: Jayanti, Rohit, et al.
Published: (2025)
by: Jayanti, Rohit, et al.
Published: (2025)
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
by: Ehsanpour, Mahsa, et al.
Published: (2024)
by: Ehsanpour, Mahsa, et al.
Published: (2024)
Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation
by: Dai, Yinpei, et al.
Published: (2023)
by: Dai, Yinpei, et al.
Published: (2023)
Open-World Panoptic Segmentation
by: Sodano, Matteo, et al.
Published: (2024)
by: Sodano, Matteo, et al.
Published: (2024)
Similar Items
-
TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals
by: Podgorski, Stefan, et al.
Published: (2025) -
Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for Robotics
by: Abou-Chakra, Jad, et al.
Published: (2024) -
ObjectReact: Learning Object-Relative Control for Visual Navigation
by: Garg, Sourav, et al.
Published: (2025) -
To Ask or Not to Ask? Detecting Absence of Information in Vision and Language Navigation
by: Abraham, Savitha Sam, et al.
Published: (2024) -
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
by: Wang, Wenze, et al.
Published: (2026)