RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Garg, Sourav, Rana, Krishan, Hosseinzadeh, Mehdi, Mares, Lachlan, Sünderhauf, Niko, Dayoub, Feras, Reid, Ian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals
von: Podgorski, Stefan, et al.
Veröffentlicht: (2025)
von: Podgorski, Stefan, et al.
Veröffentlicht: (2025)
Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for Robotics
von: Abou-Chakra, Jad, et al.
Veröffentlicht: (2024)
von: Abou-Chakra, Jad, et al.
Veröffentlicht: (2024)
ObjectReact: Learning Object-Relative Control for Visual Navigation
von: Garg, Sourav, et al.
Veröffentlicht: (2025)
von: Garg, Sourav, et al.
Veröffentlicht: (2025)
To Ask or Not to Ask? Detecting Absence of Information in Vision and Language Navigation
von: Abraham, Savitha Sam, et al.
Veröffentlicht: (2024)
von: Abraham, Savitha Sam, et al.
Veröffentlicht: (2024)
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
von: Wang, Wenze, et al.
Veröffentlicht: (2026)
von: Wang, Wenze, et al.
Veröffentlicht: (2026)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2026)
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2026)
SceneEdited: A City-Scale Benchmark for 3D HD Map Updating via Image-Guided Change Detection
von: Lin, Chun-Jung, et al.
Veröffentlicht: (2025)
von: Lin, Chun-Jung, et al.
Veröffentlicht: (2025)
Robust Scene Change Detection Using Visual Foundation Models and Cross-Attention Mechanisms
von: Lin, Chun-Jung, et al.
Veröffentlicht: (2024)
von: Lin, Chun-Jung, et al.
Veröffentlicht: (2024)
BEVPose: Unveiling Scene Semantics through Pose-Guided Multi-Modal BEV Alignment
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2024)
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2024)
Learn 2 Rage: Experiencing The Emotional Roller Coaster That Is Reinforcement Learning
von: Mares, Lachlan, et al.
Veröffentlicht: (2024)
von: Mares, Lachlan, et al.
Veröffentlicht: (2024)
LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation
von: Miao, Bo, et al.
Veröffentlicht: (2026)
von: Miao, Bo, et al.
Veröffentlicht: (2026)
3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
von: Deng, Jiajun, et al.
Veröffentlicht: (2025)
PoIFusion: Multi-Modal 3D Object Detection via Fusion at Points of Interest
von: Deng, Jiajun, et al.
Veröffentlicht: (2024)
von: Deng, Jiajun, et al.
Veröffentlicht: (2024)
Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams
von: Holden, Lachlan, et al.
Veröffentlicht: (2026)
von: Holden, Lachlan, et al.
Veröffentlicht: (2026)
Segment Beyond View: Handling Partially Missing Modality for Audio-Visual Semantic Segmentation
von: Wu, Renjie, et al.
Veröffentlicht: (2023)
von: Wu, Renjie, et al.
Veröffentlicht: (2023)
Temporal Attention for Cross-View Sequential Image Localization
von: Yuan, Dong, et al.
Veröffentlicht: (2024)
von: Yuan, Dong, et al.
Veröffentlicht: (2024)
Wasserstein Distance-based Expansion of Low-Density Latent Regions for Unknown Class Detection
von: Mallick, Prakash, et al.
Veröffentlicht: (2024)
von: Mallick, Prakash, et al.
Veröffentlicht: (2024)
LHManip: A Dataset for Long-Horizon Language-Grounded Manipulation Tasks in Cluttered Tabletop Environments
von: Ceola, Federico, et al.
Veröffentlicht: (2023)
von: Ceola, Federico, et al.
Veröffentlicht: (2023)
Open-Set Recognition in the Age of Vision-Language Models
von: Miller, Dimity, et al.
Veröffentlicht: (2024)
von: Miller, Dimity, et al.
Veröffentlicht: (2024)
Learning from 10 Demos: Generalisable and Sample-Efficient Policy Learning with Oriented Affordance Frames
von: Rana, Krishan, et al.
Veröffentlicht: (2024)
von: Rana, Krishan, et al.
Veröffentlicht: (2024)
Detecting Precise Hand Touch Moments in Egocentric Video
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026)
von: Nguyen, Huy Anh, et al.
Veröffentlicht: (2026)
G3Splat: Geometrically Consistent Generalizable Gaussian Splatting
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2025)
von: Hosseinzadeh, Mehdi, et al.
Veröffentlicht: (2025)
QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries
von: Chapman, Nicolas Harvey, et al.
Veröffentlicht: (2025)
von: Chapman, Nicolas Harvey, et al.
Veröffentlicht: (2025)
Embodied Domain Adaptation for Object Detection
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
Multi-Modal 3D Scene Graph Updater for Shared and Dynamic Environments
von: Olivastri, Emilio, et al.
Veröffentlicht: (2024)
von: Olivastri, Emilio, et al.
Veröffentlicht: (2024)
AIMC-Spec: A Benchmark Dataset for Automatic Intrapulse Modulation Classification under Variable Noise Conditions
von: Cocks, Sebastian L., et al.
Veröffentlicht: (2026)
von: Cocks, Sebastian L., et al.
Veröffentlicht: (2026)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2025)
Improving Online Source-free Domain Adaptation for Object Detection by Unsupervised Data Acquisition
von: Shi, Xiangyu, et al.
Veröffentlicht: (2023)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2023)
QueSTMaps: Queryable Semantic Topological Maps for 3D Scene Understanding
von: Mehan, Yash, et al.
Veröffentlicht: (2024)
von: Mehan, Yash, et al.
Veröffentlicht: (2024)
Category Level 6D Object Pose Estimation from a Single RGB Image using Diffusion
von: Bethell, Adam, et al.
Veröffentlicht: (2024)
von: Bethell, Adam, et al.
Veröffentlicht: (2024)
AARK: An Open Toolkit for Autonomous Racing Research
von: Bockman, James, et al.
Veröffentlicht: (2024)
von: Bockman, James, et al.
Veröffentlicht: (2024)
Changes in Real Time: Online Scene Change Detection with Multi-View Fusion
von: Galappaththige, Chamuditha Jayanga, et al.
Veröffentlicht: (2025)
von: Galappaththige, Chamuditha Jayanga, et al.
Veröffentlicht: (2025)
"Which LLM should I use?": Evaluating LLMs for tasks performed by Undergraduate Computer Science Students
von: Agarwal, Vibhor, et al.
Veröffentlicht: (2024)
von: Agarwal, Vibhor, et al.
Veröffentlicht: (2024)
GeoWorld: Geometric World Models
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
ItTakesTwo: Leveraging Peer Representations for Semi-supervised LiDAR Semantic Segmentation
von: Liu, Yuyuan, et al.
Veröffentlicht: (2024)
von: Liu, Yuyuan, et al.
Veröffentlicht: (2024)
Team Samsung-RAL: Technical Report for 2024 RoboDrive Challenge-Robust Map Segmentation Track
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2024)
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2024)
SegMASt3R: Geometry Grounded Segment Matching
von: Jayanti, Rohit, et al.
Veröffentlicht: (2025)
von: Jayanti, Rohit, et al.
Veröffentlicht: (2025)
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
von: Ehsanpour, Mahsa, et al.
Veröffentlicht: (2024)
von: Ehsanpour, Mahsa, et al.
Veröffentlicht: (2024)
Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation
von: Dai, Yinpei, et al.
Veröffentlicht: (2023)
von: Dai, Yinpei, et al.
Veröffentlicht: (2023)
Open-World Panoptic Segmentation
von: Sodano, Matteo, et al.
Veröffentlicht: (2024)
von: Sodano, Matteo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals
von: Podgorski, Stefan, et al.
Veröffentlicht: (2025) -
Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for Robotics
von: Abou-Chakra, Jad, et al.
Veröffentlicht: (2024) -
ObjectReact: Learning Object-Relative Control for Visual Navigation
von: Garg, Sourav, et al.
Veröffentlicht: (2025) -
To Ask or Not to Ask? Detecting Absence of Information in Vision and Language Navigation
von: Abraham, Savitha Sam, et al.
Veröffentlicht: (2024) -
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
von: Wang, Wenze, et al.
Veröffentlicht: (2026)