Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams
Fuente:
arXiv
Salvato in:
| Autori principali: | Holden, Lachlan, Dayoub, Feras, Candela, Alberto, Harvey, David, Chin, Tat-Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries
di: Chapman, Nicolas Harvey, et al.
Pubblicazione: (2025)
di: Chapman, Nicolas Harvey, et al.
Pubblicazione: (2025)
Robust Scene Change Detection Using Visual Foundation Models and Cross-Attention Mechanisms
di: Lin, Chun-Jung, et al.
Pubblicazione: (2024)
di: Lin, Chun-Jung, et al.
Pubblicazione: (2024)
Embodied Domain Adaptation for Object Detection
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
di: Hosseinzadeh, Mehdi, et al.
Pubblicazione: (2026)
di: Hosseinzadeh, Mehdi, et al.
Pubblicazione: (2026)
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
di: Wang, Wenze, et al.
Pubblicazione: (2026)
di: Wang, Wenze, et al.
Pubblicazione: (2026)
Enhancing Embodied Object Detection through Language-Image Pre-training and Implicit Object Memory
di: Chapman, Nicolas Harvey, et al.
Pubblicazione: (2024)
di: Chapman, Nicolas Harvey, et al.
Pubblicazione: (2024)
Temporal Attention for Cross-View Sequential Image Localization
di: Yuan, Dong, et al.
Pubblicazione: (2024)
di: Yuan, Dong, et al.
Pubblicazione: (2024)
TANGO: Traversability-Aware Navigation with Local Metric Control for Topological Goals
di: Podgorski, Stefan, et al.
Pubblicazione: (2025)
di: Podgorski, Stefan, et al.
Pubblicazione: (2025)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
di: Shi, Xiangyu, et al.
Pubblicazione: (2025)
SceneEdited: A City-Scale Benchmark for 3D HD Map Updating via Image-Guided Change Detection
di: Lin, Chun-Jung, et al.
Pubblicazione: (2025)
di: Lin, Chun-Jung, et al.
Pubblicazione: (2025)
Physically Embodied Gaussian Splatting: A Realtime Correctable World Model for Robotics
di: Abou-Chakra, Jad, et al.
Pubblicazione: (2024)
di: Abou-Chakra, Jad, et al.
Pubblicazione: (2024)
RoboHop: Segment-based Topological Map Representation for Open-World Visual Navigation
di: Garg, Sourav, et al.
Pubblicazione: (2024)
di: Garg, Sourav, et al.
Pubblicazione: (2024)
MAG-VLAQ: Multi-modal Aerial-Ground Query Aggregation for Cross-View Place Recognition
di: Xu, Zhengyi, et al.
Pubblicazione: (2026)
di: Xu, Zhengyi, et al.
Pubblicazione: (2026)
Towards Bridging the Space Domain Gap for Satellite Pose Estimation using Event Sensing
di: Jawaid, Mohsi, et al.
Pubblicazione: (2022)
di: Jawaid, Mohsi, et al.
Pubblicazione: (2022)
ObjectReact: Learning Object-Relative Control for Visual Navigation
di: Garg, Sourav, et al.
Pubblicazione: (2025)
di: Garg, Sourav, et al.
Pubblicazione: (2025)
LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation
di: Miao, Bo, et al.
Pubblicazione: (2026)
di: Miao, Bo, et al.
Pubblicazione: (2026)
Event-RGB Fusion for Spacecraft Pose Estimation Under Harsh Lighting
di: Jawaid, Mohsi, et al.
Pubblicazione: (2025)
di: Jawaid, Mohsi, et al.
Pubblicazione: (2025)
Folding Knots Using a Team of Aerial Robots
di: D'Antonio, Diego S., et al.
Pubblicazione: (2022)
di: D'Antonio, Diego S., et al.
Pubblicazione: (2022)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
di: Huang, Haifeng, et al.
Pubblicazione: (2025)
di: Huang, Haifeng, et al.
Pubblicazione: (2025)
HOTFormerLoc: Hierarchical Octree Transformer for Versatile Lidar Place Recognition Across Ground and Aerial Views
di: Griffiths, Ethan, et al.
Pubblicazione: (2025)
di: Griffiths, Ethan, et al.
Pubblicazione: (2025)
Semantic-Aware Particle Filter for Reliable Vineyard Robot Localisation
di: de Silva, Rajitha, et al.
Pubblicazione: (2025)
di: de Silva, Rajitha, et al.
Pubblicazione: (2025)
GA3T: A Ground-Aerial Terrain Traversability Dataset for Heterogeneous Robot Teams in Unstructured Environments
di: Cai, Siwei, et al.
Pubblicazione: (2026)
di: Cai, Siwei, et al.
Pubblicazione: (2026)
AARK: An Open Toolkit for Autonomous Racing Research
di: Bockman, James, et al.
Pubblicazione: (2024)
di: Bockman, James, et al.
Pubblicazione: (2024)
To Ask or Not to Ask? Detecting Absence of Information in Vision and Language Navigation
di: Abraham, Savitha Sam, et al.
Pubblicazione: (2024)
di: Abraham, Savitha Sam, et al.
Pubblicazione: (2024)
Multi-vision-based Picking Point Localisation of Target Fruit for Harvesting Robots
di: Beldek, C., et al.
Pubblicazione: (2025)
di: Beldek, C., et al.
Pubblicazione: (2025)
Collaborative Exploration with a Marsupial Ground-Aerial Robot Team through Task-Driven Map Compression
di: Zacharia, Angelos, et al.
Pubblicazione: (2025)
di: Zacharia, Angelos, et al.
Pubblicazione: (2025)
A Novel Methodology for Autonomous Planetary Exploration Using Multi-Robot Teams
di: Swinton, Sarah, et al.
Pubblicazione: (2024)
di: Swinton, Sarah, et al.
Pubblicazione: (2024)
Paired-CSLiDAR: Height-Stratified Registration for Cross-Source Aerial-Ground LiDAR Pose Refinement
di: Hoover, Montana, et al.
Pubblicazione: (2026)
di: Hoover, Montana, et al.
Pubblicazione: (2026)
ReVLA: Reverting Visual Domain Limitation of Robotic Foundation Models
di: Dey, Sombit, et al.
Pubblicazione: (2024)
di: Dey, Sombit, et al.
Pubblicazione: (2024)
Effective Tuning Strategies for Generalist Robot Manipulation Policies
di: Zhang, Wenbo, et al.
Pubblicazione: (2024)
di: Zhang, Wenbo, et al.
Pubblicazione: (2024)
RobotPan: A 360$^\circ$ Surround-View Robotic Vision System for Embodied Perception
di: Ma, Jiahao, et al.
Pubblicazione: (2026)
di: Ma, Jiahao, et al.
Pubblicazione: (2026)
Griffin: Aerial-Ground Cooperative Detection and Tracking Dataset and Benchmark
di: Wang, Jiahao, et al.
Pubblicazione: (2025)
di: Wang, Jiahao, et al.
Pubblicazione: (2025)
Test-Time Certifiable Self-Supervision to Bridge the Sim2Real Gap in Event-Based Satellite Pose Estimation
di: Jawaid, Mohsi, et al.
Pubblicazione: (2024)
di: Jawaid, Mohsi, et al.
Pubblicazione: (2024)
On the Application of Efficient Neural Mapping to Real-Time Indoor Localisation for Unmanned Ground Vehicles
di: Holder, Christopher J., et al.
Pubblicazione: (2022)
di: Holder, Christopher J., et al.
Pubblicazione: (2022)
Scalable Aerial GNSS Localization for Marine Robots
di: Wen, Shuo, et al.
Pubblicazione: (2025)
di: Wen, Shuo, et al.
Pubblicazione: (2025)
Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration
di: Wang, Jun, et al.
Pubblicazione: (2026)
di: Wang, Jun, et al.
Pubblicazione: (2026)
Physically Grounded Vision-Language Models for Robotic Manipulation
di: Gao, Jensen, et al.
Pubblicazione: (2023)
di: Gao, Jensen, et al.
Pubblicazione: (2023)
RoboLLM: Robotic Vision Tasks Grounded on Multimodal Large Language Models
di: Long, Zijun, et al.
Pubblicazione: (2023)
di: Long, Zijun, et al.
Pubblicazione: (2023)
Australian Supermarket Object Set (ASOS): A Benchmark Dataset of Physical Objects and 3D Models for Robotics and Computer Vision
di: Cosgun, Akansel, et al.
Pubblicazione: (2025)
di: Cosgun, Akansel, et al.
Pubblicazione: (2025)
Decentralized Uncertainty-Aware Active Search with a Team of Aerial Robots
di: Tabib, Wennie, et al.
Pubblicazione: (2024)
di: Tabib, Wennie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
QueryAdapter: Rapid Adaptation of Vision-Language Models in Response to Natural Language Queries
di: Chapman, Nicolas Harvey, et al.
Pubblicazione: (2025) -
Robust Scene Change Detection Using Visual Foundation Models and Cross-Attention Mechanisms
di: Lin, Chun-Jung, et al.
Pubblicazione: (2024) -
Embodied Domain Adaptation for Object Detection
di: Shi, Xiangyu, et al.
Pubblicazione: (2025) -
KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis
di: Hosseinzadeh, Mehdi, et al.
Pubblicazione: (2026) -
A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring
di: Wang, Wenze, et al.
Pubblicazione: (2026)