Autonomous Frontier-Based Exploration with VLM Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Aitha, Aarush, Zakhor, Avideh |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Semantic-Aware Guided Drone Exploration for Language-Conditioned 3D Indoor Mapping
by: Vegesna, Nitin, et al.
Published: (2026)
by: Vegesna, Nitin, et al.
Published: (2026)
Adapting Segment Anything Model to Melanoma Segmentation in Microscopy Slide Images
by: Liu, Qingyuan, et al.
Published: (2024)
by: Liu, Qingyuan, et al.
Published: (2024)
OPAL: Omnidirectional Path-efficient Aerial 3D expLoration
by: Chappidi, Yoga Satwik, et al.
Published: (2026)
by: Chappidi, Yoga Satwik, et al.
Published: (2026)
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
by: Song, Sangmim, et al.
Published: (2024)
by: Song, Sangmim, et al.
Published: (2024)
Simulation to Rules: A Dual-VLM Framework for Formal Visual Planning
by: Hao, Yilun, et al.
Published: (2025)
by: Hao, Yilun, et al.
Published: (2025)
Autonomous Implicit Indoor Scene Reconstruction with Frontier Exploration
by: Zeng, Jing, et al.
Published: (2024)
by: Zeng, Jing, et al.
Published: (2024)
Developing Autonomous Robot-Mediated Behavior Coaching Sessions with Haru
by: Jelínek, Matouš, et al.
Published: (2024)
by: Jelínek, Matouš, et al.
Published: (2024)
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
by: Li, Pengxiang, et al.
Published: (2025)
by: Li, Pengxiang, et al.
Published: (2025)
InfraGPT Smart Infrastructure: An End-to-End VLM-Based Framework for Detecting and Managing Urban Defects
by: Mohamed, Ibrahim Sheikh, et al.
Published: (2025)
by: Mohamed, Ibrahim Sheikh, et al.
Published: (2025)
Can an Embodied Agent Find Your "Cat-shaped Mug"? LLM-Guided Exploration for Zero-Shot Object Navigation
by: Dorbala, Vishnu Sashank, et al.
Published: (2023)
by: Dorbala, Vishnu Sashank, et al.
Published: (2023)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
by: Chen, Kaiyuan, et al.
Published: (2025)
by: Chen, Kaiyuan, et al.
Published: (2025)
VLURes: Benchmarking VLM Visual and Linguistic Understanding in Low-Resource Languages
by: Atuhurra, Jesse, et al.
Published: (2025)
by: Atuhurra, Jesse, et al.
Published: (2025)
GMLM: Bridging Graph Neural Networks and Language Models for Heterophilic Node Classification
by: Sinha, Aarush
Published: (2025)
by: Sinha, Aarush
Published: (2025)
VERDI: VLM-Embedded Reasoning for Autonomous Driving
by: Feng, Bowen, et al.
Published: (2025)
by: Feng, Bowen, et al.
Published: (2025)
Semantic-Drive: Democratizing Long-Tail Data Curation via Open-Vocabulary Grounding and Neuro-Symbolic VLM Consensus
by: Guillen-Perez, Antonio
Published: (2025)
by: Guillen-Perez, Antonio
Published: (2025)
Enhancing Surgical Robots with Embodied Intelligence for Autonomous Ultrasound Scanning
by: Xu, Huan, et al.
Published: (2024)
by: Xu, Huan, et al.
Published: (2024)
Code-as-Symbolic-Planner: Foundation Model-Based Robot Planning via Symbolic Code Generation
by: Chen, Yongchao, et al.
Published: (2025)
by: Chen, Yongchao, et al.
Published: (2025)
Next-Generation LLM for UAV: From Natural Language to Autonomous Flight
by: Yuan, Liangqi, et al.
Published: (2025)
by: Yuan, Liangqi, et al.
Published: (2025)
Embodied Agents for Efficient Exploration and Smart Scene Description
by: Bigazzi, Roberto, et al.
Published: (2023)
by: Bigazzi, Roberto, et al.
Published: (2023)
A Language Agent for Autonomous Driving
by: Mao, Jiageng, et al.
Published: (2023)
by: Mao, Jiageng, et al.
Published: (2023)
Vision-Guided Outdoor Flight and Obstacle Evasion via Reinforcement Learning
by: Dutta, Shiladitya, et al.
Published: (2026)
by: Dutta, Shiladitya, et al.
Published: (2026)
Generative AI for Autonomous Driving: Frontiers and Opportunities
by: Wang, Yuping, et al.
Published: (2025)
by: Wang, Yuping, et al.
Published: (2025)
Diffusion-ES: Gradient-free Planning with Diffusion for Autonomous Driving and Zero-Shot Instruction Following
by: Yang, Brian, et al.
Published: (2024)
by: Yang, Brian, et al.
Published: (2024)
IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks
by: Lu, Xiaoya, et al.
Published: (2025)
by: Lu, Xiaoya, et al.
Published: (2025)
LLM4AD: Large Language Models for Autonomous Driving -- Concept, Review, Benchmark, Experiments, and Future Trends
by: Cui, Can, et al.
Published: (2024)
by: Cui, Can, et al.
Published: (2024)
See, Point, Fly: A Learning-Free VLM Framework for Universal Unmanned Aerial Navigation
by: Hu, Chih Yao, et al.
Published: (2025)
by: Hu, Chih Yao, et al.
Published: (2025)
Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
by: Agarwal, Rajan, et al.
Published: (2025)
by: Agarwal, Rajan, et al.
Published: (2025)
Hybrid Decision Making via Conformal VLM-generated Guidance
by: Banerjee, Debodeep, et al.
Published: (2026)
by: Banerjee, Debodeep, et al.
Published: (2026)
Fake or Real, Can Robots Tell? Evaluating VLM Robustness to Domain Shift in Single-View Robotic Scene Understanding
by: Tavella, Federico, et al.
Published: (2025)
by: Tavella, Federico, et al.
Published: (2025)
Work Zones challenge VLM Trajectory Planning: Toward Mitigation and Robust Autonomous Driving
by: Liao, Yifan, et al.
Published: (2025)
by: Liao, Yifan, et al.
Published: (2025)
Indoor Asset Detection in Large Scale 360° Drone-Captured Imagery via 3D Gaussian Splatting
by: Tang, Monica, et al.
Published: (2026)
by: Tang, Monica, et al.
Published: (2026)
Swap Path Network for Robust Person Search Pre-training
by: Jaffe, Lucas, et al.
Published: (2024)
by: Jaffe, Lucas, et al.
Published: (2024)
CADENet: Condition-Adaptive Asynchronous Dual-Stream Enhancement Network for Adverse Weather Perception in Autonomous Driving
by: Khairy, Sherif, et al.
Published: (2026)
by: Khairy, Sherif, et al.
Published: (2026)
X-Blocks: Linguistic Building Blocks of Natural Language Explanations for Automated Vehicles
by: Zadeh, Ashkan Y., et al.
Published: (2026)
by: Zadeh, Ashkan Y., et al.
Published: (2026)
SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models
by: Torres-Fonseca, Josue, et al.
Published: (2026)
by: Torres-Fonseca, Josue, et al.
Published: (2026)
Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents
by: Wang, Hanlin, et al.
Published: (2026)
by: Wang, Hanlin, et al.
Published: (2026)
Limited Linguistic Diversity in Embodied AI Datasets
by: Wanna, Selma, et al.
Published: (2026)
by: Wanna, Selma, et al.
Published: (2026)
Red-Teaming Vision-Language-Action Models via Quality Diversity Prompt Generation for Robust Robot Policies
by: Srikanth, Siddharth, et al.
Published: (2026)
by: Srikanth, Siddharth, et al.
Published: (2026)
Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform
by: Alaswad, Feisal, et al.
Published: (2026)
by: Alaswad, Feisal, et al.
Published: (2026)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
by: Wang, Qiuyue, et al.
Published: (2026)
by: Wang, Qiuyue, et al.
Published: (2026)
Similar Items
-
Semantic-Aware Guided Drone Exploration for Language-Conditioned 3D Indoor Mapping
by: Vegesna, Nitin, et al.
Published: (2026) -
Adapting Segment Anything Model to Melanoma Segmentation in Microscopy Slide Images
by: Liu, Qingyuan, et al.
Published: (2024) -
OPAL: Omnidirectional Path-efficient Aerial 3D expLoration
by: Chappidi, Yoga Satwik, et al.
Published: (2026) -
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
by: Song, Sangmim, et al.
Published: (2024) -
Simulation to Rules: A Dual-VLM Framework for Formal Visual Planning
by: Hao, Yilun, et al.
Published: (2025)