ImagineUAV: Aerial Vision-Language Navigation via World-Action Modeling and Kinodynamic Planning
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Xuchen, Huang, Jiawei, Xia, Shihao, Liu, Bingxi, Cui, Jinqiang, Yang, Jiankun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CityNavAgent: Aerial Vision-and-Language Navigation with Hierarchical Semantic Planning and Global Memory
di: Zhang, Weichen, et al.
Pubblicazione: (2025)
di: Zhang, Weichen, et al.
Pubblicazione: (2025)
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
di: Zhao, Baining, et al.
Pubblicazione: (2026)
di: Zhao, Baining, et al.
Pubblicazione: (2026)
Kinodynamic Motion Planning for Mobile Robot Navigation across Inconsistent World Models
di: Damm, Eric R., et al.
Pubblicazione: (2025)
di: Damm, Eric R., et al.
Pubblicazione: (2025)
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models
di: Zhang, Qiyao, et al.
Pubblicazione: (2026)
di: Zhang, Qiyao, et al.
Pubblicazione: (2026)
AutoFly: Vision-Language-Action Model for UAV Autonomous Navigation in the Wild
di: Sun, Xiaolou, et al.
Pubblicazione: (2026)
di: Sun, Xiaolou, et al.
Pubblicazione: (2026)
AerialVLA: A Vision-Language-Action Model for UAV Navigation via Minimalist End-to-End Control
di: Xu, Peng, et al.
Pubblicazione: (2026)
di: Xu, Peng, et al.
Pubblicazione: (2026)
VISTAv2: World Imagination for Indoor Vision-and-Language Navigation
di: Huang, Yanjia, et al.
Pubblicazione: (2025)
di: Huang, Yanjia, et al.
Pubblicazione: (2025)
FloorPlan-VLN: A New Paradigm for Floor Plan Guided Vision-Language Navigation
di: Chen, Kehan, et al.
Pubblicazione: (2026)
di: Chen, Kehan, et al.
Pubblicazione: (2026)
VISTA: Generative Visual Imagination for Vision-and-Language Navigation
di: Huang, Yanjia, et al.
Pubblicazione: (2025)
di: Huang, Yanjia, et al.
Pubblicazione: (2025)
Vision-Language Navigation for Aerial Robots: Towards the Era of Large Language Models
di: Xia, Xingyu, et al.
Pubblicazione: (2026)
di: Xia, Xingyu, et al.
Pubblicazione: (2026)
IndoorUAV: Benchmarking Vision-Language UAV Navigation in Continuous Indoor Environments
di: Liu, Xu, et al.
Pubblicazione: (2025)
di: Liu, Xu, et al.
Pubblicazione: (2025)
OpenVLN: Open-world Aerial Vision-Language Navigation
di: Lin, Peican, et al.
Pubblicazione: (2025)
di: Lin, Peican, et al.
Pubblicazione: (2025)
ImagineNav: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
di: Zhao, Xinxin, et al.
Pubblicazione: (2024)
di: Zhao, Xinxin, et al.
Pubblicazione: (2024)
ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
di: Wang, Teng, et al.
Pubblicazione: (2025)
di: Wang, Teng, et al.
Pubblicazione: (2025)
Imaginative World Modeling with Scene Graphs for Embodied Agent Navigation
di: Hu, Yue, et al.
Pubblicazione: (2025)
di: Hu, Yue, et al.
Pubblicazione: (2025)
SwordRiding: A Unified Navigation Framework for Quadrotors in Unknown Complex Environments via Online Guiding Vector Fields
di: Liu, Xuchen, et al.
Pubblicazione: (2025)
di: Liu, Xuchen, et al.
Pubblicazione: (2025)
Planning from Imagination: Episodic Simulation and Episodic Memory for Vision-and-Language Navigation
di: Pan, Yiyuan, et al.
Pubblicazione: (2024)
di: Pan, Yiyuan, et al.
Pubblicazione: (2024)
TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification
di: Tao, Huaqi, et al.
Pubblicazione: (2025)
di: Tao, Huaqi, et al.
Pubblicazione: (2025)
SIMPACT: Simulation-Enabled Action Planning using Vision-Language Models
di: Liu, Haowen, et al.
Pubblicazione: (2025)
di: Liu, Haowen, et al.
Pubblicazione: (2025)
AIR-VLA: Vision-Language-Action Systems for Aerial Manipulation
di: Sun, Jianli, et al.
Pubblicazione: (2026)
di: Sun, Jianli, et al.
Pubblicazione: (2026)
UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents
di: Xiao, Jianqiang, et al.
Pubblicazione: (2025)
di: Xiao, Jianqiang, et al.
Pubblicazione: (2025)
HTNav: A Hybrid Navigation Framework with Tiered Structure for Urban Aerial Vision-and-Language Navigation
di: Fan, Chengjie, et al.
Pubblicazione: (2026)
di: Fan, Chengjie, et al.
Pubblicazione: (2026)
World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems
di: Li, Runze, et al.
Pubblicazione: (2026)
di: Li, Runze, et al.
Pubblicazione: (2026)
VLA-AN: An Efficient and Onboard Vision-Language-Action Framework for Aerial Navigation in Complex Environments
di: Wu, Yuze, et al.
Pubblicazione: (2025)
di: Wu, Yuze, et al.
Pubblicazione: (2025)
Ultrafast Sampling-based Kinodynamic Planning via Differential Flatness
di: Duong, Thai, et al.
Pubblicazione: (2026)
di: Duong, Thai, et al.
Pubblicazione: (2026)
Online Time-Informed Kinodynamic Motion Planning of Nonlinear Systems
di: Meng, Fei, et al.
Pubblicazione: (2024)
di: Meng, Fei, et al.
Pubblicazione: (2024)
Safety Assurance for Quadrotor Kinodynamic Motion Planning
di: Tavoulareas, Theodoros, et al.
Pubblicazione: (2025)
di: Tavoulareas, Theodoros, et al.
Pubblicazione: (2025)
TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic Environments
di: Huang, Zhiyu, et al.
Pubblicazione: (2026)
di: Huang, Zhiyu, et al.
Pubblicazione: (2026)
When to Trust Imagination: Adaptive Action Execution for World Action Models
di: Wang, Rui, et al.
Pubblicazione: (2026)
di: Wang, Rui, et al.
Pubblicazione: (2026)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
Towards Realistic UAV Vision-Language Navigation: Platform, Benchmark, and Methodology
di: Wang, Xiangyu, et al.
Pubblicazione: (2024)
di: Wang, Xiangyu, et al.
Pubblicazione: (2024)
NavThinker: Action-Conditioned World Models for Coupled Prediction and Planning in Social Navigation
di: Hu, Tianshuai, et al.
Pubblicazione: (2026)
di: Hu, Tianshuai, et al.
Pubblicazione: (2026)
NaVILA: Legged Robot Vision-Language-Action Model for Navigation
di: Cheng, An-Chieh, et al.
Pubblicazione: (2024)
di: Cheng, An-Chieh, et al.
Pubblicazione: (2024)
Self-Correcting VLA: Online Action Refinement via Sparse World Imagination
di: Liu, Chenyv, et al.
Pubblicazione: (2026)
di: Liu, Chenyv, et al.
Pubblicazione: (2026)
Dream to Recall: Imagination-Guided Experience Retrieval for Memory-Persistent Vision-and-Language Navigation
di: Xu, Yunzhe, et al.
Pubblicazione: (2025)
di: Xu, Yunzhe, et al.
Pubblicazione: (2025)
Survey of Vision-Language-Action Models for Embodied Manipulation
di: Li, Haoran, et al.
Pubblicazione: (2025)
di: Li, Haoran, et al.
Pubblicazione: (2025)
TrackingMiM: Efficient Mamba-in-Mamba Serialization for Real-time UAV Object Tracking
di: Liu, Bingxi, et al.
Pubblicazione: (2025)
di: Liu, Bingxi, et al.
Pubblicazione: (2025)
Object-Centric Kinodynamic Planning for Nonprehensile Robot Rearrangement Manipulation
di: Ren, Kejia, et al.
Pubblicazione: (2024)
di: Ren, Kejia, et al.
Pubblicazione: (2024)
Accelerating db-A* for Kinodynamic Motion Planning Using Diffusion
di: Franke, Julius, et al.
Pubblicazione: (2025)
di: Franke, Julius, et al.
Pubblicazione: (2025)
Multi-layer Motion Planning with Kinodynamic and Spatio-Temporal Constraints
di: Chatrola, Jeel, et al.
Pubblicazione: (2025)
di: Chatrola, Jeel, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CityNavAgent: Aerial Vision-and-Language Navigation with Hierarchical Semantic Planning and Global Memory
di: Zhang, Weichen, et al.
Pubblicazione: (2025) -
WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation
di: Zhao, Baining, et al.
Pubblicazione: (2026) -
Kinodynamic Motion Planning for Mobile Robot Navigation across Inconsistent World Models
di: Damm, Eric R., et al.
Pubblicazione: (2025) -
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models
di: Zhang, Qiyao, et al.
Pubblicazione: (2026) -
AutoFly: Vision-Language-Action Model for UAV Autonomous Navigation in the Wild
di: Sun, Xiaolou, et al.
Pubblicazione: (2026)