UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Sautenkov, Oleg, Yaqoot, Yasheerah, Lykov, Artem, Mustafa, Muhammad Ahsan, Tadevosyan, Grik, Akhmetkazy, Aibek, Cabrera, Miguel Altamirano, Martynov, Mikhail, Karaf, Sausar, Tsetserukou, Dzmitry |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
UAV-VLPA*: A Vision-Language-Path-Action System for Optimal Route Generation on a Large Scales
por: Sautenkov, Oleg, et al.
Publicado: (2025)
por: Sautenkov, Oleg, et al.
Publicado: (2025)
MorphoMove: Bi-Modal Path Planner with MPC-based Path Follower for Multi-Limb Morphogenetic UAV
por: Mustafa, Muhammad Ahsan, et al.
Publicado: (2024)
por: Mustafa, Muhammad Ahsan, et al.
Publicado: (2024)
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
por: Yaqoot, Yasheerah, et al.
Publicado: (2025)
por: Yaqoot, Yasheerah, et al.
Publicado: (2025)
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
por: Lykov, Artem, et al.
Publicado: (2025)
por: Lykov, Artem, et al.
Publicado: (2025)
MorphoNavi: Aerial-Ground Robot Navigation with Object Oriented Mapping in Digital Twin
por: Karaf, Sausar, et al.
Publicado: (2025)
por: Karaf, Sausar, et al.
Publicado: (2025)
UAV-CodeAgents: Scalable UAV Mission Planning via Multi-Agent ReAct and Vision-Language Reasoning
por: Sautenkov, Oleg, et al.
Publicado: (2025)
por: Sautenkov, Oleg, et al.
Publicado: (2025)
FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
SafeSwarm: Decentralized Safe RL for the Swarm of Drones Landing in Dense Crowds
por: Tadevosyan, Grik, et al.
Publicado: (2025)
por: Tadevosyan, Grik, et al.
Publicado: (2025)
FlightAR: AR Flight Assistance Interface with Multiple Video Streams and Object Detection Aimed at Immersive Drone Control
por: Sautenkov, Oleg, et al.
Publicado: (2024)
por: Sautenkov, Oleg, et al.
Publicado: (2024)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
por: Mahmoud, Yara, et al.
Publicado: (2026)
por: Mahmoud, Yara, et al.
Publicado: (2026)
ImpedanceDiffusion: Diffusion-Based Global Path Planning for UAV Swarm Navigation with Generative Impedance Control
por: Batool, Faryal, et al.
Publicado: (2026)
por: Batool, Faryal, et al.
Publicado: (2026)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
por: Serpiva, Valerii, et al.
Publicado: (2025)
por: Serpiva, Valerii, et al.
Publicado: (2025)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
por: Gbagbe, Koffivi Fidèle, et al.
Publicado: (2024)
por: Gbagbe, Koffivi Fidèle, et al.
Publicado: (2024)
GestLLM: Advanced Hand Gesture Interpretation via Large Language Models for Human-Robot Interaction
por: Kobzarev, Oleg, et al.
Publicado: (2025)
por: Kobzarev, Oleg, et al.
Publicado: (2025)
GestOS: Advanced Hand Gesture Interpretation via Large Language Models to control Any Type of Robot
por: Lykov, Artem, et al.
Publicado: (2025)
por: Lykov, Artem, et al.
Publicado: (2025)
MARLander: A Local Path Planning for Drone Swarms using Multiagent Deep Reinforcement Learning
por: Aschu, Demetros, et al.
Publicado: (2024)
por: Aschu, Demetros, et al.
Publicado: (2024)
AirNeRF: 3D Reconstruction of Human with Drone and NeRF for Future Communication Systems
por: Kotcov, Alexey, et al.
Publicado: (2024)
por: Kotcov, Alexey, et al.
Publicado: (2024)
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
DroneVLA: VLA based Aerial Manipulation
por: Mehboob, Fawad, et al.
Publicado: (2026)
por: Mehboob, Fawad, et al.
Publicado: (2026)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
por: Serpiva, Valerii, et al.
Publicado: (2024)
por: Serpiva, Valerii, et al.
Publicado: (2024)
DiffusionCinema: Text-to-Aerial Cinematography
por: Serpiva, Valerii, et al.
Publicado: (2026)
por: Serpiva, Valerii, et al.
Publicado: (2026)
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
por: Konenkov, Mikhail, et al.
Publicado: (2024)
por: Konenkov, Mikhail, et al.
Publicado: (2024)
MorphoGear: An UAV with Multi-Limb Morphogenetic Gear for Rough-Terrain Locomotion
por: Martynov, Mikhail, et al.
Publicado: (2024)
por: Martynov, Mikhail, et al.
Publicado: (2024)
Evolution 6.0: Robot Evolution through Generative Design
por: Khan, Muhammad Haris, et al.
Publicado: (2025)
por: Khan, Muhammad Haris, et al.
Publicado: (2025)
FlightDiffusion: Revolutionising Autonomous Drone Training with Diffusion Models Generating FPV Video
por: Serpiva, Valerii, et al.
Publicado: (2025)
por: Serpiva, Valerii, et al.
Publicado: (2025)
MissionGPT: Mission Planner for Mobile Robot based on Robotics Transformer Model
por: Berman, Vladimir, et al.
Publicado: (2024)
por: Berman, Vladimir, et al.
Publicado: (2024)
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
por: Bazhenov, Artem, et al.
Publicado: (2024)
por: Bazhenov, Artem, et al.
Publicado: (2024)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
por: Guo, Ziang, et al.
Publicado: (2024)
por: Guo, Ziang, et al.
Publicado: (2024)
TiltXter: CNN-based Electro-tactile Rendering of Tilt Angle for Telemanipulation of Pasteur Pipettes
por: Cabrera, Miguel Altamirano, et al.
Publicado: (2024)
por: Cabrera, Miguel Altamirano, et al.
Publicado: (2024)
DeepXPalm: Tilt and Position Rendering using Palm-worn Haptic Display and CNN-based Tactile Pattern Recognition
por: Miguel, Altamirano Cabrera, et al.
Publicado: (2022)
por: Miguel, Altamirano Cabrera, et al.
Publicado: (2022)
SwarmVLM: VLM-Guided Impedance Control for Autonomous Navigation of Heterogeneous Robots in Dynamic Warehousing
por: Zafar, Malaika, et al.
Publicado: (2025)
por: Zafar, Malaika, et al.
Publicado: (2025)
CognitiveDog: Large Multimodal Model Based System to Translate Vision and Language into Action of Quadruped Robot
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
AttentionSwarm: Reinforcement Learning with Attention Control Barier Function for Crazyflie Drones in Dynamic Environments
por: Tadevosyan, Grik, et al.
Publicado: (2025)
por: Tadevosyan, Grik, et al.
Publicado: (2025)
FADet: A Multi-sensor 3D Object Detection Network based on Local Featured Attention
por: Guo, Ziang, et al.
Publicado: (2024)
por: Guo, Ziang, et al.
Publicado: (2024)
ImpedanceGPT: VLM-driven Impedance Control of Swarm of Mini-drones for Intelligent Navigation in Dynamic Environment
por: Batool, Faryal, et al.
Publicado: (2025)
por: Batool, Faryal, et al.
Publicado: (2025)
Glove2UAV: A Wearable IMU-Based Glove for Intuitive Control of UAV
por: Habel, Amir, et al.
Publicado: (2026)
por: Habel, Amir, et al.
Publicado: (2026)
HoverAI: An Embodied Aerial Agent for Natural Human-Drone Interaction
por: Jin, Yuhua, et al.
Publicado: (2026)
por: Jin, Yuhua, et al.
Publicado: (2026)
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
por: Khan, Muhamamd Haris, et al.
Publicado: (2025)
por: Khan, Muhamamd Haris, et al.
Publicado: (2025)
PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
por: Lykov, Artem, et al.
Publicado: (2025)
por: Lykov, Artem, et al.
Publicado: (2025)
Ejemplares similares
-
UAV-VLPA*: A Vision-Language-Path-Action System for Optimal Route Generation on a Large Scales
por: Sautenkov, Oleg, et al.
Publicado: (2025) -
MorphoMove: Bi-Modal Path Planner with MPC-based Path Follower for Multi-Limb Morphogenetic UAV
por: Mustafa, Muhammad Ahsan, et al.
Publicado: (2024) -
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
por: Yaqoot, Yasheerah, et al.
Publicado: (2025) -
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
por: Lykov, Artem, et al.
Publicado: (2025) -
MorphoNavi: Aerial-Ground Robot Navigation with Object Oriented Mapping in Digital Twin
por: Karaf, Sausar, et al.
Publicado: (2025)