RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
Fuente:
arXiv
Guardado en:
| Autores principales: | Serpiva, Valerii, Lykov, Artem, Myshlyaev, Artyom, Khan, Muhammad Haris, Abdulkarim, Ali Alridha, Sautenkov, Oleg, Tsetserukou, Dzmitry |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
por: Lykov, Artem, et al.
Publicado: (2025)
por: Lykov, Artem, et al.
Publicado: (2025)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
por: Serpiva, Valerii, et al.
Publicado: (2024)
por: Serpiva, Valerii, et al.
Publicado: (2024)
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
por: Yaqoot, Yasheerah, et al.
Publicado: (2025)
por: Yaqoot, Yasheerah, et al.
Publicado: (2025)
Evolution 6.0: Robot Evolution through Generative Design
por: Khan, Muhammad Haris, et al.
Publicado: (2025)
por: Khan, Muhammad Haris, et al.
Publicado: (2025)
H-RINS: Hierarchical Tightly-coupled Radar-Inertial Navigation via Smoothing and Mapping
por: Abdulkarim, Ali Alridha, et al.
Publicado: (2026)
por: Abdulkarim, Ali Alridha, et al.
Publicado: (2026)
FlightDiffusion: Revolutionising Autonomous Drone Training with Diffusion Models Generating FPV Video
por: Serpiva, Valerii, et al.
Publicado: (2025)
por: Serpiva, Valerii, et al.
Publicado: (2025)
DiffusionCinema: Text-to-Aerial Cinematography
por: Serpiva, Valerii, et al.
Publicado: (2026)
por: Serpiva, Valerii, et al.
Publicado: (2026)
GazeRace: Revolutionizing Remote Piloting with Eye-Gaze Control
por: Tokmurziyev, Issatay, et al.
Publicado: (2024)
por: Tokmurziyev, Issatay, et al.
Publicado: (2024)
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
por: Serpiva, Valerii, et al.
Publicado: (2026)
por: Serpiva, Valerii, et al.
Publicado: (2026)
GestLLM: Advanced Hand Gesture Interpretation via Large Language Models for Human-Robot Interaction
por: Kobzarev, Oleg, et al.
Publicado: (2025)
por: Kobzarev, Oleg, et al.
Publicado: (2025)
SafeSwarm: Decentralized Safe RL for the Swarm of Drones Landing in Dense Crowds
por: Tadevosyan, Grik, et al.
Publicado: (2025)
por: Tadevosyan, Grik, et al.
Publicado: (2025)
GestOS: Advanced Hand Gesture Interpretation via Large Language Models to control Any Type of Robot
por: Lykov, Artem, et al.
Publicado: (2025)
por: Lykov, Artem, et al.
Publicado: (2025)
DroneVLA: VLA based Aerial Manipulation
por: Mehboob, Fawad, et al.
Publicado: (2026)
por: Mehboob, Fawad, et al.
Publicado: (2026)
AgilePilot: DRL-Based Drone Agent for Real-Time Motion Planning in Dynamic Environments by Leveraging Object Detection
por: Khan, Roohan Ahmed, et al.
Publicado: (2025)
por: Khan, Roohan Ahmed, et al.
Publicado: (2025)
FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
por: Sautenkov, Oleg, et al.
Publicado: (2025)
por: Sautenkov, Oleg, et al.
Publicado: (2025)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
por: Gbagbe, Koffivi Fidèle, et al.
Publicado: (2024)
por: Gbagbe, Koffivi Fidèle, et al.
Publicado: (2024)
CognitiveDog: Large Multimodal Model Based System to Translate Vision and Language into Action of Quadruped Robot
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
SwarmDiffusion: End-To-End Traversability-Guided Diffusion for Embodiment-Agnostic Navigation of Heterogeneous Robots
por: Zhura, Iana, et al.
Publicado: (2025)
por: Zhura, Iana, et al.
Publicado: (2025)
UAV-VLPA*: A Vision-Language-Path-Action System for Optimal Route Generation on a Large Scales
por: Sautenkov, Oleg, et al.
Publicado: (2025)
por: Sautenkov, Oleg, et al.
Publicado: (2025)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
por: Guo, Ziang, et al.
Publicado: (2024)
por: Guo, Ziang, et al.
Publicado: (2024)
MorphoNavi: Aerial-Ground Robot Navigation with Object Oriented Mapping in Digital Twin
por: Karaf, Sausar, et al.
Publicado: (2025)
por: Karaf, Sausar, et al.
Publicado: (2025)
AttentionSwarm: Reinforcement Learning with Attention Control Barier Function for Crazyflie Drones in Dynamic Environments
por: Tadevosyan, Grik, et al.
Publicado: (2025)
por: Tadevosyan, Grik, et al.
Publicado: (2025)
Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
UAV-CodeAgents: Scalable UAV Mission Planning via Multi-Agent ReAct and Vision-Language Reasoning
por: Sautenkov, Oleg, et al.
Publicado: (2025)
por: Sautenkov, Oleg, et al.
Publicado: (2025)
FlightAR: AR Flight Assistance Interface with Multiple Video Streams and Object Detection Aimed at Immersive Drone Control
por: Sautenkov, Oleg, et al.
Publicado: (2024)
por: Sautenkov, Oleg, et al.
Publicado: (2024)
PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
por: Lykov, Artem, et al.
Publicado: (2025)
por: Lykov, Artem, et al.
Publicado: (2025)
HumanDiffusion: A Vision-Based Diffusion Trajectory Planner with Human-Conditioned Goals for Search and Rescue UAV
por: Batool, Faryal, et al.
Publicado: (2026)
por: Batool, Faryal, et al.
Publicado: (2026)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
por: Gubernatorov, Konstantin, et al.
Publicado: (2025)
por: Gubernatorov, Konstantin, et al.
Publicado: (2025)
Shake-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Manipulations and Liquid Mixing
por: Khan, Muhamamd Haris, et al.
Publicado: (2025)
por: Khan, Muhamamd Haris, et al.
Publicado: (2025)
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
por: Konenkov, Mikhail, et al.
Publicado: (2024)
por: Konenkov, Mikhail, et al.
Publicado: (2024)
VLH: Vision-Language-Haptics Foundation Model
por: Fuentes, Luis Francisco Moreno, et al.
Publicado: (2025)
por: Fuentes, Luis Francisco Moreno, et al.
Publicado: (2025)
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
por: Lykov, Artem, et al.
Publicado: (2024)
por: Lykov, Artem, et al.
Publicado: (2024)
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
por: Mahmoud, Yara, et al.
Publicado: (2025)
por: Mahmoud, Yara, et al.
Publicado: (2025)
FADet: A Multi-sensor 3D Object Detection Network based on Local Featured Attention
por: Guo, Ziang, et al.
Publicado: (2024)
por: Guo, Ziang, et al.
Publicado: (2024)
LLM-Glasses: GenAI-driven Glasses with Haptic Feedback for Navigation of Visually Impaired People
por: Tokmurziyev, Issatay, et al.
Publicado: (2025)
por: Tokmurziyev, Issatay, et al.
Publicado: (2025)
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
por: Bazhenov, Artem, et al.
Publicado: (2024)
por: Bazhenov, Artem, et al.
Publicado: (2024)
SwarmPath: Drone Swarm Navigation through Cluttered Environments Leveraging Artificial Potential Field and Impedance Control
por: Khan, Roohan Ahmed, et al.
Publicado: (2024)
por: Khan, Roohan Ahmed, et al.
Publicado: (2024)
ImpedanceGPT: VLM-driven Impedance Control of Swarm of Mini-drones for Intelligent Navigation in Dynamic Environment
por: Batool, Faryal, et al.
Publicado: (2025)
por: Batool, Faryal, et al.
Publicado: (2025)
MissionGPT: Mission Planner for Mobile Robot based on Robotics Transformer Model
por: Berman, Vladimir, et al.
Publicado: (2024)
por: Berman, Vladimir, et al.
Publicado: (2024)
Ejemplares similares
-
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
por: Lykov, Artem, et al.
Publicado: (2025) -
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
por: Serpiva, Valerii, et al.
Publicado: (2024) -
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
por: Yaqoot, Yasheerah, et al.
Publicado: (2025) -
Evolution 6.0: Robot Evolution through Generative Design
por: Khan, Muhammad Haris, et al.
Publicado: (2025) -
H-RINS: Hierarchical Tightly-coupled Radar-Inertial Navigation via Smoothing and Mapping
por: Abdulkarim, Ali Alridha, et al.
Publicado: (2026)