GestOS: Advanced Hand Gesture Interpretation via Large Language Models to control Any Type of Robot
Fuente:
arXiv
Saved in:
| Main Authors: | Lykov, Artem, Kobzarev, Oleg, Tsetserukou, Dzmitry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GestLLM: Advanced Hand Gesture Interpretation via Large Language Models for Human-Robot Interaction
by: Kobzarev, Oleg, et al.
Published: (2025)
by: Kobzarev, Oleg, et al.
Published: (2025)
CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
by: Konenkov, Mikhail, et al.
Published: (2024)
by: Konenkov, Mikhail, et al.
Published: (2024)
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
by: Gbagbe, Koffivi Fidèle, et al.
Published: (2024)
by: Gbagbe, Koffivi Fidèle, et al.
Published: (2024)
LLM-MARS: Large Language Model for Behavior Tree Generation and NLP-enhanced Dialogue in Multi-Agent Robot Systems
by: Lykov, Artem, et al.
Published: (2023)
by: Lykov, Artem, et al.
Published: (2023)
MissionGPT: Mission Planner for Mobile Robot based on Robotics Transformer Model
by: Berman, Vladimir, et al.
Published: (2024)
by: Berman, Vladimir, et al.
Published: (2024)
UAV-VLPA*: A Vision-Language-Path-Action System for Optimal Route Generation on a Large Scales
by: Sautenkov, Oleg, et al.
Published: (2025)
by: Sautenkov, Oleg, et al.
Published: (2025)
CognitiveDog: Large Multimodal Model Based System to Translate Vision and Language into Action of Quadruped Robot
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
by: Yaqoot, Yasheerah, et al.
Published: (2025)
by: Yaqoot, Yasheerah, et al.
Published: (2025)
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
by: Serpiva, Valerii, et al.
Published: (2026)
by: Serpiva, Valerii, et al.
Published: (2026)
DiffusionCinema: Text-to-Aerial Cinematography
by: Serpiva, Valerii, et al.
Published: (2026)
by: Serpiva, Valerii, et al.
Published: (2026)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
by: Guo, Ziang, et al.
Published: (2024)
by: Guo, Ziang, et al.
Published: (2024)
Evolution 6.0: Robot Evolution through Generative Design
by: Khan, Muhammad Haris, et al.
Published: (2025)
by: Khan, Muhammad Haris, et al.
Published: (2025)
FADet: A Multi-sensor 3D Object Detection Network based on Local Featured Attention
by: Guo, Ziang, et al.
Published: (2024)
by: Guo, Ziang, et al.
Published: (2024)
FlightDiffusion: Revolutionising Autonomous Drone Training with Diffusion Models Generating FPV Video
by: Serpiva, Valerii, et al.
Published: (2025)
by: Serpiva, Valerii, et al.
Published: (2025)
UAV-CodeAgents: Scalable UAV Mission Planning via Multi-Agent ReAct and Vision-Language Reasoning
by: Sautenkov, Oleg, et al.
Published: (2025)
by: Sautenkov, Oleg, et al.
Published: (2025)
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
by: Lykov, Artem, et al.
Published: (2025)
by: Lykov, Artem, et al.
Published: (2025)
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
by: Bazhenov, Artem, et al.
Published: (2024)
by: Bazhenov, Artem, et al.
Published: (2024)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
by: Serpiva, Valerii, et al.
Published: (2025)
by: Serpiva, Valerii, et al.
Published: (2025)
PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
by: Lykov, Artem, et al.
Published: (2025)
by: Lykov, Artem, et al.
Published: (2025)
FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
DiffusionRL: Efficient Training of Diffusion Policies for Robotic Grasping Using RL-Adapted Large-Scale Datasets
by: Makarova, Maria, et al.
Published: (2025)
by: Makarova, Maria, et al.
Published: (2025)
CONTHER: Human-Like Contextual Robot Learning via Hindsight Experience Replay and Transformers without Expert Demonstrations
by: Makarova, Maria, et al.
Published: (2025)
by: Makarova, Maria, et al.
Published: (2025)
Echo: An Open-Source, Low-Cost Teleoperation System with Force Feedback for Dataset Collection in Robot Learning
by: Bazhenov, Artem, et al.
Published: (2025)
by: Bazhenov, Artem, et al.
Published: (2025)
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
by: Mahmoud, Yara, et al.
Published: (2025)
by: Mahmoud, Yara, et al.
Published: (2025)
METDrive: Multi-modal End-to-end Autonomous Driving with Temporal Guidance
by: Guo, Ziang, et al.
Published: (2024)
by: Guo, Ziang, et al.
Published: (2024)
MorphoNavi: Aerial-Ground Robot Navigation with Object Oriented Mapping in Digital Twin
by: Karaf, Sausar, et al.
Published: (2025)
by: Karaf, Sausar, et al.
Published: (2025)
GraspSense: Physically Grounded Grasp and Grip Planning for a Dexterous Robotic Hand via Language-Guided Perception and Force Maps
by: Semenyakina, Elizaveta, et al.
Published: (2026)
by: Semenyakina, Elizaveta, et al.
Published: (2026)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
by: Sautenkov, Oleg, et al.
Published: (2025)
by: Sautenkov, Oleg, et al.
Published: (2025)
Quadrupedal Robot Skateboard Mounting via Reverse Curriculum Learning
by: Belov, Danil, et al.
Published: (2025)
by: Belov, Danil, et al.
Published: (2025)
FiDTouch: A 3D Wearable Haptic Display for the Finger Pad
by: Trinitatova, Daria, et al.
Published: (2025)
by: Trinitatova, Daria, et al.
Published: (2025)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
by: Serpiva, Valerii, et al.
Published: (2024)
by: Serpiva, Valerii, et al.
Published: (2024)
H-RINS: Hierarchical Tightly-coupled Radar-Inertial Navigation via Smoothing and Mapping
by: Abdulkarim, Ali Alridha, et al.
Published: (2026)
by: Abdulkarim, Ali Alridha, et al.
Published: (2026)
GrainGrasp: Dexterous Grasp Generation with Fine-grained Contact Guidance
by: Zhao, Fuqiang, et al.
Published: (2024)
by: Zhao, Fuqiang, et al.
Published: (2024)
Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
by: Lykov, Artem, et al.
Published: (2024)
by: Lykov, Artem, et al.
Published: (2024)
HyperSurf: Quadruped Robot Leg Capable of Surface Recognition with GRU and Real-to-Sim Transferring
by: Satsevich, Sergei, et al.
Published: (2024)
by: Satsevich, Sergei, et al.
Published: (2024)
Low-Cost Teleoperation Extension for Mobile Manipulators
by: Belov, Danil, et al.
Published: (2026)
by: Belov, Danil, et al.
Published: (2026)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
by: Gubernatorov, Konstantin, et al.
Published: (2025)
by: Gubernatorov, Konstantin, et al.
Published: (2025)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
by: Mahmoud, Yara, et al.
Published: (2026)
by: Mahmoud, Yara, et al.
Published: (2026)
Similar Items
-
GestLLM: Advanced Hand Gesture Interpretation via Large Language Models for Human-Robot Interaction
by: Kobzarev, Oleg, et al.
Published: (2025) -
CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
by: Lykov, Artem, et al.
Published: (2024) -
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
by: Konenkov, Mikhail, et al.
Published: (2024) -
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
by: Lykov, Artem, et al.
Published: (2024) -
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
by: Gbagbe, Koffivi Fidèle, et al.
Published: (2024)