GestLLM: Advanced Hand Gesture Interpretation via Large Language Models for Human-Robot Interaction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kobzarev, Oleg, Lykov, Artem, Tsetserukou, Dzmitry |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GestOS: Advanced Hand Gesture Interpretation via Large Language Models to control Any Type of Robot
von: Lykov, Artem, et al.
Veröffentlicht: (2025)
von: Lykov, Artem, et al.
Veröffentlicht: (2025)
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
von: Konenkov, Mikhail, et al.
Veröffentlicht: (2024)
von: Konenkov, Mikhail, et al.
Veröffentlicht: (2024)
LLM-MARS: Large Language Model for Behavior Tree Generation and NLP-enhanced Dialogue in Multi-Agent Robot Systems
von: Lykov, Artem, et al.
Veröffentlicht: (2023)
von: Lykov, Artem, et al.
Veröffentlicht: (2023)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
von: Gbagbe, Koffivi Fidèle, et al.
Veröffentlicht: (2024)
von: Gbagbe, Koffivi Fidèle, et al.
Veröffentlicht: (2024)
MissionGPT: Mission Planner for Mobile Robot based on Robotics Transformer Model
von: Berman, Vladimir, et al.
Veröffentlicht: (2024)
von: Berman, Vladimir, et al.
Veröffentlicht: (2024)
UAV-VLPA*: A Vision-Language-Path-Action System for Optimal Route Generation on a Large Scales
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
CognitiveDog: Large Multimodal Model Based System to Translate Vision and Language into Action of Quadruped Robot
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
von: Yaqoot, Yasheerah, et al.
Veröffentlicht: (2025)
von: Yaqoot, Yasheerah, et al.
Veröffentlicht: (2025)
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
von: Serpiva, Valerii, et al.
Veröffentlicht: (2026)
von: Serpiva, Valerii, et al.
Veröffentlicht: (2026)
DiffusionCinema: Text-to-Aerial Cinematography
von: Serpiva, Valerii, et al.
Veröffentlicht: (2026)
von: Serpiva, Valerii, et al.
Veröffentlicht: (2026)
Evolution 6.0: Robot Evolution through Generative Design
von: Khan, Muhammad Haris, et al.
Veröffentlicht: (2025)
von: Khan, Muhammad Haris, et al.
Veröffentlicht: (2025)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
von: Serpiva, Valerii, et al.
Veröffentlicht: (2025)
von: Serpiva, Valerii, et al.
Veröffentlicht: (2025)
FADet: A Multi-sensor 3D Object Detection Network based on Local Featured Attention
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
FlightDiffusion: Revolutionising Autonomous Drone Training with Diffusion Models Generating FPV Video
von: Serpiva, Valerii, et al.
Veröffentlicht: (2025)
von: Serpiva, Valerii, et al.
Veröffentlicht: (2025)
UAV-CodeAgents: Scalable UAV Mission Planning via Multi-Agent ReAct and Vision-Language Reasoning
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
von: Lykov, Artem, et al.
Veröffentlicht: (2025)
von: Lykov, Artem, et al.
Veröffentlicht: (2025)
CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
von: Bazhenov, Artem, et al.
Veröffentlicht: (2024)
von: Bazhenov, Artem, et al.
Veröffentlicht: (2024)
CONTHER: Human-Like Contextual Robot Learning via Hindsight Experience Replay and Transformers without Expert Demonstrations
von: Makarova, Maria, et al.
Veröffentlicht: (2025)
von: Makarova, Maria, et al.
Veröffentlicht: (2025)
PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
von: Lykov, Artem, et al.
Veröffentlicht: (2025)
von: Lykov, Artem, et al.
Veröffentlicht: (2025)
FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
DiffusionRL: Efficient Training of Diffusion Policies for Robotic Grasping Using RL-Adapted Large-Scale Datasets
von: Makarova, Maria, et al.
Veröffentlicht: (2025)
von: Makarova, Maria, et al.
Veröffentlicht: (2025)
Echo: An Open-Source, Low-Cost Teleoperation System with Force Feedback for Dataset Collection in Robot Learning
von: Bazhenov, Artem, et al.
Veröffentlicht: (2025)
von: Bazhenov, Artem, et al.
Veröffentlicht: (2025)
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
von: Mahmoud, Yara, et al.
Veröffentlicht: (2025)
von: Mahmoud, Yara, et al.
Veröffentlicht: (2025)
METDrive: Multi-modal End-to-end Autonomous Driving with Temporal Guidance
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
von: Guo, Ziang, et al.
Veröffentlicht: (2024)
MorphoNavi: Aerial-Ground Robot Navigation with Object Oriented Mapping in Digital Twin
von: Karaf, Sausar, et al.
Veröffentlicht: (2025)
von: Karaf, Sausar, et al.
Veröffentlicht: (2025)
GraspSense: Physically Grounded Grasp and Grip Planning for a Dexterous Robotic Hand via Language-Guided Perception and Force Maps
von: Semenyakina, Elizaveta, et al.
Veröffentlicht: (2026)
von: Semenyakina, Elizaveta, et al.
Veröffentlicht: (2026)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
von: Sautenkov, Oleg, et al.
Veröffentlicht: (2025)
Quadrupedal Robot Skateboard Mounting via Reverse Curriculum Learning
von: Belov, Danil, et al.
Veröffentlicht: (2025)
von: Belov, Danil, et al.
Veröffentlicht: (2025)
FiDTouch: A 3D Wearable Haptic Display for the Finger Pad
von: Trinitatova, Daria, et al.
Veröffentlicht: (2025)
von: Trinitatova, Daria, et al.
Veröffentlicht: (2025)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
von: Serpiva, Valerii, et al.
Veröffentlicht: (2024)
von: Serpiva, Valerii, et al.
Veröffentlicht: (2024)
H-RINS: Hierarchical Tightly-coupled Radar-Inertial Navigation via Smoothing and Mapping
von: Abdulkarim, Ali Alridha, et al.
Veröffentlicht: (2026)
von: Abdulkarim, Ali Alridha, et al.
Veröffentlicht: (2026)
GrainGrasp: Dexterous Grasp Generation with Fine-grained Contact Guidance
von: Zhao, Fuqiang, et al.
Veröffentlicht: (2024)
von: Zhao, Fuqiang, et al.
Veröffentlicht: (2024)
Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
von: Lykov, Artem, et al.
Veröffentlicht: (2024)
HyperSurf: Quadruped Robot Leg Capable of Surface Recognition with GRU and Real-to-Sim Transferring
von: Satsevich, Sergei, et al.
Veröffentlicht: (2024)
von: Satsevich, Sergei, et al.
Veröffentlicht: (2024)
Low-Cost Teleoperation Extension for Mobile Manipulators
von: Belov, Danil, et al.
Veröffentlicht: (2026)
von: Belov, Danil, et al.
Veröffentlicht: (2026)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
HumanoidVLM: Vision-Language-Guided Impedance Control for Contact-Rich Humanoid Manipulation
von: Mahmoud, Yara, et al.
Veröffentlicht: (2026)
von: Mahmoud, Yara, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GestOS: Advanced Hand Gesture Interpretation via Large Language Models to control Any Type of Robot
von: Lykov, Artem, et al.
Veröffentlicht: (2025) -
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
von: Lykov, Artem, et al.
Veröffentlicht: (2024) -
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
von: Konenkov, Mikhail, et al.
Veröffentlicht: (2024) -
LLM-MARS: Large Language Model for Behavior Tree Generation and NLP-enhanced Dialogue in Multi-Agent Robot Systems
von: Lykov, Artem, et al.
Veröffentlicht: (2023) -
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
von: Guo, Ziang, et al.
Veröffentlicht: (2024)