CognitiveDog: Large Multimodal Model Based System to Translate Vision and Language into Action of Quadruped Robot
Fuente:
arXiv
Salvato in:
| Autori principali: | Lykov, Artem, Litvinov, Mikhail, Konenkov, Mikhail, Prochii, Rinat, Burtsev, Nikita, Abdulkarim, Ali Alridha, Bazhenov, Artem, Berman, Vladimir, Tsetserukou, Dzmitry |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
di: Bazhenov, Artem, et al.
Pubblicazione: (2024)
di: Bazhenov, Artem, et al.
Pubblicazione: (2024)
MissionGPT: Mission Planner for Mobile Robot based on Robotics Transformer Model
di: Berman, Vladimir, et al.
Pubblicazione: (2024)
di: Berman, Vladimir, et al.
Pubblicazione: (2024)
H-RINS: Hierarchical Tightly-coupled Radar-Inertial Navigation via Smoothing and Mapping
di: Abdulkarim, Ali Alridha, et al.
Pubblicazione: (2026)
di: Abdulkarim, Ali Alridha, et al.
Pubblicazione: (2026)
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
di: Konenkov, Mikhail, et al.
Pubblicazione: (2024)
di: Konenkov, Mikhail, et al.
Pubblicazione: (2024)
LLM-MARS: Large Language Model for Behavior Tree Generation and NLP-enhanced Dialogue in Multi-Agent Robot Systems
di: Lykov, Artem, et al.
Pubblicazione: (2023)
di: Lykov, Artem, et al.
Pubblicazione: (2023)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
CognitiveOS: Large Multimodal Model based System to Endow Any Type of Robot with Generative AI
di: Lykov, Artem, et al.
Pubblicazione: (2024)
di: Lykov, Artem, et al.
Pubblicazione: (2024)
FlockGPT: Guiding UAV Flocking with Linguistic Orchestration
di: Lykov, Artem, et al.
Pubblicazione: (2024)
di: Lykov, Artem, et al.
Pubblicazione: (2024)
GestLLM: Advanced Hand Gesture Interpretation via Large Language Models for Human-Robot Interaction
di: Kobzarev, Oleg, et al.
Pubblicazione: (2025)
di: Kobzarev, Oleg, et al.
Pubblicazione: (2025)
GestOS: Advanced Hand Gesture Interpretation via Large Language Models to control Any Type of Robot
di: Lykov, Artem, et al.
Pubblicazione: (2025)
di: Lykov, Artem, et al.
Pubblicazione: (2025)
PhysicalAgent: Towards General Cognitive Robotics with Foundation World Models
di: Lykov, Artem, et al.
Pubblicazione: (2025)
di: Lykov, Artem, et al.
Pubblicazione: (2025)
RaceVLA: VLA-based Racing Drone Navigation with Human-like Behaviour
di: Serpiva, Valerii, et al.
Pubblicazione: (2025)
di: Serpiva, Valerii, et al.
Pubblicazione: (2025)
Robots Can Feel: LLM-based Framework for Robot Ethical Reasoning
di: Lykov, Artem, et al.
Pubblicazione: (2024)
di: Lykov, Artem, et al.
Pubblicazione: (2024)
Bi-VLA: Vision-Language-Action Model-Based System for Bimanual Robotic Dexterous Manipulations
di: Gbagbe, Koffivi Fidèle, et al.
Pubblicazione: (2024)
di: Gbagbe, Koffivi Fidèle, et al.
Pubblicazione: (2024)
HyperSurf: Quadruped Robot Leg Capable of Surface Recognition with GRU and Real-to-Sim Transferring
di: Satsevich, Sergei, et al.
Pubblicazione: (2024)
di: Satsevich, Sergei, et al.
Pubblicazione: (2024)
HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
Echo: An Open-Source, Low-Cost Teleoperation System with Force Feedback for Dataset Collection in Robot Learning
di: Bazhenov, Artem, et al.
Pubblicazione: (2025)
di: Bazhenov, Artem, et al.
Pubblicazione: (2025)
Evolution 6.0: Robot Evolution through Generative Design
di: Khan, Muhammad Haris, et al.
Pubblicazione: (2025)
di: Khan, Muhammad Haris, et al.
Pubblicazione: (2025)
OmniRace: 6D Hand Pose Estimation for Intuitive Guidance of Racing Drone
di: Serpiva, Valerii, et al.
Pubblicazione: (2024)
di: Serpiva, Valerii, et al.
Pubblicazione: (2024)
DiffusionCinema: Text-to-Aerial Cinematography
di: Serpiva, Valerii, et al.
Pubblicazione: (2026)
di: Serpiva, Valerii, et al.
Pubblicazione: (2026)
FADet: A Multi-sensor 3D Object Detection Network based on Local Featured Attention
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
UAV-VLPA*: A Vision-Language-Path-Action System for Optimal Route Generation on a Large Scales
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
UAV-VLA: Vision-Language-Action System for Large Scale Aerial Mission Generation
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
UAV-VLRR: Vision-Language Informed NMPC for Rapid Response in UAV Search and Rescue
di: Yaqoot, Yasheerah, et al.
Pubblicazione: (2025)
di: Yaqoot, Yasheerah, et al.
Pubblicazione: (2025)
Industry 6.0: New Generation of Industry driven by Generative AI and Swarm of Heterogeneous Robots
di: Lykov, Artem, et al.
Pubblicazione: (2024)
di: Lykov, Artem, et al.
Pubblicazione: (2024)
DreamToNav: Generalizable Navigation for Robots via Generative Video Planning
di: Serpiva, Valerii, et al.
Pubblicazione: (2026)
di: Serpiva, Valerii, et al.
Pubblicazione: (2026)
CognitiveDrone: A VLA Model and Evaluation Benchmark for Real-Time Cognitive Task Solving and Reasoning in UAVs
di: Lykov, Artem, et al.
Pubblicazione: (2025)
di: Lykov, Artem, et al.
Pubblicazione: (2025)
FlightDiffusion: Revolutionising Autonomous Drone Training with Diffusion Models Generating FPV Video
di: Serpiva, Valerii, et al.
Pubblicazione: (2025)
di: Serpiva, Valerii, et al.
Pubblicazione: (2025)
METDrive: Multi-modal End-to-end Autonomous Driving with Temporal Guidance
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
Learning Elementary Cellular Automata with Transformers
di: Burtsev, Mikhail
Pubblicazione: (2024)
di: Burtsev, Mikhail
Pubblicazione: (2024)
UAV-CodeAgents: Scalable UAV Mission Planning via Multi-Agent ReAct and Vision-Language Reasoning
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
di: Sautenkov, Oleg, et al.
Pubblicazione: (2025)
Closed-Loop Verbal Reinforcement Learning for Task-Level Robotic Planning
di: Plotnikov, Dmitrii, et al.
Pubblicazione: (2026)
di: Plotnikov, Dmitrii, et al.
Pubblicazione: (2026)
HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing
di: Gubernatorov, Konstantin, et al.
Pubblicazione: (2026)
di: Gubernatorov, Konstantin, et al.
Pubblicazione: (2026)
Quadrupedal Robot Skateboard Mounting via Reverse Curriculum Learning
di: Belov, Danil, et al.
Pubblicazione: (2025)
di: Belov, Danil, et al.
Pubblicazione: (2025)
SafeHumanoid: VLM-RAG-driven Control of Upper Body Impedance for Humanoid Robot
di: Mahmoud, Yara, et al.
Pubblicazione: (2025)
di: Mahmoud, Yara, et al.
Pubblicazione: (2025)
MorphoNavi: Aerial-Ground Robot Navigation with Object Oriented Mapping in Digital Twin
di: Karaf, Sausar, et al.
Pubblicazione: (2025)
di: Karaf, Sausar, et al.
Pubblicazione: (2025)
Biologically Inspired Deep Learning Approaches for Fetal Ultrasound Image Classification
di: Prochii, Rinat, et al.
Pubblicazione: (2025)
di: Prochii, Rinat, et al.
Pubblicazione: (2025)
SwarmDiffusion: End-To-End Traversability-Guided Diffusion for Embodiment-Agnostic Navigation of Heterogeneous Robots
di: Zhura, Iana, et al.
Pubblicazione: (2025)
di: Zhura, Iana, et al.
Pubblicazione: (2025)
MorphoGear: An UAV with Multi-Limb Morphogenetic Gear for Rough-Terrain Locomotion
di: Martynov, Mikhail, et al.
Pubblicazione: (2024)
di: Martynov, Mikhail, et al.
Pubblicazione: (2024)
Loss Patterns of Neural Networks
di: Skorokhodov, Ivan, et al.
Pubblicazione: (2019)
di: Skorokhodov, Ivan, et al.
Pubblicazione: (2019)
Documenti analoghi
-
DogSurf: Quadruped Robot Capable of GRU-based Surface Recognition for Blind Person Navigation
di: Bazhenov, Artem, et al.
Pubblicazione: (2024) -
MissionGPT: Mission Planner for Mobile Robot based on Robotics Transformer Model
di: Berman, Vladimir, et al.
Pubblicazione: (2024) -
H-RINS: Hierarchical Tightly-coupled Radar-Inertial Navigation via Smoothing and Mapping
di: Abdulkarim, Ali Alridha, et al.
Pubblicazione: (2026) -
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications
di: Konenkov, Mikhail, et al.
Pubblicazione: (2024) -
LLM-MARS: Large Language Model for Behavior Tree Generation and NLP-enhanced Dialogue in Multi-Agent Robot Systems
di: Lykov, Artem, et al.
Pubblicazione: (2023)